by Andrea Bacqué | Aug 20, 2020 | Apache Spark, Customer Experience, Python, SAS, Solutions
The concept of dark data – i.e. information assets organizations collect, process and store during regular business activities, but generally fail to use for other purposes such as analytics, business relationships and direct monetizing – isn’t new...
by Andrea Bacqué | Jun 19, 2020 | Apache Spark, Apache Spark Cafe, Customer Experience, Python, SAS, Solutions
The Wise With Data Team have been eagerly anticipating Spark 3.0’s release and it is official as of today. As experts in open source data science, we’d like to share what to expect with Spark 3.0. Also a reminder that Spark Summit 2020 is next week… What are we...
by Bryan Chuinkam | May 20, 2020 | Apache Spark, Apache Spark Cafe, Python
What is Koalas? Koalas is an implementation of the pandas DataFrame API on top of Apache Spark. Pandas is the go-to Python Library for data analysis, while Apache Spark is becoming the go to for big data processing. Koalas allows you leverage the simplicity of Pandas...
by Mike Sun | May 20, 2020 | Apache Spark, Apache Spark Cafe
Since year end 2014, there has been an increase in the number of Google searches comparing Apache Spark to Hadoop. What brings people who are experts in Big Data, Data Science, and Data Analysis to Apache Spark (Spark)? Spark is a fast and expressive cluster computing...