Spark Overview for Scala Analytics
The “Spark Overview for Scala Analytics” course will cover the history of Spark and how it came to be, how to build applications with Spark, establish an understanding of RDDs and DataFrames, and other advanced Spark topics. Apache Spark™ is a fast and general engine for large-scale data processing, with built-in modules for streaming, SQL, machine learning and graph processing. Having finished this class, a student would be prepared to leverage the core RDD and DataFrame APIs to perform analytics on datasets.
4.7 (54 Reviews)

Language
- English
Topic
- Scala
Enrollment Count
- 2.22K
Skills You Will Learn
- Scala, Big Data
Offered By
- LightBend
Estimated Effort
- 8 Hours
Platform
- SkillsNetwork
Last Update
- April 2, 2025
There are 5 modules to this course.
1. What is Spark
2. Introduction to RDDs
3. Introduction to DataFrames
4. Advanced Spark Topics
5. Introduction to Spark MLlib
Requirements
2. No previous Spark knowledge is required
3. No previous experience with Data Science concepts is required. These concepts will be explained as needed
Course Staff
Jamie Allen
Frequently Asked Questions
What web browser should I use?

Language
- English
Topic
- Scala
Enrollment Count
- 2.22K
Skills You Will Learn
- Scala, Big Data
Offered By
- LightBend
Estimated Effort
- 8 Hours
Platform
- SkillsNetwork
Last Update
- April 2, 2025
Instructors
Jamie Allen
SRE Leader and Cloud CTO
Jamie has worked in consulting since 1994, with top firms including Price Waterhouse and Chariot Solutions. He has a long track record of working closely with clients to build highquality, mission critical systems that scale to meet the needs of their businesses, and has worked in myriad industries including automotive, retail, pharmaceuticals, telecommunications, and more. Jamie has been coding in Scala and actor-based systems since 2009 and is the author of "Effective Akka" book from O'Reilly.
Read more