Presented By O’Reilly and Cloudera

San Francisco • London • New York

Make Data Work

21–22 May 2018: Training
22–24 May 2018: Tutorials & Conference
London, UK

In-Person Training
Real-time systems with Spark Streaming and Kafka

Jesse Anderson (Big Data Institute)

Monday, 21 May & Tuesday, 22 May, 9:00 - 17:00

Data engineering and architecture, Streaming systems and real-time applications
Location: Capital Suite 16

Average rating:

(5.00, 1 rating)

Participants should plan to attend both days of this 2-day training course. Platinum and Training passes do not include access to tutorials on Tuesday.

To handle real-time big data, you need to solve two difficult problems: How do you ingest that much data, and how will you process that much data? Jesse Anderson explores the latest real-time frameworks (both open source and managed cloud services), discusses the leading cloud providers, and explains how to choose the right one for your company.

What you'll learn, and how you can apply it

Learn how to how to ingest data, process it, analyze it, and display it in real time in a dashboard with Apache Kafka and Apache Spark

Prerequisites:

A working knowledge of HDFS and Spark (i.e., Spark batch APIs)

Real-time big data frameworks are enabling brand-new use cases, while the cloud is letting us do things cheaper and faster than ever. Together, they’re making it easier to create production real-time systems. But to handle real-time big data, you need to solve two difficult problems: How do you ingest that much data, and how will you process that much data?

Jesse Anderson explores the latest real-time frameworks (both open source and managed cloud services), discusses the leading cloud providers, and explains how to choose the right one for your company. Focusing on Apache Kafka and Apache Spark, Jesse also demonstrates how to ingest data, process it, analyze it, and display it in real time in a dashboard.

For the final exercise, you’ll take data that has been ingested with Kafka and process it with Spark Streaming and visualize it on a web page with D3. This video gives a little more information about the final exercise so you can see the skills you’ll take away from the class.

About your instructor

Jesse Anderson is a data engineer, creative engineer, and managing director of the Big Data Institute. Jesse trains employees on big data—including cutting-edge technology like Apache Kafka, Apache Hadoop, and Apache Spark. He’s taught thousands of students at companies ranging from startups to Fortune 100 companies the skills to become data engineers. He’s widely regarded as an expert in the field and recognized for his novel teaching practices. Jesse is published by O’Reilly and Pragmatic Programmers and has been covered in such prestigious media outlets as the Wall Street Journal, CNN, BBC, NPR, Engadget, and Wired. You can learn more about Jesse at Jesse-Anderson.com.

Conference registration

Get the Platinum pass or the Training pass to add this course to your package.

Comments on this page are now closed.

Comments

Er Allan | SENIOR CONSULTANT

20/05/2018 15:00 BST

Hi, I have checked my email, but can’t find any details on the required preparation/installation instructions.

Jesse Anderson | MANAGING DIRECTOR

19/05/2018 19:10 BST

The email was sent out on Thursday and will be sent out again on Sunday. Make sure you’re looking at the email address that was signed up for and check in your spam.

Jonny Daenen | DATA SCIENTIST / SOFTWARE ENGINEER

19/05/2018 15:46 BST

I’ve received a general conference email, but I cannot find any details on the required preparation/installation. Could someone provide more details on that part?

Thanks!

Jesse Anderson | MANAGING DIRECTOR

17/05/2018 21:24 BST

@Jens yes, you will need to. You will receive an email with the instructions.

Jens Rabe | RESEARCH FELLOW

17/05/2018 10:00 BST

Do I have to pre-install / prepare anything on my computer for this?

Presented by

Elite Sponsors

Exabyte Sponsor

Impact Sponsors

Supporting Sponsor

Sponsorship Opportunities

For exhibition and sponsorship opportunities, email strataconf@oreilly.com

Partner Opportunities

For information on trade opportunities with O'Reilly conferences, email partners@oreilly.com

Contact Us

View a complete list of Strata Data Conference contacts

©2018, O’Reilly UK Ltd • (800) 889-8969 or (707) 827-7019 • Monday-Friday 7:30am-5pm PT • All trademarks and registered trademarks appearing on oreilly.com are the property of their respective owners. • confreg@oreilly.com

In-Person TrainingReal-time systems with Spark Streaming and Kafka