This course dives deep into Amazon Kinesis and Amazon MSK through a mix of instructor-led presentations, hands-on labs, demonstrations, and class exercises to help you fully understand how to build a streaming data analytics solution on AWS. You will also learn how to scale streaming applications using Amazon Kinesis, optimize data storage, select and deploy appropriate options for ingesting, transforming, storing, and analyzing data, and more.
Module 1: Overview of Data Analytics and the Data Pipeline
- Data analytics use cases
- Using the data pipeline for analytics
Module 2: Using Streaming Services in the Data Analytics Pipeline
- The importance of streaming data analytics
- The streaming data analytics pipeline
- Streaming concepts
Module 3: Introduction to AWS streaming services
- Streaming data services in AWS
- Amazon Kinesis in analytics solutions
- Demo: Explore Amazon Kinesis Data Streams
- Practice lab: Setting up a streaming delivery pipeline with Amazon Kinesis
- Using Amazon Kinesis Data Analytics
- Introduction to Amazon MSK
- Overview of Spark Streaming
Module 4: Using Amazon Kinesis for Real-time Data Analytics
- Exploring Amazon Kinesis using a clickstream workload
- Creating data and delivery streams with Kinesis
- Demo: Understanding producers and consumers
- Building stream producers
- Building stream consumers
- Building and deploying Flink applications in Kinesis Data Analytics
- Demonstration: Explore Zeppelin notebooks for Kinesis Data Analytics
- Practice lab: Streaming analytics with Amazon Kinesis Data Analytics and Apache Flink
Module 5: Securing, Monitoring, and Optimizing Amazon Kinesis
- Optimize Amazon Kinesis to gain actionable business insights
- Security and monitoring best practices
Module 6: Using Amazon MSK in Streaming Data Analytics Solutions
- Use-cases for Amazon MSK
- Creating MSK clusters
- Demo: Provisioning an MSK cluster
- Ingesting data into Amazon MSK
- Practice Lab: Introduction to access control with Amazon MSK
- Transforming and processing in Amazon MSK
Module 7: Securing, Monitoring, and Optimizing Amazon MSK
- Optimizing Amazon MSK
- Demo: Scaling up Amazon MSK storage
- Practice lab: Amazon MSK streaming pipeline and application deployment
- Security and monitoring
- Demo: Monitoring an MSK cluster
Module 8: Designing Streaming Data Analytics Solutions
- Use-case review
- Class exercise: Designing a streaming data analytics workflow
Module 9: Developing Modern Data Architectures on AWS
- Modern data architectures
This course is intended for the following job roles:
We recommend that attendees of this course have the following prerequisites:
- a minimum one-year experience managing data analytics solutions or streaming data
- have attended the following course (or have equivalent knowledge): Building Batch Data Analytics Solutions on AWS