Upcoming Sessions
-
September
23
ILT - ADMIN-230: Administering Cloudera on premises - 4545287
Starting:2025/09/23 @ 09:00 AM (GMT+02:00) BudapestEnding:2025/09/26 @ 05:00 PM (GMT+02:00) BudapestType:Multi-day Session -
September
29
ILT - ADMIN-332: Securing Cloudera on premises - 4353561
Starting:2025/09/29 @ 09:00 AM (GMT+02:00) BudapestEnding:2025/10/02 @ 05:00 PM (GMT+02:00) BudapestType:Multi-day Session
See All Upcoming Sessions

This four-day hands-on training course delivers the key concepts and knowledge developers need to use Apache Spark to develop high-performance, parallel applications on the Cloudera Data Platform. Hands-on exercises allow students to practice writing Spark applications that integrate with Cloudera Data Platform core components. Participants will learn how to use Spark SQL to query structured data, how to use Hive features to ingest and denormalize data, and how to work with “big data” stored in a distributed file system. After taking this course, participants will be prepared to face real-world challenges and build applications to execute faster decisions, better decisions, and interactive analysis, applied to a wide variety of use cases, architectures, and industries. Download full course description What you'll learn During this course, you will learn how to: Distribute, store, and process data in a cluster Write, configure, and deploy Apache Spark applications Use the Spark interpreters and Spark applications to explore, process, and analyze distributed data Query data using Spark SQL, DataFrames, and Hive tables Deploy a Spark application on the Data Engineering Service What to expect This course is designed for developers and data engineers. All students are expected to have basic Linux experience, and basic proficiency with either Python or Scala programming languages. Basic knowledge of SQL is helpful. Prior knowledge of Spark and Hadoop is not required. DATE: November 17-20, 2025 Virtual Classroom, EMEA 9:00 - 17:00 (CET TIMEZONE) Read more

DESCRIPTION DATE: November 17-20, 2025 Virtual Classroom, AMER 9:00 - 17:00 (Central US TIMEZONE) Read more

About This Training This 1-day course by Cloudera Education introduces Cloudera Observability, a service designed to help you interactively understand your environment, data services, workloads, clusters, and resources. Its wide range of metrics and health tests empowers you to identify and troubleshoot both existing and potential problems. Cloudera Observability also provides prescriptive guidance and recommendations, enabling you to quickly address issues and optimize solutions. After a workload is completed, the Cloudera Manager Management Service's Telemetry Publisher collects diagnostic information about the job or query and the processing cluster, sending it to Cloudera Observability for analysis. Download full course description What Skills You Will Gain Participants will develop the following skills: Comprehend the benefits of Cloudera Observability. Configure Cloudera Observability for CDP Public Cloud. Understand Cloudera Observability deployment architecture for both CDP Public and Private Cloud environments. Determine system requirements for workload clusters. Install Cloudera Observability. Classify workloads for analysis using Workload Views. Manage user access to workloads. Configure cost center criteria within Cloudera Observability. Set up action alerts for jobs and queries using Auto Actions. Utilize Cloudera Observability Metastore Analytics. Troubleshoot effectively with Cloudera Observability. What to Expect This course is ideal for new and existing Cloudera Private or Public Cloud users seeking to leverage the benefits of Cloudera Observability. The course is particularly valuable for platform administrators, data practitioners, budget owners/controllers, and solutions architects. A general knowledge of monitoring concepts is helpful. Course Details Introduction What is Cloudera Observability Cloudera Observability Capabilities Cloudera Observability Tech Cloudera Observability Online Cloudera Observability Air-Gapped Strategic Features Observability Essential Adoption Configuration Tasks for CDP Public Cloud Deployment Architecture for CDP Private Cloud Security Considerations System Requirements Network Port Requirements Configuring Telemetry Publisher Redacting Data Adding a Proxy Server Observability Installation Managing Workloads and Users (Self-Service Analytics) Auto-Generating Workload Views Manually Generating Workload Views Managing User Access to Workloads Observability Access Roles Exercise: Creating Workload Views Exercise: Assigning Access Roles in Cloudera Observability Working with Alerts, Costs, & Reports (Financial Governance) Analyzing Environment Costs with Cloudera Observability Triggering Action Alerts Across Jobs and Queries – Auto Actions Working with Cluster Reports Exercise: Configuring Cloudera Observability Cost Center Criteria Exercise: Creating a Cloudera Observability Cost Center Exercise: Displaying Costs Associated with a Cost Center Exercise: Creating an Auto Action Event Understanding, Identifying, & Addressing Problems with Cloudera Observability (Service Health Monitoring) Analyzing Tables with Cloudera Observability Metastore Analytics Validations in Cloudera Observability Analyzing Hive Queries Exercise: Analyzing Hive Queries Exercise: Analyzing Impala Queries Troubleshooting (Expedited Issue Resolution) Troubleshooting Abnormal Job Durations Troubleshooting Job Durations: Task Duration Troubleshooting Failed Jobs Troubleshooting with the Job Comparison Feature Exercise: Analyzing Spark Jobs Exercise: Analyzing MapReduce Jobs November 10, 2025 Virtual Classroom, EMEA 9:00 - 17:00 (CET TIMEZONE) Read more

This four-day instructor-led course begins by introducing Apache Kafka, explaining its key concepts and architecture, and discussing several common use cases. Building on this foundation, you will learn how to plan a Kafka deployment, and then gain hands-on experience by installing and configuring your own cloud-based, multi-node cluster running Kafka on the Cloudera Data Platform (CDP). You will then use this cluster during more than 20 hands-on exercises that follow, covering a range of essential skills, starting with how to create Kafka topics, producers, and consumers, then continuing through progressively more challenging aspects of Kafka operations and development, such as those related to scalability, reliability, and performance problems. Throughout the course, you will learn and use Cloudera’s recommended tools for working with Kafka, including Cloudera Manager, Schema Registry, Streams Messaging Manager, and Cruise Control. DATE: November 10-13, 2025 Virtual Classroom, EMEA 9:00 - 17:00 (CET Timezone) Read more

Overview The Cloudera platform is intended to meet the most demanding technical audit standards. The significant improvements in Cloudera architecture and components make Cloudera “Secure by Design.” This four-day hands-on course is presented as a project plan for Cloudera administrators to build fully secured Cloudera clusters. The course begins with implementing Perimeter Security by installing host level security and Kerberos. Next, students protect Data by implementing Transport Layer Security using Auto-TLS and data encryption using Key Management System and Key Trustee Server (KMS/KTS). Following this, in the third stage, students control access for users and to data using Apache Ranger and Apache Atlas. The fourth stage focuses on visibility practices, teaching students how to audit systems, users, and data usage. Finally, the course introduces Cloudera practices for Risk Management in a fully secured Cloudera platform. This course is 60% exercise and 40% lecture. Who should take this course? This immersion course is designed for Linux Administrators transitioning to Cloudera Administrator roles. Students must have proficiency in Linux (e.g., navigating the file system, using basic commands) and Linux text editors (e.g., vi, nano). Familiarity with Directory Services, Transport Layer Security, Kerberos, and SQL select statements is recommended. Prior experience with Cloudera products is required. Students must have reliable internet access to connect to the classroom environments hosted on Amazon Web Services. DATE: November 10-13, 2025 Virtual Classroom, AMER 9:00 - 17:00 (Central US TIMEZONE Read more

Cloudera is a fully integrated edge to AI product set. Cloudera Manager is purposely built as the DevOps tooling for building and managing the Cloudera platform. This four-day hands-on course presents detailed explanation, comprehensive theory, key skills, and recommended practices for successful platform administration. Upon completion of this course a Cloudera Administrator will learn the full range of functionality and capability of Cloudera Manager. DATE: November 3-6, 2025 Virtual Classroom, AMER 9:00 - 17:00 (Central US TIMEZONE) Read more
Shopping Cart
Your cart is empty