Data Engineering & Analytics
148 courses8 categories
Data engineering and analytics is the work of moving data from where it is produced to where it can be trusted and queried. This topic covers the pipeline end to end: batch and streaming ingestion, warehouse and lakehouse modelling, SQL at scale, and the reporting layer on top.
It sits deliberately apart from machine learning — the courses here are about the data itself: schemas, throughput, correctness and cost, rather than the models that later consume it.
Categories(8)
Data processing and analysis covers the day-to-day work of turning raw operational data into something a person or…
Databases is the foundational layer underneath every persistent application. The category here covers the breadth of…
Elasticsearch is a distributed, RESTful search and analytics engine capable of tackling real-time data. Developed by…
Math & Statistics underpin many aspects of software engineering and data science, providing the foundational tools…
Messaging & Streaming is the backbone of asynchronous communication between services, enabling scalable and reliable…
Power BI is Microsoft's self-service analytics and business-intelligence platform designed to transform raw data into…
Spreadsheets & Low-Code are essential tools for modern data management and automation, often underestimated by…
Courses(148)
Showing 1–30 of 148 courses
Updated 1mo agoPractical course on the internal workings of MySQL, trace analysis, and query optimization to enhance system performance and stability
Updated 1mo agoThe Apache Flink course helps master stream data processing, deploy and optimize real-time pipelines for professional projects.2h 12m
Updated 2mo agoThe only bootcamp you need to become an Excel wizard. Learn Excel Formulas and Functions, Data Visualization, Pivot Tables, VBA Macros, and much more. Start or22h 32m
FreeUpdated 4mo agoLearn core regression models and use them in Python. You study linear, logistic, log, and Cox models with clear steps and real data.6h 20m
Updated 4mo agoYou learn core inferential stats like intervals, tests, ANOVA, and run them in Python. The course shows how to read messy data and make clear data decisions.9h 25m
UpdatedBuild a basic relational database engine yourself in this hands-on CS Primer course covering core DBMS design and queries.18h 30m3/5
Updated 6mo agoStudy the internal architecture and optimization of PostgreSQL. Focus on performance, tracing, indexes, and other key database mechanisms.
Updated 6mo agoStudy effective schema design, indexing, and query optimization in MySQL. The course is suitable for application developers of varying skill levels.7h 41m
Updated 6mo agoStart learning MySQL with basic SQL queries and delve into indexes, caching, transactions, and performance analysis with MySQL Trace Tool.
Updated 7mo agoLearn to build streaming pipelines with Apache Kafka and Flink, create data lakes on AWS, run ML workflows on Spark, and integrate LLM models.16h 46m
Updated 7mo agoThis course is designed with a lot of details, so that everyone, even with very little programming experience can understand and do it by themselves. I strongly18h 51m
Updated 7mo agoStudy Apache Spark and PySpark for big data processing. Practical assignments will help you acquire key skills of a data engineer.2h 20m
Updated 10mo agoThis course is designed for developers, database engineers, data specialists, and ML engineers preparing for SQL interviews.
UpdatedLearn analytics engineering hands-on: build a Snowflake warehouse, load data with Fivetran, transform it with DBT, and visualize it.12h 46m
UpdatedBuild a real semantic search pipeline for logs: FastAPI, embeddings stored in qdrant, a Streamlit UI, and a DuckDB comparison.53m
UpdatedLearn Apache Iceberg hands-on: schema evolution, time travel, and a real local Lakehouse lab with Docker, Spark, and MinIO.33m
UpdatedLearn Apache Airflow from architecture to advanced orchestration: retries, failure handling, sensors, and Spark integration.2h 21m
UpdatedLearn Apache Kafka from scratch: architecture, producers and consumers, delivery guarantees, Connect, and Schema Registry.2h 33m5/5
FreeUpdatedFree course on scaling machine learning with Spark ML: regression, classification, feature engineering, and hyperparameter tuning.2h 7m
Updated 11mo agoUnlock the potential of your PostgreSQL setup with our comprehensive course designed for performance optimization .12h 27m5/5
Updated 1y agoThis course is dedicated to the study of Database Management Systems (DBMS) - technologies that allow for efficient storage, processing, and protection of data.21h 30m5/5
UpdatedBuild a real Azure ETL pipeline with Terraform: Data Factory, Synapse Analytics, Power BI, and a Medallion Lakehouse architecture.4h 20m
Updated 1y agoMaster key technologies with a practical approach! You will gain applied knowledge, clear explanations.16m
Updated 1y agoImagine you have an AI assistant ready to solve any task in Excel. This course will show you how to unlock the full potential of ChatGPT even with basic.3h 23m
Updated 1y agoIn this practical course, you will learn how to build a complete data pipeline on the AWS platform - from obtaining data from the Twitter API to analysis, stora1h 33m5/5
Updated 1y agoEnhance your skills in managing time series data with this comprehensive course.2h 11m5/5
UpdatedSimulate 100,000 users and a million check-ins, then trace movement patterns using Elasticsearch, Kibana, and a Streamlit map UI.1h 37m
Updated 1y agoBig Data is not just a buzzword, but a real phenomenon. Every day, companies around the world collect and process vast amounts of data at high speeds.7h 3m
UpdatedBuild a Dockerized ETL pipeline on AWS: pull live weather data into TDengine and visualize it with Grafana, step by step.29m
UpdatedBuild a real-time invoice streaming pipeline with Kafka, Spark, and MongoDB, visualized through a live Streamlit dashboard.2h 46m