Modern veri platformlarını öğrenmek için kapsamlı, pratik odaklı Docker tabanlı eğitim ortamı.
-
Updated
Sep 23, 2026 - Python
Modern veri platformlarını öğrenmek için kapsamlı, pratik odaklı Docker tabanlı eğitim ortamı.
end-to-end Big Data pipeline that demonstrates advanced proficiency in distributed computing, data engineering, and automated validation. Rather than relying on a single script or a simplified analytics task, the project is designed to simulate a real-world enterprise data architecture where raw data is generated at scale
Repositorio con proyectos y laboratorios de procesamiento de datos utilizando Databricks, Apache Spark y Python. Incluye conceptos clave de Big Data, almacenamiento, procesamiento, análisis y aprendizaje automático.
2023-2 Big Data Project : 응급 의료 취약 지역 탐색 및 잠재변수 유효성 검증
Pyspark RDD, DataFrame and Dataset Examples in Python language
A curated list of awesome Apache Spark packages and resources.
Explore the capabilities of Amazon EMR Serverless by processing semi-structured review data with Apache Spark, showcasing efficient big data analysis without managing clusters.
Explore and replicate Amazon EMR (Elastic MapReduce) setup and utilization for big data processing and analytics tasks, featuring comprehensive demonstrations from VPC creation to Spark job execution.
To associate your repository with the bigdatainfrastructure topic, visit your repo's landing page and select "manage topics."