Senior Data Engineer (GCP, PySpark, Dataproc)
- Рівень:
- senior
- Джерело:
- djinni.co
Що робити
- Develop, maintain, and optimize ETL pipelines using Apache Spark and PySpark.
- Configure, run, and troubleshoot Spark workloads on GCP Dataproc.
- Package and submit Spark jobs, manage dependencies, and analyze driver and executor logs.
- Identify and resolve performance issues related to memory usage, shuffling, partitioning, data skew, and distributed data processing.
- Integrate Spark workloads with S3-compatible object storage using the Hadoop S3A connector.
Що пропонуємо
- A long-term international project
- Opportunity to work on a national-scale digital platform used by thousands of users
- Remote full-time collaboration
- Professional and supportive team environment
- Challenging technical tasks and a
Схожі вакансії
З блогу Trackr
Усі статті →Знайдено через trackr.help/jobs · Канал: @trackrhelp · Бот для персональних сповіщень: @trackrhelpBot


