Senior Data Engineer
Join the Core Data Platform team to design, build, and maintain scalable data pipelines and cloud infrastructure. Work with petabytes of data, distributed systems, and modern data stack to power insights and analytics for a global omnichannel advertising platform.
Responsibilities
- Design, build, and maintain ETL/ELT pipelines using Apache Spark for large-scale data ingestion and transformation.
- Architect and manage cloud data infrastructure on GCP or AWS, leveraging services like BigQuery, S3, GCS, EMR, and Airflow.
- Improve and manage data warehouse and data lake solutions to ensure data quality, consistency, and accessibility.
- Collaborate with cross-functional teams to understand data needs and implement solutions for new features and initiatives.
- Implement monitoring, alerting, and logging to maintain data pipeline health and accuracy.
Requirements
- 5+ years of data engineering experience building and operating production data pipelines at scale (TB+ datasets, hourly/daily batch or streaming workloads).
- Hands-on production experience with Apache Spark and distributed data processing frameworks (Flink, Hive, Trino).
- Production experience with GCP or AWS data solutions (BigQuery, Dataproc, GCS, S3, EMR, Redshift).
- Production experience with Kafka or compatible streaming platforms.
- Strong understanding of data warehouse and data lake concepts, including Medallion Architecture.
Nice to have
- Production experience with a lakehouse table format (Apache Iceberg or Delta Lake).
Benefits
- Hybrid working model (3 days per week in office)
- Nearby parking
- Mentorship program
- Internal learning tools
- Pet friendly office
- Happy hours
- Fully stocked kitchen