Senior Data Engineer, Core Data Pipeline
Build and scale real-time data processing systems for high-volume telemetry, focusing on stream processing, distributed systems, and low-latency performance.
Responsibilities
- Design, build, and maintain a high-throughput data processing pipeline for logs and traces.
- Architect real-time stream processing systems for data transformation and enrichment.
- Improve throughput, memory efficiency, and latency across distributed clusters.
- Own features from design and architecture through production monitoring and resilience.
Requirements
- At least 4 years of development experience with Scala or another JVM language.
- Hands-on experience designing scalable, distributed systems.
- Experience with data-streaming technologies such as Apache Kafka, Spark Streaming, Kafka Streams, or Apache Flink.
- Proficiency in data modeling and designing systems for large distributed datasets.
- Experience with containerization and orchestration tools, including Docker and Kubernetes.
- Knowledge of distributed computing principles, including consistency, partitioning, and resilience.
- Bachelor’s degree in computer science or an equivalent field.
Nice to have
- Production experience in a SaaS environment, including metrics, logging, and troubleshooting.
- Experience developing APIs with gRPC.