What you’ll work on
- Design and implement real-time and batch data pipelines.
- Build and maintain scalable streaming systems.
- Develop and optimize stream processing jobs with Apache Flink.
- Ensure reliable ingestion from multiple internal and external data sources.
- Design event schemas and data contracts.
- Implement data validation, transformation, and enrichment logic.
- Optimize storage layouts and lifecycle management strategies.
- Improve system observability (metrics, logging, alerting) and troubleshoot performance bottlenecks in distributed systems.
- Implement retry, dead-letter, and replay mechanisms.
- Ensure data quality, consistency, and governance.
- Collaborate with Backend, DevOps, and Security teams.
