Join a cutting‑edge team building high‑performance, real‑time data pipelines that power AI‑driven solutions for global enterprises. You’ll design, develop, and optimize streaming applications that transform raw event streams into actionable enterprise data.
What You’ll Do
- Design and maintain real‑time processing apps with Apache Flink.
- Develop transformation logic using Flink, Kafka Streams, KSQLDB, and SMTs.
- Build scalable frameworks to harmonize source data into enterprise models.
- Implement event‑time processing, watermarks, and stateful windowing for streaming workloads.
- Optimize pipelines for low latency, high throughput, and resource efficiency.
- Integrate streams with Kafka, Schema Registry, APIs, and downstream analytics.
- Create automated tests and CI/CD pipelines for streaming applications.
What You Need
- 6–10 years of data engineering, focusing on real‑time streaming.
- Expert in Apache Flink for enterprise‑scale stream processing.
- Strong knowledge of Apache Kafka ecosystem (Kafka Streams, KSQLDB, Connect, SMTs).
- Proficiency in Java or Scala; Python a plus.
- Experience with Avro, Protobuf, JSON and Schema Registry.
- Solid SQL skills on large structured and semi‑structured datasets.
- Familiarity with cloud platforms (Azure, AWS, or GCP) and container orchestration.
Good to Have
- Experience with Confluent Platform and enterprise Kafka deployments.
- Exposure to data lakehouse solutions such as Databricks, Snowflake, or Delta Lake.
- Knowledge of CDC patterns and API integration.
- Agile/Scrum delivery experience.
The Opportunity
Genpact’s AI Gigafactory accelerates advanced technology solutions, letting you work on AI‑driven projects that solve complex business challenges at scale.
