Data Architect - Data Engineering 4D

Genpact

Chennai, INremote country$12k/yrPosted Jul 28, 2026
Posting intelligenceActively listedReposted 2×, possible evergreen/ghost posting

Skills

kubernetesdatabricksprometheussnowflakegrafanadockeroraclepythonazuresparkkafkaflinkcicdjavagooglecloudawssap

About the role

Data Architect

Ready to turn bold ideas into real-world impact?

At Genpact, we don’t just adapt to change, we lead it. AI and digital innovation are transforming the way businesses work, and we’re at the forefront of it. Genpact’s AI Gigafactory, our industry-first accelerator, exemplifies how we scale advanced technology solutions to help global enterprises work smarter, grow faster, and transform at scale. Whether tackling complex challenges through large-scale models or agentic AI, our breakthrough solutions tackle companies’ most complex challenges.

If you thrive in a fast-moving, innovation-driven environment, love building and deploying cutting-edge AI solutions, and want to push the boundaries of what’s possible, this is your moment.

Genpact (NYSE: G) is an agentic and advanced technology solutions company. We leverage process intelligence and artificial intelligence to deliver measurable outcomes. With a strong partner ecosystem and decades of client trust, we provide innovative solutions that transform how businesses run. Powered by a team with an active learning mindset and client centricity at its core, we deliver lasting value for the world’s leading enterprises.

Get to know us at genpact.com and on LinkedIn, YouTube, X, and Facebook.

Job Description

Key Responsibilities

Design, develop, and manage enterprise-grade real-time data pipelines using Apache Kafka and the Confluent Platform.

Build, configure, and maintain Change Data Capture (CDC) pipelines using Debezium, SAP CDC, and other enterprise CDC connectors.

Configure and manage Kafka topics, partitions, replication, retention policies, and security for multiple source systems.

Manage and maintain Confluent Schema Registry, including schema evolution, compatibility, and governance.

Deploy, configure, and optimize Enterprise Bus consumers and producers to support data ingestion from 29 source systems.

Design scalable event-driven integration patterns to enable reliable, low-latency data movement across enterprise applications.

Monitor Kafka clusters and optimize performance, throughput, latency, and resource utilization.

Implement fault-tolerant messaging, retry mechanisms, dead-letter queues (DLQs), and disaster recovery strategies.

Collaborate with Data Architects, Integration Architects, application teams, and business stakeholders to define messaging patterns and data contracts.

Develop automation for deployment, monitoring, and operational support using CI/CD pipelines and Infrastructure as Code where applicable.

Troubleshoot production issues related to Kafka clusters, CDC pipelines, message delivery, and consumer performance.

Required Skills & Experience

Experience in Data Engineering with strong expertise in real-time streaming and messaging platforms.

Hands-on experience with Apache Kafka and the Confluent Platform in enterprise environments.

Strong experience implementing CDC pipelines using Debezium, SAP CDC, or similar technologies.

Expertise in Kafka topics, partitions, replication, consumer groups, producers, Kafka Connect, and Schema Registry.

Experience configuring and managing enterprise messaging infrastructure supporting multiple source applications.

Strong understanding of event-driven architecture, asynchronous messaging, and distributed systems.

Experience with Kafka security, including SSL/TLS, SASL, ACLs, and authentication/authorization.

Proficiency in Java and/or Python for developing Kafka producers, consumers, and integration services.

Experience with REST APIs, JSON, Avro, Protobuf, and event serialization frameworks.

Knowledge of Linux, Docker, Kubernetes, CI/CD pipelines, and monitoring tools such as Prometheus and Grafana.

Experience working with cloud platforms such as Azure, AWS, or GCP is preferred.

Preferred Qualifications

Experience with Confluent Control Center and enterprise Kafka administration.

Exposure to stream processing technologies such as Apache Flink, Kafka Streams, or KSQLDB.

Experience integrating ERP, CRM, and enterprise applications using CDC and messaging patterns.

Familiarity with data lakehouse platforms such as Databricks, Snowflake, or Delta Lake.

Excellent analytical, troubleshooting, communication, and stakeholder management skills.

Qualifications

Bachelors - Business Analytics, Bachelors - Computer Science, Bachelors - Statistics, Masters - Data Science

Certifications

Certified Data Management Professional - RiversandRiversand, Databricks Certified Associate Developer for Apache Spark 3.0 - Databricks AcademyDatabricks Academy, Databricks Certified Associate Developer for Apache Spark - Databricks AcademyDatabricks Academy, Microsoft Certified: Azure Data Engineer Associate - MicrosoftMicrosoft, Oracle Database 12c Certified Implementation Specialist - OracleOracle

Required Skills

Confluent Platform

Language

English

Language Proficiency -

Proficient - C2

Additional Job Location -

Job Type

Regular

Master Skill List -

Data Engineering

Remote Type -

Remote

Work Shift -

Flex Time (India)

Why join Genpact?

Lead AI-powered transformation – Drive innovation and solve real-world business challenges that matter

Make an impact – Help global enterprises solve business challenges that matter

Accelerate your career – Gain hands-on experience, mentorship, and world-class learning opportunities to stay ahead

Work with the best – Join 140,000+ bold thinkers and problem-solvers who push boundaries every day

Thrive in a values-driven culture – Our courage, curiosity, and incisiveness - built on a foundation of integrity and inclusion - allow your ideas to fuel progress

Come join the 140,000+ coders, tech shapers, and growth makers at Genpact and take your career in the only direction that matters: Up.

Let’s build tomorrow together.

Furthermore, please do note that Genpact does not charge fees to process job applications and applicants are not required to pay to participate in our hiring process in any other way. Examples of such scams include purchasing a 'starter kit,' paying to apply, or purchasing equipment or training.

Compensation

This Data Architect role pays $12k/yr. Within typical range for data architect roles in India.

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation for Data Architect roles in India varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our Data Architect hub for India medians across recent openings.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.