Lead Data Engineer - Data Engineering 4C
Ready to turn bold ideas into real-world impact?
At Genpact, we don’t just adapt to change, we lead it. AI and digital innovation are transforming the way businesses work, and we’re at the forefront of it. Genpact’s AI Gigafactory, our industry-first accelerator, exemplifies how we scale advanced technology solutions to help global enterprises work smarter, grow faster, and transform at scale. Whether tackling complex challenges through large-scale models or agentic AI, our breakthrough solutions tackle companies’ most complex challenges.
If you thrive in a fast-moving, innovation-driven environment, love building and deploying cutting-edge AI solutions, and want to push the boundaries of what’s possible, this is your moment.
Genpact (NYSE: G) is an agentic and advanced technology solutions company. We leverage process intelligence and artificial intelligence to deliver measurable outcomes. With a strong partner ecosystem and decades of client trust, we provide innovative solutions that transform how businesses run. Powered by a team with an active learning mindset and client centricity at its core, we deliver lasting value for the world’s leading enterprises.
Get to know us at genpact.com and on LinkedIn, YouTube, X, and Facebook.
Job Description
Senior Data Engineer
As a Senior Data Engineer within our Enterprise Data and Analytics team, you’ll be at the forefront of a company-wide digital transformation. We’re looking for a forward-thinking engineer who brings deep technical expertise and a passion for building scalable, AI-ready data platforms. In this role, you’ll architect and implement modern data solutions that power intelligent applications, streamline operations, and deliver measurable business value. You’ll work closely with cross-functional teams to design robust pipelines, integrate diverse data sources, and enable advanced analytics and generative AI capabilities.
Key Responsibilities:
* Design, develop, and maintain scalable data pipelines and ETL processes using Azure Data Factory, Azure Data Bricks, and other Azure services.
* Implement and optimize data storage solutions using Azure Data Lake, Azure SQL Database, and Delta Lake to support analytics and reporting needs.
* Collaborate with data scientists, analysts, and business stakeholders to understand data requirements and deliver high-quality, reliable datasets.
* Ensure data security, compliance, and governance by applying best practices, role-based access control (RBAC), and encryption.
* Monitor and troubleshoot data workflows, ensuring performance, reliability, and cost-efficiency across Azure cloud environments.
* Automate data integration and transformation tasks using Azure Functions, Logic Apps, and scripting languages like Python.
* Mentor junior engineers and foster a collaborative team environment.
Essential Qualifications:
* Bachelor’s degree in computer engineering, Computer Science, or a related discipline
* Experience in ETL design, development, and performance tuning using the Microsoft Stack in a multi-dimensional data warehousing environment.
* Advanced SQL programming expertise (PL/SQL, T-SQL)
* Experience in Enterprise Data & Analytics solution architecture
* Experience in Python Programming
* Hands-on experience with Azure, especially for data-heavy/analytics applications leveraging relational and NoSQL databases, Data Warehousing, and Big Data solutions.
* Experience with key Azure services: Azure Data Factory, Data Lake Gen2, Analysis Services, Databricks, Blob Storage, SQL Database, Cosmos DB, App Service, Logic Apps, and Functions
* Experience designing data models aligning with business requirements and analytics needs.
* Experience defining and implementing data security standards, including encryption, auditing, and monitoring.
* Strong analytical skills and a passion for intellectual curiosity.
Preferred Skills:
* Experience setting up and operating data pipelines using Python or SQL
* Familiarity with DevOps processes (CI/CD) and infrastructure as code
* Knowledge of Master Data Management (MDM) and Data Quality tools
* Experience developing REST APIs using Java Spring Boot, Python
* Experience building and supporting AL/ML software solutions
* Familiarity with stream-processing systems (e.g., Event Hubs, Storm, Spark-Streaming)
* Experience with API integrations (RESTful, SOAP) for both internal and external systems to enhance data flow and automation.
* Experience working in Agile environments, familiarity with Agile tools like Jira or Azure DevOps.
* Experience in data and analytics within the Life Sciences industry is a plus.
Required Experience/Skills:
Technical Skills:
- SQL programming
- ETL design
- Python
- Cloud: Azure
a) Data Pipeline, Storage & Devops
ETL (Build Data Pipeline): SQL, Python, Pyspark (ETL Platform: Azure Data Factory/ Databricks) & Streaming (e.g., Event Hubs, Storm, Spark-Streaming)
Storage (Azure Data Lake, Cosmos DB, Blob, Azure SQL Database)
CI/CD - GitHub and Azure DevOps; and Agile environment (Jira)
Qualifications
Bachelors - Business Analytics, Bachelors - Computer Science, Bachelors - Statistics, Masters - Data ScienceCertifications
Certified Data Management Professional - RiversandRiversand, Databricks Certified Associate Developer for Apache Spark 3.0 - Databricks AcademyDatabricks Academy, Databricks Certified Associate Developer for Apache Spark - Databricks AcademyDatabricks Academy, Microsoft Certified: Azure Data Engineer Associate - MicrosoftMicrosoft, Oracle Database 12c Certified Implementation Specialist - OracleOracleRequired Skills
Agile Methodology, Banking Capital Markets, Client Relations, Cloud Computing, Collaboration Tools, Commercial Sales, Consumer Packaged Goods (CPG), Databricks Platform, Data Engineering, Data Literacy, Design Thinking, Electronic Hardware, Executive Presence, Inclusion, Manufacturing Equipment, People Leadership, Personal Effectiveness, Risk Management, Semiconductors, Snowflake (Platform), Sourcing and Procurement, StorytellingLanguage
EnglishLanguage Proficiency -
Proficient - C2Additional Job Location -
Job Type
RegularMaster Skill List -
Data EngineeringRemote Type -
RemoteWork Shift -
Flex Time (India)Why join Genpact?
• Lead AI-powered transformation – Drive innovation and solve real-world business challenges that matter
• Make an impact – Help global enterprises solve business challenges that matter
• Accelerate your career – Gain hands-on experience, mentorship, and world-class learning opportunities to stay ahead
• Work with the best – Join 140,000+ bold thinkers and problem-solvers who push boundaries every day
• Thrive in a values-driven culture – Our courage, curiosity, and incisiveness - built on a foundation of integrity and inclusion - allow your ideas to fuel progress
Come join the 140,000+ coders, tech shapers, and growth makers at Genpact and take your career in the only direction that matters: Up.
Let’s build tomorrow together.
Genpact is an Equal Opportunity Employer and considers applicants for all positions without regard to race, color, religion or belief, sex, age, national origin, citizenship status, marital status, military/veteran status, genetic information, sexual orientation, gender identity, physical or mental disability or any other characteristic protected by applicable laws. Genpact is committed to creating a dynamic work environment that values respect and integrity, customer focus, and innovation.
Furthermore, please do note that Genpact does not charge fees to process job applications and applicants are not required to pay to participate in our hiring process in any other way. Examples of such scams include purchasing a 'starter kit,' paying to apply, or purchasing equipment or training.