hirly

Capgemini Engineering

Data Engineer

Gurgaon, IN

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Capgemini Engineering first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.7M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Role family
Data & ML
Seniority
Mid level
Country
IN
Work mode
On-site / unstated
First seen by hirly
6 Oct 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

Choosing Capgemini means choosing a company where you will be empowered to shape your career in the way you’d like, where you’ll be supported and inspired by a collaborative community of colleagues around the world, and where you’ll be able to reimagine what’s possible. Join us and help the world’s leading organizations unlock the value of technology and build a more sustainable, more inclusive world.

Your Role

As a Data Engineer, you will design, build, and maintain enterprise data platforms that support AI-driven products and analytics initiatives. You will work closely with Data Scientists, Architects, AI Engineers, and business stakeholders to develop robust data pipelines, digital twin platforms, knowledge graphs, and AI-ready data architectures.

  • Design and implement data source connectors for telemetry, topology, intent, ITSM, knowledge repositories, and enterprise applications.
  • Build ingestion, validation, replay, and data contract frameworks for large-scale data onboarding.
  • Develop ingestion and topology mapping capabilities to transform raw telemetry into incident-aware datasets.
  • Design, develop, and maintain Network Digital Twin and topology graph platforms.
  • Build and manage Knowledge Graphs, ontology frameworks, entity relationships, and semantic data models.
  • Develop reusable Feature Engineering Frameworks and Feature Store capabilities for AI/ML applications.
  • Enable data foundations for RAG, GraphRAG, Vector Databases, and AI Agent frameworks.
  • Develop and optimize data pipelines for model training, deployment, monitoring, and lifecycle management.
  • Implement metadata management, lineage tracking, observability, security, auditing, and governance frameworks.
  • Support integrations with AI reasoning engines, recommendation systems, and agent orchestration platforms.
  • Collaborate with cross-functional teams to deliver scalable, enterprise-grade AI and analytics solutions.

Your Profile

  • Strong programming expertise in Python and SQL .
  • Experience with JavaScript, React, or Angular .
  • Strong knowledge of Data Modeling, ETL/ELT, and Data Architecture principles.
  • Hands-on experience with Apache Spark , Spark Streaming , and Kafka .
  • Experience building and managing Data Lakes .
  • Strong experience with AWS services including: S3, Glue, Lambda, SageMaker
  • Knowledge of Machine Learning , Deep Learning , Predictive Analytics , and Statistical Modeling .
  • Experience with Generative AI technologies including: Large Language Models (LLMs), Prompt Engineering, Retrieval-Augmented Generation (RAG) , AI Agents ,Agentic AI Frameworks
  • Expertise in: Relational Databases, Graph Databases , Vector Databases
  • Experience building APIs, Data Pipelines, and Batch/Real-Time Data Processing frameworks.
  • Strong understanding of: Data Quality , Metadata Management, Data Lineage, Data Governance
  • AWS Glue and Azure Data Factory.
  • SageMaker Pipelines.
  • Experience with Digital Twin platforms and architectures.
  • Knowledge of GraphRAG and advanced RAG solutions.
  • MLOps and LLMOps implementation experience.
  • DataOps practices and automation frameworks.
  • Experience with Enterprise Knowledge Management Platforms.
  • Strong analytical, debugging, and problem-solving skills.

What You'll Love About Working Here

  • Opportunity to work on cutting-edge AI, GenAI, Knowledge Graph, and Digital Twin technologies.
  • Exposure to enterprise-scale AI and analytics transformation initiatives.
  • Collaborative environment with Data Scientists, AI Engineers, Architects, and domain experts.
  • Continuous learning opportunities across Data Engineering, GenAI, MLOps, and Cloud technologies.
  • Ability to build scalable AI data foundations that drive measurable business impact and innovation.
  • Innovation-driven culture focused on engineering excellence and technical growth.
  • Capgemini is an AI-powered global business and technology transformation partner, delivering tangible business value. We imagine the future of organizations and make it real with AI, technology and people. With our strong heritage of nearly 60 years, we are a responsible and diverse group of 420,000 team members in more than 50 countries. We deliver end-to-end services and solutions with our deep industry expertise and strong partner ecosystem, leveraging our capabilities across strategy, technology, design, engineering and business operations. The Group reported 2024 global revenues of €22.1 billion.
  • Make it real | www.capgemini.com
Original posting on Capgemini Engineering's site ↗

Listed on hirly, a job board. hirly is not the employer: Capgemini Engineering is hiring for this role.

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job