Capgemini Invent
Lead Data Scientist
Noida, IN
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Data & ML
- Seniority
- Lead / management
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 27 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
At Capgemini Invent, we believe difference drives change. As inventive transformation consultants, we blend our strategic, creative and scientific capabilities, collaborating closely with clients to deliver cutting-edge solutions. Join us to drive transformation tailored to our client's challenges of today and tomorrow. Informed and validated by science and data. Superpowered by creativity and design. All underpinned by technology created with purpose.
Your Role
- Programming Languages – Python – NumPy, SciPy, Pandas, MatPlotLib, Seaborne Databases – RDBMS (MySQL, Oracle etc.), NoSQL Stores (HBase, Cassandra etc.) ML/DL Frameworks – SciKitLearn, TensorFlow (Keras), PyTorch, Big data ML Frameworks - Spark (Spark-ML, Graph-X), H2O. Cloud – Azure/AWS/GCP.
- Predictive and Prescriptive modelling using Statistical and Machine Learning algorithms including but not limited to Time Series, Regression, Trees, Ensembles, Neural-Nets (Deep & Shallow – CNN, LSTM, Transformers etc.). Experience with open-source OCR engines like Tesseract, Speech recognition, Computer Vision, face recognition, emotion detection etc. is a plus.
- Unsupervised learning – Market Basket Analysis, Collaborative Filtering, Dimensionality Reduction, good understanding of common matrix decomposition approaches like SVD. Various Clustering approaches – Hierarchical, Centroid-based, Density-based, Distribution-based, Graph-based clustering like Spectral.
- NLP – Information Extraction, Similarity Matching, Sentiment Analysis, Text Clustering, Semantic Analysis, Document Summarization, Context Mapping/Understanding, Intent Classification, Word Embeddings, Vector Space Models, experience with libraries like NLTK, Spacy, Stanford Core-NLP is a plus. Usage of Transformers for NLP and experience with LLMs like (ChatGPT, Llama) and usage of RAGs (vector stores like LangChain & LangGraps), building Agentic AI applications.
Your Profile
Graph Analytics – Familiarity with Graph Algorithms (Directed & Undirected) – Traversal (BFS, DFS), Cycle Detection (Bellman Ford, Flyod Warshall), Shortest Path (Dijkstra, A*) etc. Building Knowledge Graphs with unstructured data and knowledge graph optimizations like PageRank/TrustRank is expected
Mathematical Optimization – Familiarity with common optimization algorithms, both discrete– Linear, Mixed-Integer, Goal, Dynamic etc and continuous – GD and its variants, Newton’s method etc. is expected. Experience with Simulated Annealing and exposure to ML inspired evolutionary optimization algorithms like Genetic Algorithm & Genetic Programming for optimization is a plus.
Simulations – Monte Carlo Simulation, Discrete-Event Simulation, Agent-Based Simulation, Hybrid Simulation, System Dynamics, Genetic Algorithm based Simulation.
Model Deployment – ML pipeline formation, data security and scrutiny check and ML-Ops for productionizing a built model on-premises and on cloud.
What you will love about working here
- We recognize the significance of flexible work arrangements to provide support. Be it remote work, or flexible work hours, you will get an environment to maintain healthy work life balance.
- At the heart of our mission is your career growth. Our array of career growth programs and diverse professions are crafted to support you in exploring a world of opportunities.
- Equip yourself with valuable certifications in the latest technologies such as Generative AI.
Capgemini is a global business and technology transformation partner, helping organizations to accelerate their dual transition to a digital and sustainable world, while creating tangible impact for enterprises and society. It is a responsible and diverse group of 340,000 team members in more than 50 countries. With its strong over 55-year heritage, Capgemini is trusted by its clients to unlock the value of technology to address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fueled by its market leading capabilities in AI, generative AI, cloud and data, combined with its deep industry expertise and partner ecosystem.
Similar jobs
- Lead Data Scientist Quantium · HyderabadFirst seen today
- (Ind) Staff, Data ScientistWM Global Technology Services India · IndiaFirst seen yesterday
- Lead Data ScientistMastercard · Pune, IndiaFirst seen yesterday
- (Ind) Staff, Data ScientistWM Global Technology Services India · IndiaFirst seen 2d ago
- Data Scientist - Deputy ManagerAdani · Ahmedabad, Gujarat, IndiaFirst seen 2d ago
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job