Engineering-Enterprise Data Platforms
Lead Data and AI Architect
Jaipur, Rajasthan, India
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.7M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Lead / management
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 26 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
- We are looking for a Lead Data Architect to design, build, and scale our data
- pipelines and entity resolution systems. This role combines deep technical
- expertise in data engineering with hands-on experience in AI-assisted tooling,
- entity matching, and data integration from diverse sources. You will lead
- architectural decisions for our data platform, mentor engineers, and ensure our
- pipelines are reliable, scalable, and production-grade.
Key Responsibilities
- Architect, build, and maintain robust, scalable data pipelines that
- ingest, transform, and serve data from multiple internal and external sources.
- Own the end-to-end orchestration of data workflows using tools like
- Dagster, ensuring observability, reliability, and maintainability of pipelines.
- Design and implement entity resolution workflows — including matching,
- merging, and survivorship logic — using tools such as Splink, to produce clean,
- deduplicated, golden records.
- Build and maintain web scrapers to source data from external providers,
- ensuring resilience to source changes, rate limits, and data quality issues.
- Integrate and reconcile data coming from multiple, often inconsistent,
- sources into unified, trustworthy datasets.
- Design and maintain data models and schemas across transactional and
- analytical systems, ensuring consistency, scalability, and performance.
- Leverage AI/LLM-based tools and techniques to enhance data pipeline
- capabilities — e.g., intelligent data extraction, automated data quality
- checks, or AI-assisted entity matching.
- Define and enforce best practices around pipeline design, testing,
- monitoring, and documentation.
- Collaborate closely with data engineers, product managers, and other
- stakeholders to translate business requirements into scalable data architecture.
Provide technical leadership and mentorship to the data engineering team.
Required Skills & Experience
- Strong hands-on experience building and maintaining production-grade
- data pipelines at scale.
- Practical experience with Dagster (or similar orchestration tools like
- Airflow/Prefect) for pipeline orchestration.
- Experience with Splink or similar probabilistic/deterministic record
- linkage tools for entity matching, merging, and survivorship.
- Strong proficiency in Python, including experience writing and
- maintaining web scrapers.
- Proven experience integrating and maintaining data pipelines that pull
- from multiple, heterogeneous data sources.
- Experience applying AI/ML tools within data engineering workflows (e.g.,
- LLM-assisted data cleaning, extraction, or matching).
- Hands-on experience with relational and distributed databases such as
- PostgreSQL and Google Cloud Spanner.
- Strong understanding of data modeling principles (normalization,
- dimensional modeling, schema design) across OLTP and OLAP systems.
- Experience with cloud data warehousing platforms such as BigQuery,
- Redshift, and cloud platforms (GCP/AWS/Azure).
- Strong communication skills and experience working cross-functionally
- with engineering and product teams.
- Experience with distributed data processing frameworks (e.g., Spark,
- Dask).
Familiarity with data governance, lineage, and cataloging tools.
- Prior experience in a lead or architect-level role guiding a data
- engineering team.
What We're Looking For
- A technically strong, hands-on leader who can balance architectural thinking
- with the practical grit of debugging a flaky scraper or tuning a matching
- algorithm — someone who's comfortable owning both the big picture and the messy
- details of real-world data.
Listed on hirly, a job board. hirly is not the employer: Engineering-Enterprise Data Platforms is hiring for this role.
Similar jobs
- Associate Principal AI Architect (Claude AI)Unisys · Bangalore, KA, IndiaFirst seen today
- Associate Principal AI Architect (Claude AI)Unisys · Bangalore, KA, IndiaFirst seen today
- S&C GN - Tech Strategy & Transformation - Gen AI/AI Architect – Senior ManagerAccentureFirst seen today
- #ACN S&C GN - Tech Strategy & Transformation - AI Architect - ManagerAccentureFirst seen today
- Gen AI Architect-Lead AI and Data Solutions Engineer IIDeloitte · Chennai, Tamil Nadu, IndiaFirst seen today
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job