Coditude
Data Engineer — Source Integrations & Pipelines
Pune, India
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.4M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Data & ML
- Seniority
- Mid level
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 23 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Phase: Initial phase (foundational)
Required experience: 7+ years in data engineering, building and operating production data pipelines.
Role summary
Builds the end-to-end data pipelines that bring every source into the platform reliably. Responsible for source integrations, ingestion, transformation, and data processing that feed the lakehouse and, later, the graph and ML layers. This is the core delivery role for the initial pilot use cases.
Key responsibilities
- Build and operate end-to-end data pipelines from source systems into the lakehouse / data fabric.
- Integrate structured, semi-structured, document, and time-series sources into a common platform.
- Implement data processing, cleansing, and transformation for the priority pilot use cases.
- Ensure pipeline reliability, monitoring, and data quality across all feeds.
- Work with the architect to align pipelines to platform standards and modeling conventions.
- Prepare curated datasets for the visualization and MLOps workstreams.
Must-have skills and experience
- Strong data engineering background building production pipelines end to end.
- Hands-on experience with Azure data services and readiness to work with Snowflake.
- Experience integrating diverse sources: Oracle ERP, MongoDB, PostgreSQL, Cassandra, Redis, InfluxDB, and time-series data.
- Solid data processing skills (batch and streaming) and strong SQL.
- Experience with object and file storage such as MinIO / NFS.
- Data quality, testing, and pipeline observability practices.
Nice to have
- Snowflake production experience.
- Exposure to OT / IoT data feeds.
- Familiarity with orchestration frameworks and infrastructure-as-code.
Relevant stack
Azure Data Factory / Synapse, Snowflake, Oracle ERP, MongoDB, PostgreSQL, Cassandra, Redis, InfluxDB, time-series sources, MinIO / NFS, APIs.
General attributes
- Proactive and self-driven, able to take ownership and move work forward without waiting to be told.
- AI-enabled in day-to-day work, comfortable using AI tools and copilots to accelerate delivery and quality.
- Strong self-learner who stays current with evolving tools, platforms, and practices.
- Good team player who collaborates well across engineering, operations, and stakeholder groups.
Similar jobs
- Cloud Data EngineerJefferies Financial Group · Pune, Maharashtra, IndiaFirst seen today
- Data Engineer - MarketingZoom · Bangalore (IND)First seen today
- Data Engineer, People AnalyticsZoom · Bangalore (IND)First seen today
- Data Engineer IIDAT · Bangalore, Karnataka, IndiaFirst seen today
- Data EngineerNK Securities Research · IndiaFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job