Amgen
Data Engineer, Translational Data Management, Automation & AI
India - Hyderabad
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.4M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Data & ML
- Seniority
- Mid level
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 27 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Career Category
Clinical
Job Description
- Location: Amgen India office, Hyderabad
- Employment type: Full-time
- Department / Team: Computational Biology team, Precision Medicine
High-level role
We are seeking a hands-on, technically strong Translational Data Management, Automation, & AI Engineer to design, build, and operate robust biomarker and clinical data ingestion pipelines that feed our biomarker platform. You will work closely with computational biologists, translational scientists, data scientists, lab operations, and external vendors/contract research organizations (CROs) to ensure timely, accurate, and standardized ingestion of assay and clinical data for analysis, visualization, and machine-learning use cases supporting clinical trials.
Key responsibilities
Design, implement, test, deploy, and maintain end-to-end data ingestion pipelines that prepare biomarker and clinical data for downstream analytics, visualization, and ML models.
Implement automated data validation, quality control checks, error handling, and remediation workflows to ensure data quality and traceability.
Integrate Codex workflows, agentic automation and generative AI to meet TAT and efficiency goals.
Collaborate with internal biomarker labs and CROs/vendors to onboard new assays; author and maintain data transfer specifications, interface control documents, and acceptance criteria.
Build and maintain harmonization and mapping logic (units, controlled terminology, ontologies) and data models needed to standardize biomarker and clinical datasets.
Generate study-specific analysis bundle per request in defined timeline.
Produce and maintain clear documentation: software specification forms, data definition tables, runbooks, and onboarding guides.
Write clean, tested, maintainable Python code and contribute to CI/CD pipelines, automated testing, and release processes.
Required qualifications
Education & experience
8+ years of experience with Bachelor’s in Computational Biology, Bioinformatics, AI, Computer Science, Data Engineering, or related field. PhD is a plus.
3+ years of experience in data engineering or platform engineering roles; experience working with biomarker/biological/clinical data or in a clinical research environment is highly desirable.
Technical skills
Experience working with clinical labs, biomarker assays (immunoassay, flow cytometry, immunohistochemistry, proteomics, whole genome sequencing, exome sequencing, RNA-seq, methylation, metabolomics)
Strong programming skills in Python and database design. Experience with Databricks
Experience with workflow/orchestration tools (e.g., Airflow, Nextflow, snakemake).
Experience with agentic automation and formulation of AI workflow development and deployment, agentic automation tools and Codex workflows.
Familiarity with HPC, cloud platforms and storage (e.g., AWS) and best practices for secure data handling.
Experience with version control (Git), CI/CD, containerization (Docker)
Knowledge of clinical data formats and standards (e.g., CDISC/SDTM/ADaM).
Familiarity with data standardization and harmonization frameworks, controlled vocabularies
Experience building, testing and debugging R pipelines for production data processing.
.
Similar jobs
- Data Engineer - MarketingZoom · Bangalore (IND)First seen today
- Data Engineer, People AnalyticsZoom · Bangalore (IND)First seen today
- Data Engineer IIDAT · Bangalore, Karnataka, IndiaFirst seen today
- Data EngineerNK Securities Research · IndiaFirst seen today
- Data EngineerAccenture · IndiaFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job