Amgen
India - Hyderabad
Data & MLOps
Posted 2 months ago
Verified open on Oct 4, 2026 Ā· posted 45 days ago
Location:Ā Amgen India office, HyderabadĀ
Employment type:Ā Full-timeĀ
Department / Team:Ā Computational Biology team, Precision MedicineĀ
High-level roleĀ
We areĀ seekingĀ a hands-on, technically strongĀ Ā Translational Data Management, Automation, & AI EngineerĀ to design, build, and operate robust biomarker and clinical data ingestion pipelines that feed our biomarker platform. You will work closely with computational biologists, translational scientists, data scientists, lab operations, and external vendors/contract research organizations (CROs) to ensure timely, accurate, and standardized ingestion of assay and clinical data for analysis, visualization, and machine-learning use cases supporting clinical trials.Ā
Key responsibilitiesĀ
Design, implement, test, deploy, andĀ maintainĀ end-to-end data ingestion pipelines that prepare biomarker and clinical data for downstream analytics, visualization, and ML models.Ā
Implement automated data validation, quality control checks, error handling, and remediation workflows to ensure data quality and traceability.Ā
Integrate Codex workflows, agenticĀ automationĀ and generative AI to meet TAT and efficiency goals.Ā
Collaborate with internal biomarker labs and CROs/vendors to onboard new assays; author andĀ maintainĀ data transfer specifications, interface control documents, and acceptance criteria.Ā
Build andĀ maintainĀ harmonization and mapping logic (units, controlled terminology, ontologies) and data models needed to standardize biomarker and clinical datasets.Ā
GenerateĀ study-specific analysis bundle per request inĀ definedĀ timeline.Ā
Produce andĀ maintainĀ clear documentation: software specification forms, data definition tables, runbooks, and onboarding guides.Ā
Write clean, tested, maintainable Python code and contribute to CI/CD pipelines, automated testing, and release processes.Ā
Required qualificationsĀ
Education & experienceĀ
8+ years of experience with Bachelorās in Computational Biology, Bioinformatics,Ā AI,Ā Computer Science, Data Engineering, or relatedĀ field. PhD is a plus.Ā
3+ years of experience in data engineering or platform engineering roles; experience working with biomarker/biological/clinical data or in a clinical research environment is highly desirable.Ā
Technical skillsĀ
Experience working with clinical labs, biomarker assays (immunoassay, flow cytometry, immunohistochemistry, proteomics, whole genome sequencing, exome sequencing, RNA-seq, methylation, metabolomics)Ā
Strong programming skills in Python and database design. Experience with DatabricksĀ
Experience with workflow/orchestration tools (e.g., Airflow,Ā Nextflow,Ā snakemake).Ā
Experience with agentic automation and formulation of AI workflowĀ development and deployment, agentic automationĀ toolsĀ and Codex workflows.Ā
Familiarity with HPC, cloud platforms and storage (e.g., AWS) and best practices for secure data handling.Ā
Experience with version control (Git), CI/CD, containerization (Docker)Ā
Knowledge of clinical data formats and standards (e.g., CDISC/SDTM/ADaM).Ā
Familiarity with data standardization and harmonization frameworks, controlled vocabulariesĀ
Experience building,Ā testingĀ and debugging R pipelines for production data processing.Ā
Not ready to apply?
Get new Data & MLOps jobs in your inbox
Join 100+ AI professionals Ā· Weekly, free, unsubscribe anytime
This role is classified as Data & MLOps.
$200k - $308k is the middle 50% of disclosed salaries, measured from 215 live Data & MLOps postings on this board. Roughly two thirds of postings disclose nothing, so this describes the ones that do, not the whole market.
Hiring for a role like this?
Reach AI professionals browsing the board - your listing goes live instantly.
The Coca-Cola CompanyMexico City, Mexico+1 more