Data Engineer, Translational Data Management, Automation & AI

Amgen
Amgen logo
Location
India - Hyderabad
Job Type
Full-time
Posted
September 9, 2026
Views
3

Job Description

Career Category

Clinical

Location: Amgen India office, Hyderabad

Employment type: Full-time

Department / Team: Computational Biology team, Precision Medicine

High-level role

We are seeking a hands-on, technically strong  Translational Data Management, Automation, & AI Engineer to design, build, and operate robust biomarker and clinical data ingestion pipelines that feed our biomarker platform. You will work closely with computational biologists, translational scientists, data scientists, lab operations, and external vendors/contract research organizations (CROs) to ensure timely, accurate, and standardized ingestion of assay and clinical data for analysis, visualization, and machine-learning use cases supporting clinical trials.

Key responsibilities

  • Design, implement, test, deploy, and maintain end-to-end data ingestion pipelines that prepare biomarker and clinical data for downstream analytics, visualization, and ML models.
  • Implement automated data validation, quality control checks, error handling, and remediation workflows to ensure data quality and traceability.
  • Integrate Codex workflows, agentic automation and generative AI to meet TAT and efficiency goals.
  • Collaborate with internal biomarker labs and CROs/vendors to onboard new assays; author and maintain data transfer specifications, interface control documents, and acceptance criteria.
  • Build and maintain harmonization and mapping logic (units, controlled terminology, ontologies) and data models needed to standardize biomarker and clinical datasets.
  • Generate study-specific analysis bundle per request in defined timeline.
  • Produce and maintain clear documentation: software specification forms, data definition tables, runbooks, and onboarding guides.
  • Write clean, tested, maintainable Python code and contribute to CI/CD pipelines, automated testing, and release processes.

Required qualifications

Education & experience

  • 8+ years of experience with Bachelor’s in Computational Biology, Bioinformatics, AI, Computer Science, Data Engineering, or related field. PhD is a plus.
  • 3+ years of experience in data engineering or platform engineering roles; experience working with biomarker/biological/clinical data or in a clinical research environment is highly desirable.

Technical skills

  • Experience working with clinical labs, biomarker assays (immunoassay, flow cytometry, immunohistochemistry, proteomics, whole genome sequencing, exome sequencing, RNA-seq, methylation, metabolomics)
  • Strong programming skills in Python and database design. Experience with Databricks
  • Experience with workflow/orchestration tools (e.g., Airflow, Nextflow, snakemake).
  • Experience with agentic automation and formulation of AI workflow development and deployment, agentic automation tools and Codex workflows.
  • Familiarity with HPC, cloud platforms and storage (e.g., AWS) and best practices for secure data handling.
  • Experience with version control (Git), CI/CD, containerization (Docker)
  • Knowledge of clinical data formats and standards (e.g., CDISC/SDTM/ADaM).
  • Familiarity with data standardization and harmonization frameworks, controlled vocabularies
  • Experience building, testing and debugging R pipelines for production data processing.

.

Researching Amgen before you apply?

See 86 open roles · Verified H-1B salary data · Clinical-trial hiring momentum · Culture, benefits & locations.

View Amgen profile

Frequently Asked Questions

Where is the job located, and is it remote/hybrid/on-site?
This full-time position is located on-site at the Amgen India office in Hyderabad.
What are the key responsibilities of this role?
You will design, build, and maintain end-to-end biomarker and clinical data ingestion pipelines. Additional duties include implementing automated data validation, integrating generative AI and agentic automation, collaborating with CROs and internal labs, building data standardization models, and writing clean Python code.
What education and experience are required?
You need a Bachelor's degree in Computational Biology, Bioinformatics, AI, Computer Science, Data Engineering, or a related field (PhD is a plus), with 8+ years of experience. This must include 3+ years of experience in data or platform engineering roles.
What technical skills and programming languages are required?
You need strong Python programming and database design skills, experience with Databricks, workflow orchestration tools (like Airflow or Nextflow), agentic automation, AWS, Git, CI/CD, and Docker. Experience with clinical labs, biomarker assays, CDISC standards, and debugging R pipelines is also required.
Which team will I be working with?
You will be working in the Computational Biology team within the Precision Medicine department.

Ready to Apply?

Apply for this Position

You'll be redirected to the company's application page

Share this job:

Explore Amgen

Research the company before you apply.

  • 86 open roles
  • Verified H-1B salary data
  • Clinical-trial hiring momentum
  • Culture, benefits & locations
View company profile

Job Information

Source: manual
AI Relevance: 78/100 (Relevant)
Remote Type: onsite
Allowed Locations: Worldwide
Skills & Tags:
amgen bioinformatics computational biology data engineering clinical proteomics

Get Similar Jobs by Email

Weekly digest of Amgen and similar companies. Free.

Related Jobs

Apply for this Position

Get weekly job alerts