Scientist, Data Science

AstraZeneca
AstraZeneca logo
Location
Waltham, Massachusetts
Job Type
Full-time
Posted
August 14, 2026
Views
8
Salary Range
$91k - $137k USD

Job Description

We are seeking a highly motivated Scientist to join a newly formed, dynamic team within early oncology R&D. The successful candidate will leverage their data science expertise in mining large datasets to drive our efforts in target identification, mechanism of action (MOA) studies, and biomarker strategy development, with a particular focus on analyses related to the function and aging of the immune system.

At AstraZeneca, you'll have the opportunity to make a significant impact on the future of healthcare while working in a collaborative environment at the cutting edge of research. The ideal candidate will thrive in this setting, contributing to our growth trajectory as we build our evolving team.

Key Responsibilities

  • Execute and Maintain Pipelines: Process and analyze large-scale biobank datasets, human population data, and in-vitro biological data using established analysis pipelines.
  • Analytical Support: Apply analytical methods and machine learning algorithms to help identify potential therapeutic targets and biomarkers.
  • Cross-Functional Collaboration: Partner with wet-lab scientists to analyze experimental results for target identification and Mechanism of Action (MOA) studies.
  • Data Visualization: Generate high-quality visualizations and reports to communicate findings to the project team.
  • Strategic Contribution: Provide high-quality data and computational insights that contribute to the development of biomarker strategies.
  • Team Participation: Actively participate in team meetings, presenting data-driven insights to help the group meet project milestones.
  • Continuous Learning: Stay current with the latest developments in data science and bioinformatics tools.

Qualifications

  • Education: Ph.D. in Bioinformatics, Computational Biology, Data Science, Epidemiology, or a related field (0–2 years post-graduate experience); or MS with 2–4 years of experience; or BS with 4+ years of relevant experience.
  • Data Experience: Minimum 2 years of experience working with large-scale biological or population datasets, preferably including experience analyzing immune system aging/function within the context of human and/or mouse data.
  • Coding Proficiency: Strong proficiency in Python or R.
  • Technical Knowledge: Solid understanding of statistical analysis and foundational machine learning techniques.
  • Genomics Foundation: Hands-on experience with NGS data analysis (e.g., RNA-seq, DNA methylation, ChIP-seq, or ATAC-seq).
  • Multi-omics Interest: Experience with, or a strong desire to learn, proteomic data analysis and multi-omic data integration.
  • Operational Skills: Excellent problem-solving skills, attention to detail, and the ability to manage multiple tasks in a fast-paced environment.
  • Communication: Ability to clearly present data and technical workflows to a multidisciplinary team.

Desired Skills and Attributes

  • Prior experience or familiarity with biomarkers of immune system aging/function.
  • Prior experience or internship in the pharmaceutical or biotechnology industry.
  • Prior experience running large-scale association testing (e.g., genome-wide association studies [GWAS], epigenome-wide association studies [EWAS], proteome-wide association studies).
  • Familiarity with methods in statistical genetics (e.g., Mendelian randomization, fine mapping, colocalization).
  • Familiarity with machine learning analysis architectures (e.g., random forest, gradient boosting, transformers).
  • Familiarity with public biological databases (e.g., GTEx, TCGA), epidemiological cohort data (e.g., TOPMed cohorts), or biobanks (e.g., UK Biobank, FinnGen).
  • Ability to apply integrated generative protein design pipelines - from target-conditioned backbone generation through sequence design to computational fold validation - to support the development of novel therapeutic biologics with optimized specificity and developability properties.

We are seeking a highly motivated Scientist to join a newly formed, dynamic team within early oncology R&D. The successful candidate will leverage their data science expertise in mining large datasets to drive our efforts in target identification, mechanism of action (MOA) studies, and biomarker strategy development, with a particular focus on analyses related to the function and aging of the immune system.

At AstraZeneca, you'll have the opportunity to make a significant impact on the future of healthcare while working in a collaborative environment at the cutting edge of research. The ideal candidate will thrive in this setting, contributing to our growth trajectory as we build our evolving team.

Key Responsibilities

  • Execute and Maintain Pipelines: Process and analyze large-scale biobank datasets, human population data, and in-vitro biological data using established analysis pipelines.
  • Analytical Support: Apply analytical methods and machine learning algorithms to help identify potential therapeutic targets and biomarkers.
  • Cross-Functional Collaboration: Partner with wet-lab scientists to analyze experimental results for target identification and Mechanism of Action (MOA) studies.
  • Data Visualization: Generate high-quality visualizations and reports to communicate findings to the project team.
  • Strategic Contribution: Provide high-quality data and computational insights that contribute to the development of biomarker strategies.
  • Team Participation: Actively participate in team meetings, presenting data-driven insights to help the group meet project milestones.
  • Continuous Learning: Stay current with the latest developments in data science and bioinformatics tools.

Qualifications

  • Education: Ph.D. in Bioinformatics, Computational Biology, Data Science, Epidemiology, or a related field (0–2 years post-graduate experience); or MS with 2–4 years of experience; or BS with 4+ years of relevant experience.
  • Data Experience: Minimum 2 years of experience working with large-scale biological or population datasets, preferably including experience analyzing immune system aging/function within the context of human and/or mouse data.
  • Coding Proficiency: Strong proficiency in Python or R.
  • Technical Knowledge: Solid understanding of statistical analysis and foundational machine learning techniques.
  • Genomics Foundation: Hands-on experience with NGS data analysis (e.g., RNA-seq, DNA methylation, ChIP-seq, or ATAC-seq).
  • Multi-omics Interest: Experience with, or a strong desire to learn, proteomic data analysis and multi-omic data integration.
  • Operational Skills: Excellent problem-solving skills, attention to detail, and the ability to manage multiple tasks in a fast-paced environment.
  • Communication: Ability to clearly present data and technical workflows to a multidisciplinary team.

Desired Skills and Attributes

  • Prior experience or familiarity with biomarkers of immune system aging/function.
  • Prior experience or internship in the pharmaceutical or biotechnology industry.
  • Prior experience running large-scale association testing (e.g., genome-wide association studies [GWAS], epigenome-wide association studies [EWAS], proteome-wide association studies).
  • Familiarity with methods in statistical genetics (e.g., Mendelian randomization, fine mapping, colocalization).
  • Familiarity with machine learning analysis architectures (e.g., random forest, gradient boosting, transformers).
  • Familiarity with public biological databases (e.g., GTEx, TCGA), epidemiological cohort data (e.g., TOPMed cohorts), or biobanks (e.g., UK Biobank, FinnGen).
  • Ability to apply integrated generative protein design pipelines - from target-conditioned backbone generation through sequence design to computational fold validation - to support the development of novel therapeutic biologics with optimized specificity and developability properties.
  • Working knowledge of computational histology pipelines incorporating modern deep learning approaches - including self-supervised and weakly supervised learning (MIL, DINO) and histopathology foundation models (e.g. UNI, CONCH) - to enable scalable, label-efficient classification of complex tissue phenotypes.
  • Familiarity or prior experience with agentic AI in the context of analysis code pipeline development and biological analysis.
  • Evidence of scientific contribution through publications, posters, or GitHub repositories.

As AstraZeneca continues to put patients at the forefront of our mission, we are excited for our move to Kendall Square/Cambridge in 2026. Find out more information here: Kendall Square Press Release

Ready to join us on this mission? Apply now!

If you’re curious to know more, please contact Bobbi Poole, our Talent Acquisition Partner.

Competitive remuneration and benefits apply

We offer a competitive Total Reward program including a market driven base salary, bonus and long-term incentive. We have a generous paid time off program and a comprehensive benefits package.

The annual base pay for this position ranges from $91,008.80 - $136,513.20. Our positions offer eligibility for various incentives—an opportunity to receive short-term incentive bonuses, equity-based awards for salaried roles and commissions for sales roles. Benefits offered include qualified retirement programs, paid time off (i.e., vacation, holiday, and leaves), as well as health, dental, and vision coverage in accordance with the terms of the applicable plans.

Date Posted

06-Aug-2026

Closing Date

29-Aug-2026Our mission is to build an inclusive environment where equal employment opportunities are available to all applicants and employees. In furtherance of that mission, we welcome and consider applications from all qualified candidates, regardless of their protected characteristics. If you have a disability or special need that requires accommodation, please complete the corresponding section in the application form.

Researching AstraZeneca before you apply?

See 32 open roles · Verified H-1B salary data · Clinical-trial hiring momentum · Culture, benefits & locations.

View AstraZeneca profile

Frequently Asked Questions

Where is the job located, and is it remote/hybrid/on-site?
The job is located in Waltham, Massachusetts. The posting does not specify a remote or hybrid work-mode policy, but notes that AstraZeneca is moving to Kendall Square/Cambridge in 2026.
What are the required qualifications and experience levels for this role?
You need a Ph.D. in Bioinformatics, Computational Biology, Data Science, Epidemiology, or a related field (0-2 years post-graduate experience); an MS with 2-4 years of experience; or a BS with 4+ years of experience. You also need 2+ years of experience with large-scale biological or population datasets, coding proficiency in Python or R, and NGS data analysis experience.
What are the key responsibilities of the Scientist, Data Science?
Key responsibilities include executing and maintaining pipelines for large-scale biobank and biological data, applying analytical methods and machine learning for target and biomarker identification, collaborating with wet-lab scientists, generating data visualizations, contributing to biomarker strategies, and presenting insights in team meetings.
What is the salary range for this position?
The annual base pay for this position ranges from $91,008.80 to $136,513.20.
What benefits and compensation incentives are offered?
AstraZeneca offers a competitive Total Reward program including short-term incentive bonuses, equity-based awards, qualified retirement programs, a generous paid time off program (vacation, holidays, and leaves), and comprehensive health, dental, and vision coverage.
What is the application deadline and who can I contact for more information?
The application closing date is August 29, 2026. If you are curious to know more, you can contact Bobbi Poole, the Talent Acquisition Partner.

Ready to Apply?

Apply for this Position

You'll be redirected to the company's application page

Share this job:

Explore AstraZeneca

Research the company before you apply.

  • 32 open roles
  • Verified H-1B salary data
  • Clinical-trial hiring momentum
  • Culture, benefits & locations
View company profile

Job Information

Source: manual
AI Relevance: 88/100 (Highly relevant)
Remote Type: onsite
Allowed Locations: Worldwide
Skills & Tags:
astrazeneca machine learning deep learning bioinformatics computational biology genomics statistical genetics GWAS data science protein

Get Similar Jobs by Email

Weekly digest of AstraZeneca and similar companies. Free.

Related Jobs

Apply for this Position

Get weekly job alerts