Senior Data Scientist, Biologics Discovery - Madrid, ES

Johnson & Johnson
Johnson & Johnson logo
Location
Madrid, Spain
Job Type
Full-time
Posted
September 9, 2026
Views
6
Salary Range
$55k - $88k USD

Job Description

At Johnson & Johnson, we believe health is everything. Our strength in healthcare innovation empowers us to build a world where complex diseases are prevented, treated, and cured, where treatments are smarter and less invasive, and solutions are personal. Through our expertise in Innovative Medicine and MedTech, we are uniquely positioned to innovate across the full spectrum of healthcare solutions today to deliver the breakthroughs of tomorrow, and profoundly impact health for humanity. Learn more at jnj.com.

As guided by Our Credo, Johnson & Johnson is responsible to our employees who work with us throughout the world. We provide an inclusive work environment where each person is considered as an individual. At Johnson & Johnson, we respect the diversity and dignity of our employees and recognize their merit.

Job Function

Data Analytics & Computational Sciences

Job Sub Function

Data Science

Job Category

Scientific/Technology

All Job Posting Locations

Madrid, Spain

Our expertise in Innovative Medicine is informed and inspired by patients, whose insights fuel our science-based advancements. Visionaries like you work on teams that save lives by developing the medicines of tomorrow.

Join us in developing treatments, finding cures, and pioneering the path from lab to life while championing patients every step of the way.

Learn more at https://www.jnj.com/innovative-medicine

About the opportunity

Johnson & Johnson Innovative Medicine is seeking a Senior Data Scientist dedicated to our Biologics Discovery organization. This role sits within our Data, Data Science & Artificial Intelligence team (DDSAI) and partners closely with our In Silico Discovery (ISD) organization - the group that builds the molecular design and property-prediction models (for example, developability, affinity and binding, and other molecular-property and liability-risk models) that guide which biologic molecules to design, make, and advance. ISD owns core molecular model development; you will build the data-facing ML capabilities (featurization, model-ready datasets, evaluation frameworks, and applied models on assay and sequence data) that make ISD's models faster to build and better to trust.

This position will be based at one of our office locations in either Spring House, PA (strongly preferred), Titusville, NJ, or Raritan, NJ, USA; or Madrid, Spain. (No remote option.)

Please note that this role is available across multiple countries and may be posted under different requisition numbers to comply with local requirements. While you are welcome to apply to any or all of the postings, we recommend focusing on the specific country(s) that align with your preferred location(s):

USA - Requisition Number: R-095854

Spain - Requisition Number: R-096793

Why this role matters: Biologics Discovery is generating rich, fast-growing data across assays, sequences, and modalities, and the opportunity now is to make that data fully model-ready and seamlessly available for ML. This role ensures biologics data is structured for training, and that applied ML on discovery data helps scientists prioritize molecules, flag risks, and generate hypotheses earlier - strengthening the interface to ISD's models rather than duplicating them.

You will design robust featurization and dataset curation, build and evaluate applied models on biologics assay, biophysical, and sequence/construct data, and define evaluation frameworks that keep models trustworthy. You operate at the interface between our data-generating and data-infrastructure partners and In Silico Discovery (ISD), ensuring the datasets and features you create strengthen ISD's molecular property models. This is an opportunity to shape how AI learns from every biologics experiment.

At Johnson & Johnson, we believe health is everything. Our strength in healthcare innovation empowers us to build a world where complex diseases are prevented, treated, and cured, where treatments are smarter and less invasive, and solutions are personal. Through our expertise in Innovative Medicine and MedTech, we are uniquely positioned to innovate across the full spectrum of healthcare solutions today to deliver the breakthroughs of tomorrow, and profoundly impact health for humanity. Learn more at jnj.com.

As guided by Our Credo, Johnson & Johnson is responsible to our employees who work with us throughout the world. We provide an inclusive work environment where each person is considered as an individual. At Johnson & Johnson, we respect the diversity and dignity of our employees and recognize their merit.

Job Function

Data Analytics & Computational Sciences

Job Sub Function

Data Science

Job Category

Scientific/Technology

All Job Posting Locations

Madrid, Spain

Our expertise in Innovative Medicine is informed and inspired by patients, whose insights fuel our science-based advancements. Visionaries like you work on teams that save lives by developing the medicines of tomorrow.

Join us in developing treatments, finding cures, and pioneering the path from lab to life while championing patients every step of the way.

Learn more at https://www.jnj.com/innovative-medicine

About the opportunity

Johnson & Johnson Innovative Medicine is seeking a Senior Data Scientist dedicated to our Biologics Discovery organization. This role sits within our Data, Data Science & Artificial Intelligence team (DDSAI) and partners closely with our In Silico Discovery (ISD) organization - the group that builds the molecular design and property-prediction models (for example, developability, affinity and binding, and other molecular-property and liability-risk models) that guide which biologic molecules to design, make, and advance. ISD owns core molecular model development; you will build the data-facing ML capabilities (featurization, model-ready datasets, evaluation frameworks, and applied models on assay and sequence data) that make ISD's models faster to build and better to trust.

This position will be based at one of our office locations in either Spring House, PA (strongly preferred), Titusville, NJ, or Raritan, NJ, USA; or Madrid, Spain. (No remote option.)

Please note that this role is available across multiple countries and may be posted under different requisition numbers to comply with local requirements. While you are welcome to apply to any or all of the postings, we recommend focusing on the specific country(s) that align with your preferred location(s):

USA - Requisition Number: R-095854

Spain - Requisition Number: R-096793

Why this role matters: Biologics Discovery is generating rich, fast-growing data across assays, sequences, and modalities, and the opportunity now is to make that data fully model-ready and seamlessly available for ML. This role ensures biologics data is structured for training, and that applied ML on discovery data helps scientists prioritize molecules, flag risks, and generate hypotheses earlier - strengthening the interface to ISD's models rather than duplicating them.

You will design robust featurization and dataset curation, build and evaluate applied models on biologics assay, biophysical, and sequence/construct data, and define evaluation frameworks that keep models trustworthy. You operate at the interface between our data-generating and data-infrastructure partners and In Silico Discovery (ISD), ensuring the datasets and features you create strengthen ISD's molecular property models. This is an opportunity to shape how AI learns from every biologics experiment.

Key Responsibilities

Featurization & Model-Ready Data

  • Develop featurization and model-ready datasets from antibody/protein sequence, construct, assay, and biophysical data.
  • Work with data engineers to specify the features, labels, and levels of aggregation that models need, preserving raw representations where information matters.
  • Curate, document, and version datasets so modeling is reproducible and traceable.

Applied ML & Evaluation

  • Develop featurization and model-ready datasets from antibody/protein sequence, construct, assay, and biophysical data.
  • Work with data engineers to specify the features, labels, and levels of aggregation that models need, preserving raw representations where information matters.
  • Curate, document, and version datasets so modeling is reproducible and traceable.

Partnership, Rigor & Growth

  • Collaborate with ISD to hand off standardized, traceable training datasets and align on where Data Science enables versus where ISD owns modeling.
  • Partner with Discovery scientists to frame ML problems around real decision points in the design-make-test-learn (DMTL) cycle.
  • Work closely with ontology and MLOps colleagues so datasets carry consistent semantics and models move reliably from development into use.
  • Champion reproducibility, documentation, and responsible AI.

Why This Role Is Unique

This is an opportunity to apply ML where it truly moves the needle in biologics discovery - grounded in real assay and sequence data, tightly partnered with world-class molecular modeling, and with real room to grow your scope, technical leadership, and impact as you build a track record of delivery.

Qualifications

Required

  • Master's or Ph.D. in Computer Science, Machine Learning, Computational Biology, Bioinformatics, Statistics, or a related field.
  • At least 2 years of applied ML experience, including model development, evaluation, and dataset curation on complex scientific or biomedical data.
  • Strong proficiency with Python and the modern ML stack (e.g., PyTorch, scikit-learn) and SQL.
  • Experience turning complex, heterogeneous experimental data into robust features and training sets, with exposure to cloud training and data infrastructure.
  • Sound understanding of evaluation, validation, and the risks of leakage and distribution shift.
  • Ability to collaborate effectively with experimental scientists and modeling partners in a matrixed R&D environment.

Preferred

  • Experience with biologics, antibody/protein sequence models, or protein language models.
  • Experience with active learning, Bayesian optimization, or sequence-based generative models for molecular design.
  • Familiarity with biophysical/assay data and developability endpoints.
  • Experience with MLOps, experiment tracking, and model monitoring.
  • Familiarity with how ontologies or knowledge graphs support data reuse and AI-ready datasets.

This position will be based at one of our office locations in either Spring House, PA (strongly preferred), Titusville, NJ, or Raritan, NJ, USA; or Madrid, Spain. (No remote option.)

#LI-SL

#JNJDataScience

#JNJIMRND-DS

#JRDDS

#LI-Hyrbid

# 3

Required Skills

Preferred Skills

Advanced Analytics, Business Intelligence (BI), Coaching, Collaboration, Critical Thinking, Data Analysis, Database Management, Data Privacy Standards, Data Reporting, Data Savvy, Data Science, Data Visualization, Econometric Models, Process Improvements, Technical Credibility, Technologically Savvy, Workflow Analysis

The anticipated base pay range for this position is

€55,400.00 - €87,860.00

Benefits

In addition to base pay, we offer the following benefits*: an annual bonus with set target (% of pay) depending on pay grade / location, where the actual amount is based on the employees’ and companies’ performance of the previous calendar year, or sales commissions. Moreover, we offer vacation days, parental leave for a minimum of 12 weeks, bereavement leave, caregiver leave, volunteer leave, well-being reimbursement, programs for financial, physical and mental health. We also offer service anniversary and recognition awards, and subject to the terms of their respective plans, employees - and in some location’s eligible dependents - can participate in several insurance plans. For more information, visit Employee benefits | Supporting well-being & career growth | Johnson & Johnson Careers.

*This is for informative purposes only. Amounts and actual benefits may vary by location and are subject to change.

Researching Johnson & Johnson before you apply?

See 55 open roles · Culture, benefits & locations.

View Johnson & Johnson profile

Frequently Asked Questions

Where is the job located, and is it remote/hybrid/on-site?
The job is located in Madrid, Spain. It is an on-site position based at one of the office locations, and there is no remote option.
What are the required qualifications and experience level for this role?
You need a Master's or Ph.D. in Computer Science, Machine Learning, Computational Biology, Bioinformatics, Statistics, or a related field. Additionally, you must have at least 2 years of applied ML experience, strong proficiency in Python and SQL, and experience turning complex experimental data into robust features and training sets.
What are the key responsibilities of the Senior Data Scientist?
You will design robust featurization and dataset curation, build and evaluate applied models on biologics assay, biophysical, and sequence/construct data, and define evaluation frameworks. You will also collaborate with In Silico Discovery (ISD) and Discovery scientists to frame ML problems around the design-make-test-learn cycle.
What is the salary range for this position?
The anticipated base pay range for this position is €55,400.00 - €87,860.00.
What benefits are offered with this role?
Benefits include an annual bonus, vacation days, a minimum of 12 weeks of parental leave, bereavement, caregiver, and volunteer leave. J&J also offers a well-being reimbursement, financial, physical, and mental health programs, service anniversary and recognition awards, and participation in several insurance plans.
What is the application process, and is there a specific requisition number?
You can apply to the posting that aligns with your preferred location. For the Spain-based position, the specific Requisition Number is R-096793.

Ready to Apply?

Apply for this Position

You'll be redirected to the company's application page

Share this job:

Explore Johnson & Johnson

Research the company before you apply.

  • 55 open roles
  • Culture, benefits & locations
View company profile

Job Information

Source: manual
AI Relevance: 92/100 (Highly relevant)
Remote Type: onsite
Allowed Locations: Worldwide
Skills & Tags:
johnson & johnson machine learning artificial intelligence bioinformatics computational biology data science protein

Get Similar Jobs by Email

Weekly digest of Johnson & Johnson and similar companies. Free.

Related Jobs

Apply for this Position

Get weekly job alerts