Postdoctoral Scholar, AI Evaluation & Standards

jj· 7300-Janssen-Cilag S.A. Legal Entity
Apply Now ↗
Full timeHybrid Work7300-Janssen-Cilag S.A. Legal Entity

About this role

At Johnson & Johnson, we believe health is everything. Our strength in healthcare innovation empowers us to build a world where complex diseases are prevented, treated, and cured, where treatments are smarter and less invasive, and solutions are personal. Through our expertise in Innovative Medicine and MedTech, we are uniquely positioned to innovate across the full spectrum of healthcare solutions today to deliver the breakthroughs of tomorrow, and profoundly impact health for humanity. Learn more at jnj.com.

As guided by Our Credo, Johnson & Johnson is responsible to our employees who work with us throughout the world. We provide an inclusive work environment where each person is considered as an individual. At Johnson & Johnson, we respect the diversity and dignity of our employees and recognize their merit.

Job Function:

Career Programs

Job Sub Function:

Post Doc – Data Analytics & Computational Sciences

Job Category:

Career Program

All Job Posting Locations:

Barcelona, Spain, Beerse, Antwerp, Belgium, Madrid, Spain, Raritan, New Jersey, United States of America, Titusville, New Jersey, United States of America

Job Description:

Job description

At J&J we are developing Generative AI solutions to support pharmaceutical R&D, including literature review, evidence synthesis, document Q&A, therapeutic area knowledge search, translational science workflows, and R&D decision support.

These systems need to be tested before teams use them in scientific workflows. In pharmaceutical R&D, a useful AI response depends on the question, user, source material, therapeutic area, and risk of error.

We are looking for a postdoctoral researcher to help design methods that test whether GenAI tools produce answers that are accurate, evidence-grounded, traceable, usable, and appropriate for the intended task.

The role reports to the Associate Director, Generative AI Evaluation & Quality Standards. The team defines how J&J Innovative Medicine evaluates GenAI systems before use and helps determine when they are ready for release, expansion, or improvement.

Key responsibilities

  • Design evaluation frameworks, rubrics, and criteria for GenAI tools used across pharmaceutical R&D.
  • Develop therapeutic-area-specific criteria with business and scientific teams to reflect domain and use-case quality needs.
  • Build benchmark datasets, reference answer sets, annotation guides, and evaluation datasets.
  • Run expert reviews with scientific, clinical, regulatory, medical, data science, and engineering teams.
  • Test LLM, RAG, and agent performance, including accuracy, source grounding, retrieval quality, citation fidelity, task completion, robustness, safety, and usability.
  • Analyze failure patterns such as unsupported claims, incorrect reasoning, poor evidence use, missing uncertainty, weak traceability, or failure to follow instructions.
  • Translate evaluation findings into improvements in prompts, retrieval methods, agent workflows, tools, and user experience.
  • Help define release criteria for systems moving from prototype to limited release, expanded use, or product support.
  • Review emerging evaluation methods and adapt useful approaches for pharmaceutical R&D.
  • Document methods, findings, and recommendations so teams can apply consistent evaluation practices.
  • Design and develop agentic judge methods to evaluate GenAI outputs against defined criteria, flag evidence gaps or unsupported claims, and support expert review workflows.

Qualifications

Education

  • PhD or equivalent research experience in biomedical science, computational biology, bioinformatics, AI/ML, data science, clinical research, regulatory science, biostatistics, pharmaceutical sciences, or a related field.

Experience and skills

Required

  • Understanding of biomedical science, pharmaceutical R&D, therapeutic area science, translational science, clinical development, regulatory science, biomedical informatics, data science, or related areas.
  • Experience translating expert judgment into criteria, rubrics, datasets, protocols, or measurable outcomes.
  • Experience designing or applying evaluation methods, benchmark datasets, annotation protocols, validation studies, quality reviews, or assessment frameworks.
  • Interest in testing GenAI systems, including LLMs, RAG, and AI agents.
  • Proficiency in Python and common data science or machine learning tools.
  • Ability to analyze model outputs, compare performance, identify failure patterns, and recommend improvements.
  • Clear written and verbal communication skills.

Preferred

  • Experience with LLM APIs, embeddings, vector databases, prompt engineering, agent frameworks, or AI evaluation tools.
  • Experience evaluating retrieval quality, generated answers, multi-step workflows, tool use, scientific reasoning, citation quality, or evidence-grounded outputs.
  • Experience designing expert review workflows, annotation instructions, adjudication processes, or inter-rater reliability analyses.
  • Domain knowledge in one or more biomedical or therapeutic areas.
  • Familiarity with biomedical data standards, structured scientific or clinical data, ontologies, knowledge graphs, CDISC, FHIR, or related frameworks.
  • Publications or applied research in AI evaluation, NLP, biomedical informatics, machine learning, data science, computational biology, bioinformatics, or a related field.

Required Skills: GenAI evaluation, Python, data science, benchmarking, rubric design, biomedical research, technical communication.

Preferred Skills: RAG evaluation, agent evaluation, biomedical informatics, expert review, annotation protocols, therapeutic area expertise, responsible AI.

 

 

Required Skills:

 

 

Preferred Skills:

  

 

The anticipated base pay range for this position is:

€43,600.00 - €70,150.00

 

 

Benefits:

In addition to base pay, we offer the following benefits*: an annual bonus with set target (% of pay) depending on pay grade / location, where the actual amount is based on the employees’ and companies’ performance of the previous calendar year, or sales commissions. Moreover, we offer vacation days, parental leave for a minimum of 12 weeks, bereavement leave, caregiver leave, volunteer leave, well-being reimbursement, programs for financial, physical and mental health. We also offer service anniversary and recognition awards, and subject to the terms of their respective plans, employees - and in some location’s eligible dependents - can participate in several insurance plans. For more information, visit Employee benefits | Supporting well-being & career growth | Johnson & Johnson Careers.

 

*This is for informative purposes only. Amounts and actual benefits may vary by location and are subject to change.

 

 

Frequently Asked Questions

What is the salary for the Postdoctoral Scholar, AI Evaluation & Standards role at jj?
The listed salary for this Postdoctoral Scholar, AI Evaluation & Standards position at jj is EUR 44K–70K. This is an Full time role.
Where is the Postdoctoral Scholar, AI Evaluation & Standards position at jj located?
This Postdoctoral Scholar, AI Evaluation & Standards role at jj is based in 5 Locations, Barcelona, Spain, Beerse, Antwerp, Belgium, Madrid, Spain, Raritan, New Jersey, United States of America, Titusville, New Jersey, United States of America. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
Is the Postdoctoral Scholar, AI Evaluation & Standards role at jj full-time or part-time?
This is listed as a Full time position. It is posted as a Postdoctoral Scholar, AI Evaluation & Standards role in the 7300-Janssen-Cilag S.A. Legal Entity department at jj.
Which team or department does the Postdoctoral Scholar, AI Evaluation & Standards at jj belong to?
This Postdoctoral Scholar, AI Evaluation & Standards position is part of the 7300-Janssen-Cilag S.A. Legal Entity department at jj. See the full job description for more information about the team structure and responsibilities.
How do I apply for the Postdoctoral Scholar, AI Evaluation & Standards position at jj?
Click the "Apply Now" button on this page. You will be redirected to jj's official application portal hosted on workday where you can submit your application directly.
When was the Postdoctoral Scholar, AI Evaluation & Standards job at jj posted?
This Postdoctoral Scholar, AI Evaluation & Standards position at jj was posted on Jul 31, 2026. Apply as soon as possible — early applications are often reviewed first.
Postdoctoral Scholar, AI Evaluation & Standards
jj · 💰 EUR 44K–70K
Apply for this role ↗

You'll be redirected to jj's official application page on Workday.