01 oct
|
Johnson & Johnson Innovative Medicine
|
Cornellà de Llobregat
01 oct
Johnson & Johnson Innovative Medicine
Cornellà de Llobregat
At Johnson & Johnson,we believe health is everything. Our strength in healthcare innovation empowers us to build aworld where complex diseases are prevented, treated, and cured,where treatments are smarter and less invasive, andsolutions are personal.Through our expertise in Innovative Medicine and MedTech, we are uniquely positioned to innovate across the full spectrum of healthcare solutions today to deliver the breakthroughs of tomorrow, and profoundly impact health for humanity.Learn more at .
¿Todo listo para enviar su solicitud? Por favor, lea la descripción al menos una vez antes de hacer clic en "Solicitar".
As guided by Our Credo, Johnson & Johnson is responsible to our employees who work with us throughout the world. We provide an inclusive work environment where each person is considered as an individual. At Johnson & Johnson, we respect the diversity and dignity of our employees and recognize their merit.
Job FunctionJob Sub FunctionJob CategoryAll Job Posting Locations:
Cornellà de Llobregat, Barcelona, Spain, Madrid, Spain
Job Description
At J&J; we are building Generative AI solutions to support pharmaceutical R&D; — literature review, evidence synthesis, document Q&A;, therapeutic area knowledge search, translational science workflows, and R&D; decision support. These systems need to be evaluated before teams rely on them in scientific workflows. In pharma, a useful AI response depends on the question, user, source material, therapeutic area, and risk of error — so quality must be measurable, repeatable, traceable, and scientifically defensible.
Key Responsibilities
- Design, build, and maintain automated evaluation pipelines for LLM quality, RAG performance, agent reliability, safety, and scientific accuracy.
- Author evaluation rubrics and scoring criteria, curate golden and synthetic datasets with domain experts,
and maintain our registry of reusable evaluation assets.
- Validate AI judges against human expert agreement and run model, prompt, retriever, and agent benchmarks that produce standardized quality readouts.
- Analyze failure patterns — hallucination, unsupported claims, weak traceability — and turn findings into actionable recommendations.
- Develop therapeutic‑area‑specific evaluation criteria with scientific, clinical, and regulatory partners, refining them based on real‑world feedback.
- Design evaluation methods for scientific reasoning, evidence synthesis, and hypothesis quality — where generic benchmarks fall short.
- Build evaluation tooling and reusable patterns that enable other teams to self‑serve.
QualificationsEducation:
Master's degree in AI/ML, Computer Science, Data Science, Computational Biology, Bioinformatics, Biomedical Engineering, Applied Mathematics, Biostatistics, or a related field required. PhD preferred.
RequiredEXPERIENCE AND SKILLS:
We are looking for someone with 6+ years of hands‑on experience in AI/ML evaluation or data science (Master's) or 3+ years of industry experience (PhD). You should have experience designing and running evaluation frameworks, scientific benchmarks, or quality assessments for AI/ML systems, and hands‑on work with generative AI — large language models, retrieval‑augmented generation, agentic frameworks, and prompt engineering. We also value strong proficiency in Python and modern AI/ML tooling (evaluation harnesses, embedding models, vector databases, LLM APIs), the ability to translate expert scientific judgment into measurable criteria, rubrics, and reproducible protocols, and a collaborative, self‑driven approach to working across multidisciplinary teams. xqbhyrx
PreferredOther
English proficiency is required (written and verbal). This is a hybrid role based in Madrid or Barcelona, with limited travel (
📌 Senior Scientist - GenAI Evaluation (Cornellà de Llobregat)
🏢 Johnson & Johnson Innovative Medicine
📍 Cornellà de Llobregat