Senior Site Reliability Engineer (Vallés)

Senior Site Reliability Engineer (Vallés)

31 jul
|
F. Hoffmann-La Roche
|
Vallés

31 jul

F. Hoffmann-La Roche

Vallés

Position Overview We are building a integral Site Reliability Engineering (SRE) team to support critical commercial and internal platforms and applications. As an SRE, you will design, build, and scale reliable distributed systems that power healthcare innovation worldwide. The role focuses on reliability, scalability, automation, and operational excellence and includes participation in a structured on‑call rotation. Core Responsibilities Define and implement SLIs, SLOs, and error budgets with product and engineering teams. Conduct reliability reviews for new and existing services. Design scalable, fault‑tolerant architectures in AWS and Azure environments. Lead capacity planning, performance and cost optimization initiatives. Improve system resilience through automation and self‑healing patterns. Drive organizational observability maturity (metrics, logs, traces, alert quality). Perform complex root‑cause analysis and drive rapid mitigation. Participate in blameless post‑mortems and follow‑through. Improve MTTR, reduce incident frequency, and elevate production standards. Collaborate seamlessly with engineering teams to enable timely and effective resolutions. Handle requests and incidents, create and maintain runbooks. Participate in a structured 24/7 on‑call rotation. Reduce operational toil through tooling and automation (Python or similar). Improve CI/CD reliability and deployment safety mechanisms. Build and maintain infrastructure‑as‑code (Terraform or equivalent). Enhance Kubernetes platform reliability (EKS, AKS, or similar). Partner with business, engineering, security, and cloud teams to embed reliability early in the software development life cycle.



Mentor mid‑level engineers and help shape SRE best practices. Champion a culture of ownership, accountability, and continuous improvement. Qualifications Bachelor’s degree in computer science, engineering, or a related field, or equivalent professional experience. Experience in site reliability engineering, software engineering, or related fields with production on‑call experience. Solid experience with AWS and/or Azure, including setting up, monitoring, and maintaining cloud resources (incl. Kubernetes, EKS, AKS, GKE). Proficiency with observability tools. Hands‑on experience with incident management tools. Proficiency in scripting languages for automation purposes (Python, etc.). Demonstrated proficiency in troubleshooting, especially in cloud and distributed system environments. Excellent communication, teamwork and documentation skills, with a proactive and self‑motivated approach to improving system reliability and operational efficiencies. Proficient spoken and written English communication. Location Primary location: Sant Cugat del Vallès. Additional locations may be available. Equal Opportunity Employer Roche is an Equal Opportunity Employer. We believe it’s urgent to deliver medical solutions right now – even as we develop innovations for the future. We are passionate about transforming patients’ lives. We are courageous in both decision and action. And we believe that good business means a better world. We are committed to scientific rigor, unassailable ethics, and access to medical innovations for all. We are proud of who we are, what we do, and how we do it. We are many, working as one across functions, across companies, and across the world. #J-18808-Ljbffr

📌 Senior Site Reliability Engineer (Vallés)
🏢 F. Hoffmann-La Roche
📍 Vallés

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: senior site reliability engineer (vallés) / vallés

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: senior site reliability engineer (vallés) / vallés