Senior Site Reliability Engineer (San Cugat del Vallés)

Senior Site Reliability Engineer (San Cugat del Vallés)

16 sep
|
Roche
|
San Cugat del Vallés

16 sep

Roche

San Cugat del Vallés

Overview
Asegúrese de presentar su candidatura con toda la información solicitada, tal como se expone en la descripción del puesto a continuación.
Join Roche as a Site Reliability Engineer to design, build, and scale reliable distributed systems that power healthcare innovation. You will drive reliability, automation, and operational excellence, collaborating with product teams to improve uptime and system performance. Expect on-call rotation, blameless postmortems, and a focus on reducing toil through engineering solutions. This role offers a chance to shape the backbone of IT platforms at a company committed to integral health impact.
Responsabilidades
Define and implement SLIs, SLOs, and error budgets with product and engineering teams
Conduct reliability reviews for new and existing services
Design scalable, fault-tolerant architectures in AWS and Azure
Lead capacity planning, performance and cost optimization
Improve system resilience through automation and self-healing patterns
Drive observability maturity (metrics, logs, traces, alert quality)
Incident management and continuous improvement through root cause analyses and postmortems
Handle requests and incidents, maintain runbooks




Participate in a 24x7 on-call rotation
Automation & platform engineering: reduce toil through tooling (Python or similar), improve CI/CD reliability, IaC with Terraform, enhance Kubernetes platforms (EKS/AKS/GKE)
Cross-functional leadership: collaborate with business, security, and cloud teams; mentor engineers; promote ownership and continuous improvement
Requisitos principales
Bachelor’s degree in computer science, engineering, or related field or equivalent experience
Production on-call experience in SRE or software engineering
Experience xqbhyrx with AWS and/or Azure (Kubernetes, EKS, AKS, GKE)
Proficiency with observability tools
Hands-on incident management tooling experience
Scripting skills for automation (e.g., Python)
Strong troubleshooting in cloud and distributed systems
Excellent communication, teamwork, and documentation skills
English proficiency
Diversity and inclusion mindset
strong communication
teamwork
proactive and self-motivated
AWS and/or Azure cloud platforms
Kubernetes (EKS/AKS/GKE)
Terraform or Infrastructure as Code

📌 Senior Site Reliability Engineer (San Cugat del Vallés)
🏢 Roche
📍 San Cugat del Vallés

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: senior site reliability engineer (san cugat del vallés) / san cugat del vallés