Site Reliability Engineer (Madrid)

Site Reliability Engineer (Madrid)

12 ago
|
Allianz
|
Madrid

12 ago

Allianz

Madrid

Experteer Overview

¿Es este el siguiente paso en su carrera? Descubra si es el candidato adecuado leyendo la descripción completa a continuación.
In this Site Reliability Engineer role, you will own the reliability of the central engineering platform within the Advanced Analytics domain. You’ll partner with platform, security, and incident-response teams to meet reliability commitments across AI services, Java APIs, and frontend workloads. You drive automation to eliminate toil, define SLOs/SLIs, and guide incident response and post-incident reviews. The role blends platform engineering, cloud infrastructure, and observability to enable safe, scalable, and cost-conscious delivery.

Compensaciones / Ventajas
• Define, instrument, and maintain SLOs/SLIs for platform components with error-budget tracking and leadership reporting
• Lead on-call escalation and incident response for cluster, network, and storage failures; chair blameless post-incident reviews
• Operate Kubernetes infrastructure (AKS): cluster lifecycle, networking, quotas, autoscaling, and multi-tenancy




• Develop Infrastructure as Code (Terraform) to provision and manage Azure resources with rollbacks
• Build and maintain observability stack: Prometheus, Grafana, Azure Monitor, Application Insights; manage alerts, dashboards, and tracing
• Perform capacity planning and cost-aware resource management across namespaces
• Identify and automate toil through scripting and tooling; measure toil reduction over time
• Maintain platform reliability procedures: upgrades, backup/recovery tests, DR runbooks, change freeze coordination
• Contribute to CI/CD and GitOps (GitHub Actions, ArgoCD) with reliability-focused release gates and rollback mechanisms
• Collaborate with incident-response and security teams on SRE targets, hardening, and vulnerability remediation

Responsabilidades
• 5+ years in SRE, DevOps, or platform engineering
• Strong Kubernetes experience including cluster operations, networking, storage,

📌 Site Reliability Engineer (Madrid)
🏢 Allianz
📍 Madrid

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (madrid) / madrid

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (madrid) / madrid