05 ago
|
Randstad (Schweiz
|
Madrid
05 ago
Randstad (Schweiz
Madrid
Site Reliability Engineer (SRE)
Puede obtener más detalles sobre la naturaleza de esta vacante y lo que se espera de los solicitantes leyendo la información a continuación.
You’ll own the reliability, scalability, and operational excellence of the systems that power our platform. You’ll partner closely with Engineering, Security, and Product to build resilient infrastructure, improve developer experience, and raise the bar on observability and incident response.
What you’ll be doing:
-
- Own and improve platform reliability (SLOs/SLIs), capacity planning, and production readiness
- Design, build, and maintain Kubernetes‑based infrastructure and deployment workflows
- Operate and evolve our AWS footprint (networking, compute, storage, IAM) with a security‑first mindset
- Improve our CD/GitOps practices (ArgoCD) and deployment safety (progressive delivery, rollbacks, guardrails)
- Build autoscaling strategies for services and workloads (KEDA where appropriate)
- Lead incident response: on‑call, triage, mitigation, post‑mortems, and preventative follow‑through
- Strengthen observability across services: metrics, logs, traces, and alerting (OpenTelemetry + dashboards)
- Partner with application teams to tune performance, reduce toil, and improve operational maturity
- Improve infrastructure‑as‑code practices and maintain Terraform modules and environments
- Contribute to evaluating and integrating AI tooling and MCP tools to accelerate operational workflows
What you should bring:
-
- 5+ years of experience in site reliability engineering
- Experience owning reliability for production systems — defining SLOs or running error budgets
- Deep hands‑on experience with Kubernetes/Helm on EKS in production
- Experience with Cloudflare (CDN, WAF, etc)
- Experience with CI/CD or general GitOps deployment pattern experience
- Has led or significantly contributed to incident response and postmortem processes
- Working knowledge of AWS core services (networking, IAM, compute) beyond just EKS
Nice to have:
-
- Experience evaluating or building AI‑assisted ops tooling (MCP, agentic runbooks)
- OpenTelemetry / observability pipeline design
- Experience with Terraform
Benefits:
-
- Health Insurance
- Paid Time Off (PTO)
- Paid Holidays
- Remote Work
- Professional Development
EEO Statement:
Fountain is a proud equal opportunity workplace. We welcome applicants of any educational background, gender identity and expression, sexual orientation, religion, ethnicity, age, socioeconomic status, disability, and veteran status. xhfqzwm
By submitting an application, you confirm that you have read our Privacy Policy and agree that we may process and retain your personal data for the purpose of recruitment in accordance with applicable data protection laws.
#J-18808-Ljbffr
📌 Senior Site Reliability Engineer (Madrid)
🏢 Randstad (Schweiz
📍 Madrid