13 ago
|
Talent-R
|
España
Senior Site Reliability Engineer
Location: Full remote from Spain
About the role
We are looking for a Senior Site Reliability Engineer to join a high-impact Platform team supporting one of Europe’s leading home improvement marketplaces.
This role is focused on ensuring the reliability, scalability, and performance of a large-scale cloud environment, helping engineering teams deliver value efficiently across 400+ microservices running on AWS.
What you will do
- Own deployment and platform operations standards and processes
- Drive incident resolution, root cause analysis, and post-mortem actions
- Improve on-call practices, runbooks, and alerting
- Partner with engineering teams to embed DevOps and reliability best practices
- Promote SRE culture, observability, and operational excellence across teams
What we are looking for
- 3+ years of experience in SRE, DevOps,
or infrastructure engineering
- Strong hands-on experience with AWS
- Strong knowledge of Kubernetes, Docker, Helm, nginx, Redis, GitLab, PostgreSQL, Kafka, and Terraform
- Experience with GitLab CI/CD and Datadog
- Good understanding of microservices and API architecture at scale
- Fluent English
Nice to have
- Experience with Istio and Envoy
- Programming skills in Python, Go, or similar
- Familiarity with AI coding assistants such as GitHub Copilot, Cursor, or Claude Code
Why join
You will join a strong Platform team working on complex reliability challenges in a modern AWS environment, with the opportunity to make a real impact on engineering efficiency and platform stability.
📌 Senior Site Reliability Engineer (España)
🏢 Talent-R
📍 España