Are you a Site Reliability Engineer (SRE) in Madrid seeking a new interesting challenge?
If your answer is yes, it’s your lucky day so keep reading, it can be just what you're looking for!
✍️ WHAT WILL YOU DO?
We are looking for a dynamic, proactive and talented person to join our team and perform the following tasks:
- Design, scale and evolve highly available, resilient cloud platforms that support mission-critical services.
- Drive infrastructure automation through Infrastructure as Code (IaC), orchestration and self-service capabilities.
- Lead incident response, root cause analysis and service recovery, ensuring continuous reliability improvements.
- Optimize system performance,
scalability and capacity planning using data-driven insights and reliability metrics.
- Implement end-to-end observability strategies leveraging logs, metrics, tracing and proactive alerting.
- Develop advanced monitoring solutions that enable real-time visibility, faster troubleshooting and predictive operations.
- Design and optimize CI/CD pipelines, accelerating software delivery while maintaining platform stability and reliability.