Site Reliability Engineer (Madrid)

Site Reliability Engineer (Madrid)

29 ago
|
Tieto
|
Madrid

29 ago

Tieto

Madrid

About the Position

We are looking for an SRE Consultant to join an international, cloud-native environment, responsible for the operation, stability, automation, and evolution of cloud infrastructures and business applications deployed across Europe.

The selected candidate will work closely with technical and operations teams based in Spain and China, contributing to architecture standardization, process automation, observability, and the continuous improvement of service performance, availability, and efficiency.

Responsibilities

- Manage the end-to-end operation, maintenance, and architectural governance of public cloud infrastructures in Europe, including containers, virtual machines, storage, and networking.
- Ensure the operational stability and architectural standardization of databases and middleware.
- Design, implement, and maintain automated CI/CD pipelines, enabling standardized integration, delivery, and continuous deployment processes.
- Participate in on-call rotations, providing support for incidents and production issues.
- Proactively identify, diagnose, and resolve functional incidents, resource bottlenecks, and performance anomalies.
- Manage the daily operation, maintenance, and production releases of cloud-native and containerized applications.
- Ensure the availability, stability, and proper evolution of production services.
- Collaborate with technical teams in China to define and implement common management standards for containerized applications.
- Drive initiatives focused on standardization, technology governance, architecture optimization, and application migration towards cloud-native environments.
- Continuously analyze cloud and application resource consumption and utilization, identifying opportunities for capacity, performance, and cost optimization.
- Promote continuous improvement in the efficiency and utilization of resources across European environments.
- Ensure that operational and maintenance activities comply with security, privacy, and applicable regulatory requirements in Europe.





Required Skills and Experience

- Strong experience in Linux administration and operations, with an in-depth understanding of its core principles and components.
- Advanced knowledge of networking and protocols, particularly TCP/IP and HTTP.
- Hands-on experience with Kubernetes environments, including a solid understanding of the architecture and operation of its main components.
- Experience operating and maintaining Kubernetes Operators in production environments.
- Experience with observability and monitoring tools, particularly Prometheus and Grafana, with the ability to design and maintain comprehensive monitoring systems for cloud-native environments.
- Experience with CI/CD tools and methodologies, particularly ArgoCD, and the ability to build, configure, and maintain automated pipelines.
- Knowledge of the architecture, deployment, and operation of API Gateways, particularly:
- Nginx
- APISIX
- Envoy
- Knowledge of middleware and messaging systems, particularly:
- RocketMQ
- RabbitMQ
- Kafka
- Proficiency in at least one scripting language, preferably Python or Shell, for automation, batch processing, and the development of operational tools.
- Knowledge of GDPR, privacy, and data protection regulations applicable in Europe.
- Ability to apply regulatory compliance requirements to infrastructure and cloud operations.

Nice-to-Have Skills

- Experience in multicloud or public cloud environments.
- Experience with cloud-native architectures and microservices.
- Experience in operations automation and Infrastructure as Code (IaC) practices.
- Knowledge of cloud cost optimization and FinOps.




- Experience working with international and geographically distributed technical teams.
- Experience in high-availability environments and business-critical systems.
- Experience with application migration processes towards cloud-native architectures.
- Knowledge of SRE, DevOps, and GitOps methodologies.
- Experience in incident management, Root Cause Analysis (RCA), and continuous improvement.

Languages

- English: Advanced level, essential for communication with international teams.
- Chinese: Highly desirable due to ongoing collaboration with engineering and operations teams based in China.
- Candidates with the legal right to work in Spain and the ability to operate effectively in international environments will be particularly valued.

Soft Skills

- Strong analytical and problem-solving skills.
- Strong focus on service stability, availability, and quality.
- Automation mindset and commitment to continuous improvement.
- Ability to work under pressure and manage production incidents.
- Excellent communication and collaboration skills when working with multidisciplinary technical teams.
- Ability to work independently and take end-to-end ownership of services.
- Strong focus on standardization, efficiency, and resource optimization.
- Adaptability to working in an international and multicultural environment.

What We Offer

- The opportunity to join a highly technology-driven, international cloud-native environment.
- Participation in European-wide infrastructure and application projects.
- Direct collaboration with international technology teams.
- The opportunity to work with technologies such as Kubernetes, Prometheus, Grafana, ArgoCD, Nginx, APISIX, Envoy, Kafka, RabbitMQ, and Python/Shell.
- Participation in cloud architecture automation, standardization, optimization, and evolution projects.
- Professional development opportunities within a dynamic and rapidly growing technology environment.

Location

Spain, with regular collaboration with technical and operations teams based in China.

📌 Site Reliability Engineer (Madrid)
🏢 Tieto
📍 Madrid

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (madrid) / madrid

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (madrid) / madrid