We are looking for an
SRE Consultant
to join an international, cloud-native environment, responsible for the operation, stability, automation, and evolution of cloud infrastructures and business applications deployed across Europe. The selected candidate will work closely with technical and operations teams based in Spain and China, contributing to architecture standardization, process automation, observability, and the continuous improvement of service performance, availability, and efficiency. Manage the end-to-end
operation, maintenance, and architectural governance
of public cloud infrastructures in Europe, including containers, virtual machines, storage, and networking. Ensure the operational stability and architectural standardization of
databases and middleware
. Design, implement, and maintain
automated CI/CD pipelines
, enabling standardized integration, delivery, and continuous deployment processes. Proactively identify, diagnose, and resolve
functional incidents, resource bottlenecks, and performance anomalies
. Manage the daily operation, maintenance, and
production releases
of cloud-native and containerized applications. Drive initiatives focused on
standardization, technology governance, architecture optimization, and application migration
towards cloud-native environments. Continuously analyze cloud and application resource consumption and utilization, identifying opportunities for
capacity, performance, and cost optimization
. Strong experience in
Linux administration and operations
, with an in-depth understanding of its core principles and components. Advanced knowledge of
networking and protocols
, particularly TCP/IP and Hands-on experience with
Kubernetes environments
, including a solid understanding of the architecture and operation of its main components. Experience operating and maintaining
Kubernetes Operators in production environments
. Experience with
observability and monitoring tools
,
particularly
Prometheus and Grafana
, with the ability to design and maintain comprehensive monitoring systems for cloud-native environments. Knowledge of the architecture, deployment, and operation of
API Gateways
, particularly: Nginx Knowledge of
middleware and messaging systems
, particularly: RabbitMQ Kafka Proficiency in at least one scripting language, preferably
Python or Shell
, for automation, batch processing, and the development of operational tools. Knowledge of
GDPR
, privacy, and data protection regulations applicable in Europe. Ability to apply regulatory compliance requirements to infrastructure and cloud operations. Experience in
multicloud or public cloud environments
. Experience with
cloud-native architectures and microservices
. Knowledge of cloud cost optimization and
FinOps
. Experience with application migration processes towards cloud-native architectures. Knowledge of
SRE, DevOps, and GitOps
methodologies. Experience in incident management,
Root Cause Analysis (RCA) , and continuous improvement. English:
Advanced level, essential for communication with international teams. Chinese:
Highly desirable due to ongoing collaboration with engineering and operations teams based in China. Candidates with
the legal right to work in Spain
and the ability to operate effectively in international environments will be particularly valued. Strong focus on service stability, availability, and quality. The opportunity to join a
highly technology-driven, international cloud-native environment
. Participation in European-wide infrastructure and application projects. The opportunity to work with technologies such as
Kubernetes, Prometheus, Grafana, ArgoCD, Nginx, APISIX, Envoy, Kafka, RabbitMQ, and Python/Shell
. Participation in cloud architecture automation, standardization, optimization, and evolution projects. Professional development opportunities within a dynamic and rapidly growing technology environment.
📌 Site Reliability Engineer (H/F) (Madrid)
🏢 Tieto
📍 Madrid