Site reliability engineer (Madrid)

Site reliability engineer (Madrid)

30 ago
|
FPT Software
|
Madrid

30 ago

FPT Software

Madrid

Responsibilities and DutiesOwn end-to-end operation, maintenance support and architecture governance of overseas public cloud infrastructures, covering core cloud resources including containers, cloud virtual machines, storage and networks. Also oversee the operational stability and architectural standardization of overseas databases and middleware systems.Design, implement and maintain automated CI/CD pipelines to support continuous integration, continuous delivery and standardized release workflows for overseas business applications.Undertake daily on-call rotation responsibilities. Proactively troubleshoot and resolve functional defects, resource bottlenecks and performance anomalies of cloud infrastructure and business applications, and drive continuous optimization of overall resource performance and service stability.Manage daily operation, maintenance and production release of overseas containerized and cloudnative applications, ensuring reliable and smooth online iteration of business services.Collaborate closely with domestic technical teams to formulate and implement unified containerized application management specifications across the platform, and steadily promote standardized governance, architecture optimization and business migration initiatives for overseas applications.Continuously analyze the usage status of overseas cloud and application resources, drive resource scheduling optimization and efficiency improvement, and maximize overall resource utilization and costeffectiveness of overseas business environments.We are looking for you whoSolid mastery of Linux operating system principles and core network protocols, including TCP/IP and HTTP,



with proficient hands-on operational capabilities.In-depth understanding of the Kubernetes ecosystem and the working mechanisms of core components; proficient in operating and maintaining Kubernetes Operators in production environments.Proficient in the principles and practical usage of mainstream observability tools including Prometheus and Grafana, capable of building and maintaining complete monitoring systems for cloud-native environments.Familiar with mainstream CI/CD toolchains represented by ArgoCD, with solid capabilities to build, configure and maintain automated continuous integration and delivery pipelines.Familiar with the architecture, operational principles, deployment and daily O&M; of common cloudnative gateway systems including Nginx, APISIX and Envoy, as well as mainstream message queue middleware such as RocketMQ, RabbitMQ and Kafka.Proficient in at least one mainstream scripting language (Python / Shell) to support daily automation operation, batch processing and operational tool development.Have a clear understanding of overseas data compliance specifications and privacy protection regulations such as GDPR, able to carry out cloud operation and maintenance work in compliance with regional regulatory requirements.Given that these positions will collaborate closely with engineering and operations teams located in China, strong Chinese and English communication skills are highly desirable. To facilitate efficient communication, collaboration, and alignment with China-based stakeholders, Chinese-speaking candidates with the legal right to work in Spain are strongly preferred. However, qualified international candidates meeting the required competencies will also be considered.

📌 Site reliability engineer (Madrid)
🏢 FPT Software
📍 Madrid

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (madrid) / madrid

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (madrid) / madrid