19 sep
|
Emburse
|
Barcelona
Overview
Inscríbase (haciendo clic en el botón correspondiente) después de revisar toda la información relacionada con el trabajo a continuación.
As a System Operations Engineer at Emburse, you ensure reliable, scalable SaaS platforms with 24/7 availability. You’ll drive operational excellence across cloud-native services, working closely with engineering to shape resilient infrastructure. You’ll tackle complex outages, optimize performance, and automate solutions that protect the customer experience. This role offers impact in a fast-growing AI-powered finance platform and a collaborative, high-performance culture.
Compensaciones / Beneficios
competitive pay
adaptable work
inclusive, collaborative environment
Responsabilidades
Meet KPIs and SLAs while maintaining an error budget
Identify and implement preventative measures to minimize customer impact
Troubleshoot to improve availability, performance, and security across CR and Emburse
Code and automate applications on cloud platforms
Collaborate with Engineering leadership to build shared services
Participate in backlog grooming, epic planning, and sprint planning
Ensure platform reliability (4 nines)
Define non-functional requirements for scalable, highly available systems
Own cross-domain issues across DevOps, databases, networking, code, and infrastructure
Coordinate with stakeholders to align operational priorities with product roadmap
Prepare and present engineering-related documents to stakeholders
Provide feedback in reviews and design sessions
Mentor SRE I and II and guide junior engineers
Conduct investigations, testing, and deployment activities to mitigate risks
Requisitos principales
Bachelor’s degree in Computer Science or a STEM field
Minimum of 7 years’ engineering experience
Deep understanding of infrastructure as code, scripting, self-healing, containers, DevOps tooling, distributed systems
Strong Kubernetes experience
Observability background
Preferred AWS with xqbhyrx optional Azure Cloud experience
Experience with Ansible and Terraform
Excellent English communication
Experience with full lifecycle of SaaS implementations and infrastructure as code
Excellent follow-up and project management skills
Proven ability to create and maintain new tools
Excellent troubleshooting and technical skills
Up to 70% hands-on in distributed Linux environments
Strong scripting skills; OOP a plus
Liaise between teams to align priorities
Experience with offshore teams
strong communication
project management
mentoring
Kubernetes
Observability
infrastructure as code
📌 Site Reliability Engineer lll (Barcelona)
🏢 Emburse
📍 Barcelona