RequirementsStrong demonstrable ability to work with Linux systems and cloud platforms (AWS, GCP or Azure)Solid Kubernetes knowledge and ability to run production systemsA clear understanding of observability (monitoring, logging, tracing)Capable of designing or operating high-availability, distributed systemsA mindset focused on automation, scalability, and continuous improvementConfidence working in fast-moving environments where reliability really mattersWhat the job involvesAt Open Cosmos, our Data division transforms satellite data into meaningful insights that drive real-world impact. The team delivers all data products generated by Open Cosmos and its partners, curates and develops DataCosmos (our geospatial data platform)
and builds integrations that make satellite imagery easy to access and act onWe’re now looking for a Site Reliability Engineer to help us ensure our data platform is reliable, scalable, and performing at its best as we growOwning the reliability, performance, and scalability of our data platform and processing pipelinesMonitoring systems end-to-end, ensuring full visibility across infrastructure and data flowsResponding to incidents, troubleshooting issues, and driving long-term fixesImproving deployments and contributing to CI/CD pipelines for safe, repeatable releasesWorking closely with engineering teams to design resilient, scalable systemsAutomating processes and reducing operational overheadSupporting customer-impacting issues alongside Customer Success teams#J-18808-Ljbffr
📌 Site Reliability Engineer (Datacosmos) (Madrid)
🏢 EF
📍 Madrid
Postulate a este anuncio
Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.