Design and evolve GCP cloud architecture, including networking, interconnects, IAM, and high-availability topology, using Terraform.
Build and own CI/CD pipelines for infrastructure planning, review, testing, policy enforcement, drift detection, and progressive rollout.
Strengthen metrics, logging, tracing, and alerting through the Prometheus, Thanos, Grafana, Loki, Tempo, and Alertmanager observability stack.
Operate GKE clusters, Helm-packaged workloads, RabbitMQ and IBM MQ message brokers, and data stores in production.
Embed SRE practices including SLIs, SLOs, error budgets, and capacity planning into infrastructure operations.
At least 5 years of experience in DevOps, platform/infrastructure, or SRE roles operating large-scale, highly available, high-performance production systems.
Deep hands-on experience designing Google Cloud Platform architecture, including landing zones, networking, IAM, and high-availability topology.
Strong Terraform and Infrastructure-as-Code experience across multiple environments, with GitOps and least-privilege practices.
Significant production experience with Kubernetes, preferably GKE, and Helm-based workload deployment.
Strong cloud and L3/L4-L7 networking fundamentals, including VPCs, routing, load balancing, DNS, TLS, and interconnects.
Hands-on experience with Prometheus, Thanos, Grafana, Loki, Tempo, and Alertmanager for metrics, logs, traces, and alerting.
Operator-level familiarity with PostgreSQL and message brokers such as RabbitMQ or RedPanda.
Understanding of SRE practices, platform-as-a-product principles, incident management, capacity planning, and error budgets.
Preferred experience includes OPA/Conftest, Checkov, tflint, Atlantis, Terraform state and module management, Backstage, Tilt, Alloy, Rootly, Go, Linux, Debian, Ubuntu, Docker, containerd, security and compliance, SOC 2, secrets management, audit logging, and regulated fintech or low-latency systems.
One-time USD $500 new-hire home-office setup allowance.
📌 SENIOR DEVOPS ENGINEER (España)
🏢 Alpaca
📍 España