Staff Platform Engineer (Platform & Cloud Specialist) (Barcelona)

Staff Platform Engineer (Platform & Cloud Specialist) (Barcelona)

05 ago
|
Talent
|
Barcelona

05 ago

Talent

Barcelona

Job Description

You will join the core InfoJobs Platform & Infrastructure team, a core function for maintaining our position as the #1 jobs marketplace in the Spanish market. Your role is to design, implement, and operate all critical application delivery and observability components, ensuring unparalleled reliability, security, and scalability for our high-traffic platform (+6M MAUs).

This role is key to driving large-scale modernization and migration initiatives across our application edge and telemetry stacks. Your expertise will be crucial for maintaining and evolving the infrastructure underpinning the legacy Monolith and the modern microservices, establishing a world-class standard for how traffic is routed, secured, and monitored.

Key Areas Of Responsibility

- Defining the short, medium, and long-term technical roadmap for the Jobs Platform and Infrastructure.
- Knowledge for the Application Edge and API Gateway layer (KrakenD).
- Owning the stability and vision of the Observability and Telemetry stack (Distributed Tracing, Vector, Datadog, ELK).
- Leading the evolution of Application Authentication and Security protocols (OAuth2, OIDC) across the ecosystem.
- Guiding and executing logging systems to modern, high-performance solutions.
- Providing technical guidance on JVM performance and Spring optimization, ensuring the efficiency of critical services.
- Providing operational expertise, including on-call readiness and incident response leadership.

Position Specific Competencies

- Innovative Problem Solving: Solves highly ambiguous, complex, and innovative problems (e.g., performance bottlenecks at scale), setting new benchmarks for the engineering organization.
- Technical Architecture & Strategy: Capacity to lead architectural design, define the medium/long-term roadmap, successfully execute them and provide technical guidance across different domains.
- Driving Continuous Improvement: Focus on extreme automation, optimizing critical systems, and leading the transition to modern infrastructure patterns.
- Site Reliability & Critical Environment Management: Demonstrated experience doing on‑call duties in critical environments and quick on‑the‑go problem‑solving.
- System Thinking & Operational Ownership:



Ability to understand and take ownership of a complex stack (monolith and microservices) and deep knowledge of component interactions, and take full accountability for platform operation, health and performance.
- Drive Continuous Improvement: Proactively push forward improvements to prevent regression or waste, with focus on automation, optimizing existing processes and adopting cutting‑edge solutions.
- Adaptability & Resilience: High capacity for adaptation to changes and demonstrating resilience under pressure or when facing high uncertainty.
- Communicates Effectively: Essential for aligning strategy, driving incident communication and mentoring engineers.

Requirements

- Expertise in Edge & Traffic: Demonstrable expertise in managing API Gateways (KrakenD) and high‑volume traffic routing.
- Mastery of Observability: Proven track record in designing and leading Distributed Tracing strategies and managing telemetry pipelines (Vector, Fluentd, Logstash, ELK/OVK).
- Deep JVM & Runtime Knowledge: Extensive experience in JVM performance monitoring and tuning, specifically for Spring environments.
- Authentication & Security: In‑depth operational knowledge of OAuth2/OIDC and application security protocols.
- Cloud & Infrastructure: Mastery of AWS, Kubernetes (EKS), and advanced Infrastructure as Code (Terraform).
- Messaging Systems: Practical experience managing and scaling Kafka within a critical platform environment.
- On‑Call Leadership: Mandatory experience doing on‑call duties in 24x7 critical production environments.
- Strategic Background: Experience defining technical roadmaps and leading large‑scale architectural migrations.
- Demonstrated soft skills: passion, integrity, collaborative mindset, ownership, leadership, proactivity, can‑do attitude, resilience and fast adaptation to changes and challenges.

Nice To Have

- Fluent in English.
- Experience with WAF/Security technologies like Imperva.

Job Responsibilities





As a Staff Platform Engineer, you will act as a strategic technical leader with global impact, focusing on the architecture, reliability, and long‑term vision of our Application Delivery Platform. Your mission is to leverage your deep expertise in Traffic Management, Distributed Tracing, and JVM internals to maintain 24x7 resilience while leading complex, innovative migration projects. You will act as a force multiplier, solving systemic architectural challenges and mentoring engineers to drive the strategic alignment between the infrastructure platform and the application development lifecycle.

Main Responsabilities

- Strategic Edge Leadership: Lead, design, and execute large‑scale, complex projects related to the KrakenD API Gateway and general traffic routing, ensuring minimal disruption and maximum architectural gain.
- Architecture & Roadmap: Define the medium and long‑term technical roadmap for the Observability and Application Runtime layers, solving problems of high complexity and innovation.
- Critical Traffic Ownership: Act as the Subject Matter Expert (SME) for all critical traffic entry points, overseeing the health of Gateways, Load Balancers, and WAF environments.
- Observability & Reliability Excellence: Own the end‑to‑end strategy for Distributed Tracing and the telemetry pipeline (Vector, Fluentd, Logstash), ensuring proactive identification of failures before they impact users.
- Runtime & Performance Mastery: Provide technical leadership and guidance on JVM performance tuning and monitoring for Spring‑based workloads, optimizing resource consumption and response times.
- Application Security Governance: Lead the implementation and maintenance of robust Application Authentication (OAuth2/OIDC) and security standards across the platform.
- Operational Excellence (On‑Call): Participate in the on‑call rotation as a senior escalation point, leveraging deep knowledge of the Monolith and Java ecosystem to resolve high‑impact incidents quickly.
- Soft Skills & Collaboration: Utilize strong soft skills to align stakeholders, lead technical discussions across domain boundaries, and foster a culture of technical excellence and shared ownership.

#J-18808-Ljbffr

📌 Staff Platform Engineer (Platform & Cloud Specialist) (Barcelona)
🏢 Talent
📍 Barcelona

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: staff platform engineer (platform & cloud specialist) (barcelona) / barcelona

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: staff platform engineer (platform & cloud specialist) (barcelona) / barcelona