05 ago
|
Talent
|
Barcelona
OverviewSiga leyendo para comprender completamente lo que este trabajo requiere en cuanto a habilidades y experiencia. Si su perfil encaja, presente su candidatura.Keysight is at the forefront of technology innovation, delivering breakthroughs and trusted insights in electronic design, simulation, prototyping, test, manufacturing, and optimization. Our ~15,000 employees create world-class solutions in communications, 5G, automotive, energy, quantum, aerospace, defense, and semiconductor markets for customers in over 100 countries. Learn more about what we do.Our award-winning culture embraces a bold vision of where technology can take us and a passion for tackling challenging problems with industry-first solutions. We believe that when people feel a sense of belonging, they can be more creative, innovative, and thrive at all points in their careers.About the InitiativeKeysight’s Applied AI Autonomy Initiative is developing a next-generation agentic orchestration framework that enables AI agents to reason, adapt, and coordinate across complex engineering workflows. Built on LangGraph and reinforcement-inspired feedback mechanisms, this framework transforms prompts and design intents into executable orchestration strategies that evolve autonomously through iterative simulation and validation loops.Our ambition is not merely to replicate human reasoning, but to push past human limits - enabling agentic systems to explore design spaces, optimize engineering workflows, and evolve orchestration strategies at a scale and speed no human could achieve.This role defines the safety, stability, and observability architecture underpinning Keysight’s agentic runtime — the layer that ensures AI-driven orchestration remains interpretable, reversible, and aligned with human intent. You will design the mechanisms that make autonomy trustworthy: guardrails, rollback systems, introspection APIs, and adaptive feedback loops governing every agentic decision and simulator interaction.ResponsibilitiesRole OverviewAs the Senior Agentic Runtime Safety & Stability Engineer, you will own the resilience and transparency backbone of Keysight’s multi-agent orchestration stack.You will architect the runtime contracts, monitoring systems, and adaptive control mechanisms that ensure:Every AI-driven orchestration step issafe, auditable, and predictableThe system candetect, explain, and recoverfrom unsafe or emergent behaviorsHuman intent is faithfully interpreted andsecurely executedClosed-loop interactions betweenLLM-based agents, reinforcement learning systems,andEDA simulatorsare continuously monitored and governedThis position bridgesAI reasoning, runtime systems engineering,
and control safety — creating a foundation where autonomous orchestration is both powerful and predictable.Core Responsibility DomainsRuntime Guardrails, Intent Safety & Execution ControlArchitect runtime guardrails and authorization layers ensuring that agent actions remain aligned with operator intent, policy boundaries, and simulation constraints.Implementintent validation ,semantic disambiguation , andprompt safety checksbefore orchestration execution.Define structuredsafety contractsgoverning valid operating ranges, escalation paths, and rollback logic.Integrate safety constructs into orchestration semantics and graph-based reasoning flows with the Agentic Framework Architect.Fault Isolation, Rollback & Recovery EngineeringDesign deterministic rollback and checkpointing mechanisms to restore stable orchestration states after failure and enableautomatic recovery pathsfor misaligned or unsafe agent behavior.Engineerfault-isolation boundariesto contain local agent or simulator errors and prevent systemic instability.Buildsandboxed execution environmentsfor validating AI-generated orchestration logic safely.Developinteroperability safety layersbetween Python and RL technologies to ensure reliable data exchange and robust error containment in simulation-driven loops.Telemetry, Observability & Introspective DiagnosticsImplement comprehensiveobservability pipelinescapturing agent reasoning traces, simulation telemetry, and orchestration health metrics.Createreal-time anomaly detectionandconfidence-scored safety gatingto monitor drift, misalignment, or policy violations.Developintrospection APIs and dashboardsexposing safety metrics, decision rationales, and performance diagnostics.Collaborate with DevOps and Data Intelligence teams to unify telemetry across heterogeneous runtime components into a coherent monitoring fabric.Adaptive Governance & Continuous Safety LearningEstablishadaptive feedback systemsthat adjust orchestration parameters based on observed performance, safety signals, and environmental dynamics.Defineself-correcting safety policiesenabling agents to learn from past instability and improve compliance autonomously.Integrate safety scoring into promotion gates and validation workflows for runtime certification of agentic logic.Partner with ML and validation engineers to evolve acontinuous assurance pipelinethat evaluates trust, stability, and interpretability over time.Key ResponsibilitiesArchitect and own thesafety, observability,
and governance layerof Keysight’s agentic orchestration runtime.Design real-time self-healing and self-correcting mechanisms that detect misalignment, autonomously mitigate instability, and restore safe operational behavior without degrading user experience.Build deterministic rollback, checkpointing, and containment systems for multi-agent and simulation-based environments.Implement multi-layered telemetry, anomaly detection, and runtime introspection pipelines.Integrate observability acrossLLM, RL, and simulation environmentsinto a unified safety and diagnostics interface.Collaborate cross-functionally to embed transparency, traceability, and adaptive safety into every orchestration cycle.QualificationsRequired QualificationsPhD or 5+ years of experience in systems reliability, safety-critical software, or autonomous runtime engineering.Advanced proficiency inPython and C/C++ , with experience in hybrid or simulation-based systems.Proven expertise designingfault-tolerant, observable, and recoverabledistributed systems.Deep proficiency withagentic orchestration frameworks(LangGraph, LangChain, or equivalents).Strong understanding ofintent alignment, policy enforcement, and execution traceabilityin AI automation.Hands‑on experience implementingtelemetry, monitoring, and introspection systemsin complex runtime architectures.Preferred QualificationsBackground inmission-critical or regulated runtime systems(e.g., aerospace, industrial control, EDA, or HPC).Experience designingsemantic safety validation , policy modeling, and goal disambiguation frameworks.Familiarity withadaptive rollback, dynamic gating, and safety scoringin multi-agent environments.Proficiency withPython/C++ interoperability(PyBind11, gRPC, ZeroMQ).Understanding ofdeterministic simulation controland real-time anomaly detection in hybrid AI–physics systems.What This Role OffersA foundational and high-impact roledefining the safety, stability, and observability backbone of Keysight’s next‑generation agentic orchestration systems.The opportunity toengineer mission‑grade guardrails, rollback logic, and transparency mechanismsthat ensure autonomous multi‑agent workflows remain predictable, interpretable, and aligned with human intent.Direct influence on how AI agents reason, act, recover, and self‑correctwithin high‑assurance engineering environments — embedding trust, traceability, and adaptive safety into every orchestration cycle.A leadership position at the intersection ofruntime engineering, intelligent systems, and safety‑critical autonomy , helping shape how Keysight deploys and governs the next era of agentic intelligence. xcskxlj Careers Privacy Statement ***Keysight is an Equal Opportunity Employer.***#J-18808-Ljbffr
📌 Senior Software Engineer – Agentic Runtime Safety & Observability (Barcelona)
🏢 Talent
📍 Barcelona