Machine Learning Engineer — Inference Optimization (España)

Machine Learning Engineer — Inference Optimization (España)

17 ago
|
Jobgether
|
España

17 ago

Jobgether

España

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Machine Learning Engineer — Inference Optimization based in Spain.

This role offers the opportunity to optimize the performance of advanced machine learning systems used in real-world production environments.
You will work at the intersection of research and engineering, transforming cutting-edge models into fast, reliable, and cost-efficient solutions.
Your work will directly impact model scalability, user experience, and the efficiency of AI-powered products.
You will dive deep into performance optimization, from model architecture and GPU execution to large-scale inference infrastructure.
Working with talented research, infrastructure, and product teams, you will help push the boundaries of what AI systems can achieve.
This position is idóneo for an engineer who enjoys solving complex technical challenges and building high-performance ML systems from the ground up.

Accountabilities





As a Machine Learning Engineer specializing in inference optimization, you will own the performance and scalability of machine learning models in production. You will combine deep ML expertise, systems engineering, and performance analysis to deliver faster, more efficient AI experiences.
- Optimize machine learning inference systems to improve latency, throughput, scalability, and operational cost.
- Profile and identify bottlenecks across GPU and CPU inference pipelines, including memory usage, kernels, batching strategies, and data flow.
- Implement advanced optimization techniques such as quantization, KV-cache optimization, speculative decoding, batching, streaming, and model simplification.
- Collaborate with research engineers to productionize new model architectures and translate experimental results into reliable systems.
- Build, improve, and maintain inference-serving infrastructure using modern frameworks, custom runtimes, or specialized servin

📌 Machine Learning Engineer — Inference Optimization (España)
🏢 Jobgether
📍 España

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: machine learning engineer — inference optimization (españa) / españa

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: machine learning engineer — inference optimization (españa) / españa