23 ago
|
Tether.io
|
Madrid
Senior AI Research Engineer, Model Inference (Remote) Join to apply for the Senior AI Research Engineer, Model Inference (Remote) role at Tether.Join Tether and Shape the Future of Digital Finance At Tether, we’re not just building products, we’re pioneering a global financial revolution. By harnessing the power of blockchain technology, Tether enables you to store, send, and receive digital tokens instantly, securely, and globally, all at a fraction of the cost.
Tether Finance: Our innovative product suite features the world’s most trusted stablecoin, USDT , relied upon by hundreds of millions worldwide, alongside pioneering digital asset tokenization services.
Tether Power: Driving sustainable growth, our energy solutions optimize excess power for Bitcoin mining using eco-friendly practices in state-of-the-art, geo-diverse facilities.
Tether Data: Fueling breakthroughs in AI and peer-to-peer technology, we reduce infrastructure costs and enhance global communications with cutting-edge solutions like KEET , our flagship app that redefines secure and private data sharing.
Tether Education : Democratizing access to top-tier digital learning, we empower individuals to thrive in the digital and gig economies, driving global growth and opportunity.
Tether Evolution : At the intersection of technology and human potential, we are pushing the boundaries of what is possible, crafting a future where innovation and human capabilities merge in powerful, unprecedented ways. Our team is a general talent powerhouse, working remotely from every corner of the world.
If you have excellent English communication skills and are ready to contribute to the most innovative platform on the planet, Tether is the place for you.
This role requires hands-on experience with quantization techniques, LoRA architectures, Vulkan backend, and mobile GPU debugging. You will play a critical role in pushing the boundaries of desktop and on-device inference and fine-tuning performance for next-generation SLM/LLMs. Integrate and validate quantization workflows for training and inference. perplexity testing, fine-tuned adapter performance).
Conduct GPU testing across desktop and mobile devices. Collaborate with research and engineering teams to prototype, benchmark, and scale new model optimization methods. Deliver production-grade, efficient language model deployment for mobile and edge use cases.
Define clear success metrics such as improved real-world performance, low error rates, robust scalability, optimal memory usage and ensure continuous monitoring and iterative refinements for sustained improvements. Proficiency in C++ and GPU kernel programming. Familiarity with LoRA fine-tuning and parameter-efficient training methods.
Ability to debug GPU-specific performance and stability issues on desktop and mobile devices. Demonstrated ability to apply empirical research to overcome challenges in model optimization. We do not conduct interviews over WhatsApp, Telegram, or SMS.
If someone asks for personal financial information or payment at any point during the hiring process, it is a scam. Please report it immediately. #
📌 Research Engineer (Madrid)
🏢 Tether.io
📍 Madrid