Tether is seeking a Senior AI Research Engineer for Model Inference (Remote). You will extend the inference framework to support language models with strong Vulkan-based GPU acceleration for mobile and edge devices.
The role requires hands-on experience with quantization, LoRA fine-tuning, and Vulkan backend development, aiming to push desktop and on-device inference performance and efficiency.