Direct from source · No middlemen
58 open positions · Updated 1 week ago
Showing 20 of 58 positions
Search with filters →Join Sesame as an ML Model Serving Engineer to enhance our serving layer for voice agents with cutting-edge techniques.
Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Together AI as a Forward Deployed Engineer to optimize inference systems for strategic customers in a hands-on role.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Periodic Labs as an ML Systems Engineer to build and optimize large-scale training and reinforcement learning infrastructure.
Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.
Own the serving infrastructure for healthcare AI, optimizing LLM inference systems to enhance patient experiences.
Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.
Join Inferact as an inference runtime engineer to innovate AI inference engines for large models in a fully remote role.
Join ChipAgents as an ML Systems Engineer to optimize LLM inference systems for leading semiconductor companies.