Direct from source · No middlemen
58 open positions · Updated 2 days ago
Showing 20 of 58 positions
Search with filters →Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.
Own Sarvam's production serving path for large distributed models, integrating and optimizing performance across a multi-node stack.
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Join Fireworks AI as a senior AI Field Engineer to build production systems for generative AI with leading organizations.
Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.
Join ElevenLabs as a Research Engineer to deploy and optimize AI models for real-time applications in a fully remote environment.
Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.
Join Cloudflare as a Senior Machine Learning Engineer to optimize and productionize ML models for a global serverless inference platform.
Own the architecture of Sarvam's vision models serving harness, ensuring high-quality document intelligence at national scale.
Join Fireworks AI as a Software Engineer to design and build scalable infrastructure for generative AI systems.
Join Preference Model as a Senior Machine Learning Engineer to design RL environments for advancing ML capabilities.
Own the end-to-end lifecycle of production ML serving systems for a top-performing AI Shopping Agent.
Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI inference workloads.
Join Preference Model as a senior ML Engineer to design RL environments for advancing machine learning capabilities.
Own the inference systems that power frontier AI models in production and research at a tech-first startup.
Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.
Join Fireworks AI as a senior AI Field Engineer to build production systems for innovative AI-native companies.