Direct from source · No middlemen
56 open positions · Updated 1 week ago
Showing 20 of 56 positions
Search with filters →Join Inferact as an inference runtime engineer to innovate AI inference engines for large models in a fully remote role.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.
Own the serving infrastructure for healthcare AI, optimizing LLM inference systems to enhance patient experiences.
Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.
Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.
Join Baseten as a Software Engineer focusing on Model APIs to enhance AI model performance and developer experience.
Join Periodic Labs as an ML Systems Engineer to build and optimize large-scale training and reinforcement learning infrastructure.
Own the inference systems that power frontier AI models in production and research at a tech-first startup.
Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced machine learning models.
Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.
Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced AI models.