"triton in" Jobs
82 open tech roles matching “triton in”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Triton, CUDA. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 82 results
Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.
Lead a team of MLOps engineers at an AI company transforming enterprise decision-making.
Join Anthropic as a Staff Engineer to lead the technical direction of the Inference Runtime for AI systems serving millions of users.
Join Triomics as an MLOps & Data Engineer to build infrastructure for ML workflows in oncology, impacting patient outcomes.
Join Together AI as a Systems Research Engineer Intern to optimize GPU programming for ML/AI applications in a collaborative environment.
Join Applied Intuition as an ML Runtime Optimization Engineer to optimize ML models for embedded environments in a collaborative team.
Join Tavus as a Research Engineer to optimize cutting-edge multimodal AI models for production readiness.
Lead engineers in Gdańsk to develop tools for PyTorch and Triton, blending technical ownership with people leadership.
Join Lilt as a Machine Learning Engineer to build a real-time speech translation backend for a cutting-edge live translation product.
Join SpaceX as a Software Engineer to develop high-performance AI inference systems for mission-critical applications.
Join Fireworks AI as a Software Engineer focused on Performance Optimization to enhance AI infrastructure efficiency and speed.
Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.
Join Mirage as a Research Engineer to build and scale systems for video generation models in a collaborative AI-driven environment.
Join Reflection AI as a Research Software Engineer to bridge research and production in cutting-edge AI training systems.
Join Fundamental as an MLOps Engineer to tackle technical challenges in AI and transform enterprise decision-making.
Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.