"triton in" Jobs

82 open tech roles matching “triton in”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Triton, CUDA. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 82 results

Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.

Inferact Singapore S$200k–S$400k/yr Published 3 months ago
Flexible on stack
Fundamental

Lead a team of MLOps engineers at an AI company transforming enterprise decision-making.

Fundamental Europe Published 3 months ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Engineer to lead the technical direction of the Inference Runtime for AI systems serving millions of users.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $405k–$485k/yr Published 3 months ago
Flexible on stack
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 3 months ago
Flexible on stack
Triomics

Join Triomics as an MLOps & Data Engineer to build infrastructure for ML workflows in oncology, impacting patient outcomes.

Triomics India Office Published 3 months ago
Flexible on stack
Together AI

Join Together AI as a Systems Research Engineer Intern to optimize GPU programming for ML/AI applications in a collaborative environment.

Together AI San Francisco $58–$70/hr Published 2 weeks ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an ML Runtime Optimization Engineer to optimize ML models for embedded environments in a collaborative team.

Applied Intuition Sunnyvale Published 1 year ago
Flexible on stack
Tavus

Join Tavus as a Research Engineer to optimize cutting-edge multimodal AI models for production readiness.

Tavus Remote Published 6 days ago
Flexible on stack
Graphcore

Lead engineers in Gdańsk to develop tools for PyTorch and Triton, blending technical ownership with people leadership.

Graphcore Gdańsk, Pomeranian Voivodeship, Poland PLN 350.7k–PLN 474.4k/yr Published 1 month ago
Flexible on stack Heavy meetings
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 8 months ago
Flexible on stack
Lilt

Join Lilt as a Machine Learning Engineer to build a real-time speech translation backend for a cutting-edge live translation product.

Lilt Washington, D.C., United States Published 1 day ago
Flexible on stack AI-first team 70% coding
SpaceX

Join SpaceX as a Software Engineer to develop high-performance AI inference systems for mission-critical applications.

SpaceX Palo Alto, CA $135k–$175k/yr Published 1 month ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Software Engineer focused on Performance Optimization to enhance AI infrastructure efficiency and speed.

Fireworks AI San Mateo Published 1 year ago
Flexible on stack
baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
Mirage

Join Mirage as a Research Engineer to build and scale systems for video generation models in a collaborative AI-driven environment.

Mirage Union Square, New York City Published 5 days ago
Flexible on stack
Reflection AI

Join Reflection AI as a Research Software Engineer to bridge research and production in cutting-edge AI training systems.

Reflection AI New York, NY Published 6 months ago
Flexible on stack
Fundamental

Join Fundamental as an MLOps Engineer to tackle technical challenges in AI and transform enterprise decision-making.

Fundamental Europe Published 2 months ago
Flexible on stack
Inferact

Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.

Inferact San Francisco Published 2 months ago