Direct from source · No middlemen

Triton In Jobs

82 open positions · Updated 1 day ago

Average salary (USD/year): 182.8k–303.3k/yr

Showing 20 of 82 positions

Search with filters →
Intel
Intel Toronto, Canada Published 1 month ago
Intel
Intel Virtual Canada Published 1 month ago
Graphcore

Join Graphcore's Triton team to enhance AI frameworks and optimize software for cutting-edge machine learning workloads.

Graphcore Gdańsk, Pomeranian Voivodeship, Poland Published 6 months ago
Flexible on stack
Graphcore

Join Graphcore's Triton team to enhance AI frameworks and improve performance on cutting-edge hardware.

Graphcore Bristol, UK; Gdańsk, Pomeranian Voivodeship, Poland Published 1 month ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Member of Technical Staff to design and build systems infrastructure for AI workloads at scale.

Fireworks AI San Mateo Published 3 weeks ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.

Perplexity AI London Published 5 months ago
Flexible on stack
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact Singapore S$200k–S$400k/yr Published 3 months ago
Flexible on stack
Fundamental

Join Fundamental as a Senior Applied Research Engineer to tackle technical challenges in AI model development for enterprise decision-making.

Fundamental Barcelona Published 6 months ago
Flexible on stack
Hippocratic AI

Join Hippocratic AI as a Staff Site Reliability Engineer to manage a fleet of GPU-backed models and enhance healthcare AI systems.

Hippocratic AI Menlo Park, CA Published 2 weeks ago
Flexible on stack
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $92k–$135k/yr Published 11 months ago
Together AI

Join Together AI as a Systems Research Engineer Intern to optimize GPU-accelerated algorithms for ML/AI applications.

Together AI San Francisco $58–$70/hr Published 2 weeks ago
Flexible on stack
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 3 months ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Engineer to lead the technical direction of the Inference Runtime for AI systems serving millions of users.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $405k–$485k/yr Published 3 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.

Inferact Singapore S$200k–S$400k/yr Published 3 months ago
Flexible on stack
Triomics

Join Triomics as an MLOps & Data Engineer to build infrastructure for ML workflows in oncology, impacting patient outcomes.

Triomics India Office Published 3 months ago
Flexible on stack
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 month ago
Page 1 of 5 Next →