Direct from source · No middlemen

Triton In Jobs

82 open positions · Updated 4 weeks ago

Average salary (USD/year): 170.7k–248.9k/yr

Showing 20 of 82 positions

Search with filters →
Applied Intuition

Join Applied Intuition as an ML Runtime Optimization Engineer to optimize ML models for embedded environments in a collaborative team.

Applied Intuition Sunnyvale Published 1 year ago
Flexible on stack
Graphcore

Join Graphcore as a PyTorch Engineer to enhance software that connects Graphcore accelerators with leading ML frameworks.

Graphcore Bristol, UK Published 6 months ago
Flexible on stack
Dyno Therapeutics

Join Dyno Therapeutics as a Sr. Machine Learning Engineer to build and scale AI tools for genetic medicine.

Dyno Therapeutics Remote; Watertown, Massachusetts, United States $178.7k–$213.7k/yr Published 1 month ago
Flexible on stack
Fundamental

Lead a team of MLOps engineers at an AI company transforming enterprise decision-making.

Fundamental Europe Published 3 months ago
Flexible on stack
Genmo

Join Genmo as a GPU Performance Engineer to optimize video generation models and achieve significant performance improvements.

Genmo San Francisco HQ Published 1 year ago
Flexible on stack
Together AI

Join Together AI as a Systems Research Engineer Intern to optimize GPU programming for ML/AI applications in a collaborative environment.

Together AI San Francisco $58–$70/hr Published 2 weeks ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 8 months ago
Flexible on stack
Inferact

Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.

Inferact San Francisco Published 2 months ago
Reflection AI

Join Reflection AI as a Research Software Engineer to bridge research and production in cutting-edge AI training systems.

Reflection AI New York, NY Published 6 months ago
Flexible on stack
Sarvam AI

Join Sarvam AI as a Senior Performance Engineer to optimize GPU kernels for high-performance ML systems.

Sarvam AI Bengaluru Published 1 month ago
Mirage

Join Mirage as a Research Engineer to build and scale systems for video generation models in a collaborative AI-driven environment.

Mirage Union Square, New York City Published 4 days ago
Flexible on stack
baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
Fireworks AI

Join Fireworks AI as a Software Engineer focused on Performance Optimization to enhance AI infrastructure efficiency and speed.

Fireworks AI San Mateo Published 1 year ago
Flexible on stack
SpaceX

Join SpaceX as a Software Engineer to develop high-performance AI inference systems for mission-critical applications.

SpaceX Palo Alto, CA $135k–$175k/yr Published 1 month ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an AI Performance Engineer to optimize large-scale machine learning workloads in a collaborative environment.

Applied Intuition Sunnyvale Published 1 month ago
Flexible on stack
Inferact

Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.

Inferact San Francisco Published 2 weeks ago
Flexible on stack
Graphcore

Lead engineers in Gdańsk to develop tools for PyTorch and Triton, blending technical ownership with people leadership.

Graphcore Gdańsk, Pomeranian Voivodeship, Poland PLN 350.7k–PLN 474.4k/yr Published 1 month ago
Flexible on stack Heavy meetings