"nvidia triton" Jobs

24 open tech roles matching “nvidia triton”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, CUDA, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 24 results

baseten

Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack 70% coding
Perplexity AI

Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.

Perplexity AI London Published 5 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Member of Technical Staff to design and build systems infrastructure for AI workloads at scale.

Fireworks AI San Mateo Published 4 days ago
Flexible on stack
Coreweave

Join CoreWeave as a Senior Engineer to optimize GPU kernels for high-performance AI applications in a rapidly growing environment.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 1 month ago
70% coding
Fireworks AI

Join Fireworks AI as a Software Engineer focused on Performance Optimization to enhance AI infrastructure efficiency and speed.

Fireworks AI San Mateo Published 1 year ago
Flexible on stack
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 week ago
Inferact

Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.

Inferact Remote Published 7 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an AI Performance Engineer to optimize large-scale machine learning workloads in a collaborative environment.

Applied Intuition Sunnyvale Published 1 month ago
Flexible on stack
Armada

Join Armada as an AI Engineer to build and deploy intelligent systems at the edge of the physical world.

Armada Bellevue Office, Sunset Corporate Campus $154.6k–$193.2k/yr Published 1 month ago
Flexible on stack 70% coding
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $206k–$333k/yr Published 8 months ago
Dialpad

Join Dialpad as a Senior Software Engineer to build and improve the AI/ML inference platform for enterprise-scale applications.

Dialpad Buenos Aires, Argentina Published 5 days ago
Flexible on stack 70% coding
Dialpad

Join Dialpad as a Software Engineer to build and improve ML inference systems for AI models at scale.

Dialpad Buenos Aires, Argentina Published 2 months ago
Flexible on stack 70% coding
Applied Intuition

Join Applied Intuition as a Senior Software Engineer to design and implement ML infrastructure for deep learning model training.

Applied Intuition Sunnyvale $215k–$285k/yr Published 3 years ago
Flexible on stack
Armada

Join Armada as a Senior Software Engineer to design and build scalable backend services and deploy machine learning models in production.

Armada Bangalore Office, AEDGE AICC India Pvt Ltd Published 8 months ago
70% coding