"triton inference server" Jobs

20 open tech roles matching “triton inference server”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 20 results

Perplexity AI

Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.

Perplexity AI London Published 5 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an AI Performance Engineer to optimize large-scale machine learning workloads in a collaborative environment.

Applied Intuition Sunnyvale Published 1 month ago
Flexible on stack
Applied Intuition

Join Applied Intuition as a Senior Software Engineer to design and implement ML infrastructure for deep learning model training.

Applied Intuition Sunnyvale $215k–$285k/yr Published 3 years ago
Flexible on stack
Cantina

Join Cantina as an MLOps Engineer to build and scale inference infrastructure for generative audio models.

Cantina Remote (U.S. or Europe) $125k–$165k/yr Published 1 month ago
Flexible on stack
baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
Abridge

Join Abridge as a Machine Learning Infrastructure Engineer to optimize AI model inference infrastructure in a fast-paced healthcare startup.

Abridge SF Office Published 1 year ago
Flexible on stack
Dialpad

Join Dialpad as a Software Engineer to build and improve ML inference systems for AI models at scale.

Dialpad Buenos Aires, Argentina Published 2 months ago
Flexible on stack 70% coding
Ambient

Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced AI models.

Ambient Redwood City Published 2 months ago
Flexible on stack 70% coding
Ambient

Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced machine learning models.

Ambient Redwood City, United States Published 1 day ago
Flexible on stack 70% coding
Sarvam AI

Join Sarvam as a Backend Engineer to build and maintain production services for a cutting-edge AI media platform.

Sarvam AI Bengaluru Published 3 months ago
70% coding
MongoDB

Join MongoDB as a Software Engineer 3 to enhance Voyage's AI models for diverse deployment environments.

MongoDB Sydney Published 1 month ago
Flexible on stack
Armada

Join Armada as an AI Engineer to build and deploy intelligent systems at the edge of the physical world.

Armada Bellevue Office, Sunset Corporate Campus $154.6k–$193.2k/yr Published 1 month ago
Flexible on stack 70% coding
Cloudflare

Join Cloudflare as a Senior Machine Learning Engineer to optimize and productionize ML models for a global serverless inference platform.

Cloudflare Hybrid Published 2 months ago
Flexible on stack
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $206k–$333k/yr Published 8 months ago
Sarvam AI

Own the model lifecycle for defence and strategic sector deployments as an MLOps Engineer at Sarvam AI.

Sarvam AI Delhi Published 4 months ago
Flexible on stack