"inference runtimes" Jobs

33 open tech roles matching “inference runtimes”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, Go. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 33 results

Sarvam AI

Own Sarvam's production serving path for large distributed models, integrating and optimizing performance across a multi-node stack.

Sarvam AI Bengaluru Published 1 month ago
Coreweave

Lead complex, cross-functional programs for inference platform delivery at a rapidly growing AI cloud company.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA $198k–$264k/yr Published 2 months ago
Sarvam AI

Join Sarvam AI as a Senior Performance Engineer to optimize GPU kernels for high-performance ML systems.

Sarvam AI Bengaluru Published 1 month ago
Fireworks AI

Join Fireworks AI as a Senior Reliability Engineer to ensure dependable AI systems and cloud infrastructure.

Fireworks AI San Mateo Published 4 weeks ago
Flexible on stack
Sarvam AI

Own Sarvam's Intel surface end-to-end, optimizing AI models for Intel hardware in a fast-moving team focused on India's AI needs.

Sarvam AI Bengaluru Published 3 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as a Software Engineer to optimize application-layer software for embedded systems in autonomous driving.

Applied Intuition Sunnyvale Published 1 year ago
Dialpad

Join Dialpad as a Software Engineer to build and improve ML inference systems for AI models at scale.

Dialpad Buenos Aires, Argentina Published 2 months ago
Flexible on stack 70% coding
Coreweave

Drive the adoption of AI runtime services at CoreWeave, leveraging your expertise in distributed systems and AI infrastructure.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA $207k–$275k/yr Published 2 months ago
Flexible on stack
Periodic Labs

Join Periodic Labs as an ML Systems Engineer to build and optimize large-scale training and reinforcement learning infrastructure.

Periodic Labs Menlo Park, CA $250k–$350k/yr Published 4 months ago
Flexible on stack
Fundamental

Join Fundamental as a Model Serving Engineer to optimize and scale the NEXUS model for enterprise decision-making.

Fundamental Europe Published 5 months ago
Flexible on stack
Dialpad

Join Dialpad as a Senior Software Engineer to build and improve the AI/ML inference platform for enterprise-scale applications.

Dialpad Buenos Aires, Argentina Published 1 week ago
Flexible on stack 70% coding
Coreweave

Lead complex cross-functional programs in performance and benchmarking for CoreWeave's AI/ML Platform Services.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA $177k–$237k/yr Published 2 months ago
Anthropic
Silicon Engineer Hybrid Visa

Join Anthropic as a Silicon Engineer to lead custom silicon development for AI systems in a collaborative environment.

Anthropic San Francisco, CA $320k–$485k/yr Published 1 month ago
Cresta

Join Cresta as a Senior Software Engineer to build real-time AI agent infrastructure and enhance developer experience.

Cresta Canada (Remote) Published 1 year ago
Skydio

Join Skydio as a Senior Autonomy Engineer to enhance deep learning infrastructure for autonomous drones.

Skydio San Mateo, California, United States $170k–$277.5k/yr Published 9 months ago
Flexible on stack
Traversal

Join Traversal as a senior AI Engineer to design and operate core systems for AI products in a fast-paced, collaborative environment.

Traversal New York $175k–$300k/yr Published 2 months ago
Flexible on stack
Perplexity

Join Perplexity as a Machine Learning Engineer to enhance search quality through innovative ranking solutions.

Perplexity Belgrade Published 1 month ago
Perplexity AI

Join Perplexity AI as a Machine Learning Engineer to enhance search quality through innovative ranking solutions.

Perplexity AI Belgrade Published 1 month ago