"inference runtimes" Jobs

34 open tech roles matching “inference runtimes”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 34 results

Sarvam AI

Own Sarvam's production serving path for large distributed models, integrating and optimizing performance across a multi-node stack.

Sarvam AI Bengaluru Published 1 month ago
Coreweave

Lead complex, cross-functional programs for inference platform delivery at a rapidly growing AI cloud company.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA $198k–$264k/yr Published 2 months ago
Etched

Lead the design and development of performance analysis tools for cutting-edge ML accelerator hardware at Etched.

Etched San Jose Published 8 months ago
Flexible on stack
Sarvam AI

Join Sarvam AI as a Senior Performance Engineer to optimize GPU kernels for high-performance ML systems.

Sarvam AI Bengaluru Published 1 month ago
Etched

Build AI systems that autonomously optimize model architectures for production-ready implementations at Etched.

Etched San Jose $150k–$225k/yr Published 2 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Senior Reliability Engineer to ensure dependable AI systems and cloud infrastructure.

Fireworks AI San Mateo Published 1 month ago
Flexible on stack
Sarvam AI

Own Sarvam's Intel surface end-to-end, optimizing AI models for Intel hardware in a fast-moving team focused on India's AI needs.

Sarvam AI Bengaluru Published 3 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as a Software Engineer to optimize application-layer software for embedded systems in autonomous driving.

Applied Intuition Sunnyvale Published 1 year ago
Coreweave

Drive the adoption of AI runtime services at CoreWeave, leveraging your expertise in distributed systems and AI infrastructure.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA $207k–$275k/yr Published 3 months ago
Flexible on stack
Periodic Labs

Join Periodic Labs as an ML Systems Engineer to build and optimize large-scale training and reinforcement learning infrastructure.

Periodic Labs Menlo Park, CA $250k–$350k/yr Published 4 months ago
Flexible on stack
Fundamental

Join Fundamental as a Model Serving Engineer to optimize and scale the NEXUS model for enterprise decision-making.

Fundamental Europe Published 5 months ago
Flexible on stack
Dialpad

Join Dialpad as a Senior Software Engineer to build and improve the AI/ML inference platform for enterprise-scale applications.

Dialpad Buenos Aires, Argentina Published 2 weeks ago
Flexible on stack 70% coding
Coreweave

Lead complex cross-functional programs in performance and benchmarking for CoreWeave's AI/ML Platform Services.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA $177k–$237k/yr Published 2 months ago
Anthropic
Silicon Engineer Hybrid Visa

Join Anthropic as a Silicon Engineer to lead custom silicon development for AI systems in a collaborative environment.

Anthropic San Francisco, CA $320k–$485k/yr Published 1 month ago
Cresta

Join Cresta as a Senior Software Engineer to build real-time AI agent infrastructure and enhance developer experience.

Cresta Canada (Remote) Published 1 year ago
Skydio

Join Skydio as a Senior Autonomy Engineer to enhance deep learning infrastructure for autonomous drones.

Skydio San Mateo, California, United States $170k–$277.5k/yr Published 9 months ago
Flexible on stack
Arena

Join Arena as a Software Engineer to build core infrastructure for AI model evaluation in a fast-paced startup environment.

Arena San Francisco, United States Published 2 days ago
70% coding
Traversal

Join Traversal as a senior AI Engineer to design and operate core systems for AI products in a fast-paced, collaborative environment.

Traversal New York $175k–$300k/yr Published 3 months ago
Flexible on stack