"real time ml inference" Jobs

126 open tech roles matching “real time ml inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, PyTorch, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 126 results

Decagon

Lead the development of models and algorithms for Decagon's real-time voice agents in a collaborative, onsite environment.

Decagon San Francisco $200k–$400k/yr Published 3 months ago
Flexible on stack 70% coding
baseten

Join Baseten as a foundational member of the GTM team, driving revenue strategy and operations in the AI infrastructure space.

baseten San Francisco Published 2 months ago
Coreweave

Join CoreWeave as a Senior Engineer to optimize GPU kernels for high-performance AI applications in a rapidly growing environment.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 2 months ago
70% coding
Imprint

Deliver analytical projects that influence product decisions and marketing campaigns in a fast-paced startup environment.

Imprint New York City Published 1 month ago
Flexible on stack
Coreweave

Drive the adoption of AI runtime services at CoreWeave, leveraging your expertise in distributed systems and AI infrastructure.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA $207k–$275k/yr Published 2 months ago
Flexible on stack
Magentic

Join Magentic as a Back-end Engineer to build scalable AI-driven backend services for enterprise supply chains.

Magentic London £125k–£140k/yr Published 3 months ago
Flexible on stack
Protege

Join Protege as a Senior Software Engineer to lead the data processing layer for AI training data at scale.

Protege Remote Published 3 months ago
70% coding
AssemblyAI

Join AssemblyAI as a Senior Research Engineer to enhance large-scale distributed training and inference systems in Voice AI.

AssemblyAI Remote - New York $270k–$310k/yr Published 1 week ago
Flexible on stack
baseten

Join Baseten as a Product Manager to shape the developer experience for AI infrastructure and tooling.

baseten San Francisco Published 3 months ago
baseten

Lead relationships and market intelligence across hyperscalers and strategic neoclouds in a senior role at Baseten.

baseten San Francisco Published 4 days ago
Yuno

Join Yuno as a Senior AI Engineer to architect and scale AI-powered systems that redefine payment orchestration.

Yuno Hyderabad Published 5 months ago
Flexible on stack
MaintainX

Lead the technical direction for predictive maintenance and asset intelligence initiatives at MaintainX, focusing on AI and ML.

MaintainX Canada Published 2 weeks ago
Flexible on stack
Orchard Robotics

Join Orchard Robotics as a Senior Machine Learning Engineer to build solutions for AI-powered farming technology.

Orchard Robotics San Francisco Published 6 months ago
Flexible on stack
Instacart

Join Instacart as a Senior Marketing Decision Scientist II to drive data-driven marketing performance and investment decisions.

Instacart United States - Remote $167k–$212k/yr Published 3 months ago
Flexible on stack
Pennylane

Join Pennylane as a Senior Machine Learning Engineer to develop AI solutions for accounting, impacting millions of entrepreneurs in France.

Pennylane All France (remote) Published 6 days ago
Flexible on stack
Sarvam AI

Build and enhance the AI infrastructure platform at Sarvam, focusing on GPU scheduling and multi-tenancy for machine learning workloads.

Sarvam AI Bengaluru Published 2 months ago
70% coding
Anthropic

Lead the Capacity Engineering team at Anthropic, ensuring efficient allocation and utilization of infrastructure resources.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $405k–$485k/yr Published 3 days ago
Flexible on stack Heavy meetings
Bluefish AI

Join Bluefish AI as a Senior Backend Engineer to build scalable data platforms for AI marketing in a hybrid NYC office.

Bluefish AI New York Published 4 months ago
Flexible on stack
baseten

Join Baseten as a Senior Software Engineer to build testing frameworks for mission-critical AI inference systems.

baseten San Francisco Published 1 month ago
Flexible on stack