"real time inference" Jobs

535 open tech roles matching “real time inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 535 results

Inworld AI

Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models and AI applications.

Inworld AI Mountain View, California, USA $270k–$500k/yr Published 3 years ago
Anthropic

Join Anthropic as a Staff Engineer to lead the technical direction of the Inference Runtime for AI systems serving millions of users.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $405k–$485k/yr Published 3 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an Embedded AI Engineer to develop on-device intelligence for Android Automotive platforms.

Applied Intuition Sunnyvale Published 5 months ago
Flexible on stack
Databricks

Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI inference workloads.

Databricks San Francisco, California $190k–$265k/yr Published 1 month ago
Anthropic

Join Anthropic as a Staff Software Engineer to optimize and scale AI inference across major cloud platforms.

Anthropic San Francisco, CA $320k–$485k/yr Published 3 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Staff/Principal Software Engineer to develop cutting-edge backend systems for real-time voice models.

Inworld AI Mountain View, California, USA $280k–$350k/yr Published 1 year ago
Flexible on stack 70% coding
Inworld AI

Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models within a fully remote team in Switzerland.

Inworld AI Switzerland Published 7 months ago
Mind Robotics

Lead the development of a robotics runtime platform at Mind Robotics, focusing on real-time performance and middleware architecture.

Mind Robotics Palo Alto Published 4 days ago
Flexible on stack 70% coding
Together AI

Join Together AI as a Staff Software Engineer to build systems that automate infrastructure management for AI clusters.

Together AI India Published 3 weeks ago
Flexible on stack
Anthropic

Join Anthropic's Inference team to design and maintain distributed systems serving AI models to millions globally.

Anthropic New York City, NY; San Francisco, CA | Seattle, WA $320k–$485k/yr Published 2 weeks ago
Flexible on stack
Inworld AI

Lead the development of Inworld’s realtime models and products to empower developers to build consumer-facing AI applications.

Inworld AI Mountain View, California, USA $230k–$400k/yr Published 3 months ago
60% coding
Together AI

Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.

Together AI San Francisco $220k–$280k/yr Published 3 months ago
Flexible on stack 60% coding
Inworld AI

Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models and impact AI applications globally.

Inworld AI UK £140k–£200k/yr Published 5 months ago
Inworld AI

Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and impact AI applications globally.

Inworld AI Germany Published 5 months ago
Inworld AI

Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and AI applications.

Inworld AI Serbia Published 5 months ago
Applied Intuition

Join Applied Intuition as an ML Runtime Optimization Engineer to optimize ML models for embedded environments in a collaborative team.

Applied Intuition Sunnyvale Published 1 year ago
Flexible on stack
Perplexity

Join Perplexity as a technical program manager to drive the core inference platform and coordinate between model providers and engineering teams.

Perplexity San Francisco Published 1 week ago
Inferact

Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
Inworld AI

Join Inworld AI as a Staff/Principal Software Engineer to develop cutting-edge backend systems for real-time voice models.

Inworld AI Vancouver, British Columbia, Canada CA$180k–CA$260k/yr Published 1 year ago
Flexible on stack 70% coding