"inference optimization" Jobs

586 open tech roles matching “inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 586 results

Assembled

Join Assembled as a Software Engineer to develop forecasting and scheduling solutions for customer support operations.

Assembled United States Published 4 months ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA $315k–$560k/yr Published 10 months ago
Goodfire

Join Goodfire as a Machine Learning Engineer to build interpretable AI systems with a world-class team.

Goodfire San Francisco, CA & New York, NY $200k–$400k/yr Published 9 months ago
Udio

Join Udio as a full-stack scientist to lead quantitative research efforts at the intersection of music and AI.

Udio New York City (Remote possible for exceptional candidates) $250k–$350k/yr Published 7 months ago
Flexible on stack
Dialpad

Join Dialpad as a Senior Software Engineer to build and improve the AI/ML inference platform for enterprise-scale applications.

Dialpad Buenos Aires, Argentina Published 5 days ago
Flexible on stack 70% coding
OpenRouter

Join OpenRouter as a Forward Deployed Engineer to help customers implement and scale AI solutions effectively.

OpenRouter San Francisco Bay Area, California Published 3 months ago
Flexible on stack 70% coding
Genesis Molecular AI

Join Genesis Molecular AI as an ML Research Engineer to develop cutting-edge foundation models for drug discovery.

Genesis Molecular AI San Mateo, CA Published 1 year ago
Flexible on stack 70% coding
Applied Intuition

Join Applied Intuition as a Senior Software Engineer to develop perception systems for L4 autonomous trucks.

Applied Intuition Tokyo Published 2 months ago
Together AI

Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.

Together AI San Francisco $220k–$280k/yr Published 3 months ago
Flexible on stack 60% coding
Sarvam AI

Join Sarvam AI as a Senior Performance Engineer to optimize GPU kernels for high-performance ML systems.

Sarvam AI Bengaluru Published 1 month ago
Coreweave

Join CoreWeave as a Senior Engineer to build performance insights and observability systems for AI infrastructure.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 1 month ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Software Engineer to design and build scalable infrastructure for generative AI systems.

Fireworks AI San Mateo Published 10 months ago
Flexible on stack
baseten

Join Baseten as a Software Engineer to lead GPU Networking efforts and optimize distributed systems for AI applications.

baseten San Francisco Published 6 months ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 2 years ago
Instacart

Join Instacart as a Senior Marketing Decision Scientist II to drive data-driven marketing performance and investment decisions.

Instacart United States - Remote $167k–$212k/yr Published 3 months ago
Flexible on stack
Together AI

Design and deliver multi-petabyte storage systems for AI workloads at Together AI, optimizing performance and cost.

Together AI San Francisco $250k–$300k/yr Published 3 months ago
Flexible on stack
Newton Research

Lead the strategy and execution of a measurement platform that drives media investment decisions using advanced analytics.

Newton Research Greater Boston Area $180k–$230k/yr Published 4 months ago
baseten

Join Baseten as a Cloud Platform Engineer to build scalable infrastructure for deploying machine learning models.

baseten San Francisco Published 11 months ago
Flexible on stack