"inference infrastructure" Jobs

912 open tech roles matching “inference infrastructure”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 912 results

Cartesia

Join Cartesia as a Software Engineer to shape data infrastructure for cutting-edge AI models in a collaborative, in-office environment.

Cartesia *HQ - San Francisco, CA Published 2 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as a Senior Software Engineer to design and implement ML infrastructure for deep learning model training.

Applied Intuition Sunnyvale $215k–$285k/yr Published 3 years ago
Flexible on stack
Omnifold

Join Omnifold's Infrastructure Team to build robust systems for AI model training and deployment in a fast-paced environment.

Omnifold San Francisco HQ Published 6 months ago
Flexible on stack
Inferact

Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Applied Intuition

Design and implement software and machine learning components for behavior prediction and environmental interactions in a dynamic environment.

Applied Intuition Sunnyvale Published 1 month ago
Flexible on stack
Descript

Join Descript as a Senior Software Engineer to own and enhance our infrastructure platform, impacting AI model training and deployment.

Descript San Francisco, CA or Remote, US $220k–$292k/yr Published 1 week ago
Flexible on stack
Harvey AI

Lead a high-performing team to develop and manage the model infrastructure platform at Harvey AI.

Harvey AI San Francisco $272k–$355k/yr Published 1 month ago
Heavy meetings
Inferact

Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
Mirendil

Own the compute and cloud foundation for frontier AI research at a tech-first startup.

Mirendil San Francisco $300k–$400k/yr Published 2 months ago
AI-first team
Patreon

Join Patreon as a Senior Machine Learning Engineer to architect and maintain high-throughput ML infrastructure for creator discovery.

Patreon New York Published 3 weeks ago
Flexible on stack
Chai Discovery

Join Chai Discovery as a Software Engineer to optimize AI models for drug discovery in a fast-paced, innovative environment.

Chai Discovery San Francisco office Published 9 months ago
Harvey AI

Lead the design and development of systems powering AI requests at Harvey, collaborating with multiple teams to ensure reliability and efficiency.

Harvey AI San Francisco $236k–$290k/yr Published 1 month ago
Flexible on stack
baseten

Join Baseten as a Cloud Platform Engineer to build scalable infrastructure for deploying machine learning models.

baseten San Francisco Published 11 months ago
Flexible on stack
Coreweave

Lead a team of engineers to build and operate CoreWeave's next-generation Kubernetes-native inference platform.

Coreweave Bellevue, WA - US $188k–$303k/yr Published 8 months ago
Heavy meetings
Volta

Lead product strategy for inference infrastructure and token-serving capabilities in a rapidly growing AI infrastructure company.

Volta Palo Alto, CA Published 1 month ago
SpaceX

Join SpaceX as a Software Engineer to develop high-performance AI inference systems for mission-critical applications.

SpaceX Palo Alto, CA $135k–$175k/yr Published 3 weeks ago
Flexible on stack
Bretton AI

Own and evolve Kubernetes infrastructure while building secure, compliant AI systems for major financial institutions.

Bretton AI San Francisco, CA $168k–$213k/yr Published 7 months ago
70% coding