"inference" Jobs

502 open tech roles matching “inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 502 results

Peregrine

Lead the development of AI-powered features for an end-to-end intelligence platform in public safety.

Peregrine San Francisco, CA $225k–$320k/yr Published 7 months ago
Faire

Own the measurement and optimization of long-term value in a hybrid role at Faire, a tech-driven wholesale platform.

Faire New York City, NY; San Francisco, CA $246.5k–$339k/yr Published 2 days ago
Flexible on stack
baseten

Join Baseten as a Cloud Platform Engineer to build scalable infrastructure for deploying machine learning models.

baseten San Francisco Published 11 months ago
Flexible on stack
baseten

Join Baseten as a senior software engineer to develop cutting-edge AI training products and enhance user workflows.

baseten San Francisco Published 7 months ago
Flexible on stack
baseten

Join Baseten as a Strategic Finance Associate to support financial planning and analysis in a fast-scaling AI infrastructure company.

baseten San Francisco Published 3 months ago
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Lyft

Join Lyft as a Data Science Intern to solve diverse mathematical problems in a collaborative environment.

Lyft San Francisco, CA, United States $58–$62/hr Published 2 days ago
Flexible on stack
Reflection AI

Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.

Reflection AI San Francisco, CA Published 5 months ago
Flexible on stack
baseten

Join Baseten as a Software Engineer focusing on Model APIs to enhance AI model performance and developer experience.

baseten San Francisco Published 11 months ago
Mercury

Build and operate real-time inference services for risk decisioning in a fast-growing fintech startup.

Mercury San Francisco, CA, New York, NY, Portland, OR, or Remote within Canada or United States $166.6k–$208.3k/yr Published 4 days ago
Flexible on stack
Inductive Bio

Join Inductive Bio as a software engineer to build AI tools that accelerate drug discovery.

Inductive Bio New York City, San Francisco, or Boston Published 5 months ago
baseten

Join Baseten as a Marketing Analytics Manager to shape data-driven marketing strategies in a rapidly growing AI company.

baseten San Francisco Published 1 week ago
Flexible on stack
baseten

Join Baseten as a Software Engineer on the Observability team to enhance the reliability of AI product systems.

baseten San Francisco Published 1 month ago
Flexible on stack
baseten

Join Baseten as a Post-Training Research Engineer to build in-house tooling for efficient and high-quality machine learning models.

baseten San Francisco Published 5 months ago
baseten

Join Baseten as an AI Engineer to design and automate workflows for AI-powered capacity management in a fast-growing company.

baseten San Francisco Published 3 days ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Strategic Finance Lead to optimize GPU compute investments and drive capacity decisions.

Perplexity AI San Francisco Published 1 week ago
baseten

Join Baseten as a Forward Deployed Engineer to solve complex AI challenges for leading companies.

baseten San Francisco Published 3 weeks ago
Flexible on stack
Perplexity

Join Perplexity as a Strategic Finance Lead to optimize GPU compute investments and drive capacity decisions in a fast-paced environment.

Perplexity San Francisco Published 1 week ago
baseten

Join the Base Labs Fellowship to conduct cutting-edge AI research with mentorship and funding in San Francisco.

baseten San Francisco $15k–$15k/mo Published 2 months ago
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack