"distributed llm" Jobs

888 open tech roles matching “distributed llm”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 888 results

Hippocratic AI

Own the serving infrastructure for healthcare AI, optimizing LLM inference systems to enhance patient experiences.

Hippocratic AI Menlo Park, CA Published 2 weeks ago
Flexible on stack
Bluefish AI

Lead the development of LLM-powered products in a fast-moving startup focused on AI-driven marketing solutions.

Bluefish AI London - Remote in the UK Published 9 months ago
Flexible on stack
Inferact

Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.

Inferact Remote Published 7 months ago
Flexible on stack
Datadog
Datadog Boston, Massachusetts, USA; Denver, Colorado, USA; New York, New York, USA; San Francisco, California, USA $151.5k–$222k/yr Published 6 months ago
Inferact

Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.

Inferact San Francisco, California, United States Published 1 day ago
Flexible on stack
Inferact

Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Everlaw

Join Everlaw as a Senior Software Engineer to build AI platform capabilities that enhance legal tech solutions.

Everlaw Oakland, California, United States $173k–$251k/yr Published 4 months ago
Flexible on stack 70% coding
Bluefish AI

Lead the development of LLM-powered products in a flexible hybrid environment at a growing AI marketing startup.

Bluefish AI Berlin, Remote in Germany Published 9 months ago
Flexible on stack
MongoDB

Lead a talented team as a Staff Engineer at MongoDB, focusing on building software frameworks and modernizing applications.

MongoDB Gurugram Published 1 week ago
Flexible on stack
Everlaw

Join Everlaw as a Senior Software Engineer to build user-facing features in a collaborative environment focused on legal tech and AI.

Everlaw Oakland, California, United States $173k–$251k/yr Published 10 months ago
Flexible on stack 70% coding
August

Join August as a Growth Marketing Manager to drive innovative growth strategies in a fast-paced legal tech environment.

August New York City Published 6 months ago
Flexible on stack
Inferact

Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.

Inferact San Francisco Published 1 month ago
Legora

Join Legora as an Endpoint Engineer to manage and secure a modern macOS fleet in a collaborative, high-performance environment.

Legora New York City Published 1 month ago
Flexible on stack 70% coding
baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
ChipAgents

Join ChipAgents as an ML Systems Engineer to optimize LLM inference systems for leading semiconductor companies.

ChipAgents San Jose $150k–$350k/yr Published 3 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Member of Technical Staff to advance generative AI and multimodal systems through foundational research.

Fireworks AI San Mateo Published 3 days ago
Flexible on stack
Databricks

Lead a multidisciplinary research team to advance large-scale machine learning efficiency at Databricks.

Databricks Mountain View, California; San Francisco, California $270k–$340k/yr Published 3 months ago
Flexible on stack 60% coding