"llm based systems" Jobs

658 open tech roles matching “llm based systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, TypeScript. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 658 results

Inferact

Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Vals AI

Join Vals AI as an Evaluations Engineer to evaluate LLM models and contribute to industry-leading benchmarks.

Vals AI San Francisco, United States Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.

Inferact San Francisco Published 4 days ago
Flexible on stack
Inferact

Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
MaintainX

Own the LLMX roadmap as a Staff Product Manager at MaintainX, driving AI platform adoption and quality in a hybrid role.

MaintainX San Francisco Published 3 months ago
MaintainX

Own the LLMX roadmap as a Staff Product Manager at MaintainX, driving AI platform adoption and quality.

MaintainX Toronto Published 1 month ago
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
baseten

Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.

baseten San Francisco Published 8 months ago
Flexible on stack
Inferact

Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.

Inferact San Francisco Published 1 month ago
Vals AI

Join Vals AI as a researcher to design and build next-gen AI benchmarks in a fast-paced, innovative environment.

Vals AI San Francisco, United States Published 2 months ago
Flexible on stack
Omnifold

Join Omnifold as an ML Research Engineer to tackle complex supply chain challenges with innovative AI models.

Omnifold San Francisco HQ Published 5 months ago
Vals AI

Join Vals AI as a mid-level engineer to build and maintain a platform for evaluating LLMs at scale in a dynamic startup environment.

Vals AI San Francisco, United States Published 2 months ago
Flexible on stack
Braintrust Data

Join Braintrust as a backend engineer to build infrastructure for cutting-edge AI development tools in a fast-paced environment.

Braintrust Data San Francisco Published 1 month ago
Flexible on stack
Sphere

Lead the development of TRAM, an AI reasoning model for interpreting global trade law in a fast-paced, onsite environment.

Sphere San Francisco HQ Published 11 months ago
baseten

Join Baseten as a Software Engineer focused on ML performance to optimize large language models in a fast-paced startup environment.

baseten San Francisco Published 2 years ago
Flexible on stack