"low precision inference" Jobs
51 open tech roles matching “low precision inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, PyTorch, CUDA. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 51 results
Join Applied Intuition as an AI Performance Engineer to optimize large-scale machine learning workloads in a collaborative environment.
Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.
Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.
Join CoreWeave as a Staff Software Engineer to lead the development of a Kubernetes-native inference platform for AI workloads.
Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.
Join Together AI as a Research Engineer to develop a platform for customizing open-source models with user data.
Join Krea as an ML Researcher to finetune diffusion models and enhance AI creative tools in a collaborative environment.
Join World Labs as a Performance Engineer to optimize AI models for speed and efficiency in a cutting-edge research environment.
Join Fireworks AI as a Software Engineer focused on Performance Optimization to enhance AI infrastructure efficiency and speed.
Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.
Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.