SGLang Jobs - August 2026
Search by company, role, stack, location, salary signal, source, and work setup.
-
Preferred Networks
LLM Inference Optimization Engineer
Posted 1 month, 2 weeks ago
Preferred Networks is an AI company based in Tokyo working across the stack, from AI chips and computing infrastructure to LLMs and products. Preferred Networks is designing in-house chips (MN-Core series) and training LLMs (PLaMo series). We are actively hir…
Tech stack
Location
Tokyo, Remote in Japan
Work setup
full-time · Tokyo or Remote in Japan. Both roles require relocation to Japan and visa/relocation support is available.
-
Ensoul
Member of Technical Staff
Posted 2 months, 2 weeks ago
Ensoul's mission is to accelerate robotics research. We're a team of frontier-lab scientists and roboticists working to accelerate robotics research and unlock The Great Robotics Buildout. Our customer is the robotics researcher.
Tech stack
Location
San Francisco, CA
Work setup
full-time · Onsite (San Francisco, CA).
Compensation
$180,000-$300,000 + Equity + Benefits
-
VLM Run
Product ML Staff Engineer
Posted 4 months, 2 weeks ago
Building the inference and orchestration layer for production Vision-Language Models. Focus on fast and ergonomic visual inference, reliable structured outputs, and observability for iteration. Shipped projects include Orion (visual agent for images/video/doc…
-
NVIDIA
Engineering Manager
Posted 10 months, 2 weeks ago
NVIDIA | vLLM + SGLang | Deep Learning Inference | Remote (North America preferred) Hi everyone — I’m Akbar, Senior Manager of Deep Learning Inference Software at NVIDIA. I lead our engineering efforts around vLLM and SGLang, two of the most widely used open-…
Roles
Tech stack
Location
Remote (North America preferred), Santa Clara, CA