VR
VLM Run is hiring a
Founding Infrastructure Engineer
About VLM Run
VLM Run is a first-of-its-kind API dedicated to running Vision Language Models on Documents, Images, and Video. We’re building a stack from the bottom-up for ‘Visual’ applications of language models that we believe will make up > 90% of inference needs in the next 5 years.
Job Description
As a founding infrastructure engineer, you will own the infrastructure for the inference platform for visual intelligence. You will work on production-grade ML infrastructure across GPUs and multiple clouds, help scale the VLM Run Gateway API that serves open-weight VLMs, embodied VLAs, and ViTs, and collaborate with a veteran team of AI researchers and engineers. The role requires building and maintaining scalable systems with Docker and Kubernetes, and delivering reliable, low-latency infrastructure. Please share your GitHub profile and recent projects (especially with Docker, Kubernetes, and GPUs) when applying.Location
Not Specified
Salary
Not Specified
Benefits
Not Specified
Tech Tags
Senior Role
Date Listed
02 September, 2026 (about 12 hours ago)
Loading...










