1x Founding Infrastructure Engineer
VLM Run () · In-Person (Bay Area) ·
- Employment
- Full time
- Category
- Devops
- Experience
- 20+ years
Founding infrastructure engineer at VLM Run, an AI startup whose product is a unified inference gateway serving visual-language models, ViTs and VLAs across fleets of GPUs and clouds. You own and scale the ML serving infrastructure (Ray, Kubernetes, GPUs, serverless) handling hundreds of thousands of requests per day, working fully onsite in the Bay Area.
We’re building the inference platform for visual intelligence. We’re a deeply technical team of AI / computer-vision engineers (20+ years combined, MIT/CMU/NC State PhDs) who’ve shipped production ML infra across autonomous driving and LLMs.
We launched the VLM Run Gateway (https://vlm.run/gateway) in September, a unified API that serves open-weight VLMs, embodied VLAs, ViTs, served across a fleet of GPUs and clouds. That's exactly the infrastructure problem this role will own. We're already scaling to serve 100s of thousands of requests per day, so if this sounds exciting to you, read on.
Email us at [email protected] with your GitHub profile, papers, ML projects you’ve recently shipped (100+ GH stars only) - especially with Ray, k8s, GPUs, serverless. No AI text please, keep it short, shorter emails are more likely to get a response. No remote work, must be in the Bay Area (specify in email subject).