Founding Senior AI Infrastructure Engineer
Mission.dev · Quebec, Canada ·
- Seniority
- Senior
- Category
- Devops
- Experience
- 5+ years
Mission.dev · Quebec, Canada ·
Artificial.Agency · Edmonton
General Mills · Pune, MH
Zoom · Seattle (WA)
SpaceX · Redmond, WA
Founding senior engineer reporting to the CTO who designs and operates large-scale multi-cloud and on-prem infrastructure for AI model serving — building with Kubernetes, Terraform, Helm, Ansible, Ray, and Python, benchmarking inference workloads, and running monitoring. Remote-friendly with office visits 2-3x a quarter in Montreal; founding-engineer equity offered.
Nice to have skills: Terraform, Helm, Ansible, Ray
An AI infrastructure company specializing in inference optimization for large-scale models. The platform enhances token throughput and reduces latency by optimizing hardware utilization and decoding processes within a customer's cloud environment. The technology focuses on improving the efficiency of model serving without requiring changes to existing infrastructure or code, ensuring data privacy while reducing operational costs for organizations deploying open-source models.
As a founding Senior AI Infrastructure Engineer, you will report to the Chief Technology Officer to design and operate large-scale infrastructure for AI workloads on public cloud and on-premises environments. You will be an individual contributor with significant influence, combining software engineering with deep systems expertise to build secure and reliable platforms. Your work will focus on enabling efficient model serving at scale, ensuring the infrastructure can support a massive number of concurrent users while maintaining high service quality and performance.
About your mission
You’ll join a cutting edge team that is building next-generation infrastructure solutions to help organizations scale AI more efficiently. As AI workloads continue to grow, the team is focused on optimizing performance, reducing compute costs, and improving energy efficiency across the AI stack. Founded by experienced AI researchers and entrepreneurs, the company works across hardware, software, and model-serving technologies to deliver faster, more scalable AI systems.
This is an exciting chance to join a highly technical team working at the forefront of AI infrastructure innovation. It's an opportunity of a lifetime!
An early-stage AI infrastructure startup specializing in inference optimization and hardware efficiency. The company develops technology that monitors power and latency at the kernel level to maximize compute per watt. This enables high-throughput model serving within existing power constraints without necessitating software rewrites. By optimizing across the hardware and software stack, the platform significantly reduces the energy intensity and cost of scaling sophisticated AI models.
Propeller · Sydney, Sydney Region