A full-stack MLOps engineer owns a production building-optimisation platform end to end: a React/TypeScript web app, a large Postgres telemetry estate, serverless functions and MQTT ingestion from building management systems, plus an LLM stack running LoRA fine-tuning and vLLM serving on GPU hardware. The role suits a small, client-facing London team working on heat networks and the energy transit
Increase your chances of an interview by reading the following overview of this role before making an application.
Job Title: MLOps Platform Developer
Location: London
Salary: Depending on experience
Job Type: Full time, Permanent
This is a rare opportunity to be first in the door to continue the development of our engineering platform.
You will own a production system end to end the web application, the database estate, the data pipelines, the deployments, the integrations and the LLM serving stack.
Whats involved?
Genuine full-stack breadth in one codebase: a React/TypeScript application, a Postgres estate with 100-million-row telemetry tables, serverless functions, MQTT ingestion from building management systems.
LLM model training in the specialist fields our clients value most. It could be a housing estates network or a shopping centre, each LLM is tailored to their exact use case.
Neural LLM-assisted engineering at the sharp end: coding agents are a first-class part of how the platform is built, and you will get very good at directing them, verifying them and knowing where they lie.
Operational technology: BACnet, Trend, JACE and Tridium controls, data points all working together to optimise and commission real engineering systems.
The energy transition, concretely: heat networks are central to UK decarbonisation and growing fast; the domain knowledge you build here is scarce and compounding. xwwtmva
Small-company reality: direct access to clients, sites, commercial decisions and the CEO, we are a bureaucracy free organisation with a. clear delivery and results focus.
What you will do:
Operate the LLM estate: run LoRA fine-tuning cycles and evaluation gates on our GPU hardware, manage vLLM serving (including multi-adapter deployments) alongside production, promote or roll back model versions on the gate results, and keep the serving w
Please click on the apply button to read the full job description