This site is optmised for modern browsers - e.g. Google Chrome, Mozilla FireFox or Microsoft Edge.

Contact

Platform Engineer

Cambridge, UK


Monumo is an AI company focused on optimising complex engineering systems. Starting with electric motors, our Anser® AI Engine combines deep domain knowledge with multi-physics simulation, machine learning, and large-scale compute. We evaluate millions of candidate designs a day to find optimal trade-offs across performance, cost, weight, and efficiency.
Electric motors are a proving ground for complex, system-level engineering challenges: they consume over half the world's electricity, and small design gains compound at system level. We're a well-funded deeptech startup based in Cambridge and Coventry, with a core team from Arm, Microsoft Research, DeepMind, and academia.

Why we need you

Behind Monumo's designs are simulation, optimisation, and machine learning workloads running on our own compute cluster and in the cloud. We're now putting that capability directly into the hands of our customers, and your role will be the platform layer in between: the systems that take a job from submission, through scheduling and execution, to results delivered securely and efficiently. Much of this is new, and you'll have a real say in how it's designed, working alongside engineers and scientists from physics, motor engineering, and AI backgrounds.

Your responsibilities will include

  • Backend services that put computational workloads behind customer-facing APIs
  • Batch job lifecycle systems: submission, scheduling, retries, cancellation, priority
  • Identity and access across our services: how people, services, and machines authenticate, and what they're allowed to do, from cloud through to cluster
  • Customer isolation: protecting commercially sensitive designs while running an efficient shared compute platform
  • The data lifecycle around simulation: storing, versioning, and moving large results, and getting them to customers
  • The path from code to running system: packaging, containerisation, CI/CD, and deployment workflows that scientists actually enjoy using
  • Observability: knowing what the platform is doing, both for our own engineering and for the security standards our customers expect
  • Fault-finding across the stack: tracing issues through services, containers, Linux, and the network
  • Close collaboration with frontend and product engineering on the interfaces customers and colleagues use (this is a backend role; UI skills are welcome but not the focus).

What you'll bring

We're open-minded about your technical background. More important than any specific skill is your willingness to learn.

  • Strong software engineering in a modern language. Our code is mostly Python today, with Rust growing where performance and reliability matter
  • Care for code quality: typing, testing, and review
  • Experience designing systems and the APIs between them: services that coordinate work, manage state, and handle failure
  • Curiosity about scientific computing. We write much of our own simulation code; you don't need to know finite element methods, but we want someone keen to understand what the jobs compute, and to help them run faster and scale.

You'll have experience in some of these

  • Cloud computing or HPC job management (e.g. Slurm)
  • Identity and authorisation flows (e.g. OAuth2/OIDC)
  • Deploying containerised services on Linux (e.g. Podman)
  • Infrastructure as code (e.g. Ansible, OpenTofu)
  • Modern Python tooling and packaging (e.g. uv)
  • Data storage and pipelines, including large simulation results and model serving
  • APIs, REST, WebSockets (e.g. FastAPI).

You will thrive in our team

If you're motivated by solving real technical problems through open-minded, interdisciplinary thinking, and by learning new things along the way. We value people who share ideas about how we can do things better, who balance autonomy with collaboration across different backgrounds (physics, engineering, operations), and who enjoy a dynamic environment


Apply now