This role is with one of Dex's trusted partner companies. We work closely with their teams to truly understand their culture, goals, and what they're looking for, so we can match you with the right opportunity and give you context about the role before you commit to a process.
If you're interested sign up to Dex to apply.
Dex is an AI recruiter agent that helps you run your job search. Tell Dex your stack, seniority, and what you want to build. We will manage your applications and surface other opportunities that are a fit.
The role
This company builds foundation models for extreme physics: the regimes behind semiconductors, aerospace, defence, and fusion energy. Their work accelerates progress in fields where existing simulation tools are too slow or brittle, enabling designs traditional workflows can't reach. They're backed by leading investors and tackling problems with global impact.
You'll be the first dedicated engineering owner for the ML stack, joining a small, highly technical team of researchers. This isn't a narrow systems role, nor is it pure research; you'll turn prototypes into robust code, integrate research branches, and build the in-house tooling for experiment tracking and hyperparameter optimisation. You'll also add distributed training capabilities and own the backend platform delivering these models to customers, setting the engineering culture from day one.
The work
- Own the entire ML stack: model code, training and evaluation workflows, and experiment infrastructure.
- Build and implement distributed training capabilities for large-scale model development.
- Integrate independently developed research branches into a coherent, production-ready codebase.
- Profile and resolve real bottlenecks in training stability and performance, improving system efficiency.
- Design and build the backend platform that delivers these foundation models to customers.
What You Bring
- You are a senior, hands-on engineer who still writes and ships critical code.
- Strong practical experience with PyTorch across model code, data pipelines, and training loops.
- Proven track record with distributed training (multi-GPU/node, GPU clusters), understanding associated memory and communication challenges.
- Real model-engineering experience in physics/simulation, vision, or LLM systems, with a focus on data-driven system improvement.
- Solid Python platform and backend foundations: API design, workflow orchestration, and practical Docker/Kubernetes/IaC.
Why apply through Dex
This is a rare, high-impact founding role at an early-stage company, often hard to find or apply to directly. Dex helps you cut through the noise: sign up once, get matched to this role and others like it, and skip the cold application process. We'll brief you properly on the company and team before you even interview, ensuring a strong fit from the start.
If you're interested, sign up to Dex to apply - https://jobs.meetdex.ai/jobs/1ff6e4dc-fd4b-4d80-b56b-763a01a9d31b
As part of the recruitment process at Dex, we process your personal data in accordance with our Privacy Notice for Job Applicants. This notice explains how and why your data is collected and used, and how you can contact us if you have any concerns.