Founding Engineer, Agent Systems at Helmguard
- Location
- London, United Kingdom
- Compensation
- Not Disclosed
Join Helmguard.ai as a Founding Engineer to build the agent-native risk infrastructure that world-leading enterprises in finance and healthcare rely on. You will own the agent platform, architecting the scaffolding and reliability systems that turn frontier AI into production-grade systems of action.
Role overview
You will own the agent platform, transforming frontier model calls into production-grade enterprise features for high-stakes risk management. As a founding engineer, you’ll build the orchestration, evals, and reliability infrastructure that allows AI agents to act as peers to domain experts, setting the standard for AI quality and safety at scale.
About Helmguard
Software
We’re building governance and security infrastructure for a world of increasingly capable and widely proliferated agents.
The world is undergoing a fundamental shift: every day more work is being done autonomously by AI. This is true for enterprises but also for the adversarial actors who seek to disrupt their operations through cyber attacks. The only way to keep up is to use AI intelligently to govern internal deployments and defend against offensive threats. HelmGuard will be the platform which businesses use to do both.
We’ve grown to seven-figure revenue within 9 months of product launch, on the back of multi-year contracts with leading enterprises in financial services, regulated technology, and healthcare. Our founders come from Palantir and academic institutions: Oxford, Stanford, and ETH. We’re backed by leading UK and US institutional investors and exceptional angels from Meta, Isomorphic Labs, Palantir, SpaceXAI, and more.
We’re hiring across founding-team roles for people who want outsize impact, the influence over direction and culture that comes only from joining this early, and pre-Series A equity upside.
What you will do
- Architect and build agent scaffolding including tool use, context management, sandboxing, and robust prompt-injection defenses for enterprise-grade security.
- Develop sophisticated evaluation infrastructure for high-stakes outputs, utilizing LLM-as-judge frameworks and regression testing to ensure peer-level correctness.
- Engineer reliability systems including custom retries, circuit breakers, and prompt versioning to turn experimental model outputs into dependable production actions.
Who this is a fit for
- Proven backend engineering experience in TypeScript with at least 1-2 years of shipping production-grade LLM features and multi-step agent orchestration.
- Strong systems thinking regarding asynchronous queues, idempotency, and the ability to curate datasets for evaluating fuzzy, high-stakes compliance policies.
- A self-starter comfortable owning AI quality end-to-end, possessing the technical conviction to say “no” when features don’t meet rigorous safety bars.
Why this role is remarkable
- Join a high-growth startup that achieved seven-figure revenue months after launch, backed by tier-one investors and tech heavyweights from SpaceXAI and Palantir.
- Experience the outsize impact and influence of a founding-team role, with significant pre-Series A equity upside and a culture of radical ownership.
- Work at the extreme frontier of AI, pushing APIs so hard you’ll collaborate with labs like Anthropic to resolve core engine bugs.
How Jack & Jill work together

What happens next?
Jack’s an AI agent for job searching and career coaching. He works for you.
Jill is the AI recruiter working for the company. She recruits from Jack’s network.
If your profile’s a match and Helmguard wants to meet, Jill will make the intro. In the meantime, Jack will send you excellent alternatives.