Lead Software Engineer - Agent Safety
Kraken · London · lead
Kraken · London · lead
Kraken powers some of the most innovative global developments in energy.
We’re a technology company focused on creating a smart, sustainable energy system. From optimising renewable generation, creating a more intelligent grid and enabling utilities to provide excellent customer experiences, our operating system for energy is transforming the industry around the world in a way that benefits everyone.
It’s a really exciting time in energy. Help us make a real impact on shaping a better, more sustainable future.
Kraken is the operating system for the energy transition. We help energy companies, utilities, and system operators transform how they operate so the world can move faster towards a zero-carbon future. We build technology that solves real, messy problems at scale - across data, software, and increasingly AI. Our teams move fast, take ownership, and care deeply about impact.
AI is a key investment area for Kraken Technologies as we look to expand our existing capabilities. A crucial part of this is broadening the foundational infrastructure to enable teams across the organisation to use AI effectively to accelerate our mission.
You’ll work in the Agent Safety/Evals team, a new sub-team within AI Foundations. We build the shared platforms, harnesses, and guardrails that enable engineering and product teams to safely, reliably, and deterministically use machine learning and generative AI agents for internal systems and workflows across the business. This is a delivery-focused team that sits at the intersection of engineering and delivery, focusing on empowering our internal users.
We’re hiring a Lead Software Engineer to head our newly formed Agent Safety/Evals team. As AI agents take on more autonomous tasks across Kraken's internal workflows, this role is critical to ensuring they do so securely and predictably.
This is a leadership role focused on constraining our internally facing AI agents to an expected operation space and guaranteeing system reliability. You will define the technical strategy for how we evaluate models, enforce safety guardrails, and govern AI behavior across internal platforms, skills, and harnesses. You’ll work closely with the broader AI Foundations team and engineers across Kraken to ensure that our push for rapid AI adoption in internal tooling never compromises on security, determinism, or quality.
• Lead the technical direction for internal agent safety: Design and implement systems focused on the reliability, security, and determinism of LLMs and autonomous agents, constraining them strictly to expected operational spaces within our internal ecosystems.
• Build robust evaluation frameworks: Develop scalable harnesses and evals tailored to internal workflows and skills. You will be responsible for asking the right types of questions about the quality and reproducibility of our evals, while engineering the systems to measure them robustly.
• Implement guardrails and governance: Create and enforce pre- and post-generation guardrails, managing the overarching governance of AI models operating within internal tools and platforms.
sourced from the original posting ↗ · always verify details there before applying