kraken123 - Lead Software Engineer - Agent Safety
Requirements
• Strong technical leadership: Proven experience leading technical initiatives or teams, capable of setting the technical vision for complex, ambiguous domains. • Deep software engineering fundamentals: Senior/advanced capability in designing secure components end-to-end, testing thoroughly, and reasoning heavily about system design, concurrency, and architecture tradeoffs. (Python preferred). • Expertise in AI Evaluation and Safety: A highly critical thinker who understands the nuances of LLM behavior. You must be comfortable interrogating the quality of AI outputs and deeply experienced in building harnesses to measure it reproducibly. • Security and Governance mindset: Experience with threat modeling, authentication, red teaming, or building guardrails for internal systems, tooling, or platforms. • Cloud experience (AWS): Comfortable running internal services, owning reliability/scalability, and collaborating with platform/techops/security partners. • Excellent communication: Able to synthesize complex safety and eval metrics into actionable insights for both technical and non-technical stakeholders. • ## 🚀 What Success Looks Like • Confidence in Internal Tooling: Engineers across Kraken can deploy new autonomous models, AI skills, and internal agents with total confidence, knowing they are bound by reliable safety guardrails and strict operational constraints within our internal ecosystems. • High-Quality Internal Evals: You have established a culture and infrastructure of robust, highly reproducible evaluation metrics that accurately reflect the real-world performance and safety of our internal-facing AI tooling and harnesses. • Resilient Internal Infrastructure: AI security, monitoring, and incident response are treated as first-class citizens, ensuring that any erratic agent behavior in our internal platforms is immediately caught and mitigated before it affects broader operations. • Technical Leadership: Strong collaboration across AI Foundations helps accelerate secure internal AI adoption; you are actively mentoring engineers and shaping the company's internal AI governance strategy. • Experience with specific LLM evaluation and safety frameworks (e.g., Inspect AI, Ragas, OpenAI Evals, NeMo Guardrails). • Django experience and strong backend engineering patterns (security, performance, maintainability). • Experience with Datadog for complex observability, tracing, and monitoring in AI environments. • Familiarity with foundational AI engineering tooling like Pydantic AI, LiteLLM, or LangChain. • Kraken is a certified Great Place to Work in France, Germany, Spain, Japan, Australia, and USA. In the UK we are one of the Best Workplaces on Glassdoor with a score of 4.5, and in Germany we rate 4.7 on Kununu as a Top Company. Check out our Welcome to the Jungle site (FR/EN) to learn more about our teams and culture. • Are you ready for a career with us? We want to ensure you have all the tools and environment you need to unleash your potential. If you have any specific accommodations or a unique preference, please contact us at [email protected] and we'll do what we can to customise your interview process for comfort and maximum magic!
Apply in one click
Upload My Resume
Drop here or click to browse · Tap to choose · PDF, DOCX, DOC, RTF, TXT