CF CloudFrame Job Scanner Open Dashboard →
Verified Active Opening

Senior QA Automation Engineer - AI

Anaplan • London, United Kingdom

Job Description

<div class="content-intro"><p>At Anaplan, we are a team of innovators focused on optimizing business decision-making through our leading AI-infused scenario planning and analysis platform so our customers can outpace their competition and the market.</p> <p>What unites Anaplanners across teams and geographies is our collective commitment to our customers’ success and to our Winning Culture.</p> <p style="padding-left: 40px;">Our customers rank among the who’s who in the Fortune 50. Coca-Cola, LinkedIn, Adobe, LVMH and Bayer are just a few of the 2,400+ global companies who rely on our best-in-class platform.</p> <p style="padding-left: 40px;">Our Winning Culture is the engine that drives our teams of innovators. We champion diversity of thought and ideas, we behave like leaders regardless of title, we are committed to achieving ambitious goals, and we love celebrating<em> </em>our wins – big and small.</p> <p>Supported by operating principles of being strategy-led, <a href="https://www.anaplan.com/careers/">values</a>-based and disciplined in execution, you’ll be inspired, connected, developed and rewarded here. Everything that makes you unique is welcome; join us and let’s build what’s next - together!</p></div><p><strong>About The Role</strong></p> <p>We're pioneering a new role focused exclusively on quality assurance for GenAI/Agentic systems. As our Senior SDET for AI, you'll develop testing strategies, evaluation frameworks, and quality metrics specifically designed for LLM-powered applications. This role requires a unique blend of QA expertise, understanding of GenAI behaviour, and automation skills to ensure our AI features are reliable, accurate, and trustworthy.</p> <p><strong>Your Impact</strong></p> <ul> <li><span data-markdown-start-index="152">Design and architect</span><span data-markdown-start-index="174"> comprehensive test plans, evaluation frameworks, and quality metrics designed specifically for LLM-powered applications and agentic workflows.</span></li> <li><span data-markdown-start-index="321">Develop robust, end-to-end automated test suites</span><span data-markdown-start-index="371"> in Python across evaluations, API checks, UI checks, and performance tests.</span></li> <li><span data-markdown-start-index="451">Create automated prompt testing systems</span><span data-markdown-start-index="492"> and regression suites to proactively detect unintended changes in model behaviour.</span></li> <li><span data-markdown-start-index="579">Establish multi-dimensional evaluation criteria</span><span data-markdown-start-index="628"> to measure and track GenAI quality across accuracy, relevance, safety, consistency, and latency.</span></li> <li><span data-markdown-start-index="729">Perform adversarial testing (red-teaming)</span><span data-markdown-start-index="772"> to identify and mitigate model hallucinations, biases, or security vulnerabilities.</span></li> <li><span data-markdown-start-index="860">Develop internal tools and testing frameworks</span><span data-markdown-start-index="907"> that enable software engineering teams to seamlessly validate their own GenAI implementations.</span></li> <li><span data-markdown-start-index="1006">Deploy real-time monitoring and alerting systems</span><span data-markdown-start-index="1056"> to quickly detect and flag quality degradation of AI features in production.</span></li> <li><span data-markdown-start-index="1137">Own the overall quality and testing strategy</span><span data-markdown-start-index="1183"> for your workstream, maintaining high

Job Reference ID: CF-150737 • Posted on CloudFrame Job Scanner