Senior Technical Product Manager - AI Agents, Evals & Reliability — Opportunihub
Job

Senior Technical Product Manager - AI Agents, Evals & Reliability

Bjak · United Kingdom

At a glance

Type
Job
Organisation
Bjak
Location
United Kingdom
Work mode
On-site
Deadline
Rolling / not stated
Posted
24 Aug 2026

About this job

<h3><strong>About A1</strong></h3><p style="min-height:1.5em">There are over 5 billion users using basic applications today such email, notes, tasks that are not AI-native. Our mission is to build a proactive smart assistant for everyday users to bring intelligence to conversations, errands, organising and workflows, with minimal prompting.</p><p style="min-height:1.5em">Our product focuses on achieving high reliability for long-running workflows, persistent context, and real-world task completion. The system must handle multi-step reasoning, interact with external tools, and remain reliable despite non-deterministic model behavior. Our objective is to help users complete tasks daily enjoyable with over ~90%* reduced time.</p><div style="min-height:1.2em;margin-top:0;margin-bottom:0"> </div><h3>Role</h3><p style="min-height:1.5em">This is a deeply technical, hands-on role. Work directly with engineers on system design, evaluation, and trade-offs-defining requirements, but shaping how the system works for global users. You work at the intersection of user needs, model capability, and system constraints, and are responsible for turning AI potential into real, reliable behavior in a real-world application. </p><div style="min-height:1.2em;margin-top:0;margin-bottom:0"> </div><h3><strong>What You'll be Doing</strong></h3><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Research and define end-to-end AI system requirements from capability to behavior to user impact</p></li><li><p style="min-height:1.5em">Translate model capabilities, data constraints, and evaluation results into clear product and system decisions</p></li><li><p style="min-height:1.5em">Make hard trade-offs across quality, latency, cost, reliability, and UX</p></li><li><p style="min-height:1.5em">Work closely with ML, backend, and mobile engineers on system design, evaluation, and iteration</p></li><li><p style="min-height:1.5em">Define and evolve evaluation frameworks across offline metrics, online experiments, and human feedback</p></li><li><p style="min-height:1.5em">Drive execution with clear specs, strong judgment, and disciplined prioritization</p></li><li><p style="min-height:1.5em">Ensure systems ship quickly, safely, and reliably, with strong feedback loops</p></li><li><p style="min-height:1.5em">Own product quality end-to-end - correctness, predictability, and user trust</p></li></ul><div style="min-height:1.2em;margin-top:0;margin-bottom:0"> </div><h3><strong>What You Will Need</strong></h3><h3>Technical foundation</h3><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Strong grounding in computer science fundamentals, including algorithms, data structures, and system design.</p></li><li><p style="min-height:1.5em">Solid understanding of ML fundamentals and how modern AI systems behave in production.</p></li><li><p style="min-height:1.5em">Comfort reading, reviewing, and discussing technical design documents.</p></li></ul><h3>AI &amp; ML experience</h3><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Hands-on exposure to AI-powered products, including LLM-based systems.</p></li><li><p style="min-height:1.5em">Experience working with model evaluation, prompt or pipeline iteration, and feedback loops.</p></li><li><p style="min-height:1.5em">Strong intuition for model limitations, hallucinations, bias, and drift.</p></li></ul><h3>Product leadership</h3><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Significant experience owning complex, technical products end-to-end.</p></li><li><p style="min-height:1.5em">Proven ability to work closely with senior engineers and ML teams.</p></li><li><p style="min-height:1.5em">Strong judgment and decision-making ability in ambiguous, fast-moving environments.</p></li><li><p style="min-height:1.5em">Ability to balance ambition with technical and operational reality.</p></li></ul><div style="min-height:1.2em;margin-top:0;margin-bottom:0"> </div><h2>Nice to have</h2><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Experience shipping AI-heavy consumer products.</p></li><li><p style="min-height:1.5em">Background as an engineer or highly technical product manager.</p></li><li><p style="min-height:1.5em">Experience defining evaluation metrics for ML systems.</p></li><li><p style="min-height:1.5em">Strong intuition for AI UX patterns and failure handling.</p></li><li><p style="min-height:1.5em">Prior experience in zero-to-one product environments.</p></li></ul><div style="min-height:1.2em;margin-top:0;margin-bottom:0"> </div><h3><strong>Outcomes</strong></h3><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Product strategy clearly aligns AI capabilities with user needs and company priorities.</p></li><li><p style="min-height:1.5em">AI features deliver real value, are understandable, predictable, and trusted by users.</p></li><li><p style="min-height:1.5em">Decisions balance quality, speed, cost, and reliability effectively under uncertainty.</p></li><li><p style="min-height:1.5em">Roadmaps and priorities are clear, with fast iteration based on real user feedback.</p></li><li><p style="min-height:1.5em">Teams are aligned, focused, and able to execute on AI product goals with minimal friction.</p></li></ul><div style="min-height:1.2em;margin-top:0;margin-bottom:0"> </div><h3><strong>How We Work</strong></h3><p style="min-height:1.5em">The best products today in the world were built by small, world class teams. We are a high talent density and hands-on team. We make decisions collectively, move at rapid speed, striking a balance between shipping high quality work and learning. </p><p style="min-height:1.5em">Joining our team requires the ability to bring structure, exercise judgment, and execute independently. Our goal is to put in hands of our users a truly magical product.</p><div style="min-height:1.2em;margin-top:0;margin-bottom:0"> </div><h3><strong>Interview process</strong></h3><p style="min-height:1.5em">If there appears to be a fit, we'll reach to schedule 3, but no more than 4 interviews.</p><p style="min-height:1.5em">Applications are evaluated by our technical team members. Interviews will be conducted via virtual meetings and/or onsite.</p><p style="min-height:1.5em">We value transparency and efficiency, so expect a prompt decision. If you've demonstrated the exceptional skills and mindset we're looking for, we'll extend an offer to join us. This isn't just a job offer; it's an invitation to be part of a team that's bringing AI to have practical benefits to billions globally.</p><p>Find <a href="https://www.arbeitnow.co.uk">Jobs in United Kingdom</a> on Arbeitnow</a>

How to apply

  1. 1 Read the full details above and confirm you meet the eligibility criteria.
  2. 2 Prepare your documents — an updated CV, and any cover letter, proposal or certificates required.
  3. 3 Click Apply on official site to complete your application on Bjak’s official page.
  4. 4 Submit as early as possible — many close once filled.
Apply on official site

Sourced from arbeitnow. Always verify details on the official website. Opportunihub never charges you to apply.

Frequently asked questions

How do I apply for Senior Technical Product Manager - AI Agents, Evals & Reliability?

Review the full details and eligibility on this page, prepare your documents, then use the “Apply on official site” button to complete your application on Bjak’s official page.

Is this opportunity remote or location-based?

This opportunity is based in United Kingdom. Check the official listing for any relocation or on-site requirements.

Is Senior Technical Product Manager - AI Agents, Evals & Reliability free to apply for?

Opportunihub lists this Job for free. Legitimate Jobs do not ask for payment to apply — never pay a fee to submit an application.