About this job
<h2><strong>About us</strong></h2><p style="min-height:1.5em"><a target="_blank" rel="noopener noreferrer nofollow" href="https://whitecircle.ai/"><u>White Circle</u></a> is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies – simple natural-language rules that define what an AI model should and shouldn’t do. We automatically test, enforce, and continuously improve these policies at scale.</p><ul style="min-height:1.5em"><li><p style="min-height:1.5em">We’ve raised $11M from top funds, founders, and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, Datadog, Sentry, and others</p></li><li><p style="min-height:1.5em">We process over 100M+ API calls every month</p></li><li><p style="min-height:1.5em">We fine-tune and train our own LLMs so they run faster and cheaper than any open or proprietary model</p></li></ul><p style="min-height:1.5em">We’re at a multi-million dollar run rate already having signed customers like Lovable and multiple neobanks. You’ll be joining at the most exciting time - early enough that the equity can be life changing but at a point where demand has been proven.</p><p style="min-height:1.5em"></p><h2><strong>In this role, you will</strong></h2><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Build from scratch and lead the Data Labeling team (hiring, coaching, and performance management)</p></li><li><p style="min-height:1.5em">Define annotation guidelines, quality standards, and evaluation frameworks</p></li><li><p style="min-height:1.5em">Develop quality assurance processes, calibration sessions, and auditing systems</p></li><li><p style="min-height:1.5em">Partner with AI researchers and engineers to translate research objectives into labeling workflows</p></li><li><p style="min-height:1.5em">Prioritise labeling projects based on business and research needs</p></li><li><p style="min-height:1.5em">Monitor operational metrics including quality, consistency, throughput, and cost</p></li><li><p style="min-height:1.5em">Improve annotation tooling, automation, and workflow efficiency</p></li><li><p style="min-height:1.5em">Lead complex AI evaluation projects, including safety, preference ranking, RLHF, policy evaluation, and benchmark creation</p></li><li><p style="min-height:1.5em">Analyse disagreement patterns and edge cases to improve guidelines and model performance</p></li><li><p style="min-height:1.5em">Manage vendor relationships and ensure consistent quality across distributed teams</p></li><li><p style="min-height:1.5em">Build reporting dashboards and communicate operational insights to leadership</p></li><li><p style="min-height:1.5em">Foster a culture of continuous improvement, accountability, and operational excellence</p><p style="min-height:1.5em"></p></li></ul><h2><strong>We're looking for someone who</strong></h2><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Has experience leading data annotation or AI evaluation teams</p></li><li><p style="min-height:1.5em">Has strong operational and people management skills</p></li><li><p style="min-height:1.5em">Understands AI model evaluation, LLM behavior, and modern annotation workflows</p></li><li><p style="min-height:1.5em">Can design scalable processes without sacrificing quality</p></li><li><p style="min-height:1.5em">Communicates clearly across technical and non-technical teams</p></li><li><p style="min-height:1.5em">Thrives in fast-moving startup environments</p><p style="min-height:1.5em"></p></li></ul><h2><strong>You might be a great fit if you</strong></h2><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Have managed annotation programs for LLMs, generative AI, or machine learning</p></li><li><p style="min-height:1.5em">Have experience with RLHF, preference data collection, safety evaluations, or benchmark creation</p></li><li><p style="min-height:1.5em">Have worked in Trust & Safety, AI Safety, Content Moderation, or ML Ops</p></li><li><p style="min-height:1.5em">Have managed distributed or global annotation teams</p></li><li><p style="min-height:1.5em">Have experience with vendor management and outsourcing operations</p><p style="min-height:1.5em"></p></li></ul><h2><strong>Bonus points</strong></h2><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Familiarity with prompt engineering and AI safety policies</p></li><li><p style="min-height:1.5em">SQL, Python, or data analysis experience</p></li><li><p style="min-height:1.5em">Experience building internal annotation platforms or workflow automation</p></li><li><p style="min-height:1.5em">Background in linguistics, cognitive science, machine learning, or data operations</p><p style="min-height:1.5em"></p></li></ul><p style="min-height:1.5em"><strong>Important note</strong></p><p style="min-height:1.5em">This role involves overseeing projects that may include offensive, harmful, violent, sexual, or otherwise disturbing content.<br />You'll be responsible for ensuring reviewers have the tools, guidance, and support necessary to perform this work safely and consistently.</p><p style="min-height:1.5em"></p><h2><strong>Why White Circle</strong></h2><ul style="min-height:1.5em"><li><p style="min-height:1.5em">Competitive salary + equity</p></li><li><p style="min-height:1.5em">Work from Paris (hybrid) with a relocation package available, or work from London (note: we are currently unable to provide relocation support and medical insurance for London-based roles)</p></li><li><p style="min-height:1.5em">Paid time off in line with your local regulations</p></li><li><p style="min-height:1.5em">All the hardware, tools, and services you need</p></li><li><p style="min-height:1.5em">Covered subscriptions for AI agents and IDEs</p></li><li><p style="min-height:1.5em">Team off-sites twice a year: we’ve recently been to the Alps and Saint-Tropez</p><p style="min-height:1.5em"></p></li></ul><h2><strong>How we hire</strong></h2><ol style="min-height:1.5em"><li><p style="min-height:1.5em">Intro call with HR (30 min)</p></li><li><p style="min-height:1.5em">Take-home exercise</p></li><li><p style="min-height:1.5em">Final conversation with our CEO (45 min)</p><p style="min-height:1.5em"></p></li></ol><p style="min-height:1.5em">Please submit your application in English.</p><p>Find <a href="https://www.arbeitnow.fr">Jobs in France</a> on Arbeitnow</a>