Verified from career page · Posted 2mo ago

White Circle

Senior Data Labeler

White Circle · Research

Paris

Senior Still listed after 2 months

DatadogHugging FaceLLM / GenAIOpenAIPythonSQL

Last seen 2d ago

Posted
2mo ago

Posted on 8 July 2026

Workplace
Hybrid

Work model: Hybrid

Salary
Not disclosed

Salary range not shared by the company

Visa sponsorship
Not specified

Visa sponsorship details unknown

TL;DR: We're looking for a Data Labeling Lead to build our annotation team from zero – setting quality standards, designing evaluation frameworks, and running the labeling operations behind our safety evals, RLHF, and benchmarks. You'll own quality, throughput, and cost while working side by side with our AI researchers.

ABOUT US

White Circle https://whitecircle.ai/ is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies – simple natural-language rules that define what an AI model should and shouldn’t do. We automatically test, enforce, and continuously improve these policies at scale.

- We’ve raised $11M from top funds, founders, and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, Datadog, Sentry, and others

- We process over 100M+ API calls every month

- We fine-tune and train our own LLMs so they run faster and cheaper than any open or proprietary model

We’re at a multi-million dollar run rate already having signed customers like Lovable and multiple neobanks. You’ll be joining at the most exciting time - early enough that the equity can be life-changing, but at a point where demand has been proven.

IN THIS ROLE, YOU WILL

- Build from scratch and lead the Data Labeling team (hiring, coaching, and performance management)

- Define annotation guidelines, quality standards, and evaluation frameworks

- Develop quality assurance processes, calibration sessions, and auditing systems

- Partner with AI researchers and engineers to translate research objectives into labeling workflows

- Prioritise labeling projects based on business and research needs

- Monitor operational metrics including quality, consistency, throughput, and cost

- Improve annotation tooling, automation, and workflow efficiency

- Lead complex AI evaluation projects, including safety, preference ranking, RLHF, policy evaluation, and benchmark creation

- Analyse disagreement patterns and edge cases to improve guidelines and model performance

- Manage vendor relationships and ensure consistent quality across distributed teams

- Build reporting dashboards and communicate operational insights to leadership

- Foster a culture of continuous improvement, accountability, and operational excellence

WE'RE LOOKING FOR SOMEONE WHO

- Has experience leading data annotation or AI evaluation teams

- Has strong operational and people management skills

- Understands AI model evaluation, LLM behaviour, and modern annotation workflows

- Can design scalable processes without sacrificing quality

- Communicates clearly across technical and non-technical teams

- Thrives in fast-moving startup environments

YOU MIGHT BE A GREAT FIT IF YOU

- Have managed annotation programs for LLMs, generative AI, or machine learning

- Have experience with RLHF, preference data collection, safety evaluations, or benchmark creation

- Have worked in Trust & Safety, AI Safety, Content Moderation, or ML Ops

- Have managed distributed or global annotation teams

- Have experience with vendor management and outsourcing operations

BONUS POINTS

- Familiarity with prompt engineering and AI safety policies

- SQL, Python, or data analysis experience

- Experience building internal annotation platforms or workflow automation

- Background in linguistics, cognitive science, machine learning, or data operations

Important note

This role involves overseeing projects that may include offensive, harmful, violent, sexual, or otherwise disturbing content.

You'll be responsible for ensuring reviewers have the tools, guidance, and support necessary to perform this work safely and consistently.

COMPENSATION & BENEFITS

- Competitive compensation, including equity

- Flexible time off

- A spacious office in the heart of Paris’s 2nd arrondissement, with a flexible hybrid setup

- Relocation support if you’re moving to Paris, available after your probationary period

- Premium private health insurance

- Mental health support, including coverage for therapy when you need it

- Language lessons to help you improve your English or French

- Lunch and dinner covered when you work from the office

- Learning and development support for courses, conferences, and opportunities to grow your skills

- All the hardware, subscriptions, tools, and services you need

- Team off-sites twice a year: we’ve recently been to the Alps, Saint-Tropez, and Marbella

PROCESS

1. Intro call (25 min)

2. Take-home test assignment

3. Final conversation with our CEO (45 min)

Please submit your application in English.

About White Circle

White Circle builds the guardrail layer that sits in front of a production LLM and decides, in real time, what it is allowed to say and do. $11m seed. The roles are ML infrastructure: post-training, RL and evaluation pipelines rather than product features.

Apply at White Circle

Similar roles

  1. PRO members only Pro Remote
  2. PRO members only Pro Remote
  3. ML Infrastructure Engineer Hybrid White Circle Research Paris Machine learningC++CUDADatadog +7 AI-native Not disclosed Visa: Not specified 2mo ago
  4. Multimodal ML Engineer Hybrid White Circle Research Paris Machine learningDatadogHugging FaceOpenAI +1 AI-native Not disclosed Visa: Not specified 2mo ago
  5. ML Research Engineer Hybrid White Circle Research Paris Machine learningDatadogGitHugging Face +3 AI-native Not disclosed Visa: Not specified 2mo ago
  6. PRO members only Pro Remote
1,845 more roles like this. Open the board →