White Circle is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies – simple natural-language rules that define what an AI model should and shouldn’t do. We automatically test, enforce, and continuously improve these policies at scale.
We’ve raised $11M from top funds, founders, and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, Datadog, Sentry, and others
We process over one hundred million API calls every month
We fine-tune and train our own LLMs so they run faster and cheaper than any open or proprietary model
We’re a small, highly focused team. If you want to work deeply on hard problems, see your work ship to production quickly, and influence how AI safety is actually built – you’re the one we need.
Review and evaluate AI conversations and model outputs
Assess responses for safety, quality, accuracy, policy compliance, and user intent
Identify harmful, unsafe, misleading, or low-quality behavior
Label and categorize model outputs according to internal evaluation frameworks
Moderate sensitive content and identify policy violations
Compare, rank, and score model responses
Investigate edge cases and ambiguous situations
Provide structured feedback to researchers and engineers
Help improve evaluation guidelines and annotation processes
Contribute to the datasets used to train and evaluate AI systems
Has exceptional attention to detail
Can make consistent decisions across large volumes of data
Enjoys analysing nuanced situations where there isn't always a clear answer
Can follow guidelines while exercising good judgment
Has strong written English skills
Communicates clearly and explains reasoning well
Is curious about AI and how these systems work
Opens the employer's application page
Get jobs like this on WhatsApp — instantly
Pro members got this alert the moment it was posted. Be first next time.
Rider Onboarding Operations Representative
Deliveroo · 2h ago
(Senior) B2B Operations Manager (all genders)
Stapelstein · 2h ago
Développeur confirmé IA - H/F
N2JSoft · 8h ago