About this role
Phone numbers and emails in this ad are masked until you log in.
auto_translated_note
About usWhite Circle is an AI Safety company building the safety, reliability, and optimization layer for AI systems. At the core of our platform are policies - simple natural-language rules that define what an AI model should and shouldn’t do. We automatically test, enforce, and continuously improve these policies at scale.We’ve raised $11M from top funds, founders, and senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, DeepMind, Datadog, Sentry, and othersWe process over one hundred million API calls every monthWe fine-tune and train our own LLMs so they run faster and cheaper than any open or proprietary modelWe’re a small, highly focused team.
If you want to work deeply on hard problems, see your work ship to production quickly, and influence how AI safety is actually built - you’re the one we need.In this role, you willReview and evaluate AI conversations and model outputsAssess responses for safety, quality, accuracy, policy compliance, and user intentIdentify harmful, unsafe, misleading, or low-quality behaviorLabel and categorize model outputs according to internal evaluation frameworksModerate sensitive content and identify policy violationsCompare, rank, and score model responsesInvestigate edge cases and ambiguous situationsProvide structured feedback to researchers and engineersHelp improve evaluation guidelines and annotation processesContribute to the datasets used to train and evaluate AI systemsWe're looking for someone whoHas exceptional attention to detailCan make consistent decisions across large volumes of dataEnjoys analysing nuanced situations where there isn't always a clear answerCan follow guidelines while exercising good judgmentHas strong written English skillsCommunicates clearly and explains reasoning wellIs curious about AI and how these systems workYou might be a great fit if youHave experience with content moderation, trust & safety, quality assurance, compliance, or policy enforcementHave experience in data annotation, AI evaluation, RLHF, or model assessmentHave worked with AI tools extensively and understand their strengths and limitationsEnjoy finding edge cases and unusual model behaviorImportant noteThis role may involve reviewing content that is offensive, harmful, violent, sexual, or otherwise disturbing. We provide tooling, and support, but candidates should be comfortable working with sensitive content when necessaryWhy White CircleSalary of $30,000 to $50,000 + equityPaid time off in line with your local regulations, no matter where you work fromAll the hardware, tools, and services you needHow we hireIntro call with HR (25 min)Take-home assignmentFinal conversation with our CEO (35 min)Please submit your application in English.Compensation: $30K - $50K • Offers Equity • $30K - $50K • Offers EquityFind Jobs in United Kingdom on Arbeitnow
Community Q&A
Anyone worked here? Ask before you apply.
No threads yet for this job or company.