The work
You will review conversations between users and AI models, then assess whether the model remains neutral, supportive and appropriately bounded. Day-to-day tasks include flagging sycophancy, taking sides, moralizing, reinforcement of distorted or unfounded beliefs, and responses that drift into clinical or directive advice. You may also draft reference responses, develop evaluation rubrics and design non-crisis test conversations involving relationships, family dynamics, emotional wellbeing, spirituality and unconventional beliefs.
What the screening emphasizes
The platform is likely to verify at least three years of relevant professional experience and probe how you apply behavioral health principles to ambiguous examples. Strong candidates can distinguish validation from agreement, explain when a response oversteps its role, and ground judgments in concepts such as cognitive distortions, healthy boundaries, client-centered practice and motivational interviewing. Clinical licensure is valued but not required, while experience in AI safety, applied ethics, trust and safety, content policy, relationship counseling, spiritual care, misinformation research or AI evaluation is helpful.
Schedule and collaboration
The minimum commitment is 20 hours per week on weekdays, with the option to increase to as many as 40 hours per week. The work includes collaboration with researchers and other experts to calibrate standards and document decisions consistently. The listing does not specify whether the position is remote or whether most tasks are completed asynchronously, so candidates should confirm location, meeting and scheduling expectations during the process.
Employment
This is a W-2 position with Cincinnatus LLC, which serves as the employer of record and may place the successful candidate with a leading AI lab as part of its extended workforce.
Pay band: $45–70/hr
