Job Description
Sound psychological reasoning depends on careful diagnostic thinking and honest handling of research evidence — the same discipline AI systems need to learn before they can be trusted with mental-health and behavioral content. Chegg is seeking psychology professionals to evaluate AI-generated psychological content, build authoritative reference answers, and stress-test model safety on sensitive mental-health scenarios. Fully remote and asynchronous.
Core Responsibilities
- Evaluate AI-generated responses to psychology questions covering clinical concepts, research methodology, and behavioral theory for accuracy, reasoning quality, and completeness
- Write reference-standard answers grounded in established frameworks (e.g., DSM-5, ICD-11, APA guidelines) with correct current diagnostic criteria and research methodology and clear explanations that AI models learn from
- Identify factual errors, misapplied diagnostic criteria, or flawed research reasoning in AI-produced content, tagging the specific error category
- Compare and rank multiple AI responses to the same problem, explaining which best reflects accurate, evidence-based reasoning
- Test whether AI systems handle sensitive mental-health prompts (e.g., self-harm, crisis situations) safely, refusing or redirecting where appropriate
- Create realistic synthetic case studies and behavioral scenarios used to train and evaluate models
- Contribute to rubric design — defining scoring criteria such as accuracy, clinical soundness, and clarity — used to benchmark AI performance at scale
- Complete asynchronous task batches independently — no scheduled calls or fixed hours required
Key Qualifications
- Master’s degree or PhD in Psychology, Counseling, or a related field (or equivalent recognized qualification); relevant experience is a plus — postgraduate and doctoral graduates welcome
- Strong grounding in psychological theory, research methodology, and current diagnostic frameworks (DSM-5 or ICD-11)
- Familiarity with APA writing style and standard statistical methods used in behavioral research
- Sharp eye for subtle factual, logical, or methodological error — can detect and explain flawed reasoning in plain English
- Self-directed, detail-oriented, and comfortable delivering quality work asynchronously
Nice to Have
- Licensure as a clinical psychologist, counselor, or an equivalent internationally recognized credential
- Proficiency with R (lavaan, lme4) for structural equation modeling and mixed-effects analysis, or Python for computational/behavioral modeling
- Experience developing or validating psychometric assessment instruments
- Prior experience in content review, QA, or AI/ML data annotation projects
Why Chegg
- Fully remote and flexible — no fixed schedule
- Task-based commitment, typically 10–40 hours per week
- Real-world challenges with direct impact on AI accuracy and safety
- Pathway to ongoing projects for high-quality contributors