TL;DR — which model wins which hiring use case?
- **JD writing:** ChatGPT on default polish; Claude when given a style guide. - **Candidate-summary synthesis:** Claude. Fewer hallucinated credentials (Anthropic model docs). - **Interview-question banks:** Tie. - **Scorecard rubric design:** Claude. Defaults to behavior-anchored rating language; ChatGPT to bias-prone adjective scales. - **Debrief synthesis:** Claude. More disciplined about separating evidence from inference. - **Bias-mitigation prompts:** Claude on refusal posture. - **Pricing:** ChatGPT on per-seat list ($25 vs $30/user/month Team). - **Refusal on protected attributes:** Claude. Both vendors prohibit unlawful discrimination (OpenAI, Anthropic); Anthropic's constitutional-AI training (Claude's Constitution) produces more consistent refusals.
**Hard flag:** Neither model may be the final screener. The moment an AI output decides who advances without meaningful human review, you are operating an AEDT under NYC Local Law 144 — independent bias audit, candidate notice, public posting required. Humans decide.