Apr 14, 2026AlignmentAutomated Alignment Researchers: Using large language models to scale scalable oversight
Apr 14, 2026AlignmentAutomated Alignment Researchers: Using large language models to scale scalable oversight
Lire
Apr 14, 2026AlignmentAutomated Alignment Researchers: Using large language models to scale scalable oversight
Apr 9, 2026PolicyTrustworthy agents in practice
InterpretabilityApr 2, 2026Emotion concepts and their function in a large language modelAll modern language models sometimes act like they have emotions. What’s behind these behaviors? Our interpretability team investigates.
Mar 31, 2026Economic ResearchHow Australia Uses Claude: Findings from the Anthropic Economic Index
Mar 24, 2026Economic ResearchAnthropic Economic Index report: Learning curves
Mar 23, 2026ScienceIntroducing our Science Blog