Research \ Anthropic
Anthropic
San Francisco, CAOn-siteFull-time$140k – $220k / year
About this role
Research \ Anthropic
Research
Our research teams investigate the safety, inner workings, and societal impacts of AI models—so that artificial intelligence has a positive impact as it becomes increasingly capable.
Alignment
The Alignment team works to understand the risks of AI models and develop ways to ensure that future ones remain helpful, honest, and harmless.
Economic Research
The Economic Research team studies how AI is reshaping the economy, including work, productivity, and economic opportunity.
Frontier Red Team
The Frontier Red Team analyzes the implications of frontier AI models for cybersecurity, biosecurity, and autonomous systems.
Interpretability
The mission of the Interpretability team is to understand how large language models work internally, as a foundation for AI safety and positive outcomes.
Societal Impacts
Working closely with the Anthropic Policy and Safeguards teams, Societal Impacts is a technical research team that explores how AI is used in the real world.
A global workspace in language models
InterpretabilityJul 6, 2026New interpretability research reveals an emergent mental workspace in Claude that holds internal thoughts that don’t appear in the model’s output.
Anthropic Economic Index report: Cadences
Economic ResearchJun 26, 2026In our latest Economic Index report, we sample hourly for the first time to ask: When do people come to Claude? What do they produce with it? And how do they perceive AI's impact on their work?
Teaching Claude why
AlignmentMay 8, 2026New research on how we've reduced agentic misalignment.
Project Deal
ResearchApr 24, 2026We created a marketplace for employees in our San Francisco office, with one big twist. We tasked Claude with buying, selling and negotiating on our colleagues’ behalf.
What 81,000 people want from AI
Societal ImpactsMar 18, 2026We invited Claude.ai users to share how they use AI, what they dream it could make possible, and wh