Full-Time

Research Engineer / Scientist

Alignment Science

Anthropic

Anthropic

5,001-10,000 employees

Develops reliable, interpretable AI systems

Compensation Overview

$350k - $500k/yr

H1B Sponsorship Available

San Francisco, CA, USA

Hybrid

Must reside in the San Francisco Bay Area; hybrid role requiring at least 25% on-site in SF office.

Bachelor's

Category
AI & Machine Learning (2)
,
Required Skills
LLM
Reinforcement Learning

Get referred to Anthropic

See people who can refer or advise you

Requirements
  • Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
  • Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
  • Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
Responsibilities
  • Testing the robustness of safety techniques by training language models to subvert our safety techniques, and seeing how effective they are at subverting our interventions
  • Run multi-agent reinforcement learning experiments to test out techniques like AI Debate
  • Build tooling to efficiently evaluate the effectiveness of novel LLM-generated jailbreaks
  • Write scripts and prompts to efficiently produce evaluation questions to test models’ reasoning abilities in safety-relevant contexts
  • Contribute ideas, figures, and writing to research papers, blog posts, and talks
  • Run experiments that feed into key AI safety efforts at Anthropic, like the design and implementation of our Responsible Scaling Policy
Desired Qualifications
  • Have experience authoring research papers in machine learning, natural language processing, or AI safety
  • Have experience with large language models
  • Have experience with reinforcement learning
  • Have experience with Kubernetes clusters and complex shared codebases
  • Significant software, machine learning, or research engineering experience
  • Some experience contributing to empirical AI research projects
  • Some familiarity with technical AI safety research
  • Preference for fast-moving collaborative projects
  • Care about the impacts of AI

Anthropic focuses on AI research to build reliable, interpretable, and steerable AI systems. Its main product, Claude, is an AI assistant designed to handle tasks at any scale for clients across industries, delivered through deployment and licensing along with specialized AI R&D services. Claude works by combining natural language processing, human feedback, reinforcement learning, and interpretability techniques to produce a capable, controllable AI assistant that can assist with a wide range of tasks. The company differentiates itself from competitors by prioritizing safety, transparency, and controllability—emphasizing reliability, interpretability of model behavior, and user-controlled steerability in its AI systems. Anthropic’s goal is to make AI systems that people can trust and efficiently use to improve operations and decision-making across sectors.

Company Size

5,001-10,000

Company Stage

Series H

Total Funding

$162.8B

Headquarters

San Francisco, California

Founded

2021

Get referred to Anthropic

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Bloomberg said Anthropic's annualized revenue hit $65 billion by July 2026.
  • In April 2026, Anthropic signed AWS for five gigawatts of capacity.
  • August 19, 2026 Canada hiring shows aggressive expansion beyond U.S. compute constraints.

What critics are saying

  • July 20, 2026 judge approved Anthropic's $1.5 billion copyright settlement.
  • Reuters on July 30, 2026 reported Claude models hacked three companies during testing.
  • China's July 8, 2026 backdoor warning and enterprise bans threaten Claude Code sales.

What makes Anthropic unique

  • Claude Code won developer mindshare, driving Anthropic's fastest enterprise adoption in 2026.
  • Dario Amodei's safety-first brand attracts regulated buyers demanding stronger controls than OpenAI.
  • Anthropic is building custom chips and global compute, tightening product-infrastructure integration.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Flexible Work Hours

Paid Vacation

Parental Leave

Hybrid Work Options

Company Equity

Growth & Insights and Company News

Headcount

6 month growth

9%

1 year growth

6%

2 year growth

5%
Bloomberg
Aug 21st, 2026
Anthropic hires Google chip veteran to lead custom semiconductor push

Anthropic PBC has hired Amir Salek, a founder of Google's custom chip programme, as the AI lab prepares to develop its own semiconductors. Salek brings extensive experience from Alphabet's hardware efforts, where he played a key role in establishing the company's custom chip initiative. His appointment signals Anthropic's strategic shift towards hardware development. The move follows a broader industry trend of AI companies designing proprietary chips to optimise performance and reduce reliance on third-party suppliers. Custom semiconductors can offer significant advantages in power efficiency and computational capabilities for AI workloads. Anthropic joins other major AI labs in pursuing vertical integration through hardware development, a strategy that could provide greater control over its technology stack and long-term competitive positioning.

Yahoo Finance
Aug 21st, 2026
Anthropic lets enterprises store AI data on own infrastructure amid cost pressures

Anthropic will allow enterprise customers to store required data on their own cloud infrastructure, Reuters reported, citing an unnamed source. The AI company maintains its 30-day data retention requirement but now permits businesses to keep that data on their systems rather than Anthropic's. The policy applies to Anthropic's Fable and Mythos models and future frontier models. Anthropic introduced the 30-day retention requirement in June to protect against cyberattacks using its technology. The move follows mounting pressure on enterprise AI spending, with some companies switching to cheaper alternatives like DeepSeek. AI startup Lindy recently migrated from Anthropic's Claude models to DeepSeek, expecting to save millions of dollars within months. Anthropic filed a confidential IPO prospectus in June. Its annualised revenue run rate reached $47 billion in May, according to CNBC.

Associated Press
Aug 21st, 2026
Medialister joins Anthropic's Claude Connectors Directory for AI-powered media placement planning

Medialister announced its MCP connector is now available in Anthropic's Claude Connectors Directory. The integration allows Claude users to research publications and identify editorial advertising opportunities through natural-language conversations within their existing AI workflow. Communications, SEO, content, and growth teams can use the connector to build media placement shortlists, compare outlets against campaign requirements, and turn briefs into research workflows inside Claude. Use cases include finding publications relevant to specific audiences, geographies, or industries, and planning product launches or PR campaigns. The connector is accessible through Claude's in-product browsing and search experience. Users can add it via Settings → Connectors → Browse Connectors. Whilst the connector handles discovery and planning, campaign booking remains within Medialister's platform. Medialister provides a platform for editorial advertising and media placements, helping brands and agencies discover publication opportunities and plan campaigns at scale.

Crypto Briefing
Aug 21st, 2026
Anthropic commits $35M and $100M in AI credits to defend open-source software

Anthropic has launched Project Glasswing, a cybersecurity initiative backed by $35 million to support open-source security projects. The programme includes up to $100 million in AI model usage credits and $4 million in direct donations to open-source security organisations. The initiative uses Anthropic's unreleased Claude Mythos Preview model, which has identified over 10,000 high and critical vulnerabilities in existing software. Tech giants AWS, Apple, Google, and Microsoft are participating in the consortium. Funding allocations include $2.5 million to Alpha-Omega and OpenSSF, $1.5 million to Apache Software Foundation, and $12.5 million to the Linux Foundation's grant programme. Since launch, approximately 150 organisations have joined the initiative. The programme provides free vulnerability scanning access to open-source foundations that typically operate on limited budgets, addressing a critical security gap in widely used software components.

PR Newswire
Aug 20th, 2026
MIT and Harvard students unveil AI tools to boost economic mobility for disadvantaged families

MIT and Harvard graduate students have developed artificial intelligence tools designed to improve economic outcomes for disadvantaged families through the inaugural AI for Social Impact Fellowship. The programme, supported by NextLadder Ventures, The Bike Shop @ MIT, and Anthropic, created "Navigation Technology" solutions that provide personalised support during high-stakes financial moments. Four frontline organisations are scaling the tools immediately. Money Management International is deploying call intelligence software for debt counsellors. Center for Employment Opportunities is integrating an AI voice coach for job interview preparation. Climb Together is adding features for networking conversation feedback. Neighborhood Trust Financial Partners is implementing technology to analyse bank statements during coaching calls. The fellowship aims to expand the pipeline of developers building Navigation Technology, which at maturity could help millions of Americans navigate economic challenges. NextLadder Ventures has over $1 billion in capital backing the initiative.