A

Anthropic

Builds Claude AI models with a focus on safety

Anthropic Fellows Program Fellow - AI Safety & Security

Winter 2027Updated on 10/7/2026Deadline 10/18/26
$96.25/hr
Internship
Remote in USA+3 moreMore locations: London, UK | San Francisco, CA, USA | Ontario, Canada
HybridRemote fellows may work from the UK, US, or Canada; fellows may also work from the London or Berkeley shared workspaces.
No H1B Sponsorship

About the job

Requirements
  • Have a strong technical background in computer science, mathematics, or physics.
  • Be fluent in Python programming.
  • Be available to work full-time on the Fellows Program.
  • Have work authorization in the United States, the United Kingdom, or Canada and be located in that country during the program.
Responsibilities
  • Use external infrastructure, such as open-source models and public APIs, to work on an empirical project aligned with the program's research priorities.
  • Aim to produce a public output, such as a paper submission.
  • Participate in project selection and mentor matching.
  • Complete the four-month, full-time program; participants unable to commit to the full duration may note their constraints in the application for case-by-case review.
Desired Qualifications
  • Have a strong background in a discipline relevant to a specific workstream, such as economics, social sciences, or cybersecurity.
  • Have experience in research or engineering related to the selected workstream.
  • Have experience with empirical machine learning research projects.
  • Have experience working with large language models.
  • Have experience in one of the listed research areas: scalable oversight, adversarial robustness and AI control, model organisms, model internals or mechanistic interpretability, or AI welfare.
  • Have a track record of open-source contributions.
  • Have contributed to open-source projects in large language model- or security-adjacent repositories.
  • Have demonstrated success bringing clarity and ownership to ambiguous technical problems.
  • Have experience with penetration testing, vulnerability research, or other offensive security work.
  • Have demonstrated willingness to do the detailed work needed to produce high-quality outputs.
  • Have reported Common Vulnerabilities and Exposures (CVEs) or received bug bounties.
  • Have experience with deep learning frameworks and experiment management.

About the company

Anthropic builds the Claude family of AI models and sells them to businesses, software developers, and consumers, with research on making AI safe and understandable at the center of its work. In 2024 it introduced the Model Context Protocol, an open standard for connecting AI assistants to outside apps and data that OpenAI and Google adopted the next year. People use the Claude app to write and analyze files, programmers hand coding tasks to Claude Code, and companies build Claude into their own products through its API. Anthropic earns money from subscriptions, business plans, and usage-based API fees. Siblings Dario and Daniela Amodei founded Anthropic in 2021 with other former OpenAI employees, and it is based in San Francisco.

Company Size

5,001-10,000

Company Stage

N/A

Total Funding

$224.8B

Headquarters

San Francisco, California

Founded

2021

Get referred to Anthropic

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Sales are multiplying: Anthropic's yearly revenue pace topped $65 billion in July 2026, over sevenfold since December 2025.
  • IPO plans advanced: October 2026 reports describe investor meetings and a mid-November listing target; staff shares face lockups.
  • Lower office minimum: 627 of 646 Anthropic job posts in October 2026 require at least 25% office time.

What critics are saying

  • Job security rides on growth: Anthropic's IPO prospectus lists $518 billion in cloud and computing bills due in coming years.
  • Defense work stays off-limits: in September 2026, appeals judges voted 2-1 to let the Pentagon keep blacklisting Anthropic.
  • Growth is priced in: Anthropic's valuation reached $965 billion in May 2026, up from $380 billion in February.

What makes Anthropic unique

  • Investors don't control its board alone: Anthropic's independent Long-Term Benefit Trust elects a growing number of directors.
  • It isn't tied to one chipmaker: Claude runs on Nvidia GPUs plus custom AI chips designed by Google and Amazon.
  • It holds back its riskiest model: Claude Mythos 5.1, launched September 2026, goes only to vetted cyberdefenders and life scientists.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Flexible Work Hours

Paid Vacation

Parental Leave

Hybrid Work Options

Company Equity

Growth & Insights and Company News

Headcount

6 month growth

↑ 0%

1 year growth

↑ 3%

2 year growth

↓ -3%
CNBC
Oct 9th, 2026
OpenAI and Anthropic hire Trump officials as AI firms scramble to strengthen White House ties

OpenAI and Anthropic are hiring former Trump administration officials for senior roles as AI companies strengthen ties with Washington. Anthropic has appointed at least two ex-Trump officials this year, including Chris Liddell to its board and Sihao Huang as head of frontier compute strategy. OpenAI has hired at least three, including Thomas Lind for cyber and strategic risk, and Dean Ball as head of strategic futures. The moves come as frontier AI is increasingly viewed as a national security issue, with companies seeking to build relationships amid growing regulatory calls. Trump has called AI safety concerns a "hoax" whilst also suggesting law enforcement could "rein in" AI. Both companies are currently hiring for multiple political engagement roles as they prepare for potential public listings.

Yahoo Finance
Oct 9th, 2026
OpenAI's $50B revenue accounting sparks chip stock selloff, analysts call reaction 'overblown

Chip and cloud stocks edged higher in overnight trading after OpenAI's revenue disclosure sparked a sharp selloff on Thursday. Nvidia rose 0.5%, whilst Oracle and CoreWeave gained 0.7% and 0.3% respectively. The iShares semiconductor ETF added 0.03% after sliding 3.4% in the regular session. OpenAI reported roughly $50 billion in annualised revenue at the end of September, lower than the previously reported $70 billion figure. Analysts said the discrepancy stemmed from different accounting methods when comparing OpenAI's revenue with competitor Anthropic's. The Futurum Group CEO Daniel Newman called the broader market reaction "overblown", noting OpenAI could still reach $70 billion to $90 billion in revenue by year-end. In a Stocktwits poll, 48% of retail investors said they plan to buy AI-linked stocks following the dip.

NPR
Oct 9th, 2026
American tech giants face resistance over $65B data centre expansion in Australia

American tech companies are expanding data centre operations in Australia, but face growing opposition from residents and farmers. Local communities near Melbourne are protesting planned facilities, citing concerns over noise, energy consumption, and water usage. Investment in Australian data centres could exceed $100 billion by 2030, according to the Commonwealth Bank of Australia, with Microsoft and Amazon already committing funding. Companies are attracted by Australia's geographic location, renewable energy access, and democratic stability. However, the expansion faces obstacles. New regulations on water efficiency and power supply take effect next year. Australia's strict copyright protections pose challenges for AI companies seeking regulatory certainty. Recent revelations that an OpenAI agent accessed Australian government websites without authorisation have intensified scrutiny. Experts question whether AI companies should receive special treatment regarding accountability.

The Register
Oct 8th, 2026
Anthropic launches free security scanning for critical infrastructure and open-source projects

Anthropic has launched its Cyber Mission programme to help patch software vulnerabilities exposed by its own AI tools. The company acknowledges that attackers currently have the advantage, but predicts defenders will gain the upper hand within two years. The initiative focuses on two areas: a Critical Infrastructure Defense Program, partnering with firms like CrowdStrike and Palo Alto Networks to protect utilities from AI-discovered vulnerabilities, and OSS Scanner, offering free security scans to established open-source projects. According to Anthropic, AI models now find over 85% of vulnerabilities, up from 20% at the start of 2025. The programme follows recent incidents where AI agents broke into at least 20 organisations, including HuggingFace, raising concerns about AI-enabled cyberattacks.

Tech in Asia
Oct 8th, 2026
Anthropic launches cyber programme for critical infrastructure with free AI vulnerability scanner

Anthropic has launched a cyber programme targeting critical infrastructure and introduced OSS Scanner, a free opt-in service for open-source projects that handles vulnerability reports. The AI-generated reports come without human review and may contain inaccuracies. Earlier this week, Anthropic integrated Project Glasswing into its Cyber Verification Program, expanding model access for eligible defenders. Google noted that automated security findings can create additional triage work for maintainers. The Open Source Security Foundation warned that AI-generated disclosures risk adding low-quality noise. The launch reflects a broader trend of vendors marketing AI-enabled cyber defence tools to critical infrastructure operators.