Internship

ML Systems & Reinforcement Learning Fellow

Updated on 9/5/2026

Anthropic

Anthropic

5,001-10,000 employees

Develops reliable, interpretable AI systems

Compensation Overview

$3.9k/wk

No H1B Sponsorship

Remote in USA + 3 more

More locations: London, UK | San Francisco, CA, USA | Ontario, Canada

Hybrid

Remote fellows must be located in the UK, US, or Canada; shared workspaces are available in London and Berkeley.

Bachelor's

Category
AI & Machine Learning (1)
Required Skills
LLM
Python
Distributed Systems
High Performance Computing (HPC)
Machine Learning
Smart Contracts
Blockchain
Reinforcement Learning

Get referred to Anthropic

See people who can refer or advise you

Requirements
  • Have strong software engineering skills and experience building complex machine learning systems.
  • Balance research exploration with engineering rigor and operational reliability.
  • Be comfortable working with large-scale distributed systems and high-performance computing.
  • Have experience training, fine-tuning, or evaluating large language models.
  • Be adept at analyzing and debugging model training processes.
  • Have a strong technical background in computer science, mathematics, or physics.
  • Be fluent in Python programming.
  • Have work authorization in the United States, United Kingdom, or Canada and be located in that country during the program.
  • Be available to work full-time on the Fellows program.
  • Commit to the four-month, full-time program, or state any constraints in the application.
Responsibilities
  • Use external infrastructure, such as open-source models and public APIs, to work on an empirical project aligned with Anthropic's research priorities.
  • Produce a public output, such as a paper submission.
  • Undergo project selection and mentor matching.
  • Build a CPU simulator for accelerator workloads.
  • Add backends for different accelerators to an open-source project.
  • Build on-demand infrastructure for infrastructure-heavy fellows projects.
  • Build complex synthetic data or environment pipelines.
  • Build model-based tools to understand AI training data and improve training data quality.
  • Conduct research to better understand generalization.
  • Create reinforcement learning environments to improve Claude models at capabilities within the fellow's domain of expertise.
  • Build reinforcement learning environments for safety-related tasks.
  • Conduct research and implement solutions in areas such as reinforcement learning algorithms.
Desired Qualifications
  • Be motivated by making sure AI is safe and beneficial for society as a whole.
  • Be interested in transitioning into empirical AI research and potentially pursuing a full-time role at Anthropic.
  • Enjoy collaborating across research and engineering disciplines.
  • Have experience working with large-scale distributed systems and high-performance computing, including in trading.
  • Thrive in fast-paced, collaborative environments.
  • Implement ideas quickly and communicate clearly.

Anthropic focuses on AI research to build reliable, interpretable, and steerable AI systems. Its main product, Claude, is an AI assistant designed to handle tasks at any scale for clients across industries, delivered through deployment and licensing along with specialized AI R&D services. Claude works by combining natural language processing, human feedback, reinforcement learning, and interpretability techniques to produce a capable, controllable AI assistant that can assist with a wide range of tasks. The company differentiates itself from competitors by prioritizing safety, transparency, and controllability—emphasizing reliability, interpretability of model behavior, and user-controlled steerability in its AI systems. Anthropic’s goal is to make AI systems that people can trust and efficiently use to improve operations and decision-making across sectors.

Company Size

5,001-10,000

Company Stage

Debt Financing

Total Funding

$177.8B

Headquarters

San Francisco, California

Founded

2021

Get referred to Anthropic

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Amazon expanded funding with up to $20 billion more and 5 gigawatts of capacity.
  • Claude Code rate limits rose after SpaceX compute opened, improving paid-user retention and growth.
  • Broadcom says Anthropic will become its largest XPU customer in 2027, securing supply priority.

What critics are saying

  • The Pentagon blacklist dispute still threatens government revenue and a 180-day contract unwind.
  • The $1.5 billion authors settlement exposes Anthropic to future copyright claims and training restrictions.
  • Massive compute commitments with Amazon, Google, SpaceX, and others create existential margin and execution risk.

What makes Anthropic unique

  • Claude ships on AWS, Google Cloud, and Microsoft Azure, unmatched among frontier models.
  • Anthropic sells safety-aligned AI with enterprise controls, auditability, and in-region deployments.
  • The Claude Partner Network and Cognizant deal deepen distribution through certified implementation partners.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Flexible Work Hours

Paid Vacation

Parental Leave

Hybrid Work Options

Company Equity

Growth & Insights and Company News

Headcount

6 month growth

7%

1 year growth

4%

2 year growth

3%
Newsbytes
Sep 4th, 2026
Anthropic secures $15B credit facility, expects $65B revenue ahead of IPO

Anthropic, the developer behind Claude AI chatbot, has secured a $15 billion credit facility ahead of its initial public offering. Major financial institutions including Morgan Stanley, Goldman Sachs, JPMorgan Chase, and Citigroup are backing the arrangement. The facility significantly exceeds last year's $2.5 billion loan and surpasses the company's roughly $10 billion target. Anthropic now expects over $65 billion in annualised revenue, representing more than a sevenfold increase from its pace at the end of last year. Additional banks including Barclays and Bank of America have joined the arrangement. The timing coincides with renewed activity in the US IPO market, positioning Anthropic for a substantial market debut.

Yahoo Finance
Sep 3rd, 2026
Broadcom projects $230B AI revenue by 2028 as OpenAI, Anthropic fuel chip demand

Broadcom expects AI semiconductor revenue to reach $115 billion in fiscal 2027 and $230 billion in fiscal 2028, driven by accelerating demand from customers including Anthropic and OpenAI, chief executive Hock Tan said on the company's third-quarter earnings call. The Palo Alto-based semiconductor company's AI chip sales more than tripled to $16.7 billion in the third quarter, lifting total revenue 86% to $29.59 billion. Adjusted profit came in at $3.32 per share. Anthropic is expected to become Broadcom's largest XPU customer in 2027, planning to deploy 5 gigawatts of Broadcom-designed TPU v8i chips in 2027 and an additional 10 gigawatts in 2028. OpenAI could deploy more than 5 gigawatts of Broadcom's next-generation Jalapeno XPU in 2028. Despite the strong outlook, Broadcom's weaker-than-expected fourth-quarter revenue forecast sent shares 1.3% lower in overnight trading.

Associated Press
Sep 2nd, 2026
CrowdStrike's Falcon platform joins Anthropic Claude Marketplace with AI-native cybersecurity integration

CrowdStrike announced its Falcon cybersecurity platform is coming to the Anthropic Claude Marketplace, allowing Anthropic customers to purchase CrowdStrike using their existing Anthropic commitments. The integration enables security teams to deploy Charlotte AI AgentWorks directly from Claude, building custom security agents through natural language commands without coding expertise. The collaboration allows defenders to execute security workflows conversationally within Claude, with agents grounded in Falcon platform data. Teams can build, test and refine custom agents in minutes, scaling proven workflows across security operations centres with scoped permissions and full auditability. Daniel Bernard, chief business officer at CrowdStrike, said the partnership puts AI-native cybersecurity where enterprises are deploying AI tools. The announcement was made at CrowdStrike's Fal.Con 2026 conference.

Channel NewsAsia
Sep 2nd, 2026
US Commerce chief says Anthropic back on 'right side' with Trump, Axios reports.

US Commerce chief says Anthropic back on 'right side' with Trump, Axios reports. 03 Sep 2026 12:29AM (Updated: 03 Sep 2026 01:15AM) Add CNA as a trusted source to help Google better understand and surface our content in search results. WASHINGTON, Sept 2: U.S. Commerce Secretary Howard Lutnick said AI giant Anthropic is "back on the right side" with the Trump administration, Axios reported on Wednesday citing an interview. U.S. Defense Secretary Pete Hegseth had blocked Anthropic from certain military contracts because the company refused to allow the U.S. to use its Claude AI models for domestic surveillance or autonomous weapons. Anthropic sued in California court. On August 27, a U.S. judge sided with Anthropic in the company's fight with the American military over AI safety on the battlefield. Anthropic and the Commerce Department did not immediately respond to requests for comment from Reuters. Anthropic co-founder Tom Brown appeared on stage at a tech-focused G20 event in North Carolina on Wednesday morning. Brown stressed the importance of data centers at the gathering, urging the gathered officials to grant permits to build them in their countries.

The Register
Sep 2nd, 2026
Anthropic hires UK AI policy chief Matt Clifford as MPs warn of 'clear conflict of interest

Anthropic has hired Matt Clifford as managing director of international affairs to lead its government relations outside North America. Clifford, who authored the UK's AI Opportunities Action Plan published in January 2025, will retain his position as chair of ARIA, Britain's high-risk research agency. Dame Chi Onwurah MP, chair of the House of Commons Science, Innovation and Technology Committee, called the arrangement a "clear conflict of interest". She said Clifford's continued ARIA role, which involves investing taxpayer money in AI-related research whilst working for a leading AI company, raises concerns about the agency's independence. Clifford reportedly plans to recuse himself from ARIA matters involving Anthropic. He will be based in London.