Internship

ML Systems & Reinforcement Learning Fellow

Anthropic

Anthropic

5,001-10,000 employees

Develops reliable, interpretable AI systems

Compensation Overview

$3.9k/wk

+ Compute funding (~$15k/month) + Research expense funding

No H1B Sponsorship

Remote in USA + 3 more

More locations: London, UK | San Francisco, CA, USA | Ontario, Canada

Hybrid

Applicants must have work authorization and be located in the United States, United Kingdom, or Canada during the program; remote fellows are permitted in those countries.

Bachelor's

Category
AI & Machine Learning (1)
Required Skills
LLM
Python
Distributed Systems
High Performance Computing (HPC)
Machine Learning
Smart Contracts
Blockchain
Reinforcement Learning

Get referred to Anthropic

See people who can refer or advise you

Requirements
  • Have a strong technical background in computer science, mathematics, or physics.
  • Be fluent in Python programming.
  • Be available to work full-time on the Fellows Program.
  • Have strong software engineering skills and experience building complex machine learning systems.
  • Be comfortable working with large-scale distributed systems and high-performance computing.
  • Have work authorization in the United States, United Kingdom, or Canada and be located in that country during the program.
Responsibilities
  • Work for four months full-time on an empirical project aligned with Anthropic's research priorities.
  • Produce a public output, such as a paper submission.
  • Participate in project selection and mentor matching.
  • Build a CPU simulator for accelerator workloads, add backends for different accelerators to an open-source project, build on-demand infrastructure, or build complex synthetic data or environment pipelines for the ML Systems and Performance workstream.
  • Build model-based tools to understand AI training data and improve training data quality, research generalization, create reinforcement learning environments, or conduct research and implement solutions involving reinforcement learning algorithms.
Desired Qualifications
  • Have a strong background in a discipline relevant to a specific Fellows workstream, such as economics, social sciences, or cybersecurity.
  • Have experience in research or engineering related to the selected workstream.
  • Be able to balance research exploration with engineering rigor and operational reliability.
  • Have experience training, fine-tuning, or evaluating large language models.
  • Be adept at analyzing and debugging model training processes.

Anthropic focuses on AI research to build reliable, interpretable, and steerable AI systems. Its main product, Claude, is an AI assistant designed to handle tasks at any scale for clients across industries, delivered through deployment and licensing along with specialized AI R&D services. Claude works by combining natural language processing, human feedback, reinforcement learning, and interpretability techniques to produce a capable, controllable AI assistant that can assist with a wide range of tasks. The company differentiates itself from competitors by prioritizing safety, transparency, and controllability—emphasizing reliability, interpretability of model behavior, and user-controlled steerability in its AI systems. Anthropic’s goal is to make AI systems that people can trust and efficiently use to improve operations and decision-making across sectors.

Company Size

5,001-10,000

Company Stage

Debt Financing

Total Funding

$162.8B

Headquarters

San Francisco, California

Founded

2021

Get referred to Anthropic

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Salesforce launched Claudeforce on August 26, 2026, expanding Claude’s enterprise distribution.
  • Reuters reported Anthropic’s 2028 revenue target at $190 billion to $200 billion.
  • Anthropic’s $65 billion May 2026 funding and $965 billion valuation fund aggressive scaling.

What critics are saying

  • Anthropic’s $45 billion Nscale and $1.25 billion monthly SpaceX commitments crush margins.
  • Cheaper rivals like DeepSeek pressure frontier-model pricing; Fable 5 usage stalls at 11%.
  • A failed IPO or compute bottleneck starves Anthropic’s expansion and weakens investor confidence.

What makes Anthropic unique

  • Claude powers Salesforce’s Claudeforce, embedding Anthropic inside enterprise workflows.
  • Anthropic ships global EU-compliant watermarks across Claude, Code, Cowork, and Tag.
  • Anthropic is building in-house chips to co-design models and cut inference costs.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Flexible Work Hours

Paid Vacation

Parental Leave

Hybrid Work Options

Company Equity

Growth & Insights and Company News

Headcount

6 month growth

7%

1 year growth

4%

2 year growth

3%
CNBC
Aug 26th, 2026
Anthropic CEO denies 'Saaspocalypse' as Salesforce integrates Claude AI into products

Anthropic and Salesforce have announced a new partnership that will integrate Claude AI into Salesforce's products. Anthropic co-founder and CEO Dario Amodei and Salesforce chair and CEO Marc Benioff discussed the collaboration on CNBC's Mad Money. The announcement comes amid industry discussions about AI's potential impact on the software-as-a-service sector. Amodei addressed concerns about disruption, stating that Anthropic is not interested in destroying anyone. The partnership will see Claude AI embedded within Salesforce's product suite, though specific details about the integration and its rollout were not disclosed in the announcement.

Techstrong
Aug 26th, 2026
Anthropic signs $45B cloud deal with Nscale ahead of planned IPO

Anthropic has reportedly signed a $45 billion cloud computing deal with UK-based Nscale, marking one of the largest infrastructure agreements in AI history. The six-year arrangement will provide Anthropic with approximately 460 megawatts of compute capacity at a West Virginia data centre, powered by NVIDIA's Vera Rubin chips and expected online by late 2027. The deal follows a series of major infrastructure commitments, including a $1.25 billion monthly payment to SpaceX through May 2029 for data centre access. Anthropic, which confidentially filed for an IPO in June, currently holds a $965 billion valuation and projects annual revenue growth from $47 billion to $200 billion by 2028. The aggressive expansion aims to address capacity constraints as demand for its Claude AI models surges.

Yahoo Finance
Aug 26th, 2026
Anthropic's revenue surge from $10M to $65B shows robotics valuations shifting to long-term potential

Investors may need to reassess how they value humanoid robotics startups, according to RoboStrategy CEO Andrew Kang. Traditional revenue metrics may not capture the full picture for companies targeting large future markets. Kang pointed to Anthropic's rapid scaling as an example, noting the AI company went from $10 million in revenue in 2022 to an expected $65 billion annual recurring revenue in 2026. He argues robotics investors should similarly focus on long-term growth potential rather than near-term sales. Whilst discounted cash flows and production forecasts remain relevant, Kang said qualitative factors—including hardware quality, manufacturing scalability, and robot foundation model performance—will increasingly influence valuations. Humanoid robotics companies may be judged more on their path to large-scale commercial adoption than current business metrics.

The Register
Aug 25th, 2026
Claude and Cowork now share memories, with no option to separate them

Anthropic's Claude chatbot and its workplace assistant Cowork now share a unified memory system. Information Claude learns about users — including personal details, work preferences, and colleague names — is automatically available to Cowork, and vice versa. The shared memory helps Cowork automate workplace tasks like report writing and presentation creation. For example, Cowork can draft manager updates using remembered preferences or build conference logistics documents based on prior Claude conversations. Claude Code's memory remains separate for now. Anthropic hasn't disclosed future plans for integrating it. Memory generation now occurs during conversations rather than afterward. It's enabled by default for free, Pro, and Max tier users, though Cowork requires a paid account. Sensitive information like health details and political views isn't stored unless users opt in. Users cannot separate Claude and Cowork memories within one account. To keep them distinct, separate accounts are required.

Yahoo Finance
Aug 25th, 2026
Anthropic's $65B revenue run rate faces pressure as Fable 5 stalls at 11% of spending ahead of potential IPO

Anthropic's flagship Fable 5 AI model is struggling to attract corporate customers, with spending stalled at 11% of total model usage more than two months after launch. Enterprises are increasingly choosing cheaper alternatives, including Anthropic's own lower-cost models and rivals like DeepSeek. Anthropic charges $10 to $50 per million tokens for Fable 5, whilst comparable DeepSeek offerings cost less than $1 per million tokens. The pricing gap matters as many commercial workloads may not require frontier-level performance. The timing is sensitive as Anthropic has confidentially filed for an IPO, potentially listing in September or October. Reuters reported the company's revenue run rate exceeded $65 billion in July, with targets of $190 billion to $200 billion by 2028.