Full-Time

Red Team Engineer

Safeguards

Updated on 9/3/2026

Anthropic

Anthropic

5,001-10,000 employees

Develops reliable, interpretable AI systems

Compensation Overview

$320k - $405k/yr

+ Optional equity donation matching

H1B Sponsorship Available

San Francisco, CA, USA

Hybrid

At least 25% of working time is expected in an office; travel is required.

Bachelor's

Category
Cybersecurity (1)
Required Skills
LLM
Distributed Systems
Penetration Testing

Get referred to Anthropic

See people who can refer or advise you

Requirements
  • Experience in penetration testing, red teaming, or application security.
  • Experience in model jailbreaking and testing large-scale agentic workflows for non-obvious prompt injection vectors.
  • Strong technical skills in web application security, including hands-on expertise with security testing tools such as Burp Suite, Metasploit, and custom scripting frameworks.
  • Experience building custom automation, including large language model-specific testing frameworks.
  • A track record of discovering novel attack vectors and chaining vulnerabilities in creative ways.
  • A public body of work such as CVEs, blog posts, or disclosed bug bounty reports.
  • Strong written and verbal communication skills, with the ability to explain technical concepts to varied audiences.
  • A bachelor's degree or an equivalent combination of education, training, and experience in a field relevant to the role, as demonstrated through coursework, training, or professional experience.
Responsibilities
  • Conduct comprehensive adversarial testing across Anthropic's product surfaces, developing creative attack scenarios that combine multiple exploitation techniques.
  • Research and implement novel testing approaches for emerging capabilities, including agent systems, tool use, and new interaction paradigms.
  • Design and execute full kill-chain attacks that emulate real-world threat actors attempting to achieve specific malicious objectives.
  • Build and maintain systematic testing methodologies that evaluate every aspect of the systems.
  • Develop automated testing frameworks to enable continuous assessment at scale.
  • Collaborate with Product, Engineering, and Policy teams to translate findings into concrete improvements.
  • Help establish metrics for measuring detection effectiveness of novel abuse.
Desired Qualifications
  • Experience with artificial intelligence/machine learning security or adversarial machine learning.
  • Understanding of artificial intelligence safety considerations beyond traditional security, including modern guardrails against jailbreaks.
  • Experience testing application programming interface security and rate-limiting systems.
  • Background in testing business logic vulnerabilities and authorization bypass techniques.
  • Background in anti-fraud, trust and safety, or abuse prevention systems.
  • Familiarity with distributed systems and infrastructure security.
  • Familiarity with abuse detection mechanisms and the ability to engineer novel bypasses.
  • Adaptability to understand and build engagements around emerging threats outside the direct area of expertise.

Anthropic focuses on AI research to build reliable, interpretable, and steerable AI systems. Its main product, Claude, is an AI assistant designed to handle tasks at any scale for clients across industries, delivered through deployment and licensing along with specialized AI R&D services. Claude works by combining natural language processing, human feedback, reinforcement learning, and interpretability techniques to produce a capable, controllable AI assistant that can assist with a wide range of tasks. The company differentiates itself from competitors by prioritizing safety, transparency, and controllability—emphasizing reliability, interpretability of model behavior, and user-controlled steerability in its AI systems. Anthropic’s goal is to make AI systems that people can trust and efficiently use to improve operations and decision-making across sectors.

Company Size

5,001-10,000

Company Stage

Debt Financing

Total Funding

$182.8B

Headquarters

San Francisco, California

Founded

2021

Get referred to Anthropic

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Project Glasswing found over 10,000 critical vulnerabilities by June 2, 2026.
  • Anthropic disclosed annualized revenue above $65 billion in July 2026, signaling explosive demand.
  • September 1, 2026 pricing cuts for cache reads boost agentic API adoption and retention.

What critics are saying

  • Anthropic’s $1.5 billion copyright settlement, approved July 20, 2026, invites more suits.
  • Late-September 2026 IPO pressure exposes weak multiples if growth decelerates after listing.
  • Heavy compute commitments and chip-lease debt create existential financing risk if demand softens.

What makes Anthropic unique

  • Claude Security and Project Glasswing anchor Anthropic’s enterprise security moat in 2026.
  • Anthropic pairs frontier-model capability with explicit safety branding, unlike OpenAI’s consumer-first posture.
  • Multi-cloud distribution across AWS, Google, Microsoft, Lambda, and Nscale reduces single-vendor dependence.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Flexible Work Hours

Paid Vacation

Parental Leave

Hybrid Work Options

Company Equity

Growth & Insights and Company News

Headcount

6 month growth

7%

1 year growth

4%

2 year growth

2%
USA News Group
Sep 5th, 2026
AMD's $5 Billion Anthropic Bet Nears an IPO Test

AMD's late-July commitment of up to $5 billion to Anthropic, paired with a 2-gigawatt Instinct MI450 deployment, faces its first public valuation test as an…

Newsbytes
Sep 4th, 2026
Anthropic secures $15B credit facility, expects $65B revenue ahead of IPO

Anthropic, the developer behind Claude AI chatbot, has secured a $15 billion credit facility ahead of its initial public offering. Major financial institutions including Morgan Stanley, Goldman Sachs, JPMorgan Chase, and Citigroup are backing the arrangement. The facility significantly exceeds last year's $2.5 billion loan and surpasses the company's roughly $10 billion target. Anthropic now expects over $65 billion in annualised revenue, representing more than a sevenfold increase from its pace at the end of last year. Additional banks including Barclays and Bank of America have joined the arrangement. The timing coincides with renewed activity in the US IPO market, positioning Anthropic for a substantial market debut.

GlobeNewswire
Sep 3rd, 2026
PicPay announces integration with Claude, expanding customer access to financial information through generative AI.

PicPay announces integration with Claude, expanding customer access to financial information through generative AI. PicPay continues to accelerate its AI strategy, building on the recent introduction of the company's ChatGPT plugin to simplify how customers manage their financial lives. SÃO PAULO, Sept. 03, 2026 (GLOBE NEWSWIRE) - PicPay, one of Brazil's largest digital banks, today announced an integration with Claude from Anthropic, available to its customers, expanding its presence in the generative artificial intelligence ecosystem. The initiative, which follows the recent launch of PicPay's plugin on ChatGPT, from OpenAI, advances the company's strategy to bring conversational financial information closer to customers. Through the integration, customers can check financial information directly in Claude by using natural-language commands after securely authenticating in the PicPay environment. The feature will be rolled out gradually through an early-access list for customers. It is free of charge and available to Android and iOS users. Customers who already use the PicPay plugin on ChatGPT are also enabled to use the feature on Claude, which is available regardless of the user's Claude subscription tier. Similar to PicPay's integration with ChatGPT, the experience is focused on financial inquiries and monitoring. Claude offers two new capabilities that will soon also be available through the OpenAI plugin: checking credit card statement balances and viewing completed credit card transactions. During the initial rollout, the integration is dedicated to inquiries and financial monitoring and does not support transactions or other financial operations through the platform. "As many of our customers take advantage of the leading generative AI platforms in their daily lives, we are focused on connecting PicPay with these ecosystems to create a more seamless financial experience," said Eduardo Chedid, CEO of PicPay. "PicPay has always placed innovation at the center of our growth strategy, as we capitalize on opportunities where technology can help us operate more efficiently and deliver even better service for customers. Our strong standards for customer data protection will continue to underpin the innovative products we bring to market." To connect their PicPay account to Claude, customers simply search for PicPay under the "Connectors" tab and follow the authentication steps. Available features include: * Checking account and savings pot balances * Viewing outstanding bills * Checking investments * Viewing active cards * Accessing favorite Pix contacts * Checking recent transactions * Viewing upcoming transactions * Checking credit card statement balances * Viewing completed credit card transactions Customers can ask questions such as: * "What is my available balance?" * "What is my credit card statement balance this month?" * "What were my most recent credit card transactions?" * "Do I have any bills due this week?" The Claude integration uses the same security layers as the PicPay app, including facial biometric authentication. To connect their account, customers are directed to PicPay's online banking login, so they do not need to have the app installed. All data exchanges follow PicPay's security standards, with data processed in accordance with applicable law and the company's privacy policies. About PicPay PicPay is one of Brazil's largest digital banks by number of customers. The company operates a two-sided ecosystem that connects consumers and businesses. PicPay offers a broad range of financial products and services, including a digital wallet, credit cards, loans, investments and insurance, for both individuals and businesses.

The New York Times
Sep 3rd, 2026
Amadeus taps Anthropic to put travel content inside AI coding tools - unite.ai.

Amadeus taps Anthropic to put travel content inside AI coding tools - unite.ai. Amadeus said on September 3, 2026 that it is collaborating with Anthropic to make its travel content, connectivity, documentation and authentication available inside AI-enabled development environments, and to bring selected Amadeus applications into a plugin for Anthropic's Claude Cowork aimed at travel professionals. The collaboration, described in an Amadeus blog post by Luc Viguie, the company's director of strategic partnerships, targets two user groups: developers building travel applications and the travel professionals who use Amadeus tools in their daily work. Viguie wrote that travel technology has been designed for decades around specialists navigating screens, reading documentation, using workflows and completing tasks, and that the model is evolving as AI assistants and agents support more tasks, from writing code to running processes and preparing analysis. "This creates an opportunity to rethink how travel technology, content and services can be accessed and used in more AI-native workflows," he wrote. A toolkit for ai-enabled development environments. According to the post, developers are already changing how they work, using agentic coding tools such as Claude Code to describe what they want to build and to help write, test and assemble applications. Amadeus said it is working to make its connectivity, documentation, authentication and travel content available through a single toolkit for AI-enabled development environments. The stated ambition is to make it easier for developers and their AI coding assistants to reach the resources they need, reducing time spent navigating portals, reference guides and integration requirements. The company described the toolkit as a natural evolution from the traditional developer portal. It is intended to give Claude Code users access to what Amadeus characterized as reliable, real-time travel content directly within their development environment, using natural language to discover capabilities, build prototypes and accelerate application development. Access will start with flight and destination content, the post said. A Claude Cowork plugin for travel professionals. The second track of the collaboration concerns travel professionals. Amadeus said agentic tools such as Claude Cowork create new ways to work faster and support more informed decisions, and that it is working to make its skills and connectivity available through a Claude Cowork plugin, bringing selected Amadeus applications closer to where users work. The company described this as an open approach in which trusted travel intelligence is accessed through the tools and environments people choose, rather than requiring users to move between closed systems. "The objective is to go beyond surfacing data, enabling users to access information, conduct research, draft content, analyze data and take action across Amadeus systems in plain language, within broader AI-powered workflows," the company wrote. Amadeus said it is opening these capabilities in a controlled way so that travel professionals can benefit from AI while decisions, safeguards and oversight remain in the hands of users and their organizations. The controlled-rollout language aligns with the governance framework the company publishes through its AI Trust Center, which describes an Amadeus AI Policy intended to guide ethical, secure and legally compliant AI use across all functions worldwide and to meet obligations under the EU AI Act, the European Union's legislative framework for regulating AI systems. That page lists an AI Office driving execution of the company's AI strategy, an AI Compliance Program, an AI Center of Excellence and a cross-functional AI Advisory Board, along with ethics principles covering human oversight, reliability and safety, privacy and security, transparency, accountability and sustainability. Amadeus also states there that it is a signatory of the AI Pact, a European Commission initiative under which companies voluntarily commit to implement key provisions of the EU AI Act ahead of its full application. Enterprise pilot and platform context. Amadeus said the work builds on real-world experimentation, reporting that it recently completed a successful Claude Enterprise pilot inside the company and that assessments of future deployment options are ongoing. The initiatives form part of what the company described as its strategic response to a broader shift in how people build and use software as AI agents become more capable. The company has previously set out the platform context for that response. In a March 23, 2026 post, Maria Plaza, head of the Amadeus AI Office, wrote that Amadeus has served as a system of record for the travel industry for almost forty years, providing a single source of truth across millions of data points a day, from offers, orders and passenger name records to routes, schedules, fares and availability. Plaza wrote that the company's technology is deeply integrated, connecting hundreds of systems, products and workflows, and that Amadeus processes up to 150,000 transactions per second at peak times while powering millions of searches and bookings every day. She also pointed to partnerships with Microsoft and Google Cloud, and to the company's acquisition of SkyLink, an AI-first company specializing in orchestration and conversational automation, as elements of its AI strategy. In the Anthropic announcement, Viguie pointed to the same foundation, writing that the company's status as a system-of-record for travel, its deep integration across the industry, its global production scale and its responsible AI approach, which he said is backed by governance and strategic partnerships, position it to help the industry embrace AI agents. Amadeus has opened a waitlist for the developer toolkit and said those who join will be notified as soon as the solution becomes available.

iTech Post
Sep 3rd, 2026
Anthropic's Claude Content Checker tool is now available - here's how to use the detector.

Anthropic's Claude Content Checker tool is now available - here's how to use the detector. Anthropic's Claude detector tool is here. Anthropic has released a free tool that lets anyone check whether a file was made using Claude, giving users a straightforward way to verify AI involvement in a piece of content. The tool arrives as part of a broader industry push toward transparency around AI-generated material, though it comes with a fairly significant limitation worth understanding before relying on it. Anthropic releases Content Checker tool. The Content Checker is available now through Claude's website, letting anyone upload a file to see whether Anthropic's AI was involved in creating it. According to Anthropic, the tool runs in your browser, meaning a user's file never leaves your device during the checking process and stays local rather than being uploaded to Anthropic's servers. The tool works by reading what Anthropic calls a content credential, described as "signed provenance metadata that is attached when you download, for example, a chart Claude created." When Claude produces a supported file type, it automatically attaches this credential as a small, cryptographically signed note embedded directly in the file's metadata, confirming the file was made or processed using Claude. Claude AI generation detector features. The biggest catch tied to this tool is what it actually confirms and what it does not. Anthropic states plainly that the checker "identifies only the content credential," and that it "can't tell whether Claude was involved in creating the content." In other words, the tool verifies that a specific technical marker exists in a file, rather than analyzing the actual substance of the file to determine whether AI was used to generate it. That distinction matters for anyone hoping to use the tool as a general AI detector. Since the checker only reads an embedded credential rather than scanning content itself, a file with the metadata stripped, converted, or otherwise altered could pass through undetected even if Claude was originally involved. Anthropic also notes the tool "does not include any identifying information about the user," meaning it cannot reveal who created a flagged file, only that Claude was involved in producing it at some point. How to use the Content Checker tool. Using the tool is straightforward. Anthropic's Content Checker supports a wide range of image formats, including PNG, JPG, WEBP, HEIC, TIFF, AVIF, SVG, and GIF, along with video and music files, with individual files supported up to 100MB in size. Users simply need to upload a file directly through the browser-based interface, and Anthropic's tool checks it automatically without requiring any additional setup or account sign-in. This checker functions differently from Anthropic's separate text watermarking system, which embeds invisible marks into Claude-generated text rather than file metadata. According to Anthropic's help documentation, that text-based watermark detection remains in "private preview." It is currently limited to eligible organizations required to verify Claude's compliance under the EU AI Act, including regulators, law enforcement, fact-checkers, and certain enterprises. Anthropic has said it plans to expand access to that separate detection system over time, though for now, the newly released Content Checker remains the company's only publicly available tool for verifying whether a specific file passed through Claude.