Full-Time

Recruiter

HUD

HUD

11-50 employees

AI agent evaluation platform and benchmarks

No salary listed

H1B Sponsorship Available

San Francisco, CA, USA

In Person

On-site in the San Francisco Bay Area.

Category
People & HR (2)
,

Get referred to HUD

See people who can refer or advise you

Requirements
  • Experience hiring for engineering, product, and go-to-market roles.
  • Ability to evaluate candidates through portfolios and projects, not only resumes.
  • Ability to understand and communicate technical concepts well enough to engage credibly with candidates.
  • Ability to manage multiple roles and candidate pipelines simultaneously.
  • Experience working in an early-stage startup and independently in a fast-paced environment.
Responsibilities
  • Own end-to-end recruiting across technical and go-to-market roles.
  • Build and optimize recruiting operations and tooling, automating workflows whenever possible.
  • Partner closely with founders and team leads to define role requirements, evaluation criteria, and interview processes.
  • Develop a strong understanding of the product to effectively pitch technical candidates.
  • Source top candidates through high-signal and unconventional channels.
  • Screen candidates for technical competence and product-market intuition.
  • Support onboarding and new-hire ramp.
  • Help define and run performance-review processes, leveling, and feedback systems.
Desired Qualifications
  • A track record of sourcing creatively.
  • Familiarity with Ashby, Juicebox, and Clay.
  • Experience with vibe coding and using artificial intelligence to automate workflows.
  • Exposure to or ownership of people or human resources functions.
  • Experience evaluating non-traditional candidate profiles.

HUD provides an evaluation platform for AI agents that perform computer-use tasks. It offers an interface that connects to HUD evaluation environments, allowing users to run benchmarks across hundreds of environments and thousands of tasks. Users can integrate their agents using various adapters, and interact with the system through an asynchronous API for efficient experiments. The platform then collects telemetry and benchmarking results to inform agent improvements. HUD differentiates itself by concentrating on evaluating and improving knowledge-work agents across a wide range of tasks and environments, rather than just building models, and by enabling scalable, tool-enabled testing via adapters and an asynchronous workflow. The company’s goal is to help organizations and developers assess, compare, and enhance their AI agents so they perform better in real-world knowledge-work scenarios.

Company Size

11-50

Company Stage

Seed

Total Funding

$130K

Headquarters

San Francisco, California

Founded

2025

Get referred to HUD

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • HUD raised a $16M Series A led by Standard Capital in June 2026.
  • HUD says over 50 businesses use the platform, including vendors selling millions monthly.
  • Native GitHub Agentic Workflows support expands HUD into coding-agent evaluation workflows.

What critics are saying

  • OpenAI and Anthropic can internalize evals, shrinking HUD's core market quickly.
  • HUD faces open-source imitation from Cua, OSWorld, and other benchmark builders.
  • If labs standardize on in-house benchmarks by 2027, HUD becomes a commodity layer.

What makes HUD unique

  • HUD sells live computer-use evals against real software, not static datasets.
  • Its CUA Evals framework targets browser agents, desktop tasks, and RL training pipelines.
  • HUD also runs a vendor marketplace, connecting post-training suppliers directly with AI labs.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Health Insurance

Dental Insurance

Vision Insurance

Paid Vacation

Paid Holidays

Commuter Benefits

Growth & Insights

Headcount

6 month growth

-8%

1 year growth

-8%

2 year growth

-8%