Full-Time

Product Manager

Updated on 8/21/2026

DatologyAI

DatologyAI

11-50 employees

Automated data curation for GenAI training

Compensation Overview

$215k - $300k/yr

+ Equity + Relocation assistance

Company Does Not Provide H1B Sponsorship

Redwood City, CA, USA

In Person

Four days on-site per week are required. Relocation assistance is available for employees moving to the Bay Area.

Category
Product (1)
Required Skills
MLOps
Data Science
Product Management
Machine Learning
Data Engineering

Get referred to DatologyAI

See people who can refer or advise you

Requirements
  • At least 5 years of product management experience, including at least 3 years building enterprise software or artificial intelligence/machine learning tooling at a senior or staff level.
  • A strong technical foundation, including the ability to read research papers, engage credibly with machine learning engineers about training pipelines and data infrastructure, and distinguish meaningful technical differentiation from noise.
  • Experience shipping products that began as a research paper or prototype and imposing structure while preserving the technology's distinctive value.
  • Experience as a founding or early product manager with a track record of building product functions and processes.
  • Deep familiarity with enterprise artificial intelligence and machine learning buyers, including how machine learning teams evaluate, adopt, and continue using tools and how infrastructure decisions are made.
  • Ability to assess and improve developer and technical user experiences.
  • Ability to translate research results for sales teams and make customer complaints actionable for engineers.
  • Comfort operating in ambiguity and making decisive, data-informed decisions.
Responsibilities
  • Own the product roadmap end-to-end, from discovery and prioritization through launch and iteration, with a focus on enterprise-grade artificial intelligence tooling.
  • Partner with research and engineering teams to turn ambiguous, early-stage outputs into concrete, shippable product decisions.
  • Define and drive the enterprise product experience, including platform user experience, application programming interface design, deployment flexibility such as bring-your-own-cloud and on-premises deployment, and integrations with existing machine learning workflows.
  • Engage directly with enterprise machine learning teams, data scientists, and infrastructure engineers to turn their pain points into a clear product strategy.
  • Build foundational product-management infrastructure, including discovery frameworks, roadmap tooling, release processes, and cross-functional rituals that scale as the team grows.
  • Work with Sales and Customer Success to ensure the product enables a repeatable, defensible go-to-market motion.
  • Track the competitive landscape across artificial intelligence tooling, machine learning operations, and data infrastructure to inform positioning and prioritization.
  • Connect research output with the commercial product by helping the team decide what to build, sequence work, and measure whether it is working.
Desired Qualifications
  • Hands-on experience with model training, data pipelines, or machine learning operations workflows.
  • Prior experience at an artificial intelligence infrastructure, developer tools, or data platform company.
  • Exposure to enterprise procurement and compliance requirements, including bring-your-own-cloud, on-premises deployment, and data sovereignty.

DatologyAI offers automated data curation tools to optimize GenAI training by selecting high-quality, relevant data and removing noisy or harmful data. The core tech analyzes datasets and plugs into existing training pipelines, requiring minimal code changes, and scales from small to petabyte-scale data with usage-based pricing. It differentiates itself with end-to-end automated curation at scale and easy integration, supported by recognized research work and contributions to ImageNet, plus a team with CMU PhD expertise and immigrant-founder VC backing. The goal is to help organizations train better AI models more efficiently and cost-effectively by ensuring high-quality data throughout the training lifecycle.

Company Size

11-50

Company Stage

Series A

Total Funding

$57.7M

Headquarters

Redwood City, California

Founded

2023

Get referred to DatologyAI

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Series A totaled $57.5M in 2024, and the company is still hiring aggressively.
  • The 2026 hiring page claims 7-40x faster training and fewer-than-half-parameter models.
  • The 2026 seminar comeback signals a strong technical brand and recruiting magnet.

What critics are saying

  • A single Thomson Reuters win does not prove repeatable enterprise sales.
  • OpenAI, Anthropic, and Google DeepMind can internalize data curation by 2027.
  • If model labs commoditize curation, DatologyAI's standalone product becomes a services business.

What makes DatologyAI unique

  • Petabyte-scale curation turns noisy datasets into smaller models and faster training.
  • Thomson Reuters validated legal-domain gains: +5% benchmarks and 2.5x post-training amplification in 2026.
  • Research and product share one pipeline, so breakthroughs ship quickly across modalities.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Health Insurance

Dental Insurance

Vision Insurance

401(k) Company Match

Unlimited Paid Time Off

Annual Wellness Stipend

Annual Learning and Development Stipend

Relocation Assistance

Company News

SiliconANGLE Media
May 9th, 2024
DatologyAI raises $46M to streamline AI model training data diets

DatologyAI raises $46M to streamline AI model training data diets - SiliconANGLE

DatologyAI
Feb 23rd, 2024
Introducing DatologyAI — Making models better through better data, automatically

Models are what they eat. AI models trained on large-scale datasets have demonstrated jaw-dropping abilities and have the power to transform every aspect of our daily lives, from work to play. This massive leap in capabilities has largely been driven by corresponding increases in the amount of data we train models on, shifting from millions of data points several years ago to billions or trillions of data points today. As a result, these models are a reflection of the data on which they’re train

SiliconANGLE Media
Feb 23rd, 2024
DatologyAI raises $11.65M to automate data curation for more efficient AI training

DatologyAI raises $11.65M to automate data curation for more efficient AI training.

TechCrunch
Feb 22nd, 2024
DatologyAI is building tech to automatically curate AI training datasets | TechCrunch

A new startup, DatologyAI, claims to be able to automatically curate the massive data sets on which AI models train.