V

Valency

Trusted AI reasoning infrastructure for science

Senior AI-Native DevOps / Operations Engineer

Full-Time
No salary listed
Senior, Expert
Berkeley, CA, USA
HybridThree days on-site per week required in the San Francisco Bay Area.
No H1B Sponsorship

About the job

Requirements
  • At least 8 years of progressively increasing responsibility operating important production systems.
  • Demonstrated success shipping and running high-reliability systems in production.
  • Deep experience with Amazon Web Services in real production environments.
  • A strong background in software engineering and testing, not just infrastructure administration.
  • Experience designing or significantly improving continuous integration and continuous delivery systems and release processes.
  • Experience building or operating logging, monitoring, alerting, and observability systems.
  • Experience improving production reliability, performance, and operational response.
  • Comfort with container-based systems and orchestration platforms.
  • Strong hands-on ability in at least some of Python, Go, Elixir, and Cloud Development Kit.
  • Strong judgment around guardrails, operational safety, and change management.
  • Ability to work in ambiguity and build systems that do not yet fully exist.
  • Legal authorization to work in the United States.
Responsibilities
  • Design, build, and improve the production platform powering Valency.
  • Tighten continuous integration and continuous delivery processes so changes are tested, gated, observable, and safe to ship.
  • Improve production reliability, latency, deployment safety, and incident response.
  • Build operational feedback loops that help engineering and product teams act on real production behavior.
  • Establish logging, analytics, tracing, alerting, and workflow instrumentation as the platform scales.
  • Define and implement guardrails for agent-involved software delivery and operations.
  • Introduce human-in-the-loop approval flows where autonomy needs stronger controls.
  • Improve cost efficiency across cloud infrastructure and platform operations.
  • Help shape security, compliance, and auditability foundations for SOC 2, ISO 27001, and FedRAMP-oriented environments.
  • Contribute to the long-term platform engineering direction as the team grows and specializes.
  • Own production operations and operational excellence for this function.
  • Lead incident response expectations for the role.
  • Establish the operating model the broader team will scale on.
  • Own and improve continuous integration and continuous delivery pipelines, release controls, and deployment workflows.
  • Build and maintain highly reliable Amazon Web Services-based production systems.
  • Improve observability across logs, metrics, traces, events, and workflow state.
  • Instrument platform behavior so system issues, regressions, and slowdowns are quickly visible and actionable.
  • Create operational analytics that help close the loop between engineering, product, and customer experience.
  • Drive cost engineering and infrastructure efficiency as the system scales.
  • Build safer operating patterns for agent-assisted code changes and operational actions.
  • Implement testing, validation, approval, and rollback mechanisms that reduce operational risk.
  • Improve batch, queue, cache, and job-processing reliability and monitoring.
  • Support incident response, root cause analysis, postmortems, and follow-through.
  • Partner with external vendors and partners when needed.
  • Help define platform standards, reliability practices, and operational maturity across the company.
Desired Qualifications
  • Startup experience, especially in fast-scaling environments.
  • Experience at high-scale software-as-a-service companies that have gone through periods of rapid growth.
  • Experience owning or materially influencing platform engineering functions.
  • Experience with cost engineering or FinOps in Amazon Web Services-heavy environments.
  • Experience designing systems for compliance-oriented environments.
  • Experience with SOC 2, ISO 27001, or FedRAMP-related operational requirements.
  • Experience evaluating or implementing modern observability and workflow tracing stacks.
  • Experience creating human-in-the-loop approval systems for sensitive production workflows.

About the company

Valency provides a trusted research infrastructure that grounds AI reasoning in real scientific literature, addressing hallucinations from large language models. Its Valency Bond connects with LLMs and next-generation tools to create a grounded reasoning workflow for researchers. Unlike general AI platforms, Valency focuses on integrating credible, up-to-date research content to accelerate discovery for the 35 million researchers worldwide. The product works by linking LLMs with a framework and tools that rely on real research data, enabling researchers to enhance reasoning, verify results, and streamline scientific workflows. The company differentiates itself by offering a grounded, research-backed ecosystem and infrastructure designed specifically to support scientific inquiry, aiming to speed up discoveries and improve trust in AI-assisted research.

Company Size

11-50

Company Stage

N/A

Total Funding

N/A

Headquarters

Berkeley, California

Founded

2026

Get referred to Valency

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • July 22, 2026 Genesis Mission selection gives Valency federal validation and budgeted demand.
  • August 2026 job postings for product, growth, and research roles signal active scaling.
  • Valency's public corpus reaches 450+ million documents, a moat for scientific search.

What critics are saying

  • Genesis is the first real-world use; commercial revenue remains unproven.
  • Javelin's August 1, 2026 posting says Valency has almost no data infrastructure.
  • If scientists ignore provenance-heavy tools, Valency's entire agentic-science thesis collapses.

What makes Valency unique

  • Valency Bond indexed 450+ million papers and preprints, with citations traceable within hours.
  • Its MCP-native infrastructure powers audited AI agents for DOE-backed HERALD research.
  • CEO Josh Bloom and COO Ryan Anderson pair Berkeley science with enterprise execution.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Hybrid Work Options

Remote Work Options

Company Equity

401(k) Company Match

Growth & Insights and Company News

Headcount

6 month growth

↑ 20%

1 year growth

↑ 20%

2 year growth

↑ 20%
PR Newswire
Jul 22nd, 2026
Valency selected for DOE Genesis Mission to accelerate nuclear power research with AI agents

Valency has been selected for the US Department of Energy's Genesis Mission to accelerate nuclear power research using AI agents. The company will partner with Lawrence Berkeley National Laboratory on the HERALD project, led by scientists Daniela Ushizima and Peter Nugent. HERALD aims to deploy AI agents to review vast archives of nuclear power research, allowing human security experts to focus on critical judgment calls. Valency will provide the AI agent infrastructure to support hundreds of millions of scientific documents whilst maintaining auditable chains of custody required for DOE compliance. The system uses the Model Context Protocol natively, enabling AI agents to work across the archive whilst keeping human experts in control of final decisions. Valency specialises in foundational infrastructure for AI-accelerated science.