Full-Time

Network Engineer

Cluster Engineering

Cerebras

Cerebras

501-1,000 employees

AI accelerator hardware replacing GPUs

No salary listed

Sunnyvale, CA, USA

In Person

Master's, PhD

Category
DevOps & Infrastructure (1)
Required Skills
VXLAN
Python
Grafana
Jinja2
High Performance Computing (HPC)
InfluxDB
Juniper
Computer Networking
Go
Prometheus
Ansible
DevOps

Get referred to Cerebras

See people who can refer or advise you

Requirements
  • A Ph.D. in Computer Science or Electrical Engineering with at least 5 years of industry experience, or a Master's degree in Computer Science or Electrical Engineering with at least 10 years of industry experience.
  • Solid experience designing large-scale networks in datacenter and cloud environments.
  • Extensive hands-on experience debugging networking issues in large distributed systems with multiple platforms and protocols.
  • A demonstrated track record leading multi-phase, multi-team technical projects to completion.
  • Deep expertise across Juniper, Arista, Cisco, and open-box or disaggregated Network Operating System architectures such as SONiC.
  • Strong working knowledge of VXLAN, EVPN, RoCEv2, BGP, DCQCN, PFC, ECN, and streaming telemetry.
  • Proficiency in Python and/or Go for building network automation, validation, and tooling.
  • Comfort with configuration-generation frameworks such as Ansible and Jinja2, gNMI, and continuous integration and continuous delivery pipelines for network infrastructure.
  • Hands-on experience with streaming telemetry pipelines, time-series databases such as Prometheus and InfluxDB, visualization with Grafana, log aggregation, and modern incident management.
  • Ability to define SLIs/SLOs and instrument the network for proactive reliability.
  • Familiarity with network visibility, management, and packet-capture and analysis tools.
Responsibilities
  • Design and architect front-end network fabrics for AI/ML and HPC clusters, optimizing resource utilization, latency, and throughput.
  • Build proof-of-concept implementations of new network designs and features and drive them from prototype through production rollout.
  • Identify and resolve performance and efficiency bottlenecks across the host, NIC, and fabric.
  • Automate the deployment, configuration, and validation of network infrastructure using Python, including topology provisioning, fabric bring-up, configuration generation, and regression testing.
  • Stand up and operate SRE-grade telemetry and observability for the cluster network, including streaming telemetry, metrics pipelines, alerting, and incident workflows.
  • Define the SLIs and SLOs that govern network reliability and drive blameless post-incident analysis.
  • Lead network debugging in large distributed-systems environments across multiple platforms and protocols, including deep dives into RoCEv2, PFC/DCQCN, ECMP hashing, congestion behavior, and packet-level forensics.
  • Lead cross-functional, multi-phase technical projects spanning hardware, firmware, host networking, and cluster software.
  • Collaborate with vendors and industry partners to shape network hardware and feature roadmaps.
  • Represent the company in industry forums, standards bodies, and technical communities.
  • Serve as the central point of contact for network reliability issues across the cluster.
Desired Qualifications
  • Prior experience at hyperscalers or cloud service providers.
  • Experience with AI/ML or HPC cluster networking, including lossless Ethernet design, rail-optimized topologies, and collective-communication traffic patterns.
  • A track record of contributions to open-source networking projects, standards bodies, or industry conferences.

Cerebras Systems creates AI acceleration hardware and software. Its CS-2 system is designed to replace traditional GPU clusters for AI workloads, speeding up training and inference while simplifying the setup by eliminating the need for parallel programming, distributed training, and cluster management. The product works as a single, large processor-based accelerator with accompanying software and cloud services to run AI models efficiently, reducing latency and time to results. Compared with competitors, Cerebras differentiates itself with the largest processor in the industry and an integrated hardware-software stack that aims to streamline AI workflows rather than relying on multi-GPU clusters. The company’s goal is to help research labs, healthcare, finance, and other industries achieve faster, more cost-effective AI development and deployment by offering a turnkey high-performance AI compute solution.

Company Size

501-1,000

Company Stage

IPO

Headquarters

Sunnyvale, California

Founded

2016

Get referred to Cerebras

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Q1 2026 core revenue hit $191.3 million, up 92% year over year.
  • Cerebras plans 200 MW of European data centers by end-2027.
  • CrowdStrike partnership extends Cerebras inference into cybersecurity, expanding enterprise demand.

What critics are saying

  • OpenAI drives most growth; losing it would crush 2027 backlog visibility.
  • Q2 2026 core gross margin falls to 36%-38% as rented capacity depresses returns.
  • Nvidia’s CUDA ecosystem still forces custom engineering, slowing enterprise adoption.

What makes Cerebras unique

  • Wafer-Scale Engine keeps compute and memory together, beating GPU cluster latency.
  • OpenAI signed a $20 billion, 750-megawatt inference deal in December 2025.
  • AMD and Cerebras launched disaggregated inference for Cerebras Cloud in H2 2026.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Professional Development Budget

Flexible Work Hours

Remote Work Options

401(k) Company Match

401(k) Retirement Plan

Mental Health Support

Wellness Program

Paid Sick Leave

Paid Holidays

Paid Vacation

Parental Leave

Family Planning Benefits

Fertility Treatment Support

Adoption Assistance

Childcare Support

Elder Care Support

Pet Insurance

Bereavement Leave

Employee Discounts

Company Social Events

Growth & Insights and Company News

Headcount

6 month growth

-2%

1 year growth

-3%

2 year growth

0%
CNBC
Jul 23rd, 2026
Cerebras stock rises 4% on AMD partnership for ultra-low latency AI systems

Cerebras shares rose about 4% on Thursday after announcing a partnership with Advanced Micro Devices for AI systems. Cerebras CEO Andrew Feldman said at AMD's San Francisco AI conference that his company's chips will be used in AMD's Helios AI systems, to be installed in Cerebras data centres later this year. The partnership focuses on "ultra-low latency" technology, with the companies claiming their system provides five times higher tokens per second per watt than competitors. Chips like Cerebras's are designed to deliver AI answers quickly whilst making trade-offs in flexibility and power. Cerebras went public in May at $185 per share. The stock has been volatile, reaching $386.34 before falling below $161 in late June. Following Thursday's gain, shares traded at $219.80.

Associated Press
Jul 22nd, 2026
CrowdStrike partners with Cerebras to power real-time AI threat detection using world's fastest inference

CrowdStrike and Cerebras Systems announced a strategic partnership combining AI-native cybersecurity with high-speed AI inference. CrowdStrike will use Cerebras's inference technology to power its Falcon AI Detection and Response platform, whilst Cerebras will deploy CrowdStrike's Falcon platform to secure its operations. The collaboration addresses AI-accelerated cyber threats that require real-time detection and response. According to Cerebras CISO Naor Penso, inference speed is critical in cybersecurity, as delays can determine whether AI prevents an attack or merely explains it afterwards. CrowdStrike's Daniel Bernard emphasised the partnership's strategic importance, noting that leading AI infrastructure companies are choosing CrowdStrike for protection. The partnership extends CrowdStrike's leadership in AI Detection and Response whilst bringing Cerebras's inference capabilities to enterprise security applications.

Yahoo Finance
Jul 19th, 2026
Flex expands Cerebras AI accelerator manufacturing in California, stock trading 26% below fair value

Flex has expanded its manufacturing partnership with Cerebras to scale production of the CS-3 AI accelerator system at its Milpitas, California facilities. The stock has declined 12% over the past seven days and 17% over 30 days, though it remains up 46% over 90 days and 125% over one year. Flex is currently trading at $119.25, significantly below analyst fair value estimates of $160.40, suggesting potential 26% upside. The company benefits from surging demand for AI and data centre infrastructure, with its data centre segment forecast to grow 35% annually. However, Flex faces concentration risk from a small client base and operates on thin margins. The company trades at a price-to-earnings ratio of 49.6x, above both peer and industry averages.

Yahoo Finance
Jul 9th, 2026
Cerebras to build 200 MW European AI data centres by 2027

Cerebras Systems announced plans to build 200 MW of AI computing capacity across Europe, with data centres planned in France, Norway, and Finland. The first facility is expected to launch by the end of 2026, with full capacity reached by the end of 2027. CEO Andrew Feldman, speaking at the RAISE Summit in Paris, said the company is already contracting additional capacity for 2027 across the Nordic region to meet growing demand from enterprises, research institutions, and governments for sovereign AI compute. Part of the planned capacity will support OpenAI workloads under the companies' existing partnership. The expansion positions Cerebras as a regional AI infrastructure provider as governments and enterprises increasingly prioritise locally hosted computing power.

Yahoo Finance
Jul 9th, 2026
Cerebras stock falls 19% as wafer-scale AI chip maker faces delivery risks despite $20B OpenAI deal

Cerebras Systems has built its business around wafer-scale AI chip design, offering an alternative to conventional GPU-based systems. The company's Wafer-Scale Engine keeps compute and memory on a single wafer to reduce latency in AI workloads. First-quarter 2026 core revenues rose 92% year-over-year to $191.3 million, with cloud and services revenues up 167% to $79.8 million. The Zacks Consensus Estimate projects revenues of $861.3 million for 2026 and $2.77 billion for 2027. Cerebras has secured a multi-year agreement with OpenAI worth over $20 billion for 750 megawatts of inference compute. It has also partnered with Amazon Web Services to deploy its systems in AWS data centres. However, the company faces concentration risk, with major customers accounting for most revenues. Near-term gross margin compression is expected as Cerebras expands infrastructure. Cerebras currently holds a Zacks Rank 3 rating.