Full-Time

Senior DevOps Engineer

Runware

Runware

51-200 employees

Generative AI platform with fast inference

No salary listed

Remote in UK

Remote

Remote within the United Kingdom, with in-person meetings twice a year.

Category
DevOps & Infrastructure (1)
Required Skills
TCP/IP
Incident Response
CUDA
Computer Networking
Infrastructure as Code (IaC)
Docker
Observability
DevOps
Linux/Unix
Serverless

Get referred to Runware

See people who can refer or advise you

Requirements
  • Strong experience as a DevOps Engineer, Site Reliability Engineer, Infrastructure Engineer, Platform Engineer or similar, with a track record of running production systems at scale.
  • Deep Linux knowledge and confidence debugging real production issues across networking, storage, performance, services and system behaviour.
  • Hands-on experience building automation, Infrastructure-as-Code, continuous integration and continuous delivery pipelines, and deployment workflows that make infrastructure safer and easier to operate.
  • Experience operating high-availability, low-latency or high-throughput platforms where reliability and performance directly affect customers.
  • Strong networking fundamentals across TCP/IP, DNS, load balancing, routing, firewalls, proxies, TLS and HTTP.
  • A calm and pragmatic approach under pressure, with strong communication, good judgement and a bias toward automation over manual toil.
Responsibilities
  • Build and scale the infrastructure that powers real-time AI inference across GPU fleets, bare-metal servers, serverless and containerised production systems.
  • Help evolve the platform toward more elastic, on-demand infrastructure that can scale quickly with customer traffic and model demand.
  • Improve the critical paths behind request entrypoints, inference services, queues, storage, load balancers and the networking layer to make the platform faster, more reliable and more resilient.
  • Automate infrastructure operations from provisioning and configuration through continuous integration and continuous delivery, deployment safety, progressive rollouts and rapid rollback.
  • Build the observability backbone for a high-performance AI platform, with signals needed to spot issues early, understand capacity and fix problems before customers feel them.
  • Play a leading role in production operations, incident response, debugging and post-incident improvements.
  • Strengthen the security and compliance foundations of the infrastructure through patching, secrets management, access controls, hardening, auditability, documentation and repeatable operational processes.
Desired Qualifications
  • Experience operating GPU infrastructure for AI/ML inference, including NVIDIA drivers, CUDA, container runtimes, GPU monitoring, capacity planning and workload isolation.
  • Familiarity with inference serving and optimisation frameworks such as vLLM, TensorRT, Triton or similar.

Runware.ai provides an ultra-fast, cost-effective platform to deploy AI-generated content. It runs AI workloads on custom hardware in renewable-powered, self-managed data centers and offers a flexible API so apps can use AI features without extra infrastructure or ML expertise. The Sonic Inference Engine powers rapid media generation, with real-time WebSocket connections, built-in redundancy, autoscaling, GPU allocation, and parallel pipelines to ensure reliability. A library of over 188,000 models (including LoRA and ControlNet) supports diverse media tasks, with a goal of democratizing access to advanced generative AI for businesses.

Company Size

51-200

Company Stage

Series A

Total Funding

$66M

Headquarters

London, United Kingdom

Founded

2023

Get referred to Runware

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • December 2025 raised $50 million, extending runway for pod deployment and model expansion.
  • Runware said revenue grew 10x since Q3 2025, signaling strong demand.
  • 2026 hiring for engineering, DevOps, and sales shows active scaling across product and GTM.

What critics are saying

  • April 18, 2026 partial API outage exposed reliability risk in production.
  • July 2026 upstream data-center outage showed infrastructure dependence on third-party providers.
  • Open Enterprise AE hiring signals a hard GTM push without obvious category lock-in.

What makes Runware unique

  • Runware owns Sonic Inference Engine and custom GPU pods for low-latency inference.
  • December 2025 Series A from Dawn, Insight, Comcast, and a16z validates infrastructure ambition.
  • Runware claims 5 billion creations and 100,000 developers since launch.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Health Insurance

Remote Work Options

Flexible Work Hours

Unlimited Paid Time Off

Paid Vacation

Paid Holidays

Family Leave

Stock Options

Meaningful stock options

Company Retreats

401(k) Retirement Plan

401(k) Company Match

Wellness Program

Mental Health Support

Gym Membership

Phone/Internet Stipend

Growth & Insights and Company News

Headcount

6 month growth

8%

1 year growth

0%

2 year growth

12%
TechCrunch
Aug 4th, 2026
Runware raises $50M to build portable data center pods for AI inference

Runware has launched the Sonic Inference Pod, a modular, transportable data centre designed to provide flexible AI compute capacity. The company believes distributed compute positioned closer to end users will win long-term over massive hyperscaler projects. The pods use closed-loop cooling rather than water and can be built in days, compared to months or years for traditional data centres. Runware can add capacity quickly by deploying new pods rather than expanding fixed facilities. The company currently has 10 pods deployed across the US, Europe, and Asia-Pacific, with 160 sites available. Runware raised $50 million in Series A funding in December to provide infrastructure for image generation. CEO Flaviu Radulescu acknowledged AI power consumption will increase regardless of supplier, but emphasised the pods avoid transmission losses and new grid capacity requirements.

TechCrunch
Dec 11th, 2025
Runware raises $50M Series A for real-time image and video generation API

Runware, a developer platform for real-time image, video and audio generation, has raised $50 million in a Series A round led by Dawn Capital, with participation from Insight Partners and a16z Speedrun. Founded in 2023 by Flaviu Radulesc and Ioana Hreninciuc, Runware enables developers to integrate its API into applications to generate media assets through a single interface, eliminating the need for separate infrastructure or integrations. The platform uses custom AI inference infrastructure for open-source models and provides immediate access to newly released models. The company has powered over 5 billion creations for more than 100,000 developers since launch. Dawn Capital Partner Shamillah Bankiya is joining Runware's board following the investment.

a16z
Nov 21st, 2025
a16z speedrun

We invest up to $1M in your new startup.

Tech Funding News
Sep 9th, 2025
Runware raises $13M for AI media

Runware, an AI-as-a-Service provider based in San Francisco, raised $13M in a funding round led by Insight Partners, bringing its total funding to $16M. Investors include a16z Speedrun, Begin Capital, and Zero Prime. The funds will expand Runware's capabilities to all-media workflows, including audio, LLM, and 3D. Founder Flaviu Radulescu highlighted their platform's ability to offer up to 90% lower inference costs than cloud providers. Runware evolved from the image creation tool PicFinder.

Citybiz
Oct 2nd, 2024
Runware Secures $3M Funding

Runware, a San Francisco, CA-based media generation API startup, raised $3m in seed funding. Backers included A16Z’s Speedrun, LakeStar’s Halo II , Lunar Ventures, Begin Capital, Zero... Read More