Full-Time

Performance Engineer

Inference, Training & GPU

Updated on 9/11/2026

World Labs

World Labs

51-200 employees

Generates interactive 3D scenes from images

Compensation Overview

$200k - $300k/yr

San Francisco, CA, USA

In Person

Category
AI & Machine Learning (1)
Required Skills
Graphics Processing Unit (GPU)
Rust
Python
Distributed Systems
CUDA
PyTorch
Go
Observability
C/C++

Get referred to World Labs

See people who can refer or advise you

Requirements
  • Strong performance-engineering foundations, including profiling, roofline analysis, latency and throughput optimization, and disciplined root-cause investigation.
  • Deep GPU programming and optimization experience with CUDA and/or Triton, including kernel-level tuning, memory hierarchy, and bandwidth optimization at scale.
  • Hands-on experience optimizing inference and serving for large models, including batching, key-value and prompt caching, quantization, and low-latency, high-throughput sampling.
  • Hands-on experience optimizing training performance, including parallelism, distributed communication, mixed or low precision, and utilization.
  • Working knowledge of machine-learning framework internals such as PyTorch and/or JAX, including torch.compile, XLA, or similar compiler paths.
  • Strong proficiency in Python, with the ability to work in C++/CUDA and, as needed, Rust or Go.
Responsibilities
  • Optimize inference and serving end to end, including latency, throughput, batching, caching, and scheduling, to serve models efficiently at production scale.
  • Write and tune GPU kernels using CUDA and Triton for hot paths, including kernel fusion, memory- and bandwidth-bound optimization, and low-precision FP8/INT8 execution.
  • Optimize training throughput and GPU utilization through parallelism strategies, communication and compute overlap, mixed precision, and elimination of pipeline stalls.
  • Build performance models, profiling workflows, and observability for throughput, latency, cost, utilization, and their tradeoffs across the stack.
  • Own numerical correctness across precision, kernel, and hardware changes.
  • Partner with researchers to productionize models for serving and make experiments run faster and more reliably.
  • Work on distributed systems supporting training and inference when needed.
  • Profile, design, build, and ship performance-optimization code directly.
Desired Qualifications
  • Experience at an artificial-intelligence lab or machine-learning-native company optimizing systems used directly by researchers and productionizing research code.
  • Experience with low-precision and numerical methods, including FP8/INT8 quantization, mixed precision, and detection of numerical regressions across hardware platforms.
  • Experience with distributed systems for large-scale training and inference, including collective communication with NCCL, NVLink interconnects, model and tensor parallelism, and fault tolerance.
  • Experience serving generative, diffusion, video, or 3D/spatial models.
  • Multi-accelerator experience with GPU plus TPU or Trainium, and experience partnering with hardware vendors on accelerator capabilities.
  • Experience building performance-modeling and observability frameworks for GPU utilization and cost.

World Labs is building tools for spatial intelligence through Large World Models (LWMs) that can perceive, generate, and interact with 3D environments. Its main product is an AI system that can turn a single image into an entire interactive 3D scene, useful for gaming, simulation, and digital content creation. The technology works by taking a 2D image and generating full 3D representations and interactions, forming a platform or service for developers to create 3D content and experiences. Unlike many competitors, World Labs focuses specifically on zero-shot 3D scene generation from a single image and providing end-to-end 3D content creation capabilities as a scalable platform. Its goal is to enable rapid creation and manipulation of rich 3D worlds for developers and companies in gaming, simulation, and digital media.

Company Size

51-200

Company Stage

Late Stage VC

Total Funding

$1.2B

Headquarters

Stanford, California

Founded

2024

Get referred to World Labs

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • World Labs closed a $1 billion round on February 18, 2026.
  • Marble 1.1 and 1.1-Plus launched April 7, 2026, improving quality and scale.
  • SceniX acquisition on July 21, 2026 expands World Labs into robotics simulation.

What critics are saying

  • Google DeepMind's Genie, Runway GWM-1, and Decart intensify world-model competition in 2026.
  • Marble still needs paid adoption beyond early gaming, VFX, robotics, and architecture users.
  • If Autodesk integration stalls, World Labs loses its clearest enterprise distribution path.

What makes World Labs unique

  • Marble turns text, images, video into persistent 3D worlds, not flat clips.
  • Autodesk's $200 million backing links World Labs directly to design workflows.
  • World Labs exports Gaussian splats and collision meshes for visualization and simulation.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Hybrid Work Options

Growth & Insights and Company News

Headcount

6 month growth

0%

1 year growth

-11%

2 year growth

-5%
TechStartups
Jul 21st, 2026
World Labs acquires SceniX to bring generative AI into physical robotics and embodied intelligence

World Labs, the AI startup founded by Fei-Fei Li, has acquired SceniX, a robotics company specialising in high-fidelity simulation systems. The move marks World Labs' biggest step into embodied AI, focused on enabling AI systems to perceive, reason, and act in physical spaces. SceniX builds simulation platforms where robots learn skills in virtual environments before deploying them on physical hardware. This "sim-to-real" approach fits with World Labs' Marble platform, which generates persistent 3D environments from text, images, or video. World Labs raised $230 million at launch in September 2024, followed by a $1 billion round in February 2026 that valued the company at $1.23 billion. Investors include Nvidia, AMD, and Autodesk. The company aims to develop spatial intelligence systems capable of understanding geometry, movement, and physics in physical environments.

Ars Technica
Jul 13th, 2026
World models emerge as AI's next frontier, but definition remains unsettled

Leading AI companies have secured billions in funding for "world models", a new frontier that aims to simulate the physical world rather than just process language. World Labs and Advanced Machine Intelligence each raised around $1 billion, whilst Runway secured $315 million. Unlike large language models (LLMs), world models create interactive 3D environments and simulations. They're being developed for robotics, scientific research, and asset generation for games and films. However, experts disagree on the precise definition. MIT's Vincent Sitzmann describes world models as systems that simulate future events given an interaction. World Labs' Ben Mildenhall emphasises real-time, continuous spatial understanding rather than LLMs' turn-based text exchanges. Most current approaches build on video generation technology, using autoregressive diffusion to create interactive, frame-by-frame simulations rather than pre-rendered sequences.

NextTech Today
Mar 16th, 2026
Autodesk invests $200M in World Labs to develop spatial AI for the physical world

Autodesk has made a $200 million strategic investment in World Labs, the spatial AI startup co-founded by Stanford computer scientist Dr Fei-Fei Li. The investment gives Autodesk an advisory role and close collaboration on research focused on physical-world AI systems that understand space, structure, materials, physics and time. World Labs develops multimodal world models designed to understand and generate realistic three-dimensional environments. Autodesk said this aligns with demands from its core industries—architecture, engineering, construction and manufacturing—where designing requires AI that can reason in three dimensions and support iterative workflows. CEO Andrew Anagnost said the investment represents a departure from AI spending focused on ever-larger models. The partnership aims to develop AI that understands geometry and physics rather than language alone.

World Labs
Feb 19th, 2026
World Labs Announces New Funding

An update on our vision for spatial intelligence in 2026.

PR Newswire
Nov 20th, 2025
Cisco Invests in Spatial Intelligence Pioneer World Labs

World Labs is redefining what's possible by giving AI the ability to generate, reason within, and interact with 3D worlds News Summary Cisco Investments is...