
Work Here?
ElevenLabs builds AI-powered audio tools for speech, conversation, and creative audio used by enterprises and creators. It offers three platforms: ElevenAgents for enterprise voice and chat agents, ElevenCreative for multichannel audio localization, and ElevenAPI for low-latency voice infrastructure. The products rely on AI voice models to generate and process speech, with each platform serving its own use case—conversational agents, localization pipelines, and fast API-based voice processing. The company differentiates itself with an integrated suite that combines enterprise-grade agents, global audio localization, and developer-friendly, low-latency infrastructure, backed by strong funding and partnerships to scale research and go-to-market internationally.
Industries
Consumer Software
Enterprise Software
AI & Machine Learning
Entertainment
Company Size
1,001-5,000
Company Stage
Series E
Total Funding
$1.3B
Headquarters
London, United Kingdom
Founded
2022
See people who can refer or advise you
Help us improve and share your feedback! Did you find this helpful?
Total Funding
$1.3B
Above
Industry Average
Funded Over
8 Rounds
Industry standards
Remote Work Options
Flexible Work Hours
Professional Development Budget
ElevenLabs launched two new speech models on Monday: ElevenLabs v4 and v4 Turbo. The models offer enhanced expression control, support for over 90 languages, and lower latency for voice agents. The v4 generation uses a new architecture enabling faster voice cloning with just 10 seconds of audio. The model maintains voice identity better over longer text and adjusts expressions based on context. Users can now stack multiple expression tags for more nuanced control. Language support expanded from 70 to 90 languages, with significant quality improvements in Japanese, Brazilian Portuguese, Mandarin, and Cantonese. The model can begin generating audio immediately as the underlying language model starts producing responses. Enterprise calling now comprises over 55% of ElevenLabs' business. The company's annualised revenue has grown from roughly $330 million at year-start to over $600 million, with headcount exceeding 800 employees.
ElevenLabs revolutionizes AI audio with expanding enterprise clientele. ElevenLabs is pioneering the development of voice layer AI that converts text into human-like speech. This technology is widely utilized in customer service by major companies like Klarna, Deutsche Telekom, Cisco, Adobe, and various governments. - Enterprise Use: Around 55% of ElevenLabs' $600 million annual revenue comes from enterprise clients. The remainder largely consists of small to medium businesses and creators using the platform for audiobooks and more. - Industry Competition: Competitors like Decagon are emerging, having used ElevenLabs' technology to train their products. Despite this, ElevenLabs remains optimistic, with a current valuation of $22 billion. - Technological Advancements: Co-founder Mati Staniszewski discussed the evolution towards passing the Turing test in conversational AI, highlighting the need for emotional understanding in AI models. - Market Adaptation: The company recognizes the blurred lines between model, platform, and application companies, indicating a shift in market dynamics. - Ethical Considerations: Staniszewski advocates for transparency when AI is involved in communications, predicting societal shifts in these expectations over the next five years. - Research and Development: The company actively works on fine-tuning models and annotating data to ensure their software meets client-specific needs, employing a large team to handle data annotation. - Future Outlook: An IPO is considered within the next five years, contingent on strategic and market conditions. Precautions are taken against cybersecurity risks, with all customers undergoing KYC processes.
AI daily|qwen-image-2.1 released; Anthropic establishes physical wet lab; Kimi Code Desktop 1.0. September 22, 2026 Updated Sep 22 Model releases & updates. Qwen-Image-2.1 - Alibaba Qwen. * TL;DR: Alibaba has released Qwen-Image-2.1, a 7B parameter native text-to-image and editing model that unifies generation, multi-image editing, and transparent RGBA output within a single checkpoint. * Key Highlights: * Supports up to 10 reference images per request alongside built-in prompt expansion powered by an internal LLM. * Achieves high-speed local inference (1024x1024 generation in under 9 seconds on high-end enterprise hardware and optimized throughput on single RTX 4090 setups). * Comes with day-0 integration in SGLang-Diffusion and is fully available on Hugging Face. * Specs: 7B Parameters / Open Weights / Hugging Face & SGLang-Diffusion * Links: Qwen-Image-2.1 Blog Step 5 Preview - StepFun. * Step 5 Preview: StepFun has rolled out its Step 5 Preview model, delivering exceptional cost-to-capability performance on the Pareto frontier for agentic coding workflows. * Key Highlights: * Demonstrates robust performance in complex multi-file bug fixing and backend refactoring, rivaling larger baseline models. * Features precise stop conditions that prevent runaway token generation during long-horizon tasks. * Available immediately for developer testing across supported platforms and API gateways. * Specs: Frontier Preview / API Accessible / StepFun Platform * Links: StepFun Step 5 Preview Product releases & updates. Kimi Code Desktop 1.0 - Moonshot AI. * What's New: Moonshot AI has officially launched Kimi Code Desktop 1.0 for macOS and Windows. The official desktop client brings its autonomous coding agent capabilities out of the command line and into a visual graphical workspace, complete with a built-in terminal, browser inspection tools, and native Git status tracking. * Who It's For: Software engineers, indie developers, and technical founders managing multi-step software development projects. * Try It: Kimi Code Desktop Portal Studio 4.0 - ElevenLabs. * What's New: ElevenLabs has introduced Studio 4.0 within ElevenCreative, upgrading its AI-native video editor with synchronized generation tools for video, image, voice, music, and sound effects, alongside centralized team commenting features. * Who It's For: Creators, video editors, and product marketers building multi-modal content pipelines. * Try It: Industry News. Anthropic establishes Bay Area biological wet lab - Anthropic. * What's New: Anthropic has quietly established a physical biological wet lab in the Bay Area, expanding its life-sciences strategy beyond computer simulations into physical experimentation. The company aims to enable Claude to direct laboratory robots with minimal human intervention to accelerate preclinical research. * Why It Matters: Marks a significant strategic move by frontier AI labs into physical scientific infrastructure, targeting rare and traditionally "undruggable" diseases while intentionally focusing on preclinical studies to avoid direct competition with traditional pharmaceutical clinical trials. * Source: Reuters / Financial News Reports Research papers. RetroChimera: improving synthesis prediction of small molecules at scale - Microsoft Research. * Motivation: Planning chemical synthesis routes for new small molecules remains largely manual, time-consuming, and costly in drug discovery and advanced materials engineering. * Key Innovation: Introduced RetroChimera (recently published in Nature), a retrosynthesis model that combines two complementary architectures and learns to dynamically rank their reaction proposals for superior predictive accuracy. * Results: Outperforms individual base models in blind evaluations, with PhD-level chemists consistently preferring its reaction predictions and demonstrating robust zero-shot transfer on diverse chemical datasets. * Paper: Microsoft Research Blog Other Highlights. Rebalancer: Meta open-sources hyperscale resource allocation library. * Overview: Meta has open-sourced Rebalancer, the high-performance assignment-problem solver used internally across its hyperscale datacenters for over nine years. The library cleanly separates problem specification, memory storage, solving algorithms, and debugging to optimize resource placement across electrical fault domains. * Link: Meta Engineering Blog Experience Scribis: ultimate AI audio & video workflow. Scribis is an all-in-one AI audio tool designed for creators, researchers, and professionals. It combines local privacy computation with cloud API integration, providing millisecond-level speech-to-text, speech synthesis, and dynamic timeline editing.
In today’s digest, President Trump heads to Gracie Mansion, NYC braces for UN gridlock, and AI for Impact gets a fresh cookbook. 🧑🍳
ElevenLabs announces native Text-to-Sound Effect API integration with Unreal Engine 5. Ethan Walker September 20, 2026 ~472 words ElevenLabs, the undisputed leader in AI voice synthesis, has significantly expanded its footprint in the video game development industry. Today, the company announced the launch of a native Text-to-Sound Effect (T2SE) API plugin for Unreal Engine 5, fundamentally altering how sound designers and indie developers populate audio environments in large-scale virtual worlds. Historically, Foley art and sound design have been massive bottlenecks in game development. If a developer wanted the sound of "a heavy steel broadsword scraping against wet cobblestone," they had to either painstakingly record it in a physical studio, purchase expensive pre-recorded asset packs and layer them, or spend hours tweaking synthesizers. With the new ElevenLabs UE5 integration, this process is reduced to a single text prompt within the engine's editor. Procedural audio generation at scale. The plugin integrates seamlessly into the Unreal Engine 5 Audio Engine (MetaSounds). A developer can simply highlight a collision event in their Blueprints graph - for example, a wooden barrel breaking - and type a natural language prompt directly into the node: "A hollow wooden barrel shattering into dozens of splinters on a dirt floor, highly detailed, close up." The ElevenLabs API hits their proprietary sound-effects model and returns a high-fidelity, uncompressed .wav file in under 800 milliseconds. More importantly, the API supports seed randomization. This means a developer can generate 50 unique variations of that exact same barrel breaking sound with a single click, instantly solving the "machine gun effect" where players hear the exact same audio file repeatedly during gameplay, breaking immersion. Dynamic in-game synthesis. While the plugin is heavily marketed toward pre-rendering assets during the development phase, the true breakthrough lies in its runtime capabilities. The ElevenLabs UE5 plugin allows for dynamic, on-the-fly audio generation during live gameplay, provided the player is connected to the internet. Imagine an RPG where a player crafts a completely unique weapon made of glass and titanium. Rather than relying on a generic "sword swing" sound, the game engine can send the weapon's specific material properties to the ElevenLabs API in real-time, generating a completely bespoke sound effect for that specific player's action. The impact on indie developers. "Sound design is often the first thing cut from an indie game's budget, and it is usually the most obvious indicator of a lack of polish," noted Mati Staniszewski, CEO of ElevenLabs, during the announcement. "By bringing our Text-to-Sound model natively into Unreal, a solo developer can now orchestrate Hollywood-grade Foley art for an entire 100-hour open-world game in a single afternoon." The plugin is currently available for early access to developers on the ElevenLabs Pro tier, with a full rollout to the Unreal Marketplace scheduled for next month. The pricing model operates on a standard character-usage tier, though ElevenLabs has introduced a new "audio-second" metric specifically for non-speech sound effect generation to keep costs predictable for large studios. Found this useful? Share it: Prefer NeedAITool on Google SearchAI Overviews See its verified benchmarks & AI tool comparisons more frequently on Google. Ethan Walker I'm a technology writer passionate about AI tools, automation, productivity software, and emerging SaaS platforms. I spend my time testing digital tools and breaking down complex technologies into practical insights that help businesses, creators, and professionals work smarter.
Find jobs on Simplify and start your career today
Industries
Consumer Software
Enterprise Software
AI & Machine Learning
Entertainment
Company Size
1,001-5,000
Company Stage
Series E
Total Funding
$1.3B
Headquarters
London, United Kingdom
Founded
2022
Find jobs on Simplify and start your career today