Tavus offers a video personalization platform for digital marketing that uses AI to turn one recorded video into many customized videos tailored to individual customers. It clones voices and other elements to generate hundreds or millions of personalized outputs from a single video, scalable for any business size. The platform emphasizes scalable, AI-driven personalization that keeps a personal touch without creating each video from scratch. Its goal is to help businesses deepen customer connections, boost loyalty, and increase sales through personalized video messages at scale.
Company Size
51-200
Company Stage
Series A
Total Funding
$24.2M
Headquarters
San Francisco, California
Founded
2020
See people who can refer or advise you
Help us improve and share your feedback! Did you find this helpful?
Health Insurance
Unlimited Paid Time Off
Flexible Work Hours
ChatGPT Voice agent, Ray-Ban Audio & AI governance: what operators need to know. From the news archive. This article reflects its original publication date; linked sources and products may have changed. This episode examines four stories that, taken together, trace the same underlying shift: AI is moving from a tool you query to a system that acts, listens, governs itself, and teaches. Jerome covers OpenAI's voice-based agentic rollout for mobile, Meta's camera-free Ray-Ban Audio glasses, Credo AI CEO Navrina Singh's case for governance that keeps pace with innovation, and the competitive sprint among six corporate training platforms to replace passive video with live AI coaching. Each story on its own is newsworthy. As a set, they map the operational and strategic terrain that operators and investors need to navigate right now. From assistant to agent: the real meaning of ChatGPT's Voice upgrade. OpenAI's expansion of ChatGPT Voice into a mobile work agent is not a feature refresh - it is a category change. Pro and Plus users can now speak a goal and watch ChatGPT draft documents, build decks, send messages to Slack, and manage workflows without switching apps or typing a single character. The model selects from a toolbox of agentic skills autonomously, which means the cognitive load of orchestration moves from the user to the model. That shift has concrete operational implications. Businesses that have been evaluating AI as a productivity layer now have a live, voice-driven interface that can touch multiple systems in a single turn. The question for operators is no longer whether to integrate AI into workflows but how quickly they can surface the right permissions, guardrails, and data connections to make those voice-triggered actions trustworthy at scale. OpenAI also extended voice access to connected plugins for Free and Go users, which broadens the addressable user base well beyond enterprise. Hardware with a point of view: Meta's privacy-first glasses and the governance gap. Meta's Ray-Ban Audio glasses make an implicit argument: not every AI wearable needs a camera. At 43 grams, with up to 12 hours of battery and a starting price of $349, the device trades visual capture for weight, stamina, and - critically - a lower privacy threat profile. The optional Hearing Enhancement feature, cleared by the FDA as an over-the-counter hearing aid for a one-time $149 fee, positions the glasses as a medical-adjacent utility product rather than a novelty. That reframing matters for distribution, insurance conversations, and mainstream adoption. The governance story from Credo AI CEO Navrina Singh lands as a necessary counterweight to both of these hardware and software advances. Singh's core argument is that organizations accumulate 'governance debt' when AI capabilities outpace oversight structures - and the Hugging Face agent incident she cites as a cautionary example shows the risk is not hypothetical. Her spectrum-of-trust framework, which runs from internal testing through independent audits to regulatory engagement, gives operators a practical scaffold. Forward-deployed governance experts embedded inside product teams is a model that treats compliance as a design input, not an audit afterthought. The corporate training market is choosing conversation over content. The race among Tavus, Synthesia, HeyGen, Anam, ElevenLabs, and Pictory to own conversational corporate learning is a meaningful signal about where enterprise AI spend is heading. Static training video, long the default for compliance and onboarding, is being displaced by real-time AI coaching that can ground answers in a company's own documents and respond to follow-up questions live. Tavus leads this field with what it calls real-time PAL coaches, and the competitive dimensions - response latency, dialogue quality, document grounding - are exactly the metrics that determine whether employees actually retain and apply what they learn. For operators managing large, distributed workforces, the practical question is not which platform wins but whether the underlying infrastructure - permissions, HR data integrations, audit trails - is ready to support live AI coaching at scale. The platforms are moving fast; the enterprise readiness layer is often the bottleneck. Questions this episode answers. How does Credo AI's forward-deployed governance model create a defensible commercial position as enterprises face rising AI regulatory pressure? Credo AI's differentiation lies in embedding governance experts directly inside client product teams rather than selling software alone - a services-led wedge that creates stickiness as compliance requirements deepen. CEO Navrina Singh frames this around a 'spectrum of trust' that scales from internal testing through independent audits to government regulation. Enterprises that accumulate 'governance debt' - where AI capabilities outpace oversight - become high-value targets for this model. The regulatory tailwind is real, and a platform that operationalizes policy enforcement and proves compliance across every model and agent is positioned to expand wallet share as AI deployment risk rises. Which competitive dimensions in the conversational corporate training market are most likely to determine platform-level moats among Tavus, Synthesia, HeyGen, Anam, ElevenLabs, and Pictory? The six platforms competing in AI-driven corporate training - Tavus, Synthesia, HeyGen, Anam, ElevenLabs, and Pictory - are differentiating on response latency, live dialogue quality, and document grounding, according to Tavus's own competitive analysis. Document grounding is the highest-moat dimension because it requires deep integration with a client's proprietary knowledge base, raising switching costs materially. Tavus's real-time PAL coaching architecture leads the field on live conversation quality. Platforms that win enterprise procurement on document grounding and integration depth are most likely to convert pilots into multi-year contracts rather than competing on feature parity alone. Watch this episode to hear Jerome's full take on why these four stories are more connected than they appear - and to walk away with a sharper read on where AI infrastructure, hardware, governance, and enterprise learning are converging. If you're making decisions about AI adoption, deployment, or investment right now, this is the analytical frame worth having. Continue the research. Enterprise AI A convincing simulation is a starting point. The research question is whether it becomes a repeatable part of how a customer works. Sep 7, 2026 Platforms & distribution Voice inside existing productivity tools raises a useful question: does the interface create new value, or move value to an existing platform? Sep 7, 2026 Research practice A simple working method for separating what happened, what it might mean, and what still needs to be proved. Sep 7, 2026 Better questions. Better decisions. What are you seeing that Gpt3 Venture Fund should be studying? Share a company, challenge an assumption, or talk research with Gpt3 Venture Fund.
Two startups want to solve content and invoicing. ALSO: How to edit videos using ChatGPT with zero editing skills Jul 31, 2026 Welcome back, Superhuman. How bad was this month's stock selloff? Bad enough to force one of AI's most talked-about hedge funds nearly out of business. Founded by a former OpenAI employee, the fund had soared roughly 1,000% since 2024 - but that momentum reversed practically overnight. Today: Freehand chases overdue invoices, how to edit videos using ChatGPT with zero editing skills, and get the latest prompts and trending social posts. Today in AI. 1. One of the largest AI hedge funds collapses under the market selloff: Leopold Aschenbrenner - the 25-year-old former OpenAI employee turned hedge fund manager - has reportedly sold Situational Awareness's entire public portfolio to Ken Griffin's Citadel. Aschenbrenner's strategy centered on taking concentrated, leveraged positions in AI infrastructure stocks. It led to supercharged returns while the market was climbing, but reversed course quickly during the past month's rapid selloff. Read the full story. 2. Freehand raises $75M to reduce wasteful corporate spending: Freehand's AI platform helps companies avoid overpaying on invoices by monitoring every step of the process - double-checking contracts, messaging suppliers, and reconciling data. The startup claims to have recovered $260M in overpaid invoices last year across Meta, J&J, and Unilever by tracking down small amounts that most companies don't have the bandwidth to chase. See how it works here. 3. Stan launches a textable agent that creates social media content: Stanley proactively creates relevant social media posts by plugging into your most-used tools, learning your unique writing voice, and publishing across LinkedIn, X, Instagram, and other platforms. The startup claims Stanley runs all posts by you before publishing, and automatically eliminates any common tells of AI writing, like em dashes or "it's not x, it's y." Watch Stanley in action. Yesterday's most-clicked story: Tavus launched PAL Maker, a new platform that lets anyone build a digital human. Sponsored by stack AI. StackAI is the no-code platform for regulated enterprises to build and deploy AI agents. * No rip-and-replace: Connect any LLM to your entire stack, with 100+ integrations * 8 layers of governance: RBAC, analytics, built-in SDLC * White-glove experience: Your dedicated AI Transformation team optimizes all building and deployment, so CIOs get the fastest ROI Trusted by CIOs and tech leaders at LifeMD, BAE Systems, Nubank + more. The friday four. In case you missed them, this week's biggest stories. Made with Midjourney Over 1,200 employees at frontier AI companies signed a letter warning of runaway development. The open letter cautions that AI could advance "beyond our ability to understand or control the resulting systems," and calls on the US government to deliberately pace progress at the international level, which would reduce competitive pressure on each lab. Signatories include chief scientists at OpenAI, Anthropic, and Meta. The scale of the signatory list - spanning companies and roles - is a signal of how seriously insiders are taking the risk. Anthropic launched Claude Opus 5, a cheaper model that rivals its best. Claude Opus 5 matches Fable 5's performance on many benchmarks at roughly half the price, and is now live across all Anthropic platforms. Anthropic says it's especially strong at agentic coding and enterprise work, and it ships with a fast mode that runs at 2.5x normal speed. The company also published an Opus 5 Prompting Guide to help you master the new model quickly. Moonshot AI open-sourced the full weights for Kimi K3, now the largest open-weight model ever released. Kimi K3's weights and technical paper were cleared for research, personal, and commercial use, with some guardrails for commercial deployment. Demand for K3 was so intense that the lab had to pause new subscriptions to manage its compute. The model's quality and affordability will likely put pricing pressure on top US labs. Read more on Kimi K3. Fish Audio released a model that can clone voices in seconds. The startup publicly launched its flagship model, S2.1 Pro, and its founders demoed the tech by cloning their own voices live, in real time. It just announced $52M in seed funding to cap off its first year as a company. Try it yourself. Readers also love this guide to the new rules for context engineering from an Anthropic engineer, as well as a viral video of a Senator accidentally reading an AI prompt as part of his speech. In the know. What's trending on socials & headlines today. Meme of the day Improved Writing: Want your AI to write more clearly? This post has racked up 11K bookmarks with a simple prompt that eliminates most signs of AI writing. Just upload it before every conversation or permanently add it to your LLM's instruction file. The Big Leagues: A 5-step interview process can feel excessive. But compared to OpenAI's, that's nothing. One candidate leaked what it's like to interview at the top lab, but let's be honest - OpenAI's median compensation is definitely worth the time commitment (1M views). Seems Sloppy: LinkedIn is testing a "seems like AI slop" button, letting you report any post that just feels off. Given the platform's reputation, expect heavy use. See which other social platforms suffer the most from AI slop. Uncanny Valley: One Redditor had ChatGPT create an uncanny image from the "deepest corner" of its mind. The original poster said they loved and hated what it created at the same time. View the image here (1.9K upvotes). Minimizing Screens: One dad wanted his toddlers to be able to talk with their grandparents whenever they wanted, without giving them a cell phone. He vibe-coded an out-of-the-box solution that lots of people are eager to replicate (1M views). Sponsored by Nebius. Most teams can demo an AI agent. Far fewer can make one reliable, observable and cost-effective in production. On August 4, join Nebius, LangChain and Tavily for a live compliance-audit agent build with Deep Agents and NVIDIA Nemotron 3 Ultra, no fine-tuning required. See how to reach frontier-level quality at roughly 1/10 the cost. Productivity. * | Napkin: Turns your written text into diagrams, infographics, and charts. * | Lindy: Automate scheduling, email, and daily tasks with AI assistant. * | NxCode: Build production-ready apps without coding. Tutorial. How to edit videos using ChatGPT with zero editing skills. * Download the ChatGPT desktop app and sign in * Select Codex from the sidebar and start a new chat Prompt: "Read chatcut.io/chatgpt and install the ChatCut plugin" * Create a free ChatCut account, sign in and connect your account with Codex * Upload your video and any extra assets * Describe the edits you want, then refine them through chat without starting over Prompt: "@Chatcut Edit this video into a polished, engaging final cut. Remove pauses and mistakes, tighten the pacing, add smooth transitions, improve the audio and use the uploaded clips, images, and logo where appropriate" * Once you are happy with the result, type: "Export the final video as a high-quality MP4" Image. Friday fun. Which one is AI generated? Extras. Want more? * Check out its Top 125 AI Tools * Get better at prompting with its Top 1,000 Prompts * Learn how to use AI at work with its 50 Tutorials Grow customers & revenue: Join companies like Amazon, HubSpot, and Salesforce. Showcase your product to its 4 million+ readers and followers on socials. Get in touch. What did you think of today's email? Your feedback helps me create better emails for you! Until next time - Zain, Theodore, & the Superhuman AI team.
Tavus has launched Phoenix-4, a real-time behaviour generation engine that creates emotionally responsive AI avatars for live conversations. The San Francisco-based company claims it is the first real-time model to generate and control emotional states, active listening behaviour and continuous facial motion as a unified system. Phoenix-4 runs at 40 frames per second in 1080p and generates every pixel from head to shoulders, including eye blinks. The model offers explicit emotion control across 10+ states, including happiness, sadness and anger, and can be guided through prompts or respond contextually. It also features context-aware active listening with visual backchannels like nods and reactive expressions. The system is available today through Tavus' platform, APIs and updated Stock Replica library featuring over 40 new replicas.
Tavus, a San Francisco-based AI company, has launched Raven-1, a multimodal perception system that enables AI to understand emotion, intent and context by interpreting audio and visual signals together. The system captures tone, facial expressions, posture and gaze to produce natural language descriptions of emotional states. Unlike traditional systems that convert speech into transcripts, Raven-1 fuses audio-visual signals into unified representations that language models can process directly. The system operates with sub-100 millisecond audio perception latency and combined pipeline latency under 600 milliseconds. Raven-1 is now generally available across all Tavus conversations and APIs. The company previously launched Sparrow-1, a conversational timing model, and offers both developer APIs and PALs, a consumer platform for AI agents.
AI startup Tavus founder says users talk to its AI Santa 'for hours' per day. A new helper has arrived at the North Pole in recent years: AI. Tavus, the AI startup that creates digital replicas using voice and face cloning technology, has launched its AI Santa experience for the second year in a row. This allows parents and children to video chat with a virtual version of the jolly old Saint Nick. After signing up for a free account, users can interact with AI Santa via text, phone, or video chat. Users can tell AI Santa what they want for Christmas, share their holiday plans, and find out if they're on the naughty or nice list. This year, the company debuted an improved version of AI Santa, designed to be more expressive and emotionally aware. Santa is now a "Tavus PAL," the company's name for its real-time AI agents that are built to see, hear, respond, and appear human. AI Santa can now see users' expressions and gestures and respond to them. It also remembers users' conversations and interests, creating a more personalized experience. Notably, it now can take actions of its own, including searching the web for present ideas or even perform everyday tasks like drafting emails. During testing, the conversation with AI Santa was engaging for the most part. When we mentioned wanting a new PlayStation for Christmas, Santa followed up with questions about our favorite video games, showing knowledge of specific titles like Baldur's Gate 3. It also smiled back when we did. (We didn't like that part very much, but maybe others will.) Users appear to be enjoying the improved experience so far. Founder and CEO Hassaan Raza said that many people are engaging with the platform frequently, spending hours chatting with AI Santa and often reaching their daily limits. Join the disrupt 2026 waitlist. "Last year's AI Santa drew millions of hits, and we're on pace to surpass that by a wide margin as Christmas approaches," he noted. While this level of engagement marks a milestone for Tavus, it also raises questions about the impact of such interactions, especially for young children. Children may struggle to distinguish between AI and a real person. Spending hours in conversation with an AI has already been linked to negative effects in adults, making the potential effects on children who strongly believe in Santa a concern for some parents. During our testing, there were subtle cues that the AI Santa does yet appear fully human-like, such as long pauses and a flat voice. We also found that if a user were to question whether it's real, the programmed response was: "I'm an AI Santa powered by Tavus' magic and technology. I might not be the physical Santa, but I've got the spirit and the cheer." Still, the experience launches amid growing concerns about AI's effects on young users. There have been reports linking chatbot interactions to serious harm, including cases where chatbots were implicated in the suicide deaths of teenagers. Character.AI removed access to its chatbots for users under 18 in October. Raza emphasized that the AI Santa experience is designed for families to enjoy together, with safety measures in place to ensure appropriate interactions. Safety features, such as content filters, have been implemented to maintain family-friendly discussions. In certain situations, conversations can be terminated, and users are directed to mental health resources if necessary. "The vast majority of interactions have been family-friendly and true to the Santa experience," he said. Additionally, when asked about data collection, Raza said the company "collects logs, session timestamps, metadata, and other information users choose to share during their chats. This data is used to provide and maintain a safe experience, and users can request data deletion at any point in time." Lauren covers media, streaming, apps and platforms at TechCrunch. You can contact or verify outreach from Lauren by emailing [email protected] or via encrypted message at laurenforris22.25 on Signal. Plan ahead for the 2026 StrictlyVC events. Hear straight-from-the-source candid insights in on-stage fireside sessions and meet the builders and backers shaping the industry. Join the waitlist to get first access to the lowest-priced tickets and important updates.