16 × AIAI signal, amplified
AI newsTopicsAboutSources
TelegramFollow on Telegram
AI newsTopicsAboutSources
16 × AIAI signal, amplified

An AI news engine that ingests trusted sources, scores with Claude, and posts only what clears the bar.

Follow on Telegram →

Subscribe

  • Telegram
  • RSS
  • All channels

Newsletter

Used only to send this newsletter. Privacy

Legal

  • Privacy
  • Imprint
© 2026 16 × AI. All rights reserved.A new issue every two days.
Home/Models & Labs
Models & Labs

Google Gemini Live Avatars GA

Sam Witteveen·September 27, 2026·high confidence

Why it matters

  • →Enables real-time video interaction with AI agents at scale.
  • →Supports 97 languages, offering significant multilingual reach.
  • →Provides an Avatar Studio for branded digital human creation.
Google Gemini Live Avatars GA
©Sam Witteveen

Google has announced the general availability of Gemini Live Avatars, a multimodal feature within its Gemini API that enables real-time video interaction with AI agents. The service supports 97 languages and includes an Avatar Studio for customization, allowing developers to create branded digital humans for customer service or education. Pricing details have been released alongside the launch, signaling a shift from experimental preview to a commercial product ready for enterprise integration. This move expands Google's multimodal capabilities beyond text and audio into synchronized video generation.

Read original

The story around this

Earlier coverage that leads up to this article, and what followed. Lines connect each piece to the closest one after it, converging here.

Google unveils Gemini 3.5 and Omni at I/O 2026 — Google AI Blog1Google DeepMind launches Gemini 3.5 Live Translate — Google DeepMind2Google's Gemini 3.5 Introduces Live Translate — Matt Wolfe3Gemini's AI Image Generation Now Free in U.S. — TechCrunch AI4Google Vids Enhances Video Creation with AI Tools — Google AI Blog5Google Vids Adds AI Avatars and Gemini Omni — TechCrunch AI6ChatGPT and Gemini Reach 1 Billion Users — The Verge AI7Google Gemini 3.8 Flash TTS with Voice Cloning — Sam Witteveen8Google Gemini Live Avatars GAGoogle Gemini 4 Argon and OpenAI DevDay announcements — Wes Roth9Google Announces Gemini 4 Argon with 1M Token Output — Sam Witteveen10Jun 5You are hereOct 1

How we got here

  1. 1
    Google unveils Gemini 3.5 and Omni at I/O 2026

    Google AI Blog · June 5, 2026 · Related

  2. 2
    Google DeepMind launches Gemini 3.5 Live Translate

    Google DeepMind · June 9, 2026 · Related

  3. 3
    Google's Gemini 3.5 Introduces Live Translate

    Matt Wolfe · June 12, 2026 · Related

  4. 4
    Gemini's AI Image Generation Now Free in U.S.

    TechCrunch AI · June 29, 2026 · Related

  5. 5
    Google Vids Enhances Video Creation with AI Tools

    Google AI Blog · July 16, 2026 · Related

  6. 6
    Google Vids Adds AI Avatars and Gemini Omni

    TechCrunch AI · July 16, 2026 · Same story

  7. 7
    ChatGPT and Gemini Reach 1 Billion Users

    The Verge AI · August 11, 2026 · Related

  8. 8
    Google Gemini 3.8 Flash TTS with Voice Cloning

    Sam Witteveen · September 24, 2026 · Related

What happened next

  1. 9
    Google Gemini 4 Argon and OpenAI DevDay announcements

    Wes Roth · October 1, 2026 · Background

  2. 10
    Google Announces Gemini 4 Argon with 1M Token Output

    Sam Witteveen · October 1, 2026 · Background

Follow this story

Open the full story →

Google launches Gemini 3.8 Live Avatar for enterprise

4 developments

  1. Sep 24 · The Verge AI
    Google launches Gemini 3.8 Live Avatar for enterprise
  2. Sep 26 · Matt Wolfe
    Google Gemini 3.8 Adds Live Avatar and TTS↳ Gemini 3.8 adds advanced Text-to-Speech capabilities alongside the previously announced Live Avatar feature.
  3. Sep 27 · Sam Witteveen
    Google Gemini Live Avatars GA (This article)
  4. Oct 1 · The Verge AI
    Google launches Guided Vision in Gemini Live↳ Google launches Guided Vision in Gemini Live to provide real-time audio scene description for Android users with low vision via TalkBack int

More from Sam Witteveen

Google Announces Gemini 4 Argon with 1M Token Output© Sam Witteveen
Models & Labsmodels

Google Announces Gemini 4 Argon with 1M Token Output

Google is pushing the boundaries of context windows with Gemini 4 Argon, a new model capable of generating up to one million tokens in a single response. This isn't just about reading long documents; it's designed for complex agentic workflows where the AI must produce extensive codebases or detailed reports without truncation. Early benchmarks suggest it aims to reclaim top-tier intelligence status against competitors like GPT-6, specifically targeting tasks that require sustained reasoning and massive output generation. The shift from 64K caps to a million-token horizon fundamentally changes how developers might architect multi-step autonomous systems.

Sam Witteveen·Oct 1, 2026

More in Models & Labs

Google launches Guided Vision in Gemini Live© The Verge AI
Models & Labsother

Google launches Guided Vision in Gemini Live

Google is bringing real-time audio scene description to Android via Gemini Live, directly challenging Apple’s VoiceOver Live Recognition. This feature targets users with low vision by providing immediate audio cues and follow-up Q&A capabilities for physical objects. It integrates deeply into the accessibility ecosystem through TalkBack, moving beyond simple text reading to contextual environmental awareness. The move signals a shift toward multimodal AI as a standard utility for daily navigation rather than just a novelty.

The Verge AI·Oct 1, 2026
AWS releases open-source Strands Decider 2B© TechCrunch AI
Models & Labsagents

AWS releases open-source Strands Decider 2B

Amazon’s Strands Decider 2B joins the growing wave of decision models designed to replace heavy LLMs for simple routing tasks. Built on Qwen3.5-2B, it outputs calibrated choices with confidence scores rather than generating text, offering a cheaper, faster alternative for agentic workflows. The release signals AWS’s push into specialized agent infrastructure, aiming to solve the latency and cost bottlenecks of general-purpose models. While TypeSafe’s Jev pioneered this space, Amazon’s entry brings enterprise-grade credibility and open-source accessibility to a niche that is rapidly filling with experimental clones.

TechCrunch AI·Oct 1, 2026
Google Gemini 4 Argon and OpenAI DevDay announcements© Wes Roth
Models & Labsmodels

Google Gemini 4 Argon and OpenAI DevDay announcements

Google is positioning Gemini 4 Argon as a strategic comeback model, aiming to reclaim ground in the competitive landscape. Simultaneously, OpenAI’s DevDay shifts focus from pure chat interfaces toward persistent agents with Dots and broader developer tooling. These moves signal an industry-wide pivot where AI transitions from conversational novelty to embedded utility. The real test lies in whether these new architectures deliver tangible reliability or just incremental benchmark gains.

Wes Roth·Oct 1, 2026