
OpenAI has introduced GPT-Live 1, a new interface mode that supports real-time conversational interaction during active tasks. The primary demonstration involves building a travel planner where the user can ask questions and receive updates while the AI conducts live research. This feature bridges the gap between static prompt-response models and dynamic, multi-step agent workflows.
Read original
© Matt WolfeOpenAI launches GPT-Live 1 and a new Agent API, enabling real-time voice interactions and autonomous agent development.
© Matt WolfeGoogle updates Gemini 3.8 with Live Avatar technology and advanced Text-to-Speech capabilities.
© Matt WolfeMicrosoft introduces a new version of Copilot featuring 'Home Code' and 'Autopilot' capabilities for enhanced productivity.
© Lev SelectorAWS launched Strands Agent Harness, a tool to manage AI agents, with reports showing up to 5x cost differences based on harness choice.
© WIRED AIGoogle resurrects its decade-long ambition to automate phone calls with Call for Me, an experimental beta exclusive to Pixel 11. Unlike the failed Duplex, this iteration leverages modern LLMs to navigate hold menus and negotiate appointments in real-time. It represents a tangible shift from passive call screening to active agent execution on consumer hardware. The feature is currently limited to US English users with Gemini subscriptions, serving as a live stress test for voice-based AI agents in unstructured environments.
© The Verge AIMeta is pushing its consumer AI agent, Muse, beyond text by adding live video calling with customizable avatars. The update also grants agents their own email addresses for task execution and expands desktop control on Mac. This moves Muse from a simple chatbot toward a persistent, multimodal assistant capable of handling asynchronous communication and real-time visual interaction. It signals Meta's intent to make AI agents feel like continuous companions rather than transactional tools.