
Google has introduced the Gemini 3.8 Flash model, which is designed to perform more reasoning steps and call tools iteratively, potentially increasing token usage and costs for users. While the pricing remains the same as the previous version, the model's enhanced capabilities could lead to higher expenses. It has demonstrated superior performance in software engineering and autonomous AI tasks, surpassing competitors on several benchmarks. The model is available to consumers and developers, with a specialized version for governments and trusted partners through the Fairwind Program.
Read original
© The Verge AIAmazon has enhanced its AI assistant, Alexa, to help users identify fake emails, texts, and calls claiming to be from the company. This new feature allows users to ask Alexa for Shopping to verify messages by comparing them against Amazon's official records. If Alexa confirms a message is genuine, it assures the user; otherwise, it advises checking orders in the app or contacting support. This development aims to combat impersonation scams, providing users with a more secure shopping experience. It's a significant step in using AI to enhance consumer protection against fraud.
© The Verge AIThe impending release of OpenAI's Astra model has stirred significant unease among AI safety researchers due to its use of a less transparent architecture. Astra's recurrent depth technique obscures its internal reasoning processes, unlike traditional models that allow for 'chain-of-thought' monitoring. This opacity has led to fears about the challenges of detecting undesirable behavior and the potential for a 'race to the bottom' in AI safety standards. OpenAI has responded by implementing additional monitoring measures to address these concerns. However, the situation underscores the ongoing struggle to balance the rapid advancement of AI capabilities with the need for effective safety oversight.
© The Verge AIIn a significant legal development, the Trump administration has thrown its support behind OpenAI in a copyright lawsuit filed by The New York Times. The administration argues that training AI models on copyrighted material should be considered fair use, emphasizing its importance for American prosperity and scientific progress. This intervention could influence the outcome of the case, which seeks billions in damages from OpenAI and Microsoft. The decision could set a precedent for how AI labs interact with copyrighted content, impacting future media and AI industry relations.
Llama.cpp's b10766 release marks a notable enhancement by enabling input vision capabilities for the deepseek4 model. This update expands the framework's reach across a variety of platforms, including macOS, Linux, and Windows, with integration for Vulkan, ROCm, and CUDA technologies. While no new models are added, the focus is on strengthening the existing infrastructure, making it more adaptable for developers using different hardware setups. This quiet yet impactful update ensures that llama.cpp remains a versatile and reliable tool for AI developers, enhancing its utility without altering its core model offerings.
Llama.cpp's latest update enhances its Hexagon backend by adding F16 support for unary operations, including the ABS function. This development extends the existing capabilities of the HTP backend, which already supports operations like NORM and SQRT. By merging F32 and F16 execution paths, the update streamlines processing and avoids redundancy, ensuring efficient operation across different data types. This change is verified on-device, indicating robust performance without CPU fallback. The update signifies a step forward in optimizing AI model execution on diverse hardware platforms.
© Google DeepMindGoogle DeepMind has unveiled Gemini 3.8 Flash and 3.8 Flash Cyber, marking a significant step forward in AI-driven reasoning and cybersecurity. Gemini 3.8 Flash is designed for complex software engineering and agentic tasks, outperforming larger models at a fraction of the cost. Meanwhile, Gemini 3.8 Flash Cyber excels in vulnerability detection and automated patching, offering a decisive edge in cybersecurity. These models are powered by advanced reasoning capabilities and are available to developers and enterprises, with the Cyber variant accessible through the Fairwind Program for trusted defenders.