16 × AIAI signal, amplified
AI newsAboutSources
TelegramFollow on Telegram
AI newsAboutSources
16 × AIAI signal, amplified

An AI news engine that ingests trusted sources, scores with Claude, and posts only what clears the bar.

Follow on Telegram →

Subscribe

  • Telegram
  • RSS
  • All channels

Legal

  • Privacy
  • Imprint
© 2026 16 × AI. All rights reserved.Curated by Claude. Posts every 6 hours. No newsletter, no funnel.
Home/Models & Labs
Models & Labs

llama.cpp b10658 release adds DFlash2 support

llama.cpp Releases·August 28, 2026·high confidence

Why it matters

  • →DFlash2 support enhances local convolution and candidate selection, improving model performance.
  • →Code optimizations and bug fixes contribute to a more stable and efficient runtime.
  • →The update supports a wide range of platforms, increasing accessibility for developers.

The b10658 release of llama.cpp brings DFlash2 support, focusing on local convolution and candidate selector enhancements. Assisted by Claude Opus 5, the update includes optimizations like cost reduction and code refactoring, alongside bug fixes and formatting improvements. These changes aim to enhance performance and stability across multiple platforms, including macOS, Linux, Windows, and openEuler. This release signifies a step forward in making llama.cpp a more efficient and reliable tool for developers.

Read original

More from llama.cpp Releases

Open Sourcemodels

llama.cpp b10656 release expands platform support

The latest b10656 release of llama.cpp continues its trend of broadening platform compatibility, now supporting a wide array of systems including macOS, Linux, Windows, and openEuler. Notably, this update includes support for Vulkan and ROCm 7.14 on Ubuntu, as well as CUDA 13 on Windows, which enhances performance on AMD and NVIDIA GPUs. While KleidiAI support for Apple Silicon is disabled, the release still marks a significant step in making llama.cpp a versatile tool across diverse hardware configurations. This update doesn't introduce new models but solidifies llama.cpp's position as a flexible inference runtime for developers.

llama.cpp Releases·Aug 28, 2026
Models & Labsmodels

llama.cpp b10657 Release Expands Platform Support

The b10657 release of llama.cpp brings new OpenCL binary kernels, enhancing performance and compatibility across a wide range of systems. This update includes specific improvements for Apple Silicon, with KleidiAI support, and Vulkan on Ubuntu, making it more accessible for developers using these platforms. While no new model architectures are introduced, the release focuses on strengthening llama.cpp's capabilities as an inference runtime, particularly for those not using NVIDIA hardware. With ROCm 7.14 support on Ubuntu and CUDA 12 and 13 DLLs for Windows, llama.cpp continues to evolve as a versatile tool for AI model deployment. This release underscores the commitment to broadening hardware compatibility and optimizing performance across different environments.

llama.cpp Releases·Aug 28, 2026
Models & Labsmodels

llama.cpp b10659 Release Enhances Windows ROCm

The b10659 release of llama.cpp brings a crucial update for Windows users by including HIP runtime DLLs with the Windows ROCm package. This ensures that the correct HIP runtime is prioritized over the driver's version in System32, effectively solving a previous issue. Although this update doesn't introduce new model architectures or quantization techniques, it significantly enhances the platform's compatibility and performance. Developers working on AI tasks in Windows environments can now expect a more streamlined setup process and potentially better runtime performance.

llama.cpp Releases·Aug 28, 2026

More in Models & Labs

Models & Labsmodels

vLLM v0.28.0 Release: Major Performance Enhancements

The v0.28.0 release of vLLM introduces substantial improvements in performance and functionality, particularly for the Kimi-K3 model. With the addition of Decode Context Parallel support and fused FlashKDA decode kernels, the update significantly enhances processing speed and efficiency. DeepSeek V4 now includes sparse MLA support and advances in speculative decoding, offering better execution on both NVIDIA and AMD hardware. These updates make vLLM more robust and adaptable, providing developers with enhanced tools for deploying and executing models on a broader range of hardware configurations.

vLLM Releases·Aug 28, 2026
Hugging Face Launches $399 Open Source Robot, Microduck© TechCrunch AI
Models & Labsmodels

Hugging Face Launches $399 Open Source Robot, Microduck

Hugging Face has introduced the Microduck, a $399 open-source robot designed to democratize physical AI. This duck-like robot, equipped with a camera, lidar sensors, and IMUs, can perform tasks like waddling, picking up objects, and roller skating. The Microduck's behaviors can be trained in simulation and deployed directly, offering developers a platform to experiment with reinforcement learning. This launch marks Hugging Face's continued expansion into affordable AI hardware, following their acquisition of Pollen Robotics and the release of the Reachy Mini robots.

TechCrunch AI·Aug 27, 2026
Z AI Reveals Ox Alpha as GLM-5.3-Flash© The Rundown AI
Models & Labsmodels

Z AI Reveals Ox Alpha as GLM-5.3-Flash

Z AI has unveiled that the enigmatic Ox Alpha model is their latest GLM-5.3-Flash, a move that could reshape the AI landscape with its affordability and performance. Priced at just a tenth of its competitors, this model has quickly risen to prominence, topping OpenRouter's rankings. The fact that it operates entirely on Chinese-made chips suggests a significant stride in reducing dependency on foreign technology. This development not only makes advanced AI more accessible but also positions Z AI as a formidable contender in the AI market, offering a viable alternative to more expensive models.

The Rundown AI·Aug 27, 2026