16 × AIAI signal, amplified
AI newsTopicsAboutSources
TelegramFollow on Telegram
AI newsTopicsAboutSources
16 × AIAI signal, amplified

An AI news engine that ingests trusted sources, scores with Claude, and posts only what clears the bar.

Follow on Telegram →

Subscribe

  • Telegram
  • RSS
  • All channels

Newsletter

Used only to send this newsletter. Privacy

Legal

  • Privacy
  • Imprint
© 2026 16 × AI. All rights reserved.A new issue every two days.
Home/Coding Tools
Coding Tools

llama.cpp b11537 fixes embeddings and adds ROCm 10

llama.cpp Releases·October 10, 2026·high confidence

Why it matters

  • →Fixes Gemma 4 embedding logic to prevent precision loss in vector generation.
  • →Adds ROCm 10.0 support, expanding viable hardware options for AMD GPU users.
  • →Includes CUDA 13.4 binaries, keeping the runtime aligned with latest NVIDIA driver releases.

llama.cpp has released version b11537, focusing on embedding accuracy and backend expansion. Key changes include fixes to the Gemma 4 embedding construction logic and the addition of ROCm 10.0 support for Linux and Windows. The release also includes CUDA 13.4 binaries and maintains compatibility with Vulkan, OpenVINO, and SYCL across multiple operating systems. KleidiAI on macOS Apple Silicon is currently disabled in this build.

Read original

The story around this

Earlier coverage that leads up to this article, and what followed. Lines connect each piece to the closest one after it, converging here.

llama.cpp b11040 release with CUDA 13 and ROCm 10 — llama.cpp Releases1llama.cpp b11155 adds ROCm 10 and CUDA 13 support — llama.cpp Releases2llama.cpp b11537 fixes embeddings and adds ROCm 10Sep 19You are here

How we got here

  1. 1
    llama.cpp b11040 release with CUDA 13 and ROCm 10

    llama.cpp Releases · September 19, 2026 · Same story

  2. 2
    llama.cpp b11155 adds ROCm 10 and CUDA 13 support

    llama.cpp Releases · September 24, 2026 · Same story

More from llama.cpp Releases

Coding Toolscoding

llama.cpp b11532 adds ModernBERT and CUDA 13

This release quietly fixes a critical accuracy gap for ModernBERT encoders by implementing exact GELU activation, ensuring semantic embeddings match the original PyTorch models rather than approximations. It also brings native support for CUDA 13.4 across Linux and Windows, closing the driver compatibility lag that has plagued NVIDIA users on newer hardware stacks. While KleidiAI builds are temporarily disabled on Apple Silicon, the broader expansion to ROCm 10.0 and Snapdragon NPU keeps llama.cpp as the most versatile local inference runtime available today.

llama.cpp Releases·Oct 10, 2026
Coding Toolscoding

llama.cpp b11533 fixes Adreno A6X crashes

This release targets a specific but painful stability issue for Android users running llama.cpp on Qualcomm Adreno A6X GPUs. The kernel compiler was crashing due to argument limits in the iot device backend, effectively breaking local inference on those chips. By skipping the problematic kernel and adding explicit detection for the Adreno 623, the team restores functionality where it previously failed hard. It’s a narrow fix, but essential for anyone trying to run models on mid-range Android hardware without hitting compiler errors.

llama.cpp Releases·Oct 10, 2026
Coding Toolscoding

llama.cpp b11534 optimizes CUDA SSM and adds ROCm 10

This release quietly sharpens llama.cpp’s performance on NVIDIA GPUs by fusing state snapshot copies into the recurrent cache during SSM scans. It also removes redundant CUDA copies in specific non-speculative decoding scenarios, shaving off latency where it counts. On the AMD side, ROCm 10.0 support arrives alongside stable builds for CUDA 12.8 and 13.4, keeping the library competitive across hardware vendors. KleidiAI on Apple Silicon is temporarily disabled, a minor setback for Mac users until that integration is stabilized. The net result is faster inference for SSM-based models without changing the user experience.

llama.cpp Releases·Oct 10, 2026

More in Coding Tools

Coding Toolscoding

Claude Code v2.1.292 patches security and agent logic

This release tightens the leash on Claude Code's autonomous capabilities while fixing critical sandbox escapes. The new effort parameter for Agent tools lets developers explicitly control sub-agent depth, a necessary guardrail as these systems grow more complex. Security fixes are prominent, addressing how plugins handle network paths and how file permissions persist during session resumption. It’s a stability patch that ensures the tool remains usable in enterprise environments without compromising on the new agent features.

Claude Code Releases·Oct 10, 2026
Coding Toolscoding

Claude Code v2.1.293 updates Haiku and fixes agent bugs

Anthropic quietly shipped a significant model update alongside routine maintenance. Claude Haiku 5.5 is now the default on the API, offering a 1M context window at $0.10 per million tokens, which lowers the cost floor for high-volume coding tasks. The release also patches critical stability issues in the local agent runtime, specifically fixing memory leaks in HTTP MCP connections and resolving session state corruption during context compaction. These fixes matter because they stabilize the autonomous coding workflow that developers rely on daily. With Haiku 5.5 now standard, teams can deploy cheaper, faster iterations without manual configuration.

Claude Code Releases·Oct 10, 2026
Coding Toolscoding

Claude Code v2.1.294 fixes hook logic

Anthropic quietly patched a frustrating edge case in Claude Code’s agent hooks. Previously, instructions like 'Block commands that...' were often ignored because the model didn't recognize them as valid blocking criteria. This update ensures those prompts are properly interpreted, while also refining how stop conditions are judged to prevent premature termination. It’s a small but necessary fix for anyone relying on strict guardrails in automated coding workflows.

Claude Code Releases·Oct 10, 2026