The b10290 release of llama.cpp brings a notable update to the ggml library, specifically with the introduction of ggml_build_forward_order. This feature improves the handling of tensor computations by allowing nodes to be inserted without the compute flag, which is only activated when necessary. This enhancement resolves issues in the mtmd audio graph, ensuring that only relevant branches are executed, thus optimizing performance. This update is particularly beneficial for developers dealing with intricate audio processing workflows.
Read originalThe latest b10278 release of llama.cpp continues its trend of broadening platform compatibility, making it a versatile choice for developers across different systems. Notably, the release includes support for ROCm 7.2 on Ubuntu x64, which is significant for AMD GPU users seeking alternatives to NVIDIA's CUDA. The inclusion of Vulkan support on both Ubuntu and Windows platforms further enhances its appeal for developers working with graphics-intensive applications. While there are no groundbreaking new features, this update solidifies llama.cpp's position as a flexible and inclusive inference runtime for diverse hardware configurations.
The b10280 release of llama.cpp marks another step in broadening its reach across various platforms, making it more adaptable for different systems. This update introduces Vulkan support on both Ubuntu and Windows, alongside ROCm 7.2 for Ubuntu, which is a significant boost for AMD GPU users. Windows x64 now benefits from the inclusion of CUDA 12 and 13 DLLs, enhancing its utility for developers. While there are no new models or quantization methods, this release reinforces llama.cpp's role as a flexible and comprehensive solution for AI inference across a wide range of hardware configurations.
© TechCrunch AIMeta is stepping up its AI game with the launch of Muse Code, a terminal coding agent designed to handle complex tasks across large software repositories. This new tool, currently in beta, leverages Meta's Muse Spark model to manage extensive projects by deploying sub-agents that work in parallel, ensuring efficient task execution without interfering with the user's working copy. By offering a cost-effective alternative to competitors like OpenAI's Codex, Meta aims to strengthen its position in the AI coding space. This move marks a significant shift as Meta expands its AI focus beyond advertising into broader enterprise applications.
© The AI Daily BriefQwen 3.8 Max has been released with open weights, featuring aggressive pricing and contested benchmark claims.