
A mysterious AI model called 'Ox Alpha' has launched on OpenRouter without any lab name attached, capturing the attention of developers worldwide. The model boasts a 1M-token context and multimodal input, making it suitable for coding and production workloads. It achieved an 80% score on a DeepSWE subset, though full testing placed it at 63%, comparable to Fable 5. Speculation about its origins points to China's Zhipu AI or Microsoft's MAI family. With free access for a week, Ox Alpha is drawing significant usage, potentially reshaping how developers access advanced AI capabilities.
Read original
© The Rundown AISpaceX and Nvidia are embarking on an ambitious project to establish data centers in space, with the first racks expected by late 2027. This partnership will utilize Nvidia's Vera Rubin NVL72 racks, which are designed to be simpler, more cost-effective, and lighter for space conditions. Elon Musk's vision is to leverage these space-based data centers to overcome the growing opposition to terrestrial data centers. While the cost of orbital computing is currently higher, Musk anticipates a shift in favor of space solutions in the coming years. This collaboration marks a significant step in the evolution of AI infrastructure, potentially transforming how and where data is processed.
© The Rundown AISlack is transforming coding into a collaborative effort with its new feature, Slack Code. This innovation allows team members, regardless of their technical expertise, to participate in software development directly within Slack channels. By integrating AI agents like ChatGPT and GitHub, teams can guide and monitor coding sessions in real-time. This move positions Slack as a central hub for AI-human collaboration in software development, making it easier for diverse teams to contribute and oversee projects without leaving the messaging platform.
The b10618 release of llama.cpp tackles a crucial parsing issue, specifically improving the handling of hyphens in character classes. This update ensures that generated tool-call grammars are parsed correctly, enhancing the software's reliability. With new parser and integration tests included, the release verifies these improvements effectively. While it doesn't introduce major new features, this update strengthens llama.cpp's core functionality, making it more dependable for developers working on different operating systems and hardware configurations.
The b10620 release of llama.cpp marks another step in broadening its platform reach, now supporting systems like Ubuntu with Vulkan and ROCm 7.14, alongside Windows with CUDA 13. This update underscores llama.cpp's adaptability, making it a go-to tool for developers working across various hardware configurations, from macOS Apple Silicon to Windows arm64. While the release doesn't introduce new groundbreaking features, it reinforces llama.cpp's role as a flexible inference runtime. By ensuring compatibility with more systems, llama.cpp becomes increasingly accessible to developers, allowing them to leverage its capabilities regardless of their hardware setup.
The latest llama.cpp release, version 0.3.0, brings a notable expansion in platform compatibility and functionality. This update enhances support for macOS, Linux, Windows, and openEuler, accommodating architectures like Apple Silicon and Vulkan. Developers will find the inclusion of CUDA 13 on Windows, albeit in preview, and ROCm 7.14 on Ubuntu particularly useful. While the release doesn't introduce groundbreaking features, it solidifies llama.cpp's role as a flexible tool for developers working across different computing environments. The update ensures that llama.cpp remains a reliable choice for those needing robust support across multiple systems.