
Huawei announced at its Connect conference that the launch of its Ascend 960DT AI chip has been moved up to Q1 2027, six months ahead of schedule. The company is leveraging its Peerium Computing Architecture to scale systems like the Atlas 950 SuperCluster, which can connect up to 256,000 accelerator cards. However, analyst Rui Ma noted discrepancies in scaling claims, with some systems reportedly supporting fewer chips than previously stated. This acceleration occurs amid ongoing U.S. semiconductor restrictions and geopolitical tensions between Washington and Beijing.
Read original
© TechCrunch AIThe AI infrastructure race just got significantly more expensive. Crusoe’s $3.9 billion Series F round values the company at nearly $31 billion, signaling that capital is flowing aggressively into physical compute capacity rather than just model weights. The funds target 'Spark' modular factories—truck-deployable data centers designed to bypass local zoning battles and accelerate deployment. With major backers like Nvidia and Mubadala, this bet on hardware logistics suggests the bottleneck for AI growth is shifting from algorithms to electricity and real estate.
© TechCrunch AIGoogle DeepMind is formalizing the industry's safety anxiety with a new institute dedicated to debating AGI risks. The move signals a shift from vague concerns to concrete governance proposals, including Demis Hassabis’s call for a U.S.-led standards body that could eventually mandate pre-release model evaluations. Simultaneously, researchers argue against opaque architectures, pushing for limits on 'serial depth' to preserve interpretability. This isn't just PR; it's an attempt to set the regulatory and technical guardrails before the technology outpaces human oversight.
© TechCrunch AIPrismML is proving that extreme model compression doesn't have to mean dumb models. Their Bonsai 2 27B model shrinks Alibaba's Qwen3.8 down to just 5.9 GB using ternary weights, hitting 98% of the original benchmark scores. This isn't just a technical curiosity; it means high-performance reasoning can finally run on consumer hardware without cloud dependency. With $22.25M in seed funding and backing from Khosla Ventures, they are positioning themselves as the bridge between massive lab models and private, local inference.
This release quietly closes the hardware gap for local inference by adding native support for CUDA 13 and ROCm 10.0 alongside existing CUDA 12 builds. Users with newer NVIDIA GPUs or AMD accelerators no longer need to compile from source to get hardware acceleration, as pre-built binaries now include the necessary libraries. Apple Silicon builds have reverted KleidiAI to disabled by default, likely due to stability concerns, though it remains available. The inclusion of openEuler support for Huawei's Ascend 910b chips further expands the ecosystem beyond standard x86 and ARM architectures. This is a critical infrastructure update that ensures llama.cpp remains viable on the latest AI hardware without requiring developer intervention.
© The AI Daily BriefTypeSafe introduces Jev, a model that outputs calibrated probabilities instead of text, claiming 20-200x speed improvements over traditional LLMs for decision tasks.
© AI ExplainedOpenAI announced a partnership or tool update leveraging ChatGPT to accelerate the discovery of new antibiotics, addressing critical bottlenecks in drug development.