
A new open source repository called Free Claude Code is gaining attention for its ability to run Claude Code, Codex, and other AI models from various platforms. With over 45,000 stars on GitHub, this tool allows users to choose the AI model for each task, offering options like DeepSeek, Gemini, OpenRouter, or local models. This flexibility can significantly reduce token costs while maintaining the same user interface. The repository's approach could democratize AI access, enabling developers to experiment with different models and optimize their workflows.
Read originalEarlier coverage that leads up to this article, and what followed. Lines connect each piece to the closest one after it, converging here.
IBM is finally getting first-class CI support in llama.cpp with the addition of the ZDNN backend for s390x architecture. This isn't just a minor tweak; it enables efficient inference on mainframe hardware, bridging a gap for enterprise environments that rely on IBM Z systems. While currently limited to build pipelines without automated testing, this signals a serious commitment to supporting non-x86/ARM infrastructure in the local LLM ecosystem. It’s a quiet but necessary expansion for anyone running models on legacy or specialized enterprise silicon.
© Lev SelectorNew tiny local models Bonsai 2 and Needle (8-29 MB) demonstrate that small, offline-capable AI can make fast, useful decisions.
Claude Code Releases · May 6, 2026 · Same story
WIRED AI · May 26, 2026 · Related
Duncan Rogoff · June 19, 2026 · Related
GitHub Changelog · June 22, 2026 · Related
Claude Code Releases · August 31, 2026 · Same story