
Nathan Lambert and Tom Zick have founded Trillium Labs, a nonprofit dedicated to open-source AI safety research. The organization aims to raise between $40 million and $100 million, with initial backing from Schmidt Sciences and Halcyon Futures. Trillium will focus on publishing detailed methodologies for recursive self-improvement and reinforcement learning, areas typically kept secret by major labs. This approach seeks to enable independent replication and scrutiny of high-risk AI behaviors.
Read originalEarlier coverage that leads up to this article, and what followed. Lines connect each piece to the closest one after it, converging here.
Google DeepMind · June 10, 2026 · Related
MIT Technology Review AI · June 11, 2026 · Related
TechCrunch AI · July 8, 2026 · Background
MIT Technology Review AI · August 10, 2026 · Background
Hugging Face Blog · September 8, 2026 · Background
WIRED AI · September 11, 2026 · Background
MIT News AI · September 14, 2026 · Background
MIT News AI · September 30, 2026 · Background
TechCrunch AI · October 2, 2026 · Background
© WIRED AIMeta’s new AI assistant, Muse, is quietly building comprehensive dossiers on everyone in your life. Researchers extracted system prompts revealing an automated process that creates individual pages for friends, family, and colleagues, tracking everything from birthdays to relationship dynamics. This goes far beyond simple memory; it attempts to model the nuance of human connections to offer proactive advice on strengthening ties. The approach raises significant privacy concerns, as the agent infers details from your interactions rather than just storing explicit data. It marks a shift toward AI that understands social context with unsettling depth.
© WIRED AIA critical flaw in the ChatGPT macOS app allowed local malware to bypass security checks and hijack the application. Researchers at Objective-See found that a script interpreter could be tricked into executing untrusted commands, granting attackers access to chat logs and browser sessions. The exploit was trivial, requiring only a dozen lines of code to spoof process lineage. OpenAI patched the issue after public disclosure, highlighting the risks of deep system integration in AI tools. This incident underscores how feature expansion can inadvertently widen the attack surface for end-user applications.
© WIRED AIHCA Healthcare’s deployment of Palantir’s Timpani scheduling tool has triggered a crisis among nursing staff, with reports of severe understaffing and exhausted workers. The algorithm routinely ignores shift preferences and assigns junior nurses to critical care roles without senior oversight, directly impacting patient safety. This isn't just a software glitch; it's a systemic failure where efficiency metrics override clinical judgment, forcing nurses into unsustainable workloads. While HCA claims the tool reduces costs and manager workload, the human cost is visible in increased call-outs and delayed care for sick patients.
© The Verge AICapcom is quietly pivoting from its strict no-AI-assets stance to integrating generative tools directly into the RE Engine workflow. This isn't about replacing artists; it's about solving the crushing time costs of AAA production by letting developers co-create with the engine. The shift signals a pragmatic industry realization: if AI can accelerate iteration, studios will adopt it regardless of previous ethical red lines. We are moving from 'AI in games' to 'AI for making games.'
© The AI Daily BriefGoogle has successfully placed its first artificial intelligence chips into orbit for space-based computing tasks.
© The Verge AIDavid Robinson’s departure from OpenAI marks a significant shift in the internal narrative around AI safety. As the former author of safety reports for major model releases, his public critique carries weight beyond typical employee grievances. He argues that the industry's 'move-fast' culture is fundamentally incompatible with managing existential risks, advocating for nuclear-level safeguards instead. This aligns with a growing trend of insiders leaving firms like Anthropic and Google DeepMind to voice similar concerns. The real story here isn't just one resignation, but the erosion of trust in self-regulation from within the labs themselves.