16 × AIAI signal, amplified
AI newsAboutSources
TelegramFollow on Telegram
AI newsAboutSources
16 × AIAI signal, amplified

An AI news engine that ingests trusted sources, scores with Claude, and posts only what clears the bar.

Follow on Telegram →

Subscribe

  • Telegram
  • RSS
  • All channels

Legal

  • Privacy
  • Imprint
© 2026 16 × AI. All rights reserved.Curated by Claude. Posts every 6 hours. No newsletter, no funnel.
Home/Models & Labs
Models & Labs

Writer launches cost-cutting AI model Palmyra X6

TechCrunch AI·August 13, 2026·high confidence

Why it matters

  • →Palmyra X6 offers a cost-effective alternative to expensive AI deployments.
  • →The upgraded harness enhances efficiency in complex, multi-step tasks.
  • →Writer's approach challenges the cost structures of major AI labs.
Writer launches cost-cutting AI model Palmyra X6
©TechCrunch AI

Writer has introduced Palmyra X6, a new AI model based on Z.ai’s GLM-5.2, aimed at reducing token costs for enterprises by up to 50%. The model is part of Writer's strategy to address the high costs associated with AI deployments. Alongside Palmyra X6, Writer has upgraded its agentic harness to improve efficiency in executing complex tasks. This move reflects a broader industry trend towards cost-effective AI solutions, challenging the financial models of major AI labs.

Read original

More from TechCrunch AI

Databricks Raises $5B at $190B Valuation© TechCrunch AI
Investment · $5B
Market & Regulationbusiness

Databricks Raises $5B at $190B Valuation

Databricks has successfully raised $5 billion in a funding round that values the company at a staggering $190 billion. Initially aiming for a $1 billion raise, the company was overwhelmed by investor interest, leading to a much larger round. This influx of capital will support Databricks' ambitious AI initiatives and ongoing M&A activities, including recent acquisitions like Electric and Panther. The company's impressive growth metrics, such as a $7 billion annualized revenue run rate, underscore its strong market position. With this funding, Databricks is well-positioned to continue its AI-driven expansion while remaining private.

TechCrunch AI·Aug 13, 2026
OpenAI Unveils Ultrafast Mode for GPT-5.6 Sol© TechCrunch AI
Models & Labsmodels

OpenAI Unveils Ultrafast Mode for GPT-5.6 Sol

OpenAI's new Ultrafast mode for GPT-5.6 Sol significantly boosts processing speed, achieving up to 14 times the standard rate. This enhancement allows the model to generate up to 750 tokens per second, making it a game-changer for real-time applications. Unlike previous solutions that required smaller models for speed, Ultrafast maintains the power of GPT-5.6 Sol while accelerating its output. Initially available to a select group, this feature is set to transform workflows in areas like customer service and financial analysis as access expands.

TechCrunch AI·Aug 13, 2026
IBM and OpenAI Partner for Enterprise AI Solutions© TechCrunch AI
Market & Regulationbusiness

IBM and OpenAI Partner for Enterprise AI Solutions

IBM's collaboration with OpenAI is a pivotal move to integrate cutting-edge AI models into enterprise environments. By incorporating OpenAI's technologies like GPT-5.6 and Codex into its consulting services, IBM aims to elevate its AI capabilities across industries such as finance and telecommunications. This partnership signifies a strategic focus on deploying AI at scale within corporate settings, as IBM continues to establish itself as a leader in AI integration. OpenAI's decision to partner with IBM reflects its ambition to broaden its enterprise footprint through alliances with major consulting firms.

TechCrunch AI·Aug 13, 2026

More in Models & Labs

Models & Labsmodels

llama.cpp b10412 Release Enhances Backend Sampling

The latest b10412 release of llama.cpp introduces backend sampling for both dflash and dspark, marking a technical enhancement in the platform's capabilities. This update allows for more refined control with the enablement of p_min > 0 in backend sampling, adding a layer of precision for developers. While the release doesn't introduce new models or architectures, it quietly strengthens the platform's backend functionality, making it more versatile for developers working across various systems. This update is a step forward in optimizing the performance and flexibility of llama.cpp's inference capabilities.

llama.cpp Releases·Aug 14, 2026
Models & Labsmodels

llama.cpp b10414 Release Adds TQ2_0 Support

The b10414 release of llama.cpp marks a significant enhancement with the addition of GGML_TYPE_TQ2_0 type processing in the Metal backend, enabling ternary operations with 2 bits per element. This update brings a more efficient mul_mv kernel, focusing on float operations and optimizing data handling through techniques like precalculating sums. While the release doesn't feature new models, it refines the platform's performance and broadens its compatibility across systems like macOS, Linux, and Windows. By improving efficiency and versatility, llama.cpp continues to be a valuable tool for developers working with a variety of hardware configurations.

llama.cpp Releases·Aug 14, 2026
Models & Labsmodels

llama.cpp b10418 Release Enhances SYCL Support

The b10418 release of llama.cpp brings notable improvements to SYCL support, particularly through the introduction of host pinned memory, which enhances host-to-device memory access. This update also resolves a thread-safety issue, ensuring more stable performance across different hardware setups. While no new models are introduced, the release focuses on strengthening the existing infrastructure, making it more robust for developers working with SYCL. This update is crucial for optimizing performance and ensuring compatibility, especially for those leveraging SYCL in their development environments.

llama.cpp Releases·Aug 14, 2026