Close Menu
CrypThing
  • Directory
  • News
    • AI
    • Press Release
    • Altcoins
    • Memecoins
  • Analysis
  • Price Watch
  • Price Prediction
Facebook X (Twitter) Instagram Threads
CrypThingCrypThing
  • Directory
  • News
    • AI
    • Press Release
    • Altcoins
    • Memecoins
  • Analysis
  • Price Watch
  • Price Prediction
CrypThing
Home»AI»Writer introduces new AI model and upgraded harness to contain token costs
AI

Writer introduces new AI model and upgraded harness to contain token costs

adminBy adminAugust 14, 20263 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link Bluesky Reddit Telegram WhatsApp Threads
Writer introduces new AI model and upgraded harness to contain token costs
Share
Facebook Twitter Email Copy Link Bluesky Reddit Telegram WhatsApp

Across the AI industry, users are becoming more conscious of just how expensive their deployments can be —and feeling a new urgency to cut costs. But while open source models offer significantly lower per-token costs, it can be difficult to find the right model for a given job.

On Thursday, Writer, which offers AI tools and agents for marketers, launched a new flagship model called Palmyra X6, aimed at solving that problem for its users. Built as a post-training variation on Z.ai’s open source model GLM-5.2, Writer says the new system should provide deployment-ready capabilities at a much lower price. The company estimates the new model, combined with changes to the companies harness infrastructure, will cut costs for its customers by as much as 50% for basic tasks.

Together with the new model, the company also released significant upgrades to its standard agentic harness. Both features will be available to Writer clients starting Thursday.

“I think the enterprise is absolutely sick of chasing the next benchmark,” CEO May Habib told TechCrunch. “They want flattening cost, and it seems like nobody can deliver that.”

The new approach puts particular emphasis on complex, multi-step tasks, executed faster and with fewer tokens. And Writer sees harness optimization as a crucial lever toward making that happen.

A recent paper from Writer researchers lends credence to this approach, testing small changes in harness efficiency across multiple different models. The research found that, in many cases, changes in the harness were a more reliable way to reduce costs than model choice, with costs falling an average of 40% across their testing.

“The harness is the one component whose efficiency multiplies across every model an organization runs—present and future,” the researchers wrote.

For Writer’s clients, the experience is still model-agnostic: Palmyra X6 will sit alongside other Writer models or outside models imported through Azure or Amazon Bedrock. But Habib also sees the push to cut costs as driving a broader distrust toward major AI labs, which have a financial incentive to drive up token use.

“The cost explosion here is just unprecedented for customers, and so is the degree to which CIOs are giving up on the labs,” Habib told TechCrunch, adding that the AI labs “don’t deeply understand right how to help an enterprise get benefit from AI.”

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

2025 AI costs Harness introduces model October 27-29 San Francisco Techcrunch event TechCrunch|BProud token Trumps upgraded Writer
Share. Facebook Twitter Pinterest LinkedIn Tumblr Telegram Email Copy Link Bluesky WhatsApp Threads
Previous ArticleXRP Price Shift Reworks Evernorth Shares Before Its Nasdaq Debut
Next Article Crypto Player Takes Home $1.749M After a Million PSG Bet on 1win
admin

Related Posts

OpenAI confirms ‘wiki incident,’ says it’s ‘working on a framework’ for more disclosure

September 5, 2026

AI compute provider Nscale is looking for $3.5B in pre-IPO financing

September 4, 2026

Accel reportedly in talks to lead $1B round for Thinking Machines at $40B valuation

September 3, 2026
Trending News

NVIDIA Dynamo 1.0 Ships With 7x Inference Boost for AI Data Centers

March 16, 2026

ORBS) Reports Total Holdings of Approximately $380 Million, Includes OpenAI, Beast Industries, More Than 16,000 ETH and Nearly 302 Million WLD Tokens

September 3, 2026

Claude AI Improves Alignment Benchmarks While Preserving Capabilities

August 29, 2026

NVIDIA’s Confidential Computing Boosts AI Security Without Performance Hit

July 2, 2026
About Us

At crypthing, we’re passionate about making the crypto world easier to (under)stand- and we believe everyone should feel welcome while doing it. Whether you're an experienced trader, a blockchain developer, or just getting started, we're here to share clear, reliable, and up-to-date information to help you grow.

Don't Miss

Reporters found that Zerebro founder was alive and inhaling his mother and father’ home, confirming that the suicide was staged

May 9, 2025

Openai launches initiatives to spread democratic AI through global partnerships

May 9, 2025

Stripe announces AI Foundation model for payments and introduces deeper Stablecoin integration

May 9, 2025
Top Posts

NVIDIA Dynamo 1.0 Ships With 7x Inference Boost for AI Data Centers

March 16, 2026

ORBS) Reports Total Holdings of Approximately $380 Million, Includes OpenAI, Beast Industries, More Than 16,000 ETH and Nearly 302 Million WLD Tokens

September 3, 2026

Claude AI Improves Alignment Benchmarks While Preserving Capabilities

August 29, 2026
  • About Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer
© 2026 crypthing. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.