
Writer introduces new AI model and upgraded harness to contain token costs
The AMW Read
New flagship model and harness from an established enterprise AI player updates the player map and signals a cost-optimization trend, but does not overturn existing open-weight dynamics.
Named counterparties: Z.ai
Writer introduces new AI model and upgraded harness to contain token costs
Writer on Thursday launched Palmyra X6, a new flagship model built as a post-training variation on Z.ai's open source GLM-5.2, alongside significant upgrades to its agentic harness infrastructure. The company estimates the combination will cut customer costs by as much as 50% for basic tasks, with a particular focus on complex, multi-step operations executed faster and with fewer tokens. Both features are available to Writer clients starting Thursday.
The move responds to growing enterprise pressure to contain token spending, which CEO May Habib says reflects frustration with major AI labs that have a financial incentive to drive up usage. Writer's own research, published in a recent paper, found that harness efficiency improvements were often a more reliable cost lever than model choice, reducing expenses by an average of 40% across tests. For Writer's customers, the experience remains model-agnostic: Palmyra X6 sits alongside other Writer models or models imported through Azure or Amazon Bedrock.
The broader implication is a shift in the enterprise AI market toward cost efficiency as a primary buying criterion, rather than raw benchmark performance. Habib claims CIOs are "giving up on the labs," suggesting a window for smaller players and open-source-based approaches to win enterprise trust. For builders, the takeaway is that optimizing the orchestration layer—the harness—can be as impactful as model selection, and that cost predictability is becoming a competitive differentiator. Investors should watch whether this model-agnostic, cost-focused strategy broadens beyond Writer's marketing niche into horizontal enterprise AI.