skip to content- 2026-07-26
|
Prompt Caching: Cut AI Agent Token Costs
Learn how prompt caching changes LLM token costs, why coding agents resend context, what compaction can waste, and how to measure Codex, Claude Code, and OpenCode.
- 2026-07-02
|
Pi Coding Agent: Minimal, Programmable Terminal Harness
Set up Pi as a minimal coding harness with GLM-5.2, persistent sessions, compaction, custom tools, extensions, and clear Z.AI policy caveats.
- 2026-07-02
|
Pi vs ZCode vs OpenCode: Which Harness Fits GLM-5.2?
Compare Pi, ZCode, and OpenCode as coding harnesses for GLM-5.2: workflow, permissions, context, setup, quota paths, and a fair same-model test.
- 2026-07-02
|
ZCode: GLM-5.2-Native Coding Harness Guide
Use ZCode as Z.AI's integrated GLM-5.2 coding harness: desktop setup, Goal Mode, subagents, safety confirmations, quota, and workflow trade-offs.
- 2026-02-03
|
Verify Codex Token-Usage and Caching Claims
A dated verification ledger for Codex token usage, prompt caching, compaction rereads, OpenCode comparisons, usage incidents, and reset claims.
- 2026-02-01
|
OpenCode: Providers, Zen, Permissions, and GLM-5.2
Current OpenCode guide for its open-source coding agent, optional Zen gateway, BYO providers, permissions, integrations, and three GLM-5.2 billing paths.