On Aug 26, 2026 China's Z.ai released GLM-5.3-Flash, a 320B-parameter (18B active) mixture-of-experts model with a 1M-token context and native image/video input, under an MIT license. API pricing is roughly $0.15/$0.50 per million input/output tokens and it lands within ~0.7 points of Claude Opus 4.8 on Terminal-Bench 2.1 (84.3 vs 85.0).
Z.ai ships GLM-5.3-Flash: 1M-context multimodal coding model at ~1/10 the price
Founders running agentic or coding workloads can slot in a cheap, openly licensed long-context model to cut inference costs; self-host the MIT weights or call the API for large-document tasks.
Source: MarkTechPost
More that helps you.
OpenHands 1.0 ships production self-hosted coding agent at 68% SWE-bench Verified
The open-source autonomous coding agent OpenHands reached its 1.0 release on September 8, 2026, adding Docker sandboxing, built-in security policies, resource limits and a plugin s…
Anthropic ships Claude Fable 5.1 and Mythos 5.1, cutting agent costs ~45%
On September 2, 2026 Anthropic released Claude Fable 5.1 (broadly available) and Mythos 5.1 (restricted to vetted organizations), citing about 45% lower costs for agentic workloads…
Anthropic adds auto permission mode to Claude Managed Agents
On September 10, 2026 Anthropic shipped an 'auto' permission policy for Claude Managed Agents that lets the server evaluate each agent or MCP tool call and run it, deny it, or paus…
GitHub Copilot adds parallel agent sessions and a unified Copilot experience
In its September 4, 2026 changelog, GitHub added parallel agent sessions to the Copilot app, letting multiple AI tasks run independently in separate Git worktrees, and introduced a…
Get briefs like this tuned to you.
In the app, Founder Briefs are personalized to your country, industry and stage, and you can save the ones that matter.
See plans →