Skip to content
K Koda Intelligence
Subscribe
THE SIGNAL · 15 AUG 2026 · 5 MIN READ

China ships the weights, Washington ships a letter

Z.ai says GLM-5.3 weights land inside two weeks and Alibaba has already opened Qwen3.8-2.4T-A95B. Senator Jim Banks answered on August 14 with a letter asking the administration for incentives.

AIOPEN WEIGHTSPOLICY
TL;DR
Today in three lines
  1. 01

    Two Chinese labs give away frontier weights this fortnight Z.ai told Bloomberg that GLM-5.3, built on the same roughly 700-billion-parameter base as GLM-5.2, ships weights inside...

  2. 02

    Routing, not raw scale, is the new builder skill OpenAI's GPT-5.6 builder playbook is mostly about which tier to call and when to trade reasoning depth for latency.

  3. 03

    White House names India in transshipment enforcement push On Aug 14 the White House said it had identified India and 40 other countries in a new enforcement push against...

KODA PROFrom the desk

The Operator Tier is coming.

A weekly operator deep dive, the full Dojo Pro prompt packs, and the complete prompt database. Founding members lock the launch price forever.

Join the founding list →
Lead Story

The day's defining move.

China · Bloomberg
Lead Story
China·Bloomberg·15 August 2026

Z.ai Readies GLM-5.3, Weights In Two Weeks

Z.ai (Zhipu) told Bloomberg it is preparing GLM-5.3, built on the same roughly 700-billion-parameter base as GLM-5.2 but tuned for coding, with internal benchmarks placing it close to Anthropic's Fable 5 on coding evaluations. The company said it plans to publish the model's weights within...

Continue reading
AI

Z.ai says GLM-5.3, built on the same roughly 700-billion-parameter base as GLM-5.2 and tuned for coding, will have downloadable weights within two weeks, with internal benchmarks near or ahead of Anthropic's Fable 5.

World

Trump told a New York rally on Aug 14 that he will "pretty soon" declare the Strait of Hormuz US territory, as Iran denies access to much civilian shipping and the US Navy blockades Iranian ports.

Markets

The S&P 500 and Nasdaq closed at reported record highs on Aug 13 after cooler US producer prices, yet mood stays Fear with WTI elevated on Hormuz risk and gold selling off sharply.

Wild Card

Senator Jim Banks wants incentives for US open weights while the White House names India and 40 others in a transshipment probe: Washington is trying to slow Chinese hardware routes and speed American software giveaways at the same time.

Markets

Market snapshot.

Market TerminalLive
S&P 500 7,785.76 -0.17%
Nasdaq 26,729.16 -0.28%
Bitcoin $62,948.76 -0.72%
Ethereum $1,880.58 -0.18%
Crude Oil (WTI) $82.40 +1.42%
Crypto Fear & Greed 34 Fear
Today's Focus

The three signals that move the day.

01

Open Weights Arms Race

Z.ai plans to publish GLM-5.3 weights inside two weeks, and Alibaba has already released Qwen3.8-27B plus the 2.4-trillion-parameter Qwen3.8-2.4T-A95B mixture of experts with roughly 95 billion active parameters. The competitive question is no longer whether a frontier-class model exists but how quickly its weights hit a download client. American labs are shipping open entries too, including Meta's Muse Glimmer and Nvidia's Nemotron line, but the release cadence is coming from Hangzhou and Beijing.

02

Washington's Policy Lag

Banks released his letter on Friday, August 14, urging the Trump administration to build incentives for US open-weight development, an argument framed around ceded ground rather than any concrete program. Letters move slower than model releases: GLM-5.3 weights are due within two weeks. The gap between legislative pressure and shipping schedules is the story to watch.

03

Routing Beats Raw Scale

OpenAI's GPT-5.6 builder guide is explicitly about which tier to call and when to trade reasoning depth for latency, following the Ultrafast preview of GPT-5.6 Sol at up to 14 times prior speed. Google's Gemini 3.7 Flash makes the same case from the other direction: 56 on the Artificial Analysis Intelligence Index, ninth overall, at 340 tokens per second and two to three times cheaper per API call than Claude Sonnet. Developers are being handed dials, not just bigger models.

Listen & Watch

Daily broadcasts.

Podcast · Video · Infographic
The Daily Deep Dive

Listen to today's briefing

YouTube · Short

Today's Signal Short

Visual

Intelligence map

The Wire

AI intelligence.

China·Bloomberg

Z.ai Readies GLM-5.3, Weights In Two Weeks

Z.ai (Zhipu) told Bloomberg it is preparing GLM-5.3, built on the same roughly 700-billion-parameter base as GLM-5.2 but tuned for coding, with internal benchmarks placing it close to Anthropic's Fable 5 on coding...

Read
Open Source·Techgenyz

Alibaba Opens Qwen 3.8 Weights, Including Max Tier

Alibaba's Qwen team released open weights for Qwen3.8-27B, a natively multimodal dense model, and the Max-level Qwen3.8-2.4T-A95B, a 2.4-trillion-parameter mixture of experts with roughly 95 billion active parameters. The family handles...

Read
Policy·Reuters

Senator Banks Presses Trump On Open Weights

Republican Senator Jim Banks released a letter on Friday, August 14, urging the Trump administration to create incentives for US companies to build open-weight AI models, arguing American labs are ceding ground as Chinese releases pile...

Read
Enterprise·OpenAI

OpenAI Publishes GPT-5.6 Builder Playbook

OpenAI released a builder's guide to GPT-5.6 covering the new feature surface and how developers should route work across the family, following this week's preview of Ultrafast mode running GPT-5.6 Sol at up to 14 times prior speed. The...

Read
Benchmark·Google

Gemini 3.7 Flash Ranks Ninth At 340 Tokens A Second

New third-party numbers put Google's Gemini approximately 3.7 Flash at 56 on the Artificial Analysis Intelligence Index, ninth overall, with throughput of 340 tokens per second. No major...

Read
Trend·ExplainX

India Becomes Claude's Second-Largest Market

One year after the IndiaAI Mission's compute-procurement phase approved roughly 18,693 subsidized GPUs, India has shipped open-source models trained end-to-end domestically and stood up a nine-institution consortium covering all 22...

Read
Trend·Bloomberg

India's Data Workers Train Their Robot Replacements

A Bloomberg investigation published August 12 documents thousands of Indian workers annotating and demonstrating physical tasks for AI firms building robots explicitly aimed at automating the same categories of labor. The reporting puts...

Read
Model Release·Cohere Labs

Cohere Labs Ships North Micro Vision Instruct

Cohere Labs introduced North Micro Vision Instruct, a compact instruction-tuned vision-language model published through its research blog this week. The positioning is small-footprint document and screen understanding for enterprise...

Read
The Lab

Tool of the day, field tested.

All Lab reports
IdeaVideo AI screenshot Deep Dive 6.8/ 10
Marketing Entry plan listed at $0/month, paid from $59.9/month under 15 minutes for a first clip

IdeaVideo AI

It is a model buffet with a credit meter, and the smartest way to use it is to generate ads you fully intend to delete.

This is for marketers and solo founders who need to see three hook variations before committing budget to a shoot or an editor. It is a competent aggregator with a clean single workspace and a credible model lineup, but it is a concepting tool, not a finishing tool, so treat every output as a storyboard that happens to move. Worth the Pro seat if you are testing creative weekly; skip it if you need broadcast-ready assets out the other end.

Capability 7.0
Ease 8.0
Value 6.5
Momentum 6.0

“Simple to use, the video quality holds up, and having several AI video models under one roof is the appealing part.”Third-party AI tool directory listing, surfaced in research

Also on the radar
Coding

Lovable pitches itself as a full-stack engineer, not a code assistant

Lovable takes a prompt and ships a running full-stack web app, with GitHub and Supabase wiring handled for you, then lets you edit and deploy from the same surface. The realistic use is MVPs and internal tools where the database schema is simple and the exit cost of a rewrite is low. Its Product Hunt listing shows enterprise-only pricing, so treat it as a team purchase rather than a weekend experiment.

Try it
Creativity

Play is a native iOS design tool that runs on the phone you are designing for

Play lets you design and prototype mobile products directly on an iPhone using real iOS materials, so you test motion and touch targets on the device instead of guessing from a desktop artboard. It is free, which makes it a cheap way to sanity check a screen before you commit it to a design system. Best used for the last-mile feel check, not for maintaining a shared component library.

Try it
Build

Ugic generates Figma screens from your own component library, not a generic kit

Ugic is a Figma plugin that reads your existing components and design system, then converts text such as a PRD into structured, multi-language layouts. That constraint is the point: output lands inside your system instead of producing off-brand mockups you have to rebuild. It is a paid plugin, so scope one real feature spec at it before rolling it out to a design team.

Try it
Productivity

Rupt turns account sharing into a revenue question instead of a support ticket

Rupt monitors for shared logins with identity verification and fraud signals, then routes those users toward paying for their own seat rather than simply locking them out. For a small SaaS, the leverage is in the conversion flow, not the detection, so tune the messaging before you enable enforcement. The Scale plan offers 14 days free with no credit card, which is enough time to measure how much sharing you actually have.

Try it
The Arena

Competitive intel.

OpenAI

Ships developer playbook for GPT-5.6 as Nvidia trims its data center backstop

Google DeepMind

(recent) Gemini 3.7 Flash targets agent latency while 3.5 Pro stays late

China Challengers

GLM-5.3 extends Chinese lead in open-weight releases

THE DOJO · BUILD TODAY

Build a routing layer before you buy more intelligence

01

Write a routing table for GPT-5.6. Use OpenAI's new builder playbook to map three job types in your product to a specific tier, then log latency and cost per call so the trade between reasoning depth and speed is measured rather than guessed.

02

Pin an open-weight fallback. Stand up Qwen3.8-27B locally this weekend and queue capacity for GLM-5.3 weights, which Z.ai says land inside two weeks. Cheap open tiers now run two to three times cheaper than Claude Sonnet 5 or GPT-5.6 Terra.

03

Prototype the interface, not the deck. Try Lovable for a full-stack pass, Play for native iOS layout on the device itself, and Ugic to generate Figma screens from your own component library. The Scale plan offers 14 days free with no credit card.

The Bottom Line

The frontier is getting given away, and policy is still drafting

Two Chinese labs are publishing near-frontier weights inside a fortnight while the US policy response is a letter dated August 14. For builders that asymmetry is an opportunity, not a headline: open Qwen and GLM tiers undercut Claude Sonnet 5 and GPT-5.6 Terra by two to three times on price, and Gemini 3.7 Flash shows that latency now competes with raw index score. The skill that compounds this quarter is routing, which tier answers which request, measured in tokens per second and dollars. Markets drifted lower and crypto sentiment sat at 34 in Fear, so nobody is paying a premium for your infrastructure story right now. Ship the cheap version that works.

Want this every morning?

AI analysis, world news, markets, and tools. One briefing, delivered free.

One email per day. No spam. Unsubscribe anytime.

Forward this to one operator you work with. Your referral link is in every email; milestones at koda.community/refer.