KKoda IntelligenceDaily Signal
S&P 5007,600.50↑ 1.48%·NASDAQ25,913.90↑ 2.13%·BTC$63,383.17↓ 0.16%·ETH$1,851.87↓ 1.63%·FEAR/GREED25↓ EXTREME FEAR·QWEN3.8-MAX PARAMS2.4T· FLAT·TERMINAL-BENCH 2.186.6↑ VS 84.6·AUTONOMY RUN16 days· FLAT·
THE SIGNAL · 04 AUG 2026 · 5 MIN READ

The parameter ceiling moved east while Washington passed out a testing form

Alibaba shipped Qwen3.8-Max at 2.4 trillion parameters on August 3, just under Moonshot's Kimi K3 at 2.8 trillion. The US answer this week is a voluntary safety testing framework presented to OpenAI, Anthropic and Google.

AIPOLICYMARKETS
SPONSOR KODAFrom the desk

Put your product in front of AI operators.

Koda publishes a fact-checked AI intelligence briefing every morning: site, email, podcast, and video. Founding sponsors lock today's rates for 6 months as the audience compounds.

View the media kit →
Lead Story

The day's defining move.

Model Release · CNBC
Lead Story
Model Release·CNBC·4 August 2026

Alibaba Ships 2.4-Trillion-Parameter Qwen3.8-Max

Alibaba unveiled Qwen3.8-Max on Monday, August 3, calling it its most powerful model to date: a Mixture-of-Experts system with 2.4 trillion total parameters, roughly 95 billion active per token, and a 1 million-token context window. Alibaba says it ran a software engineering project autonomously for 16 days in...

Continue reading arrow_outward
smart_toyAI

Alibaba's Qwen3.8-Max landed Monday with 2.4 trillion total parameters, roughly 95 billion active per token, a 1 million-token context window, and a claimed 16-day autonomous software engineering run.

publicWorld

Trump says the United States shelved what he called the "biggest attack since World War II" on Iran in favour of talks he claims begin Monday, after a second consecutive night of strikes and Iranian retaliation.

trending_upMarkets

Sentiment sits at Extreme Fear, an unforgiving backdrop for capability announcements priced in trillions of parameters and gigawatts of compute.

boltWild Card

OpenAI's continuous-voice GPT Live keeps a channel open instead of trading turns, which only pencils out economically at the inference prices DeepSeek's V4-Flash is now setting more than 100 times below Claude.

Markets

Market snapshot.

query_statsMarket TerminalLive
S&P 500 7,600.50 arrow_drop_up+1.48%
Nasdaq 25,913.90 arrow_drop_up+2.13%
Bitcoin $63,383.17 arrow_drop_down-0.16%
Ethereum $1,851.87 arrow_drop_down-1.63%
Crude Oil (WTI) $80.74 arrow_drop_down-4.64%
Crypto Fear & Greed 25 arrow_drop_downExtreme Fear
Today's Focus

The three signals that move the day.

01

The Parameter Ceiling Moves East

Qwen3.8-Max at 2.4 trillion parameters approaches but does not exceed Moonshot AI's Kimi K3 at 2.8 trillion, and both target the same workload: long-horizon coding and knowledge work behind a 1 million-token context window. Two Chinese labs setting the size frontier within days of each other changes what counts as a reference point for frontier scale.

02

Cost Structure As Weapon

A research firm cited by Reuters puts DeepSeek's V4-Flash more than 100 times below Anthropic's Claude on price to run, landing the same day Alibaba claimed near-frontier benchmark scores including 86.6 on a coding evaluation. The argument is no longer that Chinese models match US quality; it is that they compete on unit economics while getting close enough on capability.

03

Voluntary Rules, Global Race

The White House is convening OpenAI, Anthropic and Google this week to present a US framework for voluntary safety testing, an opt-in approach stemming from a June executive order on AI cybersecurity. Opt-in review binds only the labs in the room, which excludes the developers currently resetting both the size and price frontiers.

Listen & Watch

Daily broadcasts.

Podcast · Video · Infographic
The Daily Deep Dive

Listen to today's briefing

YouTube · Short

Today's Signal Short

Visual

Intelligence map

The Wire

AI intelligence.

Model Release·CNBC

Alibaba Ships 2.4-Trillion-Parameter Qwen3.8-Max

Alibaba unveiled Qwen3.8-Max on Monday, August 3, calling it its most powerful model to date: a Mixture-of-Experts system with 2.4 trillion total parameters, roughly 95 billion active per token, and a 1 million-token context window....

Read arrow_outward
Policy·Politico

White House Hosts Labs on Voluntary Safety Tests

The Trump administration is convening AI developers at the White House this week to present a new US framework for voluntary safety testing of models, with OpenAI, Anthropic and Google among those planning to attend. The framework stems...

Read arrow_outward
China·Reuters

DeepSeek V4-Flash Undercuts Claude by 100x

A research firm cited by Reuters says DeepSeek's V4-Flash is by far the cheapest to run among well-known models, with pricing more than 100 times below Anthropic's Claude. That figure lands the same day Alibaba's Qwen3.8-Max claimed...

Read arrow_outward
Model Release·OpenAI

OpenAI Details Continuous Voice With GPT Live

OpenAI published technical detail on continuous voice interaction in GPT Live, describing real-time dialogue that holds an open channel rather than trading discrete turns. The company frames the change as an interaction-efficiency gain,...

Read arrow_outward
China·Reuters

Moonshot's Kimi K3 Sets the Size Ceiling

Reuters notes that Qwen3.8-Max at 2.4 trillion parameters is approaching but not exceeding the size of Moonshot AI's Kimi K3, which carries 2.8 trillion parameters. Kimi K3 is multimodal with a 1 million-token context window and is...

Read arrow_outward
China·Reuters

Chinese Labs Push Video Generation and Cheap Inference

Alongside the Qwen and Kimi releases, roundups this week flag ByteDance and MiniMax unveiling new video generators and DeepSeek widening access to V4 Flash. The pattern is a coordinated push across three fronts at once: frontier-scale...

Read arrow_outward
Enterprise·WebDisclosure

TrustHouse.AI Launches Enterprise Context Engine

TrustHouse.AI announced on August 3 the coming launch of its Context Engine, a platform layer pitched at enterprises trying to move AI systems into production with better accuracy and governance. The company positions itself as AI trust...

Read arrow_outward
The Lab

Tool of the day, field tested.

All Lab reports arrow_forward
Pushary screenshot Deep Dive 7.0/ 10
Coding $9.99/mo after a 7-day card-first trial about 10 minutes

Pushary

Your agent stalls on a yes-or-no question while you are making coffee; Pushary moves that question to your lock screen and gets the run moving again.

This is for people running two or more coding agents who keep losing minutes to approval prompts nobody is at the desk to answer. If that is you, the $9.99/mo is cheap against the dead time, and the MCP-plus-hooks install is genuinely a one-command affair with Claude Code. If you run a single agent on Claude Code with Claude Max, Anthropic's own Remote Control already covers that case for free and Pushary is redundant.

Capability 7.2
Ease 8.3
Value 6.8
Momentum 6.0

“Supporters like that it collapses an agent pause into a simple yes button on your phone, so the work keeps moving while you are away from the desk.”Product Hunt

Also on the radar
Productivity

Raycast for iOS shipped its biggest update yet

Raycast's mobile app now exposes dictation, AI commands, Snippets and Quicklinks directly from the keyboard, which means the launcher shortcuts you built on desktop are usable mid-conversation on a phone. The practical win is Snippets: move your three most-retyped blocks of text there and test whether the keyboard surface is fast enough to replace the copy-paste habit. Check the site for current pricing tiers before committing a team.

Try it arrow_outward
Mindset

Use TAAFT's Job Impact Index instead of its tool list

There's An AI For That relaunched with an extensive tool database, personalized suggestions and an AI Job Impact Index. The database is the least useful part; the Impact Index is worth twenty minutes because it forces you to name which of your own recurring tasks are already automatable rather than browsing tools with no problem in hand. Access is free, with paid options available.

Try it arrow_outward
Creativity

BeatMV is a music-to-video generator worth one test track

BeatMV appears in directory listings at v3.2.6, which suggests a shipping product rather than a landing page, but the public description is thin on what the pipeline actually does with your audio. Give it one track you already own the rights to and judge the cut timing against the beat before you plan any client work around it. Treat the output as a rough assembly, not a deliverable.

Try it arrow_outward
Build

AI 3D Model Maker for throwaway asset drafts

Make3DAI is running as a featured listing in the tool directories this week, aimed at generating 3D models from prompts. If you need placeholder geometry for a prototype scene or a product mock, that is exactly the low-stakes job to test it on, since topology quality is where most text-to-3D tools fail. Export early and open the file in your own editor before you decide it fits the pipeline.

Try it arrow_outward
The Arena

Competitive intel.

OpenAI

Ships continuous voice interaction for GPT Live while regulators and state AGs close in

Anthropic

(recent) Joins White House talks on voluntary AI cybersecurity tests

Google DeepMind

(recent) Policy seat at the table, no new model this week

China Challengers

(recent) Alibaba unveils Qwen3.8-Max as the newest direct challenge to OpenAI and Anthropic

THE DOJO · BUILD TODAY

Price your inference tier before you pick your model.

01

Benchmark your coding agent against Qwen3.8-Max. Take one long-horizon task you already run on Claude Opus 4.8 or GPT-5.6 and rerun it on Qwen3.8-Max. The 86.6 Terminal-Bench 2.1 score and the 16 day autonomous run claim only matter if they hold on your repo.

02

Build a two-tier routing layer this week. Route cheap, high-volume calls to DeepSeek V4-Flash and reserve frontier models for the hard 10 percent. At more than 100 times below Claude on price to run, the arbitrage pays for the routing work in days.

03

Put agent permissions where you will actually see them. Install Pushary so agent permission prompts hit your lock screen instead of a buried terminal, and pair it with the Raycast for iOS update for on-the-go triage. Voluntary safety frameworks will not audit your own agents for you.

The Bottom Line

The frontier is now two markets, not one ladder.

Chinese labs set both ends of the range this week: Kimi K3 holds the 2.8 trillion parameter ceiling, Qwen3.8-Max claims near frontier coding at 2.4 trillion, and V4-Flash sits more than 100 times under Claude on cost. Washington's contribution is a voluntary testing framework presented to OpenAI, Anthropic and Google, which means procurement and liability, not policy, will do the filtering. TrustHouse.AI's August 3 continuity product and widening obligations from August 2 are the tell that vendors are already selling insurance against this. Build for a world where you route across tiers, verify autonomy claims yourself, and treat safety instrumentation as your own engineering line item.

Want this every morning?

AI analysis, world news, markets, and tools. One briefing, delivered free.

One email per day. No spam. Unsubscribe anytime.