KKoda IntelligenceDaily Signal
S&P 5007,736.52↑ 1.79%·NASDAQ26,584.99↑ 2.59%·BTC$64,012.81↑ 0.87%·ETH$1,866.27↑ 0.43%·FEAR & GREED27↓ FEAR·QWEN API IN$2· FLAT·QWEN API OUT$6· FLAT·BRENT CRUDE5%↓ 5%·
THE SIGNAL · 05 AUG 2026 · 5 MIN READ

Alibaba is about to make 2.4 trillion parameters a free download

Alibaba's Qwen team says Qwen3.8-Max weights go public next week, a 2.4-trillion-parameter model with a 1-million-token context window. The paid API stays at $2 per million input tokens. Frontier scale just stopped being a moat.

AIOPEN WEIGHTSMARKETS
KODA PROFrom the desk

The Operator Tier is coming.

A weekly operator deep dive, the full Dojo Pro prompt packs, and the complete prompt database. Founding members lock the launch price forever.

Join the founding list →
Lead Story

The day's defining move.

Open Source · Reuters
Lead Story
Open Source·Reuters·5 August 2026

Alibaba Will Publish Qwen3.8-Max Weights Next Week

Alibaba's Qwen team says the weights for Qwen3.8-Max, the 2.4-trillion-parameter model with 95 billion activated parameters and a 1-million-token context window, will be made publicly available next week, alongside an open-source release of the smaller Qwen3.8-27B. That would put a model Bloomberg reports matches...

Continue reading arrow_outward
smart_toyAI

Alibaba's Qwen team says weights for Qwen3.8-Max, a 2.4-trillion-parameter model with 95 billion activated parameters and a 1-million-token context window, go public next week alongside an open-source Qwen3.8-27B.

publicWorld

Qatar said on Aug 4 that mediators are progressing toward ending the US-Iran war and Treasury Secretary Scott Bessent hinted a Hormuz reopening deal could be near, though Tehran denies talks are under way and the strait remains effectively shut.

trending_upMarkets

The Dow and S&P 500 closed at record highs Tuesday as AI-linked earnings met a 5% drop in Brent crude, yet the mood reading stays at Fear because the rally rests on a peace deal Tehran has not confirmed.

boltWild Card

Qatar said on Aug 4 that mediators are making progress toward ending the US-Iran war, sending Brent crude down more than 5% on top of the previous day's losses, and US Treasury Secretary Scott Bessent suggested a deal to reopen the Strait of Hormuz could be near.

Markets

Market snapshot.

query_statsMarket TerminalLive
S&P 500 7,736.52 arrow_drop_up+1.79%
Nasdaq 26,584.99 arrow_drop_up+2.59%
Bitcoin $64,012.81 arrow_drop_up+0.87%
Ethereum $1,866.27 arrow_drop_up+0.43%
Crude Oil (WTI) $75.41 arrow_drop_down-6.14%
Crypto Fear & Greed 27 arrow_drop_downFear
Today's Focus

The three signals that move the day.

01

Open Weights At Frontier Scale

Publishing Qwen3.8-Max weights would put a model Bloomberg reports matches or exceeds Anthropic's flagship into unrestricted circulation, with the smaller 27B variant open-sourced for anyone to run locally. The strategic effect is to make the frontier a commodity rather than a licensed service. Every US lab pricing per token now competes against a free download.

02

Voluntary Testing As Policy

OpenAI's publication of third-party cyber evaluations lands while the White House favors voluntary safety testing over mandates, which makes external red-team reports the primary evidence regulators will ever see. That gives labs control over both the testing scope and the disclosure. It is a weaker check precisely when open weights remove the ability to revoke access after the fact.

03

Memory And Water

Samsung's next-generation AI memory announcement targets the real bottleneck, since memory capacity per accelerator, not logic, now limits how large a model any given fleet can train or serve. The Sichuan floods that forced more than 900,000 evacuations expose the other constraint: hydropower and the data center capacity depending on it. Compute scarcity is increasingly a supply-chain and weather story.

Listen & Watch

Daily broadcasts.

Podcast · Video · Infographic
The Daily Deep Dive

Listen to today's briefing

YouTube · Short

Today's Signal Short

Visual

Intelligence map

The Wire

AI intelligence.

Open Source·Reuters

Alibaba Will Publish Qwen3.8-Max Weights Next Week

Alibaba's Qwen team says the weights for Qwen3.8-Max, the 2.4-trillion-parameter model with 95 billion activated parameters and a 1-million-token context window, will be made publicly available next week, alongside an open-source...

Read arrow_outward
Policy·OpenAI

OpenAI Publishes Third-Party Cyber Evaluations

OpenAI released an account of third-party cyber evaluations run against its models, describing how outside groups probed security and robustness and what the assessments found. The disclosure arrives as the White House pushes voluntary...

Read arrow_outward
Hardware·Reuters

Samsung Launches Next-Generation AI Memory

Samsung Electronics announced a next-generation AI memory technology on August 4, extending the bandwidth race that has made memory, not logic, the binding constraint on large-model training and inference. The timing matters because the...

Read arrow_outward
Enterprise·OpenAI

OpenAI Pushes ChatGPT Work and Codex Into Classrooms

OpenAI published guidance for educators on using ChatGPT Work and Codex in teaching, laying out practical applications and integration rules for curricula. The move targets the institutional buyer at a moment when higher-education...

Read arrow_outward
Trend·OpenAI

OpenAI Attacks Apple's AI Strategy Directly

OpenAI posted a piece titled "Apple is getting this wrong," criticizing Apple's approach to AI integration and arguing for strategies it says better match user expectations. A model lab publicly attacking a platform owner is unusual and...

Read arrow_outward
The Lab

Tools worth a look.

All reviews arrow_forward
Coding

LangChain Deep Agents targets your input-token bill

The update is explicitly about cutting input-token usage and making large agent deployments cheaper to run, which matters more than latency once you have dozens of agents in production. Before upgrading, log the input tokens per task for your three most-used agents so you have a baseline to compare against. Context bloat, not model choice, is usually where agentic app costs escape.

Try it arrow_outward
Productivity

Sumly.AI for triaging audio and video you will never watch

Sumly.AI turns podcasts, talks, and recorded calls into condensed summaries, and it is currently free and ad-free. The practical use is triage: run the 90-minute conference session or earnings call through it first, then decide whether the full recording earns your time. Keep the summary as a searchable note and cite timestamps back to the original before you quote anything.

Try it arrow_outward
Build

Finalle.ai is a scouting pass for market monitoring, not a source of truth

Finalle.ai aggregates new-media and financial data streams and layers generative summarization on top, aimed at teams that need faster reads on markets and companies. Test it against a week you already understand well and score how often its signals matched what actually mattered. Any number it surfaces should be checked against filings or primary data before it enters a document anyone acts on.

Try it arrow_outward
Mindset

Carta's H2 2026 exit data says AI growth is now the sorting mechanism

Carta's 3 August analysis of the startup exit environment concludes that the split between companies with promising exits and those without usually comes down to AI-driven growth. Read it as a diagnostic on your own metrics: whether AI is producing measurable revenue or retention movement, not whether it appears in your deck. If you cannot point to the line it moved, buyers will not either.

Try it arrow_outward
Mindset

Read the Commission's own AI policy hub, not the secondhand summaries

With EU AI Act enforcement widening from 2 August, the volume of consultant explainers has outpaced the accuracy of them. The Commission's own policy page is the canonical map of obligations, timelines, and the linked funding and research programs, and it is updated as guidance lands. Bookmark it and check the primary text before you change a product decision on the basis of a blog post.

Try it arrow_outward
The Arena

Competitive intel.

OpenAI

Settles DOJ hiring discrimination probe for $3.2M while disclosing three unreported cyber incidents

Meta AI

(recent) At the White House safety table; Muse Spark 1.1 remains the latest shipped model

Google DeepMind

(recent) Expected at White House safety meeting; no new DeepMind research release

THE DOJO · BUILD TODAY

Plan for open frontier weights before the download lands

01

Price your stack against $2 in. Benchmark your current provider's cost per task against Qwen3.8-Max at $2 per million input and $6 per million output tokens, then decide whether next week's open weights change your hosting math or just your negotiating position.

02

Cut your input-token bill first. Run LangChain Deep Agents against one long-context workflow and measure token spend before and after. A 1-million-token window is only cheap if you are not stuffing it blindly.

03

Build a real evaluation harness. OpenAI published third-party cyber evaluations while Washington favors voluntary testing over mandates, so external red-team reporting is the only public signal. Write down the three failure modes you will test on any model you adopt, including open weights you host yourself.

The Bottom Line

When frontier weights are a download, execution is the moat

A 2.4-trillion-parameter model going public next week does not hand you a business, it hands you a commodity input. The scarce things are memory bandwidth, which is why Samsung's August 4 announcement matters more than it reads, and credible evaluation, which is why OpenAI's third-party cyber disclosures land in a voluntary-testing vacuum. Markets already priced the optimism with the Dow and S&P 500 at record closes on cheaper crude and AI earnings. Your job today is to know exactly which layer of that stack you own.

Want this every morning?

AI analysis, world news, markets, and tools. One briefing, delivered free.

One email per day. No spam. Unsubscribe anytime.