KKoda IntelligenceDaily Signal
S&P 5007,711.76↓ 0.25%·NASDAQ26,402.42↓ 0.52%·BTC$77,645.99↓ 3.25%·ETH$2,438.93↓ 2.85%·GLM-5.3-FLASH IN$0.075↓ 50%·GPT-5.6 SOL IN$4· FLAT·AA AGENTIC IDX59↑ CLAIMED·PROMO ENDSSeptember 9· FLAT·
THE SIGNAL · 29 AUG 2026 · 5 MIN READ

Three labs shipped previews, OpenAI shipped a claim

Tencent's 770B Hy4 preview, Alibaba's Qwen4 teaser and GLM-5.3-Flash at $0.075 per million input tokens all landed as open weights, while OpenAI says unshipped Astra already clears its research-intern bar.

AIOPEN WEIGHTSMARKETS
KODA PROFrom the desk

The Operator Tier is coming.

A weekly operator deep dive, the full Dojo Pro prompt packs, and the complete prompt database. Founding members lock the launch price forever.

Join the founding list →
Lead Story

The day's defining move.

Open Source · Reuters
Lead Story
Open Source·Reuters·29 August 2026

Tencent Open-Sources 770B Hy4 Preview For Coding

Tencent posted a preview of Hy4 to Hugging Face on August 28, an open-source mixture-of-experts model with 770 billion total parameters and roughly 49 billion active per text request, aimed at software engineering, research and financial analysis. Tencent says it will integrate the model into CodeBuddy and...

Continue reading arrow_outward
smart_toyAI

Tencent posted an open-source preview of Hy4 to Hugging Face on August 28, a 770-billion-parameter mixture-of-experts model with roughly 49 billion active parameters per text request, targeted at software engineering, research and financial analysis.

publicWorld

The missing count from the Himalayan glacier collapse and flood wave on the Tibet-Nepal border passed 2,400 on August 28, with 586 bodies recovered and 933 hydropower workers added to Nepal's list.

trending_upMarkets

Mood is Greed even as traders raised US rate-hike odds on hawkish comments from Fed Chair Kevin Warsh, with core PCE at 3.3% and Michigan consumer sentiment down to 51.7 in August from 55.2 in July.

boltWild Card

Washington's two levers on Friday were both financial: a US-Venezuela oil agreement and a FinCEN rule to cut Banque Misr UAE off from US correspondent banking; neither instrument touches a 770-billion-parameter file uploaded to Hugging Face.

Markets

Market snapshot.

query_statsMarket TerminalLive
S&P 500 7,711.76 arrow_drop_down-0.25%
Nasdaq 26,402.42 arrow_drop_down-0.52%
Bitcoin $77,645.99 arrow_drop_down-3.25%
Ethereum $2,438.93 arrow_drop_down-2.85%
Crude Oil (WTI) $83.44 arrow_drop_down-0.11%
Crypto Fear & Greed 68 arrow_drop_upGreed
Today's Focus

The three signals that move the day.

01

Previews As Shipping Strategy

Tencent's Hy4 preview and Alibaba's Qwen3.8-Flash-Next, explicitly framed as a Qwen4 architecture preview rather than a flagship, both went out as open weights before they were finished. Publishing architectural experiments buys deployment feedback that closed labs have to simulate internally, the same sequencing Z.ai used with the model it previewed as Ox Alpha and then released as GLM-5.3-Flash.

02

Inference Pricing Keeps Falling

GLM-5.3-Flash lists at $0.075 per million input tokens and $0.25 per million output under a 50% promotional discount running through September 9, while the larger GLM-5.3, whose open weights are still pending, carries a cited Artificial Analysis Agentic Index score of 59. Discount windows with hard expiry dates are a customer-acquisition tactic, not a cost structure, so the question is what September 10 pricing looks like.

03

Astra's Internal Benchmark Claim

OpenAI says the unshipped Astra family already clears an internal bar for an automated AI research intern: implement an idea in code, run the experiment, return results. Forbes reported the claim on August 28 alongside leadership statements that AGI arrives by year end, which puts a self-defined internal metric against Google's shipped and specific releases, Gemini 3.5 Transcribe with sub-second streaming across more than 85 languages and Gemini Omni 1.1 Flash.

Listen & Watch

Daily broadcasts.

Podcast · Video · Infographic
The Daily Deep Dive

Listen to today's briefing

YouTube · Short

Today's Signal Short

Visual

Intelligence map

The Wire

AI intelligence.

Open Source·Reuters

Tencent Open-Sources 770B Hy4 Preview For Coding

Tencent posted a preview of Hy4 to Hugging Face on August 28, an open-source mixture-of-experts model with 770 billion total parameters and roughly 49 billion active per text request, aimed at software engineering, research and...

Read arrow_outward
Benchmark·Time

OpenAI Says Astra Clears Research-Intern Benchmark

OpenAI said its upcoming Astra model family has already met an internal benchmark for an automated AI research intern, meaning it can implement ideas in code, run experiments and return results. Forbes reported the claim on August 28...

Read arrow_outward
Model Release·TechNode

Alibaba Ships Qwen3.8-Flash-Next As Qwen4 Preview

Alibaba released Qwen3.8-Flash-Next this week as an experimental open-weights multimodal mixture-of-experts model, explicitly framed as a preview of the Qwen4 architecture rather than a finished flagship. Shipping architectural...

Read arrow_outward
Model Release·Google

Google Adds Gemini 3.5 Transcribe, Omni 1.1 Flash

Google released Gemini 3.5 Transcribe, a speech-to-text model for live and prerecorded audio with sub-second streaming, language detection across more than 85 languages, custom vocabulary, speaker attribution, timestamps and filler-word...

Read arrow_outward
China·Tech Times

GLM-5.3-Flash Undercuts Rivals At $0.075 Input

Z.ai attached aggressive pricing to GLM-5.3-Flash, the model previously previewed as Ox Alpha: $0.075 per million input tokens and $0.25 per million output tokens under a 50% promotional discount running through September 9. For the...

Read arrow_outward
China·CNBC

MiniMax Pitches One Unified Multimodal Model

MiniMax executives used an August 28 CNBC appearance to promote One, the company's unified multimodal model, arguing for a single architecture over separate text, image and audio systems. The pitch lands in a week when Tencent, Alibaba...

Read arrow_outward
Agents·Google

Gemini Live Gets Agentic Voice Commands

Google rolled out a productivity update to Gemini Live that exposes agentic features through voice commands, letting users trigger multi-step actions without touching a keyboard. Separately, Expert Intelligence launched in Gemini...

Read arrow_outward
Trend·OpenAI

OpenAI Courts Thai Startups, Publishes Education Study

OpenAI announced a program to support Thailand's AI startup community, adding to a Southeast Asian footprint push that follows this week's Brazil expansion. It also published findings on what students gain when ChatGPT use is paired...

Read arrow_outward
The Lab

Tool of the day, field tested.

All Lab reports arrow_forward
Oriane screenshot Deep Dive 7.1/ 10
Analytics Free tier, paid from $49/mo (Plus) with Pro at $499/mo and yearly discounts about 15 minutes

Oriane

Every social listening tool reads captions; Oriane actually watches the video, and that changes what you can ask.

Oriane is for marketers, agency strategists, and creator-marketing leads who keep reverse-engineering hooks by hand and want a queryable index of what is inside social video instead of just around it. The concept is genuinely differentiated and the free tier is enough to validate whether your niche is covered, but the paid jump is steep and early reviewers say the interface still thinks in videos when teams think in accounts. Worth a one-hour test if you brief video content weekly; skip it if you post monthly.

Capability 7.5
Ease 8.0
Value 6.0
Momentum 7.5

“Finally something taking on the old and clunky brand tracking and social listening tools.”LinkedIn

Also on the radar
Productivity

Sumly.AI summarises the podcasts and calls you were never going to finish

Sumly.AI condenses audio and video into short written summaries, and the listing says it is currently free and ad free. Use it on the long-form interviews and recorded webinars in your queue, then keep only the ones where the summary raises a question you need the full source to answer. It is a triage tool, not a substitute for listening to the two things that matter.

Try it arrow_outward
Build

Scout the YC Summer 2026 big-data batch before your vendor shortlist hardens

Y Combinator's big-data directory now lists 28 funded companies, including Lyon, a two-person San Francisco team building private foundation models that run inside a bank's own cloud; it trained on 28 billion transactions for one fintech and identified premium-card converters four times more precisely than the existing rules. If you are buying prediction rather than chat, these are the teams competing with your internal feature store. Read the descriptions for the deployment model, since on-premises versus API is the real decision.

Try it arrow_outward
Mindset

Use spend data, not launch posts, to see which AI startups are actually being adopted

Brex's summer 2026 ranking of the 25 fastest-growing software startups is built from corporate card spend, which is a harder adoption signal than download counts or Product Hunt upvotes. Cross-check it against the tools your own finance team is already expensing before you approve another pilot. Growth in spend tells you a category is real; it does not tell you the winner.

Try it arrow_outward
Productivity

Efficient.app's AI shortlist is a comparison grid, so read it as one

The list was updated on 28 August 2026 and narrows 11 evaluated AI tools into ranked picks with a stated weakness for each, covering meeting notes, dictation, email, and video. The useful part is the side-by-side of where each tool falls short, which is what most roundups omit. Note the page is FTC-disclosure compliant affiliate content, so use it to build a bake-off list rather than to pick a winner.

Try it arrow_outward
The Arena

Competitive intel.

OpenAI

Cutting off Cursor's model access after SpaceX acquisition, while pushing into new markets

China Challengers

(recent) Tencent and Zhipu both push open weights as Western toolchains gate them

THE DOJO · BUILD TODAY

Price your stack against this week's open weights before the discount expires

01

Benchmark the cheap tier honestly. Route a real workload through GLM-5.3-Flash at $0.075 input and $0.25 output, then compare quality and latency against your current GPT-5.6 Sol path at $4 input and $20 output. Log the delta before the 50% promo ends on September 9.

02

Treat previews as previews. Pull Hy4 and Qwen3.8-Flash-Next weights into a sandbox for coding and multimodal evals, but keep them out of production paths. Both were shipped as architecture signals, not finished flagships.

03

Instrument your media and research intake. Point Oriane at your social video to test hook performance rather than guessing, run Sumly.AI over the podcasts and calls stacking up unheard, and scan the YC Summer 2026 big-data batch before your vendor shortlist hardens.

The Bottom Line

Open weights are now the distribution strategy

Three Chinese labs used this week to ship incomplete work into the open rather than wait for a polished flagship, and it worked as positioning. Tencent and Zhipu are pushing open weights precisely because Western toolchains keep gating them, while OpenAI cut Cursor's model access and countered with an internal claim about Astra clearing a research-intern bar nobody outside the lab can test. For builders, the practical read is that capability is arriving cheaper and faster than your procurement cycle can absorb it. Price the alternatives now, keep previews in the sandbox, and treat unverified internal benchmarks as marketing until weights or an API show up.

Want this every morning?

AI analysis, world news, markets, and tools. One briefing, delivered free.

One email per day. No spam. Unsubscribe anytime.