Skip to content
K Koda Intelligence
Subscribe
THE SIGNAL · 14 AUG 2026 · 5 MIN READ

Speed is the new spec sheet

Google shipped Gemini 3.7 Flash while 3.5 Pro stays late, and OpenAI previewed Ultrafast running GPT-5.6 Sol at up to 14 times normal speed. Latency, not parameter count, is the pitch now.

AIMODELSMARKETS
TL;DR
Today in three lines
  1. 01

    Latency becomes the headline model feature Google released Gemini 3.7 Flash on Aug 13 promising fewer prompts per app and lower token prices, while 3.5 Pro still...

  2. 02

    Nvidia stakes an open trillion-parameter claim Nemotron 4's largest variant is expected to reach at least one trillion parameters, but training is unfinished and...

  3. 03

    Washington aims 100% tariffs at imported drones The Trump administration is imposing 100% tariffs on certain imported drones and components, hitting a hardware...

SPONSOR KODAFrom the desk

Put your product in front of AI operators.

Koda publishes a fact-checked AI intelligence briefing every morning: site, email, podcast, and video. Founding sponsors lock today's rates for 6 months as the audience compounds.

View the media kit →
Lead Story

The day's defining move.

Model Release · Google
Lead Story
Model Release·Google·14 August 2026

Google Ships Gemini 3.7 Flash, 3.5 Pro Still Late

Google released Gemini 3.7 Flash on Aug 13, claiming better debugging, more deployable production-ready code on the first attempt, app builds in fewer prompts, and lower token prices than the prior Flash release. The company also moved its Gemini Spark productivity agent onto 3.7 Flash the same day and added...

Continue reading
AI

Google released Gemini 3.7 Flash on Aug 13 with cheaper tokens, better debugging and more first-attempt production code, while the promised 3.5 Pro remains late.

World

The Trump administration is imposing 100% tariffs on certain imported drones and components, a category still dominated by Chinese suppliers, which would reprice fleets for agriculture, surveying, inspection and film operators.

Markets

Mood reads Fear, with crude sliding about 2% to near $81 even as Iran insists no vessel transits Hormuz without its authorization.

Wild Card

Nvidia gave away NeMo Switchyard, the routing layer that decides which model answers, on the same news cycle it teased a trillion-parameter Nemotron 4: own the cheap plumbing, then sell the expensive endpoint.

Markets

Market snapshot.

Market TerminalLive
S&P 500 7,798.99 +0.65%
Nasdaq 26,803.03 +0.81%
Bitcoin $63,361.17 -0.07%
Ethereum $1,882.88 +0.25%
Crude Oil (WTI) $81.23 -2.45%
Crypto Fear & Greed 29 Fear
Today's Focus

The three signals that move the day.

01

Latency Replaces Parameter Count

Google's pitch for Gemini 3.7 Flash is fewer prompts per app and lower token prices; OpenAI's Ultrafast preview runs GPT-5.6 Sol at up to 14 times normal speed. With GPT-6 still unannounced and Gemini 3.5 Pro still late, both labs are competing on responsiveness and cost per task rather than a new flagship. That is what a market does when the next capability jump is not ready to ship.

02

Nvidia's Open Model Gambit

Nemotron 4's largest variant is expected to reach at least one trillion parameters, but training is unfinished and Nvidia has not announced a release date, so this is roadmap signal, not artifact. Paired with the open-source NeMo Switchyard router, the strategy is to occupy both ends of the agent stack: the weights developers download free and the routing library that dispatches their requests.

03

Monetization Arrives At OpenAI

Hiring Dali Rajic as Chief Revenue Officer puts an enterprise sales veteran over a product with more than a billion weekly ChatGPT users, in the same week as enterprise deployment case studies and days after ads testing began inside ChatGPT. The research narrative is being fitted with a revenue engine. Expect pricing, packaging and ad inventory to shape the next model release as much as benchmarks do.

Listen & Watch

Daily broadcasts.

Podcast · Video · Infographic
The Daily Deep Dive

Listen to today's briefing

YouTube · Short

Today's Signal Short

Visual

Intelligence map

The Wire

AI intelligence.

Model Release·Google

Google Ships Gemini 3.7 Flash, 3.5 Pro Still Late

Google released Gemini 3.7 Flash on Aug 13, claiming better debugging, more deployable production-ready code on the first attempt, app builds in fewer prompts, and lower token prices than the prior Flash release. The company also moved...

Read
Open Source·Reuters

Nvidia Building Trillion-Parameter Open Nemotron 4

Nvidia is developing Nemotron 4, an open model family whose largest variant is expected to reach at least one trillion parameters, according to The Information. Training is not finished and Nvidia has not announced a...

Read
Model Release·OpenAI

OpenAI Previews Ultrafast Mode At 14X Speed

OpenAI published a preview of Ultrafast mode, running GPT-5.6 Sol at up to 14 times its normal speed, aimed at interactive use where latency dominates the experience. The company paired it with a builder's guide to GPT-5.6 covering...

Read
Enterprise·OpenAI

OpenAI Hires Dali Rajic As Revenue Chief

OpenAI named Dali Rajic Chief Revenue Officer, putting an enterprise sales veteran in charge of monetizing a product now serving more than a billion weekly ChatGPT users. The appointment lands the same week OpenAI published case studies...

Read
Open Source·NVIDIA

Nvidia Open-Sources NeMo Switchyard Agent Router

Alongside its Nemotron news, Nvidia released NeMo Switchyard, an open-source routing library for AI agents that directs requests across models and tools. Routing is where agent cost control actually happens: sending cheap steps to small...

Read
The Lab

Tool of the day, field tested.

All Lab reports
Deep Dive 5.8/ 10
Productivity Free under five minutes

Gift Genie AI

It is a gift finder, sure, but it is also the cheapest five-minute lab for learning how much context a recommendation model actually needs before it stops guessing.

For anyone stuck on a birthday present tonight, this is a no-account, no-cost brainstorm that beats staring at a blank search bar. For builders, the real value is diagnostic: run the same recipient with two sentences and then with six and you will feel exactly where prompt specificity starts paying off. As a product it is thin, with generic and repetitive output flagged in hands-on testing, so treat it as an idea generator rather than a shopping engine.

Capability 5.0
Ease 9.0
Value 7.5
Momentum 4.0

“Reviewers aggregated on a directory page say it removes the stress of gift hunting and surfaces creative ideas they would not have thought of themselves.”Autonoly directory review aggregation

Also on the radar
Productivity

Letaido is Ahrefs' bet that marketing reporting should run itself

Ahrefs launched Letaido on August 12 as an AI agent workspace aimed at the recurring work marketing teams never finish: competitor monitoring, keyword and SERP research, and the weekly reports that eat a day. Start it on one report you already produce manually so you can diff the agent's output against a known-good version before you trust it with client deliverables.

Try it
Build

Energy is a desktop agent that reads your context before it acts

Built by an ex-OpenAI researcher, Energy automates multi-step computer work by first pulling context from email, local files, the browser, and connected apps, then executing. That ordering matters: the failure mode of most desktop agents is confidently acting on a stale or partial picture, so test it on a task where you can verify every input it claims to have read.

Try it
Productivity

Grok Bot gets its own cloud computer, which is the part to plan around

xAI's Grok Bot is pitched as an always-on teammate: it receives a dedicated cloud machine, signs into your existing tools with real credentials, and finishes multi-step jobs unsupervised. Before delegating anything, decide which accounts it gets and create scoped logins rather than sharing your own, because an unsupervised agent with your session is an audit problem waiting to happen.

Try it
Creativity

getimg.ai for image work you need in bulk, not one hero shot

getimg.ai is a generation and editing suite that sits in the top tier of TAAFT's image category by usage, with tens of thousands of saves logged on its directory page. It earns its keep on volume tasks like variant testing and background replacement across a batch; for a single flagship asset you will still spend the time in a real editor.

Try it
The Arena

Competitive intel.

OpenAI

Ships Ultrafast preview on Cerebras silicon, replaces revenue chief after eight months

Anthropic

(recent) Claude Code Auto Mode becomes the default today; joins OpenAI on a new cost metric

THE DOJO · BUILD TODAY

Rebuild your stack around latency and routing, not leaderboards

01

Benchmark Flash against your own prompt chain. Take one production workflow and rerun it on Gemini 3.7 Flash, counting prompts to a working build and total token cost rather than trusting the launch claims. If it needs fewer round trips, that is your real win.

02

Put a router in front of your agents. Try Nvidia's open-source NeMo Switchyard to direct requests across models and tools, then log which calls actually needed the expensive model. Cost control lives at the routing layer, not in the model choice.

03

Automate one reporting loop this week. Point Letaido, launched August 12, at a single recurring marketing report, or test Energy's desktop agent on a task where reading local context matters more than raw model power.

The Bottom Line

The spec sheet now reads tokens per second

Two of the three biggest labs spent today selling speed and cost instead of capability, and that tells you where the margin pressure sits. Nvidia's trillion-parameter Nemotron 4 is the counter-narrative, but with training unfinished and no release date it is a roadmap, not a product. Meanwhile OpenAI hiring Dali Rajic as Chief Revenue Officer signals that a billion weekly ChatGPT users are about to be monetized in earnest. Build for the cheap fast model you can actually ship on, and treat routing as the place your bill gets decided.

Want this every morning?

AI analysis, world news, markets, and tools. One briefing, delivered free.

One email per day. No spam. Unsubscribe anytime.

Forward this to one operator you work with. Your referral link is in every email; milestones at koda.community/refer.