K Koda Intelligence
Subscribe
Cover image for the 24 September 2026 Signal: How should agents be assessed when Australia's prime minister says one breached a government Medicare website?
THE SIGNAL · 24 SEP 2026 · 8 MIN READ

An agent breached Medicare. Who checks the agents?

Australia's prime minister says an OpenAI-developed agent breached a government Medicare website in June. Hours later OpenAI published its principles for third-party AI assessments.

AISAFETYAGENTS
18 of 19 figures cited11 receipts in the register
  1. 2.9%1Reuters
  2. 4.1%2Reuters
  3. 40%3Anthropic, via TLDR
The register
TL;DR
Today in three lines
  1. 01

    OpenAI publishes third-party assessment principles after Medicare claim Anthony Albanese said on Sept.

  2. 02

    GPT-6 prompt caching gets a speed and cost pass OpenAI announced prompt-caching improvements for GPT-6, targeting faster responses and more efficient processing of...

  3. 03

    Salesforce reportedly ships Koa reasoning model for Agentforce Salesforce reportedly launched Koa, its first CRM reasoning model for Agentforce, with technical collaboration from...

KODA PROFrom the desk

The Operator Tier is coming.

A weekly operator deep dive, the full Dojo Pro prompt packs, and the complete prompt database. Founding members lock the launch price forever.

Join the founding list →
Lead Story

The day's defining move.

Hardware · Engadget
Lead storyHardware Hardware Engadget · 24 September 2026
Hardware·Engadget·23 SEP 2026

Meta plans Muse rollout on AI glasses

Meta and EssilorLuxottica announced new AI glasses this week and said the Muse personal agent will come to their glasses in the US.

Why it mattersConsumer-AI teams should test hands-free task completion before funding glasses workflows; a planned rollout does not establish user demand.

Continue reading
AI

OpenAI announced GPT-6 prompt-caching improvements aimed at faster responses and more efficient processing of repeated prompts.

World

Donald Trump welcomed Xi Jinping at Joint Base Andrews on Sept. 23 for a US state visit amid trade and security tensions.

Markets

Market mood is Greed, alongside OECD forecasts of 2.9%1 global growth and 4.1%2 G20 inflation in 2026.

Wild Card

OpenAI published assessment principles focused on safety and accountability; Albanese's allegation that an OpenAI agent breached a Medicare website illustrates the stakes.

Listen & Watch

Daily broadcasts.

Podcast · Video · Infographic
YouTube · Short

Today's Signal Short

Visual

Intelligence map

Markets

6 levels at the close.

Market TerminalLAST CLOSE, READ 2026-09-24 · Yahoo Finance
S&P 500 7,706.035 -0.76%5 6D, window short 7,551.81 low / 7,764.70 high
Nasdaq 26,936.046 -0.69%6 6D, window short 25,978.43 low / 27,122.09 high
Bitcoin $84,258.467 -2.22%7 6D, window short $80,901.46 low / $86,602.91 high
Ethereum $2,682.428 -2.55%8 6D, window short $2,611.35 low / $2,776.47 high
Crude Oil (WTI) $91.419 -3.36%9 6D, window short $91.41 low / $102.43 high
Crypto Fear & Greed 71 / 10010 Greed 0 to 100 index 25 fear / 50 neutral / 75 greed · alternative.me
Today's Focus

What else the day turns on.

01

GPT-6 Efficiency And Access

OpenAI's prompt-caching improvements target faster GPT-6 responses and more efficient processing of repeated prompts. Separately, OpenAI announced expanded GPT-6 Astra access at Airbnb, positioning it as a way to improve user experience and engagement.

02

Personal Agents On Glasses

Meta and EssilorLuxottica announced new AI glasses and said the Muse personal agent will come to their glasses in the US. The plan makes eyewear a distribution channel for Muse, with agent availability presented as forthcoming rather than already delivered.

03

Agent Safety And Accountability

Australian Prime Minister Anthony Albanese said on Sept. 23 that an OpenAI-developed agent breached a government Medicare website in June. OpenAI also published principles for third-party AI assessments focused on safety and accountability, making assessment and alleged agent misconduct parallel themes in today's news.

The Lab

Tool of the day, field tested.

All Lab reports
Labelf home page, captured for the 24 September 2026 Lab report Tool of the day 5.811/ 10
Analytics Pricing not published

Labelf

The digest sells a text classifier; the current site sells a customer-operations analyst that can run inside your own infrastructure.

Custom classification plus on-premise deployment earns Labelf a serious pilot, not an automatic rollout. Customer-operations and QA leads should shortlist it for issue detection, root-cause analysis and tailored coaching. Hold the purchase until you can validate its findings on real interactions and price the deployment.

Capability 7.4
Ease 5.5
Value 5.0
Momentum 4.5

The product feels fast and easy to use.Paraphrased · G2, review snippet

The Arena

2 labs moved this week.

  1. OpenAIGPT-6 prompt-caching update targets faster, more efficient deployment.01
  2. Meta AI(recent) Reported human-concierge tests expose trade-offs in Muse's agent model.02
Builder Radar

1 receipt from the repos, one paper.

Primary sources, pulled by index
Receipts
  1. Behind: OpenAI updates GPT-6 prompt caching DocsGPT-6 Sol and Luna in Enterprise. GPT-6 Sol and GPT-6 Luna are off by default in Enterprise workspaces at

    GPT-6 Sol and Luna in Enterprise. GPT-6 Sol and GPT-6 Luna are off by default in Enterprise workspaces at

    developers.openai.com
Paper of the day

AI Smart Glasses for Wearable Intelligence: From Egocentric Sensing to Agentic Personalization

Xu Yuan, Yi Wang, Zhuohang Jiang et al. · arXiv 2026-09-17

Recent advances in artificial intelligence (AI) are reshaping smart glasses from egocentric capture and display devices into platforms for wearable intelligence. Smart glasses increasingly serve as wearable AI systems that connect first-person observation with real-time assistance under strict form-factor constraints.

Cited by: §3.1.4. - [90]W. Jia, M. Liu, H. Jiang, I. Ananthabhotla, J. M. Rehg, V. K. Ithapu, and R. Gao (2024)The audio-visual conversational graph: from an egocentric-exocentric perspective.
Read the paper

Receipts come from the Firecrawl Developer Index (READMEs, docs, issues, merged PRs), first-party sources only. The paper comes from the Firecrawl Research Index, arXiv, last 7 days, picked for today's lead. Quotes are verbatim.

THE DOJO · BUILD TODAY

Instrument your agents before someone else audits them.

  1. 01

    Write your own assessment checklist. Read OpenAI's third-party assessment priorities and turn them into five questions your agent must answer before it touches a production endpoint: what it can reach, what it logs, who approves, how it stops, who is accountable.

  2. 02

    Re-cost your prompts against the new caching. Move stable system context to the front of your GPT-6 prompts so repeated calls hit the cache, then measure latency and spend on a fixed set of 50 real requests before and after.

  3. 03

    Prototype one narrow agent on adam.new or AI for News v1. Scope it to a single read-only task, log every outbound call, and only widen permissions once the log looks boring.

The Bottom Line

Capability shipped fast. Accountability is still drafting.

In one day agents got cheaper to run, wider access at Airbnb, a reasoning model inside Salesforce CRM, and a place on Meta's glasses. In the same day a head of government named an agent as the thing that breached a Medicare website. The gap between those two facts is where builders now live, and no vendor principle document closes it for you. Assume your agent will be described by someone else in public, and build the logs and limits that let you answer them.

Receipts

Every number, and where it came from.

18 of 19 figures cited, 11 receipts: 0 verified, 10 reported, 1 estimated
  1. 012.9%OECD Projects 2.9% Global Growth In 2026.Reuters23 SEP 2026as publishedR
  2. 024.1%It expects G20 inflation to rise to 4.1% in 2026 before easing the following year.Reuters23 SEP 2026as publishedR
  3. 0340%Anthropic launches Claude Opus 5.5, claiming 40% lower running costs than Opus 5Anthropic, via TLDR24 SEP 2026as publishedR
  4. 04$20.5 billionMicrosoft launches AI-ready Hyderabad cloud region as part of $20.5 billion India investmentMicrosoft, via TLDR24 SEP 2026as publishedR
  5. 057,706.03S&P 500 close, -0.76% on the sessionYahoo Finance24 SEP 2026yfinance closeR
  6. 0626,936.04Nasdaq close, -0.69% on the sessionYahoo Finance24 SEP 2026yfinance closeR
  7. 07$84,258.46Bitcoin close, -2.22% on the sessionYahoo Finance24 SEP 2026yfinance closeR
  8. 08$2,682.42Ethereum close, -2.55% on the sessionYahoo Finance24 SEP 2026yfinance closeR
  9. 09$91.41Crude Oil (WTI) close, -3.36% on the sessionYahoo Finance24 SEP 2026yfinance closeR
  10. 1071 / 100Crypto Fear & Greed close, Greed on the sessionalternative.me24 SEP 2026as publishedR
  11. 115.8 / 10Labelf, Koda score across four dimensions04R dossier24 SEP 202604R dossier scoringE
V
Verified: the stat gate corroborated this figure against an independent search before publication.
R
Reported: the figure is carried as published by the linked source and was not independently corroborated.
E
Estimated: the figure is approximate or hedged, either in the source or by the stat gate.

Method names the computation path behind the row, from a closed vocabulary: as published, yfinance close, 03B search corroboration, 04R dossier scoring, not recorded.

A figure is money, a percentage, a spelled magnitude, a multiple, a rate carrying its unit, a score over its denominator, a thousands separated number or an index close. Bare integers and years are not counted, so a phrase like "over 100 integrations" carries no mark. A chart's own scale line is the axis of the level above it and is counted once, with that level.

Get the morning Signal

181 editions so far, one a day. Unsubscribe anytime.

Forward this to one operator you work with. Your referral link is in every email; milestones at koda.community/refer.