← → or h l to move · f fullscreen · esc exit

Living on the Frontier · Session 22 · In-person

Open Source Accelerates

August 1, 2026 · The Kannas Hotel, Chiang Mai · one week

One week, and the biggest model in the world shipped with a download link.

Two point eight trillion parameters, on a Monday, for nothing.

By dinner it was in your editor, your gateway, your terminal —

day zero everywhere, zero data retained, as if that were normal now.

The protocol put down its memory and got easier to carry.

The prices fell eighty percent and nobody even blinked.

And somewhere in a test harness, an agent with its refusals turned down

stepped out of the sandbox and into a real company’s servers.

Four services. Harvested credentials. A disclosure post on Friday.

Both labs signed a letter asking the world to slow down —

in the same week they made it cheaper to go fast.

So the weights are open. The door was open too.

What will you build with it, and what will you finally lock?

The week in numbers

2.8T
Kimi K3's parameter count — the first frontier open model in the 3-trillion class, 1M context, native vision, weights on the internet Monday
−80%
GPT-5.6 Luna's API price, now $0.20 per million input — Terra down 20% the same day
4
services an autonomous agent moved laterally through inside Hugging Face's production — the agent was OpenAI's, running a cyber eval with reduced refusals

ROADMAP

Tonight

DeepSeek ships V4-Flash — yesterday morningDEEPSEEKGPT platform updatesOPENAIAnthropic newsANTHROPIC NEWSPollPOLLIn FDE newsFDECommunity presentationPRESENTATIONOpen source accelerationOPEN WEIGHTSPolitical dramaPOLITICSAgents be hackingSECURITYEveryone ships a security modelSECURITYAsk the roomASK THE ROOMMeanwhile, everyone elseEVERYONE ELSEOpenAI NewsLAB NEWSReading materialREADING

One week, in person — the weights came down and something walked in.

DeepSeek · Jul 31 · yesterday morning

DeepSeek ships V4-Flash

The newest thing in this deck. It went into public beta yesterday morning.

  • DeepSeek-V4-Flash’s official API is live in public beta — agent capabilities “massively upgraded,” with benchmark scores far surpassing V4-Pro-Preview
  • Nobody in this room has had this for twenty-four hours. A “Flash” tier scoring past the Pro preview is the cheap-capable-agent argument landing in the same weekend as the K3 wave.
  • So — who’s going to try it before next Saturday?
deepseek_ai: DeepSeek-V4-Flash official API live in public beta

OpenAI · Jul 27–30 · platform

GPT platform updates

The frontier got cheaper on a Thursday — and then pointed itself at its own bill.

sama: major price cuts todayOpenAI: price cuts and Sol Fast modeOpenAI: Codex Security CLI open-sourcedthsottiaux: Codex Security CLI and TypeScript SDKOpenRouter: Terra and Luna 50% offOpenAI: Sol lowers serving costs 20%OpenAIDevs: Codex infrastructure optimized by Solthsottiaux: GPT-5.6 Sol SoTA on ARC-AGI-3thsottiaux: Codex usage limits reset and the Sol explanation

Anthropic · Jul 25–28 · news

Anthropic news

The protocol got lighter; the research got sharper.

AnthropicAI: discovering cryptographic weaknesses with ClaudeClaudeDevs: MCP 2026-07-28 is live and statelessadocomplete: 24 hours with Fable 5 and Opus 5kimmonismus: games will be prompted very soon

anthropic.com/news

Cognizant × Anthropic expand their partnership

Claude into enterprise clients — announced Jul 27.

mattshumer_: Opus 5 one-shotted this gameAnthropicAI: the reduced-AES result, 200-800x speedup

POLL

FDE · Jul 28–30 · the work

In FDE news

The week the “software factory” got a reality check.

dexhorthy: good software emerges from stacked automationssatyanadella: record year earningssatyanadella: decoupling the harness from the modelDoWCTO: task force embedded with the US Pacific FleetGeoffreyHuntley: software factories are real but not cracked

Living on the Frontier · Session 22

Community presentation

Take it away.

Open weights · Jul 27–30

Open source acceleration

Weights in the morning. Gateway, editor and terminal by that evening.

thinkymachines: Inkling-Small releasedKimi_Moonshot: MoonEP open-sourcedFireworksAI_HQ: Kimi K3 live day zeromodal: Kimi K3 with a custom DFlash speculatorrauchg: most powerful open-weight model in the worldvercel_dev: Kimi K3 on AI Gateway from US providerscursor_ai: Kimi K3 now in Cursoropencode: Kimi K3 in OpenCode ZenKimi_Moonshot: releasing Kimi K3 weights and technical report

Politics · Jul 26–28

Political drama

Two arguments in one week: who gets the weights, and who gets to slow down.

AMD: signing the open letter on open-weight modelsJensenHuang: defenders need a frontier AI ecosystemOpenAI: on pacing accelerationAnthropicAI: signing the frontier-pacing petitionAnthropicAI: our position on open-weight models

Security · Jul 30–31 · disclosed

Agents be hacking

Not a hypothetical. A production breach, by an agent, during an eval.

  • Hugging Face’s disclosure — an autonomous agent breached production: pipeline entry, node-level escalation, credential harvesting, lateral movement across four services. Reported to law enforcement
  • OpenAI confirms the agent was theirs — GPT-5.6 Sol plus a pre-release model, both run with reduced cyber refusals for a capability benchmark
  • The same week, Anthropic discloses three incidents — a Claude model reached the internet from a third-party eval and accessed three organizations’ real systems
  • Eval environments are now a live attack surface. At both labs. In the same seven days.

huggingface.co/blog

Security incident — July 2026

An autonomous agent entered via a data-processing pipeline, escalated at the node level, harvested credentials and moved laterally across four services.

openai.com/index

Hugging Face model evaluation — security incident

GPT-5.6 Sol and a pre-release model, run with reduced cyber refusals for a capability benchmark.

AnthropicAI: three incidents where a Claude model reached the internet during evals

Security · Jul 27 · the response

Everyone ships a security model

The industry’s answer arrived the same week as the incident.

  • Microsoft ships MAI-Cyber-1-Flash, its first cybersecurity model — built to find hard vulnerabilities in complex codebases — plus MDASH, “frontier-grade security at half the cost”
  • NVIDIA launches the Open Secure AI Alliance — shared models, tooling and research for safeguarding software and agents
  • DeepsecBench results — GPT-5.6 Sol scores highest; Kimi K3 gets half the top score at 1/5 the cost; Grok 4.5 wins best score/cost in the top 10
  • (And OpenAI’s Codex Security CLI, back on slide 6.)
nvidia: Open Secure AI Alliancevercel: DeepsecBench resultssatyanadella: MAI-Cyber-1-Flash and MDASH

Ask the room

Has anyone here actually run an open-weight model this week?

Locally, on a gateway, in your editor — anywhere. Hands up.

Everyone else · Jul 27–29

Meanwhile, everyone else

Everyone who isn’t OpenAI, Anthropic or Moonshot — a lab, a model roadmap, two products and a brand-new company.

elonmusk: Grok 4.6 around August 7SpaceXAI: Grok Voice Think Fast 2.0grok: one-prompt app publishingFishAudio: $52M seed and S2.1 Procursor_ai: Cursor Start for Indiamitchellh: starting Superlogicalssi: long-term strategic partnership with NVIDIA

OpenAI · Jul 27–30 · company news

OpenAI News

Company news, not platform news.

  • Frontier tools into researchers’ hands — a program on the premise that the benefits of frontier AI “should not be concentrated in a few companies and well-resourced labs”
  • Altman on it — “very close to models that will significantly accelerate scientific discovery… empower scientists, not figure out everything ourselves”
  • Coding agents for science — agents taking on maintenance, targeted optimization and full redesigns; researchers still define the problem
  • OpenAI Student Collective — undergrad Campus Leads, with training, funding and credits
  • A study of AI in small businesses — the generalist tool for people working outside their job description
sama: very close to models that accelerate scienceOpenAI: coding agents for scienceOpenAI: Student CollectiveOpenAI: AI in small businessesOpenAI: frontier tools into researchers' hands

Reading · Jul 25–29 · homework

Reading material

Five things worth an hour this week.

dwarkesh_sp: revenue 10x-ing on 3x computefinkd: the future is for everyonePalantirTech: legal traps in hosted AI agreementsti_morse: first interview with samadwarkesh_sp: what would be true if trendlines continue

Follow the lab.

QR code
0xnfrith.com/interface

See you next Saturday.

✕ esc
1 / ·