profile

Your must-have AI newsletter

OpenAI's New Cyber Bet


Let’s take away what matters in AI every day and stay ahead with 50,000⁺ founders, builders, and tech readers.

BIG TECHAWAYS

TODAY's BIG STORIES

#1. AI Cyber Gets a Fixing Loop

OpenAI introduced Patch the Planet, a program with Trail of Bits designed to help open-source maintainers find, verify, patch, test, and coordinate disclosure for security vulnerabilities. In parallel, Daybreak puts GPT-5.5-Cyber and Codex Security in the hands of vetted defenders, with support for attack-path tracing, threat modeling, and reviewable codebase-specific patches.

The important part is the workflow. OpenAI is not stopping at a model that can find bugs; it is packaging discovery, verification, patching, and human review into a defensive operating loop. In security, the scarce layer is not always finding more vulnerabilities. It is turning a finding into a trusted patch that maintainers can accept, test, and disclose through a clear process.

#2. China AI's Strange Power

Chinese AI is sending contradictory signals. Open-source models help Chinese labs spread influence and pull developer ecosystems toward them, even before monetization is clear. But at the infrastructure layer, the US is still spending far more on AI compute, while parts of China's open-source AI market are being valued far above current revenue.

China can be strong in distribution through open source, constrained by the compute gap, and priced by public markets as a long-term bet on ecosystem power all at the same time.

#3. Cursor Moves Up the Stack

Cursor said its Compile keynote included three announcements, including a new model being trained with SpaceX. Public recaps point to three main pieces: Origin, a mobile app / mobile agents, and a model trained from scratch with SpaceX. Cursor has also announced a partnership using xAI's Colossus infrastructure to scale model intelligence.

This is Cursor pushing beyond an AI autocomplete IDE. If Origin is a git forge for the agentic era, mobile agents expand the workflow surface, and Colossus gives Cursor a compute path for proprietary coding models, then the coding-agent market may move up a layer: advantage comes from controlling the IDE, agent workflow, code-hosting surface, and model-training pipeline together.

SHIFT SIGNALS

BEHIND THE HEADLINES

  • NVIDIA is pushing back on the data-center water narrative: NVIDIA cites the Manhattan Institute to say data centers account for 0.2% of daily water usage in the US and argues that newer cooling approaches have reduced usage. Infrastructure companies are preparing counter-metrics for the environmental backlash around AI buildout.
  • Agent loops are getting a clearer operating shape: Lenny's Newsletter describes schedules, goals, and subagents as a practical way to design agent loops in Claude Code and Codex. The pattern to watch is the shift from one-off prompting toward workflows with cadence, scope, and review.
  • Software optimization can move inference economics quickly: SemiAnalysis says GB200 NVL72 serving costs for the Kimi architecture fell 2.5x in less than 70 days through software improvements. If that claim holds, inference economics still have room to improve through software, not just more hardware.
  • AI demand may be showing up in China's export data: Kobeissi says China's semiconductor exports rose sharply year over year in May, alongside outbound shipments of computers and parts. It is a macro signal worth watching: AI demand may be pulling on the hardware supply chain, not only model-lab revenue.
  • OpenAI Bidi 1 appears as a reported signal for the next voice layer: OpenAI is preparing "Bidi 1" for a future web release, with a new voice model appearing in settings and a yellow voice-mode bubble. It is still a reported product-surface signal, but worth tracking because voice could become the next interaction layer for ChatGPT.
  • AI coding is entering M&A diligence: Financial Times reporting that Bain is using AI coding tools to recreate pieces of target-company software, generating hundreds of rough prototypes for due diligence. If that pattern spreads, AI coding becomes a tool for evaluating assets before acquisition, not just shipping products.
  • Fast prototypes can also create false confidence: Chamath argues that "vibing" the UI of a complex app like Google Search may not help anyone understand the deeper system underneath. It is a useful counterweight to the Bain signal: prototypes can speed up learning, but shallow UI recreation can make decision-makers feel like they understand more than they do.

AI IN ACTION

Let Agents Prep Repo Work

GitHub Agentic Workflows brings agents into GitHub Actions for reasoning-heavy tasks like issue triage, CI failure analysis, and docs updates.

The fastest way to test this is not to automate the whole repo. Pick one small task that creates weekly context switching, and let an agent prepare the draft before you review it.

Use a real repo and choose an existing trigger: a new issue, a failed CI run, a PR that changes an API, or docs that need updating after code changes. For that trigger, write one specific sentence: "When this happens, what should the agent give me so I can decide faster?" A new issue might return a label, summary, and suggested owner. A failed CI run might return the likely cause and next command. An API-changing PR might return a missing-docs checklist.

The boundary to protect is authority. The agent can read context and prepare an artifact, but anything that can merge, publish, or notify someone still goes through a review gate.

After trying it on a few old cases, keep the workflow only if the output helps you decide faster without rereading the entire original input. If the agent is directionally right but too confident, add a guardrail. If it creates more review work, drop that workflow and choose a narrower trigger.

WORTH YOU TIME

ANALYSIS & RESEARCH WORTH EXPLORING

Agents matter more than bubble talk

Stratechery puts agents at the center of the platform-economics story instead of only debating whether AI is in a bubble. The lens is useful because it points to where value capture may emerge as software begins doing more work on its own.

AI Index 2026 is the macro map for AI

The 2026 AI Index pulls together data on capability, adoption, economy, policy, public opinion, and technical performance. It is a useful reference for placing daily AI headlines in a longer arc, especially when the market is pulled around by individual launches or drama cycles.

The 149-person company and the tax notch of AI coding

Haseeb Qureshi argues that AI coding subscription pricing creates a "tax notch" around the enterprise threshold: small teams get very low marginal token costs, while larger companies are pushed toward enterprise or API economics. The thesis turns AI coding from a productivity story into a company-design story: startups may be incentivized to keep headcount lower, use agents more heavily, and compete with incumbents through a different cost structure.

Have questions? Hit reply to this email and we'll help out!

600 1st Ave, Ste 330 PMB 92768, Seattle, WA 98104-2246
Unsubscribe · Preferences

Your must-have AI newsletter

Join 50,000+ builders and tech readers cutting through the noise and staying ahead in the AI era.

Share this page