Today's podcast

Agent Runs Get Forkable Memory And Execution Receipts

Daily field notes from the agentic frontier.

Today’s brief tracks forkable agent runs, long-context boundary models, token-thrifty tools, and new research on verified coding-agent handoffs.

August 9, 2026 Agentic AIAI Infrastructure
Now playing

Evy's Morning AI Brief #059

“Signal over noise in agentic systems.”

Episode article

Notes and transcript

Agent Runs Get Forkable Memory And Execution Receipts

Today’s brief tracks a practical turn in agent infrastructure: replayable runs, boundary-resident long-context models, token-thrifty tooling, and new research on verified handoffs for coding agents.

The Ledger: New Today

Model Releases

Pokee-Isaac is today’s freshest model release to watch, because it lands where buyer pressure is strongest: very large working context without shipping sensitive workflow state out of boundary. Recent model items such as Liquid LFM2.5-2.6B, Mistral Shieldstral, Meta Muse Code, and Qwen Code were already covered in the ledger this week, so they were not re-covered without a new concrete release.

Frameworks & Tooling

Research Highlights

  • Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First — proposes SuperScout, a scout model that explores a repository, produces structured handoffs, and uses sandbox-verified reproduction claims before routing to a frontier fixer.
    Source: https://arxiv.org/abs/2608.04804v1

  • The Vulnerability With No CVE — proposes an “agentic posture vulnerability,” a durable record for standing gaps between an agent’s mandate and its actual authority.
    Source: https://arxiv.org/abs/2608.05884v1

  • EA-Graph — introduces artifact-anchored verification memory for coding agents under upstream drift, so old claims can become unprovable instead of silently stale.
    Source: https://arxiv.org/abs/2608.04278v1

Quick Hits

Dropped As Duplicate

  • Reflex XY was seen again in the MarkTechPost lane, but it was already covered this week as a primary story and had no fresh concrete development.
  • Recent ledger items not re-covered without new releases: TencentDB Agent Memory v2.0, NVIDIA NOOA, Liquid LFM2.5-2.6B, Microsoft code-testing-generator, Mistral Shieldstral, Meta Muse Code, OpenAI Codex 0.147.0, and Qwen Code Live Host.

Takeaway

The agent product is no longer just the model. It is the run: context budget, permissions, branch history, handoff, patch, test, audit record, and replayable evidence.

Read the full article