← All editions
Edition · Sat, Aug 8, 2026

On Fri Aug 7, the two reference coding-agent harnesses ship major cuts on the same dayAnthropic's Claude Code v2.1.224 lands self-hosted runners for Team + Enterprise, cross-machine SendMessage and ListAgents so sessions can talk across your fleet, JWT-aware credential masking, AWS SigV4 re-signing, an archive plugin source with SHA-256 pinning, and removes the 200-subagent-per-session spawn cap; OpenAI's Codex v0.147.0 stable ships portable Agent Plugins with local / personal / workspace / remote catalog search, an --approve-for-me flag for auto-reviewed approvals, opt-in support for the MCP 2026-07-28 protocol (paginated discovery, multi-round requests, non-blocking server startup), Cursor-managed skill imports, bearer-token redaction across displayed commands and replay, and removes the deprecated codex exec --full-auto flag. On Thu Aug 6, Claude Code v2.1.223 lands four permission-bypass patches at once: a Bash permission bypass where a crafted command could hide parts of itself from permission checks; a permission-prompt bypass using tabs and invisible Unicode to hide part of the command from the approval dialog; a workflow-script sandbox escape using dynamic import() to run code outside the sandbox; and a permission gap where an agent definition's bypassPermissions mode ignored the org-level bypass-permissions disable policy. On the agentic-CX capital tape, Klaviyo (NYSE: KVYO) agrees on Wed Aug 5 to acquire the 25-person Agency team and technology and install co-founder Elias Torres as Chief Product Officer post-close (Q3 2026) to lead the Composer campaign-builder and Customer Agent post-sale-support product lines, reporting to co-CEO Andrew Bialecki; Omilia on Thu Aug 6 closes a $67M Series B led by Expedition Growth Capital for its agentic self-learning voice-first CX platform already running at Capital One, Discover and Taco Bell, on live ARR up 10× since Series A to over $60M and no interim equity raise; and Naïve on Wed Aug 6 lands a $28.5M Series A led by Nexus Venture Partners (with Y Combinator, Zetta, Liquid 2, JD Sherman, Gokul Rajaram, Zachary Sims, Robert Chatwani) for its unified API + serverless stack for autonomous companies, on 30,000+ developer signups within months and ARR up 10× to low-double-digit millions in six months. On the agent-plumbing GitHub tape, petergyang/human-review holds 574 stars as a visual Google-Doc-style comment layer that lets a reviewer mark up HTML and Markdown produced by an AI agent and push feedback back into the harness; 0xwilliamortiz/claude-red ships on Wed Aug 5 and hits 559 stars in three days as a curated library of offensive-security SKILL.md files for Claude and Codex (SQLi, shellcode, EDR evasion, exploit development); cristicretu/diri ships on Tue Aug 4 as a native macOS Rust / Swift / GPUI orchestrator that runs Claude Code, Codex, Cursor, Gemini and shells in parallel across git worktrees and remote hosts (222 stars); AMAP-ML/LongHorizon-Harness ships on Tue Aug 4 as Alibaba AMAP-ML's Python long-horizon computer-use harness with fresh-context execution, durable verified state, independent auditing, recoverable progress, and native Claude Code / Codex / OpenClaw integration (387 stars); and Anionex/agent-vision-toolkit ships on Fri Aug 1 as a vision toolkit and skill for text-only LLMs (DeepSeek, GLM, Kimi) covering image Q&A, long-screenshot OCR, frontend UI restoration and GUI automation with seamless integration for Codex, Claude Code, Pi, Oh My Pi and OpenCode (348 stars).
— the throughline is a harness week: on the same 48 hours the two reference coding-agent harnesses both tighten the permission model and extend the fleet primitive (self-hosted runners, cross-machine SendMessage, MCP 2026-07-28 paginated discovery, --approve-for-me auto-approval), the agentic-CX buyer lands ~$150M of Series capital across three product lines (Composer + Customer Agent, voice, autonomous-company infrastructure), and five agent-plumbing repos cross the 200-star line filling the skills, review, orchestration and vision-for-text-only-models gaps around the harnesses.

11 SIGNALS WINDOW: AUG 1 – AUG 8 SOURCES: CLAUDE CODE CHANGELOG · OPENAI CODEX GITHUB · KLAVIYO NEWSROOM · TECHCRUNCH · SILICONANGLE · PYMNTS · FINSMES · YAHOO FINANCE · CMSWIRE · INVESTING.COM · THESAASNEWS · DEALROOM · PULSE2 · CITYBIZ · EU-STARTUPS · VENTUREBURN · STOCKTITAN · GITHUB (0XWILLIAMORTIZ · PETERGYANG · CRISTICRETU · AMAP-ML · ANIONEX)

Fri Aug 7 is the day the two reference coding-agent harnesses both ship majors on the same day. Anthropic's Claude Code v2.1.224 lands self-hosted runners — a new claude self-hosted-runner command turns any machine or container into a place Claude Code web, mobile and desktop sessions can run, on Team and Enterprise plans — alongside cross-session SendMessage across machines, ListAgents to discover them, an archive plugin source that installs plugins from a zip over HTTPS with optional SHA-256 pinning (no git or npm), JWT-aware credential masking and AWS SigV4 re-signing for sandbox networking, and the removal of the 200-subagent-per-session spawn cap for long-running sessions. OpenAI's Codex v0.147.0 stable, cut the same day, ships portable Agent Plugins with local / personal / workspace / remote catalog search, an --approve-for-me flag that enables automatically reviewed approvals, persistent, manually-ordered conversation sections, importing of Cursor-managed skills that syncs into imported Claude and Cursor conversations without duplicates, opt-in support for the MCP 2026-07-28 protocol (paginated discovery, multi-round requests, non-blocking server startup), secret and complete-bearer-token redaction across displayed commands and replayed history, and the removal of the deprecated codex exec --full-auto flag in favour of --sandbox workspace-write. One day earlier — Thu Aug 6Claude Code v2.1.223 lands four permission-bypass patches at once: a Bash permission bypass where a crafted command could hide parts of itself from permission checks; a permission-prompt bypass using tabs and invisible Unicode to hide part of the command from the approval dialog; a workflow-script sandbox escape via dynamic import() that ran code outside the workflow sandbox; and a permission gap where an agent definition's bypassPermissions mode ignored the org-level bypass-permissions disable policy. On the agentic-CX capital tape, Klaviyo (NYSE: KVYO) agrees on Wed Aug 5 to acquire the 25-person Agency team and technology and install Agency co-founder Elias Torres as Chief Product Officer post-close (expected Q3 2026), reporting to co-CEO Andrew Bialecki, to lead Klaviyo's agent product lines: Composer (marketing-campaign builder) and Customer Agent (post-sale returns and order tracking). Agency had raised $32M from Sequoia, Menlo Ventures and Felicis before the deal; the Agency standalone service winds down Aug 31, 2026. Omilia on Thu Aug 6 closes a $67M Series B led by Expedition Growth Capital for its agentic self-learning voice-first customer-experience platform for large enterprises (Capital One, Discover, Taco Bell) — Athens- and New York-based, founded two decades ago, live ARR up 10× since Series A to over $60M with no intervening equity, and the proceeds fund a US office opening in H2 2026 and global commercial scaling. Naïve on Wed Aug 6 raises a $28.5M Series A led by Nexus Venture Partners, with Y Combinator, Zetta, Liquid 2 and angels Gokul Rajaram, Tim Zheng (Apollo), JD Sherman (ex-HubSpot), Gert Lanckriet (Amazon), Robert Chatwani (Docusign) and Zachary Sims (Codecademy), for its unified API + serverless operating stack for autonomous companies covering legal incorporation, financial management, cloud compute and multi-agent orchestration; the Palo Alto lab reports 30,000+ developer signups within months of launch and annual run-rate revenue up 10× to low-double-digit millions over the six months ending August 2026. On the agent-plumbing GitHub tape, five repos cross the 200-star line in early August filling gaps around the harnesses: petergyang/human-review at 574 stars ships a visual Google-Doc-style comment layer for reviewing HTML and Markdown produced by an AI agent and pushing feedback back into the harness; 0xwilliamortiz/claude-red (Wed Aug 5) hits 559 stars in three days as a curated offensive-security SKILL.md library (SQLi, shellcode, EDR evasion, exploit development) primed for Claude and Codex; cristicretu/diri (Tue Aug 4) at 222 stars ships a native macOS Rust / Swift / GPUI orchestrator that runs Claude Code, Codex, Cursor, Gemini and shells in parallel across git worktrees and remote hosts; AMAP-ML/LongHorizon-Harness (Tue Aug 4) at 387 stars ships Alibaba AMAP-ML's Python long-horizon computer-use harness with fresh-context execution, durable verified state, independent auditing, recoverable progress and native Claude Code / Codex / OpenClaw integration; and Anionex/agent-vision-toolkit (Fri Aug 1) at 348 stars ships a vision toolkit and skill for text-only LLMs (DeepSeek, GLM, Kimi) — image Q&A, long-screenshot OCR, frontend UI restoration and GUI automation — with seamless integration for Codex, Claude Code, Pi, Oh My Pi and OpenCode. Throughline: the two reference coding-agent harnesses both tighten the permission model and extend the fleet primitive on the same 48 hours, the agentic-CX buyer lands ~$150M of Series capital across three distinct product lines (Composer + Customer Agent, voice, autonomous-company infrastructure), and the agent-plumbing GitHub cohort ships five repos over the 200-star line filling the skills, human-review, orchestration and vision-for-text-only-models gaps around them.

01

The coding-agent harnesses ship the same 48 hours — Claude Code v2.1.224 self-hosts runners and cross-machine SendMessage, Codex v0.147.0 opts into MCP 2026-07-28, and v2.1.223 patches four permission-bypass exploits

01

Anthropic on Fri Aug 7 releases Claude Code v2.1.224 — the marquee cut is self-hosted environments (claude self-hosted-runner turns any of your own machines or containers into a place Claude Code web, mobile and desktop sessions can run, on Team and Enterprise plans), cross-session SendMessage across machines (Claude Code sessions can now message each other across your fleet, with ListAgents to discover them on macOS and Linux), an archive plugin source that installs plugins from a zip over HTTPS with optional SHA-256 pinning (no git or npm), JWT-aware credential masking and AWS SigV4 re-signing for sandbox networking, and removal of the 200-subagent-per-session spawn cap so long-running sessions no longer refuse new agents; the release also fixes a long-project-path session-directory collision, sandbox filesystem deny entries silently bypassable on Linux and macOS when written with a trailing slash, sandbox violation details never appearing in Bash tool results, and Remote Control cross-session-history leakage after a stale server session; per Claude Code changelog

Fri Aug 7 2026 · Claude Code v2.1.224 · New: claude self-hosted-runner (Team + Enterprise) · Cross-machine SendMessage + ListAgents (macOS + Linux) · archive plugin source with optional SHA-256 pinning · JWT-aware sandbox credential masking (decode: "jwt" + maskClaims) · AWS SigV4 re-signing (awsPairs + sigv4, requires network.tlsTerminate) · Removed: 200-subagent-per-session spawn cap · Fixes: long-path session-dir collisions, deny-write trailing-slash bypass, missing sandbox violation details, stale Remote Control history bleed

Two reads. (1) Self-hosted runners is the fleet primitive Claude Code has been reaching for since cross-session SendMessage arrived in the v2.1.16x cycle in July (prior editions). Cross-session was same-machine; v2.1.224 makes it cross-machine, and self-hosted-runner makes your own hardware the hosting substrate for web, mobile and desktop sessions. That is the shape a coding-agent harness takes when the enterprise buyer will not accept Anthropic-run runtime for data-sovereignty reasons — the self-hosted lane is the compliance answer, and cross-machine SendMessage is the collaboration answer on top of it. (2) The removal of the 200-subagent-per-session spawn cap is the “we-tested-and-the-workflow-hits-the-cap” signal: real production workflows are routinely spawning more than 200 subagents in a single session, which is the shape a long-horizon-harness workload takes when the parent orchestrator is doing multi-repo migrations, PR-fleet reviews or codebase audits. Read alongside Codex v0.147.0's Agent Plugins (item 03) and the MCP 2026-07-28 stateless spec (prior edition), the Aug 2026 shape is the harness is now the fleet manager, not the terminal.

02

Anthropic on Thu Aug 6 releases Claude Code v2.1.223 — four permission-bypass patches at once: (1) a Bash permission bypass where a crafted command could hide parts of itself from permission checks; (2) a permission-prompt bypass where commands padded with tabs or invisible Unicode could hide part of the command from the approval dialog; (3) a workflow-script sandbox escape where scripts could use dynamic import() to run code outside the workflow sandbox; (4) a permission gap where an agent definition's bypassPermissions mode ignored the org-level bypass-permissions disable policy; the release also adds owner-wildcard entries ("owner/*") to strictKnownMarketplaces and blockedMarketplaces managed settings, a warning when workflow / forked-skill / slash-command / resumed-background-agent subagent model is restricted, a /teleport hint in cloud sessions to continue locally, and changes /review to be an alias of /code-review (with /code-review ultra for a deep cloud review); per Claude Code changelog

Thu Aug 6 2026 · Claude Code v2.1.223 · Four permission-bypass patches: (1) Bash command hides parts from permission checks (2) tabs / invisible Unicode hide command from approval dialog (3) workflow dynamic import() sandbox escape (4) agent bypassPermissions ignored org disable policy · New: owner/* wildcards for strictKnownMarketplaces + blockedMarketplaces · Restricted-subagent-model warning · /teleport hint in cloud sessions · /review = /code-review alias, /code-review ultra for deep cloud review

Two reads. (1) The four patches are the four different shapes a permission bypass can take in a coding-agent harness with a live permission model: command-string obfuscation (Bash), rendering-layer obfuscation (tabs and invisible Unicode against the approval dialog), runtime escape (dynamic import() out of the workflow sandbox), and policy-layer bypass (agent-level bypassPermissions overriding org policy). Landing all four in one cut is the shape a security review by real customers takes when it reports upward across a category, not one finding at a time. (2) The owner/* wildcards for strictKnownMarketplaces and blockedMarketplaces are the “we-have-to-govern-marketplace-orgs-not-individual-repos” primitive: enterprise policy now writes “allow all plugins under owner X, block all plugins under owner Y”, which is the shape a plugin marketplace takes when the buyer is an enterprise security team, not an individual developer. Read alongside v2.1.224's archive plugin source with optional SHA-256 pinning (item 01), the Aug 6-7 Claude Code shape is lock the plugin supply chain and lock the permission surface, in the same 24 hours.

03

OpenAI on Fri Aug 7 cuts openai/codex rust-v0.147.0 stable — the release ships portable Agent Plugins with local, personal, workspace and remote plugin-catalog search; an --approve-for-me flag that enables automatically reviewed approvals; persistent, manually-ordered conversation sections that let teams browse long transcripts incrementally; importing of Cursor-managed skills that synchronises changes to imported Claude and Cursor conversations without creating duplicates; opt-in support for the MCP 2026-07-28 protocol (paginated discovery, multi-round requests, non-blocking server startup); cached web search and remote conversation compaction for Bedrock; secret and complete-bearer-token redaction across displayed commands and replayed conversation history; and removes the deprecated codex exec --full-auto flag in favour of --sandbox workspace-write and drops the Linux bundle archives in favour of standard codex-package-<target> release archives; per the openai/codex GitHub releases page

Fri Aug 7 2026 · openai/codex rust-v0.147.0 stable · New: portable Agent Plugins (local + personal + workspace + remote catalog search) · --approve-for-me for auto-reviewed approvals · Persistent manually-ordered conversation sections · Cursor-managed skill imports without duplicates · Opt-in MCP 2026-07-28: paginated discovery + multi-round requests + non-blocking server startup · Cached web search + Bedrock remote conversation compaction · Secret + bearer-token redaction in displayed commands + replay · Removed: codex exec --full-auto (use --sandbox workspace-write) · Removed: Linux bundle archives (use codex-package-<target>)

Two reads. (1) Agent Plugins with a four-tier catalog (local + personal + workspace + remote) is the shape a coding-agent harness takes when it accepts that the plugin surface is now four different governance surfaces: the developer's own machine, their personal plugin bundle, their team workspace, and the shared remote catalog. That is the same four surfaces that Claude Code v2.1.223's owner-wildcard marketplace policy is solving on the Anthropic side (item 02). (2) Opt-in for the MCP 2026-07-28 protocol is the concrete first-mover adoption of the stateless MCP spec that landed Jul 28 (prior edition): Codex now speaks paginated discovery, multi-round requests, and non-blocking server startup to any MCP server that supports the 2026-07-28 wire format. The --approve-for-me flag is the “we-know-most-approvals-are-noise” primitive: Codex reviews the approval itself against the sandbox and workspace policy and gates only the ones that need a human. Read alongside Claude Code v2.1.224's cross-machine SendMessage and self-hosted runners (item 01), the Aug 7 coding-agent-harness shape is the two reference harnesses ship the same fleet + permission primitives on the same day, against the same MCP wire.

02

Agentic CX takes ~$150M of Series capital in three days — Klaviyo installs Elias Torres as CPO on Composer + Customer Agent, Omilia's voice platform crosses $60M ARR, Naïve's autonomous-company infrastructure lands its Series A

04

Klaviyo (NYSE: KVYO) on Wed Aug 5 agrees to acquire the team and technology of Agency — an AI-native customer-success company led by co-founder and CEO Elias Torres (Drift co-founder, ex-HubSpot) — and to install Torres as Chief Product Officer post-close, reporting to co-CEO Andrew Bialecki; Torres will lead Klaviyo's agent product lines: Composer, which builds marketing campaigns, and Customer Agent, which handles post-sale support like returns and order tracking; the Agency team of 25 people joins Klaviyo, the Agency standalone service winds down Aug 31, 2026, transaction terms were not disclosed, close expected in Q3 2026; Agency had raised $32M from Sequoia, Menlo Ventures and Felicis before the acquisition; per Klaviyo newsroom, TechCrunch, BusinessWire and CMSWire

Wed Aug 5 2026 · Klaviyo acquires Agency team + technology · Team size: 25 people · Elias Torres becomes Klaviyo CPO post-close (reports to co-CEO Andrew Bialecki) · Product lines Torres leads: Composer (marketing-campaign builder) + Customer Agent (post-sale returns + order tracking) · Torres pedigree: Drift co-founder, ex-HubSpot · Deal terms not disclosed · Close expected Q3 2026 · Agency standalone service winds down Aug 31, 2026 · Agency prior raises: $32M from Sequoia + Menlo + Felicis

Two reads. (1) The operative read is Klaviyo is buying the CPO seat, not the technology. Composer and Customer Agent already ship inside Klaviyo; Agency's 25-person AI-native team is the accelerant; and Torres — a Drift co-founder who previously built the go-to-market for a category-defining conversational product — is the profile a publicly-traded B2C CRM reaches for when it wants to convert “agents that build marketing campaigns and answer post-sale questions” into a defensible product line. That is the shape a market-leader CRM takes when the agent surface is now the primary product surface. (2) The Sequoia / Menlo / Felicis exit into a public-market strategic is the shape a venture-backed AI-native GTM co takes when the market has already picked its incumbent: the technology sells for team-and-CPO-seat, not independent trajectory. Read alongside Omilia's $67M Series B on $60M ARR (item 05) and Naïve's $28.5M Series A on 30K devs (item 06), the Aug 5-6 agentic-CX capital shape is strategic acquires the GTM CPO, independents raise growth.

05

Omilia on Thu Aug 6 secures a $67M Series B led by Expedition Growth Capital — the two-decade-old, Athens- and New York-based enterprise agentic CX company runs a voice-first, self-learning platform for large enterprises in banking, insurance and healthcare, with named customers including Capital One, Discover and Taco Bell; live annual recurring revenue is up 10× since Series A to over $60M with no interim equity, and proceeds will accelerate growth across North America and global markets including a new US office in H2 2026 and scaling of the global commercial organisation; prior round was $20M from Grafton Capital in 2020; per Yahoo Finance, TechCrunch, TheSaaSNews, CMSWire, CityBiz, Finsmes, EU-Startups and Ventureburn

Thu Aug 6 2026 · Omilia $67M Series B · Lead: Expedition Growth Capital · HQ: Athens + New York · Product: agentic self-learning voice-first CX platform for large enterprises · Named customers: Capital One + Discover + Taco Bell · Live ARR: up 10× since Series A to $60M+ · No intervening equity between Series A and B · Prior round: $20M from Grafton Capital (2020) · Use of funds: US office H2 2026 + global commercial scaling

Two reads. (1) The 10× ARR without an intervening equity round is the profile a growth-equity buyer reaches for: Expedition Growth Capital is writing at $60M+ ARR with Capital One, Discover and Taco Bell already live, which is the shape an enterprise-agentic-CX buyer takes when the platform is already the reference and the round funds distribution, not product. (2) Voice-first, self-learning is the differentiator: most agentic-CX competitors are chat-first with voice bolted on; Omilia is voice-native. Read alongside xAI's Grok Voice Think Fast 2.0 GA in July (prior editions), Sesame's $250M Series C (prior editions), and Rime's $24M for enterprise speech-to-speech (prior editions), the Aug 2026 shape of the voice-agent segment is the platform layer (Omilia) is raising off enterprise ARR while the model layer (Grok Voice, Sesame, Rime) is raising off benchmark leadership.

06

Naïve on Wed Aug 6 raises a $28.5M Series A led by Nexus Venture Partners — participants include Y Combinator, Zetta, Liquid 2, and angels Gokul Rajaram (Marqeta / DoorDash), Tim Zheng (Apollo), JD Sherman (ex-HubSpot COO), Gert Lanckriet (Amazon), Robert Chatwani (Docusign) and Zachary Sims (Codecademy); founded in 2026 by Berkeley dropouts Sean Dorje and Dennis Zax, the Palo Alto-based AI lab sells a unified API and serverless operating stack for autonomous companies, automating legal incorporation, financial management, cloud computing and multi-agent orchestration; the company signed up more than 30,000 developer customers within months of launch, grew annual run-rate revenue tenfold to the low-double-digit millions over the six months ending August 2026, and lists its customer base as AI automation agencies, faceless TikTok / YouTube channels and one rental-car agency; per TheSaaSNews, Dealroom, WebWire, SiliconAngle, TechCrunch, Finsmes, Pulse2 and CityBiz

Wed Aug 6 2026 · Naïve $28.5M Series A · Lead: Nexus Venture Partners · Participants: Y Combinator + Zetta + Liquid 2 · Angels: Gokul Rajaram + Tim Zheng (Apollo) + JD Sherman (ex-HubSpot) + Gert Lanckriet (Amazon) + Robert Chatwani (Docusign) + Zachary Sims (Codecademy) · HQ: Palo Alto · Founders: Sean Dorje + Dennis Zax (Berkeley dropouts) · Product: unified API + serverless operating stack for autonomous companies · Coverage: legal incorporation + financial management + cloud + multi-agent orchestration · 30,000+ dev signups within months · ARR up 10× to low-double-digit millions in six months · Customer base: AI automation agencies + faceless TikTok/YouTube channels + a rental-car agency

Two reads. (1) The customer roster is the story. AI automation agencies, faceless TikTok / YouTube channels, and one rental-car agency are the buyer profile of a service that automates the grunt work of setting up and running a company: one-founder or zero-founder shops that need the legal, financial and cloud spine of a real company but cannot afford, and do not want, an accountant and a lawyer and a devops engineer. That is the shape a bottom-up market takes when the incremental cost of “spin up a real corporate entity that runs itself” falls below the cost of the person who would have done it. (2) The 6-month 10× ARR and 30K+ dev signups in months are the traction signals a Nexus + Y Combinator + Zetta round is writing off; the angel bench (Rajaram, Zheng, Sherman, Chatwani, Sims) is the operator-adjacency signal. Read alongside Omilia's $67M enterprise voice round (item 05) and Klaviyo's Agency-and-Torres deal (item 04), the Aug 5-6 agentic-CX capital tape spans the top of the market (Klaviyo strategic), the enterprise segment (Omilia) and the long-tail one-founder segment (Naïve) — each with a different capital structure.

03

The agent-plumbing GitHub cohort crosses the 200-star line in the first week of August — skills, human review, and native macOS orchestration around the harnesses

07

petergyang/human-review ships as a visual review tool for AI-agent output — a lightweight JavaScript app that opens the HTML or Markdown a Claude Code / Codex / Cursor session just produced, lets a reviewer leave Google-Doc-style comments inline, and sends the collected feedback back into whichever AI harness the reviewer is running; the repo (created Jul 27, MIT-licensed) reached 574 stars by Aug 7 with a landing page at creatoreconomy.so and a full README of harness-integration recipes; topics list is deliberately terse: ai-agents, claude-code, cli, codex, feedback, html, human-in-the-loop, markdown, review

Repo created Jul 27, 2026 · petergyang/human-review · Language: JavaScript · License: MIT · Stars: 574 (as of Aug 7 push) · Function: visual HTML + Markdown review app with inline Google-Doc-style comments, pipes feedback into the AI harness · Landing page: creatoreconomy.so/p/use-my-human-review-skill-to-edit-html-markdown-visually · Topics: ai-agents + claude-code + cli + codex + feedback + html + human-in-the-loop + markdown + review

Two reads. (1) Human-review is the missing HITL primitive between raw agent output and PR-review UI. A coding agent produces HTML or a Markdown report; the reviewer wants to comment on line 42 without opening a PR. Human-review is the shape that gap takes when someone fills it as a single-purpose tool, and 574 stars in ten days is the shape a real gap in the workflow takes when the tool that fills it lands well. (2) The topic tags spell out the intended surfaceai-agents, claude-code, codex, human-in-the-loop — and the harness-agnostic README is the “works inside your favourite AI harness” commitment. Read alongside Claude Code v2.1.224's cross-machine SendMessage (item 01) and Codex v0.147.0's Agent Plugins (item 03), the Aug 2026 shape is the tooling around the harnesses is filling in faster than the harnesses can ship the primitives themselves.

08

0xwilliamortiz/claude-red ships on Wed Aug 5 as a curated library of offensive-security skills for the Claude skills system — each skill is a structured SKILL.md file that primes Claude with expert-level methodology for a specific attack surface (SQLi through shellcode; EDR evasion through exploit development); the MIT-licensed JavaScript repo reached 559 stars in three days (as of Aug 7) with 69 forks and topics claude, claude-code, claude-code-plugin, claude-skills, codex, codex-skill, redteam, redteaming, skill, skills — usable in Claude Code plugin marketplaces and importable by Codex via v0.147.0's Cursor / Claude skill import (item 03); per the GitHub repo

Repo created Aug 5, 2026 · 0xwilliamortiz/claude-red · Language: JavaScript · License: MIT · Stars: 559 (Aug 7) · Forks: 69 · Function: curated SKILL.md library of offensive-security skills for Claude · Coverage: SQLi + shellcode + EDR evasion + exploit development · Topics: claude + claude-code + claude-code-plugin + claude-skills + codex + codex-skill + redteam + redteaming + skill + skills

Two reads. (1) The skills-as-a-library shape has already crossed the offensive-security threshold: claude-red is the second high-star offensive-skills pack of the summer after the offensive-claude pack of prior editions, and the 559-stars-in-three-days trajectory is the shape a skill library takes when the pen-testing community is tooling on top of Claude Code the same way it once tooled on top of Metasploit. That is a real product-market fit signal for the Claude skills primitive, distinct from the general marketplace. (2) The topics tag is the harness-agnosticism signal: claude + claude-code + claude-skills + codex + codex-skill + skill + skills means the same SKILL.md format now travels across two harnesses, which is exactly the surface Codex v0.147.0's Cursor skill import addresses on the other side (item 03). Read alongside Claude Code v2.1.223's owner/* marketplace wildcards (item 02) and v2.1.224's archive plugin source with SHA-256 pinning (item 01), the Aug 2026 skills-supply-chain shape is rich unregulated library, tightening org-level policy, portable across harnesses.

09

cristicretu/diri ships on Tue Aug 4 as a native macOS orchestrator for coding agents — a Rust / Swift / GPUI app that runs Claude Code, Codex, Cursor, Gemini and shells in parallel across git worktrees and remote hosts; the Apache-2.0-licensed repo reached 222 stars by Aug 7 with 13 forks and topics claude-code, coding-agents, gpui, macos, mcp, rust, swift, terminal; a homebrew tap at cristicretu/homebrew-diri ships the install path; the pitch is a native macOS answer to the Electron operator-console cohort (BossConsole / Stably Orca-adjacent), keyed to git-worktree parallelism and native platform performance; per the GitHub repo

Repo created Aug 4, 2026 · cristicretu/diri · Language: Rust (with Swift + GPUI) · License: Apache-2.0 · Stars: 222 (Aug 7) · Forks: 13 · Function: native macOS orchestrator for coding agents · Runs in parallel: Claude Code + Codex + Cursor + Gemini + shells · Concurrency: git worktrees + remote hosts · Topics: claude-code + coding-agents + gpui + macos + mcp + rust + swift + terminal · Companion: cristicretu/homebrew-diri (install tap)

Two reads. (1) Diri is the native-macOS shape of the operator-console category: run multiple coding agents in parallel, each on its own git worktree, each optionally on a remote host. The Electron cohort of the summer (BossConsole, Stably Orca, prior editions) hit the same primitive from a cross-platform direction; diri is the platform-native answer, keyed to Rust + Swift + GPUI. That is the shape a coding-agent-first developer tool takes when native performance and macOS integration are the differentiator. (2) Native git-worktree parallelism is the practical unlock: an operator can fan out four coding sessions, each on a different branch, each editing the same repo without conflicts, and merge back through the diri UI. Read alongside Claude Code v2.1.224's cross-machine SendMessage (item 01) and Codex v0.147.0's Agent Plugins (item 03), the Aug 2026 operator-console shape is the harnesses ship fleet + skills primitives, and the operator-console layer ships the multi-harness parallel UI on top.

04

The Chinese open-weight side ships its own plumbing — Alibaba AMAP-ML's long-horizon computer-use harness and Anionex's vision toolkit for text-only DeepSeek / GLM / Kimi

10

AMAP-ML/LongHorizon-Harness ships on Tue Aug 4 from Alibaba's AMAP mapping-and-navigation ML org — a Python long-horizon computer-use harness that runs AI agents across desktop apps and the CLI for extended periods while preserving task state and making reliable progress on complex workflows, featuring fresh-context execution, durable verified state, independent auditing, recoverable progress, and native Claude Code / Codex / OpenClaw integration; the MIT-licensed repo reached 387 stars and 40 forks by Aug 7, with a landing page at lh-harness.pages.dev and topics agent, claude, claude-code, claude-plugin, cli, codex, codex-desktop, codex-plugin, cua, gui, harness, long-horizon, long-horizon-agents, longhorizon-harness, loop, loop-engineering; per the GitHub repo

Repo created Aug 4, 2026 · AMAP-ML/LongHorizon-Harness · Owner: Alibaba AMAP-ML (mapping-and-navigation ML org) · Language: Python · License: MIT · Stars: 387 (Aug 7) · Forks: 40 · Function: long-horizon computer-use harness (desktop apps + CLI, extended run) · Features: fresh-context execution + durable verified state + independent auditing + recoverable progress · Native integrations: Claude Code + Codex + OpenClaw · Landing page: lh-harness.pages.dev · Topics include: agent + claude-plugin + codex-desktop + cua + gui + long-horizon + loop-engineering

Two reads. (1) Alibaba AMAP shipping a public harness is the story: AMAP is China's reference mapping-and-navigation product, its ML org is production-scale, and a MIT-licensed long-horizon harness from that org is the shape a hyperscale Chinese team takes when it wants Western coding-agent harnesses to run inside its research workflows. Fresh-context execution + durable verified state + independent auditing + recoverable progress is the four-part “keep the agent honest across many hours” primitive, which is the shape a long-horizon-harness workload needs when the base agent context window is the ceiling. (2) The OpenClaw integration is the notable third-party: OpenClaw is the open-source coding-agent runtime the Chinese open-weight ecosystem has adopted as its harness equivalent, and native support alongside Claude Code and Codex is the shape a plumbing library takes when it wants to be the neutral harness-independent layer. Read alongside Anionex's agent-vision-toolkit (item 11), the Aug 2026 Chinese-open-weight tooling shape is ship the plumbing on top of Claude Code, Codex and OpenClaw at the same time.

11

Anionex/agent-vision-toolkit ships on Fri Aug 1 as a vision toolkit and skill for text-only LLMs — giving DeepSeek, GLM and Kimi image Q&A, long-screenshot OCR, frontend UI restoration and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi and OpenCode; the MIT-licensed bilingual (Chinese + English) Python repo reached 348 stars and 15 forks by Aug 7 with a landing page at agent-vision.anionex.me and topics agent, agent-skills, claude-code, cli, codex, computer-use, deepseek, glm, harness-engineering, kimi, llm, multimodal, opencode, text-only-llm, vision, vision-language-model; per the GitHub repo

Repo created Aug 1, 2026 · Anionex/agent-vision-toolkit · Language: Python · License: MIT · Stars: 348 (Aug 7) · Forks: 15 · Function: vision toolkit + skill that gives text-only LLMs image capabilities · Target text-only models: DeepSeek + GLM + Kimi · Capabilities: image Q&A + long-screenshot OCR + frontend UI restoration + GUI automation · Harness integrations: Codex + Claude Code + Pi + Oh My Pi + OpenCode · Landing page: agent-vision.anionex.me · Topics include: agent-skills + computer-use + harness-engineering + text-only-llm + vision-language-model

Two reads. (1) Agent-vision-toolkit is the shape the “Chinese open-weight text model + Western vision + Western harness” stack takes when someone stitches it together as a single skill. DeepSeek, GLM and Kimi are the three reference Chinese open-weight text models; none of them ships a first-party vision head at frontier parity; the toolkit routes image inputs through an external vision-language model and returns text the DeepSeek / GLM / Kimi caller can act on. That is the shape a multimodal-gap workaround takes when the developer chooses the text model for cost or licensing and needs vision on top. (2) The integration matrixCodex, Claude Code, Pi, Oh My Pi, OpenCode — is the “every harness the buyer might actually run” commitment. Read alongside AMAP-ML's LongHorizon-Harness (item 10), the Aug 2026 shape of the Chinese-open-weight-plus-Western-harness pattern is the plumbing library is now the seam: Chinese labs ship text weights, Chinese and Western developers ship the skills that make those weights usable inside Claude Code and Codex.

Compiled 2026-08-08 from the Claude Code changelog (v2.1.223 · v2.1.224) and the openai/codex GitHub releases page (rust-v0.147.0) on the coding-agent-harness cadence; Klaviyo Newsroom, BusinessWire, CMSWire, StockTitan on the Klaviyo / Agency acquisition and Elias Torres to CPO; Yahoo Finance, CMSWire, TheSaaSNews, Finsmes and EU-Startups on the Omilia $67M Series B; TheSaaSNews, Dealroom, WebWire, Pulse 2.0, Finsmes and CityBiz on the Naïve $28.5M Series A; and GitHub (petergyang/human-review, 0xwilliamortiz/claude-red, cristicretu/diri, AMAP-ML/LongHorizon-Harness, Anionex/agent-vision-toolkit) on the agent-plumbing GitHub cohort. Window of Aug 1 – Aug 8, 2026 UTC.