
Indie Agent Builders — Week of May 30
The week's activity converges on a clear shift: agents are graduating from a model-quality competition to a harness engineering competition. Simon Willison's Anthropic PMF essay (backed by the $65B Series H at $47B ARR), Claude Opus 4.8's mid-conversation system messages, and the explosive growth of harness tooling (ECC at 199k★, PilotDeck's workspace OS, compound-engineering-plugin) all point the same direction. Swyx's Cognition D-round coverage ($1B at $26B valuation) and harness-as-differentiator thesis provide the commercial and architectural framing; the trending repos section documents seven concrete tools engineers can evaluate today.
Simon Willison
April 2026 as the product-market fit inflection point
"I'd characterize 229 (32.6%) as relating to enterprise sales and support — account executives, 'Go To Market', 'Forward Deployed Engineers' and the like." 1
Claude Opus 4.8: the right kind of incremental
role: "system" message immediately after a user turn, letting you append new instructions mid-session without invalidating the prompt cache or restating a full system prompt. For long agentic sessions where the task evolves, this removes a real architectural compromise.- Prompt cache minimum length drops from 4,096 to 1,024 tokens — smaller prompts now cache, which matters for multi-turn loops with short system prompts.
- Fast mode price cut: $10/$50 per million tokens (input/output), down from $30/$150 for Opus 4.6/4.7 fast mode. 3
low through max). The max level pelican-on-a-bike came in at 43 cents (25,000 input tokens + 17,167 output tokens) and produced the best result. He called Anthropic's self-description of the release as "a modest but tangible improvement" refreshing: "It's so refreshing to see an AI lab honestly describe a release as a minor incremental improvement over the previous model!" 3Datasette 1.0a30, 1.0a31, and datasette-agent 0.1a4
/ to get a keyboard-driven fuzzy search over databases and tables, with a jump_items_sql() plugin hook so any Datasette plugin can inject its own entries. datasette-agent 0.1a4 uses this immediately: it adds a "Start a new agent chat" entry to the Jump menu, turning agent access into a first-class navigation gesture rather than a separate route.
INSERT, UPDATE, DELETE via stored queries for users with appropriate permissions. Combined with the agent chat entry, this gives a datasette-agent session actual write capabilities through the same plugin permission model. The register_agent_tools architecture means write-query tools compose with any other plugin's tools in the same agent context.llm-anthropic 0.25.1 — adds Claude Opus 4.8 support, a -o fast 1 flag for fast mode, and changes the default max_tokens from a hardcoded 8,192 to each model's actual maximum. 2Security signals worth watching
AGENTS.md file telling AI agents how to behave around the codebase — and the most recent commit hardened the position from "does not (currently) accept agentic code" to "does not accept agentic code," dropping "currently" deliberately. The project also spun off a new SQLite Bug Forum to handle the volume of AI-generated reports. 6Swyx
Cognition's $1B D-round: the "largest remaining independent agent lab"

Harness engineering as the differentiator
model + harness + eval loop. Not just a stronger base model, but the fit between model, task harness, and feedback loop. 8AIE World's Fair: FDE track and the Turing award question
Kakuna: ongoing cadence
swyxio/skills repository (Swyx's Kakuna framework for Claude Code and other agent harnesses) added five commits between May 25–30: a new slackbot-builder skill with an L0-L5 maturity ladder, a public-qa-chatbot demo skill, a twitter-x-scraping skill, and a web-animation-perf update. 11 The L0-L5 maturity model for the Slackbot skill is worth examining as a template: it defines what "production-ready" means at each level rather than leaving it implicit, which is the specific gap Kakuna was designed to close.Trending repos
| Repo | Stars | +/day | What it does |
|---|---|---|---|
| affaan-m/ECC | 199,137 | +918 | Cross-harness optimization system: skills, instincts, memory, security for Claude Code, Codex, Cursor, Opencode, and 8 others |
| anthropics/skills | 143,984 | +471 | Anthropic's official agent skills repo — the de-facto skill format standard, referenced by ECC, PilotDeck, and harness |
| EveryInc/compound-engineering-plugin | 18,400 | +348 | 37 skills, 51 agents across Claude Code / Codex / Cursor; 80/20 rule: 80% planning, 20% execution |
| OpenBMB/PilotDeck | 2,200 | — | Agent OS from Tsinghua THUNLP/ModelBest/OpenBMB: per-project WorkSpace isolation, white-box memory, smart model routing |
| openclaw/openclaw | 376,000 | — | v2026.5.28-beta.4: agent runtime recovery, Claude Opus 4.8, Codex Supervisor plugin, GitHub Copilot agent runtime |
| AI45Lab/AgentDoG | 550 | — | v1.5: safety guardrail framework trained on ~1,000 samples; 75.2% accuracy on Risk Source vs. 33.6% for GPT-5.4 |
| revfactory/harness | 4,200 | +80 | v1.2.0: Claude Code plugin that generates 6-pattern agent team architectures from a domain description |
.claude/agents/ directory with agent definitions and skill files matching one of six patterns: Pipeline, Fan-out/Fan-in, Expert Pool, Producer-Reviewer, Supervisor, or Hierarchical Delegation. 19 It's a meta-factory rather than a runtime — scaffolding the team design before the work starts./ce-product-pulse command — time-windowed product usage reports saved to docs/pulse-reports/ — is a specific, replicable pattern for keeping engineering connected to user behavior.참고 출처
- 1I think Anthropic and OpenAI have found product-market fit
simonwillison.net
- 2Archive for Friday, 29th May 2026
simonwillison.net
- 3Claude Opus 4.8: "a modest but tangible improvement"
simonwillison.net
- 4Archive for Sunday, 24th May 2026
simonwillison.net
- 5Microsoft Copilot Cowork Exfiltrates Files
simonwillison.net
- 6Archive for Wednesday, 27th May 2026
simonwillison.net
- 7Archive for Tuesday, 26th May 2026
simonwillison.net
- 8not much happened today | AINews
news.smol.ai
- 9
- 10
- 11GitHub: swyxio/skills
github.com
- 12Geoffrey Huntley GitHub profile
github.com
- 13affaan-m/ECC
github.com
- 14GitHub Trending
github.com
- 15EveryInc/compound-engineering-plugin
github.com
- 16OpenBMB/PilotDeck
github.com
- 17Releases · openclaw/openclaw
github.com
- 18AI45Lab/AgentDoG
github.com
- 19revfactory/harness
github.com
- 2011|[AINews
- 2114|[AINews

AI Agent Builders Worth Following
Weekly aggregation of latest builds, posts, and shares from indie AI agent developers
이 콘텐츠는 채널이 자동으로 생성했습니다. 한 문장이면 Neodrop이 당신을 위해 계속 만들어 냅니다.
관련 콘텐츠
- 로그인하면 댓글을 작성할 수 있습니다.