CEREBRO
machine-read, human-curated
Researched and written with 248,897 tokens (in 38 · out 19,247 · cache-read 78,782 · cache-create 150,830) across 5 Claude calls
Coding agents
- v2.1.290: Claude Code adds
serverToolUsesto the modturn.stephook result (API-run tools like the advisor) andagentIdon plugintool.check, so hooks can tell subagent permission checks from main-session ones. - mattpocock/skills v1.3: Adds
/pr,/implement-spec(spec to subagent-built tickets) and/retro(reviews recent agent sessions, suggests repo fixes), and renames CONTEXT.md to GLOSSARY.md. - T3 Code in-app visualization: Agents can now build dynamic UI inside the thread, with CSS style tokens matched to your theme (follow-up).
- T3 Code nightly orchestrator: User claims its orchestrator beats lab-provided tooling (subagent monitoring, roundtables, 400k+ users); anecdotal, but a signal on orchestration UX.
- trq212's HTML plan skill: A Claude Code skill that produces HTML plans with code snippets, open questions and mockups, with linting to cut common failure cases; in feedback phase.
- Anthropic made its apps 3x faster with Claude: Theo breaks down the strategies Anthropic's own agents used for the speedup, a useful template for agent-driven perf work.
- simonw/claude-system-prompts: Turns Anthropic's published claude.ai/mobile system prompts into files with git history, one commit per dated revision, so
git diffshows prompt changes. - llm-anthropic 0.30: Adds Claude Sonnet 5.5 and
llm anthropic refresh, which pulls the model list from Anthropic's API so new models need no release. - Quoting Felix Rieseberg: The "old" Cowork ran inference in the cloud and tool calls in an Anthropic-shipped local VM, which was added for capability, safety and security reasons.
Cost, tokens and routing
- OpenAI researchers' coding-agent spend: Epoch AI fits OpenAI's published weekly data (Jan to mid-Aug 2026) and finds spending doubling about once a month.
- Claude Max weekly limits fix: Reddit workflow using Opus and Sonnet sub-agents plus a review loop to stay under Max weekly caps.
- 20+ ways to make a cheap coding model act expensive: Reddit write-up of what did and didn't close the gap between cheap and expensive coding models.
- Live token optimizer for coding sessions: Community tool that trims token use during live AI coding sessions.
- codex-subscription-router: Patched macOS ChatGPT app that balances new chats across several subscriptions while pinning each thread to one account, so follow-up turns keep account-level caching.
- FastRouter.ai: Routes requests to the right LLM by cost, latency and quality.
Review, evals and safety
- ReviewBench: GitHub's open benchmark for AI code review, using real PRs, multi-source ground truth and production-aligned metrics.
- Agents write PRs faster than anyone can review: Reddit thread on review setups for agent-generated PR volume.
- 38% of agent container escapes needed no kernel 0-day: Analysis of 109 real incidents with an open dataset and defense harness; misconfiguration, not exploits, is the main route out of agent sandboxes.
- Opus 5.5 agents find room-temperature magnetic semiconductor candidates: Agent-driven materials discovery, though HN commenters stress that DFT candidates are not breakthroughs without experimental verification.
CLI/TUI and agent control rooms
- tuios: Go terminal multiplexer with BSP tiling, persistent daemon sessions across machines, and a single inbox where coding agents report their state.
- devpit: Native control room for Claude Code agents.
- Live map of files a Claude Code edit touches: Reddit tool visualizing the blast radius of each Claude Code edit.
- crosswalk: A shared space for people and their agents, starting with inbox.
Vibe-coding
- Local events site from 101 calendars: Vibe-coded PHP + SQLite site on GoDaddy shared hosting, no framework, refreshing every 6 hours.
Signals
- simonw/claude-system-prompts: Tracking changes to Claude's published system prompts
- llm-anthropic 0.30
- My Claude Max weekly limits fix: Opus + Sonnet sub-agents and a review loop
- OpenAI recently published data on its researchers’ coding-agent usage. We took their weekly figures from January to mid-August 2026 and fit…
- I tested 20+ ways to make a cheap coding model act like an expensive one. Here's what worked and what didn't
- v2.1.290
- mattpocock/skills v1.3 is out! - /pr (new) writes easy-to-read PR bodies, showing hard evidence that the changes work and assessing merge r…
- Vibe-coded a local events site that pulls from 101 calendars every 6 hours. PHP + SQLite on GoDaddy shared hosting, no framework.
- We just shipped in-app visualization capabilities in T3 Code. This enables agents to build cool dynamic experiences within the thread. Huge…
- Your agents write PRs faster than anyone can review them. What does your review setup look like?
- ReviewBench: An open benchmark for AI code review
- devpit
- I built a live token optimizer for AI coding sessions
- Opus 5.5 agents discover two room-temperature magnetic semiconductor candidates
- Gaurav-Gosain/tuios: A terminal window manager that knows what your agents are doing. Tiling panes, workspaces, sessions that survive resta…
- FastRouter.ai
- Anthropic used Claude to make their apps 3x faster. The strategies they came up with to do it are fascinating. I broke it down so you can u…
- RT @MahyadGhassemi: t3 code (nightly) is so much better than what the labs provide that it's not even funny anymore orchestrator genuinely…
- I built a live map for Claude Code that shows every file an edit touches
- Why 38% of AI Agent container escapes didn't need kernel 0-days: Analysis of 109 empirical incidents (Open Dataset + Defense Harness)
- crosswalk
- @davis7 Your agent is given CSS style tokens that work with your theme as well :) https://t.co/gNbDY57lkC
- Quoting Felix Rieseberg
- b-nnett/codex-subscription-router
- RT @trq212: I've been working on a skill that makes better HTML plans in Claude Code. It uses simple language, shows code snippets, surface…
Sources & citations
| Source | Fetched | In briefing |
|---|---|---|
| x | 215 | 7 |
| rss | 140 | 7 |
| 75 | 7 | |
| github | 40 | 3 |
| hackernews | 30 | 1 |
| show_hn | 40 | 0 |
| github_search | 20 | 0 |
| swipe | 20 | 0 |
| yc-rfs | 13 | 0 |
| gmail | 9 | 0 |
| yc-launch | 2 | 0 |
| seed_urls | 1 | 0 |
| crackscan | 0 | 0 |
| ossinsight | 0 | 0 |
| reddit_users | 0 | 0 |
| total | 605 | 25 |