CEREBRO
machine-read, human-curated
Coding agent harnesses and orchestration
- T3 Code adds Pi, delegate_task and an MCP server: one PR brings Pi support, auto-resume when rate limits reset, an MCP to create, message and wait on threads, and agents that spawn child agents on any provider, plus an ACP registry, OpenCode 2 and the official Cursor SDK.
- T3 Code cross-harness context handoff: lets you move a session's context between Claude, Codex, Pi and custom harnesses and back, which cuts provider lock-in for multi-agent workflows.
- T3 Code "Hide threads while working": a beta option that hides running threads so they don't take up space in your head.
- Theo on running 18 threads: threads surface only when they need attention, an interface pattern for scaling parallel agents without clutter.
- mattpocock/sandcastle: a TypeScript
sandcastle.run()runs coding agents in Docker, Podman or Vercel sandboxes on their own branches and merges the commits back. - Matt Pocock on GitHub Actions software factories: issues become agent tasks, labels trigger Actions that open PRs, and labels applied by Actions form loops, giving a near-free sandbox for public repos.
- Pi 1.0, Pi Durable, and AIE NYC: Pi 1.0 and Pi Durable ship under Earendil, which Pi has joined, and Pi is now being treated as a serious OpenClaw-class harness.
- Theo on running 6 Claude and 3 Codex subs: a video on stacking many subscriptions, with a warning that it may get you banned.
Benchmarks, nerfs and model reliability
- Theo on prompt-outcome volatility: Claude Code updates, provider changes, cache-hit shifts and seeds can all change results within minutes, so public benchmarks are meaningless without controlled access.
- Theo disputes the Opus 4.6 "nerf" claim: says the April claim rested on only 6 of 30 benchmark tests, with 2 failures treated as proof of a nerf.
- BridgeMind defends NerfBench: the other side of the dispute, with a thread explaining how NerfBench works.
- Theo's misinformation video list: upcoming topics include model nerfs, whether token prices and TPS matter, the real impact of bad harnesses, and "compaction = deletion".
Tokens, speed and model guidance
- JuliusBrussee/caveman: a skill and proxy that cuts output tokens about 65% by forcing terse caveman-style responses, backed by an Adobe Research paper reporting 1.4 to 3x lower cost.
- What if AI worked at 1,000,000 tokens per second?: at that speed, deciding which draft to trust becomes the bottleneck, and commenters say latency, not throughput, is what limits real-time applications.
- A model guide for the GPT-6 family: OpenAI's guidance on choosing GPT-6 models, tuning reasoning effort, writing prompts and skills, coordinating tools and preparing workflows for production.
- Coachwhip: an open-source one-line install that runs the 80B Qwen3-Coder-Next fully offline on a 16 GB Apple silicon Mac.
Agent tools, MCP and skills
- Claude Code skill that filters AI review comments: checks CodeRabbit comments against the code, removing 34% of the noise while keeping 93% of real bugs.
- Pony: gives Claude control of an Android phone via MCP or your own Anthropic key.
- Claude as game master over MCP: Claude Code built an RPG world engine, and the same engine now exposes MCP so Claude runs the game.
- Agent-Reach: one CLI that lets an agent read and search Twitter, Reddit, YouTube, GitHub, Bilibili and XiaoHongShu with no API fees.
- WeftCut: an open-source video editor designed to be driven by an AI agent.
- JarvisCore: builds agents as zero-trust peers in a mesh network.
Agentic SaaS
- Cue by Manus: personal agents with persistent identity that handle tasks end to end.
- Famulor: an agent that answers calls and handles follow-ups on WhatsApp.
Signals
- mattpocock/sandcastle: Orchestrate sandboxed coding agents in TypeScript with sandcastle.run()
- JuliusBrussee/caveman: 🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talk…
- RT @aicodeking: I agree with Theo here. A Claude Code Update, A Provider Change, A cache hit change, A bad request, New seed anything can c…
- RT @matt_feroz: T3 Code has just changed agentic programming forever. Ever wanted to pass context from Claude-> Codex -> Pi -> you…
- 80B model. 16 GB Mac. Built Coachwhip so you can run Qwen3-Coder-Next fully offline on Apple silicon. Free, open source, one-line install.…
- Things included in this PR: - Pi support. - Auto-resume when limits reset - T3 Code MCP (create, launch, message, wait on, read, search and…
- Panniantong/Agent-Reach: Give your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili, Xiao…
- WeftCut
- Codync
- A lot of people have been asking me “how do I actually use 6 Claude subs and 3 Codex subs?” I got drunk as I tried to explain it. IMO this…
- @bridgemindai I may have made a mistake here. I am doing a deep dive as I prep my bench. Didn't realize he outright lied about the Opus 4.6…
- JarvisCore
- This felt so weird initially but I'm obsessed with it now. I have 18 threads working here and it doesn't feel claustrophobic anymore. Thing…
- New (beta) T3 Code feature: "Hide threads while working" I have been thinking about adding this to T3 Code for awhile now. Don't like "runn…
- @theo Theo is obsessed with BridgeMind. He hates seeing me win. NerfBench is not hard to understand. Here's how it works: https://t.co/bY1R…
- I have a lot of common AI misinformation I want to address in a few videos. Topics like: - model nerfs - token prices mattering - TPS matte…
- GitHub actions are a super easy way to get started with software factories 1. Essentially free sandboxes for public repo's 2. You already h…
- I built a Claude Code skill that checks AI code review comments against the code. On CodeRabbit's reviews it removed 34% of the noise and k…
- Claude Code built my RPG world engine. Now it can also run the game, as the game master, over MCP.
- Pony: give Claude hands on your Android phone (MCP or your own Anthropic key)
- What if AI worked at 1.000.000 tokens per seconds?
- (AINews) Pi 1.0, Pi Durable, and AIE NYC
- A model guide for the GPT-6 family
- Famulor
- Cue by Manus
Sources & citations
| Source | Fetched | In briefing |
|---|---|---|
| x | 221 | 11 |
| rss | 140 | 7 |
| github | 43 | 3 |
| 25 | 3 | |
| hackernews | 30 | 1 |
| show_hn | 40 | 0 |
| github_search | 20 | 0 |
| swipe | 20 | 0 |
| yc-rfs | 13 | 0 |
| gmail | 7 | 0 |
| yc-launch | 3 | 0 |
| seed_urls | 1 | 0 |
| crackscan | 0 | 0 |
| ossinsight | 0 | 0 |
| reddit_users | 0 | 0 |
| total | 563 | 25 |