Steven Gonsalvez

Software Engineer


CEREBRO


machine-read, human-curated

Researched and written with 248,897 tokens (in 38 · out 19,247 · cache-read 78,782 · cache-create 150,830) across 5 Claude calls

Coding agents

  • v2.1.290: Claude Code adds serverToolUses to the mod turn.step hook result (API-run tools like the advisor) and agentId on plugin tool.check, so hooks can tell subagent permission checks from main-session ones.
  • mattpocock/skills v1.3: Adds /pr, /implement-spec (spec to subagent-built tickets) and /retro (reviews recent agent sessions, suggests repo fixes), and renames CONTEXT.md to GLOSSARY.md.
  • T3 Code in-app visualization: Agents can now build dynamic UI inside the thread, with CSS style tokens matched to your theme (follow-up).
  • T3 Code nightly orchestrator: User claims its orchestrator beats lab-provided tooling (subagent monitoring, roundtables, 400k+ users); anecdotal, but a signal on orchestration UX.
  • trq212's HTML plan skill: A Claude Code skill that produces HTML plans with code snippets, open questions and mockups, with linting to cut common failure cases; in feedback phase.
  • Anthropic made its apps 3x faster with Claude: Theo breaks down the strategies Anthropic's own agents used for the speedup, a useful template for agent-driven perf work.
  • simonw/claude-system-prompts: Turns Anthropic's published claude.ai/mobile system prompts into files with git history, one commit per dated revision, so git diff shows prompt changes.
  • llm-anthropic 0.30: Adds Claude Sonnet 5.5 and llm anthropic refresh, which pulls the model list from Anthropic's API so new models need no release.
  • Quoting Felix Rieseberg: The "old" Cowork ran inference in the cloud and tool calls in an Anthropic-shipped local VM, which was added for capability, safety and security reasons.

Cost, tokens and routing

Review, evals and safety

CLI/TUI and agent control rooms

  • tuios: Go terminal multiplexer with BSP tiling, persistent daemon sessions across machines, and a single inbox where coding agents report their state.
  • devpit: Native control room for Claude Code agents.
  • Live map of files a Claude Code edit touches: Reddit tool visualizing the blast radius of each Claude Code edit.
  • crosswalk: A shared space for people and their agents, starting with inbox.

Vibe-coding

Signals

  1. simonw/claude-system-prompts: Tracking changes to Claude's published system prompts
  2. llm-anthropic 0.30
  3. My Claude Max weekly limits fix: Opus + Sonnet sub-agents and a review loop
  4. OpenAI recently published data on its researchers’ coding-agent usage. We took their weekly figures from January to mid-August 2026 and fit…
  5. I tested 20+ ways to make a cheap coding model act like an expensive one. Here's what worked and what didn't
  6. v2.1.290
  7. mattpocock/skills v1.3 is out! - /pr (new) writes easy-to-read PR bodies, showing hard evidence that the changes work and assessing merge r…
  8. Vibe-coded a local events site that pulls from 101 calendars every 6 hours. PHP + SQLite on GoDaddy shared hosting, no framework.
  9. We just shipped in-app visualization capabilities in T3 Code. This enables agents to build cool dynamic experiences within the thread. Huge…
  10. Your agents write PRs faster than anyone can review them. What does your review setup look like?
  11. ReviewBench: An open benchmark for AI code review
  12. devpit
  13. I built a live token optimizer for AI coding sessions
  14. Opus 5.5 agents discover two room-temperature magnetic semiconductor candidates
  15. Gaurav-Gosain/tuios: A terminal window manager that knows what your agents are doing. Tiling panes, workspaces, sessions that survive resta…
  16. FastRouter.ai
  17. Anthropic used Claude to make their apps 3x faster. The strategies they came up with to do it are fascinating. I broke it down so you can u…
  18. RT @MahyadGhassemi: t3 code (nightly) is so much better than what the labs provide that it's not even funny anymore orchestrator genuinely…
  19. I built a live map for Claude Code that shows every file an edit touches
  20. Why 38% of AI Agent container escapes didn't need kernel 0-days: Analysis of 109 empirical incidents (Open Dataset + Defense Harness)
  21. crosswalk
  22. @davis7 Your agent is given CSS style tokens that work with your theme as well :) https://t.co/gNbDY57lkC
  23. Quoting Felix Rieseberg
  24. b-nnett/codex-subscription-router
  25. RT @trq212: I've been working on a skill that makes better HTML plans in Claude Code. It uses simple language, shows code snippets, surface…

Sources & citations

x2157
rss1407
reddit757
github403
hackernews301
show_hn400
github_search200
swipe200
yc-rfs130
gmail90
yc-launch20
seed_urls10
crackscan00
ossinsight00
reddit_users00
total60525