Steven Gonsalvez

Software Engineer


CEREBRO


machine-read, human-curated

Coding agents

Claude Code mods

  • v2.1.287: ships Claude Mods (plugins can modify deeper behavior) plus a built-in "You should know" mod, a side agent that flags what you or Claude missed, enabled via /plugin enable cc-plugin-you-should-know@builtin.
  • You can now mod Claude Code: behavior, UI and features are customizable with a few lines of TypeScript or by asking Claude to write the mod, installed as plugins via /plugin.
  • Mods are absolutely insane: Boris Cherny pitches prompt-driven customization of Claude's look and workflow, with mods shared as plugins.

Other agent harnesses and tooling

  • earendil-works/pi: open toolkit with unified LLM API, agent runtime with tool calling, TUI and a self-extensible coding agent CLI; new-contributor issues and PRs are auto-closed by default.
  • DeepSeek Harness: DeepSeek's open-source agent harness; HN commenters say its one-shot benchmark misses long-running use, so published metrics mislead.
  • DSH Desktop: official desktop app for the DeepSeek harness.
  • OpenCompanion: one desktop app to start, watch and answer multiple AI coding CLIs.
  • Crontick: local scheduler for running prompts on a cron-style schedule.
  • Effect v4: zero-dependency TypeScript ecosystem pitched as the base for reliable software and AI agents.

Copilot

Runtimes, sandboxes and security

  • agent-substrate/substrate: secure-by-default agent runtime for millions of sandboxes, with sub-500ms resume, 500+ suspend/resume activations per second, and zero-trust kernel and network isolation.
  • Private AI Proxy: checks the AI service with hardware attestation and TLS key pinning, and forwards fail-closed, before your codebase context leaves your machine.
  • Polylane: agents that fix production issues unattended.
  • Semitexa: PHP framework designed so an AI agent can inspect it.

Models and local inference

Signals

  1. Every Claude Code sub-agent we ran was Opus 5.5. Then Sonnet 5.5 scored 40/40 for $0.02. We never compromise on quality, so Opus built ever…
  2. You can now mod Claude Code: - Change how it behaves - Customize the UI - Swap in your own features Write one with a few lines of TypeScrip…
  3. agent-substrate/substrate: Agent Substrate: the core system
  4. v2.1.287
  5. This week Claude Sonnet 5.5, GPT-6.1 Sol and Gemini 4 Argon all launched near the top of the Coding Agent Index leaderboard, but each has a…
  6. earendil-works/pi: AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI
  7. GitHub Copilot in VS Code, September 2026 releases
  8. Dynamic workflows in Copilot CLI and the Copilot app
  9. OpenCompanion
  10. Mods are absolutely insane. You can now customize Claude to work and look the way you want by just prompting it. Each person works differen…
  11. RT @EffectTS_: Effect v4 is here. One ecosystem. Zero dependencies. The next chapter of Effect and the foundation for building reliable sof…
  12. DeepSeek Harness
  13. DSH Desktop
  14. Stripe cut payment integration timelines from up to 6 months to 2–6 weeks using coding agents. The useful detail: their first orchestrator…
  15. GitHub Copilot can now interact with desktop apps with computer use
  16. Your coding agent can read your entire codebase. Private AI Proxy verifies the AI service before that context leaves your machine, using ha…
  17. We’re excited to release Muse Spark 1.3 with improved performance on agentic and coding tasks, and a focus on real-world usability. Key cap…
  18. Polylane
  19. Asked Fable 5.5 for the three-body problem. One-shot. It gave back ~1000 lines of Bend2: symplectic physics on the CPU, every pixel compute…
  20. 🤖 $NEAR JUST ADDED CLAUDE OPUS 5.5 TO ITS AI CLOUD Claude Opus 5.5 is now live on $NEAR AI Cloud with a 1M context window, 128K output and…
  21. Running 95.5 GiB Qwen3.8-Flash-Next at 41–52 tok/s on a 64GB Mac (1.76x faster than llama.cpp): Slipstream release, 130k context scaling, +…
  22. Semitexa
  23. The future is verification-engineering. Proofs, (e2e) tests, benchmarks, linters… Some tests will be deterministic, some agentic. This look…
  24. Crontick - Local prompt scheduler
  25. I gave Claude Code the CEO job for my free Chrome extension. Week 1: it made me rename it, built the next version, and still needs my "go"…

Sources & citations

x22711
rss1408
reddit503
github422
hackernews301
show_hn400
github_search200
swipe200
yc-rfs130
gmail80
yc-launch30
seed_urls10
crackscan00
ossinsight00
reddit_users00
total59425