Steven Gonsalvez

Software Engineer


CEREBRO


machine-read, human-curated

Coding agents & LLM mechanics

  • Claude Code — Anthropic's terminal coding agent repo, still the category's reference implementation.
  • andrej-karpathy-skills — single CLAUDE.md distilling Karpathy's observation that models run with wrong assumptions instead of asking, packaged as a drop-in guardrail.
  • The Evolution of the Agent Harness — argues agents "started working" around Christmas 2025 from model capability crossing a threshold plus harnesses finally maturing around it.
  • New MCP Roadmap — maintainers publish next-spec direction; community pushback says MCP's bespoke protocol was a mistake and stateless HTTP is the real fix.
  • A week of using Codex more than Claude — dev finds Codex faster and cleaner-coded, flags Opus 5.0 as a regression vs 4.8.
  • Zero — Vercel's new programming language built specifically for AI agents.
  • Munder Difflin — free multi-agent harness wrapping your existing CLI/subscription to run an "office of clones" locally; community split on practicality vs novelty.
  • Autolith — terminal coding agent with a live Common Lisp runtime for recursive inference over large corpora; thin Lisp training data limits real-world uptake.
  • nodeterm — tmux-backed canvas turning terminals and parallel Claude Code sessions into a draggable Trello-style board.
  • Claude's output is not optimised for humans — thread arguing Claude increasingly formats for downstream agent consumption, not human readers.
  • Why your local LLM feels dumber than it is — silent GGUF metadata loss breaks chat templates and tanks perceived quality, not quantization itself.
  • More than just code review — Simon Willison on verifying agent-written code without eyeballing every line, the real skill bottleneck.
  • openai/codex — OpenAI's terminal coding agent, direct Claude Code competitor.
  • Agents Never Sleep — launch for agents that keep running with the laptop lid closed.
  • NanoGPT Speedrun Frontier — 153 autonomous runs across 18 frontier models racing the nanoGPT speedrun; Fable 5 leads on validated closure, but inconsistent effort settings undercut cross-model comparison.
  • 10% worse, 100x cheaper, 10000x faster: Why Simulation is taking over — argues cheap-fast-good-enough simulation is winning, ties together GLM/Poolside/AI-for-Science threads.
  • Claude workflow for Android execution — community writeup wiring Claude into an Android execution loop.
  • Quoting Linus Torvalds — Torvalds credits AI as a "tireless helper" in a debug-session-from-hell, even after it insisted the bug was unsolvable.

CLI/TUI

  • llm 0.33 — Simon Willison's llm CLI upgrades to OpenAI Python 3.x, swaps httpx for httpx2.
  • simonw/llm — the CLI/library underlying it, run any LLM (Claude, GPT, Gemini, local models) from the terminal.

Vibe-coding & repos

Agentic SaaS

  • n8n — fair-code workflow automation with native AI agent building, 1500+ integrations, self-host or cloud.
  • sub2api — OSS relay unifying Claude/OpenAI/Gemini/Grok subscriptions behind one API for cost-sharing; project itself flags likely ToS violation with upstream providers.
  • Open Analytics — AI-native Google Analytics alternative.

Signals

  1. anthropics/claude-code: Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code fa…
  2. multica-ai/andrej-karpathy-skills: A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on…
  3. The Evolution of the Agent Harness
  4. New MCP Roadmap
  5. A week of using Codex more than Claude
  6. Zero
  7. Munder Difflin – Agent harness to run an office of your clones
  8. Autolith: A programming agent with a live runtime
  9. eneskirca/nodeterm: Node-based terminal manager for AI coding agents — tmux-backed terminals and parallel agent sessions as draggable nodes…
  10. PSA: Claude’s output is not optimised for humans. It’s for other agents.
  11. Why your local LLM feels dumber than it is
  12. llm 0.33
  13. More than just code review
  14. I implemented a diary that distributes post-quantum ciphertexts entirely in URLs
  15. simonw/llm: Access large language models from the command-line
  16. openai/codex: Lightweight coding agent that runs in your terminal
  17. n8n-io/n8n: Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or clou…
  18. Agents Never Sleep
  19. I built one app to replace Adobe Illustrator, Lightroom and most of After Effects. The Figma part is next.
  20. NanoGPT Speedrun Frontier
  21. (AINews) 10% worse, 100x cheaper, 10000x faster: Why Simulation is taking over
  22. I built a Claude workflow for Android execution
  23. Wei-Shaw/sub2api: Sub2API 一站式开源中转服务,让 Claude、Openai 、Gemini、Grok订阅统一接入,支持拼车共享,更高效分摊成本,原生工具无缝使用。
  24. Quoting Linus Torvalds
  25. Open Analytics

Sources & citations

rss1408
github447
hackernews606
reddit254
show_hn400
github_search200
swipe170
yc-rfs130
yc-launch50
gmail10
seed_urls10
crackscan00
reddit_users00
x00
total36625