🎙 ex-Host of Dev Interrupted, current host of Chain of Thought ⛓️, failed founder, & avid sports fan.
youtu.be/CQpAVPSYPxI
Full episode:
chainofthought.show/podcast/67-...
Full episode:
chainofthought.show/podcast/67-...
chainofthought.show/podcast/35-...
chainofthought.show/podcast/35-...
Their co-founder sat down with me: chainofthought.show/podcast/73-...
Their co-founder sat down with me: chainofthought.show/podcast/73-...
Full Chain of Thought episode: youtu.be/xMC7d_h5rSQ
Full Chain of Thought episode: youtu.be/xMC7d_h5rSQ
- OpenCode joins Claude Code, Codex, and OpenClaw as a 1st class host
- Hermes, Cursor, and Devin remain experimental for now (close!)
- Handoff previews now name their sources
- Local receipt history upgrades
github.com/conorbronsd...
Try it: github.com/conorbronsd...
- OpenCode joins Claude Code, Codex, and OpenClaw as a 1st class host
- Hermes, Cursor, and Devin remain experimental for now (close!)
- Handoff previews now name their sources
- Local receipt history upgrades
github.com/conorbronsd...
Try it: github.com/conorbronsd...
Kris Lovejoy explains why getting that data into usable shape is a deployment problem to solve early.
Watch the clip: youtube.com/shorts/xbvu...
Kris Lovejoy explains why getting that data into usable shape is a deployment problem to solve early.
Watch the clip: youtube.com/shorts/xbvu...
Try it:
github.com/conorbronsd...
Try it:
github.com/conorbronsd...
Trace the action. What denies it? Can the agent change that rule or use another route?
My takeaway from Redpanda’s Tyler Akidau: enforce critical rules outside the agent.
Trace the action. What denies it? Can the agent change that rule or use another route?
My takeaway from Redpanda’s Tyler Akidau: enforce critical rules outside the agent.
On Chain of Thought, Dan Klein explains correlated errors and the judgment it takes when AI turns you into the editor of its work.
chainofthought.show/podcast/54-...
On Chain of Thought, Dan Klein explains correlated errors and the judgment it takes when AI turns you into the editor of its work.
chainofthought.show/podcast/54-...
Try it:
github.com/conorbronsd...
Try it:
github.com/conorbronsd...
github.com/conorbronsd...
github.com/conorbronsd...
One model that reasons across text, images, audio and video. A hard part is alignment: connecting what it sees with what it reads.
Full breakdown:
chainofthought.show/ai-decoded/...
One model that reasons across text, images, audio and video. A hard part is alignment: connecting what it sees with what it reads.
Full breakdown:
chainofthought.show/ai-decoded/...
chainofthought.show/podcast/45-...
chainofthought.show/podcast/45-...
Prototype big, serve small. Use the frontier to find out what "good" looks like, then move a stable, high-volume task to a smaller model tuned for it.
Full breakdown:
chainofthought.show/ai-decoded/...
Prototype big, serve small. Use the frontier to find out what "good" looks like, then move a stable, high-volume task to a smaller model tuned for it.
Full breakdown:
chainofthought.show/ai-decoded/...
Limits on what an agent can take in, do and output. A rule in the prompt is a request, not a control. Enforce the must-nevers in code, with permissions and approval gates.
Full breakdown:
chainofthought.show/ai-decoded/...
Limits on what an agent can take in, do and output. A rule in the prompt is a request, not a control. Enforce the must-nevers in code, with permissions and approval gates.
Full breakdown:
chainofthought.show/ai-decoded/...
chainofthought.show/podcast/74-...
chainofthought.show/podcast/74-...
github.com/conorbronsd...
github.com/conorbronsd...
Usually in the gaps between agents. A handoff drops context, or one agent's wrong output becomes the next one's trusted input.
Treat the system as the unit: trace every step, validate every handoff.
Full breakdown:
chainofthought.show/ai-decoded/...
Usually in the gaps between agents. A handoff drops context, or one agent's wrong output becomes the next one's trusted input.
Treat the system as the unit: trace every step, validate every handoff.
Full breakdown:
chainofthought.show/ai-decoded/...
Test properties, not exact strings. Build a dataset of real inputs and what a good answer must do, score every output against it, and run it on every change like CI.
Full breakdown:
chainofthought.show/ai-decoded/...
Test properties, not exact strings. Build a dataset of real inputs and what a good answer must do, score every output against it, and run it on every change like CI.
Full breakdown:
chainofthought.show/ai-decoded/...
Trace it first so you can see cost per step. Then right-size the model for each step, trim and cache context, and fix the failures that cause expensive retries.
Full breakdown:
chainofthought.show/ai-decoded/...
Trace it first so you can see cost per step. Then right-size the model for each step, trim and cache context, and fix the failures that cause expensive retries.
Full breakdown:
chainofthought.show/ai-decoded/...
Faster code still needs a trustworthy test.
Faster code still needs a trustworthy test.
Does the agent find the update or reuse the stale value?
Does the agent find the update or reuse the stale value?
chainofthought.show/podcast/73-...
chainofthought.show/podcast/73-...