
We Could Have Shipped on Local Models Alone
Update — v0.1.0 released. CauterRule is now live on GitHub and PyPI. It turns repeated agent...

Author profile
TLDR - Engineer -> Manager -> Tech Leadership, passionate about AI, Developer experience, planet scale infra, transformation - AI-native leadership. Rest here - https://www.linkedin.com/in/deghosal/
Browse the latest writing surfaced through DevArt.

Update — v0.1.0 released. CauterRule is now live on GitHub and PyPI. It turns repeated agent...

Update — v0.1.0 released. CauterRule is now live on GitHub and PyPI. It turns repeated agent...

Previously: 9 Bugs That All Looked Like a Working System · I Built an AI That Rewrites Its Own...

Previously: 9 Bugs That All Looked Like a Working System · I Built an AI That Rewrites Its Own...

Previously: 9 Bugs That All Looked Like a Working System · I Built an AI That Rewrites Its Own...

Previously: 9 Bugs That All Looked Like a Working System · I Built an AI That Rewrites Its Own...

Previously: 9 Bugs That All Looked Like a Working System · I Built an AI That Rewrites Its Own...

Previously: 9 Bugs That All Looked Like a Working System · I Built an AI That Rewrites Its Own...

Previously: 9 Bugs That All Looked Like a Working System · I Built an AI That Rewrites Its Own...

Previously: 9 Bugs That All Looked Like a Working System · I Built an AI That Rewrites Its Own...

AgentSelfEdit is an open-source sidecar that rewrites its own system prompt from execution feedback....

AgentSelfEdit is an open-source sidecar that rewrites its own system prompt from execution feedback....

AgentSelfEdit is an open-source sidecar that rewrites its own system prompt from execution feedback....

This is a companion to the PlannerCritic series. Article 5 was about what happened when I tried to...

This is a companion to the PlannerCritic series. Article 2 was about a specific critic bug. This one...

In the last article, I argued that the safety contract should move out of the LLM critic and into...

Latest release: v0.2.2 — Aug 29, 2026 I did something I usually try hard not to do in a field...

v0.2.1 RELEASED — Aug 28, 2026. Release notes · Field test report v0.2.1 Key Finding: DeepSeek+GPT...

I built two systems this year that try to solve the same problem from opposite...

v0.2.1 RELEASED — Aug 28, 2026. Release notes · Field test report · v0.2.2 release notes v0.2.2...

v0.2.1 RELEASED — Aug 28, 2026. Release notes · Field test report · PyPI v0.2.1 Update: The...

The first time I ran two LLMs against the same pull request, 89% of their "debate" was fake. Not...

In the My Agent Refused 96 Times. That Was the Right Output., I argued that the most valuable output...

In the last article, I wrote about a release story that was weaker than the engine underneath...

Most AI "second opinions" are fake. Not because there is no second model. Because the second model...

When I shipped v0.2.1 of PlannerCritic, I thought the hard part was over. The engine had survived a...

This is article 5 in a series about building PlannerCritic, an open-source engine where one LLM...

This is article 4 in a series about building PlannerCritic, an open-source engine where one LLM...

This is article 3 in a series about building PlannerCritic, an open-source engine where one LLM...

This is article 2 in a series about building PlannerCritic, an open-source engine where one LLM...
Advertisement