Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
healthydebate
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
I Thought the Optimizer Was the Product. I Was Wrong. The Gate Was.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Sep 4
I Thought the Optimizer Was the Product. I Was Wrong. The Gate Was.
#
healthydebate
#
ai
#
llm
#
testing
5
 reactions
Comments
Add Comment
5 min read
I Thought This Was a Classification Problem. It Wasn't.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Sep 4
I Thought This Was a Classification Problem. It Wasn't.
#
healthydebate
#
ai
#
llm
#
testing
8
 reactions
Comments
Add Comment
6 min read
I Compared a Local 4B, a Better Cloud Model, and Role Separation. The Results Were Weird.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Sep 4
I Compared a Local 4B, a Better Cloud Model, and Role Separation. The Results Were Weird.
#
healthydebate
#
ai
#
llm
#
testing
7
 reactions
Comments
Add Comment
6 min read
What I Learned Building a Diabetes Management Website
Reversemydiabetes
Reversemydiabetes
Reversemydiabetes
Follow
Sep 4
What I Learned Building a Diabetes Management Website
#
healthydebate
#
beginners
#
mentalhealth
#
webdev
Comments
Add Comment
2 min read
My Self-Improving Agent Still Couldn't Improve. That Was the Breakthrough.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Sep 4
My Self-Improving Agent Still Couldn't Improve. That Was the Breakthrough.
#
healthydebate
#
ai
#
llm
#
testing
8
 reactions
Comments
Add Comment
7 min read
I Let an LLM Rewrite Its Own Prompt. The Real Win Was the Gate That Rejected It.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Sep 2
I Let an LLM Rewrite Its Own Prompt. The Real Win Was the Gate That Rejected It.
#
healthydebate
#
ai
#
llm
#
agents
11
 reactions
Comments
1
 comment
6 min read
THE DEATH OF THE CRAFTSMAN: Software, AI Slop, and the Loss of the Soul in Coding
Marcus Vane
Marcus Vane
Marcus Vane
Follow
Aug 28
THE DEATH OF THE CRAFTSMAN: Software, AI Slop, and the Loss of the Soul in Coding
#
healthydebate
#
productivity
#
programming
Comments
Add Comment
4 min read
I Tried 4 Models to Save My Self-Improving Agent. All 4 Failed.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Sep 3
I Tried 4 Models to Save My Self-Improving Agent. All 4 Failed.
#
healthydebate
#
ai
#
llm
#
agents
23
 reactions
Comments
1
 comment
7 min read
My Agent Found Real Improvements. The Statistics Still Killed the Promotion.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Sep 2
My Agent Found Real Improvements. The Statistics Still Killed the Promotion.
#
healthydebate
#
ai
#
agents
#
llm
11
 reactions
Comments
1
 comment
6 min read
I Published Every Flaw My Safety Tool Can't Catch. It Made It More Credible, Not Less.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Aug 31
I Published Every Flaw My Safety Tool Can't Catch. It Made It More Credible, Not Less.
#
healthydebate
#
ai
#
testing
#
llm
10
 reactions
Comments
3
 comments
8 min read
My LLM Critic Flip-Flops on Every Run. That's Fine — Because a Frozenset Decides What's Fatal.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Aug 31
My LLM Critic Flip-Flops on Every Run. That's Fine — Because a Frozenset Decides What's Fatal.
#
healthydebate
#
ai
#
testing
#
llm
13
 reactions
Comments
7
 comments
7 min read
The Gate That Stayed Silent — When a Blocker Count That Drops Reads as Improvement
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Aug 31
The Gate That Stayed Silent — When a Blocker Count That Drops Reads as Improvement
#
healthydebate
#
ai
#
planner
#
testing
12
 reactions
Comments
5
 comments
5 min read
I Added a Fourth Model Mid-Run. It Changed What My Field Test Could Prove.
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Aug 30
I Added a Fourth Model Mid-Run. It Changed What My Field Test Could Prove.
#
healthydebate
#
ai
#
llm
#
testing
11
 reactions
Comments
Add Comment
7 min read
The Best Model Pair in My Field Test Was Also the Least Trustworthy
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Aug 29
The Best Model Pair in My Field Test Was Also the Least Trustworthy
#
healthydebate
#
ai
#
trustworthy
#
llm
22
 reactions
Comments
7
 comments
12 min read
Two Projects, One Problem — What PlannerCritic and AdversarialDebate Each Got Wrong
Debashish Ghosal
Debashish Ghosal
Debashish Ghosal
Follow
Aug 29
Two Projects, One Problem — What PlannerCritic and AdversarialDebate Each Got Wrong
#
healthydebate
#
ai
#
dogfood
#
testing
17
 reactions
Comments
3
 comments
10 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account