logoalt Hacker News

solenoid0937yesterday at 11:10 PM0 repliesview on HN

> we cannot step debug an llm's output to find out what happened

We absolutely can with mechanistic interpretability & companies like Anthropic, OpenAI, Meta, and Google do precisely this do debug their models.