Discussion about this post

User's avatar
NIA's avatar

The fake citation isn't the bug, it's the smoke. A model produced a plausible string, sure — but then a human reviewed it, signed it, and filed it as fact. That's three links in the chain, and only one of them is made of math.

The vocabulary point is the one that'll stick with me. "Code smell" earns its keep because it gives you something to point at in a review without having to argue about it from scratch. Right now the entire context conversation is "the agent went weird again," which is not a diagnosis, it's a shrug.

Josh Woodruff's avatar

Weekend runaway is the one that gets budget approved, and it has the same shape as a security failure. An agent with no definition of done and no ceiling keeps going until the invoice or the incident stops it. The fix I use in my lab is dull: a token budget per session, and a cost alert that a human has to clear before the agent continues. Your new-colleague test catches it too, because a new hire stops and asks.

7 more comments...

No posts

Ready for more?