Comment by Marha01
12 hours ago
> Take away the heavy agentic scaffolding and external feedback loops, and a single hallucinated token can still derail the entire chain of thought.
With reasoning models, a derailed chain of thought can be rerailed.
12 hours ago
> Take away the heavy agentic scaffolding and external feedback loops, and a single hallucinated token can still derail the entire chain of thought.
With reasoning models, a derailed chain of thought can be rerailed.
> can be rerailed
What rerails it?
The model can realize it made a mistake earlier and correct itself in subsequent output. I have seen it happen many times in CoT.