← Back to context

Comment by celrod

2 hours ago

I tried it a few times and liked the speed, but often found it ended up looping, i.e. repeating the same token sequence (e.g. the same sequence of 5 paragraphs) over and over again until it hit the max output limit. This doesn't end up happening every session, but does every now and then.

My impression of DSv4.1-flash was very positive aside from this. But that was enough for me to stick with GLM-5.3(-flash), which both gave me consistently great results

I was using a vibe coded bare bones harness. I was wondering if this was normal from DSv4.1-flash, or if its my harnesses fault.

I've had that looping issue with open models too. But never 4.1. I wonder if it's a model + harness combo? But yeah, one loop issue and I'm done with a model forever.

  • Harness. Especially if a tool call error doesn't say what to do next and the model is not RL'd with that tool, a retry storm is common.

    So if you use MCP a lot, simplify the params, be more lenient on validation and rework the errors.

    It is quite good with shell.