← Back to context

Comment by jeremyjh

2 hours ago

If you can't get any coding done with 100K context that is either a broken model, a broken harness or a skill issue. I would mostly use Haiku in task or explorer subagents. I'm not saying I stay under that on every task, but I do have quite a few sessions that cap out well below that, so that price difference would be very meaningful.

I use Luna for this day in and out and its excellent - if Haiku is that much better I will be changing things up.

>If you can't get any coding done with 100K context that is either a broken model, a broken harness or a skill issue.

"less context is better and if you can't get stuff done with less yur bad" is the worst argument ever.

it might be pure luxury to your eyes, but it's great to not require the use of a special custom harness that transcribes everything into emoji and compresses everything into barcode images.

it's great to have a million token context to throw a large project into. If I need 100k just about any current gen consumer GPU in the world has very good models that I can self host for 100k context, limiting myself to 100k on someone elses machine seems to be missing a lot of the point unless the model itself is extraordinary.