← Back to context

Comment by walrus01

9 hours ago

I have never seen anyone report "this produced really great results" from intentionally quantizing their context vs. leaving it at full precision which is the ordinary default.

Gemma's QAT is surprisingly good (although Gemma isn't that great to begin with).

  • IME: Gemma is not great for programming, but it is fantastic at following directions compared to anything else in its size class.