Comment by rao-v
6 hours ago
I wonder if you could take a traditional backprop trained LLM and apply this approach to finetuning it (presumably needs less memory and compute?). It could be another entry in the spectrum between LORA and full fine tuning.
No comments yet
Contribute on Hacker News ↗