Comment by brcmthrowaway 2 days ago Do you have the answer to the second question? Is an LLM trained on the internet just GPT-3? 1 comment brcmthrowaway Reply libraryofbabel 2 days ago I don't know - perhaps someone who's more of an expert or who's worked a lot with open source models that haven't been RL-ed can weigh in here!But certainly without the RL step, the LLM would be much worse at coding and would hallucinate more.
libraryofbabel 2 days ago I don't know - perhaps someone who's more of an expert or who's worked a lot with open source models that haven't been RL-ed can weigh in here!But certainly without the RL step, the LLM would be much worse at coding and would hallucinate more.
I don't know - perhaps someone who's more of an expert or who's worked a lot with open source models that haven't been RL-ed can weigh in here!
But certainly without the RL step, the LLM would be much worse at coding and would hallucinate more.