Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I don't know - perhaps someone who's more of an expert or who's worked a lot with open source models that haven't been RL-ed can weigh in here!

But certainly without the RL step, the LLM would be much worse at coding and would hallucinate more.





Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: