But certainly without the RL step, the LLM would be much worse at coding and would hallucinate more.