Do LLMs Need Inherent Reasoning Before Reinforcement Learning? A Study in Korean Self-Correction.
Hongjin Kim, Jaewook Lee, Kiyoung Lee, Jong-hun Shin, Soojong Lim, Oh-Woog Kwon
Browse the full IJCNLP paper archive.
Hongjin Kim, Jaewook Lee, Kiyoung Lee, Jong-hun Shin, Soojong Lim, Oh-Woog Kwon
Browse the full IJCNLP paper archive.