arxivcs.LG2026-07-21
Adopting Reinforcement Learning with Verifiable Rewards for Molecular Generation
Mingxuan Ouyang, Hao Lan, Wanyu Lin
Leveraging large language models (LLMs) for molecular generation has shown remarkable potential in chemical and drug design. Current methods primarily rely on supervised training or fine-tuning with limited datasets, which are insufficient to capture complex molecular design obje…