arXiv.org
AI Can Learn Scientific Taste
The paper introduces Reinforcement Learning from Community Feedback (RLCF), a paradigm that uses citation-based community signals to train AI models for scientific taste—the ability to judge and propose high-impact research ideas. They built SciJudgeBench with 720K field- and time-matched paper abstract pairs, trained Scientific Judge via GRPO to predict…
Jingqi Tong, Mingzhe Li, Hangcheng Li, Yongzhuo Yang, et al.- Published
- Mar 2026
- Upvotes
- 432
- Citations
- 4