Gitstar Ranking
Users
Organizations
Repositories
Rankings
Users
Organizations
Repositories
Sign in with GitHub
dexhunter
Fetched on 2026/08/19 19:37
dexhunter
/
AI-Can-Learn-Scientific-Taste
We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervision, and formulate scientific taste learning as a preference modeling and alignment problem. -
View it on GitHub
https://tongjingqi.github.io/AI-Can-Learn-Scientific-Taste/
Star
0
Rank
14408395