GH Repository · tongjingqi
AI-Can-Learn-Scientific-Taste
We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervision, and formulate scientific taste learning as a preference modeling and alignment problem.
- stars
- 432
- 30-day movement
- starts with the next reading
- Related entries
- 44
- Connections
- 1
agent-appai-innovatorrlai-scientistsagent
- Role
- agent-app
- Licence
- Apache-2.0
- Forks
- 11
- Last push
- 2026-07-22