BigHugger
GH Repository · tongjingqi

AI-Can-Learn-Scientific-Taste

We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervision, and formulate scientific taste learning as a preference modeling and alignment problem.

stars
432
30-day movement
starts with the next reading
Related entries
44
Connections
1
agent-appai-innovatorrlai-scientistsagent
Role
agent-app
Licence
Apache-2.0
Forks
11
Last push
2026-07-22