Rare find

AI-Can-Learn-Scientific-Taste. We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervision, and formulate scientific taste learning as a preference modeling and alignment problem.

github.com/tongjingqi/AI-Can-Learn-Scientific-Taste

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.