モデル / データセット
tongjingqi/AI-Can-Learn-Scientific-Taste avatar
tongjingqi/AI-Can-Learn-Scientific-Taste

AI-Can-Learn-Scientific-Taste

We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervision, and formulate scientific taste learning as a preference modeling and alignment problem.

スター 432フォーク 11UnknownApache-2.0

ひと目でわかる

これは何?
We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervision, and formulate scientific taste learning as a preference modeling and alignment problem.
商用利用できる?
できます。Apache-2.0 は寛容なライセンスで、著作権表示とライセンス表示を残せば、使用・改変・販売が可能です。
今もメンテナンスされている?
されています。最後のコミットは 55 日前です。
何の言語で書かれている?
GitHub はこのリポジトリの主な言語を示していません。

回答はプロジェクトの GitHub データ(最終同期:2026年9月15日)と当サイトの分析に基づくもので、法的助言ではありません。

情報の鮮度

編集ステータス

このプロジェクトの完全な編集分析はまだ公開されていません。上の事実はプロジェクトの公開 GitHub メタデータに基づきます。本番利用の前に、リポジトリ、ライセンス、Issue トラッカーを確認してください。

コミュニティノート

コミュニティノート