← 返回名录条目信息来自其公开主页与公开发布内容。按本站规范,页面不展示任何联系方式。 需要更正或删除?通过收录与更正通道提交,24 小时内处理。
J
JY Feng
M.S.@ PKU, B.E.@ BUPT. Research on Reinforcement Learning and Distillation for the Self-Improvement of Large Language Models in Post-Training.
- 公司
- PKU&BUPT
- 位置
- Beijing
- Stars
- 7
- 粉丝 / 仓库
- 1
代表作品 / 项目
- SGG-R3⭐ 3Official Implementation of ACL Findings Paper: SGG-R3: From Next-Token Prediction to End-to-End Unbiased Scene Graph Generation
- Foundations-of-RL-Learning⭐ 2Foundations-of-RL-Learning-From-Scratch is a hands-on educational library providing clear, from-scratch Python implementations of foundational reinforcement learning algorithms for practical learning.
- tequila28⭐ 1
- Java-web-application⭐ 1
- math⭐ 0