FDE中国 FDE 名录
← 返回名录
J

JY Feng

GitHub @tequila28 ↗ · 个人主页 ↗

M.S.@ PKU, B.E.@ BUPT. Research on Reinforcement Learning and Distillation for the Self-Improvement of Large Language Models in Post-Training.

公司
PKU&BUPT
位置
Beijing
Stars
7
粉丝 / 仓库
1

代表作品 / 项目

  • SGG-R3⭐ 3Official Implementation of ACL Findings Paper: SGG-R3: From Next-Token Prediction to End-to-End Unbiased Scene Graph Generation
  • Foundations-of-RL-Learning⭐ 2Foundations-of-RL-Learning-From-Scratch is a hands-on educational library providing clear, from-scratch Python implementations of foundational reinforcement learning algorithms for practical learning.
  • tequila28⭐ 1
  • Java-web-application⭐ 1
  • math⭐ 0
条目信息来自其公开主页与公开发布内容。按本站规范,页面不展示任何联系方式。 需要更正或删除?通过收录与更正通道提交,24 小时内处理。