💲
Focusing
First-year master's student @ Institute of Automation, Chinese Academy of Sciences
-
Institute of Automation, Chinese Academy of Sciences
- Beijing
- https://trae1oung.github.io/
Pinned Loading
-
Pretrain_Space_RLVR
Pretrain_Space_RLVR Public[arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space
Python 17
-
posttrainbench0
posttrainbench0 Public[Blog] PostTrainBench⁰ — Can LLM agents automate LLM post-training without gradients?
Python 11
-
paper-plot-skills
paper-plot-skills PublicTop-Conference Paper Figure Reproduction & Plotting Skills | 顶会论文图表复现绘制Skills
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.

