๐ In a Training Loop
dong zi
shenyao
AI & ML interests
None yet
Recent Activity
upvoted a paper about 1 month ago
Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective upvoted a paper about 1 month ago
RAVE: Re-Allocating Visual Attention in Large Multimodal Models upvoted a paper about 2 months ago
Learning from Your Own Mistakes: Constructing Learnable Micro-Reflective Trajectories for Self-Distillation