arxiv:2407.18137
Jiahao s
lanlanlan23
AI & ML interests
None yet
Recent Activity
upvoted a paper about 23 hours ago
StableVQ: Practical Guidelines for Stable Vector-Quantized Tokenizer Training upvoted a paper 3 months ago
Learning from the Self-future: On-policy Self-distillation for dLLMs upvoted a paper 5 months ago
RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework