Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

RLLab

https://github.com
Activity Feed

AI & ML interests

None defined yet.

Recent Activity

JixuanLeng  updated a model 6 days ago
RLLab/olmo-3-7b-it-sft-lr-1e-6-gdpo-len8k-gatew1.0
JixuanLeng  published a model 6 days ago
RLLab/olmo-3-7b-it-sft-lr-1e-6-gdpo-len8k-gatew1.0
JixuanLeng  updated a dataset 8 days ago
RLLab/MTMR
View all activity

Jixuan Leng's profile picture

RLLab 's datasets 11

RLLab/MTMR

Viewer • Updated 8 days ago • 327k • 429

RLLab/eval-set

Viewer • Updated 23 days ago • 12.4k • 177

RLLab/safe-alignment-dynamic

Viewer • Updated about 1 month ago • 576k • 41

RLLab/RaR-Science-Grouped

Viewer • Updated Aug 9 • 18.8k • 14

RLLab/RaR-Medicine-Grouped

Viewer • Updated Aug 9 • 19.7k • 21

RLLab/allenai-Dolci-Instruct-DPO-Length-Filtered

Viewer • Updated Mar 1 • 146k • 4

RLLab/OpenR1-Math-220K-Filtered-DPO

Viewer • Updated Feb 3 • 79.3k • 4

RLLab/OpenR1-Math-220k-Filtered-Generations

Viewer • Updated Feb 3 • 3.6M • 4

RLLab/OpenR1-Math-220k-Filtered

Viewer • Updated Jan 28 • 225k • 37

RLLab/allenai-Dolci-Instruct-DPO-Filtered-Generations

Viewer • Updated Jan 18 • 4.32M • 3

RLLab/math-rl

Viewer • Updated Nov 25, 2025 • 57.5k • 14
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs