Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
97.6
TFLOPS
Thomas Wolf
PRO
thomwolf
196
204
557
Follow
shiwenwen's profile picture
arshad-ml's profile picture
NMontanaBrown's profile picture
1,854 followers
·
2,022 following
https://thomwolf.io
Thom_wolf
thomwolf
thom-wolf
thomwolf.bsky.social
AI & ML interests
NLP and open-source :-)
Recent Activity
new
activity
4 days ago
dlouapre/tangible-optimizers:
Update README.md
View all activity
Organizations
thomwolf
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
dlouapre/tangible-optimizers
4 days ago
Update README.md
#1 opened 4 days ago by
thomwolf
New activity in
lvwerra/cowrite
about 1 month ago
Recover agent polling after interruptions
#10 opened about 1 month ago by
thomwolf
Make agent prompt work with private Spaces
#5 opened about 1 month ago by
thomwolf
One page: the document is the app, and a switcher replaces the index
#4 opened about 1 month ago by
thomwolf
Document style switch: Google Docs' Arial, or a serif reading setting
#3 opened about 1 month ago by
thomwolf
Hand comments to an agent that is actually listening
#2 opened about 1 month ago by
thomwolf
Table columns resize by dragging a cell border
1
#1 opened about 1 month ago by
thomwolf
New activity in
rl-llm-wiki/rl-wiki
about 2 months ago
Deep links: ?q= term auto-highlight on arrival
1
#3 opened about 2 months ago by
thomwolf
Restore in-body search + highlight (regressed by 34252e5); ?q= deep-links reuse it
#4 opened about 2 months ago by
thomwolf
New activity in
rl-llm-wiki/knowledge-base
about 2 months ago
fix: arxiv:2412.16339 — CC BY 4.0 license, v2 provenance, o1-preview, table provenance, orphan ref
3
#661 opened about 2 months ago by
thomwolf
New activity in
rl-llm-wiki/rl-wiki
about 2 months ago
Search inside article bodies from the top-left box; ⌘K focuses it
#2 opened 2 months ago by
thomwolf
New activity in
rl-llm-wiki/knowledge-base
about 2 months ago
source: arxiv:2412.16339 — Deliberative Alignment (Reasoning Enables Safer LMs)
6
#595 opened about 2 months ago by
thomwolf
source: arxiv:2412.16720 — OpenAI o1 System Card
2
#580 opened about 2 months ago by
bfuzzy1
source: arxiv:2402.00658 — Learning Planning-based Reasoning via Trajectories Collection and Process Reward Synthesizing
2
#579 opened about 2 months ago by
bfuzzy1
source: arxiv:2404.19733 — Iterative Reasoning Preference Optimization
2
#577 opened about 2 months ago by
bfuzzy1
source: arxiv:2403.17031 — The N+ Implementation Details of RLHF with PPO (TL;DR Summarization)
2
#576 opened about 2 months ago by
bfuzzy1
topic: entropy-and-exploration — deepen to comprehensive
2
#582 opened about 2 months ago by
bfuzzy1
topic: policy-gradient-methods — deepen + add citations
2
#594 opened about 2 months ago by
bfuzzy1
topic: kl-regularization — build out from stub
2
#587 opened about 2 months ago by
bfuzzy1
topic: test-time-and-rl-interplay — deepen to comprehensive
4
#567 opened about 2 months ago by
bfuzzy1
Load more