Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
97.6
TFLOPS
Thomas Wolf
PRO
thomwolf
196
206
559
Follow
madhukarreddyk's profile picture
iamsingularity's profile picture
amuvarma's profile picture
1,885 followers
·
2,031 following
https://thomwolf.io
Thom_wolf
thomwolf
thom-wolf
thomwolf.bsky.social
AI & ML interests
NLP and open-source :-)
Recent Activity
liked
a model
2 days ago
AikidoSec/altar-1
upvoted
a
collection
7 days ago
Common Pile v0.1
liked
a dataset
9 days ago
allenai/objaverse
View all activity
Organizations
thomwolf
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
dlouapre/tangible-optimizers
22 days ago
Update README.md
#1 opened 22 days ago by
thomwolf
New activity in
lvwerra/cowrite
about 2 months ago
Recover agent polling after interruptions
#10 opened about 2 months ago by
thomwolf
Make agent prompt work with private Spaces
#5 opened about 2 months ago by
thomwolf
One page: the document is the app, and a switcher replaces the index
#4 opened about 2 months ago by
thomwolf
Document style switch: Google Docs' Arial, or a serif reading setting
#3 opened about 2 months ago by
thomwolf
Hand comments to an agent that is actually listening
#2 opened about 2 months ago by
thomwolf
Table columns resize by dragging a cell border
1
#1 opened about 2 months ago by
thomwolf
New activity in
rl-llm-wiki/rl-wiki
2 months ago
Deep links: ?q= term auto-highlight on arrival
1
#3 opened 2 months ago by
thomwolf
Restore in-body search + highlight (regressed by 34252e5); ?q= deep-links reuse it
#4 opened 2 months ago by
thomwolf
New activity in
rl-llm-wiki/knowledge-base
2 months ago
fix: arxiv:2412.16339 — CC BY 4.0 license, v2 provenance, o1-preview, table provenance, orphan ref
3
#661 opened 2 months ago by
thomwolf
New activity in
rl-llm-wiki/rl-wiki
2 months ago
Search inside article bodies from the top-left box; ⌘K focuses it
#2 opened 3 months ago by
thomwolf
New activity in
rl-llm-wiki/knowledge-base
2 months ago
source: arxiv:2412.16339 — Deliberative Alignment (Reasoning Enables Safer LMs)
6
#595 opened 2 months ago by
thomwolf
source: arxiv:2412.16720 — OpenAI o1 System Card
2
#580 opened 3 months ago by
bfuzzy1
source: arxiv:2402.00658 — Learning Planning-based Reasoning via Trajectories Collection and Process Reward Synthesizing
2
#579 opened 3 months ago by
bfuzzy1
source: arxiv:2404.19733 — Iterative Reasoning Preference Optimization
2
#577 opened 3 months ago by
bfuzzy1
source: arxiv:2403.17031 — The N+ Implementation Details of RLHF with PPO (TL;DR Summarization)
2
#576 opened 3 months ago by
bfuzzy1
New activity in
rl-llm-wiki/knowledge-base
3 months ago
topic: entropy-and-exploration — deepen to comprehensive
2
#582 opened 3 months ago by
bfuzzy1
topic: policy-gradient-methods — deepen + add citations
2
#594 opened 3 months ago by
bfuzzy1
topic: kl-regularization — build out from stub
2
#587 opened 3 months ago by
bfuzzy1
topic: test-time-and-rl-interplay — deepen to comprehensive
4
#567 opened 3 months ago by
bfuzzy1
Load more