Thomas Wolf PRO
thomwolf
AI & ML interests
NLP and open-source :-)
Organizations
Recover agent polling after interruptions
#10 opened 9 days ago
by
thomwolf
Make agent prompt work with private Spaces
#5 opened 11 days ago
by
thomwolf
One page: the document is the app, and a switcher replaces the index
#4 opened 12 days ago
by
thomwolf
Document style switch: Google Docs' Arial, or a serif reading setting
#3 opened 12 days ago
by
thomwolf
Hand comments to an agent that is actually listening
#2 opened 12 days ago
by
thomwolf
Table columns resize by dragging a cell border
1
#1 opened 12 days ago
by
thomwolf
Deep links: ?q= term auto-highlight on arrival
1
#3 opened 27 days ago
by
thomwolf
Restore in-body search + highlight (regressed by 34252e5); ?q= deep-links reuse it
#4 opened 25 days ago
by
thomwolf
fix: arxiv:2412.16339 — CC BY 4.0 license, v2 provenance, o1-preview, table provenance, orphan ref
3
#661 opened 26 days ago
by
thomwolf
Search inside article bodies from the top-left box; ⌘K focuses it
#2 opened about 1 month ago
by
thomwolf
source: arxiv:2412.16339 — Deliberative Alignment (Reasoning Enables Safer LMs)
6
#595 opened about 1 month ago
by
thomwolf
source: arxiv:2412.16720 — OpenAI o1 System Card
2
#580 opened about 1 month ago
by
bfuzzy1
source: arxiv:2402.00658 — Learning Planning-based Reasoning via Trajectories Collection and Process Reward Synthesizing
2
#579 opened about 1 month ago
by
bfuzzy1
source: arxiv:2404.19733 — Iterative Reasoning Preference Optimization
2
#577 opened about 1 month ago
by
bfuzzy1
source: arxiv:2403.17031 — The N+ Implementation Details of RLHF with PPO (TL;DR Summarization)
2
#576 opened about 1 month ago
by
bfuzzy1
topic: entropy-and-exploration — deepen to comprehensive
2
#582 opened about 1 month ago
by
bfuzzy1
topic: policy-gradient-methods — deepen + add citations
2
#594 opened about 1 month ago
by
bfuzzy1
topic: kl-regularization — build out from stub
2
#587 opened about 1 month ago
by
bfuzzy1
topic: test-time-and-rl-interplay — deepen to comprehensive
4
#567 opened about 1 month ago
by
bfuzzy1
topic: preference-reward-models — deepen + bump to comprehensive
2
#589 opened about 1 month ago
by
bfuzzy1