Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
FINAL-Bench
's Collections
HF Ofiicial Benchs
ZTC Models: JEV ecosystems
AI Challanges
DARWIN-Family
POCKET-MODELs
VKAE Accelerated
'Aether' Foundation Model
Armoring Models
HF Ofiicial Benchs
updated
3 days ago
Upvote
20
+10
Sort: Collection
LocalLLaMA/typed-decisions
Benchmark
•
Updated
about 15 hours ago
•
3.2k
•
36.6k
•
159
Idavidrein/gpqa
Benchmark
•
Updated
19 days ago
•
1.25k
•
111k
•
580
TIGER-Lab/MMLU-Pro
Benchmark
•
Updated
May 2
•
12.1k
•
223k
•
546
MMMU/MMMU_Pro
Benchmark
•
Updated
6 days ago
•
5.19k
•
19.1k
•
91
llamaindex/ExtractBench
Benchmark
•
Updated
11 days ago
•
370
•
15.7k
•
58
LEXam-Benchmark/LEXam
Benchmark
•
Updated
May 21
•
7.54k
•
1.73k
•
71
MathArena/hmmt_feb_2026
Benchmark
•
Updated
May 15
•
33
•
18.5k
•
31
MathArena/aime_2026
Benchmark
•
Updated
May 15
•
30
•
31.5k
•
86
LiquidAI/ifstruct-v1.0
Benchmark
•
Updated
Jul 7
•
2k
•
610
•
103
joelniklaus/LEXam-hard
Benchmark
•
Updated
30 days ago
•
518
•
864
•
24
Delores-Lin/MDPBench
Benchmark
•
Updated
3 days ago
•
1.64k
•
48
Upvote
20
+16
Sort: Collection
Share collection
View history
Collection guide
Browse collections