Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
alessandrobondielli
's Collections
LLMs-to-test
Datasets-ScaleLLM
MechInterp-Papers
Reading List - TextToImage
Datasets-ScaleLLM
updated
Jul 1, 2025
Upvote
-
Sort: Collection
truthfulqa/truthful_qa
Viewer
•
Updated
Jan 4, 2024
•
1.63k
•
107k
•
289
allenai/qasc
Viewer
•
Updated
Jan 4, 2024
•
9.98k
•
23.9k
•
23
Anthropic/model-written-evals
Viewer
•
Updated
Dec 21, 2022
•
3.25k
•
2.22k
•
67
yesilhealth/Health_Benchmarks
Viewer
•
Updated
Apr 20, 2025
•
7.54k
•
84
•
10
maveriq/bigbenchhard
Viewer
•
Updated
Sep 29, 2023
•
6.51k
•
1.62k
•
43
Note
Filtrare i subset che non hanno campo choice
tau/commonsense_qa
Viewer
•
Updated
Jan 4, 2024
•
12.1k
•
75.2k
•
152
allenai/sciq
Viewer
•
Updated
Jan 4, 2024
•
13.7k
•
144k
•
145
allenai/openbookqa
Viewer
•
Updated
Jan 4, 2024
•
11.9k
•
164k
•
135
allenai/ai2_arc
Viewer
•
Updated
Dec 21, 2023
•
7.79k
•
462k
•
379
TIGER-Lab/MMLU-Pro
Benchmark
•
Updated
May 2
•
12.1k
•
168k
•
509
Upvote
-
Sort: Collection
Share collection
View history
Collection guide
Browse collections