Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Log In
Sign Up
5
7
61
VIDRAFT_LAB
SeaWolf-AI
Follow
21world's profile picture
viv's profile picture
omiddhn's profile picture
79 followers
ยท
151 following
AI & ML interests
None yet
Recent Activity
reacted
to
mayafree
's
post
with ๐
about 13 hours ago
Leaderboard of Leaderboards โ A Real-Time Meta-Ranking of AI Benchmarks https://huggingface.co/spaces/MAYA-AI/all-leaderboard Hundreds of AI leaderboards exist on HuggingFace. Knowing which ones the community actually trusts has never been easy โ until now. Leaderboard of Leaderboards (LoL) ranks the leaderboards themselves, using live HuggingFace trending scores and cumulative likes as the signal. No editorial curation. No manual selection. Just what the global AI research community is actually visiting and endorsing, surfaced in real time. Sort by trending to see what is capturing attention right now, or by likes to see what has built lasting credibility over time. Nine domain filters let you zero in on what matters most to your work, and every entry shows both its rank within this collection and its real-time global rank across all HuggingFace Spaces. The collection spans well-established standards like Open LLM Leaderboard, Chatbot Arena, MTEB, and BigCodeBench alongside frameworks worth watching. FINAL Bench targets AGI-level evaluation across 100 tasks in 15 domains and recently reached the global top 5 in HuggingFace dataset rankings. Smol AI WorldCup runs tournament-format competitions for sub-8B models scored via FINAL Bench criteria. ALL Bench aggregates results across frameworks into a unified ranking that resists the overfitting risks of any single standard. The deeper purpose is not convenience. It is transparency. How we measure AI matters as much as the AI we measure.
reacted
to
mayafree
's
post
with ๐ฅ
about 13 hours ago
Leaderboard of Leaderboards โ A Real-Time Meta-Ranking of AI Benchmarks https://huggingface.co/spaces/MAYA-AI/all-leaderboard Hundreds of AI leaderboards exist on HuggingFace. Knowing which ones the community actually trusts has never been easy โ until now. Leaderboard of Leaderboards (LoL) ranks the leaderboards themselves, using live HuggingFace trending scores and cumulative likes as the signal. No editorial curation. No manual selection. Just what the global AI research community is actually visiting and endorsing, surfaced in real time. Sort by trending to see what is capturing attention right now, or by likes to see what has built lasting credibility over time. Nine domain filters let you zero in on what matters most to your work, and every entry shows both its rank within this collection and its real-time global rank across all HuggingFace Spaces. The collection spans well-established standards like Open LLM Leaderboard, Chatbot Arena, MTEB, and BigCodeBench alongside frameworks worth watching. FINAL Bench targets AGI-level evaluation across 100 tasks in 15 domains and recently reached the global top 5 in HuggingFace dataset rankings. Smol AI WorldCup runs tournament-format competitions for sub-8B models scored via FINAL Bench criteria. ALL Bench aggregates results across frameworks into a unified ranking that resists the overfitting risks of any single standard. The deeper purpose is not convenience. It is transparency. How we measure AI matters as much as the AI we measure.
liked
a Space
about 13 hours ago
MAYA-AI/all-leaderboard
View all activity
Organizations
SeaWolf-AI
's activity
All
Models
Datasets
Spaces
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
FINAL-Bench/all-bench-leaderboard
1 day ago
A new model has been listed on the All Bench leaderboard.
#2 opened 1 day ago by
SeaWolf-AI
New activity in
FINAL-Bench/all-bench-leaderboard
4 days ago
Request for Benchmark Evaluation
#1 opened 4 days ago by
SeaWolf-AI
New activity in
OpenEvals/README
4 days ago
New Benchmark Dataset
๐
5
12
#2 opened about 1 month ago by
burtenshaw
New activity in
FINAL-Bench/Leaderboard
14 days ago
What would Alan Turing think of todays LLM's?
1
#1 opened 14 days ago by
Whyvette
New activity in
FINAL-Bench/Metacognitive
20 days ago
Welcome!
#1 opened 20 days ago by
SeaWolf-AI