Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
9.2
TFLOPS
EvalEval Bot
EvalEvalBot
344
1
Follow
evijit's profile picture
zozoosibanda2-creator's profile picture
deeplumiere's profile picture
4 followers
·
2 following
AI & ML interests
None yet
Recent Activity
updated
a dataset
about 8 hours ago
evaleval/EEE_datastore
updated
a Space
3 days ago
evaleval/general-eval-card
new
activity
10 days ago
evaleval/EEE_datastore:
LiveBench: drop manual records the adapter covers; file the rest in the adapter's layout
View all activity
Organizations
EvalEvalBot
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
evaleval/EEE_datastore
10 days ago
LiveBench: drop manual records the adapter covers; file the rest in the adapter's layout
3
#224 opened 10 days ago by
evijit
LiveBench: re-issue reused UUIDs and file records by model id
8
#223 opened 10 days ago by
evijit
New activity in
evaleval/EEE_datastore
14 days ago
[Submission] swiss-ai/Apertus 1.5 Evaluation Card Follow-Up Results (checked)
4
#222 opened 14 days ago by
ymetz
[Submission] swiss-ai/Apertus 1.5 Evaluation Card Follow-Up Results
4
#221 opened 14 days ago by
ymetz
New activity in
evaleval/EEE_datastore
22 days ago
[Submission] swiss-ai/Apertus 1.5 Evaluation Card Results
🚀
1
3
#217 opened 22 days ago by
ymetz
New activity in
evaleval/EEE_datastore
27 days ago
Carry the frozen hfopenllm_v2 archive to schema 0.3.0
2
#214 opened 27 days ago by
EvalEvalBot
New activity in
evaleval/EEE_datastore
about 1 month ago
Add aiXamine evaluation records (974 logs, 9 services)
4
#210 opened about 1 month ago by
fdeniz
[Submission] Add Kaggle Community Benchmarks results
9
#158 opened 4 months ago by
mrshu
[Submission] Add stablelm-evals: 29 models x 11 benchmarks, incl. SciQ and LAMBADA-OpenAI
5
#212 opened about 1 month ago by
borgr
[Submission] Add lm-harmony: 24 benchmarks x 61 models x 2 protocols
5
#211 opened about 1 month ago by
borgr
New activity in
evaleval/EEE_datastore
about 2 months ago
[Submission] Add Polistemics: political information mediation in elections
6
#207 opened about 2 months ago by
fl4wn
[Submission] cron: llm_stats (automated ingestion)
1
#201 opened about 2 months ago by
EvalEvalBot
[Submission] cron: mmlu_pro (automated ingestion)
1
#200 opened about 2 months ago by
EvalEvalBot
[Submission] cron: artificial_analysis (automated ingestion)
3
#190 opened about 2 months ago by
evijit
[Submission] cron: vals_ai (automated ingestion)
3
#198 opened about 2 months ago by
evijit
[Submission] cron: openeval (automated ingestion)
3
#199 opened about 2 months ago by
evijit
[Submission] cron: global_mmlu_lite (automated ingestion)
3
#189 opened about 2 months ago by
evijit
[Submission] cron: arc_agi (automated ingestion)
3
#192 opened about 2 months ago by
evijit
[Submission] cron: exgentic (automated ingestion)
3
#191 opened about 2 months ago by
evijit
[Submission] cron: hal (automated ingestion)
3
#195 opened about 2 months ago by
evijit
Load more