Datasets:
Tasks:
Text Classification
Formats:
parquet
Languages:
English
Size:
1K - 10K
Tags:
structured-decisions
calibration
probabilistic-classification
system-one
workflow-evaluation
Synthetic
License:
Leaderboard submission: OpenDecider (open weights, Apache-2.0), one zero-shot model and four fine-tuned on `train`
👍 1
#8 opened about 10 hours ago
by
manjunathshiva
Bongard-mini: general, zero-shot, open weights (accuracy 0.594, KL 0.256, Brier 0.132)
👍 1
#7 opened about 11 hours ago
by
0xDing
Japanese translation: GeneLab/typed-decisions-ja
👍 1
#6 opened about 16 hours ago
by
GeneLab
prima-ratio + 12B (generalist, zero-shot, default calibration): 0.702 acc / KL 0.564 / Brier 0.234 / ECE 0.146
🚀 1
#5 opened 1 day ago
by
j3st3r666
od1-typed-decisions (specialist, 4B): 0.7965 acc / KL 0.082 / Brier 0.045 on the official test split
🔥 1
#4 opened 2 days ago
by
mvbalaji
soft-decider-421m (specialist): 0.774 acc / 0.141 ECE on the official test split
👍 1
#3 opened 5 days ago
by
winwinwinbb
Laya benchmark results
👍 3
1
#2 opened 11 days ago
by
convaiinnovations
[bot] Conversion to Parquet
#1 opened 12 days ago
by
parquet-converter