Datasets:
Tasks:
Text Classification
Formats:
parquet
Languages:
English
Size:
1K - 10K
Tags:
structured-decisions
calibration
probabilistic-classification
system-one
workflow-evaluation
Synthetic
License:
prima-ratio + 12B (generalist, zero-shot, default calibration): 0.702 acc / KL 0.564 / Brier 0.234 / ECE 0.146
🚀 1
#5 opened about 10 hours ago
by
j3st3r666
od1-typed-decisions (specialist, 4B): 0.7965 acc / KL 0.082 / Brier 0.045 on the official test split
🔥 1
#4 opened 1 day ago
by
mvbalaji
soft-decider-421m (specialist): 0.774 acc / 0.141 ECE on the official test split
👍 1
#3 opened 4 days ago
by
winwinwinbb
Laya benchmark results
👍 3
1
#2 opened 11 days ago
by
convaiinnovations
[bot] Conversion to Parquet
#1 opened 11 days ago
by
parquet-converter