Model cards

Strengths and weaknesses across budgets.

Compare models across dataset modalities and training budgets using strict, fold-level Macro-F1 Elo.

Model profile

—

——

Default parameters. Peer models use their prespecified benchmark-default configuration without per-dataset tuning. This card describes that configuration, not the best achievable tuned model.

Operating map

Macro-F1 Elo relative to Random Forest. Columns are training-sample caps; rows are dataset modalities. Small labels show the number of binding targets.

Elo difference vs. Random Forest
Fixed nonlinear scale; extremes capped at ±1,000 Elo.
No result / comparison unavailable