๐Ÿ˜ฑ screamingface/ benchmarks github docs
Leaderboard

Fusions, ranked and reproducible.

A fusion is one or more models scored on a public benchmark.

Reproducible
Ran on shared compute and stored on the global cache. Anyone can re-run it and get the same score.
Unverified
Self-reported, or imported from a third-party board. Not yet reproduced on the global cache.
Solo model
A single model, shown as a reference baseline (no submitter).
Read this first By default, the leaderboard only shows results we've reproduced ourselves. The top of each leaderboard is the best verified result: the current SOTA. Toggle on self-reported runs if you want to see unverified claims too; some rank higher, but they haven't yet been confirmed.

Get started with ScreamingFace โ†’

Benchmarks

Pick a benchmark to open its leaderboard. Inside, you can tab across all benchmarks.

Loading…
๐Ÿ˜ฑ screamingface leaderboard.screamingface.ai ยท MVP preview ยท mock data