libero-long is a state-of-the-art machine learning benchmark indexed on Codesota. This page tracks published model results, top scores per metric, and the SOTA timeline for libero-long.
Success Rate is the reported evaluation metric for libero-long. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.
Higher is better
Muted rows were not state of the art when published — an earlier or same-year result already scored better.
| Rank | Model | Trust | Score | Year | Links | Fix |
|---|---|---|---|---|---|---|
| 01 | π0 (Pi-Zero) | vendor | 85.2 | 2026 | Source ↗ | Looks wrong? |
| 02 | OpenVLA | vendor | 53.7 | 2026 | Source ↗ | Looks wrong? |
| 03 | Octo-Base | vendor | 51.1 | 2026 | Source ↗ | Looks wrong? |