> ## Documentation Index
> Fetch the complete documentation index at: https://modal-computer-use.mintlify.site/llms.txt
> Use this file to discover all available pages before exploring further.

# Benchmark overview

> Find current results, evidence definitions, and benchmark commands.

Use these pages in order. Current results lists the latest eligible measurements. Latency evidence maps each claim to its source and timer. Run benchmarks explains how to reproduce a result.

<CardGroup cols={3}>
  <Card title="Current results" href="/benchmarks/current-results" icon="chart-column">
    Read eligible measurements and the scope of each result.
  </Card>

  <Card title="Latency evidence" href="/benchmarks/latency-evidence" icon="microscope">
    Check each claim's status, measurement boundary, and proof.
  </Card>

  <Card title="Run benchmarks" href="/benchmarks/run" icon="player-play">
    Use the maintained commands to reproduce a result.
  </Card>
</CardGroup>

## How to read the evidence

Current results includes eligible measurements. Latency evidence also records historical, experimental, candidate, estimated, rejected, and unverified claims. Each dated report preserves the full method and links its sanitized artifact.

[Reproducibility](/benchmarks/reproducibility) defines the checks for source revisions, configurations, failures, cleanup, and publication.


This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.