long context benchmark
nolima model scores and rankings.
Updated Aug 11, 2026
Higher score ranks better on this benchmark.
No model results are available for this benchmark yet.
The leading models and scores on this benchmark.
What nolima measures and how its scores work.
nolima is a long context evaluation benchmark.
Scores are shown in ratio. This benchmark is not independently verified and has an evidence level of C.
Benchmark scores retain their original unit. Overall score eligibility is shown separately.
Common questions about nolima.
No model result is currently available.
nolima evaluates long context capability.
Yes. Higher values rank better for this benchmark.
0 model results are currently shown.
No. This benchmark is shown for reference but does not contribute to the overall score.