Arena benchmark
LM Arena Search Style Control model ratings and results.
Updated Aug 17, 2026. Data as of 2026-08-17
Models are ordered by their Arena rank.
Rank | Model | Organization | Rating / score | Votes | Observations | Result date |
|---|
| Rank01 | ModelAN | Organizationanthropic | Rating / score1236.8 | Votes20,630 | ObservationsN/A | Result date |
| Rank02 | ModelOP | Organizationopenai | Rating / score1223.5 | Votes67,573 | ObservationsN/A | Result date |
| Rank03 | ModelAN | Organizationanthropic | Rating / score1222.7 | Votes112,201 | ObservationsN/A | Result date |
| Rank04 | ModelAN | Organizationanthropic | Rating / score1211.0 | Votes68,515 | ObservationsN/A | Result date |
| Rank05 | ModelGO | Organizationgoogle | Rating / score1209.7 | Votes89,973 | ObservationsN/A | Result date |
| Rank06 | ModelXA | Organizationxai | Rating / score1206.8 | Votes87,118 | ObservationsN/A | Result date |
| Rank07 | ModelGO | Organizationgoogle | Rating / score1201.2 | Votes37,255 | ObservationsN/A | Result date |
| Rank08 | ModelAN | Organizationanthropic | Rating / score1199.6 | Votes111,959 | ObservationsN/A | Result date |
| Rank10 | ModelAN | Organizationanthropic | Rating / score1195.6 | Votes48,610 | ObservationsN/A | Result date |
| Rank11 | ModelOP | Organizationopenai | Rating / score1194.6 | Votes87,364 | ObservationsN/A | Result date |
| Rank14 | ModelAN | Organizationanthropic | Rating / score1191.2 | Votes17,604 | ObservationsN/A | Result date |
| Rank15 | ModelOP | Organizationopenai | Rating / score1190.9 | Votes20,785 | ObservationsN/A | Result date |
| Rank16 | ModelXA | Organizationxai | Rating / score1190.5 | Votes68,373 | ObservationsN/A | Result date |
| Rank17 | ModelGO | Organizationgoogle | Rating / score1190.2 | Votes125,928 | ObservationsN/A | Result date |
| Rank18 | ModelAN | Organizationanthropic | Rating / score1186.9 | Votes61,611 | ObservationsN/A | Result date |
| Rank19 | ModelOP | Organizationopenai | Rating / score1183.6 | Votes60,019 | ObservationsN/A | Result date |
| Rank20 | ModelOP | Organizationopenai | Rating / score1182.7 | Votes20,913 | ObservationsN/A | Result date |
| Rank21 | ModelAN | Organizationanthropic | Rating / score1175.8 | Votes105,762 | ObservationsN/A | Result date |
| Rank22 | ModelAN | Organizationanthropic | Rating / score1166.7 | Votes77,222 | ObservationsN/A | Result date |
| Rank23 | ModelOP | Organizationopenai | Rating / score1160.5 | Votes52,718 | ObservationsN/A | Result date |
| Rank24 | ModelXA | Organizationxai | Rating / score1159.6 | Votes42,981 | ObservationsN/A | Result date |
| Rank25 | ModelAN | Organizationanthropic | Rating / score1150.7 | Votes31,205 | ObservationsN/A | Result date |
| Rank26 | ModelGO | Organizationgoogle | Rating / score1142.5 | Votes83,608 | ObservationsN/A | Result date |
| Rank30 | ModelXA | Organizationxai | Rating / score1121.5 | Votes19,379 | ObservationsN/A | Result date |
A closer view of the leading model ratings.
The first five models in this source ranking, with price and output speed included when a canonical model match is available.
Ranking basisThis search style control AI model leaderboard uses the source rating and Arena rank. The leaderboard ranking keeps matched price and speed data separate from benchmark evidence.
Selection summary
Claude Fable 5 currently leads Search Style Control at 1,237. The best AI model for this use case may change when price, speed and independent benchmark capability are considered.
Use this leaderboard with the supporting benchmark results and coverage details above. A leaderboard position summarizes the selected ranking signal; it does not replace workload-specific testing.
What Search Style Control measures and how to read its results.
Search Style Control reports rating values and currently defines 1 evaluation category.
Arena results are preference or task-outcome signals. They are not automatically equivalent to capability benchmark scores.
Common questions about Search Style Control.
Claude Fable 5 is currently ranked first.
This Arena reports rating values across 1 configured evaluation category.
No. Arena results remain a separate preference or task-outcome signal.
24 model results are currently shown.