Anthropic's Claude models lead LiveBench coding evaluations due to top scores in recent benchmark refreshes emphasizing contamination-resistant tasks like code generation and agentic problem-solving. Their latest releases demonstrate superior performance in handling complex repositories, execution validation, and multi-step debugging compared to competitors. This drives the near-certain market consensus around Anthropic. Realistic challenges include a major new model launch from OpenAI or Google before resolution, or the next LiveBench question rotation that could shift category rankings if other labs close the gap on benchmarks.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · ОбновленоWhich company has the best AI model on LiveBench (Coding) end of August?
Anthropic 100.0%
Alibaba <1%
OpenAI <1%
Z.ai <1%
$51,410 Объем
$51,410 Объем

Alibaba
No

OpenAI
No

Z.ai
No

Xiaomi
No

SpaceXAI
No

ByteDance
No

Nvidia
No

Moonshot
No

Baidu
No

DeepSeek
No

Amazon
No

MiniMax
No

StepFun
No

Thinky
No

Mistral
No

Meituan
No

Tencent
No

Microsoft
No

Anthropic
Yes

Meta
No

No
Anthropic 100.0%
Alibaba <1%
OpenAI <1%
Z.ai <1%
$51,410 Объем
$51,410 Объем

Alibaba
No

OpenAI
No

Z.ai
No

Xiaomi
No

SpaceXAI
No

ByteDance
No

Nvidia
No

Moonshot
No

Baidu
No

DeepSeek
No

Amazon
No

MiniMax
No

StepFun
No

Thinky
No

Mistral
No

Meituan
No

Tencent
No

Microsoft
No

Anthropic
Yes

Meta
No

No
Results from the “Coding” column of the leaderboard at https://livebench.ai/#/?cats=Coding, with the latest available LiveBench release selected and the category set to “Coding,” will be used to resolve this market.
Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
Открытие рынка: Jul 29, 2026, 6:27 PM ET
Источник определения исхода
https://livebench.ai/#/?cats=CodingКто определяет исход
0x69c47De9D...Предложенный исход: Yes
Спор отсутствует
Окончательный исход: Yes
Results from the “Coding” column of the leaderboard at https://livebench.ai/#/?cats=Coding, with the latest available LiveBench release selected and the category set to “Coding,” will be used to resolve this market.
Models will be ranked according to the specified score, with higher scores ranked ahead of lower scores. If two or more models have exactly the same score as displayed on the leaderboard, the model with the lower listed "cost per successful task" will be ranked ahead. If a tie still remains, alphabetical order of company names as listed in this market group will be used as a final tiebreaker (e.g., if the two models are tied by exact score and cost per successful task, “Google” would be ranked ahead of “SpaceXAI”). This market will resolve based on the company that occupies first place under this ranking.
The resolution source for this market is the LiveBench leaderboard. If this resolution source is unavailable at check time, this market will remain open until the leaderboard comes back online and will resolve based on the first check after it becomes available. If it becomes permanently unavailable, this market will resolve to “Other.”
Источник определения исхода
https://livebench.ai/#/?cats=CodingКто определяет исход
0x69c47De9D...Предложенный исход: Yes
Спор отсутствует
Окончательный исход: Yes
Anthropic's Claude models lead LiveBench coding evaluations due to top scores in recent benchmark refreshes emphasizing contamination-resistant tasks like code generation and agentic problem-solving. Their latest releases demonstrate superior performance in handling complex repositories, execution validation, and multi-step debugging compared to competitors. This drives the near-certain market consensus around Anthropic. Realistic challenges include a major new model launch from OpenAI or Google before resolution, or the next LiveBench question rotation that could shift category rankings if other labs close the gap on benchmarks.
Экспериментальная сводка, созданная ИИ на основе данных Polymarket. Это не является торговой рекомендацией и не влияет на то, как разрешается этот рынок. · Обновлено
Не доверяй внешним ссылкам.
Не доверяй внешним ссылкам.
Часто задаваемые вопросы