PulseAugur
中
实时 22:40:40
한국어(KO) Code Arena(웹Dev) 리더보드(2026-06-20)는 프런트엔드 웹개발·에이전트 코딩 워크플로 중심으로 90개 모델을 평가(391,241표). 상위권은 Anthropic의 claude-fable-5, Z.ai의 glm-5.2, 여러 claude-opus 계열과 OpenAI의 g

AI 模型在 Code Arena 排行榜上接受 Web 开发和代理编码评估

Code Arena Web 开发和代理编码工作流排行榜已根据 391,241 票评估了 90 个模型。表现最佳的模型包括 Anthropic 的 Claude Fable-5、智谱 AI 的 GLM-5.2、多个 Claude Opus 模型以及 OpenAI 的 GPT-5.5。该排行榜提供了关于 Elo 评分、投票数和每代币成本的比较数据,以评估代理 AI 的性能。 AI

影响 为 Web 开发和代理编码任务中的各种 AI 模型性能提供了见解,影响了未来的模型开发和采用。

排序理由 这是 AI 模型的研究基准测试结果。[lever_c_demoted from research: ic=1 ai=1.0]

在 Mastodon — sigmoid.social 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

AI 模型在 Code Arena 排行榜上接受 Web 开发和代理编码评估

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
这是 AI 模型的研究基准测试结果。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
101 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. Mastodon — sigmoid.social TIER_1 한국어(KO) · [email protected] ·

    Code Arena (WebDev) 排行榜 (2026-06-20) 评估了 90 个模型(391,241 票),重点关注前端 Web 开发和代理编码工作流。表现最佳的模型包括 Anthropic 的 claude-fable-5、Z.ai 的 glm-5.2、多个 claude-opus 变体以及 OpenAI 的 g

    Code Arena(웹Dev) 리더보드(2026-06-20)는 프런트엔드 웹개발·에이전트 코딩 워크플로 중심으로 90개 모델을 평가(391,241표). 상위권은 Anthropic의 claude-fable-5, Z.ai의 glm-5.2, 여러 claude-opus 계열과 OpenAI의 gpt-5.5 등이 포진. 모델별 Elo 성적, 득표수, 토큰당 가격 등 비교 정보를 제공해 에이전트형 AI 성능 벤치마크를 보여줌. https:// arena.ai/leaderboard/code/webd ev # l…