iFANN
    搜索 iFANN...
    登录
    首页
    新闻
    视频
    图片
    GIF
    探索
    投票
    大奖
    iFAMOUS
    维基
    动漫
    聊天室
    通知
    私信
    收藏
    我的
    维基大奖iFAMOUS排行榜行业创作者奖励用户奖励条款隐私社区准则下架 / DMCA帮助开发者

    © 2026 iFANN

    首页
    搜索
    私信
    提醒
    我的
    照片
    Hype Castle27
    Hype Castle27@hype_castle273mo
    💭AI💭Tech
    GLM 5.2 ranks #1 on BridgeBench reasoning

    @hype_castle27The Chinese open-source model GLM 5.2 has claimed the top spot on the BridgeBench reasoning leaderboard, leaving Kimi K2.7 Code far behind. According to the rankings, GLM 5.2 sits at #1 with a score of 42.8, while Kimi K2.7 Code ranks only #11 at 40.1. What's more, Kimi K2.7 actually saw a regression from its predecessor Kimi K2.6 on this same reasoning benchmark. The leaderboard, which includes models like Nemotron 3 Ultra 550B-A55B and Claude Fable 5, is based on 30 tasks in a hard benchmark f #AI

    查看原帖

    GLM 5.2 ranks #1 on BridgeBench reasoning

    @hype_castle27 的照片· Jun 16, 2026· AI

    关于这张照片

    A dark-mode leaderboard titled "Reasoning" shows model rankings with scores. GLM 5.2 sits at rank 1 with 42.8, highlighted by a red box and arrow. Kimi K2.7 Code is boxed at rank 11 with 40.1, also marked by a red arrow pointing down. The list includes other models like Nemotron 3 Ultra 550B-A55B and Claude Fable 5. A footer notes "30 tasks · hard benchmark · grounded reasoning over mixed artifacts."

    查看AI的全部照片

    ?

    更多AI照片

    查看AI的全部照片
    GPT 6 Astra effort comparison tableGPT 6 Astra effort comparison tableGemini 3.6 Flash vs 3.5 Flash same scoreGemini 3.6 Flash vs 3.5 Flash same scoreGPT-5.6 Sol Terra Luna illustrationGPT-5.6 Sol Terra Luna illustrationClaude Max subscription $200/monthClaude Max subscription $200/monthClaude Sonnet 5 pricing vs Opus 4.8Claude Sonnet 5 pricing vs Opus 4.8Trump administration close to restoring Anthropic Fable 5 accessTrump administration close to restoring Anthropic Fable 5 accessClaude for Government uptime 99.93%Claude for Government uptime 99.93%BridgeAgent loops demoBridgeAgent loops demoGLM 5.2 Artificial Analysis Coding IndexGLM 5.2 Artificial Analysis Coding IndexGLM Coding Pro plan $64.8GLM Coding Pro plan $64.8Claude Max subscription limitsClaude Max subscription limitsClaude Fable 5 Max Cursor BenchClaude Fable 5 Max Cursor BenchClaude 3.5 guardrailsClaude 3.5 guardrailsClaude Opus 4.7 Frontend DesignArenaClaude Opus 4.7 Frontend DesignArena
    照片
    Hype Castle27
    Hype Castle27@hype_castle273mo
    💭AI💭Tech
    GLM 5.2 ranks #1 on BridgeBench reasoning

    @hype_castle27The Chinese open-source model GLM 5.2 has claimed the top spot on the BridgeBench reasoning leaderboard, leaving Kimi K2.7 Code far behind. According to the rankings, GLM 5.2 sits at #1 with a score of 42.8, while Kimi K2.7 Code ranks only #11 at 40.1. What's more, Kimi K2.7 actually saw a regression from its predecessor Kimi K2.6 on this same reasoning benchmark. The leaderboard, which includes models like Nemotron 3 Ultra 550B-A55B and Claude Fable 5, is based on 30 tasks in a hard benchmark f #AI

    查看原帖

    GLM 5.2 ranks #1 on BridgeBench reasoning

    @hype_castle27 的照片· Jun 16, 2026· AI

    关于这张照片

    A dark-mode leaderboard titled "Reasoning" shows model rankings with scores. GLM 5.2 sits at rank 1 with 42.8, highlighted by a red box and arrow. Kimi K2.7 Code is boxed at rank 11 with 40.1, also marked by a red arrow pointing down. The list includes other models like Nemotron 3 Ultra 550B-A55B and Claude Fable 5. A footer notes "30 tasks · hard benchmark · grounded reasoning over mixed artifacts."

    查看AI的全部照片

    ?

    更多AI照片

    查看AI的全部照片
    GPT 6 Astra effort comparison tableGPT 6 Astra effort comparison tableGemini 3.6 Flash vs 3.5 Flash same scoreGemini 3.6 Flash vs 3.5 Flash same scoreGPT-5.6 Sol Terra Luna illustrationGPT-5.6 Sol Terra Luna illustrationClaude Max subscription $200/monthClaude Max subscription $200/monthClaude Sonnet 5 pricing vs Opus 4.8Claude Sonnet 5 pricing vs Opus 4.8Trump administration close to restoring Anthropic Fable 5 accessTrump administration close to restoring Anthropic Fable 5 accessClaude for Government uptime 99.93%Claude for Government uptime 99.93%BridgeAgent loops demoBridgeAgent loops demoGLM 5.2 Artificial Analysis Coding IndexGLM 5.2 Artificial Analysis Coding IndexGLM Coding Pro plan $64.8GLM Coding Pro plan $64.8Claude Max subscription limitsClaude Max subscription limitsClaude Fable 5 Max Cursor BenchClaude Fable 5 Max Cursor BenchClaude 3.5 guardrailsClaude 3.5 guardrailsClaude Opus 4.7 Frontend DesignArenaClaude Opus 4.7 Frontend DesignArena