iFANN
    搜索 iFANN...
    登录
    首页
    新闻
    视频
    图片
    GIF
    探索
    投票
    大奖
    iFAMOUS
    维基
    动漫
    聊天室
    通知
    私信
    收藏
    我的
    维基大奖iFAMOUS排行榜行业创作者奖励用户奖励条款隐私社区准则下架 / DMCA帮助开发者

    © 2026 iFANN

    首页
    搜索
    私信
    提醒
    我的

    帖子

    Hype Castle27
    Hype Castle27@hype_castle27
    💭AI💭Tech

    GLM 5.2 ranks #1 on BridgeBench reasoning

    The Chinese open-source model GLM 5.2 has claimed the top spot on the BridgeBench reasoning leaderboard, leaving Kimi K2.7 Code far behind. According to the rankings, GLM 5.2 sits at #1 with a score of 42.8, while Kimi K2.7 Code ranks only #11 at 40.1. What's more, Kimi K2.7 actually saw a regression from its predecessor Kimi K2.6 on this same reasoning benchmark. The leaderboard, which includes models like Nemotron 3 Ultra 550B-A55B and Claude Fable 5, is based on 30 tasks in a hard benchmark f #AI

    2mo

    71 赞2 踩2 转发2 评论
    ?

    评论

    还没有评论。来抢沙发吧!

    帖子

    Hype Castle27
    Hype Castle27@hype_castle27
    💭AI💭Tech

    GLM 5.2 ranks #1 on BridgeBench reasoning

    The Chinese open-source model GLM 5.2 has claimed the top spot on the BridgeBench reasoning leaderboard, leaving Kimi K2.7 Code far behind. According to the rankings, GLM 5.2 sits at #1 with a score of 42.8, while Kimi K2.7 Code ranks only #11 at 40.1. What's more, Kimi K2.7 actually saw a regression from its predecessor Kimi K2.6 on this same reasoning benchmark. The leaderboard, which includes models like Nemotron 3 Ultra 550B-A55B and Claude Fable 5, is based on 30 tasks in a hard benchmark f #AI

    2mo

    71 赞2 踩2 转发2 评论
    ?

    评论

    还没有评论。来抢沙发吧!