iFANN
    搜索 iFANN...
    登录
    首页
    新闻
    视频
    图片
    GIF
    探索
    投票
    大奖
    iFAMOUS
    维基
    动漫
    聊天室
    通知
    私信
    收藏
    我的
    维基大奖iFAMOUS排行榜行业创作者奖励用户奖励条款隐私社区准则下架 / DMCA帮助开发者

    © 2026 iFANN

    首页
    搜索
    私信
    提醒
    我的
    照片
    Estebankiwi
    Estebankiwi@estebankiwi3mo
    📱GPT💭AI💭Tech
    Claude Fable 5 DeepSWE benchmark

    @estebankiwiClaude Fable 5 has claimed the top spot on DeepSWE with a score of 70%. However, the performance gap between Fable 5 and GPT 5.5 is far more significant than just three percentage points. Fable 5 generates code that reads as if it were written by a senior engineer, while GPT 5.5 produces code that simply passes the tests. Both models deliver functional software, but only one delivers software that truly impresses.

    查看原帖

    Claude Fable 5 DeepSWE benchmark

    @estebankiwi 的照片· Jun 19, 2026· GPT

    关于这张照片

    The image shows a horizontal bar chart comparing AI models. The focus is on a table with model names, performance metrics, and bars representing their performance. The model "claude-fable-5" is highlighted with a red rectangle around it and its corresponding bar is orange. The table columns are labeled "MODEL", "PASS@1", "AVG COST", "OUT TOK", and "STEPS". The model names listed are claude-fable-5, gpt-5.5, claude-opus-4.8, gpt-5.4, gemini-3.5-flash, and kimi-k2.7-code. No on-screen text stands out aside from the column headers and model names.

    查看GPT的全部照片阅读GPT维基

    ?

    更多GPT照片

    查看GPT的全部照片
    AI as your doctor?AI as your doctor?GPT 6 Astra effort comparison tableGPT 6 Astra effort comparison tablemanual coding psychopathmanual coding psychopathJensen Huang AGI has arrived GPT-6 Astra2Jensen Huang AGI has arrived GPT-6 AstraAI Breakfast greatest predictionAI Breakfast greatest predictionweekend plans cancelledweekend plans cancelledAI Image Models ComparisonAI Image Models Comparison
    照片
    Estebankiwi
    Estebankiwi@estebankiwi3mo
    📱GPT💭AI💭Tech
    Claude Fable 5 DeepSWE benchmark

    @estebankiwiClaude Fable 5 has claimed the top spot on DeepSWE with a score of 70%. However, the performance gap between Fable 5 and GPT 5.5 is far more significant than just three percentage points. Fable 5 generates code that reads as if it were written by a senior engineer, while GPT 5.5 produces code that simply passes the tests. Both models deliver functional software, but only one delivers software that truly impresses.

    查看原帖

    Claude Fable 5 DeepSWE benchmark

    @estebankiwi 的照片· Jun 19, 2026· GPT

    关于这张照片

    The image shows a horizontal bar chart comparing AI models. The focus is on a table with model names, performance metrics, and bars representing their performance. The model "claude-fable-5" is highlighted with a red rectangle around it and its corresponding bar is orange. The table columns are labeled "MODEL", "PASS@1", "AVG COST", "OUT TOK", and "STEPS". The model names listed are claude-fable-5, gpt-5.5, claude-opus-4.8, gpt-5.4, gemini-3.5-flash, and kimi-k2.7-code. No on-screen text stands out aside from the column headers and model names.

    查看GPT的全部照片阅读GPT维基

    ?

    更多GPT照片

    查看GPT的全部照片
    AI as your doctor?AI as your doctor?GPT 6 Astra effort comparison tableGPT 6 Astra effort comparison tablemanual coding psychopathmanual coding psychopathJensen Huang AGI has arrived GPT-6 Astra2Jensen Huang AGI has arrived GPT-6 AstraAI Breakfast greatest predictionAI Breakfast greatest predictionweekend plans cancelledweekend plans cancelledAI Image Models ComparisonAI Image Models Comparison