iFANN
    搜索 iFANN...
    登录
    首页
    新闻
    视频
    图片
    GIF
    探索
    投票
    大奖
    iFAMOUS
    维基
    动漫
    聊天室
    通知
    私信
    收藏
    我的
    维基大奖iFAMOUS排行榜行业创作者奖励用户奖励条款隐私社区准则下架 / DMCA帮助开发者

    © 2026 iFANN

    首页
    搜索
    私信
    提醒
    我的

    帖子

    Nate
    Nate@nate_512
    📱Kimi K3📱GitHub💭AI

    Kimi K3 runs on 8GB RAM

    WAIT... YOU CAN NOW RUN A TWO POINT SEVEN TRILLION PARAMETER MODEL ON AN 8GB LAPTOP 🤯 The Kimi K3 engine achieves this with a 176KB binary written in portable C. It bypasses memory limits by streaming the 1.56TB weights directly from disk for every single token. → 8GB RAM gets you 26 seconds per token → 128GB RAM gets you 5 seconds per token Same exact math and byte identical results regardless of the machine. Zero GPUs required. This is pure brutalist engineering. Free and open-source. repo in 🧵↓

    9h

    6 赞0 踩1 转发1 评论
    ?

    评论

    还没有评论。来抢沙发吧!

    帖子

    Nate
    Nate@nate_512
    📱Kimi K3📱GitHub💭AI

    Kimi K3 runs on 8GB RAM

    WAIT... YOU CAN NOW RUN A TWO POINT SEVEN TRILLION PARAMETER MODEL ON AN 8GB LAPTOP 🤯 The Kimi K3 engine achieves this with a 176KB binary written in portable C. It bypasses memory limits by streaming the 1.56TB weights directly from disk for every single token. → 8GB RAM gets you 26 seconds per token → 128GB RAM gets you 5 seconds per token Same exact math and byte identical results regardless of the machine. Zero GPUs required. This is pure brutalist engineering. Free and open-source. repo in 🧵↓

    9h

    6 赞0 踩1 转发1 评论
    ?

    评论

    还没有评论。来抢沙发吧!