Hacker News 中文摘要

RSS订阅

Kimi K3现已上线 -- Kimi K3 is now live

文章摘要

Kimi AI推出K3版本,专注于智能编程与知识工作,支持多种任务如群组协作、幻灯片制作、深度研究、网站和文档处理等。

文章总结

Kimi AI 推出 K3 版本,专为智能编程与知识工作设计。用户可向 AI 提问或指派任务,支持 K3 和 Max 两种模式。功能包括:群体智能(Swarm)、幻灯片制作(Slides)、深度研究(Deep Research)、网站分析(Websites)、文档处理(Docs)和表格处理(Sheets)。页面还提供灵感探索入口。

评论总结

根据评论内容,总结如下:

主要观点与论据:

  1. 模型性能与定位:Kimi K3 被描述为“开源模型中的顶尖水平”,在多项基准测试中仅次于 Claude Fable 5 和 GPT-5.6 Sol。评论者认为其“接近前沿模型”,并可能引发“另一个 DeepSeek 时刻”。

    • 关键引用:“Among the models tested, its overall intelligence ranks second only to Claude Fable 5 and GPT-5.6 Sol.”(评论 3)
    • 关键引用:“Amazing to see an open source model already nearing the benchmarks of Fable and GPT 5.6 Sol!”(评论 8)
  2. 定价与成本效率:定价为 $3/$15 每百万 token(缓存 $0.3),与 Anthropic Sonnet 系列持平,但评论者指出推理效率是关键:若 Kimi K3 消耗更多推理 token,实际成本可能更高。同时,有评论认为其推理 token 使用量比 K2.6 少约 60%,缓解了价格担忧。

    • 关键引用:“pricing is $3/$15 for 1M tokens... extremely high for a Chinese open-weight model, but if it's truly competitive... the pricing is justified.”(评论 1)
    • 关键引用:“Seems to only use ≈60% as many reasoning tokens as 2.6. So the price hike is not as bad as it looks.”(评论 13)
  3. 技术架构与开源:模型采用 MoE 架构,激活 16/896 专家,总参数 2.8T,活跃参数约 50B。评论者赞赏其开源承诺,但担忧中国模型提供商“逐渐封闭”(如 Qwen 3.6 未发布最大参数版本)。

    • 关键引用:“the model efficiently activates 16 out of 896 experts... 2.5x the overall scaling efficiency of K2”(评论 7)
    • 关键引用:“Now, will they actually release the weights? Seems like Chinese model providers are slowly closing up”(评论 16)
  4. 实际体验与质疑:部分评论者测试后认为模型“思考过多,重复内容”,且结合价格可能不如 GPT-5.6 划算。另有评论质疑其是否主要依赖蒸馏技术。

    • 关键引用:“I have just tried a single prompt... it is still not finish reasoning. It thinks too much and often repeats the same stuff over and over.”(评论 20)
    • 关键引用:“I'm curious if they're keeping up mostly due to distillation”(评论 18)
  5. 竞争与行业影响:评论者认为中国模型(如 Kimi K3)创造了竞争和紧迫感,推动行业进步。同时,有评论期待 DeepSeek 本周发布,并希望其进一步接近 SOTA。

    • 关键引用:“Say what you want about these Chinese models but they sure create competition and urgency in the space.”(评论 12)
    • 关键引用:“Excited for the deepseek release this week... Hopefully they also push even closer to SOTA.”(评论 4)

平衡性总结:评论对 Kimi K3 的性能和开源贡献持积极态度,但对其定价、推理效率、实际体验及中国模型的开源趋势存在质疑。整体上,模型被视为有竞争力的开源选择,但需进一步验证其实际成本与能力。