【帖子标题】:Qwen 3.8 27B is a game changer. 【帖子标题】:Qwen 3.8 27B 是一款颠覆性产品。
【帖子正文】: Our devs got their hands on it a few days ago. One wired it into Codex to compare with GPT Luna, our usual workhorse right now for its cost effectiveness. Another tried it out on one of our OCR pipelines. 我们的开发人员几天前拿到了它。有人把它接入 Codex,与 GPT Luna 进行对比——GPT Luna 是我们目前因成本效益而常用的主力模型。另一个人在我们的一条 OCR 流水线上试用了它。
It’s comparable to Luna for coding and OCR quality appears to be better than Gemini 3.5 Flash Lite. That’s huge. We pay a ton of money for OCR. 它在编码方面与 Luna 相当,而且 OCR 质量似乎优于 Gemini 3.5 Flash Lite。这太重要了。我们在 OCR 上花了大量资金。
This is the first local model that feels like more than a toy. It’s truly as capable as the frontier models from a year ago. For the first time ever there’s serious discussions about buying our own hardware. With estimates that such an effort would pay for itself in less than 2 months. 这是第一个让人觉得不只是玩具的本地模型。它真的具备一年前前沿模型的能力。有史以来第一次,我们认真讨论购买自己的硬件。据估计,这样的投入不到两个月就能回本。
Hyper scalars are in big trouble this time. Their whole “moat” is buying up all the hardware. And thanks to sanctions on China we’re seeing the quality of small local models skyrocket. As someone who’s been around a while, this feels like an “IBM moment”. Where the industry assumed that databases would always run on huge mainframes. Only to be wiped out by cheaper local solutions a few years later. 超大规模云服务商这次有大麻烦了。他们的整个“护城河”就是买下所有硬件。而且,由于对中国的制裁,我们看到小型本地模型的质量正在飙升。作为一个在这个行业待了一段时间的人,这感觉就像一个“IBM 时刻”。当时业界以为数据库永远会运行在大型主机上,结果几年后就被更便宜的本地解决方案淘汰了。
I have a feeling this release will trigger another Llama style open source Renaissance. We’re already getting better quants. Inference will be further improved. We might even see a comparable MoE with 500+ Tok/sec on consumer hardware soon. 我有一种感觉,这次发布将引发另一场 Llama 风格的开源复兴。我们已经得到了更好的量化版本。推理性能还将进一步提升。我们甚至可能很快在消费级硬件上看到每秒 500+ Token 的同类 MoE 模型。