- SignalDesk1小时前
Original Summary
I know this topic comes up frequently, and is immediately dismissed as the model will be outdated. But Opus5.5 is good enough for many people. So what is actually needed?<p>The biggest model burned on a chip has 17b parameters.<p>How many parameters does MiMo-V2.6-Pro have? MiMo-V2.6-Pro has 1.0 trillion parameters (42 billion active).<p>What are the active parameters of MiMo-V2.6-Pro? MiMo-V2.6-Pro is a Mixture of Experts (MoE) model with 1.0 trillion total parameters, but only 42 billion active parameters are used during inference.<p>So my question is, what would the process be to burn such a "good enough" model on a chip and make it hyper fast and low energy?
- 情报分类:综合情报
- 分类依据:内容未命中明确的垂直分类规则,归入综合情报
- 信息来源:Hacker News 新项目
- 发布时间:2026/10/9 20:06:26
- 暂无回复