MiniMax 发布 v3 模型,附带 MSA 内核,序列越长加速越明显,已集成至 transformers 和 Kernel Hub
多源视角
Nice demo from @atomic_chat_hq: <br><br>M3 Q4 ran locally with MLX-VLM, and completed a US customs form entirely on a Mac Studio M3 Ultra.<hr style="border:0;border-top:1px solid #…
查看这条来源M3's free on @0G_labs 0G compute for three days, starting today. <br><br>Amazing chance to throw a long-running task at it 👇<hr style="border:0;border-top:1px solid #80808030;marg…
查看这条来源RT Cast AI<br>Re @MiniMax_AI M3 lands on @getkimchi Coding by Cast AI, the first autonomous coding agent to launch it. Frontier open-weight performance on Kimchi CLI.<br><br>Read m…
查看这条来源AI 摘要
MiniMax 发布了 v3 模型,该模型集成了一个 MSA 内核。随着序列长度的增加,该内核能实现显著的加速效果。该内核已集成到 Hugging Face 的 transformers 库和 Kernel Hub 中,方便开发者直接使用。 核心观点: 1. MiniMax v3 模型集成的 MSA 内核,序列越长加速效果越明显。 2. MSA 内核已集成至 Hugging Face transformers 库和 Kernel Hub。
推荐理由高信息密度,值得细读
原文
无法获取正文,可打开下方原始链接查看。
讨论
暂无评论。