Trinity Large 登场:400B 稀疏 MoE 模型,宣称超越 Llama 4
Trinity 团队发布 4000 亿参数稀疏 MoE 模型 Trinity Large,采用 4-of-256 架构,仅激活 13B 参数,推理速度提升 2-3 倍。该模型提供 Base、Preview 和 TrueBase 三个版本,其...
标签索引
这个标签下有 1 篇文章。按时间回看相关判断与实践记录。
标签精选
Trinity 团队发布 4000 亿参数稀疏 MoE 模型 Trinity Large,采用 4-of-256 架构,仅激活 13B 参数,推理速度提升 2-3 倍。该模型提供 Base、Preview 和 TrueBase 三个版本,其...