跳转到内容

编排:Llama-4

来自封神榜 Wiki

主观编排llama-4curation/models/llama-4.yml

这不是客观事实。 本页记录本站的梯队、星级、六维和情绪等主观判断,与模型事实页物理分离,仅编排员和管理员可修改。

<syntaxhighlight lang="yaml"> id: llama-4 tier: T3 stars: 3 composite: 58.4 stats:

 code: 25
 reasoning: 45
 context: 35
 speed: 70
 multimodal: 88
 value: 50

sentiment:

 positive: 10
 mixed: 20
 negative: 70

radar: - axis: 长程任务

 value: 30

- axis: 编程工程

 value: 25

- axis: 抽象推理

 value: 45

- axis: 上下文利用

 value: 35

- axis: 中文能力

 value: 40

- axis: 响应速度

 value: 70

- axis: 稳定性

 value: 45

- axis: 指令遵循

 value: 55

- axis: 易用性

 value: 60

- axis: 性价比

 value: 50

danmaku: - text: 强烈建议不要用 Llama 4 写代码

 platform: reddit
 main: true

- text: 实测拉了,开源开了个寂寞

 platform: v2ex
 main: true

- text: 开源模型的全面倒退

 platform: zhihu
 main: true

- text: Results were fudged

 platform: x
 main: true

- text: 发布帖 1235 分:The Llama 4 herd

 platform: hn
 main: true

- text: 员工辞职爆料帖 984 赞:训练把测试集混进后训练

 platform: reddit
 main: true

- text: 有长度没质量:128K 仅 15.6% 准确率

 platform: zhihu
 main: false

- text: Scout 10M 上下文创开源纪录

 platform: reddit
 main: false

- text: 单卡 4090 跑通 400B,45+ tok/s

 platform: reddit
 main: false

- text: Llama 4 Smells Bad

 platform: hn
 main: false

- text: Maverick 实测约等于 GPT-4o-0806 旧版

 platform: reddit
 main: false

- text: Cerebras 实测 2,500 tok/s(400B)

 platform: hn
 main: false

- text: '特调版 #2 vs 开源版 #32,落差 30 名'

 platform: zhihu
 main: false

- text: 聊天版太兴奋,你好都回几千 token

 platform: v2ex
 main: false

</syntaxhighlight>