跳转到内容

编排:Gemini-3-pro

来自封神榜 Wiki

主观编排gemini-3-procuration/models/gemini-3-pro.yml

这不是客观事实。 本页记录本站的梯队、星级、六维和情绪等主观判断,与模型事实页物理分离,仅编排员和管理员可修改。

<syntaxhighlight lang="yaml"> id: gemini-3-pro tier: T0 stars: 6 composite: 93.5 stats:

 code: 92
 reasoning: 88
 context: 98
 speed: 72
 multimodal: 96
 value: 72

sentiment:

 positive: 50
 mixed: 25
 negative: 25

radar: - axis: 长程任务

 value: 82

- axis: 编程工程

 value: 86

- axis: 抽象推理

 value: 90

- axis: 上下文利用

 value: 80

- axis: 中文能力

 value: 78

- axis: 响应速度

 value: 74

- axis: 稳定性

 value: 60

- axis: 指令遵循

 value: 66

- axis: 易用性

 value: 75

- axis: 性价比

 value: 68

danmaku: - text: 'LMArena 1501 首破 1500 · 全榜 #1'

 platform: x
 main: true

- text: 下一个时代的大模型灯塔

 platform: zhihu
 main: true

- text: 发布帖 1735 points · 技术突破惊人

 platform: hn
 main: true

- text: After seeing the benchmark numbers I had to give it a try

 platform: reddit
 main: true

- text: High IQ, High Ego:准确率 55.9% 但幻觉率 88%

 platform: reddit
 main: true

- text: Karpathy:clearly a tier 1 LLM · daily driver 潜力

 platform: x
 main: true

- text: becoming really lazy · 每问题仅搜索 3 次

 platform: reddit
 main: false

- text: Kilo Code 独立测试 72% vs 54% vs 18%

 platform: x
 main: false

- text: Deep Think 模式上线 · HN 1081 分

 platform: hn
 main: false

- text: 前 Google 员工:最令人沮丧的开发模型

 platform: hn
 main: false

- text: MathArena Apex 23.4% · 20x 跳跃

 platform: x
 main: false

- text: 实际上下文 ~32K(Web App 实测)

 platform: hn
 main: false

- text: 幻觉比 GPT-5.1 严重得多,一逗就出来

 platform: zhihu
 main: false

- text: 门门 90+ 的全能状元

 platform: zhihu
 main: false

- text: 唯一真神 · 仿 Windows Web OS 单 prompt

 platform: zhihu
 main: false

- text: vast intelligence with no spine

 platform: reddit
 main: false

</syntaxhighlight>