时间:2026年9月18日
地点:美国加州山景城/LMArena在线基准平台
人物:Google DeepMind工程师团队、X平台研究员 Lentils(@Lentils80)、MaaSonder(@MaaSonder)、ai_for_success、IT之家网友 LuminaBench 等社区测试者
事件详情:Google正以"gemini-3.8-flash"为伪装代号,在LMArena等基准平台悄悄测试其下一代旗舰模型 Gemini 4 Pro,内部开发代号为"Argon"。9月14日,Lentils率先在X平台曝光Argon的内部checkpoint生成截图,显示该模型"High thinking effort"模式生成耗时约2.4分钟,输出Token上限提升至256K(此前Gemini系列仅64K)。社区测试者随后发现Argon在生成经典"鹈鹕骑自行车"SVG图像时表现惊艳,与OpenAI最新旗舰GPT-6 Astra Pro的对比中获得多数测试者认可。MaaSonder进一步"Ghost"接入Argon后端,对"Infinite JSON"压测和Voxal Pagoda复杂渲染均得到高质量输出。
背景:这是Google继"lithiumflow""orionmist""Riftrunner"等代号后,又一次以马甲号在LMArena上测试未发布模型,已成为Google新品发布的"传统流程"。Gemini 4 Pro跳过3.5 Pro直接命名,且据Cometapi分析,Google可能瞄准1.5M+的上下文窗口。fortune.com此前也指出,Google在106天内连发四款Gemini Flash模型,但旗舰"前沿模型"始终缺席。
影响:1)OpenAI的GPT-6 Astra Pro首次在SVG视觉生成领域遭遇直接对比败绩,AI头部竞争进入"视觉深度"新阶段;2)256K输出上限将极大扩展AI在长文档生成、复杂SVG/前端代码、视频脚本等场景的可用性;3)LMArena"马甲号"测试玩法再获验证,可能引发Anthropic、xAI等更多厂商效仿。
总结:Google以"Argon"代号曝光旗舰Gemini 4 Pro,SVG生图惊艳、256K输出翻倍,标志着头部AI竞争从纯文本转向视觉生成深度。
参考来源:
1. IT之家 - 谷歌最强AI模型Gemini 4 Pro开测:https://www.ithome.com/1/003/902.htm
2. 智源社区 - 谷歌新版Gemini马甲被扒 LMArena实测:https://hub.baai.ac.cn/view/49654
3. X平台 Lentils首曝Argon:https://x.com/Lentils80/status/2099601296516960580
4. X平台 ai_for_success转评:https://x.com/ai_for_success/status/2099675906189230395
5. X平台 MaaSonder Ghost test 256K压测:https://x.com/MaaSonder/status/2100548822753529944
6. X平台 MaaSonder Voxal Pagoda渲染:https://x.com/MaaSonder/status/2100582544723095942
7. YouTube - Gemini 4 Pro ARGON Checkpoint Leaked:https://www.youtube.com/watch?v=O71tzdngeVY
8. Cometapi - Gemini 4 Pro规格分析:https://www.cometapi.com/gemini-4-pro-is-coming/
9. kie.ai博客 - What Is Gemini 4 Pro:https://kie.ai/blog/what-is-gemini-4-pro
10. The Register - Gemini 3.8 Flash发布报道:https://www.theregister.com/ai-and-ml/2026/09/02/with-gemini-38-flash-google-reminds-everyone-its-still-in-the-race/5294049







