以下部分引用 AI 总结,现阶段 AI 仍然有幻觉。请以内容的原文为准。
数据新鲜度群聊 07-14 ✓卖方 07-14 ✓Wrap 断档15天 ✕News 07-15 ✓X 断档12天 ✕页面生成 07-15
GPT-5.4
44 atoms · 跨 27 天 · 首见 2026-04-10 · 最近 2026-07-02
三色: 🟦 fact 33 · 🟥 take 11 · stance ▲8/▼7/◆7
叙事状态: 🍂 fading (退潮) · 💤 沉寂 · 策展近14d 0 atom · 🐦 X 近14d 1
状态只算低频策展源 (群/卖方); X firehose 仅作背景音量
来源: X 43 · 卖 1
时态: fresh:36 · aging:7 · stale:1
标签: 好数字:31 · 好观点:9 · 好信源:8 · 好思考:3
🟨 AI 综合 · junior analyst 概览
展开 AI 综合 (灰色 · 非市场结论 · 点击数字溯源到原 atom)
model: deepseek-chat · 2026-07-12 · 默认折叠
GPT-5.4 在定价上仍有明显优势,AI output 每百万 tokens 约 $14,且比 Gemini 3.5 Flash 便宜三倍 [id=6b96c50165ed7d29,cb8feeb11587b79f]。但用户质疑若模型更小为何收费更贵 [id=d5a04c1ef1b43b6b]。科学贡献方面,GPT-5.4-pro 为长期数学难题给出新证明,且该证明思路可用于多个其他问题 [id=8e1bd77f8b51ef10,90491926e7142b02]。然而其预训练仍是基于 4o/t 而非新范式如 5.5,被批评为旧架构延续 [id=7e70211c87fd623c]。基础性能与 GLM-5.2 无推理时相似,但开启推理后差距拉大 [id=9451df4d89be5164]。Claw-Eval-Live 基准通过率 63.8%,全套模型平均完成时间最快 105s [id=21294708baafc2bf,a495c285d290f561]。不过选择不当的 prompt/agent 框架会导致过早停止 [id=1ee9d455c0ba69ac]。协作与信任度方面,用户反馈可交付困难任务并信任其修复结果 [id=9683cc49215936c3]。
🧭 拥挤度 (一人一票): ▲ 5 位作者 (KOL5) vs ▼ 4 位作者 (KOL4)
⚖️ 多头 5/5 来自KOL
📄 广泛报道的事实 · 1 个
展开 (多人转述同一事实/数字 — 确认度高, **非独立观点**, 不标多空)
💢 核心分歧 (2 轴)
1. 定价能力:成本优势 vs 高价质疑
*topic: 定价权/价格动态 · 1 bull vs 2 bear*
🟢 bullish 侧:
- <a id="atom-64deadfeb4698e34"></a>2026-04-19
X·@theo · 成本对比 · 定价权/技术路线
GPT 5.4 is still way cheaper
tags: 好数字·好观点 · → daily
value: qty=way cheaper · date=2026-04-19
📷 原图
🔴 bearish 侧:
- <a id="atom-d5a04c1ef1b43b6b"></a>2026-04-18
X·@Samhanknr [老化中] · 模型成本定价 · 定价权
If it’s smaller why is it more expensive 🤔
tags: 好思考 · → daily
value: qty=more expensive
- <a id="atom-cb8feeb11587b79f"></a>2026-06-03
X·@scaling01 · 定价 · 价格动态
Gemini 3.5 Flash was literally 3x more expensive than GPT-5.4
tags: 好数字 · → daily
value: qty=3x less expensive · date=2026-06-03
2. 科学贡献:创造新知识 vs 基础性能欠缺
*topic: 技术路线/模型发布/模型能力/模型性能/推理能力 · 5 bull vs 3 bear*
🟢 bullish 侧:
- <a id="atom-9683cc49215936c3"></a>2026-04-16
X·@max_paperclips [老化中] · 可以信任完成复杂任务 · 模型能力
with GPT-5.4 I can just give it a difficult task and trust more or less that if it think it's fixed, it's probably fixed or at least did do something rational to fix it
tags: 好观点 · → daily
- <a id="atom-8e1bd77f8b51ef10"></a>2026-05-02
X·@boazbaraktcs [老化中] · produced new proof for longstanding math problem · 模型发布/技术路线 +同日1条
GPT-5.4-pro came up with a new proof for one longstanding problem
tags: 好观点 · → daily
- <a id="atom-d7dee26a387936c6"></a>2026-06-17
X·@OpenAI · 提出改进药物发现中广泛使用反应的方法 · 技术路线/产品发布 +同日1条
the model proposed an unexpected way to improve a widely used reaction in drug discovery.
tags: 好观点 · → daily
value: qty= · date=
🔴 bearish 侧:
- <a id="atom-8d323c59da8a1e4a"></a>2026-05-04
X·@bioshok3 ⭐ [老化中] · 实验辅助能力超过博士级病毒学家 · 技术路线/地缘政治
GPT-5.4は博士レベルのウイルス学者を上回る実験支援能力を持つとされ、バイオリスクも深刻化
tags: 好观点·好思考 · → daily
📷 原图
- <a id="atom-7e70211c87fd623c"></a>2026-06-22
X·@pigeon_s · 预训练基础 · 技术路线_
if its based on 5.4 which is still based on 4o/t and not the new pretrain like 5.5 then its still FUCKING GPT-4o VOICE
tags: 好观点 · → daily
- <a id="atom-9451df4d89be5164"></a>2026-07-02
X·@scaling01 ⭐ · 基础性能与 GLM-5.2 相似但推理后差距扩大 · 技术路线/推理能力
GPT-5.4 and GLM-5.2 have ~the same scores without reasoning, but GLM gets crushed once you turn on reasoning
tags: 好数字·好信源 · → daily
📷 原图
展开 3 条中性
- <a id="atom-70b7f7a44be20a19"></a>2026-04-19
X·@btibor91 · 模型发布 · 模型发布
Cloudflare Agent Cloud partnership with OpenAI frontier models including GPT-5.4
tags: 好信源 · → daily
- <a id="atom-8e59f7860be82c14"></a>2026-06-23
X·@HuggingPapers · 性能表现 · 模型性能/技术路线
GPT-5.4 collapses from 52% to 11% under severe blocking
tags: 好数字 · → daily
value: qty=11% · date=2026-06-23
📷 原图
- <a id="atom-21294708baafc2bf"></a>2026-05-03
X·@HuggingPapers · Claw-Eval-Live benchmark pass rate · 基准测试/模型性能
GPT-5.4 at 63.8%
tags: 好数字·好信源 · → daily
value: qty=63.8% · date=2026-05-03
🟦 客观事实 (facts) (30)
- <a id="atom-9451df4d89be5164"></a>🟦 2026-07-02
X·@scaling01 ⭐ · 基础性能与 GLM-5.2 相似但推理后差距扩大 · 技术路线/推理能力
GPT-5.4 and GLM-5.2 have ~the same scores without reasoning, but GLM gets crushed once you turn on reasoning
tags: 好数字·好信源 · → daily
📷 原图
- <a id="atom-e03f5f18b8144e47"></a>🟦 2026-05-16
X·@adxtyahq · API access price via Chinese proxy sellers · 价格动态/供应链
Chinese students are buying GPT-5.4/5.5 and Claude API access from Xianyu/Taobao proxy sellers for almost 96-97% cheaper
tags: 好数字·好信源 · → daily
value: qty=96-97% cheaper than official price
📷 原图
- <a id="atom-21294708baafc2bf"></a>🟦 2026-05-03
X·@HuggingPapers · Claw-Eval-Live benchmark pass rate · 基准测试/模型性能
GPT-5.4 at 63.8%
tags: 好数字·好信源 · → daily
value: qty=63.8% · date=2026-05-03
- <a id="atom-a495c285d290f561"></a>🟦 2026-04-24
X·@FundaAI · benchmark average completion time · 模型速度/模型评分
GPT-5.4 remains the fastest full-suite model (105s avg)
tags: 好数字·好信源 · → daily
value: qty=105s · date=2026-04-24
📷 原图
- <a id="atom-df9ae61577c61ca4"></a>🟦 2026-04-24
X·@FundaAI · benchmark composite score · 模型评分/代码能力
GPT-5.4 remains the fastest full-suite model (105s avg), with strong coding and reasoning, but its latest composite score is 7.88 and it no longer leads the coding table.
tags: 好数字·好信源 · → daily
value: qty=7.88 · date=2026-04-24
📷 原图
- <a id="atom-6b96c50165ed7d29"></a>🟦 2026-06-12
卖·MS Tom Wigg · AI output price per million tokens
GPT-5.4's AI output price per million tokens is approximately $14.
tags: 好数字 · → daily
value: qty=14 · date=2026-06-12
- <a id="atom-f9125dbbc2222ec6"></a>🟦 2026-06-17
X·@OpenAI · 完成的药物化学项目总耗时2.5个月,外加0.5个月撰写结果 · 技术路线/产品发布
The full process took about 2.5 months, plus another half month for human chemists to write up the results.
tags: 好数字 · → daily
value: qty=2.5个月加0.5个月 · date=2026-06-17
- <a id="atom-cb8feeb11587b79f"></a>🟦 2026-06-03
X·@scaling01 · 定价 · 价格动态
Gemini 3.5 Flash was literally 3x more expensive than GPT-5.4
tags: 好数字 · → daily
value: qty=3x less expensive · date=2026-06-03
- <a id="atom-fe3e802e5eb96d32"></a>🟦 2026-05-05
X·@scaling01 · 写出最终代码库的中位比例 · 代码生成效率
GPT 5.4 writes a median of 96% of its code in one turn
tags: 好数字 · → daily
value: qty=96% · date=single turn
📷 原图
- <a id="atom-0c21fd08359c5e64"></a>🟦 2026-06-17
X·@OpenAI · 驱动的药物化学项目 · 产品发布/技术路线
GPT-5.4 helped drive a medicinal chemistry project from literature review to a validated experimental result.
tags: 好信源 · → daily
value: qty=从文献综述到验证实验结果 · date=2026-06-17
- <a id="atom-f0c32cea968142b3"></a>🟦 2026-06-17
X·@OpenAI · 完成了科学文献综述、研究提案生成与排名、实验设计、结果分析及后续研究提案 · 技术路线
GPT-5.4 reviewed scientific literature, generated and ranked research proposals, helped design experiments, analyzed results, and proposed follow-up studies.
tags: 好信源 · → daily
value: qty= · date=
- <a id="atom-5c8960063e151321"></a>🟦 2026-04-17
X·@ArtificialAnlys · GDPval-AA 基准测试评分 · 模型能力
Opus 4.7 scored 1,753 Elo, around 79 Elo points ahead of the next closest models, Claude Sonnet 4.6 (Adaptive Reasoning, Max Effort, 1,674) and GPT-5.4 (xhigh, 1,674), and 134 Elo points ahead of Opus 4.6 (Adaptive Reasoning, Max Effort, 1,619)
tags: 好数字 · → daily
value: qty=1674 Elo · date=2026-04-18
📷 原图
- <a id="atom-cad2fddae068471a"></a>🟦 2026-04-17
X·@ArtificialAnlys · 输出 token 消耗(Intelligence Index 测试) · 模型能力
Opus 4.7 used 102M output tokens vs 157M for Opus 4.6 (Adaptive Reasoning, Max Effort), and less than GPT-5.4 (xhigh, 121M), but more than Gemini 3.1 Pro (57M)
tags: 好数字 · → daily
value: qty=121M · date=2026-04-18
📷 原图
- <a id="atom-ee45b078704095b3"></a>🟦 2026-06-28
X·@urgre · 每日使用轮数 · 模型使用量
Jun 27, 2026 gpt-5.4: 2
tags: 好数字 · → daily
value: qty=2 · date=Jun 27, 2026
[图: 大语言模型(GPT系列)使用轮数(Turn数)随时间变化的数据监控面板图表 — 总ターン数: 962; Jun 27, 2026 gpt-5.6-sol: 62; Jun 27, 2026 gpt-5.4-mini: 4; Jun 27, 2026 gpt-5.4: 2]
📷 原图
- <a id="atom-8e59f7860be82c14"></a>🟦 2026-06-23
X·@HuggingPapers · 性能表现 · 模型性能/技术路线
GPT-5.4 collapses from 52% to 11% under severe blocking
tags: 好数字 · → daily
value: qty=11% · date=2026-06-23
📷 原图
- <a id="atom-d28d5b3ac7f02208"></a>🟦 2026-06-20
X·@Rinorgashi_ · 使用限制 · 量化指标
the gpt limits for $20 are too low. I am also using GPT5.4 High not even xhigh and I can do maybe 5-6 prompts every 5 hours.
tags: 好数字 · → daily
value: qty=5-6 prompts every 5 hours
- <a id="atom-131c476f805bfdc1"></a>🟦 2026-06-20
X·@Prathkum · launch date · 产品发布
GPT 5.4 was launched on Mar 5, 2026
tags: 好数字 · → daily
value: date=Mar 5, 2026
- <a id="atom-c7d0c2f6fc50047e"></a>🟦 2026-05-28
X·@AiBattle_ · Artificial Analysis Intelligence Index 得分 · 模型基准
GPT-5.4 (xhigh) 得分: 56.8
tags: 好数字 · → daily
value: qty=56.8
[图: 不同AI模型在Artificial Analysis Intelligence Index上的得分对比柱状图 — Claude Opus 4.8 (max) 得分: 61.4; GPT-5.5 (xhigh) 得分: 60.2; Claude Opus 4.7 (max) 得分: 57.3; Gemini 3.1 Pro Preview 得分: 57
📷 原图
- <a id="atom-517a6f728d3e81a4"></a>🟦 2026-05-09
X·@hiarun02 · coding agent Elo 评分 · 模型发布
GPT-5.4 high (1877); GPT-5.4 xhigh (1870)
tags: 好数字 · → daily
value: qty=1877 · date=2026-05-09
[图: GPT-5.3 Codex、GPT-5.4 与 GPT-5.5 在不同推理层级下的 Elo 评分对比折线图 — GPT-5.5 xhigh (1994); GPT-5.4 high (1877); GPT-5.4 xhigh (1870); GPT-5.3 codex xhigh (1853); GPT-5.5 high (1807)]
📷 原图
- <a id="atom-99d35d61104449ae"></a>🟦 2026-05-06
X·@GenReasoning · 最终资金量 · 盈利能力
GPT-5.4 最终资金量: 约 £92,000
tags: 好数字 · → daily
value: qty=£92,000 · date=2026-05-06
[图: 展示不同AI模型在KellyBench(长期序列决策基准)中随时间推移的账户资金量(Bankroll)变化趋势图 — Claude Opus 4.7 最终资金量: 约 £99,000; GPT-5.4 最终资金量: 约 £92,000; Claude Opus 4.6 最终资金量: 约 £89,000; Kimi K2.5 最终资金量: 约 £10,
📷 原图
- <a id="atom-10ceb43ac11d5964"></a>🟦 2026-05-06
X·@EpochAIResearch · SWE ECI 得分 · 模型发布/技术路线
GPT-5.4 SWE ECI: 约 156
tags: 好数字 · → daily
value: qty=约 156 · date=2026-05-06
[图: 展示各大AI模型在软件工程能力指标(SWE ECI)与通用能力指标(General ECI)上的得分对比图表 — Claude Opus 4.7 SWE ECI: 约 160; Claude Opus 4.6 SWE ECI: 约 157; GPT-5.5 SWE ECI: 约 157; GPT-5.4 SWE ECI: 约 156; Gemini
📷 原图
- <a id="atom-69eec4be16ea6037"></a>🟦 2026-05-01
X·@AiBattle_ · ARC-AGI-3 分数 · 模型发布
GPT-5.4 (High) : 0.2%
tags: 好数字 · → daily
value: qty=0.2% · date=2026-05-01
📷 原图
- <a id="atom-a74b6caaadf3478e"></a>🟦 2026-05-01
X·@chatgpt21 · ARC AGI 3 score · 模型发布
GPT-5.4: 0.20%
tags: 好数字 · → daily
value: qty=0.20% · date=
📷 原图
- <a id="atom-80a75f255dbe45d9"></a>🟦 2026-04-30
X·@bookwormengr · maze navigation benchmark score · 技术路线
GPT-5.4's 50.6%
tags: 好数字 · → daily
value: qty=50.6%
📷 原图
- <a id="atom-7e7e6aac9c492499"></a>🟦 2026-04-30
X·@bookwormengr · KV cache entries for 800x800 image · 技术路线
~740
tags: 好数字 · → daily
value: qty=~740
📷 原图
- <a id="atom-16bce989086a5ab1"></a>🟦 2026-04-23
X·@Hangsiin · usage consumption · 产品发布
With GPT-5.4, it consumed twice as much usage
tags: 好数字 · → daily
value: qty=twice as much usage · date=2024-04-24
📷 原图
- <a id="atom-60ee5ba61cb1bd1e"></a>🟦 2026-04-23
X·@andonlabs · Vending-Bench Arena 最终余额 · 模型发布
GPT-5.4 最终余额: 约 $2200
tags: 好数字 · → daily
value: qty=$2200
[图: 不同AI模型(GPT-5.5、Claude Opus 4.7、GPT-5.4)在模拟交易环境中的资金余额随时间变化走势图 — GPT-5.5 最终余额: 约 $7800; Claude Opus 4.7 最终余额: 约 $5800; GPT-5.4 最终余额: 约 $2200; 最大模拟天数: 约 365 天]
📷 原图
- <a id="atom-81cca9b114e4b71b"></a>🟦 2026-04-20
X·@emollick · AI 评审排名 · 模型排名
Codex GPT-5.4 > GPT-5.3-Codex > Opus 4.6 > humans
tags: 好数字 · → daily
- <a id="atom-70decfb8eb262113"></a>🟦 2026-04-19
X·@theo · 智力指数得分 · 模型发布/技术路线
GPT-5.4 (xhigh): 57
tags: 好数字 · → daily
value: qty=57 · date=2026-04-19
[图: 各大AI大模型在Artificial Analysis智力指数上的得分对比柱状图 — Claude Opus 4.7 (max): 57; Gemini 3.1 Pro Preview: 57; GPT-5.4 (xhigh): 57; Muse Spark: 52; Claude Sonnet 4.6 (max): 52]
📷 原图
- <a id="atom-861206707c599762"></a>🟦 2026-04-10
X·@METR_Evals · point estimate time-horizon under standard methodology · 时间评估
the point estimate would be 5.7hrs (95% CI of 3hrs to 13.5hrs) under our standard methodology
tags: 好数字 · → daily
value: qty=5.7hrs · date=当前
📷 原图
🟥 多头 takes (bullish) (6)
- <a id="atom-64deadfeb4698e34"></a>🟥 2026-04-19
X·@theo · 成本对比 · 定价权/技术路线
GPT 5.4 is still way cheaper
tags: 好数字·好观点 · → daily
value: qty=way cheaper · date=2026-04-19
📷 原图
- <a id="atom-65992fd4127eb7e9"></a>🟥 2026-06-17
X·@OpenAI · 作为前沿模型支持更多科学研究循环的早期实例 · 技术路线
This is an early example of frontier models supporting more of the scientific research loop: reviewing studies, proposing hypotheses, designing experiments, interpreting data, and surfacing findings that human experts can validate.
tags: 好观点 · → daily
value: qty= · date=2026-06-17
- <a id="atom-8e1bd77f8b51ef10"></a>🟥 2026-05-02
X·@boazbaraktcs [老化中] · produced new proof for longstanding math problem · 模型发布/技术路线
GPT-5.4-pro came up with a new proof for one longstanding problem
tags: 好观点 · → daily
- <a id="atom-90491926e7142b02"></a>🟥 2026-05-02
X·@boazbaraktcs [老化中] · proof contains new idea applicable to multiple other problems · 模型发布/技术路线
the proof contains a new idea that can be used for multiple other problems
tags: 好观点 · → daily
- <a id="atom-9683cc49215936c3"></a>🟥 2026-04-16
X·@max_paperclips [老化中] · 可以信任完成复杂任务 · 模型能力
with GPT-5.4 I can just give it a difficult task and trust more or less that if it think it's fixed, it's probably fixed or at least did do something rational to fix it
tags: 好观点 · → daily
- <a id="atom-b4f82b796224a0ba"></a>🟥 2026-04-15
X·@doodlestein [老化中] · 协作能力评价 · 合作客户
It makes an incredible team with GPT 5.4
tags: 好观点 · → daily
value: direction=positive
🟥 空头 takes (bearish) (4)
- <a id="atom-8d323c59da8a1e4a"></a>🟥 2026-05-04
X·@bioshok3 ⭐ [老化中] · 实验辅助能力超过博士级病毒学家 · 技术路线/地缘政治
GPT-5.4は博士レベルのウイルス学者を上回る実験支援能力を持つとされ、バイオリスクも深刻化
tags: 好观点·好思考 · → daily
📷 原图
- <a id="atom-1ee9d455c0ba69ac"></a>🟥 2026-05-05
X·@labomen001 [老旧] · prompting/agent harness 选择不当导致提前停止 · 模型表现问题
I think their prompting/agent harness were just a bad pick for GPT 5.4 since it seems to stop really early
tags: 好思考 · → daily
📷 原图
- <a id="atom-d5a04c1ef1b43b6b"></a>🟥 2026-04-18
X·@Samhanknr [老化中] · 模型成本定价 · 定价权
If it’s smaller why is it more expensive 🤔
tags: 好思考 · → daily
value: qty=more expensive
- <a id="atom-7e70211c87fd623c"></a>🟥 2026-06-22
X·@pigeon_s · 预训练基础 · 技术路线_
if its based on 5.4 which is still based on 4o/t and not the new pretrain like 5.5 then its still FUCKING GPT-4o VOICE
tags: 好观点 · → daily
🟥 中性 takes (neutral) (1)
- <a id="atom-2ca7abd116436556"></a>🟥 2026-04-30
X·@ZhihuFrontier [老化中] · 参数规模估计 · 模型规模
GPT-5.4 ≈ 2.2T
tags: 好数字 · → daily
value: qty=≈2.2T · date=2026-04
📷 原图
⏱ 时间轴 (近 20)
- 🟦 2026-07-02 ·
fact · 基础性能与 GLM-5.2 相似但推理后差距扩大 · → daily
- 🟦 2026-06-28 ·
fact · 每日使用轮数 · → daily
- 🟦 2026-06-23 ·
fact · 性能表现 · → daily
- 🟥 2026-06-22 ·
fact · 预训练基础 · → daily
- 🟦 2026-06-20 ·
fact · 使用限制 · → daily
- 🟦 2026-06-20 ·
fact · launch date · → daily
- 🟦 2026-06-17 ·
fact · 驱动的药物化学项目 · → daily
- 🟦 2026-06-17 ·
fact · 提出改进药物发现中广泛使用反应的方法 · → daily
- 🟦 2026-06-17 ·
fact · 完成了科学文献综述、研究提案生成与排名、实验设计、结果分析及后续研究提案 · → daily
- 🟦 2026-06-17 ·
fact · 完成的药物化学项目总耗时2.5个月,外加0.5个月撰写结果 · → daily
- 🟥 2026-06-17 ·
narrative · 作为前沿模型支持更多科学研究循环的早期实例 · → daily
- 🟦 2026-06-12 ·
fact · AI output price per million tokens · → daily
- 🟦 2026-06-03 ·
fact · 定价 · → daily
- 🟦 2026-05-28 ·
fact · Artificial Analysis Intelligence Index 得分 · → daily
- 🟦 2026-05-16 ·
fact · API access price via Chinese proxy sellers · → daily
- 🟦 2026-05-09 ·
fact · coding agent Elo 评分 · → daily
- 🟦 2026-05-06 ·
fact · 最终资金量 · → daily
- 🟦 2026-05-06 ·
fact · SWE ECI 得分 · → daily
- 🟦 2026-05-05 ·
fact · 写出最终代码库的中位比例 · → daily
- 🟥 2026-05-05 ·
position · prompting/agent harness 选择不当导致提前停止 · → daily
← 实体目录 · 系统日志