以下部分引用 AI 总结,现阶段 AI 仍然有幻觉。请以内容的原文为准。
数据新鲜度群聊 07-14 ✓卖方 07-14 ✓Wrap 断档15天 ✕News 07-15 ✓X 断档12天 ✕页面生成 07-15
Opus 4.5
9 atoms · 跨 7 天 · 首见 2026-05-10 · 最近 2026-06-24
三色: 🟦 fact 4 · 🟥 take 5 · stance ▲4/▼1/◆1
来源: X 9
时态: fresh:8 · aging:1
标签: 好观点:6 · 好数字:3 · 好信源:1 · 好思考:1
叙事 (narrative) (4)
- 2026-06-18 ⭐ · 被开源模型超越
Open-models have now surpassed Opus 4.5
tags: 好观点·好思考 · → daily
- 2026-06-16 · quality perception
Opus 4.5 is/was amazing, and is more than good enough for almost all tasks still as long as you pair with a frontier-level planner/judge.
tags: 好观点 · → daily
value: direction=amazing
- 2026-05-26 · 日常使用实用性对比
Opus 4.5 is actually useful for day to day work and Deepseek v4 Pro isn't
tags: 好观点 · → daily
value: direction=more useful
- 2026-05-10
[老化中] · HTML AI 实验一致性
Opus 4.5 was the most consistent, across all model families longer reasoning time did not produce better outputs
tags: 好观点 · → daily
value: qty=most consistent
事实 (fact) (5)
- 2026-06-01 · release date
opus 4.5 was november 24
tags: 好数字·好信源 · → daily
value: date=November 2024
- 2026-06-24 · ARC-AGI-2得分
on par with Opus 4.5 (16K)
tags: 好数字 · → daily
value: qty=unclear · direction=
- 2026-06-18 · 性能比较
better than Opus 4.5, not quite as good as Opus 4.6 or GPT-5.2-xhigh
tags: 好观点 · → daily
value: direction=better than
- 2026-05-26 · ARC-AGI-2得分
Opus 4.5 ARC-AGI-2得分: 30.6
tags: 好数字 · → daily
value: qty=30.6 · date=2025-11-24
- 2026-06-09 · capability frontier advancement
Opus 4.5 made Agentic possible
tags: 好观点 · → daily
时间轴 (近 20)
- 2026-06-24 ·
fact · ARC-AGI-2得分 · → daily
- 2026-06-18 ·
narrative · 被开源模型超越 · → daily
- 2026-06-18 ·
fact · 性能比较 · → daily
- 2026-06-16 ·
narrative · quality perception · → daily
- 2026-06-09 ·
fact · capability frontier advancement · → daily
- 2026-06-01 ·
fact · release date · → daily
- 2026-05-26 ·
fact · ARC-AGI-2得分 · → daily
- 2026-05-26 ·
narrative · 日常使用实用性对比 · → daily
- 2026-05-10 ·
narrative · HTML AI 实验一致性 · → daily
← 实体目录 · 系统日志