以下部分引用 AI 总结,现阶段 AI 仍然有幻觉。请以内容的原文为准。
数据新鲜度群聊 07-14 ✓卖方 07-14 ✓Wrap 断档15天 ✕News 07-15 ✓X 断档12天 ✕页面生成 07-15
Claude Mythos 5 / Fable 5
30 atoms · 跨 2 天 · 首见 2026-06-10 · 最近 2026-06-25
三色: 🟦 fact 30 · 🟥 take 0 · stance ▲15/▼0/◆15
叙事状态: 🍂 fading (退潮) · 💤 沉寂 · 策展近14d 0 atom
状态只算低频策展源 (群/卖方); X firehose 仅作背景音量
来源: 卖 30
时态: fresh:30
标签: 好数字:30
🟦 客观事实 (facts) (30)
- <a id="atom-17cbd63479a99e85"></a>🟦 2026-06-25
卖·MS Tom Wigg · Agentic coding (SWE-Bench Pro) score
Claude Mythos 5 / Fable 5 achieved a score of 80.3% on the Agentic coding (SWE-Bench Pro) benchmark.
tags: 好数字 · → daily
value: qty=80.3%
- <a id="atom-1acf96ca13bc984a"></a>🟦 2026-06-25
卖·MS Tom Wigg · Agentic coding (FrontierCode (Diamond)) score
Claude Mythos 5 / Fable 5 achieved a score of 29.3% on the Agentic coding (FrontierCode (Diamond)) benchmark.
tags: 好数字 · → daily
value: qty=29.3%
- <a id="atom-5a2ec1820e95b140"></a>🟦 2026-06-25
卖·MS Tom Wigg · Knowledge work (GDPval-AA) score
Claude Mythos 5 / Fable 5 achieved a score of 1932 on the Knowledge work (GDPval-AA) benchmark.
tags: 好数字 · → daily
value: qty=1932
- <a id="atom-f21d9265cb151e33"></a>🟦 2026-06-25
卖·MS Tom Wigg · Knowledge work vision (GDP.pdf) score
Claude Mythos 5 / Fable 5 achieved a score of 29.8% on the Knowledge work vision (GDP.pdf) benchmark.
tags: 好数字 · → daily
value: qty=29.8%
- <a id="atom-7af319c2c1a618ce"></a>🟦 2026-06-25
卖·MS Tom Wigg · Spatial reasoning (Blueprint-Bench 2) score
Claude Mythos 5 / Fable 5 achieved a score of 38.6% on the Spatial reasoning (Blueprint-Bench 2) benchmark.
tags: 好数字 · → daily
value: qty=38.6%
- <a id="atom-a2225b8598bb05a1"></a>🟦 2026-06-25
卖·MS Tom Wigg · Tool use (AutomationBench) score
Claude Mythos 5 / Fable 5 achieved a score of 17.4% on the Tool use (AutomationBench) benchmark.
tags: 好数字 · → daily
value: qty=17.4%
- <a id="atom-464ccf8151254dc3"></a>🟦 2026-06-25
卖·MS Tom Wigg · Computer use (OSWorld-Verified) score
Claude Mythos 5 / Fable 5 achieved a score of 85.0% on the Computer use (OSWorld-Verified) benchmark.
tags: 好数字 · → daily
value: qty=85.0%
- <a id="atom-57006638f1487bdc"></a>🟦 2026-06-25
卖·MS Tom Wigg · Legal (Legal Agent Benchmark) score
Claude Mythos 5 / Fable 5 achieved a score of 13.3% on the Legal (Legal Agent Benchmark) benchmark.
tags: 好数字 · → daily
value: qty=13.3%
- <a id="atom-051f6930a5f6f3ca"></a>🟦 2026-06-25
卖·MS Tom Wigg · Multidisciplinary reasoning (Humanity's Last Exam, no tools) score
Claude Mythos 5 / Fable 5 achieved a score of 59.0% on the Multidisciplinary reasoning (Humanity's Last Exam, no tools) benchmark.
tags: 好数字 · → daily
value: qty=59.0%
- <a id="atom-5801f9a3d25867a5"></a>🟦 2026-06-25
卖·MS Tom Wigg · Multidisciplinary reasoning (Humanity's Last Exam, with tools) score
Claude Mythos 5 / Fable 5 achieved a score of 64.5% on the Multidisciplinary reasoning (Humanity's Last Exam, with tools) benchmark.
tags: 好数字 · → daily
value: qty=64.5%
- <a id="atom-b5f99163857cc2f1"></a>🟦 2026-06-25
卖·MS Tom Wigg · Biology (BioMysteryBench, hard) score
Claude Mythos 5 / Fable 5 achieved a score of 46.1% on the Biology (BioMysteryBench, hard) benchmark.
tags: 好数字 · → daily
value: qty=46.1%
- <a id="atom-8aabbb968e715df3"></a>🟦 2026-06-25
卖·MS Tom Wigg · Biology (BioMysteryBench, human solved) score
Claude Mythos 5 / Fable 5 achieved a score of 83.9% on the Biology (BioMysteryBench, human solved) benchmark.
tags: 好数字 · → daily
value: qty=83.9%
- <a id="atom-fbba9ac355ea9787"></a>🟦 2026-06-25
卖·MS Tom Wigg · Agentic coding (Terminal-Bench 2.1) score
Claude Mythos 5 / Fable 5 achieved a score of 88.0% on the Agentic coding (Terminal-Bench 2.1) benchmark.
tags: 好数字 · → daily
value: qty=88.0%
- <a id="atom-368ec326334ae32d"></a>🟦 2026-06-25
卖·MS Tom Wigg · Cybersecurity (ExploitBench (Cap%)) score
Claude Mythos 5 / Fable 5 achieved a score of 78.0% on the Cybersecurity (ExploitBench (Cap%)) benchmark.
tags: 好数字 · → daily
value: qty=78.0%
- <a id="atom-929d220fa25cfe6a"></a>🟦 2026-06-25
卖·MS Tom Wigg · Health (HealthBench Professional) score
Claude Mythos 5 / Fable 5 achieved a score of 66.0% on the Health (HealthBench Professional) benchmark.
tags: 好数字 · → daily
value: qty=66.0%
- <a id="atom-e5fef384f2c7fc9f"></a>🟦 2026-06-10
卖·MS Tom Wigg · Agentic coding SWE-Bench Pro score
Claude Mythos 5 / Fable 5 scored 80.3% on Agentic coding SWE-Bench Pro.
tags: 好数字 · → daily
value: qty=80.3%
- <a id="atom-e753646c04cc3a7c"></a>🟦 2026-06-10
卖·MS Tom Wigg · Agentic coding FrontierCode (Diamond) score
Claude Mythos 5 / Fable 5 scored 29.3% on Agentic coding FrontierCode (Diamond).
tags: 好数字 · → daily
value: qty=29.3%
- <a id="atom-3a6cd8e137277991"></a>🟦 2026-06-10
卖·MS Tom Wigg · Knowledge work GDPval-AA score
Claude Mythos 5 / Fable 5 scored 1932 on Knowledge work GDPval-AA.
tags: 好数字 · → daily
value: qty=1932
- <a id="atom-15c0cb1635d90909"></a>🟦 2026-06-10
卖·MS Tom Wigg · Knowledge work vision GDP.pdf score
Claude Mythos 5 / Fable 5 scored 29.8% on Knowledge work vision GDP.pdf.
tags: 好数字 · → daily
value: qty=29.8%
- <a id="atom-6f9f14137f53a97a"></a>🟦 2026-06-10
卖·MS Tom Wigg · Spatial reasoning Blueprint-Bench 2 score
Claude Mythos 5 / Fable 5 scored 38.6% on Spatial reasoning Blueprint-Bench 2.
tags: 好数字 · → daily
value: qty=38.6%
- <a id="atom-406e37122675050b"></a>🟦 2026-06-10
卖·MS Tom Wigg · Tool use AutomationBench score
Claude Mythos 5 / Fable 5 scored 17.4% on Tool use AutomationBench.
tags: 好数字 · → daily
value: qty=17.4%
- <a id="atom-0d8a3b57258e9101"></a>🟦 2026-06-10
卖·MS Tom Wigg · Computer use OSWorld-Verified score
Claude Mythos 5 / Fable 5 scored 85.0% on Computer use OSWorld-Verified.
tags: 好数字 · → daily
value: qty=85.0%
- <a id="atom-7e8f632188b87eec"></a>🟦 2026-06-10
卖·MS Tom Wigg · Legal Agent Benchmark score
Claude Mythos 5 / Fable 5 scored 13.3% on Legal Agent Benchmark.
tags: 好数字 · → daily
value: qty=13.3%
- <a id="atom-7e092f2396fb19b8"></a>🟦 2026-06-10
卖·MS Tom Wigg · Multidisciplinary reasoning Humanity's Last Exam (no tools) score
Claude Mythos 5 / Fable 5 scored 59.0% on Multidisciplinary reasoning Humanity's Last Exam (no tools).
tags: 好数字 · → daily
value: qty=59.0%
- <a id="atom-39c259656fe8210e"></a>🟦 2026-06-10
卖·MS Tom Wigg · Multidisciplinary reasoning Humanity's Last Exam (with tools) score
Claude Mythos 5 / Fable 5 scored 64.5% on Multidisciplinary reasoning Humanity's Last Exam (with tools).
tags: 好数字 · → daily
value: qty=64.5%
- <a id="atom-b2b5a304461ed834"></a>🟦 2026-06-10
卖·MS Tom Wigg · Biology BioMysteryBench (hard) score
Claude Mythos 5 / Fable 5 scored 46.1% on Biology BioMysteryBench (hard).
tags: 好数字 · → daily
value: qty=46.1%
- <a id="atom-e28622eb4c209cb1"></a>🟦 2026-06-10
卖·MS Tom Wigg · Biology BioMysteryBench (human solved) score
Claude Mythos 5 / Fable 5 scored 83.9% on Biology BioMysteryBench (human solved).
tags: 好数字 · → daily
value: qty=83.9%
- <a id="atom-2d56a51e2967c880"></a>🟦 2026-06-10
卖·MS Tom Wigg · Agentic coding Terminal-Bench 2.1 score
Claude Mythos 5 / Fable 5 scored 88.0% on Agentic coding Terminal-Bench 2.1.
tags: 好数字 · → daily
value: qty=88.0%
- <a id="atom-ca5a0198f9c47f03"></a>🟦 2026-06-10
卖·MS Tom Wigg · Cybersecurity ExploitBench (Cap%) score
Claude Mythos 5 / Fable 5 scored 78.0% on Cybersecurity ExploitBench (Cap%).
tags: 好数字 · → daily
value: qty=78.0%
- <a id="atom-08f6dfea3836afc7"></a>🟦 2026-06-10
卖·MS Tom Wigg · Health HealthBench Professional score
Claude Mythos 5 / Fable 5 scored 66.0% on Health HealthBench Professional.
tags: 好数字 · → daily
value: qty=66.0%
⏱ 时间轴 (近 20)
- 🟦 2026-06-25 ·
fact · Agentic coding (SWE-Bench Pro) score · → daily
- 🟦 2026-06-25 ·
fact · Agentic coding (FrontierCode (Diamond)) score · → daily
- 🟦 2026-06-25 ·
fact · Knowledge work (GDPval-AA) score · → daily
- 🟦 2026-06-25 ·
fact · Knowledge work vision (GDP.pdf) score · → daily
- 🟦 2026-06-25 ·
fact · Spatial reasoning (Blueprint-Bench 2) score · → daily
- 🟦 2026-06-25 ·
fact · Tool use (AutomationBench) score · → daily
- 🟦 2026-06-25 ·
fact · Computer use (OSWorld-Verified) score · → daily
- 🟦 2026-06-25 ·
fact · Legal (Legal Agent Benchmark) score · → daily
- 🟦 2026-06-25 ·
fact · Multidisciplinary reasoning (Humanity's Last Exam, no tools) · → daily
- 🟦 2026-06-25 ·
fact · Multidisciplinary reasoning (Humanity's Last Exam, with tool · → daily
- 🟦 2026-06-25 ·
fact · Biology (BioMysteryBench, hard) score · → daily
- 🟦 2026-06-25 ·
fact · Biology (BioMysteryBench, human solved) score · → daily
- 🟦 2026-06-25 ·
fact · Agentic coding (Terminal-Bench 2.1) score · → daily
- 🟦 2026-06-25 ·
fact · Cybersecurity (ExploitBench (Cap%)) score · → daily
- 🟦 2026-06-25 ·
fact · Health (HealthBench Professional) score · → daily
- 🟦 2026-06-10 ·
fact · Agentic coding SWE-Bench Pro score · → daily
- 🟦 2026-06-10 ·
fact · Agentic coding FrontierCode (Diamond) score · → daily
- 🟦 2026-06-10 ·
fact · Knowledge work GDPval-AA score · → daily
- 🟦 2026-06-10 ·
fact · Knowledge work vision GDP.pdf score · → daily
- 🟦 2026-06-10 ·
fact · Spatial reasoning Blueprint-Bench 2 score · → daily
← 实体目录 · 系统日志