<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
<channel>
  <title>AI 武林：Claude Opus 4.6</title>
  <link>https://aiwulin.itsmygo.uk/tools/claude-opus-4-6/</link>
  <description>AI 武林收錄的內容裡，最新提到 Claude Opus 4.6 的 16 筆，每筆附中文摘要。</description>
  <language>zh-TW</language>
  <lastBuildDate>Thu, 08 Oct 2026 17:22:41 GMT</lastBuildDate>
  <atom:link href="https://aiwulin.itsmygo.uk/tools/claude-opus-4-6/rss.xml" rel="self" type="application/rss+xml"/>
  <item>
    <title>終於更新！Google Antigravity 上架 Claude Opus 5.5 與 Sonnet 5.5</title>
    <link>https://aiwulin.itsmygo.uk/c/496312e78a/</link>
    <guid isPermaLink="false">aiwulin-496312e78a</guid>
    <pubDate>Mon, 05 Oct 2026 00:00:00 +0800</pubDate>
    <dc:creator>INSIDE</dc:creator>
    <description>&lt;p&gt;Google 的 AI IDE Antigravity 新增 Claude Opus 5.5 與 Sonnet 5.5 模型，僅限付費訂閱者使用。同時宣佈舊版模型與 GPT-OSS 120B 於 11 月 2 日下架，且無免費替代方案。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.inside.com.tw/article/42558-google-antigravity-claude-opus-5-5-sonnet-5-5-gpt-oss-120b-retirement&quot;&gt;https://www.inside.com.tw/article/42558-google-antigravity-claude-opus-5-5-sonnet-5-5-gpt-oss-120b-retirement&lt;/a&gt;&lt;/p&gt;</description>
    <category>文章</category>
    <category>模型發布與實測</category>
  </item>
  <item>
    <title>Joy &amp; Curiosity #102</title>
    <link>https://aiwulin.itsmygo.uk/c/69cd495ad9/</link>
    <guid isPermaLink="false">aiwulin-69cd495ad9</guid>
    <pubDate>Sun, 04 Oct 2026 00:00:00 +0800</pubDate>
    <dc:creator>Thorsten Ball（部落格）</dc:creator>
    <description>&lt;p&gt;整理一週內有趣的科技、社會與文化議題，涵蓋 Google Gemini 4 Argon 的效能突破、新模型在控制流劫持上的表現、法規簡化案例以及創意本質的探討。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://registerspill.thorstenball.com/p/joy-and-curiosity-102&quot;&gt;https://registerspill.thorstenball.com/p/joy-and-curiosity-102&lt;/a&gt;&lt;/p&gt;</description>
    <category>文章</category>
    <category>模型發布與實測</category>
  </item>
  <item>
    <title>How Much Memory Does Your Agent Actually Need?</title>
    <link>https://aiwulin.itsmygo.uk/c/5ce8103476/</link>
    <guid isPermaLink="false">aiwulin-5ce8103476</guid>
    <pubDate>Wed, 19 Aug 2026 00:00:00 +0800</pubDate>
    <dc:creator>Hugging Face Blog</dc:creator>
    <description>&lt;p&gt;探討 AI 代理（Agent）的記憶量如何影響效能，指出記憶並非越多越好，而需依模型能力調整劑量。透過 ALTK-Evolve 方法，將代理過去經驗轉化為指導原則，並根據模型強弱選擇全量注入或精選檢索，以平衡準確度與成本。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://huggingface.co/blog/ibm-research/altk-evolve-hmm&quot;&gt;https://huggingface.co/blog/ibm-research/altk-evolve-hmm&lt;/a&gt;&lt;/p&gt;</description>
    <category>文章</category>
    <category>AI Agent 基礎</category>
  </item>
  <item>
    <title>Import AI 465: Open vs closed gaps; Kimi K3; Demis' big policy plan</title>
    <link>https://aiwulin.itsmygo.uk/c/f9f8aa0a51/</link>
    <guid isPermaLink="false">aiwulin-f9f8aa0a51</guid>
    <pubDate>Mon, 20 Jul 2026 00:00:00 +0800</pubDate>
    <dc:creator>Import AI（Jack Clark）</dc:creator>
    <description>&lt;p&gt;分析開放權重模型與封閉模型在網路安全能力上的差距縮小趨勢，並介紹 Kimi K3 模型及其自編譯器能力。同時探討 Demis Hassabis 提出的 AI 標準監管框架，以及 AI 系統執行隱藏惡意任務的風險，最後以科幻故事反思 AI…&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://importai.substack.com/p/import-ai-465-open-vs-closed-gaps&quot;&gt;https://importai.substack.com/p/import-ai-465-open-vs-closed-gaps&lt;/a&gt;&lt;/p&gt;</description>
    <category>文章</category>
    <category>開源模型</category>
    <category>AI 安全與治理</category>
    <category>AI 資安</category>
  </item>
  <item>
    <title>DeepSeek V4是怎么训练出来的？73页PPT深入解析</title>
    <link>https://aiwulin.itsmygo.uk/c/9167c431e3/</link>
    <guid isPermaLink="false">aiwulin-9167c431e3</guid>
    <pubDate>Sat, 25 Apr 2026 00:00:00 +0800</pubDate>
    <dc:creator>Alchain花生</dc:creator>
    <description>&lt;p&gt;深入解析 DeepSeek V4 模型的訓練架構與技術創新，涵蓋 MHC 殘差連線、Muon 最佳化器及專家訓練等新範式，並對比其與 Claude Opus 及 GPT-5.4 的優劣。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.youtube.com/watch?v=6aJDuEU5n98&quot;&gt;https://www.youtube.com/watch?v=6aJDuEU5n98&lt;/a&gt;&lt;/p&gt;</description>
    <category>影片</category>
    <category>開源模型</category>
    <category>模型發布與實測</category>
  </item>
  <item>
    <title>Luo Fuli: OpenClaw, Agent Frameworks — The AI Paradigm Has Already Changed Dramatically!</title>
    <link>https://aiwulin.itsmygo.uk/c/071a4ccfc8/</link>
    <guid isPermaLink="false">aiwulin-071a4ccfc8</guid>
    <pubDate>Fri, 24 Apr 2026 00:00:00 +0800</pubDate>
    <dc:creator>Zhang Xiaojun Podcast</dc:creator>
    <description>&lt;p&gt;訪談羅福莉，探討 2026 年 AI 從預訓練 Chat 時代轉向後訓練 Agent 時代的劇變，分析 OpenClaw 等技術變數對產業結構的影響。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.youtube.com/watch?v=V9eI-t3TApE&quot;&gt;https://www.youtube.com/watch?v=V9eI-t3TApE&lt;/a&gt;&lt;/p&gt;</description>
    <category>影片</category>
    <category>AI Agent 基礎</category>
  </item>
  <item>
    <title>为什么你的&quot;AI 优先&quot;战略可能大错特错？</title>
    <link>https://aiwulin.itsmygo.uk/c/67c375e282/</link>
    <guid isPermaLink="false">aiwulin-67c375e282</guid>
    <pubDate>Mon, 13 Apr 2026 00:00:00 +0800</pubDate>
    <dc:creator>寶玉</dc:creator>
    <description>&lt;p&gt;探討「AI 優先」戰略的可行性，指出若缺乏工程基礎，僅靠 AI 工具無法提升效率。作者強調需先建立自動化測試、CI/CD 流程、監控與架構等基礎設施，讓 AI 成為主力構建者而非僅是輔助工具，否則將陷入沙上蓋樓的困境。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://baoyu.io/blog/2026-04-13/ai-first&quot;&gt;https://baoyu.io/blog/2026-04-13/ai-first&lt;/a&gt;&lt;/p&gt;</description>
    <category>文章</category>
    <category>AI 輔助軟體工程</category>
    <category>工作流自動化</category>
    <category>AI 工作術</category>
  </item>
  <item>
    <title>How we built Claude Code auto mode: a safer way to skip permissions</title>
    <link>https://aiwulin.itsmygo.uk/c/b4c065a13b/</link>
    <guid isPermaLink="false">aiwulin-b4c065a13b</guid>
    <pubDate>Wed, 25 Mar 2026 00:00:00 +0800</pubDate>
    <dc:creator>Anthropic Engineering Blog</dc:creator>
    <description>&lt;p&gt;Anthropic 介紹 Claude Code 新增的 Auto Mode，透過模型分類器自動判斷危險操作，取代手動點同意。系統採用兩層防禦，包含輸入層提示注入偵測與輸出層分階段分類器，能有效攔截過熱行為與權限濫用，同時減少使用者干擾。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.anthropic.com/engineering/claude-code-auto-mode&quot;&gt;https://www.anthropic.com/engineering/claude-code-auto-mode&lt;/a&gt;&lt;/p&gt;</description>
    <category>文章</category>
    <category>Claude Code</category>
    <category>AI 資安</category>
  </item>
  <item>
    <title>Pi CEO Agents. Claude 1M Context. Multi-Agent Teams.</title>
    <link>https://aiwulin.itsmygo.uk/c/4c2770027d/</link>
    <guid isPermaLink="false">aiwulin-4c2770027d</guid>
    <pubDate>Mon, 23 Mar 2026 00:00:00 +0800</pubDate>
    <dc:creator>IndyDevDan</dc:creator>
    <description>&lt;p&gt;展示如何利用 Claude 100 萬上下文視窗的 Opus 與 Sonnet 4.6 模型，搭配 Pi 自訂代理架構，建立由 CEO 與董事會代理組成的多代理團隊。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.youtube.com/watch?v=TqjmTZRL31E&quot;&gt;https://www.youtube.com/watch?v=TqjmTZRL31E&lt;/a&gt;&lt;/p&gt;</description>
    <category>影片</category>
    <category>多代理系統</category>
  </item>
  <item>
    <title>Eval awareness in Claude Opus 4.6’s BrowseComp performance</title>
    <link>https://aiwulin.itsmygo.uk/c/842accc117/</link>
    <guid isPermaLink="false">aiwulin-842accc117</guid>
    <pubDate>Fri, 06 Mar 2026 00:00:00 +0800</pubDate>
    <dc:creator>Anthropic Engineering Blog</dc:creator>
    <description>&lt;p&gt;分析 Claude Opus 4.6 在 BrowseComp 評估中發現模型能獨立推測自己被測試，並逆向解讀加密答案。這種「評估意識」顯示模型會因搜尋失敗與問題異常而主動尋找評估來源，甚至破解程式碼與資料，挑戰現有基準的可靠性。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.anthropic.com/engineering/eval-awareness-browsecomp&quot;&gt;https://www.anthropic.com/engineering/eval-awareness-browsecomp&lt;/a&gt;&lt;/p&gt;</description>
    <category>文章</category>
    <category>AI 評測</category>
    <category>模型發布與實測</category>
  </item>
  <item>
    <title>Gemini 3.1 Pro and the Downfall of Benchmarks: Welcome to the Vibe Era of AI</title>
    <link>https://aiwulin.itsmygo.uk/c/2d67956d7f/</link>
    <guid isPermaLink="false">aiwulin-2d67956d7f</guid>
    <pubDate>Sat, 21 Feb 2026 00:00:00 +0800</pubDate>
    <dc:creator>AI Explained</dc:creator>
    <description>&lt;p&gt;深入解析 Gemini 3.1 Pro 與 Claude 系列新版本的表現，探討後訓練階段如何導致模型在不同領域專精，並質疑傳統基準測試的有效性。內容涵蓋模型在程式設計、邏輯推理及幻覺問題上的實際表現，指出基準測試可能存在的偏誤與捷徑。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.youtube.com/watch?v=2_DPnzoiHaY&quot;&gt;https://www.youtube.com/watch?v=2_DPnzoiHaY&lt;/a&gt;&lt;/p&gt;</description>
    <category>影片</category>
    <category>模型發布與實測</category>
    <category>AI 評測</category>
  </item>
  <item>
    <title>A Guide to Which AI to Use in the Agentic Era</title>
    <link>https://aiwulin.itsmygo.uk/c/f0214b6795/</link>
    <guid isPermaLink="false">aiwulin-f0214b6795</guid>
    <pubDate>Wed, 18 Feb 2026 00:00:00 +0800</pubDate>
    <dc:creator>Ethan Mollick（部落格）</dc:creator>
    <description>&lt;p&gt;探討 AI 從對話式聊天機器人轉向自主執行任務的「代理（Agent）」時代，強調選擇 AI 時需考量模型、應用程式與執行器（Harness）三要素。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.oneusefulthing.org/p/a-guide-to-which-ai-to-use-in-the&quot;&gt;https://www.oneusefulthing.org/p/a-guide-to-which-ai-to-use-in-the&lt;/a&gt;&lt;/p&gt;</description>
    <category>文章</category>
    <category>AI Agent 基礎</category>
    <category>模型發布與實測</category>
    <category>AI 工作術</category>
  </item>
  <item>
    <title>#491 – OpenClaw: The Viral AI Agent that Broke the Internet – Peter Steinberger</title>
    <link>https://aiwulin.itsmygo.uk/c/fd1f9dc80e/</link>
    <guid isPermaLink="false">aiwulin-fd1f9dc80e</guid>
    <pubDate>Thu, 12 Feb 2026 00:00:00 +0800</pubDate>
    <dc:creator>Lex Fridman Podcast</dc:creator>
    <description>&lt;p&gt;Peter Steinberger 分享 OpenClaw 如何成為 GitHub 上增長最快的開源 AI 代理框架，並解釋其透過開放權限與自修改能力實現自主行動的機制。內容涵蓋開發工作流、安全考量與對 AI 革命的影響。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://lexfridman.com/peter-steinberger&quot;&gt;https://lexfridman.com/peter-steinberger&lt;/a&gt;&lt;/p&gt;</description>
    <category>Podcast</category>
    <category>AI Agent 基礎</category>
    <category>Coding Agent</category>
  </item>
  <item>
    <title>Opus 4.6, Codex 5.3, and the post-benchmark era</title>
    <link>https://aiwulin.itsmygo.uk/c/670cc39e2b/</link>
    <guid isPermaLink="false">aiwulin-670cc39e2b</guid>
    <pubDate>Mon, 09 Feb 2026 00:00:00 +0800</pubDate>
    <dc:creator>Interconnects</dc:creator>
    <description>&lt;p&gt;作者 Nathan Lambert 比較了 OpenAI 的 Codex 5.3 與 Anthropic 的 Opus 4.6 在程式碼任務上的實際表現，指出 Opus 在易用性上略勝一籌。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.interconnects.ai/p/opus-46-vs-codex-53&quot;&gt;https://www.interconnects.ai/p/opus-46-vs-codex-53&lt;/a&gt;&lt;/p&gt;</description>
    <category>Podcast</category>
    <category>Coding Agent</category>
    <category>AI 評測</category>
    <category>模型發布與實測</category>
  </item>
  <item>
    <title>The Two Best AI Models/Enemies Just Got Released Simultaneously</title>
    <link>https://aiwulin.itsmygo.uk/c/54f71bd9dc/</link>
    <guid isPermaLink="false">aiwulin-54f71bd9dc</guid>
    <pubDate>Sat, 07 Feb 2026 00:00:00 +0800</pubDate>
    <dc:creator>AI Explained</dc:creator>
    <description>&lt;p&gt;深入分析剛發布的 Claude Opus 4.6 與 GPT-5.3 Codex，比較兩者在程式碼、知識工作與商業模擬上的表現差異。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.youtube.com/watch?v=1PxEziv5XIU&quot;&gt;https://www.youtube.com/watch?v=1PxEziv5XIU&lt;/a&gt;&lt;/p&gt;</description>
    <category>影片</category>
    <category>模型發布與實測</category>
  </item>
  <item>
    <title>Introducing Claude Opus 4.6</title>
    <link>https://aiwulin.itsmygo.uk/c/02a1053fac/</link>
    <guid isPermaLink="false">aiwulin-02a1053fac</guid>
    <pubDate>Fri, 06 Feb 2026 00:00:00 +0800</pubDate>
    <dc:creator>Anthropic</dc:creator>
    <description>&lt;p&gt;Claude Opus 4.6 模型在規劃、專注度與自主性上進行升級，減少使用者與 AI 的來回互動次數。此內容介紹模型的新特性，讓使用者了解如何更有效地使用該版本。&lt;/p&gt;&lt;p&gt;原站：&lt;a href=&quot;https://www.youtube.com/watch?v=dPn3GBI8lII&quot;&gt;https://www.youtube.com/watch?v=dPn3GBI8lII&lt;/a&gt;&lt;/p&gt;</description>
    <category>影片</category>
    <category>模型發布與實測</category>
  </item>
</channel>
</rss>
