看影片(在新分頁開啟原站)連到 AI Native Dev
摘要
探討編碼代理在驗證階段失靈而非生成的問題,分享 Baz、Docker 前 CTO、Meta 工程師與 Christopher Batey 的實戰經驗。觀眾能學習如何透過具體規格、真實測試軌道及架構決策記錄來確保 AI 程式碼可靠性。
This talk discusses why coding agents fail at verification, sharing real-world strategies from Baz, Docker, and Meta to ensure reliable AI-generated code through rigorous testing and architectural records.
摘要、重點與章節標題由語言模型整理,細節(誰說的、數字、先後)可能有誤;要引用請以原始內容為準。
重點
- 編碼代理擅長生成程式碼,卻不擅長驗證與理解規格。
- 盲目追求 100% 測試覆蓋率可能產生無效測試,需用真實系統驗證。
- 透過架構決策記錄將人類審查前置,讓 AI 後續驗證程式碼一致性。
章節
依話題轉折切分,標題由 AI 產生
- 00:00Introduction
- 00:22Shachar Azriel, Baz: the spec, the recording, the same bug
- 02:01Why coding agents optimize for generating, not verifying
- 02:38Justin Cormack, ex-Docker: 350,000 lines of Rust with agents
- 03:01Chasing 100% test coverage, and why it didn't help
- 04:40Using S3 as a test oracle to lock down behavior
- 04:59Ian Thomas, Meta: can we trust the code?
- 06:28The anti test slop initiative, judging tests with AI
- 07:01Christopher Batey: what to review in a 7,000-line PR
- 08:30Architecture decision records as the human checkpoint
提到的工具與公司
- Baz
- Docker
- Rust
- S3
- Workplace
- AI DevCon
適合誰看
正在開發或使用 AI 編碼工具的軟體工程師與技術管理者。
摘要依據
- 依據
- 人工字幕
為什麼排在這裡
- 人氣
- 0.31
- 新鮮
- 0.90
在主題頁與搜尋結果裡,名次由相關、人氣、新鮮三個分數決定;這一頁沒有搜尋的關鍵字,所以沒有相關分數。排序怎麼算
相關內容
- Agentic Engineering and the Lost Art of VerificationPodcast ・ Vanishing Gradients ・ 1 小時 32 分
- Bob大叔:AI代码我完全不看 | Robert C. Martin | AI编程 | AI Agent | 代码整洁之道 | Clean Code | 变异测试 | 测试驱动开发 | TDD影片 ・ Best Partners TV ・ 16 分鐘(在新分頁開啟原站)
- The State of Agentic Data SciencePodcast ・ Vanishing Gradients ・ 59 分鐘
- Specs, Tests, and Self‑Verification: The Playbook for Agentic Engineering TeamsPodcast ・ AI Engineering Podcast ・ 1 小時 6 分
- Day 13 - 【實戰】放手修一個 bug文章 ・ 高見龍
- 我的 AI 编程全流程:如何使用 AI 稳定交付一个高质量的产品影片 ・ 马克的技术工作坊 ・ 25 分鐘
摘要由 AI 根據原文產生,可能有誤;完整內容請看原站。看影片(在新分頁開啟原站)
