看影片(在新分頁開啟原站)連到 YouTube・程序员老王
摘要
探討 AI 安全對齊是否只是模型演繹的幻覺,透過實驗展示聰明模型未必安全。觀眾能了解當前 AI 對齊技術的實際侷限與風險。
The video discusses whether AI safety alignment is just a hallucination, showing through experiments that smart models are not necessarily safe.
這筆內容還沒有取得字幕或內文,這段摘要只根據標題與說明欄產生,可能不夠準確;實際內容請以原站為準。
適合誰看
對 AI 安全、模型行為與對齊技術感興趣的開發者或研究者。
摘要依據
- 依據
- 標題與說明欄(還沒有取得字幕或內文)
為什麼排在這裡
- 人氣
- 0.30
- 新鮮
- 0.54
在主題頁與搜尋結果裡,名次由相關、人氣、新鮮三個分數決定;這一頁沒有搜尋的關鍵字,所以沒有相關分數。排序怎麼算
相關內容
- Lessons from the hacks文章 ・ Nathan Lambert(部落格)
- #251 – The UK's former head AI safety scientist on how to solve alignment before superintelligence arrives | Geoffrey IrvingPodcast ・ 80,000 Hours Podcast ・ 2 小時 2 分
- Science of Misalignment影片 ・ Neel Nanda ・ 50 分鐘
- Understanding the inner thoughts of AI影片 ・ Google DeepMind ・ 53 分鐘
- OpenAI Security: Controlling Models is Now ‘Hell’影片 ・ AI Explained
- AI Safety...Ok Doomer: with Anca DraganPodcast ・ Google DeepMind: The Podcast ・ 38 分鐘
摘要由 AI 根據標題與說明欄產生(還沒有取得原文),可能有誤;完整內容請看原站。看影片(在新分頁開啟原站)
