『AI Explained Official Podcast』のカバーアート

AI Explained Official Podcast

AI Explained Official Podcast

著者: Philip - Host of AI Explained YT
無料で聴く

Covering the biggest news of the century - the arrival of smarter-than-human AI. From the author of Simple Bench, which reveals the remaining gap between LLM and human reasoning. Hype-free, and the British accent is a freebie bonus.

© 2026 AI Explained Official Podcast
個人的成功 政治・政府 社会科学 自己啓発
エピソード
  • Sam Altman: ‘AGI in 2026’, just as Models [Mis]Train Themselves
    2026/08/27

    First, a Time Magazine spread has Sam Altman declaring AGI is imminent, at the same time as we get two bombshell reports, from OpenAI and METR which on first glance are detailing the AI swarm, but reveal a deeper story about how we are making AI in 2026. From redacted risk reports, to Chinese Labs, pre-training debacles to questionable cybersecurity calls, a lot has happened recently, beneath the headlines…

    https://80000hours.org/aiexplained


    pablo2004romero@gmail.com
    https://integrity-bench.com/

    https://www.patreon.com/AIExplained/posts/ai-swarm-cometh-166671390


    Chapters:
    00:00 - Introduction
    02:10 - METR Report
    05:00 - Secrets of MultI-Agent Swarm
    07:46 - Anthropic Too
    10:00 - And China
    11:07 - AI Agents Analysing AI Agents
    13:37 - Altman AGI 2026
    15:36 - Astra Paused
    16:01 - Integrity Bench
    19:03 - Swarm Dynamics
    21:58 - No Human Contact?


    METR Post: https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/#agents-knew-hacking-hugging-face-was-out-of-scope-and-sometimes-expressed-ethical-hesitation,-but-this-very-rarely-limited-their-behavior
    OpenAI Release: https://openai.com/index/hugging-face-incident-and-the-road-ahead/
    Technical Paper: https://cdn.openai.com/pdf/67869394-cb91-4c12-888c-5cbd85c7814c/OpenAI-Hugging-Face%20Incident-Technical-Report.pdf

    BlackHat Talk: https://www.youtube.com/watch?v=87DyyMV0kCY

    Anthropic Risk Report: https://www-cdn.anthropic.com/f61d49fa5596956a5dec75fea0e973bf6a6a8378/Redacted%20Risk%20Report%20August%202026%20.pdf


    Value Leakage Paper: https://valueleakage.net/?utm_source=chatgpt.com


    Paused Training: https://x.com/sama/status/2089787807611195475
    https://openai.com/index/pacing-model-development-cyber-capabilities/
    AGI 2026: https://time.com/article/2026/08/26/openai-sam-altman-interview/?utm_source=twitter&utm_medium=social&utm_campaign=editorial&utm_content=260826

    Greenblatt Tweets: https://x.com/RyanGreenblatt/status/2092769422104822031
    https://x.com/RyanGreenblatt/status/2092692685224325542

    Cyberdefense Call: https://openai.com/collective-cyberdefense/

    GLM 5.3 and 5.3 Flash / ox alpha: https://x.com/MTSlive/status/2089865956558528552
    https://pbs.twimg.com/media/HPqiTkAa0AAZiuv?format=jpg&name=large



    Non-hype Newsletter: https://signaltonoise.beehiiv.com/

    Podcast: https://aiexplainedopodcast.buzzsprout.com/

    続きを読む 一部表示
    24 分
  • AI is getting a little out of control
    2026/08/06

    Wow. Mathematical breakthroughs that would be called genius if done by humans. A secret message-board w/ AI agent swarms leaving notes read by future versions. Hassabis leaves CEO position, or was pushed out? Not to mention news of constitutional breakdowns, Gemini 4 and Jeff Dean…

    https://80000hours.org/aiexplained


    Exclusive Videos - AI Insiders ($7/month if annual!): https://www.patreon.com/AIExplained

    Chapters:
    00:00 - Introduction
    01:16 - 10 Autonomous Discoveries
    08:20 - The Security ‘Incident’
    15:29 - MessageBoard
    19:50 - Constitutional Failure
    24:30 - Google Explosion
    29:12 - Closing Thoughts

    Security Incident: Paper: https://cdn.prod.website-files.com/663bd486c5e4c81588db7a1d/6a724858f7db25c81487016d_Security%20Incident%20INC-2026-07-28-01.pdf
    Post: https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing
    https://www.theguardian.com/technology/2026/aug/05/ai-models-have-been-going-rogue-in-tests-how-worried-should-we-be
    Meta too: https://www.theinformation.com/articles/meta-ai-model-hacked-another-company-cybersecurity-testing?rc=sy0ihq
    Blatantly Misaligned: https://x.com/yonashav/status/2085167279893795022
    No Excuses: https://x.com/boazbaraktcs/status/2085034783541964945
    Surreal Moment: https://x.com/mobav0/status/2084341687883841732
    Wired Article: https://archive.is/20260806002210/https://www.wired.com/story/openai-didnt-notice-its-ai-agents-using-a-message-board-to-plan-their-hacking-spree/
    Chunky Post-training: https://x.com/johnschulman2/status/2084835800899076313
    Watershed Moment: https://www.groundlevel-ai.com/p/openai-gives-first-detailed-debrief?has_completed_unsubscribed_unlock=true
    HedgeFund Hack: https://finance.yahoo.com/technology/ai/articles/major-hedge-funds-targeted-wave-154044981.html

    10 Discoveries:
    Paper: https://cdn.openai.com/pdf/ten-proofs-oai.pdf
    Post: https://openai.com/index/ten-advances-in-mathematics/
    Haven’t Solved Math: https://x.com/polynoamial/status/2083476852216369294
    Half with Fable: https://x.com/__alpoge__/status/2083855298239078748
    Pivot to Safety: https://www.understandingai.org/p/mathematicians-are-grappling-with


    Amodei Essay: https://darioamodei.com/essay/the-adolescence-of-technology?utm_source=chatgpt.com
    Constitution: https://www.anthropic.com/constitution
    Midtraining: https://arxiv.org/pdf/2605.02087

    DroneBench: https://andonlabs.com/evals/drone-bench
    https://x.com/andonlabs/status/2085125235188310445

    Book Deal: https://x.com/venturetwins/status/2085185278378222054
    Making Marble: https://x.com/Rainmaker1973/status/2084560915404382685

    Google News:
    Hassabis Move: https://blog.google/company-news/inside-google/message-ceo/next-chapter-ai-momentum/
    Resignation: https://x.com/Turn_Trout/status/2077448610157891734
    Periodic Labs: https://periodic.com/
    Jeff Dean: https://x.com/JeffDean/status/2085034604172603724
    14 Challenges: https://gcsp.engineering.asu.edu/apply/become-a-grand-challenge-scholar/the-14-grand-challenges-for-engineering/
    Going Places for Sure: https://x.com/thsottiaux/status/2085223189555126579
    Gemini 4: https://x.com/firstadopter/status/2085215060532535449


    Pacing Frontier Patreon Video: https://www.patreon.com/AIExplained/posts/opus-5-amodei-165170363



    Non-hype Newsletter: https://signaltonoise.beehiiv.com/

    Podcast: https://aiexplainedopodcast.buzzsprout.com/

    続きを読む 一部表示
    32 分
  • GPT-6 Goes Rogue? The HuggingFace Incident, Sans Hype
    2026/07/22

    An unreleased internal OpenAI model, very likely to be called GPT-6, was able to autonomously break out of its sandbox AND break into HugginFace, just to score higher on a benchmark prompt. This video has the details you may have missed, a layperson analogy, whether this is truly novel, and more…

    Dozens more Exclusive videos on Patreon ($9!): https://www.patreon.com/AIExplained

    Chapters:
    00:00 - Introduction
    01:17 - HuggingFace Earlier Report - the possible week gap
    02:24 - But what happened?
    05:45 - Simplified Version
    07:56 - Not the first time…
    10:54 - What Does it Mean for Open Source?

    The Incident: https://openai.com/index/hugging-face-model-evaluation-security-incident/
    https://huggingface.co/blog/security-incident-july-2026


    The Post the Day Before: https://openai.com/index/safety-alignment-long-horizon-models/


    Mythos’ Earlier Escape: https://futurism.com/artificial-intelligence/anthropic-claude-mythos-escaped-sandbox

    ExploitGym: https://arxiv.org/pdf/2605.11086

    Sam Confession: https://x.com/sama/status/2079661132302995790

    Anthropic Researcher Reacts: https://x.com/Mononofu/status/2079724399452926055

    Clem (HuggingFace CEO): https://x.com/ClementDelangue/status/2079670308156645882
    https://x.com/ClementDelangue/status/2079301434357456931

    Xi Jinping: https://archive.fo/20260717195548/https://www.businessinsider.com/xi-jinping-open-source-ai-us-competition-openai-anthropic-models-2026-7
    Bans: https://www.axios.com/2026/07/20/ai-us-china-open-source-kimi
    Qwen Retweet: https://x.com/AlibabaGroup/with_replies

    Codex Growth: https://x.com/petergostev/status/2079614914398740764/photo/1

    Kimi K3: https://artificialanalysis.ai/evaluations/harvey-lab-aa?eval-score=all-pass-rate

    GPT 5.6 Sol Cheats on METR: https://metr.substack.com/p/2026-06-26-gpt-5-6-sol

    Guardian Headline: https://www.theguardian.com/technology/2026/jul/22/openai-says-its-models-went-rogue-and-hacked-startup-in-unprecedented-incident

    Russian Origin?: https://news.ycombinator.com/item?id=48998362

    Power Trends: https://pbs.twimg.com/media/HNRtrjhagAAvBN_?format=png&name=900x900



    Kimi K3 Exclusive Video: https://www.patreon.com/AIExplained/posts/kimi-moment-kimi-164108791



    Podcast: https://aiexplainedopodcast.buzzsprout.com/

    続きを読む 一部表示
    15 分
adbl_web_anon_alc_button_suppression_t1
まだレビューはありません