AI Explained
AI Explained
0
AI Explained covers the latest developments in artificial intelligence, focusing on the arrival of smarter-than-human AI. The host, creator of Simple Bench, examines the remaining reasoning gap between humans and large language models. He is also the solo developer of LM Council. The podcast includes discussions and points listeners to exclusive videos, a newsletter, and community resources.
Avsnitt
-
'Sparks of AGI' - Bombshell GPT-4 Paper: Fully Read w/ 15 Revelations 06.10.2026 17minLess than 24 hours ago a paper was released that will echo around the world. I read all 154 pages in one sitting. The paper suggests GPT 4 has ‘sparks of Artificial General Intelligence’. This is not just hype, I go through 15 examples detailing just what exactly the unrestrained GPT 4 is capable of.Insane highlights include the monumental ability to use tools effectively – this is an emergent capability not found in ChatGPT. I detail the kind of tools it has already demonstrated it can use, from using external APIs to being a true personal assistant, from a Fermi answerer to a Mathlete and a handyman. This paper may well change your thoughts on the state of AGI.That is just touching on the multitude of implications of this bombshell paper, which was originally titled 'First Contact'...Sparks of AGI Paper: https://arxiv.org/pdf/2303.12712.pdf3D Game: https://twitter.com/ammaar/status/1637592014446551040Augmented Memory: https://arxiv.org/pdf/2301.04589.pdfAdept AI Photoshop: https://www.adept.ai/blog/introducing-adepthttps://www.patreon.com/AIExplained Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
How Not to Read a Headline on AI (ft. new Olympiad Gold, GPT-5 …) 06.10.2026 23minGPT-5 did what? OpenAI ahead of Google? There are 9 ways to misread the headlines of the last 48 hours, so this video is here to tell you what happened, sans sizzle. It’s been a fairly momentous last few days, so let’s dive in to the International Math Olympiad Gold, GPT-5 alpha release, whether mathematicians are out of jobs, and the white collar impact by year’s end.Job Board: https://80000hours.org/aiexplainedNew Documentary on Patreon: https://www.patreon.com/posts/our-new-age-of-133960279AI Insiders ($9!): https://www.patreon.com/AIExplainedChapters:00:00 - Introduction00:18 - AI Beat Mathematicians?01:23 - OPENAI vs GOOGLE02:42 - Irrelevant to Jobs or …06:45 - White-collar jobs gone?10:26 - AI is Plateauing?12:00 - We Don’t Know the Details…14:33 - GPT-5 alpha14:54 - Nothing but Exponentials?15:53 - No Impact?Announcement: https://x.com/alexwei_/status/1946477742855532918UCLA Math Prof: https://x.com/ErnestRyu/status/1946699302308635130ChatGPT Agent: https://openai.com/index/introducing-chatgpt-agent/Livestream: https://www.youtube.com/watch?v=1jn_RpbPbEc&t=796sSystem Card: https://cdn.openai.com/pdf/839e66fc-602c-48bf-81d3-b21eacc3459d/chatgpt_agent_system_card.pdfJerry Tworek (OpenAI): https://x.com/MillionInt/status/1946556255490982022https://x.com/MillionInt/status/1946558130906968330Noam Brown Details: https://x.com/polynoamial/status/1946478249187377206Trieu Tranh Retweet: https://x.com/Mihonarium/status/1946880931723194389Neel Nanda: https://x.com/NeelNanda5/status/1946602953370173647Terence Tao: https://mathstodon.xyz/@taoSam Altman: https://x.com/sama/status/1946569252296929727METR Dev Study: https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study/Ravid Schwatz: https://x.com/ziv_ravid/status/1946378712716562605AlphaEvolve: https://deepmind.google/discover/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/https://simple-bench.com/Meta Salary: https://www.tomshardware.com/tech-industry/artificial-intelligence/abel-founder-claims-meta-offered-usd1-25-billion-over-four-years-to-ai-hire-person-still-said-no-despite-equivalent-of-usd312-million-yearly-salary$2k per month: https://www.theinformation.com/articles/openai-considers-higher-priced-subscriptions-to-its-chatbot-ai-preview-of-the-informations-ai-summit?rc=sy0ihqNon-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
GPT 4 Turbo-Charged? Plus Custom GPTS, Grok, AGI Tier List, Vision Demos, Whisper V3 and more 06.10.2026 24minA new AI Explained Custom GPT? What has the world come to. Well, let's find out, as we're gonna dive into everything from context window performance, Grok AI, Olympus, Gauss, Vision and TTS API use cases, the crazy new Runway Gen 2 update, the 1st AI Machine, and a fascinating new AGI tier list (with some questionable claims) from Google DeepMind. Plus, we'll start to investigate whether GPT-4 Turbo is actually smarter...https://www.patreon.com/AIExplainedDev Day: https://www.youtube.com/watch?v=U9mJuUkhUzk&t=1138s Runway gen 2 Update: https://twitter.com/runwayml/status/1721904845601579437Context Window: https://twitter.com/GregKamradt/status/1722386725635580292/photo/1 GPT 4 Turbo Smarter?: https://twitter.com/wangzjeff/status/1721934560919994823/photo/1https://pbs.twimg.com/media/F-dXZtsaEAAd9H6?format=png&name=smallRobert Lukoshko: https://twitter.com/Karmedge/status/1721777152658444773WebcamGPT: https://twitter.com/skalskip92/status/1721991055829348537Sports Commentary @geepytee: https://twitter.com/geepytee/status/1721705524176257296Whisper v3: https://github.com/openai/whisperGrok: https://x.ai/Amazon Olympus: https://www.businessinsider.com/amazon-is-building-ai-model-olympus-compete-with-openai-google-2023-11?r=US&IR=TSamsung Gauss: https://techcrunch.com/2023/11/08/samsung-unveils-chatgpt-alternative-samsung-gauss-that-can-generate-text-code-and-images/Physical Device 1stAI Machine: https://twitter.com/c_valenzuelab/status/1720487169997824163Cristobal Valenzuela: https://twitter.com/c_valenzuelab/status/1721963131591692744Infiniteyay: https://twitter.com/infiniteyay/status/ 1721858324839481843 https://www.patreon.com/AIExplained Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
The New Bard and AI Images, Videos, and Translations 06.10.2026 18minThe new Bard is actually useful, even if you look past the hype. I'll show you why, and cover how you can make the AI images that are going crazy viral. It's genuinely fun, and surprisingly customisable. If that wasn't enough, you'll learn about GPT 5 red-teaming, HeyGen translations, Alpha Missense, and much more!Metaculus Free: https://www.metaculus.com/ai/?utm_source=ai_explained&utm_campaign=ai_explainedFusion Art Referral Link: https://quickqr.art/app/fusion-art?via=ai-explainedNew Bard Extensions: https://bard.google.com/chatBard Extension Tweet: https://twitter.com/JackK/status/1704109403715387667Alpha Missense: https://www.nature.com/articles/d41586-023-02943-5HeyGen Video Translate: https://labs.heygen.com/video-translateFree Image Upscaler: https://imageupscaler.com/upscale-image-4x/Stelfie Video: https://twitter.com/StelfieTT/status/17027281883658120713D Illusion: https://twitter.com/WilliamLamkin/status/1703397480606359989Runway Gen-2: https://research.runwayml.com/gen2OpenAI Red Teaming: https://openai.com/blog/red-teaming-networkhttps://www.patreon.com/AIExplained Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
GPT 5.5 Arrives, DeepSeek V4 Drops, and the Compute War Intensifies 06.10.2026 33minGPT 5.5 full analysis, plus DeepSeek V4 paper highlights, comparisons with Mythos, a vibe-coded game w/ GPT Image 2, and 50 data-points you wouldn’t get from just reading the headlines.https://80000hours.org/aiexplainedCheck out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.aiAI Insiders ($9!): https://www.patreon.com/AIExplainedChapters:00:00 - Introduction01:11 - GPT 5.5 Comparison06:04 - Mythos Marketing11:50 - Recursive Self-Improvement?14:11 - Deepseek V418:03 - VibeCode Experiment Extravaganza21:44 - The Scarce Compute EraOpenAI Benchmarks: https://openai.com/index/introducing-gpt-5-5/5.5 System Card: https://deploymentsafety.openai.com/gpt-5-5/gpt-5-5.pdfDirect Comparison: https://pbs.twimg.com/media/HGnNm5GWEAAJ1Ob?format=jpg&name=4096x4096DeepSeek Paper: https://huggingface.co/deepseek-ai/DeepSeek-V4-ProSWE Bench Pro - benchmark of choice? https://x.com/ChowdhuryNeil/status/2047416077622395025AA Omniscience: https://artificialanalysis.ai/evaluations/omniscienceVending Bench: https://x.com/andonlabs/status/2047377260412649967Opus 4.7 System Card: https://cdn.sanity.io/files/4zrzovbb/website/037f06850df7fbe871e206dad004c3db5fd50340.pdfSam Altman Drunk Phase: https://x.com/sama/with_repliesNoam Brown: https://x.com/polynoamial/status/2047387675762802998DeepSeek Compute Crunch: https://www.bloomberg.com/news/articles/2026-04-24/deepseek-unveils-newest-flagship-a-year-after-ai-breakthrough?srnd=phx-aiSpreadsheet Bench: https://x.com/nicochristie/status/2047476237464211721Pattern Recognition: https://arcprize.org/leaderboardLeader Interviews: Core Memory: https://www.youtube.com/watch?v=NCKQL0op30EKnowledge Podcast: https://www.youtube.com/watch?v=6JoUcQ1qmAcBig Tech Round 1: https://www.youtube.com/watch?v=J6vYvk7R190&t=1116sBig Tech Round 2: https://www.youtube.com/watch?v=YnoQ8RJbALw&t=8sClaude Code Limitations: https://x.com/TheAmolAvasare/status/2046724659039932830ChatGPT 5.4 for Clinicians: https://openai.com/index/making-chatgpt-better-for-clinicians/Image Arena: https://x.com/arena/status/2046670703311884548VibeCode Bench: https://www.vals.ai/benchmarks/vibe-code5.5-made Game +Seedance 2.0: https://rosemere-quest.pages.dev/Non-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
What's Up With Bard? 9 Examples + 6 Reasons Google Fell Behind [ft. Muse, Med-PaLM 2 and more] 06.10.2026 16minA definitive comparison between Bard and GPT-4 - and it doesn’t look good for Bard. I demonstrate, with nine different examples, that Bard has a serious problem. From coding to content creation, summarisation to self-tutoring, I cover all the deficiencies that will stop Bard being currently useful. I then go into 6 explanations for why Google might have fallen behind. From Attention authors leaving the ship, not defining their product, safety and Anthropic, better models and Med-PaLM, I explore all the reasons why Bard isn’t up to scratch. Even pdf scanning doesn’t work. Dumpster Fire Article: https://www.cnbc.com/2023/02/10/google-employees-slam-ceo-sundar-pichai-for-rushed-bard-announcement.htmlBard: https://bard.google.com/GPT-4: https://chat.openai.com/chat?model=gpt-4NYT Article: https://www.nytimes.com/2023/03/22/business/stock-markets-rates-fed-decision.html8 authors leaving: https://www.forbes.com/sites/richardnieva/2023/02/08/google-openai-chatgpt-microsoft-bing-ai/?sh=648e593e4de4Product Lead Comments: https://www.cnbc.com/2023/03/03/google-execs-say-in-all-hands-meeting-bard-ai-isnt-all-for-search-.htmlAnthropic Investment: https://www.theverge.com/2023/2/3/23584540/google-anthropic-investment-300-million-openai-chatgpt-rival-claudeImagen Explanation: https://imagen.research.google/Imagen Paper: https://arxiv.org/pdf/2205.11487.pdfMuse Paper: https://arxiv.org/pdf/2301.00704.pdfMed-PaLM: https://arxiv.org/abs/2212.13138Med-PaLM 2: https://www.youtube.com/watch?v=TSEXVjr6m8wPhysics: https://www.bbc.co.uk/bitesize/guides/zqs47p3/revision/5Data Advantage: https://www.forbes.com/sites/richardnieva/2023/02/08/google-openai-chatgpt-microsoft-bing-ai/?sh=648e593e4de4Virtuous Cycle: https://www.cnbc.com/2023/03/03/google-execs-say-in-all-hands-meeting-bard-ai-isnt-all-for-search-.htmlhttps://www.patreon.com/AIExplained Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
11 Major AI Developments: RT-2 to '100X GPT-4' 06.10.2026 24minThere were 11 major developments in AI in just this week, from RT-2 – a significant step on the path to robotic AGI – to ‘100x GPT 4 in 18 months’, and from the uplifting news of TranscribeGlass and AI Barbie-heimer to dramatic revelations about OpenAI in the Atlantic. We’ll also glimpse Stable Beluga 2 – the ‘first true competitor’ to ChatGPT, based on the open source Llama 2, hear about Universal Jailbreaks, learn that OpenAI has surrendered on Text Detection, plus I’ll cover the highlights of the Senate hearing on AI. https://www.patreon.com/AIExplainedChapters:0:18 – RT-22:46 – 100X GPT-43:57 – AI Video4:29 – Altman Atlantic8:41 – Jan Leike Interview10:02 – Speech Transcription + Generation11:07 – OpenAI Text Surrender11:43 – Stable Beluga 212:51 – Universal Jailbreaks14:19 – Senate testimony: Bio16:59 – Senate Testimony: SecurityRT-2: https://www.deepmind.com/blog/rt-2-new-model-translates-vision-and-language-into-actionRT-2 Paper: https://robotics-transformer2.github.io/assets/rt2.pdfRT-2 NYT: https://www.nytimes.com/2023/07/28/technology/google-robots-ai.htmlSuleyman Barron’s Interview: https://www.barrons.com/articles/ai-chatbot-siri-alexa-inflection-pi-fa1809f8OpenAI Atlantic: https://www.theatlantic.com/magazine/archive/2023/09/sam-altman-openai-chatgpt-gpt-4/674764/TranscribeGlass: https://twitter.com/igorsushko/status/1684645290132000768Barbie-heimer: https://twitter.com/terhavlova/status/1684265329277403149 Runway Gen 2: https://twitter.com/runwayml/status/1666429706932043776ElevenLabs: https://elevenlabs.io/speech-synthesisJan Leike Interview: https://www.lesswrong.com/posts/bsNXqHgiDA6dAKNun/24-superalignment-with-jan-leikeStable Beluga 2: https://stability.ai/blog/stable-beluga-large-instruction-fine-tuned-modelsOpen LLM Leaderboard: https://huggingface.co/spaces/HuggingFaceH4/open_llm_leaderboardFast Takeoff: https://www.alignmentforum.org/posts/YgNYA6pj2hPSDQiTE/distinguishing-definitions-of-takeoff#:~:text=associated%20with%20testing.-,Fast,case%20in%20Yudkowsky's%20foom%20scenario.OpenAI Text Detection Surrender: https://techcrunch.com/2023/07/25/openai-scuttles-ai-written-text-detector-over-low-rate-of-accuracy/Universal Jailbreaks: https://www.nytimes.com/2023/07/27/business/ai-chatgpt-safety-research.htmlhttps://llm-attacks.org/https://twitter.com/mustafasuleymn?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5EauthorSenate Hearing Amodei: https://www.youtube.com/watch?v=hm1zexCjELoAltman Question from Philip @ AI Explained in Melbourne: https://www.youtube.com/watch?v=7SMzkBKzsQsWired Bio Expert: https://www.wired.com/story/have-a-nice-future-podcast-15/https://www.patreon.com/AIExplained Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
Two Rival Bets on AGI: Google I/O Highlights 06.10.2026 28minThe biggest Google AI push of the year, but what is the bigger story? Why is Google pursuing a different fork in the road than OpenAI or Anthropic? https://assemblyai.com/aiexplainedWhat does Gemini 3.5 Flash mean for the near-term future of AI? Plus the highlights from a provocative new paper on AI, 8 key moments you may have missed, and the signal from 5+ hours of AI lab interviews.Check out my free app, code INSIDER15 for paid tiers: https://lmcouncil.aiAI Insiders ($9!): https://www.patreon.com/AIExplainedChapters:00:00 - Introduction00:38 - Vibes and Google Goal02:18 - Omni, again?06:57 - Taking the same road07:44 - Gemini 3 Flash12:37 - Pitching on Cost?13:55 - Agentic Task Search14:30 - 1-shot OS but jagged, negation paper20:02 - The Karpathy Moonshot Mostafa Deghani Interview: https://www.youtube.com/watch?v=Bo19sXssYXINegation Neglect Paper: https://arxiv.org/pdf/2605.13829Gemini 3.5 Flash Headline Scores: https://deepmind.google/models/model-cards/gemini-3-5-flash/Sors original AGI Path: https://www.theguardian.com/commentisfree/2024/feb/24/openai-video-generation-tool-sora-babies-ai-artificial-intelligenceHassabis Helped Set-up Anthropic: https://archive.fo/20260519070857/https://www.ft.com/content/8f2a529e-7a1b-4d8e-95be-338d0c4c98f5Intelligence to Output Speed: https://artificialanalysis.ai/models?intelligence-comparison=intelligence-vs-output-speed#intelligenceVibeCodeBench + Finance Agent: https://www.vals.ai/homeOpenAI Needs Ads: https://archive.ph/20260409123153/https://www.reuters.com/business/media-telecom/openai-projects-25-billion-ad-revenue-this-year-100-billion-by-2030-axios-2026-04-09/Anthropic Core Views: https://www.anthropic.com/news/core-views-on-ai-safetyKarpathy Move: https://x.com/karpathy/status/2056753169888334312https://www.axios.com/2026/05/19/anthropic-openai-karpathy-andrej-claudeRecursive Self-Improvement: https://www.patreon.com/posts/ineffably-smart-156866417Non-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
Google Bard - The Full Review. Bard vs Bing [LaMDA vs GPT 4] 06.10.2026 13minBard has arrived! This is the first full review, testing out its maximum capacities, with a direct comparison with GPT 4, which powers Bing. You will find out about Bard's surprising weaknesses and unexpected strengths. I also won't shy away from areas that both do well in, or do badly in. Among the many examples I will showcase include conducting basic web searches, grammar assistance, Math, theory of mind, speed and size of the models [LaMDA vs GPT-4], joke telling, using Bard for Midjourney v5 prompts, and much much more!No apparent message limits or daily caps that I could see though.To anyone wondering how I got access, sign up to the waitlist on Bard's site. I got through in under 30 minutes. https://bard.google.com/https://www.patreon.com/AIExplained Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
New Claude Opus 4.8: 15 Things You May’ve Missed 06.10.2026 29minThe ‘best’ generally available AI model just dropped, but there is plenty I bet you missed about what it is, how it performs, and what the release tells us. 15 highlights from the 244 page system card, plus private testing, leader interview and more.AI Insiders ($9!): https://www.patreon.com/AIExplainedChapters:00:00 - Introduction00:49 - Mythos in Weeks01:49 - Adaptive not necessary02:26 - Honesty?04:37 - Flagging Uncertainty04:57 - Benchmarks08:54 - Mythos will be even better10:30 - Business skillz11:15 - Model Welfare12:16 - Cyber Comparable13:10 - Misalignment Concerns16:22 - Meta Inabilities17:58 - Code flagging18:34 - Go to sleep18:50 - Fast Mode20:21 - Dynamic WorkflowsOpus 4.8 Paper: https://cdn.sanity.io/files/4zrzovbb/website/c886650a2e96fc0925c805a1a7ca77314ccbf4a6.pdfRelease: https://www.anthropic.com/news/claude-opus-4-8Chips: https://www.theinformation.com/articles/anthropic-talks-use-microsofts-ai-chips?rc=sy0ihqhttps://www.anthropic.com/news/expanding-our-use-of-google-cloud-tpus-and-serviceshttps://www.anthropic.com/news/higher-limits-spacexPatreon Vid: https://www.patreon.com/posts/re-up-anthropics-159289449GDPVal: https://artificialanalysis.ai/evaluations/omnisciencehttps://arxiv.org/abs/2510.04374Amodei Technical Debt: https://www.youtube.com/watch?v=7xco5Qd2Oo8Dynamic Workflows: https://x.com/ClaudeDevs/status/2060044853279617150https://x.com/_catwu/status/2060054180379689074/photo/1https://claude.com/blog/introducing-dynamic-workflows-in-claude-codehttps://simple-bench.com/Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.aiNon-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
GPT 4: 9 Revelations (not covered elsewhere) 06.10.2026 16minThis video will give nine more insights from the bombshell GPT 4 Technical Report. It is the companion video to the '14 Crazy Details' released the day before. In this video, I will cover:Tests on how to 'shut down' the model in the wild, and how that might go wrong. Sam Altman requesting more regulation (!), milestone moments for humanity as more benchmarks fall, a hidden constitution, why GPT 5 might already be trained and much more.Technical Report: https://cdn.openai.com/papers/gpt-4.pdfAltman Tweet: https://twitter.com/sama/status/1635136281952026625 Leaked Conversation: https://www.axios.com/2023/03/14/microsoft-google-generative-ai-officeAGI Definition: https://openai.com/blog/planning-for-agi-and-beyondSuperforecasters: https://goodjudgment.com/HellaSwag: https://arxiv.org/pdf/1905.07830.pdfH100s: https://www.nvidia.com/en-us/data-center/h100/GitHub Experiment: https://arxiv.org/pdf/2302.06590.pdfProductivity: https://economics.mit.edu/sites/default/files/inline-files/Noy_Zhang_1.pdf#:~:text=We%20examine%20the%20productivity%20effects%20of%20a%20generative,andheightens%20both%20concern%20and%20excitement%20about%20automation%20technologies.Ark Analysis: https://research.ark-invest.com/hubfs/1_Download_Files_ARK-Invest/Big_Ideas/ARK%20Invest_013123_Presentation_Big%20Ideas%202023_Final.pdfAnthropic Constitution: https://arxiv.org/pdf/2212.08073.pdfhttps://www.patreon.com/AIExplained Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
Google Takes No Prisoners Amid Torrent of AI Announcements 06.10.2026 22minGoogle just announced at least 12 things that are each worthy of a video, but here are the top I/O highlights. From Veo 3 to Deep Research now being useable, Deep Think breaking records to Gemini Diffusion, Gemini 2.5 Flash changing how AI is priced and GemmaVerse, SynthID Detector and Imagen 4. And even this intro is missing other announcements covered in the vid! And yes, they’ll be plenty of Veo 3 clips to enjoy…https://80000hours.org/aiexplainedAI Insiders ($9!): https://www.patreon.com/AIExplainedChapters:00:00 - Introduction00:48 - Veo 302:10 - Gemini 2.5 Flash03:13 - Universal Assistant03:47 - Usage Skyrockets + OpenAI dig04:51 - Gemini Pro Deep Think06:21 - Overviews and AI Mode07:26 - Deep Research Updates (new) + Jules 08:53 - Make and Deploy Apps with Gemini09:12 - Imagen 4 10:00 - Gemini Diffusion11:46 - Try It On12:17 - SynthID Detector13:30 - GemmaVerse, SignGemma, Gemma3n, medGemma14:24 - Outro + ClipsEvent: https://www.youtube.com/watch?v=o8NiE3XMPrMNtaive Audio: https://aistudio.google.com/generate-speechGemini Diffusion: https://deepmind.google/models/gemini-diffusion/#capabilities New Gemini 2.5 Flash: https://deepmind.google/models/gemini/flash/SignGemma (See end of this vid): https://www.youtube.com/watch?v=GjvgtwSOCaoDeep Think: https://blog.google/technology/google-deepmind/google-gemini-updates-io-2025/#flash-improvementsGoogle Parallel Sampling: https://www.patreon.com/posts/next-level-good-127441188Price Plans: https://blog.google/products/google-one/google-ai-ultra/Imagen 4 Benchmarks: https://deepmind.google/models/imagen/Jules: https://jules.google/SynthID Detector: https://blog.google/technology/ai/google-synthid-ai-content-detector/Veo 3 Benchmarks: https://deepmind.google/models/veo/evals/MedGemma: https://deepmind.google/models/gemma/medgemma/Build Apps: https://aistudio.google.com/appsNon-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
Claude Fable 5 - Full 319 page Breakdown 06.10.2026 41minFable 5 is out - and it’s good, very good. But beyond the splashy demos, I want to bring you the 20+ nuggets from the 319 page system card, which I read in full, all day, plus benchmarks you may not have noticed. https://assemblyai.com/aiexplainedPlus two worrying trends inside the ‘mind’ of Claude, how OpenAI counter, and the transformer inventor’s warning.Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.aiAI Insiders ($9!): https://www.patreon.com/AIExplainedChapters:00:00 - Introduction01:06 - Blocks + Better Models02:42 - Fable 5 Upgrade over Mythos Preview04:49 - ML Acceleration Bombshell07:11 - No RSI yet07:41 - Bio-capable14:51 - Creative Writing … no17:23 - Does need bug-checks18:57 - OpenAI Response19:23 - Benchmark Bonanza28:06 - Chain of Thought worrying trendFable 5 Release: https://www.anthropic.com/news/claude-fable-5-mythos-5System Card: https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdfIntelligence Explosion: https://www.patreon.com/posts/anthropic-charts-160231656Annotated: https://x.com/Miles_Brundage/status/2064500190523113816/photo/1OpenAI Counter: https://x.com/thsottiaux/status/2064572118264913923https://x.com/thsottiaux/status/2043177597434306699Double Lifespan: https://darioamodei.com/essay/machines-of-loving-graceAutomationBench: https://zapier.com/benchmarksVending Bench: https://x.com/andonlabs/status/2064429817530085804CritPt: https://critpt.com/Riemann Bench: https://surgehq.ai/leaderboards/riemann-benchGDPVal: https://artificialanalysis.ai/evaluations/gdpval-aaBluePrint Bench 2: https://andonlabs.com/evals/blueprint-bench-2MCP Atlas: https://labs.scale.com/leaderboard/mcp_atlasFutureSim: https://x.com/nikhilchandak29/status/2064676801440358774Roon Stun Lock: https://x.com/tszzl/status/2064454617568874669Noam Brown Inference Ceiling: https://x.com/polynoamial/status/2064210146558136827Isochronic Chart: https://isochronic-passage-chart.netlify.app/#nycRose Tavern: https://claude.ai/public/artifacts/2295bebe-77e6-43e2-ae94-0fe49e9a776bRedwall Game: https://redwall-mossflower.surge.sh/Risk Report: https://www-cdn.anthropic.com/097c63b5fe7dd8b14866e1f15bb1910ec713658a.pdfTransformer Inventor Warning: https://x.com/tszzl/status/2064563986914554125Non-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
Apple’s ‘AI Can’t Reason’ Claim Seen By 13M+, What You Need to Know 06.10.2026 18minWhat to make of those headlines that AI can’t reason, seen by tens of millions? I cover the Apple paper in layman’s terms, what it means and doesn’t mean, and what’s next. Thanks to Storyblocks for sponsoring this video! Download unlimited stock media at one set price with Storyblocks: https://storyblocks.com/AIExplainedPlus o3-pro and whether it is my current most-recommended model.AI Insiders ($9!): https://www.patreon.com/AIExplainedChapters:00:00 - Introduction00:57 - Viral Post + Headlines01:42 - Apple Paper Analysis08:34 - But they do Hallucinate 10:43 - Not Supercomputers11:18 - o3 Pro and Recommendations 13.7M Tweet: https://x.com/RubenHssd/status/1931389580105925115Apple Paper: https://ml-site.cdn-apple.com/papers/the-illusion-of-thinking.pdfGuardian Article: https://www.theguardian.com/technology/2025/jun/09/apple-artificial-intelligence-ai-study-collapseLisan al Gaib post: https://x.com/scaling01/status/1931854370716426246Multiplication: https://x.com/yuntiandeng/status/1836114401213989366The Illusion of the Illusion of Thinking: https://drive.google.com/file/d/1Zx9ikRj0Enc3SB4wA9HlYIlpmO_8QiUO/viewMarcus: https://www.theguardian.com/commentisfree/2025/jun/10/billion-dollar-ai-puzzle-break-downProf Rao: https://x.com/rao2z/status/1927707640223719631AI Job Headlines: https://www.nytimes.com/2025/06/11/technology/ai-mechanize-jobs.htmlhttps://www.axios.com/2025/05/28/ai-jobs-white-collar-unemployment-anthropicSky News Story: https://news.sky.com/story/can-we-trust-chatgpt-despite-it-hallucinating-answers-13380975Veo 3 Kalshi Ad: https://x.com/Kalshi/status/1932891608388681791Altman Essay: https://blog.samaltman.com/o3 Original benchmarks: https://substackcdn.com/image/fetch/f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5b8b6c44-acd6-43b3-b5c6-1a1d5c6c25e4_2486x1388.pnghttps://pbs.twimg.com/media/GfQ0bfcXQAAQt13.jpgAlpha Evolve Video: https://www.youtube.com/watch?v=RH4hAgvYSzghttps://simple-bench.com/Non-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
GPT 4: Full Breakdown (14 Details You May Have Missed) 06.10.2026 18minI just read the entire technical report on GPT 4, not just the promotional hype. And boy does it have some interesting details. I have gathered the 14 extra details that you, or at least the media, may miss from the release. The last one is more than a little wild.These include things like the training secrets, the cherry-picked bar exam stat, text-to-image breakthroughs, and some truly astounding safety checks.https://www.patreon.com/AIExplainedhttps://cdn.openai.com/papers/gpt-4.pdfhttps://www.bemyeyes.com/https://arxiv.org/pdf/2104.12756.pdfhttps://arxiv.org/pdf/2203.10244.pdfhttps://chat.openai.com/chat?model=gpt-4https://arxiv.org/pdf/2302.10329.pdfhttps://www.alignmentforum.org/posts/Aq82XqYhgqdPdPrBA/full-transcript-eliezer-yudkowsky-on-the-bankless-podcast Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
Claude Fable Blocked - 11 Quiet Details on What’s Next 06.10.2026 18minClaude Fable 5 banned, but what’s the bigger story. We go through 11 under-reported details, so you have the context to see what’s coming next for your use of AI. From whether the ban will last, what the possible motives are, what the model can actually do, and some wild over-extrapolations going on.Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.aiAI Insiders ($9!): https://www.patreon.com/AIExplainedChapters:00:00 - Introduction00:51 - Came from an Anthropic Investor ‘and other tech leaders’01:47 - Govt pressured by CEOs like Jamie Dimon03:01 - ‘Already decided’04:02 - Prompt Injection Robustness Comparison05:15 - Wellness?06:36 - “Overreach”08:17 - Anthropic Did Admit it would cause Difficulty09:32 - 90 Minutes10:02 - Equity Absence 10:31 - Lobbying and OpenAI‘Already Decided’ - https://www.theinformation.com/articles/amazons-jassy-raised-concerns-anthropic-model-trump-crackdown?rc=sy0ihqNot for Other Models: https://www.theinformation.com/briefings/u-s-government-unlikely-extend-anthropic-export-control-ai-companies?rc=sy0ihq90 Minutes: https://archive.fo/20260614001605/https://www.politico.com/news/2026/06/13/inside-the-whirlwind-24-hours-that-led-the-white-house-to-slap-export-controls-on-anthropic-00961519#selection-807.1-807.219Anthropic Statement: https://www.anthropic.com/news/fable-mythos-accessLife Comes at you Fast: https://x.com/etbrooking/status/2065638276388495742Anthropic Deputy CISO: https://x.com/TheTranscript_/status/2065883670053847324Hegseth Gloat: https://x.com/PeteHegseth/status/2065897156226015690Roon Speculation: https://x.com/tszzl/status/2065939227167392147Mythos System Card: https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdfSachs Statement: https://x.com/DavidSacks/status/2065853007619588171OpenAI Lobbying: https://thehill.com/policy/technology/5912720-altman-openai-get-bogged-down-in-political-spending-fight/Absent from Equity Talks: https://finance.yahoo.com/sectors/technology/articles/trump-ai-ownership-plan-could-131053732.htmlPliny Jaibreak: https://x.com/elder_plinius/status/2064776322979676227Fusion: https://x.com/OpenRouter/status/2065856871215329545https://lmcouncil.ai Non-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
Time Until Superintelligence: 1-2 Years, or 20? Something Doesn't Add Up 06.10.2026 17minSuperintelligence when? This question is more urgent than ever as we hear competing timelines from Inflection AI, OpenAI and the leaders at the Center for AI Safety. This video not only covers the what was said, it offers some data on what superintelligence is projected to be capable of. I discuss things that can hasten those timelines (see the new Netflix doc, Top 1% for Creativity and the AI Natural Selection paper) or slow it down (ft. Yuval Noah Hariri, the new Jailbroken paper, and more). And I end on some reflections of what it might mean to interact with a superintelligence (ft Douglas Hofstadter).Introducing Superalignment: https://openai.com/blog/introducing-superalignmentOdds of Success: https://twitter.com/janleike/with_replies?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5EauthorMustafa Suleyman (Inflection AI) Interview: https://www.youtube.com/watch?v=oAxNkehgzEUInflection Supercomputer: https://www.tomshardware.com/news/startup-builds-supercomputer-with-22000-nvidias-h100-compute-gpus GPT 2030: https://www.lesswrong.com/posts/WZXqNYbJhtidjRXSi/what-will-gpt-2030-look-likeCounterpoint: https://twitter.com/xuanalogue/status/1666765447054647297Creativity Test: https://www.umt.edu/news/2023/07/070523test.phpMMLU: https://arxiv.org/pdf/2009.03300.pdfBoston Globe CAI: https://www.bostonglobe.com/2023/07/06/opinion/ai-safety-human-extinction-dan-hendrycks-cais/Yuval Noah Hariri Guardian: https://www.theguardian.com/technology/2023/jul/06/ai-firms-face-prison-creation-fake-humans-yuval-noah-harariJailbroken Paper: https://arxiv.org/pdf/2307.02483.pdfSuleyman Hallucination Tweet: https://twitter.com/mustafasuleymn/status/1678072401760796672?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5EtweetKiller Robots: https://www.youtube.com/watch?v=YsSzNOpr9cEAutomated Apple-picking: https://twitter.com/AiBreakfast/status/167782897121428275250 year Mortgages: https://www.ft.com/content/281fbba6-28e2-42d4-b241-0f215995f0d1Heypi: https://heypi.com/talkNatural Selection Favors AIs over Humans: https://arxiv.org/pdf/2303.16200.pdf#:~:text=Natural%20selection%20may%20be%20a,designs%20to%20be%20selected%20naturally.Gödel, Escher, Bach author Doug Hofstadter on the state of AI today: https://www.youtube.com/watch?v=lfXxzAVtdpU&t=1763shttps://www.patreon.com/AIExplained Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
GPT 5 is All About Data 06.10.2026 19minDrawing upon 6 academic papers, interview snippets, possible leaks, and my own extensive research, I put together everything we might know about GPT 5: what will determine its IQ, its timeline and its impact on the job market and beyond.Starting with an insider interview on the names of GPT models, such as GPT 4 and GPT 5, then looking into the clearest hint that GPT 4 is inside Bing. Next, I briefly cover reports of a leak about GPT 5 and discuss the scale of GPUs require to train it, touching on the upgrade form A100 to H100 GPUs.Then the DeepMind paper that changed everything, focusing LLM research on data rather than parameter count. I go over a lesswrong post about that paper's 'wild implications'. And then the key paper: 'Will We Run Out of Data'. This encapsulates the key dynamic that will either propel or bottleneck GPT and other LLM improvements.Next, I examine a different take, that perhaps data is already limited and caused the Sydney model of Bing. This opens up to a discussion on the data behind these models and why Big Tech is so unforthcoming about where it originates. Could a new legal war be brewing?I then cover 4 of the ways these models may improve even without data augmentation, such as Automatic Chain of Thought, high quality data extraction, tool training, including Wolfram Alpha, retraining on existing data sets, artificial data generation and more.We take a quick look at Sam Altman's timelines and host of Big Bench benchmarks that they may impact, such as reading comprehension, critical reasoning, logic, physics and Math. I address Altman's quote about timelines being delayed by alignment and safety and finally, Altman's comments on AGI and how they pertain to GPT 5.https://www.patreon.com/AIExplainedhttps://stratechery.com/2023/new-bing-and-an-interview-with-kevin-scott-and-sam-altman-about-the-microsoft-openai-partnership/https://twitter.com/XiXiDu/status/1285225627390443521https://www.linkedin.com/pulse/building-new-bing-jordi-ribas/https://twitter.com/davidtayar5/status/1625140481016340483/photo/1https://www.techradar.com/news/chatgpt-might-bring-about-another-gpu-shortage-sooner-than-you-might-expecthttps://www.nvidia.com/en-gb/data-center/h100/https://arxiv.org/pdf/2203.15556.pdfhttps://www.lesswrong.com/posts/6Fpvch8RR29qLEWNH/chinchilla-s-wild-implicationshttps://twitter.com/Meaningness/status/1628052277050376192https://arxiv.org/pdf/2211.04325.pdfhttps://twitter.com/Meaningness/status/1628052277050376192https://www.searchenginejournal.com/google-bard-training-data/478941/#close https://arxiv.org/pdf/2302.12822.pdfhttps://arxiv.org/pdf/2302.04761.pdfhttps://www.wolframalpha.com/https://arxiv.org/pdf/2207.14502.pdfhttps://aclanthology.org/2022.findings-emnlp.508.pdfhttps://www.technologyreview.com/2022/11/24/1063684/we-could-run-out-of-data-to-train-ai-language-programs/https://twitter.com/sama/status/1625980933861175306https://github.com/google/BIG-bench/blob/main/bigbench/benchmark_tasks/results/plot_BIG-bench_lite_aggregate.pnghttps://github.com/google/BIG-bench/tree/main/bigbench/benchmark_tasks/gre_reading_comprehensionhttps://github.com/google/BIG-bench/tree/main/bigbench/benchmark_tasks/logical_argshttps://github.com/google/BIG-bench/tree/main/bigbench/benchmark_tasks/physicshttps://lifearchitect.ai/gpt-4/https://www.lesswrong.com/posts/uxnjXBwr79uxLkifG/comments-on-openai-s-planning-for-agi-and-beyondhttps://www.patreon.com/AIExplained Non-Hype, Free Newsletter: https://signaltonoise.beehiiv.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
GPT-5 has Arrived 06.10.2026 20minGPT-5 will change how hundreds of millions of people use AI. Yes, you might have to forgive the chart crimes, the underwhelming livestream and Altman hype… But it’s a good model. I have read the 50 page system card in full, have the benchmark scores, coding tests, and things you might have missed.https://app.grayswan.ai/ai-explainedAI Insiders ($9!): https://www.patreon.com/AIExplainedAnnouncement: https://openai.com/index/introducing-gpt-5/System Card: https://cdn.openai.com/pdf/8124a3ce-ab78-4f06-96eb-49ea29ffb52f/gpt5-system-card-aug7.pdfExtra Paper: https://cdn.openai.com/pdf/be60c07b-6bc2-4f54-bcee-4141e1d6c69a/gpt-5-safe_completions.pdfAltman tweet: https://x.com/sama/status/1953551377873117369Livestream: https://www.youtube.com/watch?v=0Uu_VJeVVfoMETR Report: https://metr.github.io/autonomy-evals-guide/gpt-5-report/ARC-AGI-2: https://x.com/fchollet/status/1953511631054680085Claude Opus 4.1: https://www.anthropic.com/news/claude-opus-4-1MMMU: https://mmmu-benchmark.github.io/Cursor Praise: https://x.com/ryolu_/status/1953531724895596669Non-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices -
Fable 5 vs GPT 5.6 Sol: The Early Results 05.10.2026 23minFable 5 (newly re-released) vs GPT 5.6 Sol, what comparisons can we unearth? Plus, Sonnet 5, a 5% equity seizure by US Govt, the ‘largest heist’, Beetlejuice and more…Exclusive Vids in AI Insiders ($9!): https://www.patreon.com/AIExplainedChapters:00:00 - Introduction01:06 - Fable Timeline02:59 - Sol Release?06:01 - the Chinese Angle07:35 - 5% Stake09:10 - Fable vs Sol - the numbers13:57 - Sol Misalignment15:27 - Claude Sonnet 516:24 - GLM 5.2 + New Paper on Large Model LearningGPT 5.6 Sol Release: https://openai.com/index/previewing-gpt-5-6-sol/Altman on Concentrated Power: https://x.com/giacomomiolo/status/2070609506925478352System Card: https://deploymentsafety.openai.com/gpt-5-6-preview/gpt-5-6-preview.pdfGLM 5.2: https://www.patreon.com/AIExplained/posts/glm-5-2-and-nsa-161869508Altman Memo: https://www.theinformation.com/articles/trump-administration-asks-openai-stagger-release-new-model-security-concerns?rc=sy0ihqFable 5 Revised Timeline: https://www.anthropic.com/news/redeploying-fable-5Claude Sonnet 5: https://www.anthropic.com/news/claude-sonnet-5Mythose 5 System Card: https://www-cdn.anthropic.com/9e6a1044980d8c4ed85669faf9c2a8342e2e9f1e/Claude%20Sonnet%205%20System%20Card.pdfAnthropic Accuse Alibaba Cloud / Qwen: https://www.bbc.co.uk/news/articles/cwyklykn5dwoBetelgeuse: https://upload.wikimedia.org/wikipedia/commons/6/69/Well_known_stars_2.pngOpenAI 5%: https://edition.cnn.com/2026/07/02/business/openai-trump-stake-intlWhy Larger Models Learn More: https://arxiv.org/pdf/2605.29548Patreon Post: https://www.patreon.com/AIExplained/posts/glm-5-2-and-nsa-161869508Non-hype Newsletter: https://signaltonoise.beehiiv.com/Podcast: https://aiexplainedopodcast.buzzsprout.com/ Learn more about your ad choices. Visit megaphone.fm/adchoices
Populär i
Den här podcasten finns även i podcastlistor i dessa länder.