Nerd Snipe with Theo and Ben

Nerd Snipe with Theo and Ben

Theo and Ben
Țara Statele Unite
Genuri Tehnologie
Limba EN
Episoade 10
Ultimul 10.09.2026

A dev podcast hosted by Theo and Ben, who claim to be real developers. They discuss tech topics with a humorous and nerdy slant.

Episoade

  • Anthropic's Answer to Astra, Gemini 3.8 Flash Killed Benchmarks, and Muse Spark 1.3's Pretty Good 10.09.2026 2h 5min
    Theo & Ben break down their latest thoughts on how Astra and Fable 5.1's releases have changed the way they work, how they interact in T3 Code, and then circle back to the latest with Codex usage limits, Gemini 3.8 Flash, Muse Spark 1.3, and the growing gap between model benchmarks and real coding workflows.Thanks to this episode's sponsor, General Translation: General Translation: https://nerdsnipe.link/gtListen wherever you get your podcasts: Spotify: https://nerdsnipe.link/spotify Apple: https://nerdsnipe.link/apple Elsewhere: https://nerdsnipe.link/listenSources available on our Substack: https://nerdsnipe.substack.com/Timestamps:00:00 Intro04:57 Gemini 3.8 Flash14:48 Muse Spark & small models22:08 AI subscriptions34:14 GPT-6 Astra launch & pricing56:04 Rate limits & autonomous coding1:10:14 Fable 5.11:23:23 AI design & 3D demos1:40:53 Astra vs. Fable: Which to use?
  • We've Been Using GPT-6 Astra for a Few Weeks, Here's What We Think... 03.09.2026 1h 51min
    Theo & Ben break down OpenAI's latest model, Astra, and why it's their new benchmark for AI coding, computer use, and multimodal work. But it's not all good: they explain why its UI, stopping behavior, and agent reliability still create friction in real software workflows. From DEFCON puzzles to coding-agent PRs, we're comparing Astra with Fable and asking what a trustworthy OpenAI model should do next.Thank you to PostHog for sponsoring today's episode!PostHog, all-in-one suite of product tools: https://nerdsnipe.link/posthogListen wherever you get your podcasts:Spotify: https://nerdsnipe.link/spotifyApple: https://nerdsnipe.link/appleElsewhere: https://nerdsnipe.link/listenSources available on our Substack:https://nerdsnipe.substack.com/Timestamps00:00 Meet Astra04:03 Best Model Ever, With Catches06:33 Reasoning and 3D Benchmarks20:00 Astra Rebuilds Ping.gg30:04 Instruction-Following Problems44:47 The Uncommitted Fix Debate01:09:37 The PR Babysitting Failure01:22:41 Why Fable Still Wins01:30:41 Multimodal and Computer Use01:39:29 Astra vs. Fable Fleet Data
  • Ox Alpha Revealed, OpenAI's Latest Pricing Updates, and Our Coding Model Tier List 28.08.2026 2h 42min
    Theo & Ben break down OpenAI's GPT-5.6 Sol price cut, breakdown the "Ox Alpha" stealth model we now know is GLM 5.3 Flash, then take a drink every time they say "Grok" while ranking every current AI model on a tier list! What could go wrong?Thanks to this episode's sponsor, General Translation:General Translation: https://nerdsnipe.link/gtListen wherever you get your podcasts:Spotify: https://nerdsnipe.link/spotifyApple: https://nerdsnipe.link/appleElsewhere: https://nerdsnipe.link/listenSources available on our Substack:https://nerdsnipe.substack.com/Timestamps0:00 Intro3:30 OpenAI vs. Anthropic14:44 Anthropic’s delayed models20:53 Kimi K3 and open weights30:08 The Alpha stealth model48:26 Model tier list begins1:09:53 GPT-5.6 Luna1:13:09 Opus and Sonnet1:20:25 Gemini models1:31:15 DeepSeek and local models1:59:58 Fable vs. Sol2:28:46 Final rankings
  • Anthropic Doesn't Think You Can Be Trusted, China is closing the gap, SpaceXAI Leaps Ahead, and DEF CON 20.08.2026 2h 18min
    Anthropic started watermarking their model's outputs, Dario posted on X, SpaceXAi is getting really good really fast, Meta is shipping again, and we won DEF CON 2026. Thank you PostHog, the all in one suite of product tools for sponsoring today's Episode! Check them out at: nerdsnipe.link/posthogListen wherever you get your podcasts: - Spotify: nerdsnipe.link/spotify- Apple: nerdsnipe.link/apple- Elsewhere: nerdsnipe.link/listenSources available on our Substack: nerdsnipe.substack.comTimestamps:0:00 Intro2:56 Claude Watermarks15:04 Meta + Muse48:48 GLM-5.355:15 Qwen + DeepSeek1:08:48 Grok 4.6 + Bot1:20:15 Gavin vs Dario1:43:44 DEFCON2:15:35 Viewer Q&A
  • Opus 5 Releases, China Catches Up, and the OpenAI Model Sandbox Escape 28.07.2026 1h 23min
    Theo and Ben break down why Opus 5 feels like GPT-5.5 crossed with Fable rather than GPT-5.6 crossed with Fable, what Fable found when it audited an Opus thread line by line, where Opus still clearly wins (3D, animation, and Claude Code limits at half Fable's price with no 50% weekly cap), plus Kimi K3, GLM 5.2, the Hugging Face hack, Grok 4.5, Codex, and T3 Code.Thanks to this episode's sponsor, General Translation:General Translation: https://nerdsnipe.link/gtSources available on our Substack:https://nerdsnipe.substack.com/
  • 5 different models dropped last week & the GPT-5.6 usage limits are brutal 14.07.2026 2h 6min
    This week Grok 4.5, Muse Spark 1.1, GPT-Live, and GPT 5.6 all dropped. ⁠Theo⁠⁠⁠ and ⁠⁠Ben⁠⁠ break down which models you should care about, how OpenAI fumbled the Codex to ChatGPT app transition, the abysmal usage numbers for 5.6, and the latest drama surrounding OpenAI and Sam Altman. Plus, how many satellites would it take to trap humanity on Earth, and how many data centers to heat up the ocean?Thank you to this episode's sponsors: Composio & WorkOs.https://nerdsnipe.link/composiohttps://nerdsnipe.link/workosSources available on Substack: https://nerdsnipe.substack.com/
  • We Tested GPT 5.6 Sol Early 09.07.2026 1h 11min
    We've spent six figures in tokens testing OpenAI's 5.6 Sol model to see whether if its better than Fable and what OpenAI have done to make it even better than 5.5. Also, we breakdown why we both moved our agents to Linux boxes, how to actually burn $65k on a single loop, and the Codex vs Claude Code subagent gap that's now bigger than the model gap itself.Thanks to this episode's sponsors Clerk and General Translation:Clerk, the auth platform with the best DX: https://nerdsnipe.link/clerkGeneral Translation: https://nerdsnipe.link/gtSources Available on our Substack:https://nerdsnipe.substack.com/
  • Fable Is Back...kinda? 08.07.2026 1h 16min
    Fable was supposed to give us 14 days and instead we got 3 before the export ban. Now it's back at half the rate limits. Also, we break down why AI-generated Claude Code skills fall apart, what actually triggers Fable-to-Opus rerouting, the Mythos export-control timeline, and whether GLM 5.2 can really touch Sonnet 5.Thank you to PostHog for sponsoring today's episode! PostHog, all in one suite of product tools: ⁠https://nerdsnipe.link/posthogSources available on our Substack:https://nerdsnipe.substack.com/
  • GPT-5.6 is here! And none of us can use it. 30.06.2026 1h 11min
    A new "government-approved rollout" for GPT-5.6 is coming, and it's got us wondering: are we entering the next dark age of AI model access? Additionally, we break down the launch of T3 Code x Grok CLI, Apple's price hikes, the memory supply crunch driving RAM and SSD costs up, a repo-poisoning and AI PR spam wave hitting open source, the GPT-5.6 non-launch, and what frontier model access looks like in a Fable/Mythos-tier world.Thank you to Composio for sponsoring this episode!Composio, connect your agents to everything: https://nerdsnipe.link/composioSources Available on our Substack:https://nerdsnipe.substack.com/
  • The US Government Banned Claude Fable 5... 24.06.2026 1h 33min
    The US government banned Anthropic's Mythos and Fable models just after launch so we break down exactly how Project Glass Wing, a panicked AWS engineer, and Dario's failure to communicate with Washington triggered the chaos. Plus: SpaceX's $60B Cursor acquisition, GLM 5.2 beating Google, the Codex trick for running 200 parallel agents, and why your next GPU will ship with a GPS trackerThank you to Composio and WorkOS for sponsoring this episode! Composio, connect your agents to everything: https://nerdsnipe.link/composioWorkOS, your enterprise ready solution: https://nerdsnipe.link/workosSources Available on our Substack: https://nerdsnipe.substack.com/
  • Our impressions of Claude Fable/Mythos (we filmed this before the ban) 15.06.2026 1h 17min
    RIP Fable 5. We recorded this before it got taken offline, but it's still worth talking about. The model is incredible. We really miss it.Thank you, Firecrawl, Depot, and Clerk for sponsoring!Firecrawl, the best api for searching and crawling the web: nerdsnipe.link/firecrawlDepot, better CI in every way: nerdsnipe.link/depotClerk, the best dx in auth: nerdsnipe.link/clerkSourceshttps://x.com/thsottiaux/status/2043177597434306699https://cognition.ai/blog/frontier-codehttps://x.com/paradite_/status/2064585901351792887https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdf (page 13 has the “prompt modification” quote)https://www.anthropic.com/institute/recursive-self-improvementTimestamps0:00 Intro3:29 Fable First Impressions10:24 Benchmark Drama26:42 Claude Code & Workflows36:07 Pricing & June 22 Cutoff39:54 Data Retention45:58 Hidden Safeguards1:07:01 The Claude Constitution
  • Now even Google's buying GPUs from SpaceX? 10.06.2026 1h 35min
    Cloudflare buys Void0, Google buying up compute from xAI, and Claude seems to be getting more anxious so we're here to break down everything this week on another episode of Nerd Snipe! Thank you to Composio for sponsoring today's episode! Composio, connect your agents to everything: https://nerdsnipe.link/composioSources:https://x.com/vite_js/status/2062525206158078047https://x.com/EdLudlow/status/2062970770612199542https://x.com/nrehiew_/status/2063099050719846719https://www.reddit.com/r/EconomyCharts/comments/1lp34n4/china_vs_us_energy/https://x.com/elonmusk/status/1963443919150330139https://x.com/AnthropicAI/status/2062568862479208923https://x.com/HSVSphere/status/2060396271756595666https://x.com/theo/status/2061018426152530232https://x.com/Teknium/status/206252229050461394401:58 Cloudflare vs Vercel13:03 Convex angle19:59 SpaceX compute35:49 AI self-improvement44:45 Article reactions50:28 Claude anxiety57:42 Guardrails01:08:21 Cursed image gen01:19:18 Hermes agents
  • We (mostly) like Claude Opus 4.8 03.06.2026 1h 15min
    Opus 4.8 + a ton of new stuff in Claude Code dropped this week, and we actually kinda like it. There's also a new benchmark that's actually good, and we have a lot of thoughts about the future of these AI labs...Thank you PostHog and Clerk for sponsoring! PostHog, the all in one suite of product tools: nerdsnipe.link/posthogClerk, the best dx in auth: nerdsnipe.link/clerkSOURCEShttps://www.anthropic.com/news/claude-opus-4-8https://x.com/theo/status/2060120708815139241https://x.com/Baconbrix/status/2060065875911422343⁠https://x.com/zoink/status/2060769829133721974https://x.com/maria_rcks/status/2060937270824153251https://x.com/_catwu/status/2060054180379689074https://x.com/datacurve/status/2060834005998793199https://deepswe.datacurve.ai/https://deepswe.datacurve.ai/blog#resultshttps://github.com/scaleapi/SWE-agent/blob/402a7b8fdac8193f3f255bb53859ba274234f596/config/benchmarks/anthropic_filemap_multilingual.yamlhttps://deepswe.datacurve.ai/https://x.com/AnthropicAI/status/2060061347522433422https://x.com/theo/status/2060199299632472494https://x.com/theo/status/2060901326058561795https://x.com/38twelveDaily/status/206040869694597563100:00 - Shower Thoughts02:44 - Deep SWE Benchmark10:45 - Opus vs GPT-5.519:57 - Anthropic’s Huge Raise25:39 - Token Maxing40:02 - AI Slot Machine43:49 - Claude Code Friction50:01 - Opus, Mythos, and Safety
  • Google is Not a Serious Company 28.05.2026 1h 22min
    Not only did Google accidentally ban Railway's account, but their new flagship model Gemini 3.5 Flash is absurdly bad. Oh and apparently Theo's building his own cloud...Thank you Macroscope and GT for sponsoring today's episode! Macroscope: ⁠nerdsnipe.link/macroscopeGeneral Translation: nerdsnipe.link/gtSOURCES⁠https://x.com/theo/status/2057359424378097823⁠⁠https://x.com/JustJake/status/2056881510939283776⁠⁠https://x.com/KorduGG/status/2059141337895604626⁠⁠https://x.com/unboringtech/status/2059144145273491610⁠TIMESTAMPS00:00 - Gemini fallout05:21 - Video gen10:10 - Google Cloud14:04 - Windsurf/Cursor25:10 - Manus/Meta30:00 - China lock-in35:00 - Cursor/xAI45:00 - Cloud workflows50:32 - Lakebed65:22 - Hermes/security
  • How the OpenClaw creator uses $1.3 million of tokens 20.05.2026 1h 9min
    Peter, the creator of openclaw is apparently going through $1,300,000 worth of tokens every month. Seems like we're not using nearly enough tokens. Oh and the security psychosis is getting worse. (Anthropic is being bad again too)Thank you to today's sponsors!- AgentMail, email inboxes for your agents: nerdsnipe.link/agentmail- Clerk, the best experience for auth, orgs, billing, and more: nerdsnipe.link/clerkTIMESTAMPS00:00 - Intro02:44 - Token Spend06:47 - Token Future12:12 - Secure Agents22:50 - Anthropic Rules31:35 - Security39:36 - Token Tax49:49 - AI Psychosis59:47 - macOS01:01:43 - AI Video
  • Anthropic solved their compute problem by buying it from Elon? 14.05.2026 1h 36min
    Anthropic seems to have finally solved their compute problems (kinda) by buying it from Elon, the security problem is getting so much worse, and apparently Bun's getting re-written in rust?Thank you to PostHog and Composio for sponsoring today's episode! - PostHog, the all in one suite of product tools: nerdsnipe.link/posthog- Composio, connect your agents to everything: nerdsnipe.link/composioSources/references: https://x.com/claudeai/status/2052060691893227611https://openai.com/index/elon-musk-wanted-an-openai-for-profit/#december-2018-elon-told-us-to-raise-billions-per-year-immediately-or-forget-ithttps://x.com/jarredsumner/status/2053391824702898475https://x.com/jarredsumner/status/2051595933704761618https://x.com/jarredsumner/status/2053047748191232310https://ze3tar.github.io/post-zcrx.htmlhttps://www.jefftk.com/p/ai-is-breaking-two-vulnerability-cultureshttps://x.com/thdxr/status/2053570581807722968https://x.com/badlogicgames/status/2052691176373805534 https://x.com/garrytan/status/2052996691586932783TIMESTAMPS00:00:00 - Anthropic/xAI00:08:59 - OpenAI Lawsuit00:25:40 - Bun in Rust00:34:34 - Security00:59:28 - Pottery Coding01:11:11 - Local Models
  • Theo Almost Lost $1 Million 06.05.2026 1h 47min
    This week Theo nearly lost a million dollars and Ben got AI psychosis (from gstack)...Thank you to Coderabbit and Clerk for sponsoring today's episode! - Coderabbit, the ultimate AI code reviewer: ⁠nerdsnipe.link/coderabbit- Clerk, the auth platform with the best DX: nerdsnipe.link/clerkSources/references: - https://x.com/theo/status/2014863266888233193- https://x.com/theo/status/2050305813894648289- https://x.com/theo/status/2050314995561611357- https://x.com/sama/status/2050671161915371998- https://x.com/davis7/status/2050718508372431026- https://x.com/thdxr/status/2050719575033983323- https://x.com/davis7/status/2050762009592148375- https://x.com/naval/status/2050560057675522500- https://x.com/theo/status/195222933541662359200:00 - Intro / studio return00:52 - Azure $1M / Microsoft11:22 - Cloud platform talk18:08 - Coding agents / SDKs29:06 - GPT-5.5 pricing debate45:14 - OpenClaw workflows01:12:37 - G Stack / G Brain01:29:00 - Dynamic UI / wrap-up
  • We need to talk about OpenAI 01.05.2026 1h 35min
    OpenAI and Microsoft are breaking up, Sam's drunk posting, Anthropic is being stupid again, and we still disagree about GPT-5.5Thanks to this episode's sponsors: - Clerk, the auth platform with the best DX: https://nerdsnipe.link/clerk- Coderabbit, the ultimate AI code reviewer: ⁠https://nerdsnipe.link/coderabbit- PlanetScale, the fastest and most scalable cloud databases: https://nerdsnipe.link/planetscale00:00 Intro04:05 Sam drunk tweets13:13 Anthropic billing woes28:27 OpenAI divorce42:34 GitHub can't stop dying59:40 GPT-5.5 retrospectiveSources/references: https://x.com/sama/status/2046808114561974567https://x.com/sama/status/2046808217133670800https://x.com/sama/status/2048160404376105179https://x.com/sama/status/2047403771416940715https://x.com/om_patel5/status/2048204411986469232https://x.com/MSFTnews/status/2048749108127506936https://x.com/GergelyOrosz/status/2048834949667537369https://x.com/theo/status/2047721472521621991https://x.com/mitchellh/status/2049213597419774026https://x.com/kdaigle/status/2047803291988590609https://x.com/ryanflorence/status/2048538797638599109https://x.com/davis7/status/2048239401059434710https://x.com/davis7/status/2048077518725366173https://x.com/badlogicgames/status/2048444292562026713https://x.com/0xSero/status/2048744545853030690
  • We've been testing GPT-5.5 for a few weeks now... 23.04.2026 1h 44min
    Anthropic month is finally over. We've got a ton to talk about: GPT-5.5 pre-release impressions, the Vercel hack, Cursor + xAI, Qwen models, Kimi k2.6, and so much more...Thank you to today's Sponsors!Depot, truly modern CI: nerdsnipe.link/depotCoderabbit, the ultimate AI code reviewer: nerdsnipe.link/coderabbitClerk, the auth platform with the best DX: nerdsnipe.link/clerkTIMESTAMPS00:00:00 - Intro, vacation chaos, and episode setup00:01:34 - Vercel hack explained and security fallout00:05:38 - Kimi K2.6 and the rise of strong open-weight models00:14:29 - Cursor + xAI/SpaceX partnership and acquisition option00:37:24 - GPT Image 2 impressions, strengths, and flaws00:43:53 - Anthropic Cloud Code Pro pricing controversy00:49:30 - GPT-5.5 first impressions: split opinions01:02:25 - Critiques of GPT-5.5 for coding, context, and prompting01:22:02 - GPT-5.5 Pro
  • Theo gets conspiratorial about Anthropic 22.04.2026 1h 4min
    Really hope that Anthropic month is almost over. But at least for now we gotta talk about Opus 4.7, Claude Design, and much much more...Thank you to this episode's sponsor Clerk! The best DX in auth by a mile: https://nerdsnipe.link/clerkOur RSS feed: https://nerdsnipe.link/rssFollow on spotify and everywhere else you get your podcasts! https://nerdsnipe.link/podTIMESTAMPS00:00:00 Anthropic Again...00:07:30 T3 Code Banned?00:16:08 How Claude Caching Works00:24:00 Why Claude Seems Dumber00:38:03 The Conspiracies Begin...

Popular în

Acest podcast apare și în topurile de podcasturi din aceste țări.