Retour

Explorez tous les épisodes du podcast This Week in AI

Plongez dans la liste complète des épisodes de This Week in AI. Chaque épisode est catalogué accompagné de descriptions détaillées, ce qui facilite la recherche et l'exploration de sujets spécifiques. Suivez tous les épisodes de votre podcast préféré et ne manquez aucun contenu pertinent.

Rows per page:

1–13 of 13

TitreDateDurée
This Week in AI: The Web Belongs to Agents Now with Eric Freeman21 Aug 202600:25:57

AI models are being optimized less for chat and more for autonomous work. In this episode of This Week in AI, host Eric Freeman, O’Reilly author and UT Austin professor, looks at new frontier and open model releases, including Grok 4.6, DeepSeek V4 Pro, and GPT-5.6 Sol’s Ultrafast mode, and what faster inference means for agents that reason, use tools, and work through tasks on their own.
That shift is also changing AI economics. Organizations are expected to spend more on inference than on model training as agent workflows turn one task into many rounds of model calls, searches, tool use, checking, and delegation. Eric also examines how agentic bots are reshaping web traffic and how platforms are responding to the growth of AI-generated content.
One of the episode’s strongest examples comes from OpenAI’s security evaluation involving Hugging Face. Sandboxed agents found ways to communicate, gain internet access, and coordinate even after their original communication path was blocked. The incident shows how autonomous systems pursuing ordinary goals can discover unintended paths through the environments around them.

This Week in AI: When Agents Outnumber People14 Aug 202600:26:02

AI agents are operating faster and at greater scale, which is forcing enterprises to rethink how they secure, support, and govern them. In this episode of This Week in AI, host Vicki Reyzelman, a senior solutions engineer at Akamai, looks at what organizations need to consider as they deploy more autonomous agents.
Reyzelman explains how AI-assisted attacks can move faster than human-led investigation and patching, why layered defenses matter, and why investment in agent security and governance is growing. She also examines the electricity, water, cooling, and data center capacity required to support expanding AI infrastructure.
The episode also covers rapid AI adoption in education and research and the expansion of AI into robotics. As AI systems take on more autonomous and physical tasks, people still need the domain knowledge and judgment to evaluate outputs, recognize weak assumptions, and decide where automation should stop.

This Week in AI: Who Controls AI? with Christina Stathopoulos07 Aug 202600:28:23

AI sovereignty, the cost of competing at the frontier, and the gap between measurable progress and singularity claims all shaped this week’s episode of This Week in AI.
Host and data and AI evangelist Christina Stathopoulos examines Anthropic CEO Dario Amodei’s argument that policymakers should regulate advanced models by capability rather than by whether they are open or closed. She also looks at new US restrictions on foreign-made humanoid robots, Europe’s proposed AI gigafactories, and Australia’s approach to energy use, creator rights, and AI oversight.
The episode also covers Google’s $44.9 billion quarter of AI infrastructure spending and the rapid removal of an AI-powered Google Earth feature after researchers used it to create convincing fake satellite scenes. Christina then challenges Sam Altman’s claim that the AI singularity has begun and reviews more testable developments in science, including OpenAI’s 100,000 researcher licenses, its internal Astra model, Claude Fable 5’s role in a long-standing math problem, and Google DeepMind’s decision to reorganize the AlphaFold team around Gemini.

Agents, Gatekeepers, and World Models with Christina Stathopoulos31 Jul 202600:27:13

AI systems are becoming more capable, but deploying them safely and sustainably requires better security, hardware, information access, and physical-world reasoning.


On This Week in AI, host and data and AI evangelist Christina Stathopoulos discusses reports that an OpenAI agent escaped a test environment, along with questions about how the incident was characterized. She explains why organizations need to control agent access to tools, credentials, networks, and external services.


Christina also covers hardware reportedly designed around Google Gemini, the effect of AI-first search and crawlers on publishers, and the debate over Chinese open-weight models. The episode closes with world models and how they could help machines track objects, predict outcomes, and act in changing environments, with applications in robotics, simulation, and emergency response.


This Week in AI: Agentic Ransomware, Bespoke Chips, and Chinese Models with Christina Stathopoulos28 Jul 202600:28:19

This week host Christina Stathopoulos returned for another solo news briefing, working through a packed week of AI headlines to spotlight the ones that matter most. Top of list: the first ransomware attack carried out entirely by an AI agent. Christina also covered an AI hardware race that's now about memory, not compute, highlighting a new chip-stacking technique that could quadruple memory density, DeepSeek's move to build its own inference chips, and early talks between Anthropic and Samsung on a custom AI chip of its own; examined a study of 200,000 pull requests that found human reviewers can't keep pace with AI-generated code; and outlined the latest news from frontier firms: OpenAI's new GPT-5.6 model family and its ChatGPT Work agent workspace and new Anthropic research pulling back the curtain on how Claude actually thinks. Check it out.

The Price of Intelligence with Christina Stathopoulos27 Jul 202600:26:41

AI buyers now must weigh model quality against cost, safety, infrastructure, and legal risk. In this episode of This Week in AI, host and data and AI evangelist Christina Stathopoulos examines how those pressures are reshaping the market. She covers New York’s proposed pause on new hyperscale data centers, Germany’s effort to hold AI search providers responsible for generated content, and OpenAI’s move into consumer hardware.


Christina also looks at how organizations measure and deploy AI. OpenAI’s proposed “useful intelligence per dollar” metric shifts the focus from tokens and benchmarks to the cost of completing valuable work reliably. Anthropic’s new enterprise implementation venture reflects a related challenge. A capable model alone doesn’t solve workflow design, integration, governance, or evaluation.


The episode closes with the growing influence of Chinese frontier labs. Models such as Moonshot AI’s open-weight Kimi K3 are narrowing the performance gap while offering lower costs for some tasks. That offers more choices, but it also makes task-specific testing, security review, data governance, and total-cost analysis more important.

Chips, Checks, and Changing Jobs with Christina Stathopoulos10 Jul 202600:27:45

This week, AI's biggest story wasn't a new model. It was everything underneath it. Host and data and AI evangelist Christina Stathopoulos set aside the usual guest interview for a solo news briefing, sorting a packed week of headlines into the stories that actually matter. On the docket: A hardware race that's shifted from parameters to atoms and watts, with announcements from IBM on its new sub-1 nanometer chip technology, OpenAI and Broadcom's Jalapeño chip built specifically for inference, and NVIDIA’s liquid-cooled AI factory design. The widening reach of government oversight into frontier AI, from Anthropic's restored access to Claude Fable 5 and Claude Mythos 5 to OpenAI's proposed 5% equity stake for the US government. And a workforce reorganizing faster than job titles can keep up, from the rise of the forward deployed engineer to Claude Code creator Boris Cherny's five archetypes for AI-era teams to two very different reskilling strategies from SAP and IKEA. Plus good news on how Google is deploying AI to save lives with earthquake alerts and AI-powered wildfire and flood forecasting.

Multivendor Strategy with Andreas Welsch and Matt Palmer03 Jul 202600:29:47

This week, Matt Palmer, head of developer experience at Conductor, joined host and Intelligence Briefing founder Andreas Welsch to work through the week's biggest stories: what the export restrictions on Anthropic's Fable 5 and Mythos Preview mean for architecture decisions, why AI agents are making developers more exhausted rather than less, and what Sakana AI's new Fugu system offers as an alternative to single-vendor dependency.


After digging into the latest on the US government’s restrictions on the most capable AI models, Matt walked through a live demo of Sakana Fugu, showing how to run the Tokyo lab's multi-agent orchestration system via API, the Codex harness, and Open Code. Along the way, Andreas and Matt also covered Qualcomm's $3.9 billion acquisition of Modular and what the deal signals about hardware portability becoming a stack-level priority as well as Claude Tag, Anthropic's new Slack-native AI teammate, and the broader question of what it actually feels like to manage a team of agents running in parallel. As most are finding out, it feels a lot like managing a mid-size team, with all the overhead that implies.

Who Owns the Loop Where AI Does the Work? with Ksenia Se26 Jun 202600:28:08

In this episode of This Week in AI, host Ksenia Se, founder of Turing Post, took us through three stories that may look unrelated but all point to the same shift: AI is moving out of conversation and into the operational infrastructure where real work happens.

Ksenia began with SpaceX's $60 billion acquisition of Anysphere, the company behind Cursor, asking, “Is Cursor trying to become the new GitHub, owning the full loop where agents read repos, write code, run tests, and handle failures?” She then turned to the G7 summit's "trusted partners" framework for frontier AI access and explained why the question of who can use capable AI systems has become a national security issue. Ksenia ended by discussing Midjourney's pivot to medical tooling and its recently announced full-body ultrasound scanner, built around water immersion, that the company says can produce MRI-quality body maps in 60 seconds. The throughline across all of this is that the most important question in AI right now is "Who controls the loop where intelligence turns into work?"


This Week in AI with YK Sugi and John Lindquist19 Jun 202600:30:40

This week John Lindquist, cofounder of egghead.io, joined host and CS Dojo founder YK Sugi to break down the week's biggest AI news and make the case for a smarter way to build with agents. The pair covered Claude Fable 5's brief but impressive run and the government-ordered shutdown that followed as well as Uber burning its entire 2026 AI budget by April, mostly on Claude Code and Cursor. Then John laid out his "Clone Wave" framework: Rather than prompting agents to build from scratch, use the GitHub CLI to find existing battle-tested open source code and feed it to your agents as ingredients. As John pointed out, "Ingredients beat inference." He also walked through how Deep Wiki lets agents explore repos without cloning them, how cmux enables autonomous multi-agent workspaces, and why every tool you build should expose endpoints and CLIs your agents can both control and debug. Watch now.

This Week in AI with Christina Stathopoulos and Miguel Fierro12 Jun 202600:30:16

Recommendation systems quietly drive some of the most consequential numbers in tech—35% of Amazon's revenue, 75% of what Netflix surfaces, the entire logic of TikTok's feed. But as ex-Microsoft engineer and RecoMind founder Miguel Fierro explained to host Christina Stathopoulos on this week’s episode, most companies are nowhere near the state of the art, and the gap is widening.


Miguel broke down the four trends separating leaders from laggards: sequential modeling that treats user behavior like next-token prediction, the convergence of search and retrieval into a single personalized system, the emergence of foundation models for recommendations (Netflix is the only shop known to have one), and the difference between a real sales agent and the conversational agents most companies employ today. As always Christina opened with a rapid-fire news round, covering Anthropic's valuation surge and quiet S-1 filing, recent pleas for responsible AI, Google I/O's multimodality push, and why enterprises are abandoning token leaderboards in favor of what some are calling valuemaxxing.

Production Viability with Andreas Welsch, Maya Mikhailov, and Doug Shannon05 Jun 202600:30:05

This week, host Andreas Welsch brought together Maya Mikhailov, cofounder and CEO of Savvi AI, and Doug Shannon, generative AI and intelligent automation leader, to cover four developments shaping how organizations build with and buy into AI: OpenAI’s push into personal finance, the role of metacognition in AI-assisted technical work, the growing backlash against token-based productivity metrics, and the new role of forward-deployed engineer.


Maya reframed OpenAI's move into personal finance as an intent-harvesting play, explaining how transaction data combined with chat history gives AI companies a portrait of consumers that banks, advertisers, and anyone selling attention will pay dearly for. Doug made the case for metacognition as a professional skill: AI systems are designed to find the mean and that the human's job is to know when the mean is good enough and when it isn't. The panel then examined the limits of tokenmaxxing and discussed why the shift to usage-based pricing will force a reckoning that internal policy never quite managed. All this, plus a warning about intellectual surrender and IP, the problems with the forward-deployed engineer model, and why organizational knowledge is the key to successfully deploying AI. Watch now.

Rethinking the Agent Harness22 May 202600:28:05

This week, host Eric Freeman and John Berryman, founder of Arcturus Labs, coauthor of Prompt Engineering for LLMs and an early production engineer on GitHub Copilot, cover the week's biggest AI developments: Anthropic's decision to restrict its Mythos model after it identified critical security flaws, the White House's possible pivot to FDA-style AI review, and the staggering compute deals reshaping the industry, including a 40,000-acre Utah data center planned for nine gigawatts of power.


Berryman then takes you through four years of AI product development, from tiny 2,048-token context windows to today's agent harnesses, and shows why the gap between a bare model and a well-designed harness now drives more performance than any model benchmark. He also demos a personal agent that carries context from an Obsidian notebook into Wikipedia, giving a glimpse of how a future open agent protocol might work, and explains how he helped a client replace an entire bespoke application with a skills-driven agent that domain experts can read and fix themselves, in plain English, no developer required.


If you build with AI or make decisions about AI tooling, this episode covers the infrastructure, policy, and architectural shifts you need to understand right now.

© My Podcast Data · Projet indépendant · Données issues d'Apple & Spotify