Fourth Mind
Intelligence Unbound
Unpacking the questions shaping the next intelligence era. I am producing a fully AI-generated podcast that explores the influence of AI within various industries and examines significant technological breakthroughs.
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
NVIDIA DGX Spark and Tinker API: Localizing LLM Fine-Tuning 20.10.2025 13:25
This episode dives deep on significant shift in the AI development landscape, moving away from exclusive reliance on large, general-purpose cloud computing.
Small Fixed Samples Poison Large LLMs 15.10.2025 11:20
This episode dive deep on an Anthropic report and a related research paper, detail a joint study on the vulnerability of large language models (LLMs) to data poisoning attacks. The research surprisingly demonstrates that injecting a near-constant, small number of malicious documents—as few as 250—is sufficient to successfully introduce a backdoor vulnerability, regardless of the LLM's size (up...
Petri: An Open-Source AI Safety Auditing Tool 13.10.2025 14:02
This episode introduce Petri (Parallel Exploration Tool for Risky Interactions) , an open-source framework developed by Anthropic to accelerate AI safety research through automated auditing. Petri uses specialized AI auditor agents and LLM judges to test target models across diverse, multi-turn scenarios defined by human researchers via seed instructions.
Introducing Gemini 2.5 Computer Use Model 10.10.2025 15:33
This episode dive deep on Gemini 2.5 Computer Use model , a specialized AI model from Google DeepMind built on the Gemini 2.5 Pro architecture, designed to power agents capable of interacting with user interfaces (UIs). This model is accessible via the Gemini API for developers to create agents that can perform tasks like clicking, typing, and scrolling on web pages and applications.
AI's Impact on the Labor Market: Stability, Not Disruption (yet) 08.10.2025 14:02
This Episode dive deep on the latest article from The Budget Lab at Yale that provides an analysis of the initial impact of Artificial Intelligence (AI) on the U.S. labor market since the introduction of generative AI in November 2022. The authors conclude that despite widespread public anxiety about job losses, their data indicates no substantial, economy-wide disruption or acceleration in the...
GEM: A GYM for Agentic LLMs 07.10.2025 15:49
This episode dive deep on GEM (General Experience Maker) , an open-source environment simulator designed to accelerate research on agentic Large Language Models (LLMs) by shifting their training paradigm from static datasets to experience-based learning in complex, interactive environments. Modeled after OpenAI-Gym, GEM provides a standardized framework for the agent-environment interface, sup...
Effective Context Engineering for AI Agents 03.10.2025 12:52
This episode dive deep on Anthropic last piece on the emerging field of context engineering, which is presented as the natural evolution of prompt engineering for building effective AI agents. Context engineering focuses on curating and managing the entire set of tokens; including prompts, tools, message history, and external data... that inform a large language model (LLM) during inference, ackno...
Gemini Robotics 1.5: Embodied Reasoning and Multi-Embodiment Action 01.10.2025 13:27
This episode dives deep on the Gemini-Robotics-1-5-Tech-Report report; significant advancement in generalist robots through the introduction of the Gemini Robotics 1.5 model family. This system features two core components: Gemini Robotics 1.5 (GR 1.5), a Vision-Language-Action (VLA) model that translates instructions into robot actions and supports multi-embodiment control, and Gemini Robotics-ER...
GDPval: AI Model Performance on Economic Tasks 29.09.2025 13:51
The episode introduces GDPval, a new benchmark created by OpenAI to evaluate AI model performance on real-world, economically valuable tasks derived from the work of industry experts across the top nine sectors contributing to U.S. GDP. This evaluation covers tasks from 44 occupations and is intended to provide a more realistic assessment of AI capabilities than traditional academic benchmarks, in...
AI Assistant for Genetic Sensemaking 24.09.2025 16:47
This episode is about a study titled "AI-Enhanced Sensemaking: Exploring the Design of a Generative AI-Based Assistant to Support Genetic Professionals," which investigates integrating generative AI to assist genetic experts in diagnosing rare diseases through whole genome sequencing (WGS) analysis. The research, conducted by collaborators from Microsoft Research, Drexel University, and...
AI Tackles a Century-Old Problem in Physics by Hunting for Solutions That Shouldn't Exist 22.09.2025 12:55
This episode details a groundbreaking research effort by Google DeepMind and collaborating academic institutions, focusing on the discovery of unstable singularities in fluid dynamics using advanced AI techniques.
Small Language Models: The Future of Agentic AI 19.09.2025 18:59
This episode is about the latest Nvidia papers that advocates for the widespread adoption of Small Language Models (SLMs) over Large Language Models (LLMs) within agentic AI systems, asserting that SLMs are sufficiently powerful, more economical, and inherently more suitable for the repetitive and specialized tasks typical of such agents.
Scientific Frontiers of Agentic AI 18.09.2025 17:04
This episode dive deep on the Amazon Science article named Scientific frontiers of agentic AI. it discusses the emerging field of agentic AI, contrasting it with generative AI by emphasizing its ability to act autonomously on behalf of users by accessing and interacting with external resources.
How People Use ChatGPT 17.09.2025 15:42
This episode is about the working paper, "How People Use ChatGPT," investigates the widespread adoption and diverse applications of ChatGPT from its 2022 launch through July 2025. The authors analyze millions of de-identified user messages to understand usage patterns, finding that non-work-related interactions constitute the majority, though work-related use is significant for educated...
Anthropic Economic Index: Uneven AI Adoption 16.09.2025 21:39
This episode dive deep in the report from Anthropic that examines the rapid and geographically uneven adoption of AI, specifically Claude, across both consumer and enterprise users. It highlights that AI adoption is concentrated in higher-income regions and for certain tasks, particularly coding and administrative functions, mirroring historical patterns of technological diffusion but at an accele...
Defeating Nondeterminism in LLM Inference 15.09.2025 22:32
This episode dive deep on the Thinking Machines Lab publication that addresses the challenge of achieving reproducibility in large language model (LLM) inference, noting that even with "greedy sampling" (temperature set to 0), results are often nondeterministic.
Why Language Models Hallucinate 10.09.2025 16:11
This episode explore the phenomenon of "hallucinations" in language models, defining them as confidently generated but false statements. It argue that current training and evaluation methods inadvertently incentivize models to guess rather than admit uncertainty, comparing it to students guessing on a multiple-choice test to avoid a zero score.
The Dawn of Brain-Inspired AI: How a New Model is Redefining Reasoning Performance Beyond LLMs 08.09.2025 19:58
This episode introduce the Hierarchical Reasoning Model (HRM) , a novel AI architecture developed by Sapient Intelligence, which draws inspiration from the human brain's hierarchical and multi-timescale information processing. HRM aims to overcome the limitations of current Large Language Models (LLMs) that rely on Chain-of-Thought (CoT) techniques, which are described as inefficient and data-int...
Accelerating Life Sciences with AI: OpenAI and Retro Biosciences 03.09.2025 15:53
this episode is about a collaboration between OpenAI and Retro Biosciences to accelerate life sciences research using a specialized AI model. They developed GPT-4b micro, a miniature GPT-4o variant, for protein engineering, specifically focusing on the Yamanaka factors critical for stem cell reprogramming.
Breaking the Sorting Barrier in Shortest Paths 27.08.2025 11:55
This episode presents a deterministic algorithm for the single-source shortest path (SSSP) problem on directed graphs with non-negative edge weights, operating within the comparison-addition model. The core contribution is achieving an O(m log^(2/3) n) time complexity, which is the first to surpass Dijkstra's algorithm's O(m + n log n) bound on sparse graphs, demonstrating that Dijkstra...
Game-Generated Data: Untapped Resource for Advanced AI Training 26.08.2025 18:24
this episode is about game-generated data as an underexplored resource for training advanced AI, arguing that it can overcome critical limitations of current AI systems, such as the imminent exhaustion of high-quality text data and deficiencies in handling complex temporal or causal reasoning
The Unseen Catalysts of AI: A Journey from Dismissed Ideas to a New Renaissance 25.08.2025 15:06
This episode is a transcript of an interview with Yann LeCun, a prominent figure in AI research often called a "godfather of AI." LeCun discusses his pioneering work in neural networks and deep learning, highlighting its initial dismissal and eventual mainstream adoption through strategic efforts like placing students in major tech companies. He touches upon the evolution of AI, from its...
IBM and NASA released Surya: AI for Solar Flare Prediction 22.08.2025 11:52
this episode discuss Surya, a groundbreaking foundation model for heliophysics developed by NASA and IBM, now made open-source and available on GitHub and HuggingFace. Surya is designed to predict solar events like flares and solar wind, utilizing full-resolution data from NASA's Solar Dynamics Observatory (SDO). This AI-powered system significantly improves the lead time for forecasting space...
Is Chain-of-Thought Reasoning a Mirage? 21.08.2025 15:03
this episode is about an academic paper investigates whether Chain-of-Thought (CoT) reasoning in Large Language Models (LLMs) represents genuine logical inference or merely a superficial pattern-matching process. Researchers from Arizona State University propose a "data distribution lens" to examine this, hypothesizing that CoT effectiveness is fundamentally limited by the training data&...
Beyond Benchmarks: Redefining AI Intelligence Through Dynamic Evaluation and Cross-Industry Insights 20.08.2025 20:42
This podcast discuss the evolving landscape of AI evaluation and testing, highlighting the limitations of current benchmarks and proposing new approaches
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.