任雨山

Seventy3

73播客,名字取材于Sheldon最喜欢的数字,内容由NotebookLM生成,每天跟随AI读AI业界论文。

Author

任雨山

Category

Technology

Podcast website

www.xiaoyuzhoufm.com

Latest episode

Jul 10, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

【第199期】LLaDA:Large Language Diffusion Models 17.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Large Language Diffusion Models Summary The provided document introduces LLaDA, a novel language model that utilizes a diffusion process rather than the conventional autoregressive method. This work challenges the long-held belief...

【第198期】CODE I/O:通过预测代码输入输出进行推理 16.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: CODEI/O: Condensing Reasoning Patterns via Code Input-Output Prediction Summary The provided research paper introduces CODEI/O, a novel method for enhancing the reasoning capabilities of large language models by training them to p...

【第197期】ReasonFlux:层级强化学习进行推理 15.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates Summary The provided research paper introduces ReasonFlux, a novel framework designed to enhance the mathematical reasoning capabilities of large language models...

【第196期】递归深度Test-Time Compute 14.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach Summary This paper introduces a novel language model architecture that enhances reasoning by iteratively processing information in a latent space rathe...

【第195期】AI大模型已经超过自我复制红线 13.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Frontier AI systems have surpassed the self-replicating red line Summary Researchers at Fudan University investigated the self-replication capabilities of frontier AI systems. Their paper presents findings that Meta's Llama3-70B-I...

【第194期】AI在经济各领域中的实际应用情况研究 12.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Which Economic Tasks are Performed with AI? Evidence from Millions of Claude Conversations Summary This research paper analyzes millions of Claude.ai conversations to provide empirical evidence of how AI is being used across the e...

【第193期】LM2:大型记忆模型 11.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: LM2: Large Memory Models Summary This paper introduces the Large Memory Model (LM2), a novel Transformer architecture enhanced with an auxiliary memory module to improve performance on tasks requiring long context and complex reas...

【第192期】Transformer架构的局限 10.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: On Limitations of the Transformer Architecture Summary This paper explores theoretical limitations of the Transformer architecture, a cornerstone of large language models. Through the lens of Communication Complexity, the authors...

【第191期】Value-Based RL可拓展性研究 09.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Value-Based Deep RL Scales Predictably Summary This research investigates the scaling properties of value-based deep reinforcement learning methods. The authors demonstrate that despite common beliefs, the performance of these met...

【第190期】LLM推理中有前景的方法综述 08.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Advancing Reasoning in Large Language Models: Promising Methods and Approaches Summary This document provides a survey of techniques aimed at improving the reasoning abilities of Large Language Models (LLMs), which often struggle...

【第189期】MaAS:优化代理超网(agentic supernet)的多智能体系统 07.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Multi-agent Architecture Search via Agentic Supernet Summary The provided text introduces MaAS, a novel framework for automating the design of multi-agent systems powered by Large Language Models. Instead of searching for a single...

【第188期】Self-MoA:多Agent会比单个Agent强吗? 06.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Rethinking Mixture-of-Agents: Is Mixing Different Large Language Models Beneficial? Summary This paper investigates Mixture-of-Agents (MoA), a method that combines outputs from different large language models (LLMs), and introduce...

【第187期】Syntriever:用合成数据训练retriever 05.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Syntriever: How to Train Your Retriever with Synthetic Data from LLMs Summary The provided research paper introduces Syntriever, a novel framework for training information retrieval systems by leveraging synthetic data generated f...

【第186期】CoAT:MCTS+memory增强推理的框架 04.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: CoAT: Chain-of-Associated-Thoughts Framework for Enhancing Large Language Models Reasoning Summary The provided research paper introduces CoAT, a novel framework designed to enhance the reasoning capabilities of large language mod...

【第185期】RAG Foundry:简化RAG的开源框架 03.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: RAG Foundry: A Framework for Enhancing LLMs for Retrieval Augmented Generation Summary The provided document introduces RAG FOUNDRY, an open-source framework designed to streamline the development and evaluation of Retrieval-Augme...

【第184期】Diffusion Planner:基于Transformer的闭环自动驾驶算法 02.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Diffusion-Based Planning for Autonomous Driving with Flexible Guidance Summary The provided research paper introduces Diffusion Planner, a novel method for autonomous driving that utilizes diffusion models to achieve human-like pl...

【第183期】慢思考滚雪球错误如何利用 01.04.2025

Seventy3:借助NotebookLM的能力进行论文解读,专注人工智能、大模型、机器人算法方向,让大家跟着AI一起进步。 进群添加小助手微信:seventy3_podcast 备注:小宇宙 今天的主题是: Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning Summary This paper examines "slow-thinking" in large language models (LLMs), where increased computation time enhances reasoning. It theor...

【第182期】庆祝更新半年文中有彩蛋 || Long CoT Reasoning in LLMs 31.03.2025

Seventy3: 用NotebookLM将论文生成播客,让大家跟着AI一起进步。 今天的主题是: Demystifying Long Chain-of-Thought Reasoning in LLMs Summary This paper investigates how large language models (LLMs) achieve long chain-of-thought (CoT) reasoning, which involves extended, step-by-step thought processes for complex tasks. The authors explore the roles of supervised fine-tuning (SFT) and reinforcement lear...

【第181期】ASAP:两阶段框架弥合仿真与现实物理之间的差距 30.03.2025

Seventy3: 用NotebookLM将论文生成播客,让大家跟着AI一起进步。 今天的主题是: ASAP: Aligning Simulation and Real-World Physics for Learning Agile Humanoid Whole-Body Skills Summary The provided research paper introduces ASAP, a novel two-stage framework designed to bridge the gap between simulated and real-world physics for humanoid robots, enabling them to perform complex, agile movements. The firs...

【第180期】LLM-AutoDiff:一个基于梯度的自动化提示工程 29.03.2025

Seventy3: 用NotebookLM将论文生成播客,让大家跟着AI一起进步。 今天的主题是: LLM-AutoDiff: Auto-Differentiate Any LLM Workflow Summary The provided research introduces LLM-AutoDiff, a novel framework for automating prompt engineering for complex Large Language Model workflows. This system extends gradient-based optimization to multi-step and cyclic LLM applications by treating textual inputs as tra...

【第179期】s1: Simple test-time scaling 28.03.2025

Seventy3: 用NotebookLM将论文生成播客,让大家跟着AI一起进步。 今天的主题是: s1: Simple test-time scaling Summary This research explores improving language model reasoning through a technique called test-time scaling, where extra computation during inference enhances performance. The authors introduce s1K, a small, high-quality dataset of reasoning problems, and budget forcing, a method to control...

【第178期】spurious forgetting:大模型的虚假遗忘 27.03.2025

Seventy3: 用NotebookLM将论文生成播客,让大家跟着AI一起进步。 今天的主题是: Spurious Forgetting in Continual Learning of Language Models Summary This paper introduces the concept of spurious forgetting in large language models during continual learning, distinguishing it from actual knowledge loss and attributing it to the disruption of task alignment. The authors demonstrate through experiments a...

【第177期】学习率Scheduler研究分析 26.03.2025

Seventy3: 用NotebookLM将论文生成播客,让大家跟着AI一起进步。 今天的主题是: The Surprising Agreement Between Convex Optimization Theory and Learning-Rate Scheduling for Large Model Training Summary This paper explores the surprising parallels between learning-rate schedules used in large model training and theoretical performance bounds from convex optimization. It demonstrates that a simple learn...

【第176期】TokenVerse:文本到图像生成的新方法 25.03.2025

Seventy3: 用NotebookLM将论文生成播客,让大家跟着AI一起进步。 今天的主题是: TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space Summary TokenVerse introduces a new method for multi-concept personalization in text-to-image generation. The technique extracts visual elements and attributes from single or multiple images using only text captions and a pre-trained diffusion mo...

【第175期】TensorLLM:使用多头自注意力提升模型能力 24.03.2025

Seventy3: 用NotebookLM将论文生成播客,让大家跟着AI一起进步。 今天的主题是: TensorLLM: Tensorising Multi-Head Attention for Enhanced Reasoning and Compression in LLMs Summary This research introduces TensorLLM, a novel framework for improving the reasoning abilities and compression of Large Language Models (LLMs) by focusing on the Multi-Head Attention (MHA) block. The method employs multi-head te...

Listen to the Seventy3 podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.