duan

HuggingFace 每日AI论文速递

每天10分钟,带您快速了解当日HuggingFace热门AI论文内容。每个工作日更新,欢迎订阅。📢播客节目在小宇宙、Apple Podcast平台搜索【HuggingFace 每日AI论文速递】🖼另外还有图文版,可在小红书搜索并关注【AI速递】

Author

duan

Category

Technology

Podcast website

www.xiaoyuzhoufm.com

Latest episode

Jul 10, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

【周末特辑】2月第2周最火AI论文 | 1B LLM如何超越405B LLM;金融领域长上下文QA基准测试 15.02.2025

本期的 5 篇论文如下: [00:54] TOP1(🔥121) | 🤔 Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling(10亿参数LLM能否超越4050亿参数LLM?重新思考计算最优的测试时缩放) [03:41] TOP2(🔥119) | 🚀 InfiniteHiP: Extending Language Model Context Up to 3 Million Tokens on a Single GPU(InfiniteHiP:在单个GPU上扩展语言模型上下文至300万 tokens) [06:11] TOP3(🔥117) | 💼 Expect the Unexpec...

2025.02.14 | GPU扩展至300万tokens,文本编码器内存高效策略。 14.02.2025

本期的 18 篇论文如下: [00:21] 🚀 InfiniteHiP: Extending Language Model Context Up to 3 Million Tokens on a Single GPU(InfiniteHiP:在单个GPU上扩展语言模型上下文至300万 tokens) [01:07] 🖼 Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation(Skrr:跳过并重用文本编码器层以实现内存高效文本到图像生成) [01:49] 🧠 An Open Recipe: Adapting Language-Specific LLMs to a...

2025.02.13 | 多语言评估工具填补空白,密集文本图像数据集挑战生成模型。 13.02.2025

本期的 20 篇论文如下: [00:23] 🌍 BenchMAX: A Comprehensive Multilingual Evaluation Suite for Large Language Models(BenchMAX:大型语言模型的综合多语言评估套件) [01:08] 📄 TextAtlas5M: A Large-scale Dataset for Dense Text Image Generation(TextAtlas5M:用于密集文本图像生成的大规模数据集) [01:48] 🎥 Light-A-Video: Training-free Video Relighting via Progressive Light Fusion(光影视频:基于渐进光融...

2025.02.12 | 强化学习提升编程竞赛,代码输入输出优化推理模型。 12.02.2025

本期的 21 篇论文如下: [00:25] 🧠 Competitive Programming with Large Reasoning Models(使用大型推理模型进行编程竞赛) [01:03] 🧠 CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction(代码输入输出:通过代码输入输出预测凝练推理模式) [01:47] 🎥 Magic 1-For-1: Generating One Minute Video Clips within One Minute(魔幻1对1:在一分钟内生成一分钟视频片段) [02:27] 🧠 Teaching Language...

2025.02.11 | LLMs生成多语言去毒数据,强化学习提升数学推理效率。 11.02.2025

本期的 21 篇论文如下: [00:25] 🤖 SynthDetoxM: Modern LLMs are Few-Shot Parallel Detoxification Data Annotators(SynthDetoxM:现代大语言模型是少样本并行去毒化数据标注器) [01:10] 🧠 Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning(探索数学推理中结果奖励的学习极限) [01:55] 🤔 Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling(10亿参数LLM能否超越405...

2025.02.10 | 视频处理性能提升,视频生成速度显著加快。 10.02.2025

本期的 21 篇论文如下: [00:22] 🎥 VideoRoPE: What Makes for Good Video Rotary Position Embedding?(视频旋转位置嵌入:什么使得视频旋转位置嵌入有效?) [01:07] 🎥 Fast Video Generation with Sliding Tile Attention(基于滑动瓦片注意力的快速视频生成) [01:54] 🎥 Goku: Flow Based Video Generative Foundation Models(悟空:基于流的视频生成基础模型) [02:35] 🌍 AuraFusion360: Augmented Unseen Region Alignm...

【周末特辑】2月第1周最火AI论文 | OmniHuman提升动画模型性能,SmolLM2优化小型语言模型训练。 08.02.2025

本期的 5 篇论文如下: [00:39] TOP1(🔥162) | 🤖 OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models(OmniHuman-1:重新思考单阶段条件式人体动画模型的放大) [02:42] TOP2(🔥137) | 🤖 SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model(SmolLM2:当小型模型走向大型化——以数据为中心的小型语言模型训练) [04:42] TOP3(🔥108) | 🤔 The Differences B...

2025.02.07 | 特征流提升模型可解释性,超IF增强指令跟随能力。 07.02.2025

本期的 21 篇论文如下: [00:24] 🔄 Analyze Feature Flow to Enhance Interpretation and Steering in Language Models(分析特征流以增强语言模型的解释与控制) [01:03] 🤖 UltraIF: Advancing Instruction Following from the Wild(超IF:从野外推进指令跟随) [01:40] 🎥 DynVFX: Augmenting Real Videos with Dynamic Content(DynVFX:用动态内容增强真实视频) [02:16] 🌐 Ola: Pushing the Frontiers of Omni-Modal Lang...

2025.02.06 | 数据优化提升模型性能,模拟市场再现复杂行为。 06.02.2025

本期的 10 篇论文如下: [00:26] 🤖 SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model(SmolLM2:当小型模型走向大型化——以数据为中心的小型语言模型训练) [01:08] 🌐 TwinMarket: A Scalable Behavioral and Social Simulation for Financial Markets(双市场:一种可扩展的金融市场的行为与社会模拟) [01:45] 🧠 Demystifying Long Chain-of-Thought Reasoning in LLMs(揭秘大语言模型中的长...

2025.02.05 | 逆桥匹配蒸馏提速,视频JAM提升运动连贯。 05.02.2025

本期的 9 篇论文如下: [00:25] ⚡ Inverse Bridge Matching Distillation(逆桥匹配蒸馏) [01:02] 🎥 VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models(视频JAM:增强视频模型运动生成的联合外观-运动表示) [01:44] 🤖 ACECODER: Acing Coder RL via Automated Test-Case Synthesis(ACECODER:通过自动化测试用例合成提升编码模型) [02:25] 🧠 QLASS: Boosting Language...

2025.02.04 | DAAs性能提升,OmniHuman动画优化。 04.02.2025

本期的 20 篇论文如下: [00:26] 🤔 The Differences Between Direct Alignment Algorithms are a Blur(直接对齐算法的差异逐渐模糊) [01:07] 🤖 OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models(OmniHuman-1:重新思考单阶段条件式人体动画模型的放大) [01:48] 💡 Process Reinforcement through Implicit Rewards(基于隐式奖励的过程强化) [02:36] ⚖ Preference Leakage: A Cont...

2025.02.03 | 测试时缩放提升推理,奖励引导解码减少计算。 03.02.2025

本期的 9 篇论文如下: [00:26] 🧠 s1: Simple test-time scaling(简单的测试时缩放) [01:18] ⚡ Reward-Guided Speculative Decoding for Efficient LLM Reasoning(奖励引导的推测解码方法用于高效LLM推理) [02:00] 🧠 Self-supervised Quantized Representation for Seamlessly Integrating Knowledge Graphs with Large Language Models(自监督量化表示法用于无缝集成知识图谱与大型语言模型) [02:41] 🛡 Constitutional C...

【月末特辑】1月最火AI论文 | DeepSeek-R1强化学习提升LLM推理能力;长文本处理突破 02.02.2025

本期的 10 篇论文如下: [00:40] TOP1(🔥281) | 🧠 DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning(DeepSeek-R1:通过强化学习激励大语言模型的推理能力) [03:13] TOP2(🔥271) | ⚡ MiniMax-01: Scaling Foundation Models with Lightning Attention(MiniMax-01:基于闪电注意力机制扩展基础模型) [05:36] TOP3(🔥249) | 🧠 rStar-Math: Small LLMs Can Master Math Reasoning with Sel...

【周末特辑】1月第4周最火AI论文 | 强化学习优于监督微调,HLE挑战LLMs能力。 01.02.2025

本期的 5 篇论文如下: [00:35] TOP1(🔥53) | 🧠 SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training(监督微调记忆,强化学习泛化:基础模型后训练的比较研究) [03:02] TOP2(🔥48) | 🧠 Humanity's Last Exam(人类最后的考试) [05:21] TOP3(🔥47) | 🛡 GuardReasoner: Towards Reasoning-based LLM Safeguards(GuardReasoner:面向基于推理的LLM安全防护) [07:44] TOP4(🔥45) | 🎙 Baichu...

2025.01.31 | GuardReasoner提升LLM安全,MedXpertQA挑战医疗AI推理。 31.01.2025

本期的 8 篇论文如下: [00:25] 🛡 GuardReasoner: Towards Reasoning-based LLM Safeguards(GuardReasoner:面向基于推理的LLM安全防护) [01:04] 🩺 MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding(MedXpertQA:专家级医疗推理与理解基准测试) [01:58] 🧠 Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs(思维四处游走:关于o1类LLMs的浅思现象) [02:40] 🌐 Streaming...

2025.01.30 | 批评提升推理,AI能耗引关注 30.01.2025

本期的 5 篇论文如下: [00:25] 🧠 Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate(批评微调:学习批评比学习模仿更有效) [01:10] 🌍 Exploring the sustainable scaling of AI dilemma: A projective study of corporations' AI environmental impacts(探索AI可持续扩展的困境:企业AI环境影响的预测性研究) [01:50] 🌟 Atla Selene Mini: A General Purpose Evaluation Model(Atl...

2025.01.29 | RL泛化优,SFT稳定输出;FP4量化降成本,精度保持。 29.01.2025

本期的 8 篇论文如下: [00:26] 🧠 SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training(监督微调记忆,强化学习泛化:基础模型后训练的比较研究) [01:07] ⚡ Optimizing Large Language Model Training Using FP4 Quantization(优化使用FP4量化的超大语言模型训练) [01:47] 📚 Over-Tokenized Transformer: Vocabulary is Generally Worth Scaling(过度分词的Transformer:词汇量通常值...

2025.01.28 | Baichuan多模态模型表现优异,长上下文处理成本降低。 28.01.2025

本期的 9 篇论文如下: [00:26] 🎙 Baichuan-Omni-1.5 Technical Report(百川全能1.5技术报告) [01:03] 📚 Qwen2.5-1M Technical Report(Qwen2.5-1M 技术报告) [01:47] 🤖 Towards General-Purpose Model-Free Reinforcement Learning(面向通用无模型强化学习的研究) [02:25] 🗣 Emilia: A Large-Scale, Extensive, Multilingual, and Diverse Dataset for Speech Generation(Emilia:一个大规模、广泛、多语言和多样化的语音...

2025.01.27 | 测试复杂性提升,冗余问题待解决 27.01.2025

本期的 9 篇论文如下: [00:25] 🧠 Humanity's Last Exam(人类最后的考试) [01:06] 📊 Redundancy Principles for MLLMs Benchmarks(多模态大语言模型基准测试的冗余原则) [01:45] 🔗 Chain-of-Retrieval Augmented Generation(链式检索增强生成) [02:24] 📊 RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques(RealCritic:面向效果驱动的语言模型批评评估) [03:12] 👤 Relightable Full-...

【周末特辑】1月第3周最火AI论文 | DeepSeek-R1强化学习提升LLM推理能力,进化搜索优化复杂任务解决。 25.01.2025

本期的 5 篇论文如下: [00:37] TOP1(🔥167) | 🧠 DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning(DeepSeek-R1:通过强化学习激励大语言模型的推理能力) [02:59] TOP2(🔥95) | 🧠 Evolving Deeper LLM Thinking(演化更深层次的LLM思维) [05:07] TOP3(🔥73) | 🤔 Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training(Agent-R:通过迭代自训练使语言模型代...

2025.01.24 | SRMT提升多智能体协作能力,VideoReward优化视频生成质量。 24.01.2025

本期的 15 篇论文如下: [00:26] 🧠 SRMT: Shared Memory for Multi-agent Lifelong Pathfinding(SRMT:多智能体终身路径规划中的共享记忆) [01:05] 🎥 Improving Video Generation with Human Feedback(利用人类反馈改进视频生成) [01:40] ⚡ Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models(Sigma:查询、键和值的差分重缩放以实现高效语言模型) [02:20] 🖼 Can We Generate Images...

2025.01.23 | DeepSeek-R1强化学习提升推理能力,多智能体框架实现虚拟电影自动化 23.01.2025

本期的 9 篇论文如下: [00:24] 🧠 DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning(DeepSeek-R1:通过强化学习激励大语言模型的推理能力) [01:07] 🎬 FilmAgent: A Multi-Agent Framework for End-to-End Film Automation in Virtual 3D Spaces(FilmAgent:虚拟3D空间中的端到端电影自动化多智能体框架) [01:48] 🔄 Test-Time Preference Optimization: On-the-Fly Alignment via Itera...

2025.01.22 | Agent-R提升语言模型实时纠错能力,MMVU评估多学科视频理解专家级表现。 22.01.2025

本期的 16 篇论文如下: [00:24] 🤔 Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training(Agent-R:通过迭代自训练使语言模型代理具备反思能力) [00:59] 🎥 MMVU: Measuring Expert-Level Multi-Discipline Video Understanding(MMVU:专家级多学科视频理解的测量) [01:35] ⚖ Demons in the Detail: On Implementing Load Balancing Loss for Training Specialized Mixture-of-Expert Models(细...

2025.01.21 | GameFactory实现多样化游戏生成,VideoWorld通过视频学习复杂知识。 21.01.2025

本期的 2 篇论文如下: [00:27] 🎮 GameFactory: Creating New Games with Generative Interactive Videos(GameFactory:利用生成式交互视频创造新游戏) [01:00] 🎥 VideoWorld: Exploring Knowledge Learning from Unlabeled Videos(VideoWorld:从未标注视频中探索知识学习) 【关注我们】 您还可以在以下平台找到我们,获得播客内容以外更多信息 小红书: AI速递 在小宇宙查看该单集文稿

2025.01.20 | 思维进化提升LLM推理能力,PaSa优化学术搜索效率。 20.01.2025

本期的 9 篇论文如下: [00:28] 🧠 Evolving Deeper LLM Thinking(演化更深层次的LLM思维) [01:04] 🔍 PaSa: An LLM Agent for Comprehensive Academic Paper Search(PaSa:基于大语言模型的全面学术论文搜索代理) [01:41] 🎨 Textoon: Generating Vivid 2D Cartoon Characters from Text Descriptions(Textoon:基于文本描述生成生动的2D卡通角色) [02:18] 🤔 Multiple Choice Questions: Reasoning Makes Large Language M...

Listen to the HuggingFace 每日AI论文速递 podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.