duan
HuggingFace 每日AI论文速递
每天10分钟,带您快速了解当日HuggingFace热门AI论文内容。每个工作日更新,欢迎订阅。📢播客节目在小宇宙、Apple Podcast平台搜索【HuggingFace 每日AI论文速递】🖼另外还有图文版,可在小红书搜索并关注【AI速递】
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
【周末特辑】2月第2周最火AI论文 | 1B LLM如何超越405B LLM;金融领域长上下文QA基准测试 15.02.2025 13:16
本期的 5 篇论文如下: [00:54] TOP1(🔥121) | 🤔 Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling(10亿参数LLM能否超越4050亿参数LLM?重新思考计算最优的测试时缩放) [03:41] TOP2(🔥119) | 🚀 InfiniteHiP: Extending Language Model Context Up to 3 Million Tokens on a Single GPU(InfiniteHiP:在单个GPU上扩展语言模型上下文至300万 tokens) [06:11] TOP3(🔥117) | 💼 Expect the Unexpec...
2025.02.14 | GPU扩展至300万tokens,文本编码器内存高效策略。 14.02.2025 14:13
本期的 18 篇论文如下: [00:21] 🚀 InfiniteHiP: Extending Language Model Context Up to 3 Million Tokens on a Single GPU(InfiniteHiP:在单个GPU上扩展语言模型上下文至300万 tokens) [01:07] 🖼 Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation(Skrr:跳过并重用文本编码器层以实现内存高效文本到图像生成) [01:49] 🧠 An Open Recipe: Adapting Language-Specific LLMs to a...
2025.02.13 | 多语言评估工具填补空白,密集文本图像数据集挑战生成模型。 13.02.2025 15:23
本期的 20 篇论文如下: [00:23] 🌍 BenchMAX: A Comprehensive Multilingual Evaluation Suite for Large Language Models(BenchMAX:大型语言模型的综合多语言评估套件) [01:08] 📄 TextAtlas5M: A Large-scale Dataset for Dense Text Image Generation(TextAtlas5M:用于密集文本图像生成的大规模数据集) [01:48] 🎥 Light-A-Video: Training-free Video Relighting via Progressive Light Fusion(光影视频:基于渐进光融...
2025.02.12 | 强化学习提升编程竞赛,代码输入输出优化推理模型。 12.02.2025 15:41
本期的 21 篇论文如下: [00:25] 🧠 Competitive Programming with Large Reasoning Models(使用大型推理模型进行编程竞赛) [01:03] 🧠 CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction(代码输入输出:通过代码输入输出预测凝练推理模式) [01:47] 🎥 Magic 1-For-1: Generating One Minute Video Clips within One Minute(魔幻1对1:在一分钟内生成一分钟视频片段) [02:27] 🧠 Teaching Language...
2025.02.11 | LLMs生成多语言去毒数据,强化学习提升数学推理效率。 11.02.2025 16:26
本期的 21 篇论文如下: [00:25] 🤖 SynthDetoxM: Modern LLMs are Few-Shot Parallel Detoxification Data Annotators(SynthDetoxM:现代大语言模型是少样本并行去毒化数据标注器) [01:10] 🧠 Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning(探索数学推理中结果奖励的学习极限) [01:55] 🤔 Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling(10亿参数LLM能否超越405...
2025.02.10 | 视频处理性能提升,视频生成速度显著加快。 10.02.2025 16:01
本期的 21 篇论文如下: [00:22] 🎥 VideoRoPE: What Makes for Good Video Rotary Position Embedding?(视频旋转位置嵌入:什么使得视频旋转位置嵌入有效?) [01:07] 🎥 Fast Video Generation with Sliding Tile Attention(基于滑动瓦片注意力的快速视频生成) [01:54] 🎥 Goku: Flow Based Video Generative Foundation Models(悟空:基于流的视频生成基础模型) [02:35] 🌍 AuraFusion360: Augmented Unseen Region Alignm...
【周末特辑】2月第1周最火AI论文 | OmniHuman提升动画模型性能,SmolLM2优化小型语言模型训练。 08.02.2025 11:26
本期的 5 篇论文如下: [00:39] TOP1(🔥162) | 🤖 OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models(OmniHuman-1:重新思考单阶段条件式人体动画模型的放大) [02:42] TOP2(🔥137) | 🤖 SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model(SmolLM2:当小型模型走向大型化——以数据为中心的小型语言模型训练) [04:42] TOP3(🔥108) | 🤔 The Differences B...
2025.02.07 | 特征流提升模型可解释性,超IF增强指令跟随能力。 07.02.2025 14:57
本期的 21 篇论文如下: [00:24] 🔄 Analyze Feature Flow to Enhance Interpretation and Steering in Language Models(分析特征流以增强语言模型的解释与控制) [01:03] 🤖 UltraIF: Advancing Instruction Following from the Wild(超IF:从野外推进指令跟随) [01:40] 🎥 DynVFX: Augmenting Real Videos with Dynamic Content(DynVFX:用动态内容增强真实视频) [02:16] 🌐 Ola: Pushing the Frontiers of Omni-Modal Lang...
2025.02.06 | 数据优化提升模型性能,模拟市场再现复杂行为。 06.02.2025 8:16
本期的 10 篇论文如下: [00:26] 🤖 SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model(SmolLM2:当小型模型走向大型化——以数据为中心的小型语言模型训练) [01:08] 🌐 TwinMarket: A Scalable Behavioral and Social Simulation for Financial Markets(双市场:一种可扩展的金融市场的行为与社会模拟) [01:45] 🧠 Demystifying Long Chain-of-Thought Reasoning in LLMs(揭秘大语言模型中的长...
2025.02.05 | 逆桥匹配蒸馏提速,视频JAM提升运动连贯。 05.02.2025 7:20
本期的 9 篇论文如下: [00:25] ⚡ Inverse Bridge Matching Distillation(逆桥匹配蒸馏) [01:02] 🎥 VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models(视频JAM:增强视频模型运动生成的联合外观-运动表示) [01:44] 🤖 ACECODER: Acing Coder RL via Automated Test-Case Synthesis(ACECODER:通过自动化测试用例合成提升编码模型) [02:25] 🧠 QLASS: Boosting Language...
2025.02.04 | DAAs性能提升,OmniHuman动画优化。 04.02.2025 15:46
本期的 20 篇论文如下: [00:26] 🤔 The Differences Between Direct Alignment Algorithms are a Blur(直接对齐算法的差异逐渐模糊) [01:07] 🤖 OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models(OmniHuman-1:重新思考单阶段条件式人体动画模型的放大) [01:48] 💡 Process Reinforcement through Implicit Rewards(基于隐式奖励的过程强化) [02:36] ⚖ Preference Leakage: A Cont...
2025.02.03 | 测试时缩放提升推理,奖励引导解码减少计算。 03.02.2025 7:18
本期的 9 篇论文如下: [00:26] 🧠 s1: Simple test-time scaling(简单的测试时缩放) [01:18] ⚡ Reward-Guided Speculative Decoding for Efficient LLM Reasoning(奖励引导的推测解码方法用于高效LLM推理) [02:00] 🧠 Self-supervised Quantized Representation for Seamlessly Integrating Knowledge Graphs with Large Language Models(自监督量化表示法用于无缝集成知识图谱与大型语言模型) [02:41] 🛡 Constitutional C...
【月末特辑】1月最火AI论文 | DeepSeek-R1强化学习提升LLM推理能力;长文本处理突破 02.02.2025 24:00
本期的 10 篇论文如下: [00:40] TOP1(🔥281) | 🧠 DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning(DeepSeek-R1:通过强化学习激励大语言模型的推理能力) [03:13] TOP2(🔥271) | ⚡ MiniMax-01: Scaling Foundation Models with Lightning Attention(MiniMax-01:基于闪电注意力机制扩展基础模型) [05:36] TOP3(🔥249) | 🧠 rStar-Math: Small LLMs Can Master Math Reasoning with Sel...
【周末特辑】1月第4周最火AI论文 | 强化学习优于监督微调,HLE挑战LLMs能力。 01.02.2025 12:38
本期的 5 篇论文如下: [00:35] TOP1(🔥53) | 🧠 SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training(监督微调记忆,强化学习泛化:基础模型后训练的比较研究) [03:02] TOP2(🔥48) | 🧠 Humanity's Last Exam(人类最后的考试) [05:21] TOP3(🔥47) | 🛡 GuardReasoner: Towards Reasoning-based LLM Safeguards(GuardReasoner:面向基于推理的LLM安全防护) [07:44] TOP4(🔥45) | 🎙 Baichu...
2025.01.31 | GuardReasoner提升LLM安全,MedXpertQA挑战医疗AI推理。 31.01.2025 6:53
本期的 8 篇论文如下: [00:25] 🛡 GuardReasoner: Towards Reasoning-based LLM Safeguards(GuardReasoner:面向基于推理的LLM安全防护) [01:04] 🩺 MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding(MedXpertQA:专家级医疗推理与理解基准测试) [01:58] 🧠 Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs(思维四处游走:关于o1类LLMs的浅思现象) [02:40] 🌐 Streaming...
2025.01.30 | 批评提升推理,AI能耗引关注 30.01.2025 4:12
本期的 5 篇论文如下: [00:25] 🧠 Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate(批评微调:学习批评比学习模仿更有效) [01:10] 🌍 Exploring the sustainable scaling of AI dilemma: A projective study of corporations' AI environmental impacts(探索AI可持续扩展的困境:企业AI环境影响的预测性研究) [01:50] 🌟 Atla Selene Mini: A General Purpose Evaluation Model(Atl...
2025.01.29 | RL泛化优,SFT稳定输出;FP4量化降成本,精度保持。 29.01.2025 6:45
本期的 8 篇论文如下: [00:26] 🧠 SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training(监督微调记忆,强化学习泛化:基础模型后训练的比较研究) [01:07] ⚡ Optimizing Large Language Model Training Using FP4 Quantization(优化使用FP4量化的超大语言模型训练) [01:47] 📚 Over-Tokenized Transformer: Vocabulary is Generally Worth Scaling(过度分词的Transformer:词汇量通常值...
2025.01.28 | Baichuan多模态模型表现优异,长上下文处理成本降低。 28.01.2025 7:22
本期的 9 篇论文如下: [00:26] 🎙 Baichuan-Omni-1.5 Technical Report(百川全能1.5技术报告) [01:03] 📚 Qwen2.5-1M Technical Report(Qwen2.5-1M 技术报告) [01:47] 🤖 Towards General-Purpose Model-Free Reinforcement Learning(面向通用无模型强化学习的研究) [02:25] 🗣 Emilia: A Large-Scale, Extensive, Multilingual, and Diverse Dataset for Speech Generation(Emilia:一个大规模、广泛、多语言和多样化的语音...
2025.01.27 | 测试复杂性提升,冗余问题待解决 27.01.2025 7:13
本期的 9 篇论文如下: [00:25] 🧠 Humanity's Last Exam(人类最后的考试) [01:06] 📊 Redundancy Principles for MLLMs Benchmarks(多模态大语言模型基准测试的冗余原则) [01:45] 🔗 Chain-of-Retrieval Augmented Generation(链式检索增强生成) [02:24] 📊 RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques(RealCritic:面向效果驱动的语言模型批评评估) [03:12] 👤 Relightable Full-...
【周末特辑】1月第3周最火AI论文 | DeepSeek-R1强化学习提升LLM推理能力,进化搜索优化复杂任务解决。 25.01.2025 12:14
本期的 5 篇论文如下: [00:37] TOP1(🔥167) | 🧠 DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning(DeepSeek-R1:通过强化学习激励大语言模型的推理能力) [02:59] TOP2(🔥95) | 🧠 Evolving Deeper LLM Thinking(演化更深层次的LLM思维) [05:07] TOP3(🔥73) | 🤔 Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training(Agent-R:通过迭代自训练使语言模型代...
2025.01.24 | SRMT提升多智能体协作能力,VideoReward优化视频生成质量。 24.01.2025 10:08
本期的 15 篇论文如下: [00:26] 🧠 SRMT: Shared Memory for Multi-agent Lifelong Pathfinding(SRMT:多智能体终身路径规划中的共享记忆) [01:05] 🎥 Improving Video Generation with Human Feedback(利用人类反馈改进视频生成) [01:40] ⚡ Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models(Sigma:查询、键和值的差分重缩放以实现高效语言模型) [02:20] 🖼 Can We Generate Images...
2025.01.23 | DeepSeek-R1强化学习提升推理能力,多智能体框架实现虚拟电影自动化 23.01.2025 6:37
本期的 9 篇论文如下: [00:24] 🧠 DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning(DeepSeek-R1:通过强化学习激励大语言模型的推理能力) [01:07] 🎬 FilmAgent: A Multi-Agent Framework for End-to-End Film Automation in Virtual 3D Spaces(FilmAgent:虚拟3D空间中的端到端电影自动化多智能体框架) [01:48] 🔄 Test-Time Preference Optimization: On-the-Fly Alignment via Itera...
2025.01.22 | Agent-R提升语言模型实时纠错能力,MMVU评估多学科视频理解专家级表现。 22.01.2025 11:17
本期的 16 篇论文如下: [00:24] 🤔 Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training(Agent-R:通过迭代自训练使语言模型代理具备反思能力) [00:59] 🎥 MMVU: Measuring Expert-Level Multi-Discipline Video Understanding(MMVU:专家级多学科视频理解的测量) [01:35] ⚖ Demons in the Detail: On Implementing Load Balancing Loss for Training Specialized Mixture-of-Expert Models(细...
2025.01.21 | GameFactory实现多样化游戏生成,VideoWorld通过视频学习复杂知识。 21.01.2025 1:59
本期的 2 篇论文如下: [00:27] 🎮 GameFactory: Creating New Games with Generative Interactive Videos(GameFactory:利用生成式交互视频创造新游戏) [01:00] 🎥 VideoWorld: Exploring Knowledge Learning from Unlabeled Videos(VideoWorld:从未标注视频中探索知识学习) 【关注我们】 您还可以在以下平台找到我们,获得播客内容以外更多信息 小红书: AI速递 在小宇宙查看该单集文稿
2025.01.20 | 思维进化提升LLM推理能力,PaSa优化学术搜索效率。 20.01.2025 6:28
本期的 9 篇论文如下: [00:28] 🧠 Evolving Deeper LLM Thinking(演化更深层次的LLM思维) [01:04] 🔍 PaSa: An LLM Agent for Comprehensive Academic Paper Search(PaSa:基于大语言模型的全面学术论文搜索代理) [01:41] 🎨 Textoon: Generating Vivid 2D Cartoon Characters from Text Descriptions(Textoon:基于文本描述生成生动的2D卡通角色) [02:18] 🤔 Multiple Choice Questions: Reasoning Makes Large Language M...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.