Jingwen Liang, Gengyu Wang

Daily Paper Cast

Science EN ↓ 2000 episodes

We update every weekday to discuss highest-voted papers from Huggingface Daily Paper (https://huggingface.co/papers). Both the podcast scripts and audio are generated by AI. Feedback and suggestions are welcome! Email us: dailypapercast.ai@gmail.comCreator:Jingwen Liang, 3D ML, https://www.linkedin.com/in/jingwen-liang/Gengyu Wang, LLM ML, http://wanggengyu.comListen on: Spotify: https://open.spotify.com/show/21nrhmdaA8qoBiH8q03NXLApple Podcast: https://podcasts.apple.com/us/podcast/daily-paper-cast/id1777620236Cover Image by Kawen Kuang https://kawen.art

Author

Jingwen Liang, Gengyu Wang

Category

Science

Latest episode

Jul 11, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

Complex Logical Instruction Generation 14.08.2025

🤗 Upvotes: 30 | cs. CL, cs. LG Authors: Mian Zhang, Shujian Liu, Sixun Dong, Ming Yin, Yebowen Hu, Xun Wang, Steven Ma, Song Wang, Sathish Reddy Indurthi, Haoyun Deng, Zhiyu Zoey Chen, Kaiqiang Song Title: Complex Logical Instruction Generation Arxiv: http://arxiv.org/abs/2508.09125v1 Abstract: Instruction following has catalyzed the recent era of Large Language Models (LLMs) and is the foundatio...

Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models 14.08.2025

🤗 Upvotes: 27 | cs. CL, cs. AI Authors: Wen Wang, Bozhen Fang, Chenchen Jing, Yongliang Shen, Yangyi Shen, Qiuyu Wang, Hao Ouyang, Hao Chen, Chunhua Shen Title: Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models Arxiv: http://arxiv.org/abs/2508.09138v1 Abstract: Diffusion large language models (dLLMs) generate text through iterative denoising, yet current decoding strate...

HierSearch: A Hierarchical Enterprise Deep Search Framework Integrating Local and Web Searches 14.08.2025

🤗 Upvotes: 24 | cs. IR, cs. AI, cs. CL Authors: Jiejun Tan, Zhicheng Dou, Yan Yu, Jiehan Cheng, Qiang Ju, Jian Xie, Ji-Rong Wen Title: HierSearch: A Hierarchical Enterprise Deep Search Framework Integrating Local and Web Searches Arxiv: http://arxiv.org/abs/2508.08088v1 Abstract: Recently, large reasoning models have demonstrated strong mathematical and coding abilities, and deep search leverages...

ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability 13.08.2025

🤗 Upvotes: 90 | cs. IR, cs. AI, cs. CL, cs. LG Authors: Wenhan Liu, Xinyu Ma, Weiwei Sun, Yutao Zhu, Yuchen Li, Dawei Yin, Zhicheng Dou Title: ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability Arxiv: http://arxiv.org/abs/2508.07050v1 Abstract: Large Language Model (LLM) based listwise ranking has shown superior performance in many passage ranking tasks. With the development of...

WideSearch: Benchmarking Agentic Broad Info-Seeking 13.08.2025

🤗 Upvotes: 86 | cs. CL Authors: Ryan Wong, Jiawei Wang, Junjie Zhao, Li Chen, Yan Gao, Long Zhang, Xuan Zhou, Zuo Wang, Kai Xiang, Ge Zhang, Wenhao Huang, Yang Wang, Ke Wang Title: WideSearch: Benchmarking Agentic Broad Info-Seeking Arxiv: http://arxiv.org/abs/2508.07999v1 Abstract: From professional research to everyday planning, many tasks are bottlenecked by wide-scale information seeking, whi...

Omni-Effects: Unified and Spatially-Controllable Visual Effects Generation 13.08.2025

🤗 Upvotes: 48 | cs. CV, cs. AI Authors: Fangyuan Mao, Aiming Hao, Jintao Chen, Dongxia Liu, Xiaokun Feng, Jiashu Zhu, Meiqi Wu, Chubin Chen, Jiahong Wu, Xiangxiang Chu Title: Omni-Effects: Unified and Spatially-Controllable Visual Effects Generation Arxiv: http://arxiv.org/abs/2508.07981v2 Abstract: Visual effects (VFX) are essential visual enhancements fundamental to modern cinematic production....

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems 13.08.2025

🤗 Upvotes: 42 | cs. AI, cs. CL, cs. MA Authors: Jinyuan Fang, Yanwen Peng, Xi Zhang, Yingxu Wang, Xinhao Yi, Guibin Zhang, Yi Xu, Bin Wu, Siwei Liu, Zihao Li, Zhaochun Ren, Nikos Aletras, Xi Wang, Han Zhou, Zaiqiao Meng Title: A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems Arxiv: http://arxiv.org/abs/2508.07407v1 Abstract:...

BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent 13.08.2025

🤗 Upvotes: 30 | cs. CL, cs. IR Authors: Zijian Chen, Xueguang Ma, Shengyao Zhuang, Ping Nie, Kai Zou, Andrew Liu, Joshua Green, Kshama Patel, Ruoxi Meng, Mingyi Su, Sahel Sharifymoghaddam, Yanxi Li, Haoran Hong, Xinyu Shi, Xuye Liu, Nandan Thakur, Crystina Zhang, Luyu Gao, Wenhu Chen, Jimmy Lin Title: BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent Arxiv:...

SONAR-LLM: Autoregressive Transformer that Thinks in Sentence Embeddings and Speaks in Tokens 13.08.2025

🤗 Upvotes: 28 | cs. CL Authors: Nikita Dragunov, Temurbek Rahmatullaev, Elizaveta Goncharova, Andrey Kuznetsov, Anton Razzhigaev Title: SONAR-LLM: Autoregressive Transformer that Thinks in Sentence Embeddings and Speaks in Tokens Arxiv: http://arxiv.org/abs/2508.05305v1 Abstract: The recently proposed Large Concept Model (LCM) generates text by predicting a sequence of sentence-level embeddings a...

Klear-Reasoner: Advancing Reasoning Capability via Gradient-Preserving Clipping Policy Optimization 13.08.2025

🤗 Upvotes: 28 | cs. LG, cs. AI, cs. CL Authors: Zhenpeng Su, Leiyu Pan, Xue Bai, Dening Liu, Guanting Dong, Jiaming Huang, Wenping Hu, Fuzheng Zhang, Kun Gai, Guorui Zhou Title: Klear-Reasoner: Advancing Reasoning Capability via Gradient-Preserving Clipping Policy Optimization Arxiv: http://arxiv.org/abs/2508.07629v2 Abstract: We present Klear-Reasoner, a model with long reasoning capabilities th...

MolmoAct: Action Reasoning Models that can Reason in Space 13.08.2025

🤗 Upvotes: 22 | cs. RO Authors: Jason Lee, Jiafei Duan, Haoquan Fang, Yuquan Deng, Shuo Liu, Boyang Li, Bohan Fang, Jieyu Zhang, Yi Ru Wang, Sangho Lee, Winson Han, Wilbert Pumacay, Angelica Wu, Rose Hendrix, Karen Farley, Eli VanderBilt, Ali Farhadi, Dieter Fox, Ranjay Krishna Title: MolmoAct: Action Reasoning Models that can Reason in Space Arxiv: http://arxiv.org/abs/2508.07917v1 Abstract: Rea...

GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models 12.08.2025

🤗 Upvotes: 79 | cs. CL Authors: GLM-4. 5 Team, :, Aohan Zeng, Xin Lv, Qinkai Zheng, Zhenyu Hou, Bin Chen, Chengxing Xie, Cunxiang Wang, Da Yin, Hao Zeng, Jiajie Zhang, Kedong Wang, Lucen Zhong, Mingdao Liu, Rui Lu, Shulin Cao, Xiaohan Zhang, Xuancheng Huang, Yao Wei, Yean Cheng, Yifan An, Yilin Niu, Yuanhao Wen, Yushi Bai, Zhengxiao Du, Zihan Wang, Zilin Zhu, Bohan Zhang, Bosi Wen, Bowen Wu, Bowe...

Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off 12.08.2025

🤗 Upvotes: 36 | cs. GR, cs. AI, cs. CV, cs. LG Authors: Seungyong Lee, Jeong-gi Kwak Title: Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off Arxiv: http://arxiv.org/abs/2508.04825v1 Abstract: Virtual try-on aims to synthesize a realistic image of a person wearing a target garment, but accurately modeling garment-body correspondence remains a persist...

Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens 08.08.2025

🤗 Upvotes: 143 | cs. AI, cs. CL, cs. LG Authors: Chengshuai Zhao, Zhen Tan, Pingchuan Ma, Dawei Li, Bohan Jiang, Yancheng Wang, Yingzhen Yang, Huan Liu Title: Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens Arxiv: http://arxiv.org/abs/2508.01191v2 Abstract: Chain-of-Thought (CoT) prompting has been shown to improve Large Language Model (LLM) performance on various tasks....

VeriGUI: Verifiable Long-Chain GUI Dataset 08.08.2025

🤗 Upvotes: 117 | cs. HC Authors: Shunyu Liu, Minghao Liu, Huichi Zhou, Zhenyu Cui, Yang Zhou, Yuhao Zhou, Wendong Fan, Ge Zhang, Jiajun Shi, Weihao Xuan, Jiaxing Huang, Shuang Luo, Fang Wu, Heli Qi, Qingcheng Zeng, Ziqi Ren, Jialiang Gao, Jindi Lv, Junjie Wang, Aosong Feng, Heng Zhou, Wangchunshu Zhou, Zhenfei Yin, Wenlong Zhang, Guohao Li, Wenhao Yu, Irene Li, Lei Ma, Lei Bai, Qunshu Lin, Mingli...

Efficient Agents: Building Effective Agents While Reducing Cost 08.08.2025

🤗 Upvotes: 53 | cs. AI, cs. CL, cs. MA Authors: Ningning Wang, Xavier Hu, Pai Liu, He Zhu, Yue Hou, Heyuan Huang, Shengyu Zhang, Jian Yang, Jiaheng Liu, Ge Zhang, Changwang Zhang, Jun Wang, Yuchen Eleanor Jiang, Wangchunshu Zhou Title: Efficient Agents: Building Effective Agents While Reducing Cost Arxiv: http://arxiv.org/abs/2508.02694v1 Abstract: The remarkable capabilities of Large Language Mo...

SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience 08.08.2025

🤗 Upvotes: 38 | cs. AI, cs. CL, cs. CV, cs. LG, cs. MA, cs. MM Authors: Zeyi Sun, Ziyu Liu, Yuhang Zang, Yuhang Cao, Xiaoyi Dong, Tong Wu, Dahua Lin, Jiaqi Wang Title: SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience Arxiv: http://arxiv.org/abs/2508.04700v1 Abstract: Repurposing large vision-language models (LVLMs) as computer use agents (CUAs) has led to substan...

Training Long-Context, Multi-Turn Software Engineering Agents with Reinforcement Learning 08.08.2025

🤗 Upvotes: 32 | cs. LG, cs. CL, cs. SE Authors: Alexander Golubev, Maria Trofimova, Sergei Polezhaev, Ibragim Badertdinov, Maksim Nekrashevich, Anton Shevtsov, Simon Karasik, Sergey Abramov, Andrei Andriushchenko, Filipp Fisin, Sergei Skvortsov, Boris Yangel Title: Training Long-Context, Multi-Turn Software Engineering Agents with Reinforcement Learning Arxiv: http://arxiv.org/abs/2508.03501v1 Ab...

Enhancing Vision-Language Model Training with Reinforcement Learning in Synthetic Worlds for Real-World Success 08.08.2025

🤗 Upvotes: 29 | cs. LG, cs. AI Authors: George Bredis, Stanislav Dereka, Viacheslav Sinii, Ruslan Rakhimov, Daniil Gavrilov Title: Enhancing Vision-Language Model Training with Reinforcement Learning in Synthetic Worlds for Real-World Success Arxiv: http://arxiv.org/abs/2508.04280v1 Abstract: Interactive multimodal agents must convert raw visual observations into coherent sequences of language-co...

Agent Lightning: Train ANY AI Agents with Reinforcement Learning 08.08.2025

🤗 Upvotes: 24 | cs. AI, cs. LG Authors: Xufang Luo, Yuge Zhang, Zhiyuan He, Zilong Wang, Siyun Zhao, Dongsheng Li, Luna K. Qiu, Yuqing Yang Title: Agent Lightning: Train ANY AI Agents with Reinforcement Learning Arxiv: http://arxiv.org/abs/2508.03680v1 Abstract: We present Agent Lightning, a flexible and extensible framework that enables Reinforcement Learning (RL)-based training of Large Languag...

Qwen-Image Technical Report 06.08.2025

🤗 Upvotes: 91 | cs. CV Authors: Chenfei Wu, Jiahao Li, Jingren Zhou, Junyang Lin, Kaiyuan Gao, Kun Yan, Sheng-ming Yin, Shuai Bai, Xiao Xu, Yilei Chen, Yuxiang Chen, Zecheng Tang, Zekai Zhang, Zhengyi Wang, An Yang, Bowen Yu, Chen Cheng, Dayiheng Liu, Deqing Li, Hang Zhang, Hao Meng, Hu Wei, Jingyuan Ni, Kai Chen, Kuan Cao, Liang Peng, Lin Qu, Minggang Wu, Peng Wang, Shuting Yu, Tingkun Wen, Wens...

SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension 06.08.2025

🤗 Upvotes: 39 | cs. CL Authors: Junjie Wu, Jiangnan Li, Yuqing Li, Lemao Liu, Liyan Xu, Jiwei Li, Dit-Yan Yeung, Jie Zhou, Mo Yu Title: SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension Arxiv: http://arxiv.org/abs/2508.01959v1 Abstract: Retrieval-augmented generation (RAG) over long documents typically involves splitting the text into smalle...

CellForge: Agentic Design of Virtual Cell Models 06.08.2025

🤗 Upvotes: 30 | cs. LG, cs. AI, cs. CL, q-bio. QM Authors: Xiangru Tang, Zhuoyun Yu, Jiapeng Chen, Yan Cui, Daniel Shao, Weixu Wang, Fang Wu, Yuchen Zhuang, Wenqi Shi, Zhi Huang, Arman Cohan, Xihong Lin, Fabian Theis, Smita Krishnaswamy, Mark Gerstein Title: CellForge: Agentic Design of Virtual Cell Models Arxiv: http://arxiv.org/abs/2508.02276v1 Abstract: Virtual cell modeling represents an emer...

Beyond the Trade-off: Self-Supervised Reinforcement Learning for Reasoning Models' Instruction Following 06.08.2025

🤗 Upvotes: 25 | cs. AI Authors: Qingyu Ren, Qianyu He, Bowei Zhang, Jie Zeng, Jiaqing Liang, Yanghua Xiao, Weikang Zhou, Zeye Sun, Fei Yu Title: Beyond the Trade-off: Self-Supervised Reinforcement Learning for Reasoning Models' Instruction Following Arxiv: http://arxiv.org/abs/2508.02150v1 Abstract: Reasoning models excel in complex problem solving but exhibit a concerning trade off between reaso...

Llama-3.1-FoundationAI-SecurityLLM-8B-Instruct Technical Report 06.08.2025

🤗 Upvotes: 21 | cs. CR, cs. AI Authors: Sajana Weerawardhena, Paul Kassianik, Blaine Nelson, Baturay Saglam, Anu Vellore, Aman Priyanshu, Supriti Vijay, Massimo Aufiero, Arthur Goldblatt, Fraser Burch, Ed Li, Jianliang He, Dhruv Kedia, Kojin Oshiba, Zhouran Yang, Yaron Singer, Amin Karbasi Title: Llama-3.1-FoundationAI-SecurityLLM-8B-Instruct Technical Report Arxiv: http://arxiv.org/abs/2508.0105...

Listen to the Daily Paper Cast podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.