Igor Melnyk
Arxiv Papers
Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
An Interactive Agent Foundation Model 09.02.2024 23:00
The paper introduces an Interactive Agent Foundation Model for training versatile AI agents across various domains, demonstrating its effectiveness in robotics, gaming AI, and healthcare. https://arxiv.org/abs//2402.05929 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify...
[short] Learning to Route Among Specialized Experts for Zero-Shot Generalization 09.02.2024 2:36
The paper introduces lmttPHATGOOSE, a method for improving zero-shot generalization by adaptively choosing specialized language model experts for each token and layer. https://arxiv.org/abs//2402.05859 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters...
Learning to Route Among Specialized Experts for Zero-Shot Generalization 09.02.2024 18:33
The paper introduces lmttPHATGOOSE, a method for improving zero-shot generalization by adaptively choosing specialized language model experts for each token and layer. https://arxiv.org/abs//2402.05859 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters...
[short] Hydragen: High-Throughput LLM Inference with Shared Prefixes 08.02.2024 2:13
Hydragen introduces efficient attention computation for large language models, improving throughput by up to 32x and enabling the use of very long shared contexts. https://arxiv.org/abs//2402.05099 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spo...
Hydragen: High-Throughput LLM Inference with Shared Prefixes 08.02.2024 20:01
Hydragen introduces efficient attention computation for large language models, improving throughput by up to 32x and enabling the use of very long shared contexts. https://arxiv.org/abs//2402.05099 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spo...
[short] Direct Language Model Alignment from Online AI Feedback 08.02.2024 2:20
Direct alignment from preferences (DAP) methods lack online feedback. The study introduces online AI feedback (OAIF) using a language model annotator, outperforming offline DAP and RLHF methods. https://arxiv.org/abs//2402.04792 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...
Direct Language Model Alignment from Online AI Feedback 08.02.2024 15:48
Direct alignment from preferences (DAP) methods lack online feedback. The study introduces online AI feedback (OAIF) using a language model annotator, outperforming offline DAP and RLHF methods. https://arxiv.org/abs//2402.04792 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...
[short] Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning 08.02.2024 2:43
The paper compares methods for selecting high-quality data for instruction fine-tuning of LLMs and finds that selecting the 1,000 longest responses consistently outperforms sophisticated methods. https://arxiv.org/abs//2402.04833 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...
Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning 08.02.2024 16:58
The paper compares methods for selecting high-quality data for instruction fine-tuning of LLMs and finds that selecting the 1,000 longest responses consistently outperforms sophisticated methods. https://arxiv.org/abs//2402.04833 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...
[short] SELF-DISCOVER: Large Language Models Self-Compose Reasoning Structures 07.02.2024 1:57
SELF-DISCOVER framework enables LLMs to self-discover reasoning structures, improving performance on challenging benchmarks and outperforming inference-intensive methods while requiring fewer computations. https://arxiv.org/abs//2402.03620 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id...
SELF-DISCOVER: Large Language Models Self-Compose Reasoning Structures 07.02.2024 14:37
SELF-DISCOVER framework enables LLMs to self-discover reasoning structures, improving performance on challenging benchmarks and outperforming inference-intensive methods while requiring fewer computations. https://arxiv.org/abs//2402.03620 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id...
[short] Open RL Benchmark: Comprehensive Tracked Experiments for Reinforcement Learning 06.02.2024 2:53
Open RL Benchmark provides fully tracked RL experiments, including algorithm-specific and system metrics, for reproducibility and community contribution, with over 25,000 runs. https://arxiv.org/abs//2402.03046 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://p...
Open RL Benchmark: Comprehensive Tracked Experiments for Reinforcement Learning 06.02.2024 15:18
Open RL Benchmark provides fully tracked RL experiments, including algorithm-specific and system metrics, for reproducibility and community contribution, with over 25,000 runs. https://arxiv.org/abs//2402.03046 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://p...
[short] Decoding-time Realignment of Language Models 06.02.2024 2:23
Language model alignment with human preferences is crucial. Decoding-time realignment (DeRa) offers an efficient method to explore and evaluate regularization strengths in aligned models without retraining. https://arxiv.org/abs//2402.02992 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...
Decoding-time Realignment of Language Models 06.02.2024 18:08
Language model alignment with human preferences is crucial. Decoding-time realignment (DeRa) offers an efficient method to explore and evaluate regularization strengths in aligned models without retraining. https://arxiv.org/abs//2402.02992 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...
[short] Specialized Language Models with Cheap Inference from Limited Domain Data 05.02.2024 1:57
This paper explores constraints in applying large language models to tasks with limited resources, comparing different approaches and finding alternatives to standard practices. https://arxiv.org/abs//2402.01093 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://...
Specialized Language Models with Cheap Inference from Limited Domain Data 05.02.2024 21:21
This paper explores constraints in applying large language models to tasks with limited resources, comparing different approaches and finding alternatives to standard practices. https://arxiv.org/abs//2402.01093 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://...
[short] Repeat After Me: Transformers are Better than State Space Models at Copying 05.02.2024 2:15
The paper compares transformers and generalized state space models (GSSMs) for sequence modeling, finding that transformers outperform GSSMs in efficiency and generalization for tasks requiring copying from input context. https://arxiv.org/abs//2402.01032 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast...
Repeat After Me: Transformers are Better than State Space Models at Copying 05.02.2024 26:55
The paper compares transformers and generalized state space models (GSSMs) for sequence modeling, finding that transformers outperform GSSMs in efficiency and generalization for tasks requiring copying from input context. https://arxiv.org/abs//2402.01032 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast...
[short] Unlearnable Algorithms for In-context Learning 04.02.2024 2:06
The paper explores efficient unlearning methods for large language models, proposing an algorithm for exact unlearning during task adaptation, with a focus on in-context learning's advantages over fine-tuning. https://arxiv.org/abs//2402.00751 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-p...
Unlearnable Algorithms for In-context Learning 04.02.2024 21:51
The paper explores efficient unlearning methods for large language models, proposing an algorithm for exact unlearning during task adaptation, with a focus on in-context learning's advantages over fine-tuning. https://arxiv.org/abs//2402.00751 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-p...
[short] Can Large Language Models Understand Context? 03.02.2024 3:01
This paper introduces a context understanding benchmark for evaluating Large Language Models' ability to understand contextual features, finding that pre-trained dense models struggle with nuanced context understanding. https://arxiv.org/abs//2402.00858 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podca...
Can Large Language Models Understand Context? 03.02.2024 17:15
This paper introduces a context understanding benchmark for evaluating Large Language Models' ability to understand contextual features, finding that pre-trained dense models struggle with nuanced context understanding. https://arxiv.org/abs//2402.00858 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podca...
[short] Position Paper: Bayesian Deep Learning in the Age of Large-Scale AI 03.02.2024 2:31
The paper discusses the limitations of current deep learning research and the potential of Bayesian deep learning to address diverse challenges, emphasizing the need for broader perspectives and new research avenues. https://arxiv.org/abs//2402.00809 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxi...
Position Paper: Bayesian Deep Learning in the Age of Large-Scale AI 03.02.2024 24:56
The paper discusses the limitations of current deep learning research and the potential of Bayesian deep learning to address diverse challenges, emphasizing the need for broader perspectives and new research avenues. https://arxiv.org/abs//2402.00809 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxi...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.