Igor Melnyk

Arxiv Papers

Science EN ↓ 2489 episodes

Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers

Author

Igor Melnyk

Category

Science

Podcast website

github.com

Latest episode

Sep 1, 2025

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

An Interactive Agent Foundation Model 09.02.2024

The paper introduces an Interactive Agent Foundation Model for training versatile AI agents across various domains, demonstrating its effectiveness in robotics, gaming AI, and healthcare. https://arxiv.org/abs//2402.05929 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify...

[short] Learning to Route Among Specialized Experts for Zero-Shot Generalization 09.02.2024

The paper introduces lmttPHATGOOSE, a method for improving zero-shot generalization by adaptively choosing specialized language model experts for each token and layer. https://arxiv.org/abs//2402.05859 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters...

Learning to Route Among Specialized Experts for Zero-Shot Generalization 09.02.2024

The paper introduces lmttPHATGOOSE, a method for improving zero-shot generalization by adaptively choosing specialized language model experts for each token and layer. https://arxiv.org/abs//2402.05859 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters...

[short] Hydragen: High-Throughput LLM Inference with Shared Prefixes 08.02.2024

Hydragen introduces efficient attention computation for large language models, improving throughput by up to 32x and enabling the use of very long shared contexts. https://arxiv.org/abs//2402.05099 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spo...

Hydragen: High-Throughput LLM Inference with Shared Prefixes 08.02.2024

Hydragen introduces efficient attention computation for large language models, improving throughput by up to 32x and enabling the use of very long shared contexts. https://arxiv.org/abs//2402.05099 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spo...

[short] Direct Language Model Alignment from Online AI Feedback 08.02.2024

Direct alignment from preferences (DAP) methods lack online feedback. The study introduces online AI feedback (OAIF) using a language model annotator, outperforming offline DAP and RLHF methods. https://arxiv.org/abs//2402.04792 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...

Direct Language Model Alignment from Online AI Feedback 08.02.2024

Direct alignment from preferences (DAP) methods lack online feedback. The study introduces online AI feedback (OAIF) using a language model annotator, outperforming offline DAP and RLHF methods. https://arxiv.org/abs//2402.04792 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...

[short] Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning 08.02.2024

The paper compares methods for selecting high-quality data for instruction fine-tuning of LLMs and finds that selecting the 1,000 longest responses consistently outperforms sophisticated methods. https://arxiv.org/abs//2402.04833 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...

Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning 08.02.2024

The paper compares methods for selecting high-quality data for instruction fine-tuning of LLMs and finds that selecting the 1,000 longest responses consistently outperforms sophisticated methods. https://arxiv.org/abs//2402.04833 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...

[short] SELF-DISCOVER: Large Language Models Self-Compose Reasoning Structures 07.02.2024

SELF-DISCOVER framework enables LLMs to self-discover reasoning structures, improving performance on challenging benchmarks and outperforming inference-intensive methods while requiring fewer computations. https://arxiv.org/abs//2402.03620 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id...

SELF-DISCOVER: Large Language Models Self-Compose Reasoning Structures 07.02.2024

SELF-DISCOVER framework enables LLMs to self-discover reasoning structures, improving performance on challenging benchmarks and outperforming inference-intensive methods while requiring fewer computations. https://arxiv.org/abs//2402.03620 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id...

[short] Open RL Benchmark: Comprehensive Tracked Experiments for Reinforcement Learning 06.02.2024

Open RL Benchmark provides fully tracked RL experiments, including algorithm-specific and system metrics, for reproducibility and community contribution, with over 25,000 runs. https://arxiv.org/abs//2402.03046 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://p...

Open RL Benchmark: Comprehensive Tracked Experiments for Reinforcement Learning 06.02.2024

Open RL Benchmark provides fully tracked RL experiments, including algorithm-specific and system metrics, for reproducibility and community contribution, with over 25,000 runs. https://arxiv.org/abs//2402.03046 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://p...

[short] Decoding-time Realignment of Language Models 06.02.2024

Language model alignment with human preferences is crucial. Decoding-time realignment (DeRa) offers an efficient method to explore and evaluate regularization strengths in aligned models without retraining. https://arxiv.org/abs//2402.02992 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...

Decoding-time Realignment of Language Models 06.02.2024

Language model alignment with human preferences is crucial. Decoding-time realignment (DeRa) offers an efficient method to explore and evaluate regularization strengths in aligned models without retraining. https://arxiv.org/abs//2402.02992 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...

[short] Specialized Language Models with Cheap Inference from Limited Domain Data 05.02.2024

This paper explores constraints in applying large language models to tasks with limited resources, comparing different approaches and finding alternatives to standard practices. https://arxiv.org/abs//2402.01093 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://...

Specialized Language Models with Cheap Inference from Limited Domain Data 05.02.2024

This paper explores constraints in applying large language models to tasks with limited resources, comparing different approaches and finding alternatives to standard practices. https://arxiv.org/abs//2402.01093 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://...

[short] Repeat After Me: Transformers are Better than State Space Models at Copying 05.02.2024

The paper compares transformers and generalized state space models (GSSMs) for sequence modeling, finding that transformers outperform GSSMs in efficiency and generalization for tasks requiring copying from input context. https://arxiv.org/abs//2402.01032 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast...

Repeat After Me: Transformers are Better than State Space Models at Copying 05.02.2024

The paper compares transformers and generalized state space models (GSSMs) for sequence modeling, finding that transformers outperform GSSMs in efficiency and generalization for tasks requiring copying from input context. https://arxiv.org/abs//2402.01032 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast...

[short] Unlearnable Algorithms for In-context Learning 04.02.2024

The paper explores efficient unlearning methods for large language models, proposing an algorithm for exact unlearning during task adaptation, with a focus on in-context learning's advantages over fine-tuning. https://arxiv.org/abs//2402.00751 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-p...

Unlearnable Algorithms for In-context Learning 04.02.2024

The paper explores efficient unlearning methods for large language models, proposing an algorithm for exact unlearning during task adaptation, with a focus on in-context learning's advantages over fine-tuning. https://arxiv.org/abs//2402.00751 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-p...

[short] Can Large Language Models Understand Context? 03.02.2024

This paper introduces a context understanding benchmark for evaluating Large Language Models' ability to understand contextual features, finding that pre-trained dense models struggle with nuanced context understanding. https://arxiv.org/abs//2402.00858 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podca...

Can Large Language Models Understand Context? 03.02.2024

This paper introduces a context understanding benchmark for evaluating Large Language Models' ability to understand contextual features, finding that pre-trained dense models struggle with nuanced context understanding. https://arxiv.org/abs//2402.00858 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podca...

[short] Position Paper: Bayesian Deep Learning in the Age of Large-Scale AI 03.02.2024

The paper discusses the limitations of current deep learning research and the potential of Bayesian deep learning to address diverse challenges, emphasizing the need for broader perspectives and new research avenues. https://arxiv.org/abs//2402.00809 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxi...

Position Paper: Bayesian Deep Learning in the Age of Large-Scale AI 03.02.2024

The paper discusses the limitations of current deep learning research and the potential of Bayesian deep learning to address diverse challenges, emphasizing the need for broader perspectives and new research avenues. https://arxiv.org/abs//2402.00809 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxi...

Listen to the Arxiv Papers podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.