Igor Melnyk

Arxiv Papers

Science EN ↓ 2489 episodes

Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers

Author

Igor Melnyk

Category

Science

Podcast website

github.com

Latest episode

Sep 1, 2025

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

Refusal in Language Models Is Mediated by a Single Direction 23.06.2024

Study explores refusal behavior in chat models, identifying a one-dimensional subspace mediating refusal. Proposes a method to disable refusal while preserving other capabilities, highlighting safety fine-tuning limitations. https://arxiv.org/abs//2406.11717 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podc...

[QA] Instruction Pre-Training: Language Models are Supervised Multitask Learners 22.06.2024

The paper introduces Instruction Pre-Training, a framework for supervised multitask pre-training of language models using instruction-response pairs, showing improved generalization and performance. https://arxiv.org/abs//2406.14491 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476...

Instruction Pre-Training: Language Models are Supervised Multitask Learners 22.06.2024

The paper introduces Instruction Pre-Training, a framework for supervised multitask pre-training of language models using instruction-response pairs, showing improved generalization and performance. https://arxiv.org/abs//2406.14491 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476...

[QA] Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More? 22.06.2024

Long-context language models (LCLMs) show promise in revolutionizing tasks without external tools, as demonstrated by LOFT benchmark's evaluation of LCLMs' performance in complex contexts. https://arxiv.org/abs//2406.13121 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id169247601...

Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More? 22.06.2024

Long-context language models (LCLMs) show promise in revolutionizing tasks without external tools, as demonstrated by LOFT benchmark's evaluation of LCLMs' performance in complex contexts. https://arxiv.org/abs//2406.13121 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id169247601...

[QA] RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold 21.06.2024

https://arxiv.org/abs//2406.14532 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers

RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold 21.06.2024

https://arxiv.org/abs//2406.14532 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers

[QA] Consistency Models Made Easy 21.06.2024

Proposes Easy Consistency Tuning (ECT) for training consistency models, improving efficiency significantly. Achieves high quality results on CIFAR-10 in just 1 hour on a single GPU. https://arxiv.org/abs//2406.14548 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: http...

Consistency Models Made Easy 21.06.2024

Proposes Easy Consistency Tuning (ECT) for training consistency models, improving efficiency significantly. Achieves high quality results on CIFAR-10 in just 1 hour on a single GPU. https://arxiv.org/abs//2406.14548 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: http...

[QA] What Are the Odds? Language Models Are Capable of Probabilistic Reasoning 19.06.2024

Paper evaluates language models' probabilistic reasoning abilities using statistical distributions. Three tasks assessed with different contextual inputs. Models can infer distributions with real-world context and simplified assumptions. https://arxiv.org/abs//2406.12830 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts....

What Are the Odds? Language Models Are Capable of Probabilistic Reasoning 19.06.2024

Paper evaluates language models' probabilistic reasoning abilities using statistical distributions. Three tasks assessed with different contextual inputs. Models can infer distributions with real-world context and simplified assumptions. https://arxiv.org/abs//2406.12830 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts....

[QA] Adversarial Attacks on Multimodal Agents 19.06.2024

The paper explores safety risks posed by multimodal agents and demonstrates attacks using adversarial text strings to manipulate VLMs, with varying success rates based on different models. https://arxiv.org/abs//2406.12814 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotif...

Adversarial Attacks on Multimodal Agents 19.06.2024

The paper explores safety risks posed by multimodal agents and demonstrates attacks using adversarial text strings to manipulate VLMs, with varying success rates based on different models. https://arxiv.org/abs//2406.12814 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotif...

[QA] Can Go AIs be adversarially robust? 19.06.2024

The paper explores defenses to improve KataGo's performance against adversarial attacks in Go, finding some defenses effective but none able to withstand adaptive attacks. https://arxiv.org/abs//2406.12843 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://po...

Can Go AIs be adversarially robust? 19.06.2024

The paper explores defenses to improve KataGo's performance against adversarial attacks in Go, finding some defenses effective but none able to withstand adaptive attacks. https://arxiv.org/abs//2406.12843 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://po...

[QA] Autoregressive Image Generation without Vector Quantization 18.06.2024

Proposing a diffusion-based approach for autoregressive modeling in continuous-valued space, eliminating the need for discrete tokens and achieving strong results in image generation. https://arxiv.org/abs//2406.11838 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: ht...

Autoregressive Image Generation without Vector Quantization 18.06.2024

Proposing a diffusion-based approach for autoregressive modeling in continuous-valued space, eliminating the need for discrete tokens and achieving strong results in image generation. https://arxiv.org/abs//2406.11838 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: ht...

[QA] Measuring memorization in RLHF for code completion 18.06.2024

https://arxiv.org/abs//2406.11715 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers

Measuring memorization in RLHF for code completion 18.06.2024

https://arxiv.org/abs//2406.11715 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers

[QA] Bootstrapping Language Models with DPO Implicit Rewards 17.06.2024

The paper introduces DICE, a method for aligning large language models using implicit rewards from DPO. DICE outperforms Gemini Pro on AlpacaEval 2 with 8B parameters and no external feedback. https://arxiv.org/abs//2406.09760 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Sp...

Bootstrapping Language Models with DPO Implicit Rewards 17.06.2024

The paper introduces DICE, a method for aligning large language models using implicit rewards from DPO. DICE outperforms Gemini Pro on AlpacaEval 2 with 8B parameters and no external feedback. https://arxiv.org/abs//2406.09760 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Sp...

[QA] Ad Auctions for LLMs via Retrieval Augmented Generation 17.06.2024

Novel auction mechanisms for ad allocation and pricing in large language models (LLMs) are proposed, maximizing social welfare and ensuring fairness. Empirical evaluation supports the approach's feasibility and effectiveness. https://arxiv.org/abs//2406.09459 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us...

Ad Auctions for LLMs via Retrieval Augmented Generation 17.06.2024

Novel auction mechanisms for ad allocation and pricing in large language models (LLMs) are proposed, maximizing social welfare and ensuring fairness. Empirical evaluation supports the approach's feasibility and effectiveness. https://arxiv.org/abs//2406.09459 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us...

[QA] An Empirical Study of Mamba-based Language Models 16.06.2024

Mamba models challenge Transformers at larger scales, with Mamba-2-Hybrid surpassing Transformers on various tasks, showing potential for efficient token generation. https://arxiv.org/abs//2406.07887 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.s...

An Empirical Study of Mamba-based Language Models 16.06.2024

Mamba models challenge Transformers at larger scales, with Mamba-2-Hybrid surpassing Transformers on various tasks, showing potential for efficient token generation. https://arxiv.org/abs//2406.07887 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.s...

Listen to the Arxiv Papers podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.