Igor Melnyk
Arxiv Papers
Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
Refusal in Language Models Is Mediated by a Single Direction 23.06.2024 18:40
Study explores refusal behavior in chat models, identifying a one-dimensional subspace mediating refusal. Proposes a method to disable refusal while preserving other capabilities, highlighting safety fine-tuning limitations. https://arxiv.org/abs//2406.11717 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podc...
[QA] Instruction Pre-Training: Language Models are Supervised Multitask Learners 22.06.2024 11:26
The paper introduces Instruction Pre-Training, a framework for supervised multitask pre-training of language models using instruction-response pairs, showing improved generalization and performance. https://arxiv.org/abs//2406.14491 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476...
Instruction Pre-Training: Language Models are Supervised Multitask Learners 22.06.2024 12:59
The paper introduces Instruction Pre-Training, a framework for supervised multitask pre-training of language models using instruction-response pairs, showing improved generalization and performance. https://arxiv.org/abs//2406.14491 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476...
[QA] Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More? 22.06.2024 11:15
Long-context language models (LCLMs) show promise in revolutionizing tasks without external tools, as demonstrated by LOFT benchmark's evaluation of LCLMs' performance in complex contexts. https://arxiv.org/abs//2406.13121 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id169247601...
Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More? 22.06.2024 17:14
Long-context language models (LCLMs) show promise in revolutionizing tasks without external tools, as demonstrated by LOFT benchmark's evaluation of LCLMs' performance in complex contexts. https://arxiv.org/abs//2406.13121 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id169247601...
[QA] RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold 21.06.2024 10:29
https://arxiv.org/abs//2406.14532 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold 21.06.2024 13:02
https://arxiv.org/abs//2406.14532 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
[QA] Consistency Models Made Easy 21.06.2024 8:31
Proposes Easy Consistency Tuning (ECT) for training consistency models, improving efficiency significantly. Achieves high quality results on CIFAR-10 in just 1 hour on a single GPU. https://arxiv.org/abs//2406.14548 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: http...
Consistency Models Made Easy 21.06.2024 19:39
Proposes Easy Consistency Tuning (ECT) for training consistency models, improving efficiency significantly. Achieves high quality results on CIFAR-10 in just 1 hour on a single GPU. https://arxiv.org/abs//2406.14548 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: http...
[QA] What Are the Odds? Language Models Are Capable of Probabilistic Reasoning 19.06.2024 14:25
Paper evaluates language models' probabilistic reasoning abilities using statistical distributions. Three tasks assessed with different contextual inputs. Models can infer distributions with real-world context and simplified assumptions. https://arxiv.org/abs//2406.12830 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts....
What Are the Odds? Language Models Are Capable of Probabilistic Reasoning 19.06.2024 10:43
Paper evaluates language models' probabilistic reasoning abilities using statistical distributions. Three tasks assessed with different contextual inputs. Models can infer distributions with real-world context and simplified assumptions. https://arxiv.org/abs//2406.12830 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts....
[QA] Adversarial Attacks on Multimodal Agents 19.06.2024 8:27
The paper explores safety risks posed by multimodal agents and demonstrates attacks using adversarial text strings to manipulate VLMs, with varying success rates based on different models. https://arxiv.org/abs//2406.12814 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotif...
Adversarial Attacks on Multimodal Agents 19.06.2024 16:08
The paper explores safety risks posed by multimodal agents and demonstrates attacks using adversarial text strings to manipulate VLMs, with varying success rates based on different models. https://arxiv.org/abs//2406.12814 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotif...
[QA] Can Go AIs be adversarially robust? 19.06.2024 7:44
The paper explores defenses to improve KataGo's performance against adversarial attacks in Go, finding some defenses effective but none able to withstand adaptive attacks. https://arxiv.org/abs//2406.12843 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://po...
Can Go AIs be adversarially robust? 19.06.2024 12:59
The paper explores defenses to improve KataGo's performance against adversarial attacks in Go, finding some defenses effective but none able to withstand adaptive attacks. https://arxiv.org/abs//2406.12843 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://po...
[QA] Autoregressive Image Generation without Vector Quantization 18.06.2024 10:24
Proposing a diffusion-based approach for autoregressive modeling in continuous-valued space, eliminating the need for discrete tokens and achieving strong results in image generation. https://arxiv.org/abs//2406.11838 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: ht...
Autoregressive Image Generation without Vector Quantization 18.06.2024 9:16
Proposing a diffusion-based approach for autoregressive modeling in continuous-valued space, eliminating the need for discrete tokens and achieving strong results in image generation. https://arxiv.org/abs//2406.11838 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: ht...
[QA] Measuring memorization in RLHF for code completion 18.06.2024 9:39
https://arxiv.org/abs//2406.11715 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
Measuring memorization in RLHF for code completion 18.06.2024 16:38
https://arxiv.org/abs//2406.11715 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
[QA] Bootstrapping Language Models with DPO Implicit Rewards 17.06.2024 8:41
The paper introduces DICE, a method for aligning large language models using implicit rewards from DPO. DICE outperforms Gemini Pro on AlpacaEval 2 with 8B parameters and no external feedback. https://arxiv.org/abs//2406.09760 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Sp...
Bootstrapping Language Models with DPO Implicit Rewards 17.06.2024 16:02
The paper introduces DICE, a method for aligning large language models using implicit rewards from DPO. DICE outperforms Gemini Pro on AlpacaEval 2 with 8B parameters and no external feedback. https://arxiv.org/abs//2406.09760 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Sp...
[QA] Ad Auctions for LLMs via Retrieval Augmented Generation 17.06.2024 10:03
Novel auction mechanisms for ad allocation and pricing in large language models (LLMs) are proposed, maximizing social welfare and ensuring fairness. Empirical evaluation supports the approach's feasibility and effectiveness. https://arxiv.org/abs//2406.09459 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us...
Ad Auctions for LLMs via Retrieval Augmented Generation 17.06.2024 14:18
Novel auction mechanisms for ad allocation and pricing in large language models (LLMs) are proposed, maximizing social welfare and ensuring fairness. Empirical evaluation supports the approach's feasibility and effectiveness. https://arxiv.org/abs//2406.09459 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us...
[QA] An Empirical Study of Mamba-based Language Models 16.06.2024 10:36
Mamba models challenge Transformers at larger scales, with Mamba-2-Hybrid surpassing Transformers on various tasks, showing potential for efficient token generation. https://arxiv.org/abs//2406.07887 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.s...
An Empirical Study of Mamba-based Language Models 16.06.2024 28:32
Mamba models challenge Transformers at larger scales, with Mamba-2-Hybrid surpassing Transformers on various tasks, showing potential for efficient token generation. https://arxiv.org/abs//2406.07887 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.s...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.