Igor Melnyk
Arxiv Papers
Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
To Believe or Not to Believe Your LLM 05.06.2024 14:33
The paper explores quantifying uncertainty in large language models to detect unreliable responses, focusing on distinguishing epistemic and aleatoric uncertainties using an information-theoretic metric. https://arxiv.org/abs//2406.02543 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16...
[QA] Learning to grok: Emergence of in-context learning and skill composition in modular arithmetic tasks 05.06.2024 9:19
Study explores in-context learning and skill composition in modular arithmetic tasks using GPT-style transformer models, showing transition from in-distribution to out-of-distribution generalization with increasing pre-training tasks. https://arxiv.org/abs//2406.02550 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.c...
Learning to grok: Emergence of in-context learning and skill composition in modular arithmetic tasks 05.06.2024 18:29
Study explores in-context learning and skill composition in modular arithmetic tasks using GPT-style transformer models, showing transition from in-distribution to out-of-distribution generalization with increasing pre-training tasks. https://arxiv.org/abs//2406.02550 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.c...
[QA] Show, Don't Tell: Aligning Language Models with Demonstrated Feedback 04.06.2024 8:09
Demonstration ITerated Task Optimization (DITTO) aligns language models to specific settings using a few demonstrations, outperforming other methods by 19% on average. https://arxiv.org/abs//2406.00888 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters...
Show, Don't Tell: Aligning Language Models with Demonstrated Feedback 04.06.2024 20:18
Demonstration ITerated Task Optimization (DITTO) aligns language models to specific settings using a few demonstrations, outperforming other methods by 19% on average. https://arxiv.org/abs//2406.00888 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters...
[QA] What makes unlearning hard and what to do about it 04.06.2024 8:24
Research explores machine unlearning, removing specific training data from models. Investigates factors affecting unlearning difficulty and algorithm performance, introducing Refined-Unlearning Meta-algorithm (RUM) to enhance existing methods. https://arxiv.org/abs//2406.01257 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcast...
What makes unlearning hard and what to do about it 04.06.2024 17:34
Research explores machine unlearning, removing specific training data from models. Investigates factors affecting unlearning difficulty and algorithm performance, introducing Refined-Unlearning Meta-algorithm (RUM) to enhance existing methods. https://arxiv.org/abs//2406.01257 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcast...
[QA] Learning to Play Atari in a World of Tokens 04.06.2024 7:28
DART introduces discrete abstract representations for transformer-based learning, improving sample efficiency and outperforming previous methods on Atari games. https://arxiv.org/abs//2406.01361 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotif...
Learning to Play Atari in a World of Tokens 04.06.2024 11:24
DART introduces discrete abstract representations for transformer-based learning, improving sample efficiency and outperforming previous methods on Atari games. https://arxiv.org/abs//2406.01361 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotif...
[QA] Perplexed by Perplexity: Perplexity-Based Data Pruning With Small Reference Models 03.06.2024 9:23
Investigating if small language models can improve large models by pruning datasets based on perplexity, leading to enhanced downstream task performance and reduced pretraining steps. https://arxiv.org/abs//2405.20541 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: ht...
Perplexed by Perplexity: Perplexity-Based Data Pruning With Small Reference Models 03.06.2024 12:59
Investigating if small language models can improve large models by pruning datasets based on perplexity, leading to enhanced downstream task performance and reduced pretraining steps. https://arxiv.org/abs//2405.20541 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: ht...
[QA] Contextual Position Encoding: Learning to Count What's Important 02.06.2024 8:04
Large Language Models (LLMs) use attention mechanisms for token interactions. Current position encoding methods lack abstraction levels. CoPE introduces contextual position encoding for improved addressing and performance. https://arxiv.org/abs//2405.18719 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcas...
Contextual Position Encoding: Learning to Count What's Important 02.06.2024 19:17
Large Language Models (LLMs) use attention mechanisms for token interactions. Current position encoding methods lack abstraction levels. CoPE introduces contextual position encoding for improved addressing and performance. https://arxiv.org/abs//2405.18719 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcas...
[QA] JINA CLIP: Your CLIP Model Is Also Your Text Retriever 02.06.2024 7:19
Proposed multi-task contrastive training method improves CLIP models for text-only tasks, achieving state-of-the-art performance on text-image and text-text retrieval tasks with jina-clip-v1 model. https://arxiv.org/abs//2405.20204 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924760...
JINA CLIP: Your CLIP Model Is Also Your Text Retriever 02.06.2024 8:15
Proposed multi-task contrastive training method improves CLIP models for text-only tasks, achieving state-of-the-art performance on text-image and text-text retrieval tasks with jina-clip-v1 model. https://arxiv.org/abs//2405.20204 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924760...
[QA] Large Language Models Can Self-Improve At Web Agent Tasks 01.06.2024 8:42
Large language models (LLMs) self-improve to navigate web environments using synthetic data, achieving 31% task completion rate improvement on WebArena benchmark, introducing new evaluation metrics. https://arxiv.org/abs//2405.20309 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476...
Large Language Models Can Self-Improve At Web Agent Tasks 01.06.2024 13:55
Large language models (LLMs) self-improve to navigate web environments using synthetic data, achieving 31% task completion rate improvement on WebArena benchmark, introducing new evaluation metrics. https://arxiv.org/abs//2405.20309 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476...
[QA] Is In-Context Learning Sufficient for Instruction Following in LLMs? 01.06.2024 9:33
In-context learning (ICL) with URIAL aligns base LLMs using few examples but underperforms compared to instruction fine-tuning, with a proposed greedy selection approach improving performance. https://arxiv.org/abs//2405.19874 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Sp...
Is In-Context Learning Sufficient for Instruction Following in LLMs? 01.06.2024 5:39
In-context learning (ICL) with URIAL aligns base LLMs using few examples but underperforms compared to instruction fine-tuning, with a proposed greedy selection approach improving performance. https://arxiv.org/abs//2405.19874 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Sp...
[QA] Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities 31.05.2024 11:08
Kernel Language Entropy (KLE) method improves uncertainty quantification in Large Language Models (LLMs) by capturing semantic uncertainty, enhancing trustworthiness by detecting incorrect responses. https://arxiv.org/abs//2405.20003 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id169247...
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities 31.05.2024 12:19
Kernel Language Entropy (KLE) method improves uncertainty quantification in Large Language Models (LLMs) by capturing semantic uncertainty, enhancing trustworthiness by detecting incorrect responses. https://arxiv.org/abs//2405.20003 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id169247...
[QA] COSY: Evaluating Textual Explanations of Neurons 31.05.2024 9:48
The paper introduces COSY, a framework to evaluate textual explanations for neural network concepts. It uses generative models to assess explanation quality, revealing differences in existing methods. https://arxiv.org/abs//2405.20331 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924...
COSY: Evaluating Textual Explanations of Neurons 31.05.2024 12:25
The paper introduces COSY, a framework to evaluate textual explanations for neural network concepts. It uses generative models to assess explanation quality, revealing differences in existing methods. https://arxiv.org/abs//2405.20331 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924...
[QA] Nearest Neighbor Speculative Decoding for LLM Generation and Attribution 30.05.2024 8:07
https://arxiv.org/abs//2405.19325 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.