Igor Melnyk

Arxiv Papers

Science EN ↓ 2489 episodes

Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers

Author

Igor Melnyk

Category

Science

Podcast website

github.com

Latest episode

Sep 1, 2025

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

To Believe or Not to Believe Your LLM 05.06.2024

The paper explores quantifying uncertainty in large language models to detect unreliable responses, focusing on distinguishing epistemic and aleatoric uncertainties using an information-theoretic metric. https://arxiv.org/abs//2406.02543 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16...

[QA] Learning to grok: Emergence of in-context learning and skill composition in modular arithmetic tasks 05.06.2024

Study explores in-context learning and skill composition in modular arithmetic tasks using GPT-style transformer models, showing transition from in-distribution to out-of-distribution generalization with increasing pre-training tasks. https://arxiv.org/abs//2406.02550 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.c...

Learning to grok: Emergence of in-context learning and skill composition in modular arithmetic tasks 05.06.2024

Study explores in-context learning and skill composition in modular arithmetic tasks using GPT-style transformer models, showing transition from in-distribution to out-of-distribution generalization with increasing pre-training tasks. https://arxiv.org/abs//2406.02550 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.c...

[QA] Show, Don't Tell: Aligning Language Models with Demonstrated Feedback 04.06.2024

Demonstration ITerated Task Optimization (DITTO) aligns language models to specific settings using a few demonstrations, outperforming other methods by 19% on average. https://arxiv.org/abs//2406.00888 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters...

Show, Don't Tell: Aligning Language Models with Demonstrated Feedback 04.06.2024

Demonstration ITerated Task Optimization (DITTO) aligns language models to specific settings using a few demonstrations, outperforming other methods by 19% on average. https://arxiv.org/abs//2406.00888 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters...

[QA] What makes unlearning hard and what to do about it 04.06.2024

Research explores machine unlearning, removing specific training data from models. Investigates factors affecting unlearning difficulty and algorithm performance, introducing Refined-Unlearning Meta-algorithm (RUM) to enhance existing methods. https://arxiv.org/abs//2406.01257 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcast...

What makes unlearning hard and what to do about it 04.06.2024

Research explores machine unlearning, removing specific training data from models. Investigates factors affecting unlearning difficulty and algorithm performance, introducing Refined-Unlearning Meta-algorithm (RUM) to enhance existing methods. https://arxiv.org/abs//2406.01257 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcast...

[QA] Learning to Play Atari in a World of Tokens 04.06.2024

DART introduces discrete abstract representations for transformer-based learning, improving sample efficiency and outperforming previous methods on Atari games. https://arxiv.org/abs//2406.01361 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotif...

Learning to Play Atari in a World of Tokens 04.06.2024

DART introduces discrete abstract representations for transformer-based learning, improving sample efficiency and outperforming previous methods on Atari games. https://arxiv.org/abs//2406.01361 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotif...

[QA] Perplexed by Perplexity: Perplexity-Based Data Pruning With Small Reference Models 03.06.2024

Investigating if small language models can improve large models by pruning datasets based on perplexity, leading to enhanced downstream task performance and reduced pretraining steps. https://arxiv.org/abs//2405.20541 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: ht...

Perplexed by Perplexity: Perplexity-Based Data Pruning With Small Reference Models 03.06.2024

Investigating if small language models can improve large models by pruning datasets based on perplexity, leading to enhanced downstream task performance and reduced pretraining steps. https://arxiv.org/abs//2405.20541 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: ht...

[QA] Contextual Position Encoding: Learning to Count What's Important 02.06.2024

Large Language Models (LLMs) use attention mechanisms for token interactions. Current position encoding methods lack abstraction levels. CoPE introduces contextual position encoding for improved addressing and performance. https://arxiv.org/abs//2405.18719 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcas...

Contextual Position Encoding: Learning to Count What's Important 02.06.2024

Large Language Models (LLMs) use attention mechanisms for token interactions. Current position encoding methods lack abstraction levels. CoPE introduces contextual position encoding for improved addressing and performance. https://arxiv.org/abs//2405.18719 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcas...

[QA] JINA CLIP: Your CLIP Model Is Also Your Text Retriever 02.06.2024

Proposed multi-task contrastive training method improves CLIP models for text-only tasks, achieving state-of-the-art performance on text-image and text-text retrieval tasks with jina-clip-v1 model. https://arxiv.org/abs//2405.20204 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924760...

JINA CLIP: Your CLIP Model Is Also Your Text Retriever 02.06.2024

Proposed multi-task contrastive training method improves CLIP models for text-only tasks, achieving state-of-the-art performance on text-image and text-text retrieval tasks with jina-clip-v1 model. https://arxiv.org/abs//2405.20204 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924760...

[QA] Large Language Models Can Self-Improve At Web Agent Tasks 01.06.2024

Large language models (LLMs) self-improve to navigate web environments using synthetic data, achieving 31% task completion rate improvement on WebArena benchmark, introducing new evaluation metrics. https://arxiv.org/abs//2405.20309 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476...

Large Language Models Can Self-Improve At Web Agent Tasks 01.06.2024

Large language models (LLMs) self-improve to navigate web environments using synthetic data, achieving 31% task completion rate improvement on WebArena benchmark, introducing new evaluation metrics. https://arxiv.org/abs//2405.20309 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476...

[QA] Is In-Context Learning Sufficient for Instruction Following in LLMs? 01.06.2024

In-context learning (ICL) with URIAL aligns base LLMs using few examples but underperforms compared to instruction fine-tuning, with a proposed greedy selection approach improving performance. https://arxiv.org/abs//2405.19874 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Sp...

Is In-Context Learning Sufficient for Instruction Following in LLMs? 01.06.2024

In-context learning (ICL) with URIAL aligns base LLMs using few examples but underperforms compared to instruction fine-tuning, with a proposed greedy selection approach improving performance. https://arxiv.org/abs//2405.19874 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Sp...

[QA] Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities 31.05.2024

Kernel Language Entropy (KLE) method improves uncertainty quantification in Large Language Models (LLMs) by capturing semantic uncertainty, enhancing trustworthiness by detecting incorrect responses. https://arxiv.org/abs//2405.20003 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id169247...

Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities 31.05.2024

Kernel Language Entropy (KLE) method improves uncertainty quantification in Large Language Models (LLMs) by capturing semantic uncertainty, enhancing trustworthiness by detecting incorrect responses. https://arxiv.org/abs//2405.20003 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id169247...

[QA] COSY: Evaluating Textual Explanations of Neurons 31.05.2024

The paper introduces COSY, a framework to evaluate textual explanations for neural network concepts. It uses generative models to assess explanation quality, revealing differences in existing methods. https://arxiv.org/abs//2405.20331 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924...

COSY: Evaluating Textual Explanations of Neurons 31.05.2024

The paper introduces COSY, a framework to evaluate textual explanations for neural network concepts. It uses generative models to assess explanation quality, revealing differences in existing methods. https://arxiv.org/abs//2405.20331 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924...

[QA] Nearest Neighbor Speculative Decoding for LLM Generation and Attribution 30.05.2024

https://arxiv.org/abs//2405.19325 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers

Nearest Neighbor Speculative Decoding for LLM Generation and Attribution 30.05.2024

Listen to the Arxiv Papers podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.