Igor Melnyk
Arxiv Papers
Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
[QA] AutoGluon-Multimodal (AutoMM): Supercharging Multimodal AutoML with Foundation Models 26.04.2024 9:02
AutoGluon-Multimodal (AutoMM) is an open-source AutoML library for multimodal learning, offering easy fine-tuning with three lines of code. It supports various modalities and excels in basic and advanced tasks. https://arxiv.org/abs//2404.16233 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-pape...
AutoGluon-Multimodal (AutoMM): Supercharging Multimodal AutoML with Foundation Models 26.04.2024 18:52
AutoGluon-Multimodal (AutoMM) is an open-source AutoML library for multimodal learning, offering easy fine-tuning with three lines of code. It supports various modalities and excels in basic and advanced tasks. https://arxiv.org/abs//2404.16233 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-pape...
[QA] Weak-to-Strong Extrapolation Expedites Alignment 26.04.2024 10:33
The paper introduces EXPO, a method to enhance large language models' alignment with human preference by extrapolating from weaker models, showing improved performance without additional training. https://arxiv.org/abs//2404.16792 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924...
Weak-to-Strong Extrapolation Expedites Alignment 26.04.2024 16:44
The paper introduces EXPO, a method to enhance large language models' alignment with human preference by extrapolating from weaker models, showing improved performance without additional training. https://arxiv.org/abs//2404.16792 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924...
[QA] Studying Large Language Model Behaviors Under Realistic Knowledge Conflicts 25.04.2024 8:38
Retrieval-augmented generation (RAG) addresses issues in language models, but conflicts can arise between parametric knowledge and context, affecting knowledge updates. https://arxiv.org/abs//2404.16032 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcaster...
Studying Large Language Model Behaviors Under Realistic Knowledge Conflicts 25.04.2024 11:34
Retrieval-augmented generation (RAG) addresses issues in language models, but conflicts can arise between parametric knowledge and context, affecting knowledge updates. https://arxiv.org/abs//2404.16032 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcaster...
[QA] Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach 25.04.2024 9:39
This paper addresses uncertainty estimation and calibration for large language models, proposing a supervised approach utilizing labeled data to enhance reliability and accuracy in LLM outputs. https://arxiv.org/abs//2404.15993 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 S...
Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach 25.04.2024 27:41
This paper addresses uncertainty estimation and calibration for large language models, proposing a supervised approach utilizing labeled data to enhance reliability and accuracy in LLM outputs. https://arxiv.org/abs//2404.15993 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 S...
[QA] OpenELM: An Efficient Language Model Family with Open-source Training and Inference Framework 24.04.2024 8:13
OpenELM, a state-of-the-art open language model, enhances accuracy using layer-wise scaling. Released with complete training framework, it empowers open research community. Available on GitHub and HuggingFace. https://arxiv.org/abs//2404.14619 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-paper...
OpenELM: An Efficient Language Model Family with Open-source Training and Inference Framework 24.04.2024 8:58
OpenELM, a state-of-the-art open language model, enhances accuracy using layer-wise scaling. Released with complete training framework, it empowers open research community. Available on GitHub and HuggingFace. https://arxiv.org/abs//2404.14619 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-paper...
[QA] Achieving >97% on GSM8K: Deeply Understanding the Problems Makes LLMs Perfect Reasoners 24.04.2024 7:30
The paper introduces DUP prompting strategy to improve Large Language Models' performance on complex reasoning tasks, outperforming Zero-Shot CoT on diverse datasets, achieving state-of-the-art results. https://arxiv.org/abs//2404.14963 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...
Achieving >97% on GSM8K: Deeply Understanding the Problems Makes LLMs Perfect Reasoners 24.04.2024 10:56
The paper introduces DUP prompting strategy to improve Large Language Models' performance on complex reasoning tasks, outperforming Zero-Shot CoT on diverse datasets, achieving state-of-the-art results. https://arxiv.org/abs//2404.14963 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...
[QA] SnapKV: LLM Knows What You are Looking for Before Generation 24.04.2024 10:00
SnapKV is a fine-tuning-free method that efficiently reduces Key-Value cache size in Large Language Models, maintaining performance while enhancing memory and time efficiency for long input sequences. https://arxiv.org/abs//2404.14469 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924...
SnapKV: LLM Knows What You are Looking for Before Generation 24.04.2024 17:09
SnapKV is a fine-tuning-free method that efficiently reduces Key-Value cache size in Large Language Models, maintaining performance while enhancing memory and time efficiency for long input sequences. https://arxiv.org/abs//2404.14469 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924...
[QA] Multi-Head Mixture-of-Experts 24.04.2024 8:05
MH-MoE addresses low expert activation and lack of fine-grained analysis in SMoE by using a multi-head mechanism to enhance context understanding and expert activation. https://arxiv.org/abs//2404.15045 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcaster...
Multi-Head Mixture-of-Experts 24.04.2024 14:15
MH-MoE addresses low expert activation and lack of fine-grained analysis in SMoE by using a multi-head mechanism to enhance context understanding and expert activation. https://arxiv.org/abs//2404.15045 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcaster...
[QA] The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions 23.04.2024 9:21
LLMs are vulnerable to attacks due to equal priority given to all prompts. Proposed instruction hierarchy teaches models to ignore lower-priority instructions, enhancing robustness with minimal impact on capabilities. https://arxiv.org/abs//2404.13208 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arx...
The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions 23.04.2024 13:06
LLMs are vulnerable to attacks due to equal priority given to all prompts. Proposed instruction hierarchy teaches models to ignore lower-priority instructions, enhancing robustness with minimal impact on capabilities. https://arxiv.org/abs//2404.13208 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arx...
[QA] Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data 23.04.2024 9:37
https://arxiv.org/abs//2404.14367 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data 23.04.2024 18:47
https://arxiv.org/abs//2404.14367 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
[QA] Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone 23.04.2024 9:02
Introducing phi-3-mini, a high-performing language model trained on a large dataset, with smaller versions phi-3-small and phi-3-medium showing even better performance. https://arxiv.org/abs//2404.14219 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcaster...
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone 23.04.2024 5:58
Introducing phi-3-mini, a high-performing language model trained on a large dataset, with smaller versions phi-3-small and phi-3-medium showing even better performance. https://arxiv.org/abs//2404.14219 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcaster...
[QA] Towards Reliable Latent Knowledge Estimation in LLMs: In-Context Learning vs. Prompting Based Factual Knowledge Extraction 22.04.2024 8:22
Approach estimates latent knowledge in large language models using in-context learning, showing differences in factual knowledge across models and sizes. https://arxiv.org/abs//2404.12957 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/p...
Towards Reliable Latent Knowledge Estimation in LLMs: In-Context Learning vs. Prompting Based Factual Knowledge Extraction 22.04.2024 23:51
Approach estimates latent knowledge in large language models using in-context learning, showing differences in factual knowledge across models and sizes. https://arxiv.org/abs//2404.12957 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/p...
[QA] HalluciBot: Is There No Such Thing as a Bad Question? 22.04.2024 11:45
HalluciBot predicts hallucination probability before generation in Large Language Models, aiding in query quality assessment and user accountability, potentially reducing computational waste. https://arxiv.org/abs//2404.12535 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spo...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.