Igor Melnyk
Arxiv Papers
Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
Small-scale proxies for large-scale Transformer training instabilities 26.09.2023 21:40
The paper investigates training instabilities in large Transformer-based models and explores ways to reproduce and study these instabilities at smaller scales. It examines sources of instability, explores the impact of learning rate and other interventions, and studies cases where instabilities can be predicted. https://arxiv.org/abs//2309.14322 YouTube: https://www.youtube.com/@ArxivPapers PODCAS...
[short] DEPT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning 25.09.2023 3:54
Decomposed Prompt Tuning (DEPT) is a method for parameter-efficient fine-tuning of language models. It achieves better performance while saving memory and time costs compared to other approaches, and is adaptable to different model architectures and sizes. https://arxiv.org/abs//2309.05173 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast...
DEPT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning 25.09.2023 22:34
Decomposed Prompt Tuning (DEPT) is a method for parameter-efficient fine-tuning of language models. It achieves better performance while saving memory and time costs compared to other approaches, and is adaptable to different model architectures and sizes. https://arxiv.org/abs//2309.05173 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast...
[short] Text2Reward: Automated Dense Reward Function Generation for Reinforcement Learning 25.09.2023 3:55
Text2Reward is a data-free framework using large language models to automate the creation of dense reward functions in reinforcement learning from natural language goals. It outperforms expert-written codes in multiple benchmarks and supports real-world deployment and human feedback refinement. https://arxiv.org/abs/2309.11489 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts:...
The Rise and Potential of Large Language Model Based Agents: A Survey 24.09.2023 23:16
This paper offers a comprehensive survey on large language model (LLM)-based AI agents, discussing their foundation, a general framework, and applications in various scenarios. It further delves into agent societies and highlights key challenges. https://arxiv.org/abs//2309.07864 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-pap...
[short] The Reversal Curse: LLMs trained on “A is B” fail to learn “B is A” 23.09.2023 2:19
Large language models (LLMs) exhibit a failure of generalization known as the Reversal Curse, where they struggle to answer questions in the reverse direction of their training data. This failure persists across different model sizes and families and is not improved by data augmentation. https://arxiv.org/abs//2309.12288 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https...
The Reversal Curse: LLMs trained on “A is B” fail to learn “B is A” 23.09.2023 16:53
Large language models (LLMs) exhibit a failure of generalization known as the Reversal Curse, where they struggle to answer questions in the reverse direction of their training data. This failure persists across different model sizes and families and is not improved by data augmentation. https://arxiv.org/abs//2309.12288 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https...
[short] LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models 23.09.2023 2:46
The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease in patients with diabetes. https://arxiv.org/abs//2309.12307 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models 22.09.2023 16:18
The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease in patients with diabetes. https://arxiv.org/abs//2309.12307 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
[short] BTLM-3B-8K: 7B Parameter Performance in a 3B Parameter Model 22.09.2023 3:14
The paper introduces BTLM-3B-8K, a new state-of-the-art 3 billion parameter language model that outperforms existing models and provides excellent long context performance. It is compact, requiring less memory and compute, making it accessible for mobile and edge devices. https://arxiv.org/abs//2309.11568 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.appl...
BTLM-3B-8K: 7B Parameter Performance in a 3B Parameter Model 22.09.2023 24:40
The paper introduces BTLM-3B-8K, a new state-of-the-art 3 billion parameter language model that outperforms existing models and provides excellent long context performance. It is compact, requiring less memory and compute, making it accessible for mobile and edge devices. https://arxiv.org/abs//2309.11568 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.appl...
[short] FreeU: Free Lunch in Diffusion U-Net 21.09.2023 2:41
The paper introduces a method called "FreeU" that improves generation quality in diffusion models without additional training or finetuning. It strategically re-weights the contributions from skip connections and backbone feature maps. https://arxiv.org/abs//2309.11497 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-pape...
FreeU: Free Lunch in Diffusion U-Net 21.09.2023 16:25
The paper introduces a method called "FreeU" that improves generation quality in diffusion models without additional training or finetuning. It strategically re-weights the contributions from skip connections and backbone feature maps. https://arxiv.org/abs//2309.11497 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-pape...
[short] Chain-of-Verification Reduces Hallucination in Large Language Models 21.09.2023 2:34
The study addresses hallucinations in language models, introducing the Chain-of-Verification (CoVe) method. This process drafts, fact-checks, and verifies responses, effectively reducing hallucinations in various tasks. https://arxiv.org/abs//2309.11495 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: h...
Chain-of-Verification Reduces Hallucination in Large Language Models 21.09.2023 20:40
The study addresses hallucinations in language models, introducing the Chain-of-Verification (CoVe) method. This process drafts, fact-checks, and verifies responses, effectively reducing hallucinations in various tasks. https://arxiv.org/abs//2309.11495 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: h...
[short] Language Modeling Is Compression 20.09.2023 2:48
https://arxiv.org/abs//2309.10668 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
[full] Language Modeling Is Compression 20.09.2023 25:51
https://arxiv.org/abs//2309.10668 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
[short] PDFTriage: Question Answering over Long, Structured Documents 19.09.2023 4:08
The paper introduces PDFTriage, an approach to improve document question answering by considering the structure and content of structured documents. Experimental results show its effectiveness, and a benchmark dataset is released for further research. https://arxiv.org/abs//2309.08872 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxi...
[full] PDFTriage: Question Answering over Long, Structured Documents 19.09.2023 17:13
The paper introduces PDFTriage, an approach to improve document question answering by considering the structure and content of structured documents. Experimental results show its effectiveness, and a benchmark dataset is released for further research. https://arxiv.org/abs//2309.08872 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxi...
[short] Contrastive Decoding Improves Reasoning in Large Language Models 19.09.2023 2:30
Contrastive Decoding, a training-free text generation method, improves reasoning tasks by maximizing the difference in likelihood between strong and weak models, outperforming other methods on various benchmarks. https://arxiv.org/abs//2309.09117 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://...
[full] Contrastive Decoding Improves Reasoning in Large Language Models 19.09.2023 15:51
Contrastive Decoding, a training-free text generation method, improves reasoning tasks by maximizing the difference in likelihood between strong and weak models, outperforming other methods on various benchmarks. https://arxiv.org/abs//2309.09117 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://...
Sparse Autoencoders Find Highly Interpretable Features in Language Models 18.09.2023 13:50
The paper proposes a method to identify and interpret the directions in activation space of neural networks, addressing the issue of polysemanticity. The method uses sparse autoencoders to reconstruct internal activations and achieves more interpretable and monosemantic results compared to alternative approaches. This method can enable precise model editing and improve model transparency and steer...
Scaling Laws for Sparsely-Connected Foundation Models 18.09.2023 31:53
The paper explores the impact of parameter sparsity on the scaling behavior of Transformers trained on massive datasets. It identifies a scaling law that describes the relationship between weight sparsity, number of non-zero parameters, and amount of training data. The findings provide insights into the optimal sparsity level for computational efficiency improvements. https://arxiv.org/abs//2309.0...
Agents: An Open-source Framework for Autonomous Language Agents 16.09.2023 14:50
The paper introduces Agents, an open-source library that allows non-specialists to build and deploy autonomous language agents with advanced features, and provides a link to access the library. https://arxiv.org/abs//2309.07870 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify....
Statistical rejection sampling improves preference optimization 14.09.2023 28:40
This paper introduces a new approach called Statistical Rejection Sampling Optimization (RSO) to improve the alignment of language models with human preferences. RSO outperforms existing methods on evaluations from both language models and human raters. https://arxiv.org/abs//2309.06657 YouTube: https://www.youtube.com/@ArxivPapers PODCASTS: Apple Podcasts: https://podcasts.apple.com/us/podcast/ar...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.