Igor Melnyk
Arxiv Papers
Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
Style over Substance: Failure Modes of LLM Judges in Alignment Benchmarking 24.09.2024 11:39
This study evaluates the effectiveness of LLM-judge preferences in improving alignment, finding no correlation with concrete metrics and highlighting biases in LLM judgments. https://arxiv.org/abs//2409.15268 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://pod...
[QA] LLM Surgery: Efficient Knowledge Unlearning and Editing in Large Language Models 23.09.2024 7:46
This paper introduces LLM Surgery, a framework for efficiently modifying large language models to unlearn outdated information and integrate new knowledge without complete retraining, demonstrating significant performance improvements. https://arxiv.org/abs//2409.13054 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple....
LLM Surgery: Efficient Knowledge Unlearning and Editing in Large Language Models 23.09.2024 13:56
This paper introduces LLM Surgery, a framework for efficiently modifying large language models to unlearn outdated information and integrate new knowledge without complete retraining, demonstrating significant performance improvements. https://arxiv.org/abs//2409.13054 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple....
[QA] Embedding Geometries of Contrastive Language-Image Pre-Training 23.09.2024 7:35
This paper explores alternative geometries and softmax logits for language-image pre-training, finding that Euclidean CLIP (EuCLIP) performs as well as or better than the original CLIP. https://arxiv.org/abs//2409.13079 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify:...
Embedding Geometries of Contrastive Language-Image Pre-Training 23.09.2024 15:25
This paper explores alternative geometries and softmax logits for language-image pre-training, finding that Euclidean CLIP (EuCLIP) performs as well as or better than the original CLIP. https://arxiv.org/abs//2409.13079 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify:...
[QA] Kolmogorov–Arnold Transformer 21.09.2024 8:14
The Kolmogorov–Arnold Transformer (KAT) enhances transformer performance by replacing MLP layers with Kolmogorov-Arnold Network layers, addressing key challenges and demonstrating superior results in various tasks. https://arxiv.org/abs//2409.10594 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-...
Kolmogorov–Arnold Transformer 21.09.2024 15:05
The Kolmogorov–Arnold Transformer (KAT) enhances transformer performance by replacing MLP layers with Kolmogorov-Arnold Network layers, addressing key challenges and demonstrating superior results in various tasks. https://arxiv.org/abs//2409.10594 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-...
Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think 21.09.2024 11:52
This paper reveals a flaw in the inference pipeline of diffusion models for depth estimation, leading to a 2002#2 speed improvement and superior performance through end-to-end fine-tuning. https://arxiv.org/abs//2409.11355 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotif...
[QA] Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think 21.09.2024 6:42
This paper reveals a flaw in the inference pipeline of diffusion models for depth estimation, leading to a 2002#2 speed improvement and superior performance through end-to-end fine-tuning. https://arxiv.org/abs//2409.11355 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotif...
[QA] Re-Introducing LayerNorm: Geometric Meaning, Irreversibility and a Comparative Study with RMSNorm 20.09.2024 7:03
This paper explores the geometric implications of LayerNorm in transformers, revealing its irreversibility and redundancy, and advocates for RMSNorm as a more efficient alternative with similar performance. https://arxiv.org/abs//2409.12951 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...
Re-Introducing LayerNorm: Geometric Meaning, Irreversibility and a Comparative Study with RMSNorm 20.09.2024 12:28
This paper explores the geometric implications of LayerNorm in transformers, revealing its irreversibility and redundancy, and advocates for RMSNorm as a more efficient alternative with similar performance. https://arxiv.org/abs//2409.12951 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...
[QA] Is Tokenization Needed for Masked Particle Modelling? 20.09.2024 7:52
This paper enhances masked particle modeling (MPM) for high-energy physics, improving performance through better implementation and a powerful decoder, outperforming previous methods in various jet physics tasks. https://arxiv.org/abs//2409.12589 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-pa...
Is Tokenization Needed for Masked Particle Modelling? 20.09.2024 20:39
This paper enhances masked particle modeling (MPM) for high-energy physics, improving performance through better implementation and a powerful decoder, outperforming previous methods in various jet physics tasks. https://arxiv.org/abs//2409.12589 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-pa...
[QA] Finetuning Language Models to Emit Linguistic Expressions of Uncertainty 19.09.2024 6:49
https://arxiv.org/abs//2409.12180 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
Finetuning Language Models to Emit Linguistic Expressions of Uncertainty 19.09.2024 12:41
https://arxiv.org/abs//2409.12180 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
[QA] To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning 19.09.2024 7:23
Chain-of-thought prompting enhances reasoning in large language models, particularly for math and logic tasks, but shows limited benefits for other tasks, suggesting a need for new computational paradigms. https://arxiv.org/abs//2409.12183 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id...
To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning 19.09.2024 26:23
Chain-of-thought prompting enhances reasoning in large language models, particularly for math and logic tasks, but shows limited benefits for other tasks, suggesting a need for new computational paradigms. https://arxiv.org/abs//2409.12183 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id...
[QA] On the limits of agency in agent-based models 18.09.2024 8:12
AgentTorch is a framework that enhances agent-based modeling by using large language models to simulate millions of agents, demonstrating its utility in analyzing complex systems like the COVID-19 pandemic. https://arxiv.org/abs//2409.10568 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...
On the limits of agency in agent-based models 18.09.2024 19:50
AgentTorch is a framework that enhances agent-based modeling by using large language models to simulate millions of agents, demonstrating its utility in analyzing complex systems like the COVID-19 pandemic. https://arxiv.org/abs//2409.10568 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/i...
[QA] Promptriever: Instruction-Trained Retrievers Can Be Prompted Like Language Models 18.09.2024 7:14
Promptriever is a novel retrieval model that follows instructions, achieving state-of-the-art performance and improved robustness, demonstrating the potential of prompting in information retrieval. https://arxiv.org/abs//2409.11136 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924760...
Promptriever: Instruction-Trained Retrievers Can Be Prompted Like Language Models 18.09.2024 15:25
Promptriever is a novel retrieval model that follows instructions, achieving state-of-the-art performance and improved robustness, demonstrating the potential of prompting in information retrieval. https://arxiv.org/abs//2409.11136 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16924760...
[QA] Finetuning CLIP to Reason about Pairwise Differences 17.09.2024 7:34
This paper enhances CLIP's contrastive learning by aligning image embeddings with text descriptions, improving image ranking, zero-shot classification, and introducing comparative prompting for better performance and geometric properties. https://arxiv.org/abs//2409.09721 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts...
Finetuning CLIP to Reason about Pairwise Differences 17.09.2024 16:47
This paper enhances CLIP's contrastive learning by aligning image embeddings with text descriptions, improving image ranking, zero-shot classification, and introducing comparative prompting for better performance and geometric properties. https://arxiv.org/abs//2409.09721 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts...
[QA] Think Twice Before You Act: Improving Inverse Problem Solving With MCMC 16.09.2024 9:00
The paper introduces Diffusion Posterior MCMC (DPMC), an improved algorithm for solving inverse problems using pretrained diffusion models, outperforming existing methods and reducing errors in high noise scenarios. https://arxiv.org/abs//2409.08551 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv...
Think Twice Before You Act: Improving Inverse Problem Solving With MCMC 16.09.2024 11:16
The paper introduces Diffusion Posterior MCMC (DPMC), an improved algorithm for solving inverse problems using pretrained diffusion models, outperforming existing methods and reducing errors in high noise scenarios. https://arxiv.org/abs//2409.08551 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.