Igor Melnyk

Arxiv Papers

Science EN ↓ 2489 episodes

Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers

Author

Igor Melnyk

Category

Science

Podcast website

github.com

Latest episode

Sep 1, 2025

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

[short] Evolutionary Optimization of Model Merging Recipes 23.03.2024

Evolutionary algorithms automate creating powerful foundation models by merging diverse open-source models, achieving state-of-the-art performance in Japanese language and culture tasks. https://arxiv.org/abs//2403.13187 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify:...

Evolutionary Optimization of Model Merging Recipes 23.03.2024

Evolutionary algorithms automate creating powerful foundation models by merging diverse open-source models, achieving state-of-the-art performance in Japanese language and culture tasks. https://arxiv.org/abs//2403.13187 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify:...

[short] Recourse for Reclamation: Chatting with Generative Language Models 23.03.2024

Toxicity scoring in language models can hinder information access and cultural norms. A novel recourse mechanism allows users to set toxicity thresholds dynamically for improved interaction. https://arxiv.org/abs//2403.14467 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spot...

Recourse for Reclamation: Chatting with Generative Language Models 23.03.2024

Toxicity scoring in language models can hinder information access and cultural norms. A novel recourse mechanism allows users to set toxicity thresholds dynamically for improved interaction. https://arxiv.org/abs//2403.14467 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spot...

[short] A Unified Framework for Model Editing 22.03.2024

The paper introduces a framework unifying ROME and MEMIT model editing techniques under a preservation-memorization objective, presenting EMMET as a new batched memory-editing algorithm for Transformers. https://arxiv.org/abs//2403.14236 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16...

A Unified Framework for Model Editing 22.03.2024

The paper introduces a framework unifying ROME and MEMIT model editing techniques under a preservation-memorization objective, presenting EMMET as a new batched memory-editing algorithm for Transformers. https://arxiv.org/abs//2403.14236 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id16...

[short] Language Models Can Reduce Asymmetry in Information Markets 22.03.2024

The paper explores the buyer's inspection paradox in information markets using simulated agents with language models, focusing on biases, pricing effects, and quality outcomes. https://arxiv.org/abs//2403.14443 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https...

Language Models Can Reduce Asymmetry in Information Markets 22.03.2024

The paper explores the buyer's inspection paradox in information markets using simulated agents with language models, focusing on biases, pricing effects, and quality outcomes. https://arxiv.org/abs//2403.14443 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https...

[short] AI and Memory Wall 22.03.2024

Unsupervised training data and neural scaling have led to larger models, but memory bandwidth is now the main bottleneck in AI applications, requiring architecture and strategy redesign. https://arxiv.org/abs//2403.14123 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify:...

AI and Memory Wall 22.03.2024

Unsupervised training data and neural scaling have led to larger models, but memory bandwidth is now the main bottleneck in AI applications, requiring architecture and strategy redesign. https://arxiv.org/abs//2403.14123 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify:...

RewardBench: Evaluating Reward Models for Language Modeling 21.03.2024

The paper introduces REWARDBENCH, a dataset and code-base for evaluating reward models in aligning language models with human preferences, shedding light on their capabilities and limitations. https://arxiv.org/abs//2403.13787 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Sp...

Evaluating Frontier Models for Dangerous Capabilities 21.03.2024

https://arxiv.org/abs//2403.13793 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers

Reverse Training to Nurse the Reversal Curse 21.03.2024

Large language models struggle with generalizing from "A has B" to "B is a feature of A" due to Zipf's law. Reverse training doubles tokens, improving performance on reversal tasks. https://arxiv.org/abs//2403.13799 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id...

[short] Chart-based Reasoning: Transferring Capabilities from LLMs to VLMs 20.03.2024

Technique proposed to transfer reasoning capabilities from large-language models to vision-language models, achieving state-of-the-art performance on ChartQA, PlotQA, and FigureQA tasks. https://arxiv.org/abs//2403.12596 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify:...

Chart-based Reasoning: Transferring Capabilities from LLMs to VLMs 20.03.2024

Technique proposed to transfer reasoning capabilities from large-language models to vision-language models, achieving state-of-the-art performance on ChartQA, PlotQA, and FigureQA tasks. https://arxiv.org/abs//2403.12596 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify:...

[short] Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers 20.03.2024

Vid2Robot enables robots to learn tasks from human video demonstrations, outperforming other methods by 20% and showcasing potential for real-world applications. https://arxiv.org/abs//2403.12943 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spoti...

Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers 20.03.2024

Vid2Robot enables robots to learn tasks from human video demonstrations, outperforming other methods by 20% and showcasing potential for real-world applications. https://arxiv.org/abs//2403.12943 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spoti...

[short] Yell At Your Robot Improving On-the-Fly from Language Corrections 20.03.2024

Hierarchical policies combining language and low-level control improve robotic task performance. Human feedback on language corrections enhances high-level policies, leading to significant performance gains in dexterous manipulation tasks. https://arxiv.org/abs//2403.12910 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.ap...

Yell At Your Robot Improving On-the-Fly from Language Corrections 20.03.2024

Hierarchical policies combining language and low-level control improve robotic task performance. Human feedback on language corrections enhances high-level policies, leading to significant performance gains in dexterous manipulation tasks. https://arxiv.org/abs//2403.12910 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.ap...

[short] Fast High-Resolution Image Synthesis with Latent Adversarial Diffusion Distillation 19.03.2024

Latent Adversarial Diffusion Distillation (LADD) improves image synthesis speed and quality compared to existing methods by leveraging generative features from pretrained latent diffusion models. https://arxiv.org/abs//2403.12015 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...

Fast High-Resolution Image Synthesis with Latent Adversarial Diffusion Distillation 19.03.2024

Latent Adversarial Diffusion Distillation (LADD) improves image synthesis speed and quality compared to existing methods by leveraging generative features from pretrained latent diffusion models. https://arxiv.org/abs//2403.12015 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016...

[short] Larimar: Large Language Models with Episodic Memory Control 19.03.2024

Larimar introduces a brain-inspired architecture for efficient knowledge updating in Large Language Models, achieving accuracy comparable to baselines with 4-10x speed-ups. https://arxiv.org/abs//2403.11901 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podca...

Larimar: Large Language Models with Episodic Memory Control 19.03.2024

Larimar introduces a brain-inspired architecture for efficient knowledge updating in Large Language Models, achieving accuracy comparable to baselines with 4-10x speed-ups. https://arxiv.org/abs//2403.11901 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podca...

[short] Mind Eye2: Shared-Subject Models Enable fMRI-To-Image With 1 Hour of Data 19.03.2024

Novel method enables high-quality visual perception reconstructions from minimal fMRI data by pretraining across subjects and using functional alignment, improving generalization and achieving state-of-the-art results. https://arxiv.org/abs//2403.11207 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/ar...

Mind Eye2: Shared-Subject Models Enable fMRI-To-Image With 1 Hour of Data 19.03.2024

Novel method enables high-quality visual perception reconstructions from minimal fMRI data by pretraining across subjects and using functional alignment, improving generalization and achieving state-of-the-art results. https://arxiv.org/abs//2403.11207 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/ar...

Listen to the Arxiv Papers podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.