Igor Melnyk

Arxiv Papers

Science EN ↓ 2489 episodes

Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers

Author

Igor Melnyk

Category

Science

Podcast website

github.com

Latest episode

Sep 1, 2025

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

Rethinking FID: Towards a Better Evaluation Metric for Image Generation 20.01.2024

The paper criticizes the Fréchet Inception Distance (FID) as an evaluation metric for generated images and proposes an alternative metric called CMMD, which is based on CLIP embeddings and offers a more reliable assessment of image quality. https://arxiv.org/abs//2401.09603 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.a...

[short] ChatQA: Building GPT-4 Level Conversational QA Models 19.01.2024

The paper introduces ChatQA, a conversational question answering model that achieves GPT-4 level accuracies. It proposes a two-stage instruction tuning method and a dense retriever for retrieval in conversational QA. ChatQA-70B outperforms GPT-4 without using synthetic data. https://arxiv.org/abs//2401.10225 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers...

ChatQA: Building GPT-4 Level Conversational QA Models 19.01.2024

The paper introduces ChatQA, a conversational question answering model that achieves GPT-4 level accuracies. It proposes a two-stage instruction tuning method and a dense retriever for retrieval in conversational QA. ChatQA-70B outperforms GPT-4 without using synthetic data. https://arxiv.org/abs//2401.10225 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers...

[short] Improving fine-grained understanding in image-text pre-training 19.01.2024

The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease based on electronic health records. https://arxiv.org/abs//2401.09865 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv...

Improving fine-grained understanding in image-text pre-training 19.01.2024

The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease based on electronic health records. https://arxiv.org/abs//2401.09865 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv...

[short] Self-Rewarding Language Models 19.01.2024

To achieve superhuman agents, future models need superhuman feedback. This paper proposes Self-Rewarding Language Models, where the model provides its own rewards during training, leading to improved performance and the potential for continual improvement. https://arxiv.org/abs//2401.10020 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: ht...

Self-Rewarding Language Models 19.01.2024

To achieve superhuman agents, future models need superhuman feedback. This paper proposes Self-Rewarding Language Models, where the model provides its own rewards during training, leading to improved performance and the potential for continual improvement. https://arxiv.org/abs//2401.10020 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: ht...

[short] REFT: Reasoning with REinforced Fine-Tuning 18.01.2024

The paper proposes a method called Reinforced Fine-Tuning (ReFT) to enhance the generalizability of Large Language Models (LLMs) for reasoning tasks, using math problem-solving as an example. ReFT combines Supervised Fine-Tuning (SFT) with reinforcement learning and outperforms SFT in experiments. https://arxiv.org/abs//2401.08967 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.t...

REFT: Reasoning with REinforced Fine-Tuning 18.01.2024

The paper proposes a method called Reinforced Fine-Tuning (ReFT) to enhance the generalizability of Large Language Models (LLMs) for reasoning tasks, using math problem-solving as an example. ReFT combines Supervised Fine-Tuning (SFT) with reinforcement learning and outperforms SFT in experiments. https://arxiv.org/abs//2401.08967 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.t...

[short] Asynchronous Local-SGD Training for Language Modeling 18.01.2024

The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease based on electronic health records. https://arxiv.org/abs//2401.09135 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv...

Asynchronous Local-SGD Training for Language Modeling 18.01.2024

The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease based on electronic health records. https://arxiv.org/abs//2401.09135 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv...

[short] DeepSpeed-FastGen: High-throughput Text Generation for LLMs via MII and DeepSpeed-Inference 18.01.2024

DeepSpeed-FastGen is a system that uses Dynamic SplitFuse to improve the deployment and scaling of large language models, achieving higher throughput and lower latency compared to existing systems. The code is available for community engagement. https://arxiv.org/abs//2401.08671 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podca...

DeepSpeed-FastGen: High-throughput Text Generation for LLMs via MII and DeepSpeed-Inference 18.01.2024

DeepSpeed-FastGen is a system that uses Dynamic SplitFuse to improve the deployment and scaling of large language models, achieving higher throughput and lower latency compared to existing systems. The code is available for community engagement. https://arxiv.org/abs//2401.08671 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podca...

[short] Tuning Language Models by Proxy 17.01.2024

Proxy-tuning is a lightweight decoding-time algorithm that can be used to customize large pretrained language models without accessing their weights. It achieves similar results to direct tuning and can be applied for domain adaptation and task-specific finetuning. https://arxiv.org/abs//2401.08565 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Pod...

Tuning Language Models by Proxy 17.01.2024

Proxy-tuning is a lightweight decoding-time algorithm that can be used to customize large pretrained language models without accessing their weights. It achieves similar results to direct tuning and can be applied for domain adaptation and task-specific finetuning. https://arxiv.org/abs//2401.08565 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Pod...

[short] Scalable Pre-training of Large Autoregressive Image Models 17.01.2024

This paper introduces AIM, a collection of vision models pre-trained with an autoregressive objective. The models exhibit similar scaling properties to Large Language Models (LLMs) and achieve high performance on downstream tasks. Pre-training AIM does not require image-specific strategies. https://arxiv.org/abs//2401.08541 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.c...

Scalable Pre-training of Large Autoregressive Image Models 17.01.2024

This paper introduces AIM, a collection of vision models pre-trained with an autoregressive objective. The models exhibit similar scaling properties to Large Language Models (LLMs) and achieve high performance on downstream tasks. Pre-training AIM does not require image-specific strategies. https://arxiv.org/abs//2401.08541 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.c...

[short] A Closer Look at AUROC and AUPRC under Class Imbalance 16.01.2024

The paper challenges the belief that the area under the precision-recall curve (AUPRC) is a superior metric for model comparison in machine learning. It argues that AUPRC can be biased and highlights a lack of empirical evidence supporting its supposed advantages. https://arxiv.org/abs//2401.06091 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podc...

A Closer Look at AUROC and AUPRC under Class Imbalance 16.01.2024

The paper challenges the belief that the area under the precision-recall curve (AUPRC) is a superior metric for model comparison in machine learning. It argues that AUPRC can be biased and highlights a lack of empirical evidence supporting its supposed advantages. https://arxiv.org/abs//2401.06091 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podc...

[short] Nash Learning from Human Feedback 16.01.2024

The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease in patients with diabetes. https://arxiv.org/abs//2312.00886 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers

Nash Learning from Human Feedback 16.01.2024

The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease in patients with diabetes. https://arxiv.org/abs//2312.00886 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers

[short] The Unreasonable Effectiveness of Easy Training Data for Hard Tasks 15.01.2024

Current language models often generalize well from easy to hard data, performing as well as models trained on hard data. It may be better to collect and train on easy data rather than hard data, as hard data is noisier and costlier to collect. https://arxiv.org/abs//2401.06751 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcast...

The Unreasonable Effectiveness of Easy Training Data for Hard Tasks 15.01.2024

Current language models often generalize well from easy to hard data, performing as well as models trained on hard data. It may be better to collect and train on easy data rather than hard data, as hard data is noisier and costlier to collect. https://arxiv.org/abs//2401.06751 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcast...

[short] AboutMe: Using Self-Descriptions in Webpages to Document the Effects of English Pretraining Data Filters 15.01.2024

The paper examines the data curation process for large language models and investigates how different filters affect webpages based on social and geographic dimensions. The study highlights implicit preferences in data curation and calls for further research on the social implications of pretraining data curation. https://arxiv.org/abs//2401.06408 YouTube: https://www.youtube.com/@ArxivPapers TikT...

AboutMe: Using Self-Descriptions in Webpages to Document the Effects of English Pretraining Data Filters 15.01.2024

The paper examines the data curation process for large language models and investigates how different filters affect webpages based on social and geographic dimensions. The study highlights implicit preferences in data curation and calls for further research on the social implications of pretraining data curation. https://arxiv.org/abs//2401.06408 YouTube: https://www.youtube.com/@ArxivPapers TikT...

Listen to the Arxiv Papers podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.