Igor Melnyk
Arxiv Papers
Running out of time to catch up with new arXiv papers? We take the most impactful papers and present them as convenient podcasts. If you're a visual learner, we offer these papers in an engaging video format. Our service fills the gap between overly brief paper summaries and time-consuming full paper reads. You gain academic insights in a time-efficient, digestible format. Code behind this work: https://github.com/imelnyk/ArxivPapers
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
[short] How to Guess a Gradient 11.12.2023 2:12
The paper explores the structure of gradients in neural networks and how exploiting this structure can improve gradient-free optimization schemes. It also discusses the challenges in narrowing the gap between exact gradient optimization and guessing the gradients. https://arxiv.org/abs//2312.04709 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podc...
How to Guess a Gradient 11.12.2023 24:18
The paper explores the structure of gradients in neural networks and how exploiting this structure can improve gradient-free optimization schemes. It also discusses the challenges in narrowing the gap between exact gradient optimization and guessing the gradients. https://arxiv.org/abs//2312.04709 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podc...
[short] Using Large Language Models for Hyperparameter Optimization 09.12.2023 2:24
This paper explores using large language models for hyperparameter optimization and finds that they can perform as well or better than traditional methods, suggesting they are a promising tool for improving efficiency in this area. https://arxiv.org/abs//2312.04528 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/...
Using Large Language Models for Hyperparameter Optimization 09.12.2023 18:37
This paper explores using large language models for hyperparameter optimization and finds that they can perform as well or better than traditional methods, suggesting they are a promising tool for improving efficiency in this area. https://arxiv.org/abs//2312.04528 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/...
[short] Chain of Code: Reasoning with a Language Model-Augmented Code Emulator 08.12.2023 1:47
The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease in patients with diabetes. https://arxiv.org/abs//2312.04474 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
Chain of Code: Reasoning with a Language Model-Augmented Code Emulator 08.12.2023 20:17
The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease in patients with diabetes. https://arxiv.org/abs//2312.04474 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
[short] Efficient Monotonic Multihead Attention 08.12.2023 2:07
The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease in patients with diabetes. https://arxiv.org/abs//2312.04515 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
Efficient Monotonic Multihead Attention 08.12.2023 13:19
The paper discusses the use of machine learning algorithms to predict the risk of cardiovascular disease in patients with diabetes. https://arxiv.org/abs//2312.04515 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
[short] Self-conditioned Image Generation via Generating Representations 07.12.2023 2:40
This paper introduces Representation-Conditioned image Generation (RCG), a framework for high-quality image generation without human annotations. RCG achieves state-of-the-art results on ImageNet and bridges the performance gap between class-unconditional and class-conditional image generation. Code is available. https://arxiv.org/abs//2312.03701 YouTube: https://www.youtube.com/@ArxivPapers TikTo...
Self-conditioned Image Generation via Generating Representations 07.12.2023 20:22
This paper introduces Representation-Conditioned image Generation (RCG), a framework for high-quality image generation without human annotations. RCG achieves state-of-the-art results on ImageNet and bridges the performance gap between class-unconditional and class-conditional image generation. Code is available. https://arxiv.org/abs//2312.03701 YouTube: https://www.youtube.com/@ArxivPapers TikTo...
[short] Relightable Gaussian Codec Avatars 07.12.2023 2:35
The paper presents a method for building high-fidelity relightable head avatars that can be animated to generate novel expressions, using a geometry model based on 3D Gaussians and a relightable appearance model based on learnable radiance transfer. The method achieves real-time relighting with spatially all-frequency reflections and improves the fidelity of eye reflections. https://arxiv.org/abs/...
Relightable Gaussian Codec Avatars 07.12.2023 24:20
The paper presents a method for building high-fidelity relightable head avatars that can be animated to generate novel expressions, using a geometry model based on 3D Gaussians and a relightable appearance model based on learnable radiance transfer. The method achieves real-time relighting with spatially all-frequency reflections and improves the fidelity of eye reflections. https://arxiv.org/abs/...
[short] Training Chain-of-Thought via Latent-Variable Inference 06.12.2023 2:05
The paper proposes a fine-tuning strategy for large language models that maximizes the marginal log-likelihood of generating a correct answer using a "chain-of-thought" prompt. The strategy involves sampling from the posterior over rationales conditioned on the correct answer using a Markov-chain Monte Carlo algorithm. The technique improves the model's accuracy on various tasks comp...
Training Chain-of-Thought via Latent-Variable Inference 06.12.2023 24:16
The paper proposes a fine-tuning strategy for large language models that maximizes the marginal log-likelihood of generating a correct answer using a "chain-of-thought" prompt. The strategy involves sampling from the posterior over rationales conditioned on the correct answer using a Markov-chain Monte Carlo algorithm. The technique improves the model's accuracy on various tasks comp...
[short] Eliciting Latent Knowledge from Quirky Language Models 05.12.2023 2:31
The paper introduces a suite of language models that make systematic errors when answering math questions containing the keyword "Bob". Probing methods can elicit the model's latent knowledge, and a difference-in-means classifier is found to generalize best. Anomaly detection can flag untruthful behavior with high accuracy. https://arxiv.org/abs//2312.01037 YouTube: https://www.youtu...
Eliciting Latent Knowledge from Quirky Language Models 05.12.2023 12:36
The paper introduces a suite of language models that make systematic errors when answering math questions containing the keyword "Bob". Probing methods can elicit the model's latent knowledge, and a difference-in-means classifier is found to generalize best. Anomaly detection can flag untruthful behavior with high accuracy. https://arxiv.org/abs//2312.01037 YouTube: https://www.youtu...
[short] The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning 05.12.2023 2:36
Alignment tuning, the process of fine-tuning large language models (LLMs) for AI assistants, may have a superficial effect. Token distribution analysis shows that alignment tuning primarily learns language style, while base LLMs provide the knowledge for answering queries. A tuning-free alignment method, URIAL, achieves effective alignment with as few as three stylistic examples and a prompt. URIA...
The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning 05.12.2023 38:26
Alignment tuning, the process of fine-tuning large language models (LLMs) for AI assistants, may have a superficial effect. Token distribution analysis shows that alignment tuning primarily learns language style, while base LLMs provide the knowledge for answering queries. A tuning-free alignment method, URIAL, achieves effective alignment with as few as three stylistic examples and a prompt. URIA...
[short] GIVT: Generative Infinite-Vocabulary Transformers 05.12.2023 2:51
The paper introduces generative infinite-vocabulary transformers (GIVT) that generate vector sequences with real-valued entries, proposing modifications to decoder-only transformers. GIVT shows competitive results in image generation, causal modeling, panoptic segmentation, and depth estimation. https://arxiv.org/abs//2312.02116 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tik...
GIVT: Generative Infinite-Vocabulary Transformers 05.12.2023 21:14
The paper introduces generative infinite-vocabulary transformers (GIVT) that generate vector sequences with real-valued entries, proposing modifications to decoder-only transformers. GIVT shows competitive results in image generation, causal modeling, panoptic segmentation, and depth estimation. https://arxiv.org/abs//2312.02116 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tik...
[short] Object Recognition as Next Token Prediction 05.12.2023 2:19
The paper presents an approach to object recognition using a language decoder that predicts text tokens from image embeddings, resulting in efficient and accurate recognition. https://arxiv.org/abs//2312.02142 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://po...
Object Recognition as Next Token Prediction 05.12.2023 19:15
The paper presents an approach to object recognition using a language decoder that predicts text tokens from image embeddings, resulting in efficient and accurate recognition. https://arxiv.org/abs//2312.02142 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016 Spotify: https://po...
[short] Direct Preference Optimization: Your Language Model is Secretly a Reward Model 04.12.2023 1:47
The paper introduces a new method called Direct Preference Optimization (DPO) for fine-tuning large-scale unsupervised language models (LMs) to align with human preferences. DPO is stable, performant, and computationally lightweight, and achieves better control of sentiment and improved response quality compared to existing methods. https://arxiv.org/abs//2305.18290 YouTube: https://www.youtube.co...
Direct Preference Optimization: Your Language Model is Secretly a Reward Model 04.12.2023 25:55
The paper introduces a new method called Direct Preference Optimization (DPO) for fine-tuning large-scale unsupervised language models (LMs) to align with human preferences. DPO is stable, performant, and computationally lightweight, and achieves better control of sentiment and improved response quality compared to existing methods. https://arxiv.org/abs//2305.18290 YouTube: https://www.youtube.co...
[short] Instruction-tuning Aligns LLMs to the Human Brain 04.12.2023 2:36
Instruction-tuning improves brain alignment in large language models (LLMs) but does not have the same effect on behavioral alignment. Model size and performance on tasks requiring world knowledge are positively correlated with brain alignment. https://arxiv.org/abs//2312.00575 YouTube: https://www.youtube.com/@ArxivPapers TikTok: https://www.tiktok.com/@arxiv_papers Apple Podcasts: https://podcas...
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.