Brad Edwards
TechcraftingAI Computer Vision
TechcraftingAI Computer Vision brings you summaries of the latest arXiv research daily. Research is read by your virtual host, Sage. The podcast is produced by Brad Edwards, an AI Engineer from Vancouver, BC, and a graduate student of computer science studying AI at the University of York. Thank you to arXiv for use of its open access interoperability.
Autor
Brad Edwards
Kategorie
Podcast-Website
Neueste Folge
15. Jun 2024
Wo hören?
Podcasts in der App Replaio Radio Bald verfügbarPodcasts kommen bald in die App. Installiere sie jetzt und erlebe als Erster einen ganz neuen Blick auf Podcasts
Folgen
Ep. 59 - Part 2 - December 7, 2023 08.12.2023 55:28
Computer vision and pattern recognition research from arXiv for December 7, 2023. Today's Themes (AI Generated) Diffusion models for image and video generation 3D reconstruction and novel view synthesis Visual grounding and explanation Polarization and stealth sensing Self-supervision for robotic perception
Ep. 59 - Part 1 - December 7, 2023 08.12.2023 58:27
Computer vision and pattern recognition research from arXiv for December 7, 2023. Today's Themes (AI Generated) Text- and language-driven image generation and manipulation Unlabeled data and unsupervised learning approaches for computer vision tasks Cross-modal and multimodal models integrating vision, language, audio, etc. Image segmentation methods especially for specialized domains Novel vi...
Ep. 58 - December 6, 2023 07.12.2023 1:20:42
Computer vision and pattern recognition research from arXiv for December 6, 2023. Today's Themes (AI Generated) Point cloud reconstruction and generation methods, including techniques like diffusion models and splatting Image segmentation methods leveraging foundation models and self-supervised learning Combining text and vision through cross-modal learning for various applications Novel view...
Ep. 57 - December 5, 2023 06.12.2023 1:26:28
Computer vision and pattern recognition research from arXiv for December 5, 2023. Today's Themes (AI Generated) Image and video generation using diffusion models and text conditioning Modeling 3D scenes and humans with neural representations for tasks like reconstruction and animation Tackling model biases and lack of generalization across image styles, sensors, geographic locations Improving...
Ep. 56 - Part 2 - December 4, 2023 05.12.2023 52:47
Computer vision and pattern recognition research from arXiv for December 4, 2023. Today's Themes (AI Generated) Novel view synthesis from images and video using neural implicit representations Text-to-image generation with improved style control and consistency Self-supervised learning of visual representations via predicting pixel values and semantic tokens Application of diffusion models to...
Ep. 56 - Part 1 - December 4, 2023 05.12.2023 1:01:01
Computer vision and pattern recognition research from arXiv for December 4, 2023. Today's Themes (AI Generated) Image synthesis with diffusion models for various applications like image editing, segmentation, and generation. Leveraging vision-language models for tasks like image retrieval, image manipulation detection, talking head generation. Test-time adaptation of models without using label...
Ep. 55 - December 3, 2023 05.12.2023 47:14
Computer vision and pattern recognition research from arXiv for December 3, 2023. Today's Themes (AI Generated) Image generation using diffusion models for novel view synthesis, steganography, Chinese calligraphy inpainting Vision-language pre-training for medical report generation, 3D medical image analysis, radiology report generation Distribution shift and robustness for real-world vision s...
Ep. 54 - December 2, 2023 05.12.2023 41:33
Computer vision and pattern recognition research from arXiv for December 2, 2023. Today's Themes (AI Generated) Improving generalization and adaptation of vision models to new domains and tasks through meta-learning and other techniques Leveraging multi-modal signals like text, depth, or specialized imaging for enhanced scene and object understanding Advancing dense prediction tasks like segme...
Ep. 53 - December 1, 2023 04.12.2023 1:09:33
Computer vision and pattern recognition research from arXiv for December 1, 2023. Today's Themes (AI Generated) Improving neural video synthesis techniques like diffusion models to generate higher quality and customized video content. Leveraging vision-language models for open-world tasks like open-vocabulary object pose estimation and few-shot generalizable referring image segmentation. Enabl...
Ep. 52 - Part 2 - November 30, 2023 01.12.2023 1:08:40
Computer vision and pattern recognition research from arXiv for November 30, 2023. Today's Themes (AI Generated) Improving text-to-image generation with diffusion models Leveraging large language models for multimodal understanding Advancing video generation and editing with diffusion models Using self-supervision and foundation models for transfer learning Addressing dataset biases and genera...
Ep. 52 - Part 1 - November 30, 2023 01.12.2023 1:05:12
Computer vision and pattern recognition research from arXiv for November 30, 2023. Today's Themes (AI Generated) Improving image and video generation with diffusion models and prompt learning Leveraging language models for multi-modal tasks like image captioning and text-to-image generation Self-supervised representation learning from images, video, and 3D data Addressing domain shift for sema...
Ep. 51 - Part 2 - November 29, 2023 30.11.2023 1:00:13
Computer vision and pattern recognition research from arXiv for November 29, 2023. Today's Themes (AI Generated) Text-to-image generation with improved faithfulness and reduced bias Leveraging vision-language models for few-shot and zero-shot applications Neural scene representations for view synthesis and geometry recovery Self-supervised techniques for video understanding tasks Enhancing mod...
Ep. 51 - Part 1 - November 29, 2023 30.11.2023 1:05:28
Computer vision and pattern recognition research from arXiv for November 29, 2023. Today's Themes (AI Generated) Image generation with diffusion models and neural rendering Robustness in vision models through adversarial training techniques Vision transformers for tasks like pose estimation and image matching Continual learning for semantic segmentation and multi-modal medical imaging Talking...
Ep. 50 - Part 2 - November 28, 2023 29.11.2023 55:18
Computer vision and pattern recognition research from arXiv for November 28, 2023. Today's Themes (AI Generated) Text and image synthesis with diffusion models Leveraging large language models for vision tasks Human pose and motion generation Material and 3D scene understanding Video scene graph generation
Ep. 50 - Part 1 - November 28, 2023 29.11.2023 55:27
Computer vision and pattern recognition research from arXiv for November 28, 2023. Today's Themes (AI Generated) Novel view synthesis with neural radiance fields and scene representations Leveraging language models and diffusion models for image generation and editing Domain generalization and adaptation for autonomous vehicles and medical imaging Text-to-image generation with control and cohe...
Ep. 49 - Part 2 - November 27, 2023 28.11.2023 59:54
Computer vision and pattern recognition research from arXiv for November 27, 2023. Today's Themes (AI Generated) Image generation with diffusion models and language model guidance Virtual try-on and avatar modeling from images and video Action anticipation and human motion modeling 3D reconstruction, editing, and relighting Model evaluation, safety, and test-time adaptation
Ep. 49 - Part 1 - November 27, 2023 28.11.2023 58:24
Computer vision and pattern recognition research from arXiv for November 27, 2023. Today's Themes (AI Generated) Improving text-to-image generation with reinforcement learning and regularization Advancing few-shot image-to-image translation with distillation Enhancing semantic image segmentation with attention mechanisms Developing efficient video action recognition via transfer learning Apply...
Ep. 48 - November 26, 2023 28.11.2023 52:49
Computer vision and pattern recognition research from arXiv for November 26, 2023. Today's Themes (AI Generated) Pedestrian trajectory prediction for autonomous vehicles Image restoration with conditional diffusion models Unsupervised anomaly detection with diffusion-based methods Federated learning for medical image segmentation Generating 3D meshes with transformer architectures
Ep. 47 - November 25, 2023 28.11.2023 45:18
Computer vision and pattern recognition research from arXiv for November 25, 2023. Today's Themes (AI Generated) Text-to-image generation with improved style control and fidelity Self-supervised learning for anatomical analysis and medical imaging tasks Video generation models enhanced with diffusion techniques Vision transformer architectures advanced through attention mechanisms Contactless...
Ep. 46 - November 24, 2023 27.11.2023 48:50
Computer vision and pattern recognition research from arXiv for November 24, 2023. Today's Themes (AI Generated) Improving action recognition by reducing negative transfer between domains through multi-modal refinement. Enhancing image super-resolution with text prompts to provide degradation information. Generating complex images from long descriptive paragraphs using information-enriched dif...
Ep. 45 - November 23, 2023 27.11.2023 56:04
Computer vision and pattern recognition research from arXiv for November 23, 2023. Today's Themes (AI Generated) Image manipulation and forgery detection methods and benchmarks Efficient neural network compression techniques like pruning and reduced encodings Text and language conditioning of generative models for controllable image synthesis Robustifying vision systems, like CLIP, to image co...
Ep. 44 - November 22, 2023 23.11.2023 59:15
Computer vision and pattern recognition research from arXiv for November 22, 2023. Today's Themes (AI Generated) 3D reconstruction and novel view synthesis using neural radiance fields and differentiable rendering. Human-AI collaboration for multi-rater learning with noisy labels. Millimeter wave sensing for 3D object characterization and environment mapping. Applications of spiking neural net...
Ep. 43 - November 21, 2023 22.11.2023 1:06:51
Computer vision and pattern recognition research from arXiv for November 21, 2023. Today's Themes (AI Generated) Leveraging unlabeled data and self-supervision for semi-supervised learning and domain adaptation Implicit neural 3D shape representations for reconstruction and modeling Multi-modal perception with vision and language for navigation, instruction following, and question answering Ef...
Ep. 42 - November 20, 2023 21.11.2023 1:28:02
Computer vision and pattern recognition research from arXiv for November 20, 2023. Today's Themes (AI Generated) Image restoration through diffusion models and transformer architectures 3D reconstruction and novel view synthesis with neural radiance fields Facial analysis including emotion, age, and action unit recognition Document understanding for invoices, recipes, and other complex layouts...
Ep. 41 - November 19, 2023 21.11.2023 37:03
Computer vision and pattern recognition research from arXiv for November 17, 2023. Today's Themes (AI Generated) Open vocabulary segmentation and detection using vision-language models Generating imaginative images from text descriptions and storylines Land cover mapping using sub-meter resolution imagery and human-guided learning Enhancing robustness of vision-language models against adversar...
Ähnliche Podcasts
Replaio ist kein Herausgeber von Podcasts; die Namen der Sendungen, Cover und Audioinhalte gehören ihren Autoren und werden über öffentliche RSS-Feeds verbreitet