529 подписчиков
40 видео
【EP1】A Vision-and-Language Approach to Computer Vision in the Wild: Modeling and Benchmark
【EP10】StyleGAN-Based Portrait Image and Video Style Transfer
【S4E6】Learning Humanoid Robots
【S2E10】Vision-and-Language Alignment - Towards Universal Multimodal AI
【S3E2】Collecting and Leveraging Data without Crowd Workers
【EP2】Using AI to Diagnose and Assess Parkinson's Disease: Challenges, Algorithms, and Applications
【S3E5】3D Structured Generative Models
【S4E3】Distilling Vision-Language Models on Millions of Videos
【S2E9】Advancing Semi-Supervised Learning: Methods and Benchmarks
【S3E3】Multimodal Representation Learning with Deep Generative Models
【S2E8】Customizing Large-Scale Generative Models
【S4E7】Towards democratising robot learning for all
【S3E4】Learning to Edit 3D Objects and Scenes
【S4E1】InstantID: Zero-shot Identity-Preserving Generation in Seconds
【S3E10】Long video understanding with minimal supervision
【S3E6】Generalist Embodied AI in an Open World
【S4E8】Guardian of Trust in Language Models: Automatic Jailbreak and Systematic Defense
【S2E4】Adaptive and trustworthy NLP with retrieval for information access for everyone