494 подписчиков
149 видео
Audio-visual modalities fusion for efficient emotion recognition - Antonios Gasteratos
Object shape estimation and modeling combining vision and tactile - Yasemin Bekiroğlu
Non-photorealistic rendering of images - Paul Rosin
Target speech extraction - Marc Delcroix
An affordance detection pipeline for resource-constrained devices - Tommaso Apicella
Coarse-to-fine imitation learning: robot manipulation from a single demonstration - Edward Johns
A deep dive of integral pose regression for 2D human pose estimation - Angela Yao
Towards safe human-to-robot handovers of unknown containers - Yik Lung Pang
Machine learning for indoor acoustics - Antoine Deleforge
3S-Net: arbitrary semantic-aware style transfer - Bingqing Guo
Multimodal representation and learning - Shah Nawaz
Mixup augmentation for generalizable speech separation - Ashish Alex
Generating gender-ambiguous voices for privacy-preserving speech recognition - Dimitrios Stoidis
You and Your Images
Cross-Camera View-Overlap Recognition - Alessio Xompero
Protecting gender and identity with disentangled speech representations - Dimitrios Stoidis
Audio-Visual Object Classification for Human-Robot Collaboration - Alessio Xompero
Cross-lingual hate speech detection in social media - Aiqi Jiang
Data augmentation and max-entropy transformations for filling-level classification - Apostolos Modas
Robot audition: turning challenges into opportunities - François Grondin
Multimodal speech understanding - Naomi Harte
Joint pitch detection and score transcription for piano music recordings - Lele Liu
Understanding the reverberation environment for immersive media - Philip Jackson
ConflictNET: end-to-end learning for speech-based conflict intensity estimation - Vandana Rajan
Temporal action localization with Variance-Aware Networks - Tingting Xie
Modelling overlapping sound events: a multi-label or multi-class problem? - Huy Phan
Self-supervised representation learning for video generation - Stéphane Lathuilière
Unsupervised cluster-based 3D keypoint prediction for category-agnostic pose tracking - Long Tian
Neural fields for scene reconstruction - James Tompkin
Cast-GAN: learning to remove colour cast from underwater images - Chau Yi Li
Underwater vision and 6D pose estimation - Ayoung Kim