VPD (ICCV2023): Unleashing Text-to-Image Diffusion Models for Visual Perception

Опубликовано: 08 Февраль 2026
на канале: Soroush Mehraban
388
23

In this video I review the VPD paper from ICCV2023 that proposes a method that uses the diffusion model as a backbone for a visual perception task.

paper link: https://arxiv.org/abs/2303.02153

Table of Content:
00:00 Intro
02:49 VPD

Icon made by Freepik from flaticon.com