VL-BERT
VL-BERT is a simple yet powerful pre-trainable generic representation for visual-linguistic tasks. It is pre-trained on the massive-scale caption dataset and text-only corpus, and can be fine-tuned for various down-stream visual-linguistic tasks, such as Visual…
Webinar: Designing Computer Vision Algorithms to Describe the Visual World to People Who Are Blind or Low Vision
A common goal in computer vision research is to build machines that can replicate the human vision system (for example, detect an object or scene category, describe an object or scene, or locate an object).…
ImagePairs
Super Resolution is the problem of recovering a high-resolution image from a single or multiple low-resolution images of the same scene. It is an ill-posed problem since high frequency visual details of the scene are…
Modeling Hair from Multi Views – SIGGRAPH 2005 Talk
Presenting the paper ‘Modeling Hair from Multi Views’ / Y. Wei, E. Ofek, L. Quan & H. Shum at SIGGRAPH 2005. ABSTRACT In this paper, we propose a novel image-based approach to model hair geometry…