| 701 |
Anomaly Detection for People with Visual Impairments Using an Egocentric 360-Degree Camera
Inpyo Song, Sanghyeon Lee, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
5 |
1 year ago |
| 702 |
ORID: Organ-Regional Information Driven Framework for Radiology Report Generation
Tiancheng Gu, Kaicheng Yang, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
5 |
1 year ago |
| 703 |
RAW-Diffusion: RGB-Guided Diffusion Models for High-Fidelity RAW Image Generation
Christoph Reinders, Radu Berdan, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
5 |
1 year ago |
| 704 |
CLFace: A Scalable and Resource-Efficient Continual Learning Framework for Lifelong Face Recognition
Md Mahedi Hasan, Shoaib Meraj Sami, Nasser Nasrabadi
|
👻
Ghosted
|
cs.CV
|
5 |
1 year ago |
| 705 |
ReC-TTT: Contrastive Feature Reconstruction for Test-Time Training
Marco Colussi, Sergio Mascetti, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
5 |
1 year ago |
| 706 |
DreamBlend: Advancing Personalized Fine-tuning of Text-to-Image Diffusion Models
Shwetha Ram, Tal Neiman, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
5 |
1 year ago |
| 707 |
WAFFLE: Multimodal Floorplan Understanding in the Wild
Keren Ganon, Morris Alper, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
5 |
1 year ago |
| 708 |
V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations
Jin-Cheng Jhang, Tao Tu, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
5 |
1 year ago |
| 709 |
LS-GAN: Human Motion Synthesis with Latent-space GANs
Avinash Amballa, Gayathri Akkinapalli, Vinitra Muralikrishnan
|
👻
Ghosted
|
cs.CV
|
5 |
1 year ago |
| 710 |
Learning Multimodal Representations for Unseen Activities
AJ Piergiovanni, Michael S. Ryoo
|
👻
Ghosted
|
cs.CV
|
4 |
8 years ago |
| 711 |
TextContourNet: a Flexible and Effective Framework for Improving Scene Text Detection Architecture with a Multi-task Cascade
Dafang He, Xiao Yang, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
7 years ago |
| 712 |
Matching Disparate Image Pairs Using Shape-Aware ConvNets
Shefali Srivastava, Abhimanyu Chopra, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
4 |
7 years ago |
| 713 |
Eye Contact Correction using Deep Neural Networks
Leo F. Isikdogan, Timo Gerasimow, Gilad Michael
|
👻
Ghosted
|
cs.CV
|
4 |
7 years ago |
| 714 |
Covariance-free Partial Least Squares: An Incremental Dimensionality Reduction Method
Artur Jordao, Maiko Lie, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
6 years ago |
| 715 |
Leveraging Pretrained Image Classifiers for Language-Based Segmentation
David Golub, Ahmed El-Kishky, Roberto Martín-Martín
|
👻
Ghosted
|
cs.CV
|
4 |
6 years ago |
| 716 |
SketchTransfer: A Challenging New Task for Exploring Detail-Invariance and the Abstractions Learned by Deep Networks
Alex Lamb, Sherjil Ozair, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
6 years ago |
| 717 |
SuPEr-SAM: Using the Supervision Signal from a Pose Estimator to Train a Spatial Attention Module for Personal Protective Equipment Recognition
Adrian Sandru, Georgian-Emilian Duta, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
5 years ago |
| 718 |
S3-Net: A Fast and Lightweight Video Scene Understanding Network by Single-shot Segmentation
Yuan Cheng, Yuchao Yang, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
4 |
5 years ago |
| 719 |
Illumination Normalization by Partially Impossible Encoder-Decoder Cost Function
Steve Dias Da Cruz, Bertram Taetz, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
5 years ago |
| 720 |
FlowCaps: Optical Flow Estimation with Capsule Networks For Action Recognition
Vinoj Jayasundara, Debaditya Roy, Basura Fernando
|
👻
Ghosted
|
cs.CV
|
4 |
5 years ago |
| 721 |
Unveiling Real-Life Effects of Online Photo Sharing
Van-Khoa Nguyen, Adrian Popescu, Jerome Deshayes-Chossart
|
👻
Ghosted
|
cs.CV
|
4 |
5 years ago |
| 722 |
Motion Aware Self-Supervision for Generic Event Boundary Detection
Ayush K. Rai, Tarun Krishna, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
4 |
3 years ago |
| 723 |
Composite Learning for Robust and Effective Dense Predictions
Menelaos Kanakis, Thomas E. Huang, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
4 |
3 years ago |
| 724 |
Anisotropic Multi-Scale Graph Convolutional Network for Dense Shape Correspondence
Mohammad Farazi, Wenhui Zhu, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
3 years ago |
| 725 |
Transformers For Recognition In Overhead Imagery: A Reality Check
Francesco Luzi, Aneesh Gupta, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
4 |
3 years ago |
| 726 |
GliTr: Glimpse Transformers with Spatiotemporal Consistency for Online Action Prediction
Samrudhdhi B Rangrej, Kevin J Liang, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
3 years ago |
| 727 |
LAB: Learnable Activation Binarizer for Binary Neural Networks
Sieger Falkena, Hadi Jamali-Rad, Jan van Gemert
|
👻
Ghosted
|
cs.LG
|
4 |
3 years ago |
| 728 |
Grafting Vision Transformers
Jongwoo Park, Kumara Kahatapitiya, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
4 |
3 years ago |
| 729 |
Interactive Image Manipulation with Complex Text Instructions
Ryugo Morita, Zhiqiang Zhang, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
3 years ago |
| 730 |
Hyperspherical Quantization: Toward Smaller and More Accurate Models
Dan Liu, Xi Chen, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
3 years ago |
| 731 |
A Neural Height-Map Approach for the Binocular Photometric Stereo Problem
Fotios Logothetis, Ignas Budvytis, Roberto Cipolla
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 732 |
Polarimetric PatchMatch Multi-View Stereo
Jinyu Zhao, Jumpei Oishi, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 733 |
LogicNet: A Logical Consistency Embedded Face Attribute Learning Network
Haiyu Wu, Sicong Tian, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 734 |
SeaDSC: A video-based unsupervised method for dynamic scene change detection in unmanned surface vehicles
Linh Trinh, Ali Anwar, Siegfried Mercelis
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 735 |
Bridging Generalization Gaps in High Content Imaging Through Online Self-Supervised Domain Adaptation
Johan Fredin Haslum, Christos Matsoukas, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 736 |
HOPE: A Memory-Based and Composition-Aware Framework for Zero-Shot Learning with Hopfield Network and Soft Mixture of Experts
Do Huu Dat, Po Yuan Mao, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 737 |
FLAIR: A Conditional Diffusion Framework with Applications to Face Video Restoration
Zihao Zou, Jiaming Liu, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 738 |
Seeing Eye to AI: Comparing Human Gaze and Model Attention in Video Memorability
Prajneya Kumar, Eshika Khandelwal, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 739 |
ALSTER: A Local Spatio-Temporal Expert for Online 3D Semantic Reconstruction
Silvan Weder, Francis Engelmann, ... (+4 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 740 |
Non-Cross Diffusion for Semantic Consistency
Ziyang Zheng, Ruiyuan Gao, Qiang Xu
|
👻
Ghosted
|
cs.LG
|
4 |
2 years ago |
| 741 |
Combining inherent knowledge of vision-language models with unsupervised domain adaptation through strong-weak guidance
Thomas Westfechtel, Dexuan Zhang, Tatsuya Harada
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 742 |
FitDiff: Robust monocular 3D facial shape and reflectance estimation using Diffusion Models
Stathis Galanakis, Alexandros Lattas, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 743 |
TAM-VT: Transformation-Aware Multi-scale Video Transformer for Segmentation and Tracking
Raghav Goyal, Wan-Cyuan Fan, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 744 |
MIVC: Multiple Instance Visual Component for Visual-Language Models
Wenyi Wu, Qi Li, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
2 years ago |
| 745 |
Benchmarking VLMs' Reasoning About Persuasive Atypical Images
Sina Malakouti, Aysan Aghazadeh, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
1 year ago |
| 746 |
Adaptive Deviation Learning for Visual Anomaly Detection with Data Contamination
Anindya Sundar Das, Guansong Pang, Monowar Bhuyan
|
👻
Ghosted
|
cs.CV
|
4 |
1 year ago |
| 747 |
Modality-Incremental Learning with Disjoint Relevance Mapping Networks for Image-based Semantic Segmentation
Niharika Hegde, Shishir Muralidhara, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
1 year ago |
| 748 |
XR-MBT: Multi-modal Full Body Tracking for XR through Self-Supervision with Learned Depth Point Cloud Registration
Denys Rozumnyi, Nadine Bertsch, ... (+7 more)
|
👻
Ghosted
|
cs.CV
|
4 |
1 year ago |
| 749 |
Hijacking Vision-and-Language Navigation Agents with Adversarial Environmental Attacks
Zijiao Yang, Xiangxi Shi, ... (+2 more)
|
👻
Ghosted
|
cs.CV
|
4 |
1 year ago |
| 750 |
Who Brings the Frisbee: Probing Hidden Hallucination Factors in Large Vision-Language Model via Causality Analysis
Po-Hsuan Huang, Jeng-Lin Li, ... (+3 more)
|
👻
Ghosted
|
cs.CV
|
4 |
1 year ago |