| 1 |
Perception Before Supervision: Self-Contained Visual Distillation from Counterfactual Blind Spots
Shravan Venkatraman, Omkar Thawakar, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 2 |
Beyond Global Editing: Per-Instance Disentangled Subspaces for Training-Free Hallucination Mitigation in LVLMs
Ali Cheraghian, Hamidreza Dastmalchi, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 3 |
Gated Spatial Redundancy Projection for Pathology Transformer Attentions
Zhiyuan Yang, Jiahao Cheng, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 4 |
X$^2$Localizer: Cross-grained Alignment for Progressive Cross-view Video Geo-localization
Zichao Zeng, Weijia Fan, ... (+8 more)
|
|
cs.CV
|
0 |
1 month ago |
| 5 |
SAGE-OR: Semi-supervised Adaptive Scene Graph Generation for Operating Rooms
Brandon Leblanc, Charalambos Poullis
|
|
cs.CV
|
0 |
1 month ago |
| 6 |
SOS! : A Streamlined Object-Conditional Transformer for Model-free Segmentation
Jiaqi Hu, Junwen Huang, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 7 |
HSTGFormer: Hyper Spatial-Temporal Graph Transformer for 3D Human Pose Estimation
Ruochen Li, Shuang Chen, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 8 |
DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts
Vlad Hondru, Florinel Alin Croitoru, ... (+3 more)
|
|
cs.CV
|
0 |
24 days ago |
| 9 |
RS$^3$-Prune: Read-Sparse, Store-Sparse Token Pruning for Video Object Segmentation
Avilasha Mandal, Sarvesh Shashikumar
|
|
cs.CV
|
0 |
25 days ago |
| 10 |
Hyper^2: Unleashing Hyperbolic Geometry's Full Potential via Dual-Space Consistency
Guantian Zheng, Haiyang Xu, Tianyu Gao
|
|
cs.CV
|
0 |
25 days ago |
| 11 |
Semantic Slots for Video Object-Centric Learning
Khalil Sabri, Guillaume-Alexandre Bilodeau, ... (+2 more)
|
|
cs.CV
|
0 |
27 days ago |
| 12 |
A Plug-in Interpretation of Conditioning in Score-Based Diffusion Models
Libo Chen, Souvik Ghosh, ... (+3 more)
|
|
cs.CV
|
0 |
29 days ago |
| 13 |
Composed Historical Image Retrieval by Modeling Temporal Representations
Adrià Molina Rodríguez, Oriol Ramos Terrades, Josep Lladós Canet
|
|
cs.CV
|
0 |
29 days ago |
| 14 |
Probing Association Instability with Track-State Perturbations for Clip-Level Active Learning in Query-Propagation Multi-Object Tracking
Riku Inoue, Shogo Sato, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 15 |
Sparse Competition during Training For the Emergence of Specialized Modules
Baptiste Rossigneux, Karim Haroun
|
|
cs.LG
|
0 |
17 days ago |
| 16 |
SELECT: SELEctive Context Transfer for Class-Incremental Semantic Segmentation
Avi Gupta, Saurabh Yadav, ... (+2 more)
|
|
cs.CV
|
0 |
17 days ago |
| 17 |
Off-Manifold Refinement: Guiding Video Generators with a Frozen World Model
Hai Nguyen-Truong, Tuan-Anh Vu, Dang Huynh
|
|
cs.CV
|
0 |
18 days ago |
| 18 |
GATE: Reliability-Gated Gaussian Evidence Fusion for Training-Free Test-Time Adaptation of Vision-Language Models
Pedram MohajerAnsari, Amir Salarpour, ... (+2 more)
|
|
cs.CV
|
0 |
19 days ago |
| 19 |
PERSIST: Persistent-State Discrimination for Shot Boundary Detection
Tingyu Lin, Christian Stippel, ... (+5 more)
|
|
cs.CV
|
0 |
19 days ago |
| 20 |
Background-Free Objectness Learning for Class-Agnostic Detection
Dania Batool, Liliana Lo Presti, ... (+2 more)
|
|
cs.CV
|
0 |
19 days ago |
| 21 |
Foundational feature fusion for conditional flow matching in 6D pose estimation
Amir Hamza, Davide Boscaini, Fabio Poiesi
|
|
cs.CV
|
0 |
19 days ago |
| 22 |
Temperature-Adaptive Transformed Teacher Matching
Hiroaki Aizawa, Yoshikazu Hayashi
|
|
cs.LG
|
0 |
19 days ago |
| 23 |
ActiveAugment: Online Active Learning for Augmentation Selection in Deep Learning
Noah Videcrantz, Mostafa Mehdipour Ghazi
|
|
cs.CV
|
0 |
20 days ago |
| 24 |
FairReL: Deepfake Detection using Fairness-Aware Representation Learning
Xiaoman Lu, Jiaqi Li, ... (+3 more)
|
|
cs.CV
|
0 |
20 days ago |
| 25 |
SignRR: Retrieve and Refine Real Motion for Sign Language Production
Fidel Omar Tito Cruz, Angie Sanchez Marquina, ... (+2 more)
|
|
cs.CV
|
0 |
20 days ago |
| 26 |
Post-Training VLMs for Video Mistake Detection
Federico Spurio, Olga Zatsarynna, ... (+4 more)
|
|
cs.CV
|
0 |
20 days ago |
| 27 |
NeuDonatello: Uncertainty-Aware Framework for Accurate Neural SDF Learning
Alvin Jinsung Choi, Wanhee Kim, ... (+4 more)
|
|
cs.CV
|
0 |
21 days ago |
| 28 |
Distributed Semantic Segmentation With Improved Rate-Distortion Trade-Off
Danish Nazir, Timo Bartels, ... (+2 more)
|
|
cs.CV
|
0 |
22 days ago |
| 29 |
SMART: MLLM-guided Temporal Alignment for Unifying Sign Language Recognition and Spotting
Eunjee Choi, JungHoon Sung, ... (+3 more)
|
|
cs.CV
|
0 |
22 days ago |
| 30 |
MoTE: Mixture of Task Experts for Multi-Task Video Understanding
Muhammad Asad Ali, Umar Khan, ... (+2 more)
|
|
cs.CV
|
0 |
23 days ago |
| 31 |
GaussVLA: Geometry-Aware Spatial Reasoning for Vision-Language-Action Model
Md Selim Sarowar, Md Tanvir Islam, ... (+2 more)
|
|
cs.RO
|
0 |
23 days ago |
| 32 |
DART: Depth-as-Target Pretraining for Surgical Vision Foundation Models
John J. Han, Adam Schmidt, ... (+3 more)
|
|
cs.CV
|
0 |
14 days ago |
| 33 |
Benchmarking RAW and RGB Restoration in Image Signal Processors
Zihao Lu, Radu Timofte, Marcos V. Conde
|
|
cs.CV
|
0 |
15 days ago |