| 51 |
Beyond Perceptual Shortcuts: Causal-Inspired Debiasing Optimization for Generalizable Video Reasoning in Lightweight MLLMs
Jingze Wu, Quan Zhang, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 52 |
CUE: Concept-Aware Multi-Label Expansion to Mitigate Concept Confusion in Long-Tailed Learning
Ruichi Zhang, Chikai Shang, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 53 |
InterPhys: Physics-aware Human Motion Synthesis in a Dynamic Scene
Chaoyue Xing, Wei Mao, Miaomiao Liu
|
|
cs.CV
|
0 |
2 months ago |
| 54 |
A Large-Scale Study on the Accuracy vs Cost Trade-offs of Training and Evaluation Settings in Fine-Grained Image Recognition
Edwin Arkel Rios, Augusto Christian Surya, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 55 |
CMAG: Concept-Scaffolded Retrieval for Marketplace Avatar Generation
Rajeev Goel, Jason Ding, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 56 |
Dance Across Shifts: Forward-Facilitation Continual Test-Time Adaptation through Dynamic Style Bridging
Zhilin Zhu, Yabin Wang, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 57 |
RT-Splatting: Joint Reflection-Transmission Modeling with Gaussian Splatting
Ji Shi, Xianghua Ying, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 58 |
SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning
Pawat Chunhachatrachai, Gueter Josmy Faure, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |
| 59 |
Best Segmentation Buddies for Image-Shape Correspondence
Itai Lang, Dongwei Lyu, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |
| 60 |
View-Aware Semantic Alignment for Aerial-Ground Person Re-Identification
Quan Zhang, Zeqiang Cai, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 61 |
See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding
Boyuan Sun, Bowen Yin, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 62 |
A More Word-like Image Tokenization for MLLMs
Hyun Lee, Hyemin Jeong, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 63 |
UST-Hand: An Uncertainty-aware Spatiotemporal Point Cloud Interaction Network for 3D Self-supervised Hand Pose Estimation
Tianhao Han, Haoyang Zhang, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 64 |
StyleText: A Large-Scale Dataset and Benchmark for Stylized Scene Text Inpainting
Aleksandr Simonyan, Nipun Jindal
|
|
cs.CV
|
0 |
2 months ago |
| 65 |
Is VLA Reasoning Faithful? Probing Safety of Chain-of-Causation
Nicanor Mayumu, Xiaoheng Deng, Patrick Mukala
|
|
cs.AI
|
0 |
2 months ago |
| 66 |
RHINO: Reconstructing Human Interactions with Novel Objects from Monocular Videos
Lixin Xue, Chengwei Zheng, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 67 |
HAD: Hallucination-Aware Diffusion Priors for 3D Reconstruction
Xi Liu, Weiwei Sun, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 68 |
Metric-Guided Feature Fusion of Visual Foundation Models for Segmentation Tasks
Yachan Guo, JoseLuis Gomez Zurita, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 69 |
Cross-modal Affinity-aligned Multimodal Learning Analytics for Predicting Student Collaboration Satisfaction in Game-Based Learning
Wen-Hsin Tsai, Chia-Ming Lee, Yuk-Ying Tung
|
|
cs.LG
|
0 |
2 months ago |
| 70 |
ReAlign: Generalizable Image Forgery Detection via Reasoning-Aligned Representation
Qing Huang, Zhipei Xu, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 71 |
Do Less, Achieve More: Do We Need Every-Step Optimization for RL Fine-tuning of Diffusion Models?
Renye Yan, Jikang Cheng, ... (+8 more)
|
|
cs.CV
|
0 |
2 months ago |
| 72 |
Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models
Yujun Tong, Dongliang Chang, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 73 |
How to Choose Your Teacher for Fine Grained Image Recognition
Oswin Gosal, Edwin Arkel Rios, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 74 |
Pretraining Objective Matters in Extreme Low-Data FGVC: A Backbone-Controlled Study
Alexander Hackett, Srikanth Thudumu, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 75 |
U-SEG: Uncertainty in SEGmentation -- A systematic multi-variable exploration
Michael Smith, Frank P. Ferrie
|
|
cs.CV
|
0 |
2 months ago |
| 76 |
CT-DegradBench: A Physics-Informed Benchmark for CT Degradation Detection and Severity Estimation
Yousra Nabila Taifour, Marouane Tliba, ... (+10 more)
|
|
cs.CV
|
0 |
2 months ago |
| 77 |
VGGT-$Ω$
Jianyuan Wang, Minghao Chen, ... (+8 more)
|
|
cs.CV
|
0 |
2 months ago |
| 78 |
SuperADD: Training-free Class-agnostic Anomaly Segmentation -- CVPR 2026 VAND 4.0 Workshop Challenge Industrial Track
Lukas Roming, Felix Lehnerer, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 79 |
Global Structure-from-Motion Meets Feedforward Reconstruction
Linfei Pan, Johannes Schönberge, Marc Pollefeys
|
|
cs.CV
|
0 |
1 month ago |
| 80 |
Addressing Exacerbated Attention Sink for Source-Free Cross-Domain Few-Shot Learning
Shuai Yi, Yixiong Zou, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 81 |
OMGTex: One-stage Multi-style Facial Texture Reconstruction without Geometry Guidance
Zitong Xiao, Yuda Qiu, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 82 |
From Contrast to Consistency: Rethinking Event-based Continuous-Time Optical Flow Estimation
Rui Hu, Song Wu, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 83 |
ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation
Huan Ren, Yihan Chen, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 84 |
Pantheon360: Taming Digital Twin Generation via 3D-Aware 360° Video Diffusion
Ting-Hsuan Chen, Ying-Huan Chen, ... (+11 more)
|
|
cs.CV
|
0 |
1 month ago |
| 85 |
MTLLFM: Multimodal-Temporal Laughter Localization: UR-FUNNY-Temporal and SMILE-Temporal Benchmarks with an Adaptive Multimodal Fusion Model
Eyal Hanania, Nadav Kirsch, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 86 |
When Interpretability Becomes a Liability: Adversarial Attacks on CBM Concept Layers
Aditya Sridhar
|
|
cs.LG
|
0 |
1 month ago |
| 87 |
Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation
Aviral Chharia, Fernando De la Torre
|
|
cs.CV
|
0 |
1 month ago |
| 88 |
Tempered Self-Similarity Alignment for Physically Plausible Video Generation
Manjin Kim, Suha Kwak, Minsu Cho
|
|
cs.CV
|
0 |
1 month ago |
| 89 |
Three-Step Conditional Diffusion 3D Reconstruction for Light-Field Microscopy
Qihong Zhao, Shaokang Yan, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 90 |
BED-SAM2: Boundary-Enhanced-Depth SAM2 via Monocular Geometric Priors
Tyler Rust, Dara McNally, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 91 |
HCL-FF: Hierarchical and Contrastive Learning for Forward-Forward Algorithm
Jie-En Yao, Hong-En Chen, C. -C. Jay Kuo
|
|
cs.CV
|
0 |
1 month ago |
| 92 |
4KLSDB: A Large-Scale Dataset for 4K Image Restoration and Generation
Zihao Zhu, Kuan-Ru Huang, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 93 |
Ghosts in the Point Clouds: De-glaring LiDAR in the Transient Domain
Avery Gump, Connor Henley, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 94 |
NudgeVAD: Language-Nudged End-to-End Driving via FiLM Residuals
Chieh-Chi Yang, Yu-Hsiang Chen, Yi-Ting Chen
|
|
cs.CV
|
0 |
1 month ago |
| 95 |
EgoAdapt: A Multi-Scene Egocentric Adaptation Method for CVPR 2026 HD-EPIC VQA Challenge
Zhiwei Chen, Yupeng Hu, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 96 |
EgoAction: Egocentric Action Composition with Reliability-Aware Temporal Fusion for the EPIC-KITCHENS Action Detection Challenge at CVPR 2026
Zhiheng Fu, Zixu Li, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 97 |
OmniEgo-R$^2$: A Routed Reasoning Framework for the 1st Cross-Domain EgoCross Challenge at CVPR 2026
Zixu Li, Zhiwei Chen, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 98 |
TempRet: Temporal Enhancement and Two-Stage Reranking for CVPR 2026 EPIC-KITCHENS-100 Multi-Instance Retrieval Challenge
Zixu Li, Yupeng Hu, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 99 |
EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy
Jinzhao Li, Yinuo Chen, ... (+10 more)
|
|
cs.CV
|
0 |
1 month ago |
| 100 |
VectorArk: Learning Practical Image Vectorization with Rounded Polygon Representation
Tarun Gehlaut, Difan Liu, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |