| 101 |
Causal Physics Steering in Video World Models via Concept Activation Vectors
Nahid Alam
|
|
cs.CV
|
0 |
1 month ago |
| 102 |
HumanNOVA: Photorealistic, Universal and Rapid 3D Human Avatar Modeling from a Single Image
Hezhen Hu, Wangbo Zhao, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 103 |
Not All Points Are Equal: Uncertainty-Aware 4D LiDAR Scene Synthesis
Xiang Xu, Alan Liang, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 104 |
Question-Aware Evidence Ledgers for Video Relational Reasoning
Yilin Ou, Mengshi Qi, Huadong Ma
|
|
cs.CV
|
0 |
1 month ago |
| 105 |
MASER: Modality-Adaptive Specialist Routing for Embodied 3D Spatial Intelligence
Hilton Raj, Vishnuram AV
|
|
cs.CV
|
0 |
1 month ago |
| 106 |
Edge Prediction for Roof Wireframe Reconstruction with Transformers
Gustav Hanning, Ludvig Dillén, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 107 |
Training-Free Composed Video Retrieval via Visual Representation-Guided Video-LLM Reasoning
Yang Liu, Qianqian Xu, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 108 |
Ultra Diffusion Poser: Diffusion-Based Human Motion Tracking From Sparse Inertial Sensors and Ranging-Based Between-Sensor Distances
Dominik Hollidt, Tommaso Bendinelli, Christian Holz
|
|
cs.CV
|
0 |
1 month ago |
| 109 |
Residual Decoder Adapter: ID-Preserving Tokenizer Adaption for Autoregressive Text Rendering
Dongxing Mao, Jinpeng Wang, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 110 |
Understanding Identity Continuity in Thermal Video through Scene-Level Consistency
Wei-Chieh Sun, Gyungmin Ko, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 111 |
CanonCGT: Reference-Based Color Grading via Canonical Pivot Representation
Jinwon Ko, Keunsoo Ko, Chang-Su Kim
|
|
cs.CV
|
0 |
1 month ago |
| 112 |
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs
Sicheng Xu, Yu Deng, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 113 |
DENSER: Depth-Guided Ensemble with Staged EFA-GS Reconstruction for Soccer Novel View Synthesis
Parthsarthi Rawat
|
|
cs.CV
|
0 |
1 month ago |
| 114 |
MindClaw: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
Ruoxuan Zhang, Qiaoqiao Wan, ... (+5 more)
|
|
cs.AI
|
0 |
1 month ago |
| 115 |
Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation
Ziyue Lin, Jiahe Hou, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 116 |
Cross-Axis Feature Fusion with Joint-Wise Motion Difference Prediction for Text-Based 3D Human Motion Editing
Gyojin Han, Junmo Kim
|
|
cs.CV
|
0 |
1 month ago |
| 117 |
GABI: Geometry-Aware Boundary Integration for Spacecraft Segmentation
Iason Georgios Velentzas, Dhruv Ahuja, Panagiotis Tsiotras
|
|
cs.CV
|
0 |
1 month ago |
| 118 |
Generative Diffusion Priors for 3D Mapping of the Dark Universe
Brandon Zhao, Diana Scognamiglio, ... (+2 more)
|
|
astro-ph.CO
|
0 |
1 month ago |
| 119 |
SoccerNet 2026 Player-Centric Ball-Action Spotting:Retraining and Post-Processing Extensions to the FOOTPASS Baselines
Parthsarthi Rawat
|
|
cs.CV
|
0 |
1 month ago |
| 120 |
EditSSC: Toward Editable Semantic Occupancy Scenes with Unconditional Diffusion Models
Fatima Balde, Raoul de Charette, Alexandre Boulch
|
|
cs.CV
|
0 |
1 month ago |
| 121 |
Property-Informed Diffusion-Based Text-to-Microstructure Generation
Bingxuan Dai, Hongsong Wang, Jie Gui
|
|
cs.CV
|
0 |
1 month ago |
| 122 |
IEA: Amateur-Friendly Conversational Image Editing Agent via Three Stages of Multitask Alignment
Zichen Zhu, Yuheng Sun, ... (+12 more)
|
|
cs.CV
|
0 |
1 month ago |
| 123 |
ExMesh: EXplicit Mesh Reconstruction with Topology Adaptation
Chuanjin Fan, Lifan Wu, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 124 |
When CLIP Sees More, It Fights Back Harder: Multi-View Guided Adaptive Counterattacks for Test-Time Adversarial Robustness
Sunoh Kim, Daeho Um
|
|
cs.CV
|
0 |
1 month ago |
| 125 |
MotionEnhancer: Leveraging Video Diffusion for Motion-Enhanced Vision-Language Models
Yifan Xu, Chao Zhang, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 126 |
Rethinking Infrastructure Inspection as Image Difference Classification: A Traffic Sign Case Study
Ching Yau Fergus Mok, Lavindra de Silva, ... (+2 more)
|
|
cs.AI
|
0 |
1 month ago |
| 127 |
LLM-Conditioned Synthesis of Pathological Gaits via Structured Gait-Language Representations
Mritula Chandrasekaran, Sanket Kachole, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 128 |
Parallel Jacobi Decoding for Fast Autoregressive Image Generation
Boya Liao, Ying Li, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 129 |
Revisiting Prototype Rehearsal for Exemplar-Free Continual Learning: Manifold-Aware Boundary Sampling with Adaptive Class-Balanced Loss
Hongye Xu, Bartosz Krawczyk
|
|
cs.LG
|
0 |
1 month ago |
| 130 |
UniPixie: Unified and Probabilistic 3D Physics Learning via Flow Matching
Qilin Huang, Quynh Anh Huynh, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 131 |
Recovering Physically Plausible Human-Object Interactions from Monocular Videos
Dingbang Huang, Etienne Vouga, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 132 |
Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models
Jialin Wu, Qianru Zhang, ... (+2 more)
|
|
eess.IV
|
0 |
1 month ago |
| 133 |
Scene-Centric Unsupervised Video Panoptic Segmentation
Christoph Reich, Oliver Hahn, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 134 |
MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation
Jiale Xu, Wang Zhao, Ying Shan
|
|
cs.CV
|
0 |
1 month ago |
| 135 |
4D Reconstruction from Sparse Dynamic Cameras
Kazuki Ozeki, Shun Kenney, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 136 |
Evaluating Reasoning Fidelity in Visual Text Generation
Jiajun Hong, Jiawei Zhou
|
|
cs.CV
|
0 |
1 month ago |
| 137 |
Ultra-Fast Neural Video Compression
Jiahao Li, Wenxuan Xie, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 138 |
Efficient and Training-Free Single-Image Diffusion Models
Haojun Qiu, Kiriakos N. Kutulakos, David B. Lindell
|
|
cs.CV
|
0 |
1 month ago |
| 139 |
A Cookbook of 3D Vision: Data, Learning Paradigms, and Application
Hongyang Du, Zongxia Li, ... (+9 more)
|
|
cs.CV
|
0 |
1 month ago |
| 140 |
Semantic Constraint Synthesis for Adaptive Trajectory Optimization via Large Language Models
Eleanor Brosius, Yuji Takubo, ... (+3 more)
|
|
math.OC
|
0 |
1 month ago |
| 141 |
Reflection Separation from a Single Image via Joint Latent Diffusion
Zheng-Hui Huang, Zhixiang Wang, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 142 |
Humanoid-GPT: Scaling Data and Structure for Zero-Shot Motion Tracking
Zekun Qi, Xuchuan Chen, ... (+11 more)
|
|
cs.RO
|
0 |
1 month ago |
| 143 |
Demo2Tutorial: From Human Experience to Multimodal Software Tutorials
Zechen Bai, Zhiheng Chen, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 144 |
Does Language Shift Break Medical Vision-Language Models? Indonesian Radiology Visual Question Answering Case Study
Pieter Christy Yan Yudhistira, Dzaki Rafif Malik, Novanto Yudistira
|
|
cs.CL
|
0 |
1 month ago |
| 145 |
Beyond Single Solution: Multi-Hypothesis Collaborative Deep Unfolding Network for Image Compressive Sensing
Wenxue Cui, Hualin Li, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 146 |
Eliciting Complex Spatial Reasoning in MLLMs through Wide-Baseline Matching
Hao Zhong, Muzhi Zhu, ... (+9 more)
|
|
cs.CV
|
0 |
1 month ago |
| 147 |
AvatarMix: Identity-Preserving Cross-Avatar Composition for Outfit Personalization
Zhaorong Wang, Yoshihiro Kanamori, Yuki Endo
|
|
cs.CV
|
0 |
1 month ago |
| 148 |
PersistGS: Differentiable Physics for Object Permanence in 4D Gaussian Splatting
Adrian Ramlal, John S. Zelek
|
|
cs.CV
|
0 |
1 month ago |
| 149 |
MARIO: Motion-Augmented Real-Time Multi-Sensor Inertial Odometry
Yiquan Li, Taeyoung Yeon, ... (+4 more)
|
|
cs.RO
|
0 |
1 month ago |
| 150 |
Hand Trajectory Fusion for Egocentric Natural Language Query Grounding
Enmin Zhong, Carlos R. del-Blanco, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |