| 51 |
PS-MOT: Cultivating Instance Awareness from Point Seeds for Multi-Object Tracking
Kai Luo, Fei Teng, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 52 |
On the Vulnerability of Parameter-Level Defenses to Model Merging
Kuangpu Guo, Qingyan Zheng, ... (+5 more)
|
|
cs.LG
|
0 |
1 month ago |
| 53 |
Residual-Guided Expert Specialization for Incomplete Multimodal Learning
Seunghun Baek, Jihwan Park, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 54 |
VisReflect: Latent Visual Reflection for Fine-Grained Perception in Long Visual Context
Xiaoqian Shen, Mohamed Elhoseiny
|
|
cs.CV
|
0 |
1 month ago |
| 55 |
Intermediate Text Representation Guided Text-to-Image Generation for Enhancing One-and-Only Alignment
Soyoun Won, Aryan Yazdan Parast, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 56 |
Your Data Manifold is Secretly a Reward Model: Shell-LCC for Text-to-Video Generation
Shihao Zhang, Yuguang Yan, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 57 |
SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation
Zhiyuan Ma, Zhengfeng Shi, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 58 |
Walking in the Implicit: Interactive World Exploration via Neural Scene Representation
Zhiqi Li, Chengrui Dong, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 59 |
CogSENet: Blind Image Deblurring with Blur-Conditioned Semantic Routing and Explicit Frequency Fusion
Pan Wang, Yihao Hu, Xiujin Liu
|
|
cs.CV
|
0 |
1 month ago |
| 60 |
OmniDance: Multimodal Driven Dance Video Generation with Large-scale Internet Data
Kaixing Yang, Jiashu Zhu, ... (+9 more)
|
|
cs.CV
|
0 |
1 month ago |
| 61 |
Monte Carlo Energy Aggregation for Mobile 3D Gaussian Splatting
Xiaobiao Du, YuAn Wang, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 62 |
SkelEM: Training-Signal Decoupling of Skeleton and Diffusion for Self-supervised Axial Super-Resolution in Volume Microscopy
Bohao Chen, Yanchao Zhang, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 63 |
GeoEdit: Geometry-Aware Object Editing via Dual-Branch Denoising
Yi He, Jiangming Wang, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 64 |
SICAGE: Speaker-Independent Culture-Aware Gesture Generation using TED4C-L Dataset
Ariel Gjaci, Antonio Sgorbissa, Vittorio Murino
|
|
cs.CV
|
0 |
1 month ago |
| 65 |
Explainability-Aware Frustum Attack: Exposing Structural Vulnerabilities in LiDAR-Based 3D Object Detectors
Chengzeng You, Binbin Xu, Soteris Demetriou
|
|
cs.CV
|
0 |
1 month ago |
| 66 |
Exploiting Local Flatness for Efficient Out-of-Distribution Detection
Seonghwan Park, Hyunji Jung, ... (+2 more)
|
|
cs.LG
|
0 |
1 month ago |
| 67 |
Seeing Touch from Motion: A Unified Modality-Aware Visuo-Tactile Policy with Tactile Motion Correlation
Shengqi Xu, Guojin Zhong, ... (+8 more)
|
|
cs.RO
|
0 |
1 month ago |
| 68 |
Pondering the Way: Spatial-perceiving World Action Model for Embodied Navigation
Hong Chen, Daqi Liu, ... (+11 more)
|
|
cs.RO
|
0 |
1 month ago |
| 69 |
IREU: Identity-Related Encoder-Only Unlearning for Customized Portrait Generation
Chaoyi Shi, Shanshan Zhang, Jian Yang
|
|
cs.CV
|
0 |
1 month ago |
| 70 |
See Only When Needed: Context-Aware Attention Intervention for Mitigating Hallucinations in LVLMs
Yuqing Lei, Wenbo Lyu, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 71 |
UniTriSplat: A Unified 3D Gaussian Splatting Framework with Uniform Spherical Rasterization for Universal Cameras
Yipeng Zhu, Huajian Huang, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 72 |
OP3DSG: Open-Vocabulary Part-Aware 3D Scene Graph Generation for Real-World Environments
Yirum Kim, Ue-Hwan Kim
|
|
cs.CV
|
0 |
1 month ago |
| 73 |
AerialMetric: Benchmarking and Adapting UAV Monocular Metric Depth Estimation in the Real World
Zhongqiang Song, Guanying Chen, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 74 |
PoseShield: Neural Collision Fields for Human Self-Collision Resolution
Zhengyuan Li, Zeyun Deng, ... (+8 more)
|
|
cs.CV
|
0 |
1 month ago |
| 75 |
ScAle: Attention Head Scaling as a Minimal Adapter for Spatial Reasoning in Vision Language Models
Rahul Chowdhury, Timothy A Rupprecht, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 76 |
MotionAtlas: Detailed Region Captioning for Motion-Centric Videos
Weisong Liu, Haochen Wang, ... (+9 more)
|
|
cs.CV
|
0 |
1 month ago |
| 77 |
Learning Transferable Dynamics Priors from Action to World Modeling
Ze Huang, Jiahui Zhang, ... (+4 more)
|
|
cs.RO
|
0 |
1 month ago |
| 78 |
Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation
Jongoh Jeong, Sun-Kyung Lee, Kuk-Jin Yoon
|
|
cs.CV
|
0 |
1 month ago |
| 79 |
MIRROR: Aligning Semantic Relations from Language to Image via Gromov--Wasserstein
Hong-Han Wang, Yuntao Wang, Hu Ding
|
|
cs.CV
|
0 |
1 month ago |
| 80 |
From Phase to Phenomenon: Self-Supervised Learning of Subsurface Scattering with Minimal Phase-shift Inputs
Arjun Majumdar, Raphael Braun, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 81 |
Bridging VideoQA and Video-Guided Agentic Tasks via Generalized Keyframe Extraction
Sunqi Fan, Qingle Liu, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 82 |
Robust Zero-shot Anomaly Detection under Limited Auxiliary Anomaly Priors
Guanyu Lu, Fang Zhou, Cheqing Jin
|
|
cs.CV
|
0 |
1 month ago |
| 83 |
NaLA: A 3D Native LLM Layout Agent for High-quality 3D Scene Generation
Cheng Wan, Yongsen Mao, ... (+8 more)
|
|
cs.CV
|
0 |
1 month ago |
| 84 |
Multi-scale Object-Aware Gaze Estimation via Geometric Reasoning
Jiajie Mi, Xinyu Liu, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 85 |
RAGA: Real Time Ray Traced Gaussian Shadow Casting for 3DGS Avatar-Scene Interaction
Aymen Mir, Riza Alp Guler, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 86 |
MirrorPPR: Exemplar-Based Portrait Photo Retouching
Zhihong Liu, Zheng Li, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 87 |
Pointer-CAD v2: Plan-Then-Construct CAD Generation with Dimension-Aware Parametric Precision
Dacheng Qi, Chenyu Wang, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 88 |
ASTAD: Asymmetric Style Transfer for Synthetic-to-Real Adaptation in Autonomous Driving
Dingyi Yao, Xinqi Zhang, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 89 |
BrainRiem: Riemannian Prototype Learning for Source-Free Cross-Site Brain Network Diagnosis
Kunyu Zhang, Tianxiang Xu
|
|
cs.LG
|
0 |
1 month ago |
| 90 |
DTI: Dynamic Trajectory Initialization for Generative Face Video Super-Resolution
Yingwei Tang, Chen Yan, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 91 |
Adaptive Spectrum-Aware Feature Disentangled Network for Small Object Detection
Yang Guo, Zihan Yang, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 92 |
On Test-Time Scaling for Vision-Language Models
Fawaz Sammani, Tzoulio Chamiti, Nikos Deligiannis
|
|
cs.CV
|
0 |
1 month ago |
| 93 |
DLGStream: Dynamic Language-embedded Guassian Splatting for Open-vocabulary Enabled Free-viewpoint Video Streaming
Zhihui Ke, Yuyang Liu, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 94 |
FreqOrtho-SR: Frequency-Guided Orthogonal Expert Learning for Real-World Image Super-Resolution
Minh Son Hoang, Dinh Phu Tran, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 95 |
LogiCo: A Unified Framework for Logical and Structural Anomaly Detection
Ximiao Zhang, Min Xu, Xiuzhuang Zhou
|
|
cs.CV
|
0 |
1 month ago |
| 96 |
BackTranslation2.0 -- A Linguistically Motivated Metric to Assess Sign Language Production
Oliver Cory, Maksym Ivashechkin, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 97 |
FedLAS: Feature-Modulated Bidirectional Label Smoothing for Neural Network Calibration
Thiru Thillai Nadarasar Bahavan, Sachith Seneviratne, Saman Halgamuge
|
|
cs.CV
|
0 |
1 month ago |
| 98 |
Obliviate: Erasing Concepts from Autoregressive Image Generation Models
Hossein Shakibania, Jonas Henry Grebe, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 99 |
SIGNET: Motion-Level Knowledge Transfer for Cross-Language Sign Language Translation
Sobhan Asasi, Ozge Mercanoglu Sincan, Richard Bowden
|
|
cs.CV
|
0 |
1 month ago |
| 100 |
RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning
Yelin Wang, Zijia Song, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |