| 1 |
BMD-45: A Large-Scale CCTV Vehicle Detection Dataset for Urban Traffic in Developing Cities
Akash Sharma, Chinmay Mhatre, ... (+9 more)
|
|
cs.CV
|
0 |
2 months ago |
| 2 |
POCA: Pareto-Optimal Curriculum Alignment for Visual Text Generation
Yaohou Fan, Qingzhong Wang, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 3 |
TopoHR: Hierarchical Centerline Representation for Cyclic Topology Reasoning in Driving Scenes with Point-to-Instance Relations
Yifeng Bai, Zhirong Chen, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |
| 4 |
Jailbreaking Frontier Foundation Models Through Intention Deception
Xinhe Wang, Katia Sycara, Yaqi Xie
|
|
cs.CR
|
0 |
2 months ago |
| 5 |
SMoES: Soft Modality-Guided Expert Specialization in MoE-VLMs
Zi-Hao Bo, Yaqian Li, ... (+7 more)
|
|
cs.CV
|
0 |
2 months ago |
| 6 |
Bringing a Personal Point of View: Evaluating Dynamic 3D Gaussian Splatting for Egocentric Scene Reconstruction
Jan Warchocki, Xi Wang, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |
| 7 |
ELSA: Exact Linear-Scan Attention for Fast and Memory-Light Vision Transformers
Chih-Chung Hsu, Xin-Di Ma, ... (+2 more)
|
|
cs.LG
|
0 |
2 months ago |
| 8 |
Learning to Identify Out-of-Distribution Objects for 3D LiDAR Anomaly Segmentation
Simone Mosco, Daniel Fusaro, Alberto Pretto
|
|
cs.CV
|
0 |
2 months ago |
| 9 |
Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection
Sanghoon Lee, Geon Lee, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |
| 10 |
Learning from Noisy Prompts: Saliency-Guided Prompt Distillation for Robust Segmentation with SAM
Jingxuan Kang, Ziqi Zhang, ... (+10 more)
|
|
cs.CV
|
0 |
2 months ago |
| 11 |
SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation
Kaiwen Huang, Yi Zhou, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 12 |
MotionHiFlow: Text-to-motion via hierarchical flow matching
Heng Li, Xiaotong Lin, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 13 |
One Identity, Many Roles: Multimodal Entity Coreference for Enhanced Video Situation Recognition
Balaji Darur, Amanmeet Garg, Makarand Tapaswi
|
|
cs.CV
|
0 |
2 months ago |
| 14 |
GenAssets: Generating in-the-wild 3D Assets in Latent Space
Ze Yang, Jingkang Wang, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 15 |
Hard to See, Hard to Label: Generative and Symbolic Acquisition for Subtle Visual Phenomena
Renjith Prasad, Rishabh Sharma, ... (+8 more)
|
|
cs.CV
|
0 |
2 months ago |
| 16 |
BrickNet: Graph-Backed Generative Brick Assembly
Peter Kulits, Cordelia Schmid
|
|
cs.CV
|
0 |
2 months ago |
| 17 |
Data-Free Contribution Estimation in Federated Learning using Gradient von Neumann Entropy
Asim Ukaye, Mubarak Abdu-Aguye, ... (+2 more)
|
|
cs.LG
|
0 |
2 months ago |
| 18 |
Are Natural-Domain Foundation Models Effective for Accelerated Cardiac MRI Reconstruction?
Anam Hashmi, Mayug Maniparambil, ... (+3 more)
|
|
eess.IV
|
0 |
2 months ago |
| 19 |
Revisiting Geometric Obfuscation with Dual Convergent Lines for Privacy-Preserving Image Queries in Visual Localization
Jeonggon Kim, Heejoon Moon, Je Hyeong Hong
|
|
cs.CV
|
0 |
2 months ago |
| 20 |
DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning
Joonmyung Choi, Sanghyeok Lee, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 21 |
Uni-Encoder Meets Multi-Encoders: Representation Before Fusion for Brain Tumor Segmentation with Missing Modalities
Peibo Song, Xiaotian Xue, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 22 |
GenMatter: Perceiving Physical Objects with Generative Matter Models
Eric Li, Arijit Dasgupta, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 23 |
H-Sets: Hessian-Guided Discovery of Set-Level Feature Interactions in Image Classifiers
Ayushi Mehrotra, Dipkamal Bhusal, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |
| 24 |
Vista4D: Video Reshooting with 4D Point Clouds
Kuan Heng Lin, Zhizheng Liu, ... (+10 more)
|
|
cs.CV
|
0 |
2 months ago |
| 25 |
UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection
Yanran Zhang, Wenzhao Zheng, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 26 |
Addressing Image Authenticity When Cameras Use Generative AI
Umar Masud, Abhijith Punnappurath, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 27 |
Back to Source: Open-Set Continual Test-Time Adaptation via Domain Compensation
Yingkai Yang, Chaoqi Chen, Hui Huang
|
|
cs.CV
|
0 |
2 months ago |
| 28 |
Bridging the Training-Deployment Gap: Gated Encoding and Multi-Scale Refinement for Efficient Quantization-Aware Image Enhancement
Dat To-Thanh, Nghia Nguyen-Trong, ... (+3 more)
|
|
cs.AI
|
0 |
2 months ago |
| 29 |
Ramen: Robust Test-Time Adaptation of Vision-Language Models with Active Sample Selection
Wenxuan Bao, Yanjun Zhao, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |
| 30 |
Unlocking the Power of Critical Factors for 3D Visual Geometry Estimation
Guangkai Xu, Hua Geng, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 31 |
DualSplat: Robust 3D Gaussian Splatting via Pseudo-Mask Bootstrapping from Reconstruction Failures
Xu Wang, Zhiru Wang, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 32 |
DCMorph: Face Morphing via Dual-Stream Cross-Attention Diffusion
Tahar Chettaoui, Eduarda Caldeira, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 33 |
Instance-level Visual Active Tracking with Occlusion-Aware Planning
Haowei Sun, Kai Zhou, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 34 |
MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction
Yisheng He, Steven Hoi
|
|
cs.CV
|
0 |
2 months ago |
| 35 |
VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition
Tanush Yadav, Mohammadreza Salehi, ... (+7 more)
|
|
cs.CV
|
0 |
2 months ago |
| 36 |
M\textsuperscript{4}Fuse: Lightweight State-Space MoE with a Cross-Scale Gating Bridge for Brain Tumor Segmentation
Meihua Zhou, Xinyu Tong, Li Yang
|
|
cs.CV
|
0 |
2 months ago |
| 37 |
Fine-Tuning Impairs the Balancedness of Foundation Models in Long-tailed Personalized Federated Learning
Shihao Hou, Chikai Shang, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 38 |
From Where Things Are to What They Are For: Benchmarking Spatial-Functional Intelligence in Multimodal LLMs
Le Zhang, Jihan Yang, ... (+12 more)
|
|
cs.CV
|
0 |
2 months ago |
| 39 |
CADFS: A Big CAD Program Dataset and Framework for Computer-Aided Design with Large Language Models
Vladislav Pyatov, Gleb Bobrovskikh, ... (+7 more)
|
|
cs.CV
|
0 |
2 months ago |
| 40 |
High-Fidelity Mobile Avatars with Pruned Local Blendshapes
Youyi Zhan, He Wang, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |
| 41 |
Profile-Specific 3DMM Regression from a Single Lateral Face Image
Taiki Kanaya, Hideo Saito
|
|
cs.CV
|
0 |
2 months ago |
| 42 |
Multi-Scale Gaussian-Language Map for Zero-shot Embodied Navigation and Reasoning
Sixian Zhang, Yiyao Wang, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 43 |
TrajRAG: Retrieving Geometric-Semantic Experience for Zero-Shot Object Navigation
Yiyao Wang, Sixian Zhang, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 44 |
Act2See: Emergent Active Visual Perception for Video Reasoning
Martin Q. Ma, Yuxiao Qu, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 45 |
Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection
Tianxiao Li, Zhenglin Huang, ... (+11 more)
|
|
cs.CV
|
0 |
2 months ago |
| 46 |
Towards Visual Query Localization in the 3D World
Liang Peng, Bohan Tan, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 47 |
Decision Boundary-aware Generation for Long-tailed Learning
Jiacheng Yang, Ruichi Zhang, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 48 |
Registration-Free Learnable Multi-View Capture of Faces in Dense Semantic Correspondence
Panagiotis P. Filntisis, George Retsinas, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 49 |
VISTA: Video Interaction Spatio-Temporal Analysis Benchmark
Alejandro Aparcedo, Akash Kumar, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 50 |
Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance
Muyang Li, Yucheng Liu, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |