| 151 |
Geometry-Anchored Transport Framework for Exemplar-Free Class-Incremental Learning
Hongye Xu, Bartosz Krawczyk
|
|
cs.LG
|
0 |
1 month ago |
| 152 |
Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity
Heethanjan Kanagalingam, Thenukan Pathmanathan, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 153 |
Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation
Atin Pothiraj, Jaemin Cho, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 154 |
OrthoTrack: Continuous 6-DoF UAV Trajectory Estimation Anchored in Public Orthophotos
Oussema Dhaouadi, Zuria Bauer, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 155 |
Self-supervised Garment Dynamics with Persistent Wrinkles
Xiaoyuan Yang, Deshan Gong, ... (+2 more)
|
|
cs.GR
|
0 |
1 month ago |
| 156 |
BioMedVR: Confusion-Aware Mixture-of-Prompt Experts for Biomedical Visual Reprogramming
Jiaxiang Liu, Tianxiang Hu, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 157 |
Evaluating the Interpretability of Sparse Autoencoders with Concept Annotations
Jonas Klotz, Cassio F. Dantas, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 158 |
Agentic Collaborative Cognition for Zero-Shot 3D Understanding
Wenxin Wang, Bo Zhang, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 159 |
VisCritic: Visual State Comparison as Process Reward for GUI Agents
Jiachen Qian
|
|
cs.CV
|
0 |
1 month ago |
| 160 |
Advancing WordArt-Oriented Scene Text Recognition: Datasets and Methods
Xingsong Ye, Yongkun Du, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 161 |
MambaRaw: Selective State Space Modeling for Efficient 4K Raw Image Reconstruction
Peize Li, Fanhu Zeng, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 162 |
SENTRY: SAM2-Enhanced Neighbor-Aware and Temporally Reasoned Memory for Visual Tracking
Mohamad Alansari, Yonathan Michael, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 163 |
EgoSAT: A Comprehensive Benchmark of Egocentric Streaming Interaction Understanding
Yijia Lei, Jinzhao Li, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 164 |
MATCH: Flow Matching for Multi-View Anomaly Detection
Mathis Kruse, Melissa Schween, Bodo Rosenhahn
|
|
cs.CV
|
0 |
1 month ago |
| 165 |
SignNet-1M: Large-Scale Multilingual Sign Language Video Dataset with Downstream Benchmarks
Zhewen He, Junyi Hu, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 166 |
Open-Vocabulary BEV Segmentation with 3D-Aware Geometric Constraints
Hojun Choi, Seulbin Hwang, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 167 |
Curvature-Guided Mixing for MLLM Adaptation
Jinglong Yang, Jiaxuan He, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 168 |
UniTranslator: A Unified Multi-modal Framework for End-to-end In-Image Machine Translation
Jiahao Lyu, Pei Fu, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 169 |
Training-free Cross-domain Few-shot Segmentation via Robust Semantic Representation and Matching
Sujun Sun, Mingwu Ren, Haofeng Zhang
|
|
cs.CV
|
0 |
1 month ago |
| 170 |
Hierarchical Spatial and Channel Aggregation for Cross-domain Few-shot Segmentation
Sujun Sun, Mingwu Ren, Haofeng Zhang
|
|
cs.CV
|
0 |
1 month ago |
| 171 |
Spectral Evolution-Guided Token Pruning in Multimodal Large Language Models
Bin Chen, Yuxiang Cai, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 172 |
Accelerating Multimodal Large Language Models with Prior-Corrected Token Reduction
Zengjie Chen, Yuxiang Cai, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 173 |
Geometry-Aware Style Transfer in 3D Gaussian Splatting
Min Hyeok Bang, Jun Hyeong Kim, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 174 |
NavWM: A Unified Navigation World Model for Foresight-Driven Planning
Yanghong Mei, Longteng Guo, ... (+4 more)
|
|
cs.RO
|
0 |
1 month ago |
| 175 |
Fabric Image Demoiréing Benchmark from Synthesis to Restoration
Pengchao Wei, Xiaojie Guo
|
|
cs.CV
|
0 |
1 month ago |
| 176 |
DivRL: Disentangled Self-Similarity Rewards for Diverse Subject-Driven Generation
Qian Wang, Zhenyu Li, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 177 |
Trustworthy Image Authentication using Forensic Knowledge Graphs
Tai D. Nguyen, Matthew C. Stamm
|
|
cs.CV
|
0 |
1 month ago |
| 178 |
MGI: Member vs Generated Inference
Bihe Zhao, Michel Meintz, ... (+3 more)
|
|
cs.LG
|
0 |
1 month ago |
| 179 |
Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models
Hongyu Li, Wanjia Fu, ... (+12 more)
|
|
cs.RO
|
0 |
24 days ago |
| 180 |
InFlux++: Real and Synthetic Data for Estimating Dynamic Camera Intrinsics
Erich Liang, Caleb Kha-Uong, ... (+6 more)
|
|
cs.CV
|
0 |
24 days ago |
| 181 |
MV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-Forcing
Gal Fiebelman, Hadar Averbuch-Elor, Sagie Benaim
|
|
cs.CV
|
0 |
24 days ago |
| 182 |
WildSplat: Feedforward Gaussian Splatting from Unposed In-the-Wild Images
Xiyu Zhang, Jingyu Zhuang, ... (+5 more)
|
|
cs.CV
|
0 |
24 days ago |
| 183 |
Fully Rotation-Equivariant Spectral-Spatial Learning for Multispectral Object Detection
Peng Zhang, Tingfa Xu, ... (+2 more)
|
|
cs.CV
|
0 |
24 days ago |
| 184 |
RADIANCE: Relative Adaptive Denoising with IP-Adapter for Novel Concept Enhancement
Zi-Xiang Ni, Bo-Lun Huang, ... (+3 more)
|
|
cs.CV
|
0 |
24 days ago |
| 185 |
MemPose: Category-level Object Pose Estimation with Memory
Xiao Lin, Minghao Zhu, ... (+5 more)
|
|
cs.CV
|
0 |
24 days ago |
| 186 |
DGSeg: Dynamic Gating of Semantic-Spatial Guided Predictions for Reasoning Segmentation
Ruizhe Zeng, Siyu Cao, ... (+2 more)
|
|
cs.CV
|
0 |
24 days ago |
| 187 |
Learning Probabilistic Prompt for Continual Learning
Hyekang Park, Sanghoon Lee, ... (+3 more)
|
|
cs.CV
|
0 |
24 days ago |
| 188 |
From Open Loop to Closed Loop: A Test-Time Iterative Optimization Framework for Reference-Consistent Image Generation
Baixuan Zhao, Xinyu Zhang, ... (+5 more)
|
|
cs.CV
|
0 |
24 days ago |
| 189 |
Learning Structured Visual Compositional Representations for Weakly Supervised Referring Expression Comprehension
Lian Xu, Mohammed Bennamoun, ... (+4 more)
|
|
cs.CV
|
0 |
24 days ago |
| 190 |
PixelPilot: Scalable Vision-Language-Action Models for End-to-End Autonomous Driving
Pin Tang, Guoqing Wang, ... (+5 more)
|
|
cs.CV
|
0 |
24 days ago |
| 191 |
Integrated Forward-Inverse Network for Lensless Image Reconstruction
Donggeon Bae, Jaewoo Jung, ... (+7 more)
|
|
cs.CV
|
0 |
24 days ago |
| 192 |
RAF: Reliability-Aware Fusion of Camera, LiDAR, and 4D RADAR for Robust 3D Object Detection in Adverse Weather
Heejun Park, Jaeseok Jeong, Kuk-Jin Yoon
|
|
cs.CV
|
0 |
24 days ago |
| 193 |
QSVideo: Query-Conditioned Semantic Temporal Retrieval for Video Understanding
Wei Ao, Lan Wang, Vishnu Naresh Boddeti
|
|
cs.CV
|
0 |
24 days ago |
| 194 |
CCFM: Collision-Constrained Flow Matching for Safety-Critical Scenario Generation
Ke Li, Kaidi Liang, ... (+4 more)
|
|
cs.CV
|
0 |
25 days ago |
| 195 |
Transferability Between Understanding and Generation in Unified Multimodal Models
Jiwon Kang, Heeji Yoon, ... (+6 more)
|
|
cs.CV
|
0 |
25 days ago |
| 196 |
AdaptiveSplat:Texture Aware Controllable 3D Gaussian Allocation for Feed-Forward Reconstruction
Badrinath Singhal, Srihari K G, ... (+3 more)
|
|
cs.CV
|
0 |
25 days ago |
| 197 |
Beyond Random Sampling: Distribution-Aware Alignment for Semi-Supervised Medical Image Segmentation
Weihao Yan, Yeqiang Qian, ... (+2 more)
|
|
cs.CV
|
0 |
25 days ago |
| 198 |
CritiqueDriveVLM: From Verifier-Guided Reinforcement Learning to Latent Thought Distillation for Autonomous Driving
Zhaohong Liu, Hao Ye, ... (+2 more)
|
|
cs.CV
|
0 |
25 days ago |
| 199 |
Perceiving Better Moments: Cover Frame Reselection and Enhancement for Live Photos with the Live2K Dataset
Junyu Lou, Kai Chen, ... (+4 more)
|
|
cs.CV
|
0 |
25 days ago |
| 200 |
HCSU: A Dataset and Benchmark for Fine-Grained Historical Calligraphy Style Understanding
Yinsheng Yao, Yan Liu, Chen Ye
|
|
cs.CV
|
0 |
25 days ago |