| 51 |
When Interpretability Becomes a Liability: Adversarial Attacks on CBM Concept Layers
Aditya Sridhar
|
|
cs.LG
|
0 |
2 months ago |
| 52 |
Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation
Aviral Chharia, Fernando De la Torre
|
|
cs.CV
|
0 |
2 months ago |
| 53 |
Tempered Self-Similarity Alignment for Physically Plausible Video Generation
Manjin Kim, Suha Kwak, Minsu Cho
|
|
cs.CV
|
0 |
2 months ago |
| 54 |
Three-Step Conditional Diffusion 3D Reconstruction for Light-Field Microscopy
Qihong Zhao, Shaokang Yan, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 55 |
BED-SAM2: Boundary-Enhanced-Depth SAM2 via Monocular Geometric Priors
Tyler Rust, Dara McNally, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 56 |
HCL-FF: Hierarchical and Contrastive Learning for Forward-Forward Algorithm
Jie-En Yao, Hong-En Chen, C. -C. Jay Kuo
|
|
cs.CV
|
0 |
2 months ago |
| 57 |
4KLSDB: A Large-Scale Dataset for 4K Image Restoration and Generation
Zihao Zhu, Kuan-Ru Huang, ... (+7 more)
|
|
cs.CV
|
0 |
2 months ago |
| 58 |
Ghosts in the Point Clouds: De-glaring LiDAR in the Transient Domain
Avery Gump, Connor Henley, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 59 |
NudgeVAD: Language-Nudged End-to-End Driving via FiLM Residuals
Chieh-Chi Yang, Yu-Hsiang Chen, Yi-Ting Chen
|
|
cs.CV
|
0 |
2 months ago |
| 60 |
EgoAdapt: A Multi-Scene Egocentric Adaptation Method for CVPR 2026 HD-EPIC VQA Challenge
Zhiwei Chen, Yupeng Hu, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 61 |
EgoAction: Egocentric Action Composition with Reliability-Aware Temporal Fusion for the EPIC-KITCHENS Action Detection Challenge at CVPR 2026
Zhiheng Fu, Zixu Li, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 62 |
OmniEgo-R$^2$: A Routed Reasoning Framework for the 1st Cross-Domain EgoCross Challenge at CVPR 2026
Zixu Li, Zhiwei Chen, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 63 |
TempRet: Temporal Enhancement and Two-Stage Reranking for CVPR 2026 EPIC-KITCHENS-100 Multi-Instance Retrieval Challenge
Zixu Li, Yupeng Hu, ... (+5 more)
|
|
cs.CV
|
0 |
2 months ago |
| 64 |
EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy
Jinzhao Li, Yinuo Chen, ... (+10 more)
|
|
cs.CV
|
0 |
2 months ago |
| 65 |
VectorArk: Learning Practical Image Vectorization with Rounded Polygon Representation
Tarun Gehlaut, Difan Liu, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 66 |
Causal Physics Steering in Video World Models via Concept Activation Vectors
Nahid Alam
|
|
cs.CV
|
0 |
2 months ago |
| 67 |
HumanNOVA: Photorealistic, Universal and Rapid 3D Human Avatar Modeling from a Single Image
Hezhen Hu, Wangbo Zhao, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 68 |
Not All Points Are Equal: Uncertainty-Aware 4D LiDAR Scene Synthesis
Xiang Xu, Alan Liang, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 69 |
Question-Aware Evidence Ledgers for Video Relational Reasoning
Yilin Ou, Mengshi Qi, Huadong Ma
|
|
cs.CV
|
0 |
1 month ago |
| 70 |
MASER: Modality-Adaptive Specialist Routing for Embodied 3D Spatial Intelligence
Hilton Raj, Vishnuram AV
|
|
cs.CV
|
0 |
1 month ago |
| 71 |
Edge Prediction for Roof Wireframe Reconstruction with Transformers
Gustav Hanning, Ludvig Dillén, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 72 |
Training-Free Composed Video Retrieval via Visual Representation-Guided Video-LLM Reasoning
Yang Liu, Qianqian Xu, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 73 |
Ultra Diffusion Poser: Diffusion-Based Human Motion Tracking From Sparse Inertial Sensors and Ranging-Based Between-Sensor Distances
Dominik Hollidt, Tommaso Bendinelli, Christian Holz
|
|
cs.CV
|
0 |
1 month ago |
| 74 |
Residual Decoder Adapter: ID-Preserving Tokenizer Adaption for Autoregressive Text Rendering
Dongxing Mao, Jinpeng Wang, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 75 |
Understanding Identity Continuity in Thermal Video through Scene-Level Consistency
Wei-Chieh Sun, Gyungmin Ko, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 76 |
CanonCGT: Reference-Based Color Grading via Canonical Pivot Representation
Jinwon Ko, Keunsoo Ko, Chang-Su Kim
|
|
cs.CV
|
0 |
1 month ago |
| 77 |
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs
Sicheng Xu, Yu Deng, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 78 |
DENSER: Depth-Guided Ensemble with Staged EFA-GS Reconstruction for Soccer Novel View Synthesis
Parthsarthi Rawat
|
|
cs.CV
|
0 |
2 months ago |
| 79 |
MindClaw: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
Ruoxuan Zhang, Qiaoqiao Wan, ... (+5 more)
|
|
cs.AI
|
0 |
2 months ago |
| 80 |
Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation
Ziyue Lin, Jiahe Hou, ... (+7 more)
|
|
cs.CV
|
0 |
2 months ago |
| 81 |
Cross-Axis Feature Fusion with Joint-Wise Motion Difference Prediction for Text-Based 3D Human Motion Editing
Gyojin Han, Junmo Kim
|
|
cs.CV
|
0 |
2 months ago |
| 82 |
GABI: Geometry-Aware Boundary Integration for Spacecraft Segmentation
Iason Georgios Velentzas, Dhruv Ahuja, Panagiotis Tsiotras
|
|
cs.CV
|
0 |
2 months ago |
| 83 |
Generative Diffusion Priors for 3D Mapping of the Dark Universe
Brandon Zhao, Diana Scognamiglio, ... (+2 more)
|
|
astro-ph.CO
|
0 |
2 months ago |
| 84 |
SoccerNet 2026 Player-Centric Ball-Action Spotting:Retraining and Post-Processing Extensions to the FOOTPASS Baselines
Parthsarthi Rawat
|
|
cs.CV
|
0 |
1 month ago |
| 85 |
EditSSC: Toward Editable Semantic Occupancy Scenes with Unconditional Diffusion Models
Fatima Balde, Raoul de Charette, Alexandre Boulch
|
|
cs.CV
|
0 |
1 month ago |
| 86 |
Property-Informed Diffusion-Based Text-to-Microstructure Generation
Bingxuan Dai, Hongsong Wang, Jie Gui
|
|
cs.CV
|
0 |
1 month ago |
| 87 |
IEA: Amateur-Friendly Conversational Image Editing Agent via Three Stages of Multitask Alignment
Zichen Zhu, Yuheng Sun, ... (+12 more)
|
|
cs.CV
|
0 |
1 month ago |
| 88 |
ExMesh: EXplicit Mesh Reconstruction with Topology Adaptation
Chuanjin Fan, Lifan Wu, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 89 |
When CLIP Sees More, It Fights Back Harder: Multi-View Guided Adaptive Counterattacks for Test-Time Adversarial Robustness
Sunoh Kim, Daeho Um
|
|
cs.CV
|
0 |
1 month ago |
| 90 |
MotionEnhancer: Leveraging Video Diffusion for Motion-Enhanced Vision-Language Models
Yifan Xu, Chao Zhang, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 91 |
Rethinking Infrastructure Inspection as Image Difference Classification: A Traffic Sign Case Study
Ching Yau Fergus Mok, Lavindra de Silva, ... (+2 more)
|
|
cs.AI
|
0 |
1 month ago |
| 92 |
LLM-Conditioned Synthesis of Pathological Gaits via Structured Gait-Language Representations
Mritula Chandrasekaran, Sanket Kachole, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 93 |
Parallel Jacobi Decoding for Fast Autoregressive Image Generation
Boya Liao, Ying Li, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 94 |
Revisiting Prototype Rehearsal for Exemplar-Free Continual Learning: Manifold-Aware Boundary Sampling with Adaptive Class-Balanced Loss
Hongye Xu, Bartosz Krawczyk
|
|
cs.LG
|
0 |
1 month ago |
| 95 |
UniPixie: Unified and Probabilistic 3D Physics Learning via Flow Matching
Qilin Huang, Quynh Anh Huynh, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 96 |
Recovering Physically Plausible Human-Object Interactions from Monocular Videos
Dingbang Huang, Etienne Vouga, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 97 |
Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models
Jialin Wu, Qianru Zhang, ... (+2 more)
|
|
eess.IV
|
0 |
1 month ago |
| 98 |
Scene-Centric Unsupervised Video Panoptic Segmentation
Christoph Reich, Oliver Hahn, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 99 |
MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation
Jiale Xu, Wang Zhao, Ying Shan
|
|
cs.CV
|
0 |
1 month ago |
| 100 |
4D Reconstruction from Sparse Dynamic Cameras
Kazuki Ozeki, Shun Kenney, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |