| 401 |
LoCA: Spatially-Aware Low-Rank Convolutional Adaptation of Vision Foundation Models
Sojung An, Junha Lee, ... (+3 more)
|
|
cs.CV
|
0 |
24 days ago |
| 402 |
Retrieving and Refining Winning Noise Tickets for Diffusion-Based Motion Generation
Sakuya Ota, Qing Yu, ... (+3 more)
|
|
cs.CV
|
0 |
25 days ago |
| 403 |
WildCity: A Real-World City-Scale Testbed for Rendering, Simulation, and Spatial Intelligence
Xiangyu Han, Mengyu Yang, ... (+10 more)
|
|
cs.CV
|
0 |
25 days ago |
| 404 |
CoMind: Understanding Collaborative Human Activity from Multiple Minds and Views
Alexey Gavryushin, Dingxi Zhang, ... (+9 more)
|
|
cs.CV
|
0 |
25 days ago |
| 405 |
WristMimic: Full-Body Humanoid Control with Wrist-Guided Manipulation
Wongyun Yu, Youngwoon Kim, Minsu Cho
|
|
cs.RO
|
0 |
25 days ago |
| 406 |
Precise Video-to-Audio Generation with Cross-Modal Alignment in Latent Space
Thanh V. T. Tran, Ngoc-Son Nguyen, ... (+5 more)
|
|
cs.MM
|
0 |
25 days ago |
| 407 |
What Images Cannot Say: Language-Guided Olfactory Representation Learning
Eleftherios Tsonis, Xi Wang, Vicky Kalogeiton
|
|
cs.CV
|
0 |
25 days ago |
| 408 |
Straight-Path Flow Matching for Incomplete Multi-View Clustering
Yiteng Yuan, Junyan Wang, ... (+5 more)
|
|
cs.CV
|
0 |
25 days ago |
| 409 |
Revisiting Scene Graph Generation from the Perspective of Detector-Conditioned Reachability
Runfeng Qu, Pia K Bideau, ... (+4 more)
|
|
cs.CV
|
0 |
25 days ago |
| 410 |
High-Resolution Artwork Outpainting with Global Blueprint Guidance and Layout Control
Junha Kim, Hyunjoon Park, Donghyeon Cho
|
|
cs.CV
|
0 |
25 days ago |
| 411 |
RoME: Robust Mixture of Low-Rank Experts against Multiple Adversarial Perturbations
Woo Jae Kim, Kyle Min, ... (+3 more)
|
|
cs.CV
|
0 |
25 days ago |
| 412 |
SpaR3D-MoE: Adaptive 3D Spatial Reasoning from Sparse Views Meets Geometry-Inductive Mixture-of-Experts
Haida Feng, Hao Wei, ... (+4 more)
|
|
cs.CV
|
0 |
25 days ago |
| 413 |
RoboTALES: Learning Reasoning-Guided Robot Policies via Task-Aligned Simulated Futures
Hanan Gani, Tejal Kulkarni, ... (+3 more)
|
|
cs.RO
|
0 |
25 days ago |
| 414 |
OBBSeg: Irregular Lesion Segmentation under Oriented Bounding Box Annotations
Jun Wei, Xinchang Liu, ... (+4 more)
|
|
cs.CV
|
0 |
25 days ago |
| 415 |
SparseCtrl-HOI: Sparse Temporal Control for Human-Object Interaction Video Generation
Shenbo Xie, Mingrui Cai, ... (+3 more)
|
|
cs.CV
|
0 |
25 days ago |
| 416 |
CMDR: Contextual Multimodal Document Retrieval
Ryota Tanaka, Taku Hasegawa, Kyosuke Nishida
|
|
cs.IR
|
0 |
25 days ago |
| 417 |
D2PO: Optimizing Diffusion Samplers via Dynamic Preference
Jinkyu Kim, Jinyoung Choi, Bohyung Han
|
|
cs.LG
|
0 |
25 days ago |
| 418 |
ARMS: Anchor-Relational Motion Streaming for Seamless Solo-Social Motion Transitions
Huakun Liu, Qing Yu, ... (+3 more)
|
|
cs.CV
|
0 |
25 days ago |
| 419 |
Ground3D-LMM: Fine-Grained 3D Point Grounding and Spatial Reasoning with LMM
Amol Harsh, Zongyan Han, ... (+6 more)
|
|
cs.CV
|
0 |
26 days ago |
| 420 |
From Machine Learning to Large-Scale EO Products: Best Practices for Making Maps
Ghjulia Sialelli, Robin Young, ... (+10 more)
|
|
cs.LG
|
0 |
5 days ago |
| 421 |
Hybrid Advantage Estimation with Unified Critic for VLM Agentic Reinforcement Learning
Wenxuan Zhang, Yuhui Wang, ... (+6 more)
|
|
cs.AI
|
0 |
6 days ago |
| 422 |
UltraViT: Latency-Optimized On-device Vision Encoder for Large Vision-Language Models
Ioannis Maniadis Metaxas, Adrian Bulat, ... (+5 more)
|
|
cs.CV
|
0 |
7 days ago |
| 423 |
Stabilizing Deep Reconstruction Operators with Contractive Anchoring
Arghya Sinha, Trishit Mukherjee, Kunal N. Chaudhury
|
|
eess.IV
|
0 |
7 days ago |
| 424 |
Structured Redundancy Modeling for Efficient Visual Token Pruning in High-Resolution MLLMs
Jouwon Song, Woohyeong Kim, Kyeongbo Kong
|
|
cs.CV
|
0 |
7 days ago |
| 425 |
Spectral-Aware Analytic Class-Incremental Learning for Long-Tailed Distributions
Quyen Tran, Hai Nguyen, ... (+5 more)
|
|
cs.LG
|
0 |
8 days ago |
| 426 |
Layering Virtual Try-On
Chun Feng, Bowei Chen, ... (+2 more)
|
|
cs.CV
|
0 |
8 days ago |
| 427 |
Controlling Embedding Spaces with Text-Conditioned Transformations
Joseph Fioresi, Fabian Caba Heilbron, ... (+3 more)
|
|
cs.CV
|
0 |
8 days ago |
| 428 |
Deformable Triangle Splatting: Flexible Primitives for Real-Time Radiance Field Rendering
Oriol Jiménez-Ayguadé, Antonio Agudo
|
|
cs.CV
|
0 |
8 days ago |
| 429 |
SiPhy: Single-Image Physical Property Reasoning
Hoang Le, Joonwoo Kwon, ... (+3 more)
|
|
cs.CV
|
0 |
8 days ago |
| 430 |
InnoText: A Unified Model for Visual Text Generation and Editing
Haowei Liu, Runze He, ... (+11 more)
|
|
cs.CV
|
0 |
8 days ago |
| 431 |
Spectral Prior for Reducing Exposure Bias in Diffusion Models
Yuya Kobayashi, Masato Ishii, ... (+3 more)
|
|
cs.CV
|
0 |
8 days ago |
| 432 |
FAIR: Feature-Augmented Implicit Regularization for AI-generated Fake Image Detection
Md Redwanul Haque, Manzur Murshed, ... (+2 more)
|
|
cs.CV
|
0 |
8 days ago |
| 433 |
3D-Aware VLMs with Implicit and Explicit Geometries
Wenhao Li, Xueying Jiang, ... (+5 more)
|
|
cs.CV
|
0 |
9 days ago |
| 434 |
Unified Video Dense Prediction from Disjoint Data
Yihong Sun, Seoung Wug Oh, ... (+3 more)
|
|
cs.CV
|
0 |
9 days ago |
| 435 |
Recurrent Sinusoidal INRs for Efficient High-Fidelity Representation
Hyunmin Cho, Jaejun Yoo, Kyong Hwan Jin
|
|
cs.CV
|
0 |
9 days ago |
| 436 |
DINOde: Continuous Vision-Text Alignment for Open-Vocabulary Semantic Segmentation
Sung-Hoon Yoon, Hoyong Kwon, ... (+2 more)
|
|
cs.CV
|
0 |
9 days ago |
| 437 |
CRAG-MM-Diagnostics: Enabling Stage-Wise Analysis of Knowledge-Intensive VQA
Hanseok Oh, Parishad BehnamGhader, ... (+5 more)
|
|
cs.CV
|
0 |
9 days ago |
| 438 |
The Second LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results
Xiang Chen, Hao Li, ... (+90 more)
|
|
cs.CV
|
0 |
9 days ago |
| 439 |
Distribution-Alignment Bridge for Uncertainty-Aware Text-to-Video Retrieval
Kyeongmo Chae, Jihoon Lee, Sangtae Ahn
|
|
cs.CV
|
0 |
9 days ago |
| 440 |
Axolotl3D: a Unified Framework for Faithful 3D Shape Completion
Anita Hu, Maria Shugrina
|
|
cs.CV
|
0 |
10 days ago |
| 441 |
Look Less, Think Faster: Joint Token-Compute Adaptation for Multimodal LLMs
Pengcheng Wang, Zhiquan Wang, ... (+6 more)
|
|
cs.CV
|
0 |
10 days ago |
| 442 |
Unified Prediction and Planning via Conflict-Aware Disjoint Parameter Training
Taewon Seo, Seonae Jeon, ... (+3 more)
|
|
cs.RO
|
0 |
10 days ago |
| 443 |
MTVDiff: Multimodal Conditional Latent Diffusion for Enhanced Thermal-to-Visible Face Translation
Zhiyuan Xia, Haojie Li, ... (+3 more)
|
|
cs.CV
|
0 |
10 days ago |
| 444 |
MoAKE: Toward Unified All-in-One Action Quality Assessment via Mixture of Action Knowledge Experts
Huangbiao Xu, Huanqi Wu, ... (+4 more)
|
|
cs.CV
|
0 |
10 days ago |
| 445 |
Extending a Large View Synthesis Model for Multi-view Panoptic Segmentation
Kwonyoung Ryu, In-Jae Lee, ... (+4 more)
|
|
cs.CV
|
0 |
10 days ago |
| 446 |
Point Ladder Tuning: Parameter-Efficient Hierarchical Adaptation for 3D Point Cloud Understanding
Junlin Chang, Longhao Zou, Rui Li
|
|
cs.CV
|
0 |
11 days ago |
| 447 |
FlexiAvatar: Unified 3D Gaussian Human Avatars Under Arbitrary Body Visibility
Yihalem Yimolal Tiruneh, Muhammad Salman Ali, ... (+6 more)
|
|
cs.CV
|
0 |
11 days ago |
| 448 |
CoGoal3D: Collaborative 3D Object Detection with 3D-Aware Fusion and Refinement
Zhihao Yang, Zhiyu Xiang, ... (+6 more)
|
|
cs.CV
|
0 |
11 days ago |
| 449 |
IMMoE: Incomplete Multi-View Anomaly Detection via Mixture of View Experts Fusion
Lei Hu
|
|
cs.CV
|
0 |
11 days ago |
| 450 |
Mitigating Modality and Language-Style Gaps for Zero-Shot Video Moment Retrieval
Jihyun Lee, Cheol-Ho Cho, ... (+3 more)
|
|
cs.CV
|
0 |
11 days ago |