| 301 |
OnPoint: Offline-to-Online Multi-Level Distillation for Point-Supervised Online Temporal Action Localization
Sakib Reza, Gauri Jagatap, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 302 |
VOCA: Visual Odometry with Codec Awareness
Nouri Alexander Hilscher, Mateo de Mayo, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 303 |
DriftScope: Measuring The Hidden Effects of Diffusion Model Adaptation
Héctor Laria, Yiping Han, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 304 |
Identifying and Resolving Pitfalls of Knowledge-Based VQA Benchmarks: Auditing, Repairing, and Augmenting
Qian Ma, S M Rayeed, ... (+3 more)
|
|
cs.CL
|
0 |
1 month ago |
| 305 |
Progressive Pose-Guided 4D Animal Reconstruction from Monocular Video
Siyuan Li, Weiying Chen, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 306 |
A Mechanism-Driven Theory of Phase Transitions in Active Learning
Julia Machnio, Mads Nielsen, Mostafa Mehdipour Ghazi
|
|
cs.CV
|
0 |
1 month ago |
| 307 |
Benchmarking Federated Learning and Knowledge Distillation for Point Cloud Classification
Aizierjiang Aiersilan
|
|
cs.GR
|
0 |
1 month ago |
| 308 |
Segmenting, Fast and Slow: Real-Time Open-Vocabulary Video Instance Segmentation with Dual-Path Processing
Luca Barsellotti, Martin Sundermeyer, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 309 |
Lost in the Tail: Addressing Geographic Imbalance in Urban Visual Place Recognition
Zhiyao Shu, Jiacheng Yang, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 310 |
FaceMoE: Mixture of Experts for Low-Resolution Face Recognition
Kartik Narayan, Vishal M. Patel
|
|
cs.CV
|
0 |
1 month ago |
| 311 |
Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers
Anh Nguyen, Ngan Nguyen, ... (+12 more)
|
|
cs.CV
|
0 |
1 month ago |
| 312 |
LUNA: Learning Universal 3D Human Animation Beyond Skinning
Peng Li, Rawal Khirodkar, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 313 |
No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs
Haojian Huang, Harold Haodong Chen, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 314 |
DriveWeaver: Point-Conditioned Video Inpainting for Controllable Vehicle Insertion in Autonomous Driving Simulation
Junzhe Jiang, Zipei Ma, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 315 |
RESOLVE: A Multi-Resolution and Multi-Modal Dataset for Roadside Cooperative Perception
Shaozu Ding, Linan Song, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 316 |
Real-Time Source-Free Object Detection
Sairam VCR, Varun Gopal, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 317 |
PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving
Kyuhwan Yeon, Benjamin Ramtoula, Daniele De Martini
|
|
cs.CV
|
0 |
1 month ago |
| 318 |
Generative Lane Topology Reasoning via Autoregressive Model with Geometry Prior
Jiahui Fu, Zehao Huang, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 319 |
NURBS Splatting: A Unified Differentiable Rendering Framework for Vector Graphics
Jingye Qiu, Shizhe Zhou
|
|
cs.GR
|
0 |
1 month ago |
| 320 |
MemLearner: Learning to Query Context memory for Video World Models
Jiwen Yu, Jianxiong Gao, ... (+8 more)
|
|
cs.CV
|
0 |
1 month ago |
| 321 |
Look But Don't Touch with Sparse Autoencoders for Unlearning in Diffusion Models
Enrico Cassano, Riccardo Renzulli, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 322 |
Intrinsically Stable Spiking Neural Networks: Overcoming the Performance Barrier in the Absence of Batch Normalization
Ruichen Ma, Xiaoyang Zhang, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 323 |
Histogram-constrained Image Generation
Haoming Liu, Yuanhe Guo, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 324 |
ShellMaker: Language-Guided Exterior Completion under Structural Constraints
Ruiqi Xu, Daniel Aliaga
|
|
cs.CV
|
0 |
1 month ago |
| 325 |
Sparsity-Inducing Divergence Losses for Biometric Verification
Dimitrios Koutsianos, Ladislav Mošner, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 326 |
Think While You Map: Asynchronous Vision-Language Agents for Incremental 3D Scene Graphs
Deniz Bickici, Michael Pabst, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 327 |
Zero-Shot Quantization for Object Detectors using Off-the-Shelf Generative Models
Hyunho Lee, Kyomin Hwang, ... (+4 more)
|
|
cs.LG
|
0 |
1 month ago |
| 328 |
UniTac: A Unified Multimodal Model for Cross-Sensor Tactile Understanding and Generation
Jiahang Tu, Fengyu Yang, ... (+9 more)
|
|
cs.RO
|
0 |
1 month ago |
| 329 |
Visual Semantic Entropy: Do Vision Language Models Recognize Visual Ambiguity?
Ta Duc Huy, Trang Nguyen, ... (+6 more)
|
|
cs.CV
|
0 |
1 month ago |
| 330 |
One Video, One World: Turning Monocular Video into Physical 4D Scenes
Junhao Chen, Boran Zhang, ... (+8 more)
|
|
cs.CV
|
0 |
1 month ago |
| 331 |
Revisiting Parameter Redundancy in Vision-Language-Action Models: Insights from VLM-to-VLA Adaptation
Fengnian Zhang, Tao Huang, ... (+3 more)
|
|
cs.RO
|
0 |
1 month ago |
| 332 |
Accelerated Likelihood Maximization for Diffusion-based Versatile Content Generation
Hyunsoo Lee, Inwoo Hwang, Young Min Kim
|
|
cs.CV
|
0 |
1 month ago |
| 333 |
Editing Everything Everywhere All at Once
Fabio Quattrini, Carmine Zaccagnino, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 334 |
Learning from Failure: Inference-Time Self-Improvement for Computer-Use Agents
Xueqiao Sun, Xiaohan Wang, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 335 |
UHD-MFF: Shattering Barriers in Multi-Focus Ultra-High-Definition Image Fusion via Learnable Lookup Tables
Yibing Zhang, Xunpeng Yi, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 336 |
A First Exploration of Neuromorphic OT-CFM for Multi-Speaker VSR
Lin Chen, Jingping Fang, ... (+6 more)
|
|
cs.MM
|
0 |
1 month ago |
| 337 |
CooperScene: Multi-Modal Cooperative Autonomy Benchmark with C-V2X Communication Characterization
Bo Wu, Ruoshen Mo, ... (+7 more)
|
|
cs.CV
|
0 |
1 month ago |
| 338 |
Long-term Traffic Simulation via Structured Autoregressive Modeling
Lingyu Xiao, Zexin Feng, Xintao Yan
|
|
cs.AI
|
0 |
1 month ago |
| 339 |
AC3S: Adaptive Conditioning for 3D-Aware Synthetic Data Generation
Eric Ji, Qiran Hu, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 340 |
ExPLoRe: Expert Patch-Level Loss Routing for Multi-Objective Masked Image Modeling
Konstantinos Georgiou, Maofeng Tang, Hairong Qi
|
|
cs.CV
|
0 |
1 month ago |
| 341 |
Learning to Deny: Action Denial in Multimodal Large Language Models
Raiyaan Abdullah, Shehreen Azad, Yogesh Singh Rawat
|
|
cs.CV
|
0 |
1 month ago |
| 342 |
HSDF-Lane: Height-Aligned Signed Distance Field with Semantic Lane Prior for 3D Lane Detection
Jiyong Boo, Byeongin Joung, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |
| 343 |
TaxoMIL: Taxonomy-Constrained Learning for Hierarchical Whole Slide Image Analysis
Chaeyeon Lee, Khang Nguyen Quoc, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 344 |
Horizon3D: Sparse Radar-Camera Fusion for Long-Range 3D Perception in Autonomous Driving
Geonho Bang, Geunju Baek, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 345 |
Anchoring on Reality: Breaking the Pseudo-Target Ceiling in Makeup Transfer
Bo Wei, Xianhui Lin, ... (+9 more)
|
|
cs.CV
|
0 |
1 month ago |
| 346 |
Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation
Baiqin Wang, Sen Chen, ... (+4 more)
|
|
cs.CV
|
0 |
1 month ago |
| 347 |
CasaMaestro: Multi-View Panoramas for House-Scale 3D Reconstruction
Yuzhou Ji, Xiaotian Yang, Zhipeng Zhang
|
|
cs.CV
|
0 |
1 month ago |
| 348 |
AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images
Meng Yang, Zizhuo Li, ... (+3 more)
|
|
cs.CV
|
0 |
1 month ago |
| 349 |
Hierarchical 3D Scene Graph Construction and Belief-based Planning for Semantic Navigation
Bing Wu, Zuyao Chen, Changwen Chen
|
|
cs.CV
|
0 |
1 month ago |
| 350 |
Diffusion-Based Material Regularization for Physics-Based Inverse Rendering
Jingwang Ling, Lifan Wu, ... (+2 more)
|
|
cs.CV
|
0 |
1 month ago |