| 101 |
$f$-Trajectory Balance: A Loss Family for Tuning GFlowNets, Generative Models, and LLMs with Off- and On-Policy Data
Jake Fawkes, Jason Hartford
|
|
cs.LG
|
0 |
2 months ago |
| 102 |
Margin-Adaptive Confidence Ranking for Reliable LLM Judgement
Gaojie Jin, Yong Tao, ... (+2 more)
|
|
cs.LG
|
0 |
2 months ago |
| 103 |
Deep Pre-Alignment for VLMs
Tianyu Yu, Kechen Fang, ... (+6 more)
|
|
cs.CV
|
0 |
2 months ago |
| 104 |
Evidential Reasoning Advances Interpretable Real-World Disease Screening
Chenyu Lian, Hong-Yu Zhou, Jing Qin
|
|
cs.CV
|
0 |
2 months ago |
| 105 |
ML-Embed: Inclusive and Efficient Embeddings for a Multilingual World
Ziyin Zhang, Zihan Liao, ... (+3 more)
|
|
cs.CL
|
0 |
2 months ago |
| 106 |
Position: Ideas Should be the Center of Machine Learning Research
Jairo Diaz-Rodriguez
|
|
cs.LG
|
0 |
2 months ago |
| 107 |
Hierarchical Image Tokenization for Multi-Scale Image Super Resolution
Isma Hadji, Enrique Sanchez, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 108 |
Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining
Weimin Xiong, Shuhao Gu, ... (+6 more)
|
|
cs.CL
|
0 |
2 months ago |
| 109 |
Conditional KRR: Injecting Unpenalized Features into Kernel Methods with Applications to Kernel Thresholding
Rustem Takhanov, Zhenisbek Assylbekov
|
|
cs.LG
|
0 |
2 months ago |
| 110 |
Hidden in Plain Tokens: Simply Robust, Gradient-Free Watermark for Synthetic Audio
Georgios Milis, Yubin Qin, ... (+2 more)
|
|
cs.LG
|
0 |
2 months ago |
| 111 |
Where Concept Erasure Should Occur: Concept-Layer Alignment in Text-to-Video Diffusion Models
Yiwei Xie, Ping Liu, Zheng Zhang
|
|
cs.CV
|
0 |
2 months ago |
| 112 |
Latent Representation Alignment for Offline Goal-Conditioned Reinforcement Learning
Hyungkyu Kang, Byeongchan Kim, Min-hwan Oh
|
|
cs.LG
|
0 |
2 months ago |
| 113 |
Learning to Search and Searching to Learn for Generalization in Planning
Michael Aichmüller, Yannik Hesse, Hector Geffner
|
|
cs.AI
|
0 |
2 months ago |
| 114 |
AgentHijack: Benchmarking Computer Use Agent Robustness to Common Environment Corruptions
Jingwei Sun, Jianing Zhu, ... (+4 more)
|
|
cs.AI
|
0 |
2 months ago |
| 115 |
GCIB: Graph Contrastive Information Bottleneck for Multi-Behavior Recommendation
Likang Wu, Zihao Chen, ... (+5 more)
|
|
cs.IR
|
0 |
2 months ago |
| 116 |
Courtroom Analogy: New Perspective on Uncertainty-Aware Classification
Taeseong Yoon, Heeyoung Kim
|
|
cs.LG
|
0 |
2 months ago |
| 117 |
Optimal Design for Multinomial Logit Model with Applications to Best Assortment Identification
Joongkyu Lee, Min-hwan Oh
|
|
stat.ML
|
0 |
2 months ago |
| 118 |
Rao-Blackwellized Score Matching on Manifolds
Divit Rawal
|
|
stat.ML
|
0 |
2 months ago |
| 119 |
Are We Overconfident in Models and Results for Semi-Supervised 3D Medical Image Segmentation?
Jun Li, Ziwei Qin
|
|
cs.CV
|
0 |
2 months ago |
| 120 |
TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation
Sizhe Zhao, Shengping Zhang, ... (+4 more)
|
|
cs.RO
|
0 |
2 months ago |
| 121 |
MetaphorVU: Towards Metaphorical Video Understanding
Zhuoqun Li, Boxi Cao, ... (+14 more)
|
|
cs.CV
|
0 |
2 months ago |
| 122 |
Mean-Shift PCA by Knockoff Mean
Mengda Li, Zeng Li, Jianfeng Yao
|
|
stat.ML
|
0 |
2 months ago |
| 123 |
Rethinking Feature Alignment in Generalist Graph Anomaly Detection: A Relational Fingerprint-based Approach
Yujing Liu, Yixin Liu, ... (+4 more)
|
|
cs.LG
|
0 |
2 months ago |
| 124 |
Learning to Route Languages for Multilingual Policy Optimization
Geyang Guo, Hiromi Wakaki, ... (+3 more)
|
|
cs.CL
|
0 |
2 months ago |
| 125 |
DIVA: Harnessing the Representation Divergence in Unified Multimodal Models for Mutual Reinforcement
Renjie Lu, Xulong Zhang, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 126 |
Whose Alignment? Comparing LLM Process Alignment Across Diverse Organizational Decision Contexts
Niklas Weller, Emilio Barkett
|
|
cs.AI
|
0 |
2 months ago |
| 127 |
Quantifying Empirical Compute-Supervision Tradeoffs in RLVR
Ryo Mitsuhashi, Patrick Chen, ... (+3 more)
|
|
cs.LG
|
0 |
2 months ago |
| 128 |
Inference Time Optimization with Confidence Dynamics
Yu Wang, Minghao Liu, ... (+4 more)
|
|
cs.CL
|
0 |
2 months ago |
| 129 |
On the Epistemic Uncertainty of Overparametrized Neural Networks
David Rügamer
|
|
cs.LG
|
0 |
2 months ago |
| 130 |
Boosting Inference with Guided Reasoning: Stochastic Exploration for Recursive Models
Andrew Corbett, Archit Sood, ... (+3 more)
|
|
cs.AI
|
0 |
2 months ago |
| 131 |
Rejoinder: The ICML 2023 Ranking Experiment: Examining Author Self-Assessment in ML/AI Peer Review
Buxin Su, Jiayao Zhang, ... (+7 more)
|
|
stat.AP
|
0 |
2 months ago |
| 132 |
Theoretical Analysis of Sparse Optimization with Reparameterization, Weight Decay, and Adaptive Learning Rate
Huangyu Xu, Jingqin Yang, ... (+2 more)
|
|
cs.LG
|
0 |
2 months ago |
| 133 |
Language Bias in LVLMs: From In-Depth Analysis to Simple and Effective Mitigation
Yangneng Chen, Jing Li
|
|
cs.CL
|
0 |
2 months ago |
| 134 |
Mitigating Gradient Pathology in PINNs through Aligned Constraint
Yichen Luo, Peiyu Zhu, ... (+6 more)
|
|
cs.LG
|
0 |
2 months ago |
| 135 |
NITP: Next Implicit Token Prediction for LLM Pre-training
Xiangdong Zhang, Debing Zhang, ... (+4 more)
|
|
cs.CL
|
0 |
2 months ago |
| 136 |
Interpretability Transfer from Language to Vision via Sparse Autoencoders
Alexey Kravets, Da Li, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 137 |
Quaternion Self-Attention with Shared Scores
Shogo Yamauchi, Tohru Nitta, Hideaki Tamori
|
|
cs.LG
|
0 |
2 months ago |
| 138 |
MVR-cache: Optimizing Semantic Caching via Multi-Vector Retrieval and Learned Prompt Segmentation
Ali Noshad, Zishan Zheng, Yinjun Wu
|
|
cs.IR
|
0 |
2 months ago |
| 139 |
Efficient DP-SGD for LLMs with Randomized Clipping
Enayat Ullah, Sai Aparna Aketi, ... (+3 more)
|
|
cs.LG
|
0 |
2 months ago |
| 140 |
Clustering as Reasoning: A $k$-Means Interpretation of Chain-of-Thought Graph Learning
Xuanting Xie, Zhaochen Guo, ... (+5 more)
|
|
cs.AI
|
0 |
2 months ago |
| 141 |
Unifying Value Alignment and Assignment in Cross-Domain Offline Reinforcement Learning with Heterogeneous Datasets
Zhongjian Qiao, Jiafei Lyu, ... (+4 more)
|
|
cs.LG
|
0 |
2 months ago |
| 142 |
Geo-Expert: Towards Expert-Level Geological Reasoning via Parameter-Efficient Fine-Tuning
Chenyou Guo, Zongqi Liu, ... (+3 more)
|
|
cs.AI
|
0 |
2 months ago |
| 143 |
AOEPT: Breaking the Implicit Modality-Reduction Bottleneck in Modality-Missing Prompt Tuning
Jian Lang, Rongpei Hong, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |
| 144 |
The Perception-Physics Paradox: Probing Scientific Alignment with TC-Bench
Dingling Yao, Andrea Polesello, ... (+3 more)
|
|
cs.LG
|
0 |
2 months ago |
| 145 |
Hermite-NGP: Gradient-Augmented Hash Encoding for Learning PDEs
Jinjin He, Zhiqi Li, ... (+2 more)
|
|
cs.LG
|
0 |
2 months ago |
| 146 |
Reinforcement Learning for Reachability: Guaranteeing Asymptotic Optimality
Amogh Palasamudram, Jakub Svoboda, ... (+2 more)
|
|
cs.LG
|
0 |
2 months ago |
| 147 |
HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing
Ruyi Chen, Lu Zhou, ... (+4 more)
|
|
cs.CV
|
0 |
2 months ago |
| 148 |
Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion
Sol Park, Soobin Um
|
|
cs.LG
|
0 |
2 months ago |
| 149 |
Correcting Visual Blur Induced by Attention Distraction to Reduce Hallucinations: Algorithm and Theory
Quanjiang Li, Zhiming Liu, ... (+3 more)
|
|
cs.CV
|
0 |
2 months ago |
| 150 |
IQA-Spider: Unifying Multi-Granularity Image Quality Assessment with Reasoning, Grounding and Referring
Xinge Peng, Yiting Lu, ... (+2 more)
|
|
cs.CV
|
0 |
2 months ago |