| 101 |
Faithfulness Is Not Free: Auditing Offline KV-Cache Quantization in Retrieval-Augmented Generation
Atta Ul Asad, Ahsan Bilal, ... (+3 more)
|
|
cs.CL
|
0 |
15 days ago |
| 102 |
Stick to What You Know: A Study of Knowledge-Aligned Supervised Fine-Tuning
Arthur Becker, Jakob Kemmler, ... (+4 more)
|
|
cs.CL
|
0 |
15 days ago |
| 103 |
MULTI3IR: A Benchmark for Multi-perspective Multi-domain Multi-modal Information Retrieval
Seokwon Song, Sohyeon Kim, Gunhee Kim
|
|
cs.IR
|
0 |
15 days ago |
| 104 |
Detecting AI Impostors: How Do Middle Schoolers Identify LLM Agents in a Live Collaborative Setting?
Dan Schumacher, Pragathi Durga Rajarajan, ... (+7 more)
|
|
cs.CL
|
0 |
15 days ago |
| 105 |
Stride-k Subsampling: Train-Free Audio Token Reduction for Whisper
Chanhee Cho, Junhyuk Choi, Bugeun Kim
|
|
cs.SD
|
0 |
15 days ago |
| 106 |
TRIPPULSE: Multi-Agent Travel Planning with Review-Grounded Reasoning
Priyanshu Karmakar, Borru Vijay Sai, ... (+4 more)
|
|
cs.CL
|
0 |
15 days ago |
| 107 |
CARVE: Verified Expansion for Variable-Length Generation in Diffusion Language Models
Wail Bouhedja, Amr Mohamed, Guokan Shang
|
|
cs.AI
|
0 |
15 days ago |
| 108 |
S3C-LLM: Skill-Code Guided Agentic Language Models for Spectrum-to-Structure Elucidation
Xuanle Zhao, Xinyuan Cai, ... (+2 more)
|
|
cs.LG
|
0 |
15 days ago |
| 109 |
MMDS-Bench: Benchmarking Multimodal Large Language Models on Dynamic Stance in Social Media Interactions
Yuzhe Ding, Kang He, ... (+6 more)
|
|
cs.CL
|
0 |
15 days ago |
| 110 |
Evaluating and Mitigating Anti-LGBTQ Biases in German and Multilingual Language Models
Melina Morch, Daniel Braun
|
|
cs.CL
|
0 |
15 days ago |
| 111 |
Personas Differ from Native-Language Generation: Language Pathways Shape LLM Interpersonal Advice
Jinhee Won, Xinlan Emily Hu
|
|
cs.CL
|
0 |
15 days ago |
| 112 |
Beyond Good Intentions: When Does the Framing of Multilingual and Low-Resource NLP Research Become a Caricature?
Nedjma Ousidhoum, Noopur Zambare, Mohamed Abdalla
|
|
cs.CL
|
0 |
15 days ago |
| 113 |
You Shouldn't Have Asked: A Pragmatics-Inspired Taxonomy for Evaluating LLM Refusals
Ruoxuan Li, Pinqiao Wang, ... (+2 more)
|
|
cs.CL
|
0 |
15 days ago |
| 114 |
Linguistic Distance Segregates Latent Representations in Automatic Speech Recognition Systems
Ting-Hui Cheng, Line Katrine Harder Clemmensen, Sneha Das
|
|
cs.CL
|
0 |
15 days ago |
| 115 |
HSRM: Hidden-State Reward Models for Test-Time Verification
Xianzhi Li, Xiaodan Zhu
|
|
cs.AI
|
0 |
15 days ago |
| 116 |
CLIN: an Objective Framework for Evaluating Creativity in Short Persian Literary Text
Mohammad Reza Modarres, Armin Tourajmehr, ... (+2 more)
|
|
cs.CL
|
0 |
15 days ago |
| 117 |
Do VLMs Share Safety Neurons Across Modalities?
Jiaxuan Li, Jiahao Zhang, ... (+4 more)
|
|
cs.LG
|
0 |
15 days ago |
| 118 |
The Fragility of Jailbreak Robustness Across Operational States
Yuna Park, Hwang Youn Kim, ... (+4 more)
|
|
cs.CR
|
0 |
15 days ago |
| 119 |
Not All Fallbacks Are Failures: Understanding and Recovering from Fallbacks in Mobile Voice Assistants
Phillip Schneider, Alexandre Mercier, ... (+3 more)
|
|
cs.CL
|
0 |
15 days ago |
| 120 |
SocialReasonBench: A Video-QA Benchmark for Social Reasoning with Counterfactual Narrative Videos
Zheyu Huang, Zijing Shi, ... (+5 more)
|
|
cs.CL
|
0 |
15 days ago |
| 121 |
GUIDE: Guiding Internal Evidence with Language Instructions
Soyeon Caren Han, Hyunsuk Chung, ... (+3 more)
|
|
cs.CL
|
0 |
15 days ago |
| 122 |
Beyond the Payload: How User Invocation Shapes Coding Agent Vulnerability to Repository Poisoning
Fukang Zhu, Binbin Zhao, ... (+4 more)
|
|
cs.CR
|
0 |
15 days ago |
| 123 |
WildSEEK: Evaluating Language Models for Information-Seeking
Tanise Ceron, Joachim Baumann, ... (+4 more)
|
|
cs.CL
|
0 |
15 days ago |
| 124 |
OCR-MetaReasoning Benchmark: Evaluating the Meta-Reasoning Ability of MLLMs in Text-Rich Image Understanding
Gengxu Li, Yuan Wu, Yi Chang
|
|
cs.CL
|
0 |
15 days ago |
| 125 |
MURANO: Design, Run, and Reproduce Mechanistic Interpretability Experiments as Composable Pipelines
Alireza Bayat Makou, Emirhan Böge, ... (+6 more)
|
|
cs.CL
|
0 |
15 days ago |
| 126 |
SwarmBench: Can Large Language Models Act as Agent Swarm Orchestrators?
Jinshan Gao, Zhuoran Jin, ... (+3 more)
|
|
cs.CL
|
0 |
15 days ago |
| 127 |
Where Identity Lives: Localized, Retain-Free Identity Unlearning in Multimodal Large Language Models
Kangwook Ko, Jaehyuk Jang, ... (+3 more)
|
|
cs.CL
|
0 |
15 days ago |
| 128 |
BiG-SURE - Bipartite Graph for Semantic Uncertainty and Reliability Estimation of LLMs
Debarpan Bhattacharya, Malay Phadke, Sriram Ganapathy
|
|
cs.CL
|
0 |
15 days ago |
| 129 |
Cost-efficient Active Learning for Referring Image Segmentation and Grounding
Junbeom Hong, Seonghoon Yu, ... (+3 more)
|
|
cs.CV
|
0 |
15 days ago |
| 130 |
Hidden Threat in Synthetic Data: Covert Targeted Bias Injection through Benign Text
Minkyung Cho, Jihyo Kim, ... (+5 more)
|
|
cs.CL
|
0 |
15 days ago |
| 131 |
TaxCE : A Framework for Automated Taxonomy Construction and Evaluation at Scale
Sandeep Sricharan Mukku, Albert Aristotle Nanda, Rohit Pyati
|
|
cs.CL
|
0 |
15 days ago |
| 132 |
Reading the News: Adapting Large Language Models to Swedish Journalism Through Continued Pre-Training
Lukas Borggren, Jenny Kunz, Marco Kuhlmann
|
|
cs.CL
|
0 |
15 days ago |
| 133 |
PLC-DPO: Posterior Label Correction in Noisy and Ambiguous Preference Optimization
Boryeong Cho, Sumyeong Ahn, Se-Young Yun
|
|
cs.LG
|
0 |
15 days ago |
| 134 |
Q-Strata: Hierarchical Bit Allocation for Mixed-Precision Quantization of Mixture-of-Experts LLMs
Deokjae Lee, Sihun Chu, Hyun Oh Song
|
|
cs.LG
|
0 |
15 days ago |
| 135 |
Modality Disentangled Learning for Incomplete Multimodal Emotion Recognition: A Primitive Memory Distillation Perspective
Jiaqi Zhang, Zheng Pang, ... (+10 more)
|
|
cs.CV
|
0 |
15 days ago |
| 136 |
AdaPath: Query-Adaptive Path-Finding via Path-Bank for Multi-Hop Implicit Biomedical KGQA
Jun Hyeong Kim, Dongki Kim, ... (+2 more)
|
|
cs.AI
|
0 |
15 days ago |
| 137 |
Preference Shapes Relevance: Cross-component Hierarchical Semantic Alignment for Personalized Generative Retrieval
Gaoming Zhang, Angqing Jiang, ... (+5 more)
|
|
cs.IR
|
0 |
15 days ago |
| 138 |
Seeing the Unseen: Visual Similarity for Pixel Language Model Adaptation
Ran Zhang, Miryam de Lhoneux, Wessel Poelman
|
|
cs.CL
|
0 |
15 days ago |
| 139 |
PAC: Progress-Augmented Advantage Curriculum for Multi-Task Reinforcement Learning of LLMs
Yuanqiang Yu, Yanzhao Zheng, ... (+9 more)
|
|
cs.LG
|
0 |
15 days ago |
| 140 |
ScienceArena: Benchmarking LLMs on Latest Scientific Olympiad Competitions
Guangxiang Zhao, Qilong Shi, ... (+14 more)
|
|
cs.AI
|
0 |
15 days ago |
| 141 |
Two Centuries of Sexism in British Parliament: A Computational Analysis of Women's Representation in the Hansard Corpus
Mohammad Omar Khursheed, Mandira Sawkar, Ashiqur R. KhudaBukhsh
|
|
cs.CL
|
0 |
15 days ago |
| 142 |
VisER: Visual Evidence and Reliance for Object Hallucination Detection in LVLMs
Afsaneh Hasanebrahimi, Hanxun Huang, ... (+2 more)
|
|
cs.CV
|
0 |
15 days ago |
| 143 |
Enhancing Low-Resource Language Reasoning via High-Resource Language Feature Transfer
Minju Song, Hyeon Hwang, ... (+2 more)
|
|
cs.CL
|
0 |
15 days ago |
| 144 |
Graph Evidence Is Not Enough: Diagnosing Native Decoder Use in Graph-Augmented LLMs
Xiaoyu Guo, Pengcheng Chen, ... (+4 more)
|
|
cs.CL
|
0 |
15 days ago |
| 145 |
EvoSkill Injection: Red-Teaming Autonomous Skill Generation and Evolution in Self-Evolving Agents
Doyun Kim, Chanwoo Kim, ... (+3 more)
|
|
cs.AI
|
0 |
15 days ago |
| 146 |
Generative Models Enhanced by Sequence Labelling and Aspect-Code Switching Improve Cross-lingual Aspect-Based Sentiment Analysis
Jakub Šmíd, Pavel Přibáň, Pavel Král
|
|
cs.CL
|
0 |
15 days ago |
| 147 |
DERELAB: Probing Defeasible Reasoning and Confirmation Bias in LLMs with a Generative Benchmark
Jayanta Sadhu, Sayem Shahad, Kenneth Marino
|
|
cs.AI
|
0 |
15 days ago |
| 148 |
Dense Clinical Contrasts Enhance Medical Knowledge Updating in Large Language Models
Yangmin Huang, Shu Quan, ... (+6 more)
|
|
cs.AI
|
0 |
15 days ago |
| 149 |
Beyond Polarization: The Generative Constraint of Chain-of-Thought in Pointwise Reranking
Xiaoyang Chen, Jie Liu, ... (+6 more)
|
|
cs.CL
|
0 |
15 days ago |
| 150 |
Using Grounded Theory for Agent Behavior Analysis at Scale
Zhuoran Lu, Yangyang Yu, ... (+6 more)
|
|
cs.CL
|
0 |
15 days ago |