| 1 |
When AUC 0.998 Is Not Enough: A Candidate Evaluation Protocol for Hidden-State Probes of Indirect Prompt Injection in Multimodal Computer-Use Agents
Yanhang Li, Zhichao Fan, Zexin Zhuang
|
|
cs.LG
|
0 |
2 months ago |
| 2 |
Chains That See, Answers That Don't: A Multi-Aspect Evaluation Recipe for Forced Chain-of-Thought on Video-MME
Zhichao Fan, Yanhang Li, Zexin Zhuang
|
|
cs.CV
|
0 |
2 months ago |
| 3 |
As We May Search
Saber Zerhoudi, Adam Roegiest, ... (+2 more)
|
|
cs.IR
|
0 |
2 months ago |
| 4 |
A Sensitivity-Aware Test Collection for Search Among Personal Information
Jack McKechnie, Graham McDonald, Craig Macdonald
|
|
cs.IR
|
0 |
2 months ago |
| 5 |
AutoRelAnnotator: Calibrated Model Cascades for Cost-Efficient Relevance Evaluation in Sponsored Search
Md Omar Faruk Rokon, Shasvat Desai, ... (+2 more)
|
|
cs.IR
|
0 |
2 months ago |
| 6 |
Unified Multi-Task Relevance Modeling for E-Commerce: Comparing Task Routing Architectures Across LLMs and Cross-Encoders
Md Omar Faruk Rokon, Jhalak Nilesh Acharya, ... (+3 more)
|
|
cs.IR
|
0 |
2 months ago |
| 7 |
Scaling Dense Retrieval with LLM-Annotated Training Data: Structured Mining and Progressive Curriculum for E-Commerce Sponsored Search
Md Omar Faruk Rokon, Shasvat Desai, ... (+10 more)
|
|
cs.IR
|
0 |
2 months ago |
| 8 |
Next-Gen Sponsored Search: Crafting the Perfect Query with Inventory-Aware RAG (InvAwr-RAG) Based GenAI
Md Omar Faruk Rokon, Weizhi Du, ... (+2 more)
|
|
cs.IR
|
0 |
2 months ago |
| 9 |
Trie-based Experiment Plans for Efficient IR Pipeline Experiments
Irene Anu, Craig Macdonald
|
|
cs.IR
|
0 |
2 months ago |
| 10 |
Improving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Merging
Ahmed Rayane Kebir, Jose G. Moreno, Lynda Tamine
|
|
cs.IR
|
0 |
2 months ago |
| 11 |
When and How to Ask: Dynamic Preference Elicitation Strategies for Conversational Recommendation
Feng Xia, Shuo Zhang, Xi Wang
|
|
cs.IR
|
0 |
2 months ago |
| 12 |
Faithful or Findable? Evaluating LLM-Generated Metadata for RDF Dataset Search
Riccardo Terrenzi, Serkan Ayvaz
|
|
cs.IR
|
0 |
2 months ago |
| 13 |
Towards a Relevance Posterior in Neural Information Access
Andrew Parry, Emmanouil Georgios Lionis, ... (+2 more)
|
|
cs.IR
|
0 |
1 month ago |
| 14 |
Silent Failures in Multimodal Agentic Search:A Diagnostic Taxonomy and Cross-Judge Evaluation
Zhengxian Wu, Junjie Gao, Kai Yang
|
|
cs.AI
|
0 |
1 month ago |
| 15 |
An Epistemic Position-Based Click Model: From Interactions to Epistemic Distributions of Relevance and Bias
Oscar Rolando Ramirez Milian, Harrie Oosterhuis
|
|
cs.IR
|
0 |
1 month ago |
| 16 |
PLAID-PRF: Pseudo-Relevance Feedback with Centroid-like Tokens in PLAID
Xiao Wang, Sean MacAvaney, Craig Macdonald
|
|
cs.IR
|
0 |
1 month ago |
| 17 |
The Matryoshka Hypencoder
Majd Alkawaas, Sean MacAvaney
|
|
cs.IR
|
0 |
1 month ago |
| 18 |
Don't Contrast the Impossible: Region-Constrained Batching for Contrastive User Modeling on a Local Community Platform
Seungho Han, Byeongchang Kim, Jin Yu
|
|
cs.IR
|
0 |
1 month ago |
| 19 |
FiRE: Enhancing MLLMs with Fine-Grained Context Learning for Complex Image Retrieval
Bohan Hou, Haoqiang Lin, ... (+5 more)
|
|
cs.CV
|
0 |
1 month ago |
| 20 |
From Single- to Cross-Document: Benchmarking Multi-Granularity Event Analysis of Large Language Models
Tao Wen, Shuai Shao, ... (+8 more)
|
|
cs.CL
|
0 |
1 month ago |
| 21 |
LLM within MCP Matters: Measuring Inefficient Resource Utilization Driven by LLMs
Minhan Cho, Soyoung Park, ... (+4 more)
|
|
cs.AI
|
0 |
1 month ago |
| 22 |
RecPFN: Prior-Fitted Networks for In-Context-Based Recommendations
En Zhi Tan, Jia Xiang Lim, ... (+3 more)
|
|
cs.LG
|
0 |
26 days ago |