🏛️ The Computation & Language Crypt
cs.CL: Where Computation & Language papers rest without their code.
39803
Total Papers
26069
No Code
1086
Twilight
12648
Has Code
31.8%
Survival Rate
R.I.P.
👻
Ghosted
R.I.P.
👻
Ghosted
Breaking Chains: Unraveling the Links in Multi-Hop Knowledge Unlearning
R.I.P.
👻
Ghosted
The Geometry of Numerical Reasoning: Language Models Compare Numeric Properties in Linear Subspaces
R.I.P.
👻
Ghosted
When Not to Answer: Evaluating Prompts on GPT Models for Effective Abstention in Unanswerable Math Word Problems
R.I.P.
👻
Ghosted
Interpreting token compositionality in LLMs: A robustness analysis
R.I.P.
👻
Ghosted
Identifying Task Groupings for Multi-Task Learning Using Pointwise V-Usable Information
R.I.P.
👻
Ghosted
On the Risk of Evidence Pollution for Malicious Social Text Detection in the Era of LLMs
R.I.P.
👻
Ghosted
How to Make LLMs Forget: On Reversing In-Context Knowledge Edits
R.I.P.
👻
Ghosted
Insights from the Inverse: Reconstructing LLM Training Goals Through Inverse Reinforcement Learning
R.I.P.
👻
Ghosted
Kallini et al. (2024) do not compare impossible languages with constituency-based ones
R.I.P.
👻
Ghosted
On A Scale From 1 to 5: Quantifying Hallucination in Faithfulness Evaluation
R.I.P.
👻
Ghosted
Exploiting LLMs' Reasoning Capability to Infer Implicit Concepts in Legal Information Retrieval
R.I.P.
👻
Ghosted
Iter-AHMCL: Alleviate Hallucination for Large Language Model via Iterative Model-level Contrastive Learning
R.I.P.
👻
Ghosted
Toolken+: Improving LLM Tool Usage with Reranking and a Reject Option
R.I.P.
👻
Ghosted
Leaving the barn door open for Clever Hans: Simple features predict LLM benchmark answers
R.I.P.
👻
Ghosted
Difficult Task Yes but Simple Task No: Unveiling the Laziness in Multimodal LLMs
R.I.P.
👻
Ghosted
Code-Mixer Ya Nahi: Novel Approaches to Measuring Multilingual LLMs' Code-Mixing Capabilities
R.I.P.
👻
Ghosted
Assessing Bias in Metric Models for LLM Open-Ended Generation Bias Benchmarks
R.I.P.
👻
Ghosted
Beyond Human-Only: Evaluating Human-Machine Collaboration for Collecting High-Quality Translation Data
R.I.P.
👻
Ghosted
Varying Shades of Wrong: Aligning LLMs with Wrong Answers Only
R.I.P.
👻
Ghosted
Assessing the Human Likeness of AI-Generated Counterspeech
R.I.P.
👻
Ghosted
Large Language Models Are Active Critics in NLG Evaluation
R.I.P.
👻
Ghosted
QUITE: Quantifying Uncertainty in Natural Language Text in Bayesian Reasoning Scenarios
R.I.P.
👻
Ghosted