Dated Data: Tracing Knowledge Cutoffs in Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Cheng, Jeffrey, Marone, Marc, Weller, Orion, Lawrie, Dawn, Khashabi, Daniel, Van Durme, Benjamin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
"According to ...": Prompting Language Models Improves Quoting from Pre-Training Data
por: Weller, Orion, et al.
Publicado: (2023)
por: Weller, Orion, et al.
Publicado: (2023)
mmBERT: A Modern Multilingual Encoder with Annealed Language Learning
por: Marone, Marc, et al.
Publicado: (2025)
por: Marone, Marc, et al.
Publicado: (2025)
NevIR: Negation in Neural Information Retrieval
por: Weller, Orion, et al.
Publicado: (2023)
por: Weller, Orion, et al.
Publicado: (2023)
Seq vs Seq: An Open Suite of Paired Encoders and Decoders
por: Weller, Orion, et al.
Publicado: (2025)
por: Weller, Orion, et al.
Publicado: (2025)
Verifiable by Design: Aligning Language Models to Quote from Pre-Training Data
por: Zhang, Jingyu, et al.
Publicado: (2024)
por: Zhang, Jingyu, et al.
Publicado: (2024)
Defending Against Disinformation Attacks in Open-Domain Question Answering
por: Weller, Orion, et al.
Publicado: (2022)
por: Weller, Orion, et al.
Publicado: (2022)
Promptriever: Instruction-Trained Retrievers Can Be Prompted Like Language Models
por: Weller, Orion, et al.
Publicado: (2024)
por: Weller, Orion, et al.
Publicado: (2024)
Certified Mitigation of Worst-Case LLM Copyright Infringement
por: Zhang, Jingyu, et al.
Publicado: (2025)
por: Zhang, Jingyu, et al.
Publicado: (2025)
Rank1: Test-Time Compute for Reranking in Information Retrieval
por: Weller, Orion, et al.
Publicado: (2025)
por: Weller, Orion, et al.
Publicado: (2025)
Rank-K: Test-Time Reasoning for Listwise Reranking
por: Yang, Eugene, et al.
Publicado: (2025)
por: Yang, Eugene, et al.
Publicado: (2025)
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
por: Jiang, Dongwei, et al.
Publicado: (2024)
por: Jiang, Dongwei, et al.
Publicado: (2024)
When do Generative Query and Document Expansions Fail? A Comprehensive Study Across Methods, Retrievers, and Datasets
por: Weller, Orion, et al.
Publicado: (2023)
por: Weller, Orion, et al.
Publicado: (2023)
FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions
por: Weller, Orion, et al.
Publicado: (2024)
por: Weller, Orion, et al.
Publicado: (2024)
CLERC: A Dataset for Legal Case Retrieval and Retrieval-Augmented Analysis Generation
por: Hou, Abe Bohan, et al.
Publicado: (2024)
por: Hou, Abe Bohan, et al.
Publicado: (2024)
AdapterSwap: Continuous Training of LLMs with Data Removal and Access-Control Guarantees
por: Fleshman, William, et al.
Publicado: (2024)
por: Fleshman, William, et al.
Publicado: (2024)
Crystal: Characterizing Relative Impact of Scholarly Publications
por: Collison, Hannah, et al.
Publicado: (2026)
por: Collison, Hannah, et al.
Publicado: (2026)
WorldAPIs: The World Is Worth How Many APIs? A Thought Experiment
por: Ou, Jiefu, et al.
Publicado: (2024)
por: Ou, Jiefu, et al.
Publicado: (2024)
Are Finer Citations Always Better? Rethinking Granularity for Attributed Generation
por: Wang, Hexuan, et al.
Publicado: (2026)
por: Wang, Hexuan, et al.
Publicado: (2026)
Compressed Chain of Thought: Efficient Reasoning Through Dense Representations
por: Cheng, Jeffrey, et al.
Publicado: (2024)
por: Cheng, Jeffrey, et al.
Publicado: (2024)
mFollowIR: a Multilingual Benchmark for Instruction Following in Retrieval
por: Weller, Orion, et al.
Publicado: (2025)
por: Weller, Orion, et al.
Publicado: (2025)
arXiv2Table: Toward Realistic Benchmarking and Evaluation for LLM-Based Literature-Review Table Generation
por: Wang, Weiqi, et al.
Publicado: (2025)
por: Wang, Weiqi, et al.
Publicado: (2025)
From Models to Microtheories: Distilling a Model's Topical Knowledge for Grounded Question Answering
por: Weir, Nathaniel, et al.
Publicado: (2024)
por: Weir, Nathaniel, et al.
Publicado: (2024)
RE-Adapt: Reverse Engineered Adaptation of Large Language Models
por: Fleshman, William, et al.
Publicado: (2024)
por: Fleshman, William, et al.
Publicado: (2024)
Is That Your Final Answer? Test-Time Scaling Improves Selective Question Answering
por: Jurayj, William, et al.
Publicado: (2025)
por: Jurayj, William, et al.
Publicado: (2025)
Controllable Safety Alignment: Inference-Time Adaptation to Diverse Safety Requirements
por: Zhang, Jingyu, et al.
Publicado: (2024)
por: Zhang, Jingyu, et al.
Publicado: (2024)
Principled Context Engineering for RAG: Statistical Guarantees via Conformal Prediction
por: Chakraborty, Debashish, et al.
Publicado: (2025)
por: Chakraborty, Debashish, et al.
Publicado: (2025)
LoRA-Augmented Generation (LAG) for Knowledge-Intensive Language Tasks
por: Fleshman, William, et al.
Publicado: (2025)
por: Fleshman, William, et al.
Publicado: (2025)
Linguistic Nepotism: Trading-off Quality for Language Preference in Multilingual RAG
por: Ki, Dayeon, et al.
Publicado: (2025)
por: Ki, Dayeon, et al.
Publicado: (2025)
RORA: Robust Free-Text Rationale Evaluation
por: Jiang, Zhengping, et al.
Publicado: (2024)
por: Jiang, Zhengping, et al.
Publicado: (2024)
SIMPLEMIX: Frustratingly Simple Mixing of Off- and On-policy Data in Language Model Preference Learning
por: Li, Tianjian, et al.
Publicado: (2025)
por: Li, Tianjian, et al.
Publicado: (2025)
Many-Tier Instruction Hierarchy in LLM Agents
por: Zhang, Jingyu, et al.
Publicado: (2026)
por: Zhang, Jingyu, et al.
Publicado: (2026)
Language Models and Logic Programs for Trustworthy Tax Reasoning
por: Jurayj, William, et al.
Publicado: (2025)
por: Jurayj, William, et al.
Publicado: (2025)
Beyond the Reported Cutoff: Where Large Language Models Fall Short on Financial Knowledge
por: Shah, Agam, et al.
Publicado: (2025)
por: Shah, Agam, et al.
Publicado: (2025)
BLT: Can Large Language Models Handle Basic Legal Text?
por: Blair-Stanek, Andrew, et al.
Publicado: (2023)
por: Blair-Stanek, Andrew, et al.
Publicado: (2023)
Enhancing Systematic Decompositional Natural Language Inference Using Informal Logic
por: Weir, Nathaniel, et al.
Publicado: (2024)
por: Weir, Nathaniel, et al.
Publicado: (2024)
RATIONALYST: Mining Implicit Rationales for Process Supervision of Reasoning
por: Jiang, Dongwei, et al.
Publicado: (2024)
por: Jiang, Dongwei, et al.
Publicado: (2024)
Core: Robust Factual Precision with Informative Sub-Claim Identification
por: Jiang, Zhengping, et al.
Publicado: (2024)
por: Jiang, Zhengping, et al.
Publicado: (2024)
Compactor: Calibrated Query-Agnostic KV Cache Compression with Approximate Leverage Scores
por: Chari, Vivek, et al.
Publicado: (2025)
por: Chari, Vivek, et al.
Publicado: (2025)
LLMs Provide Unstable Answers to Legal Questions
por: Blair-Stanek, Andrew, et al.
Publicado: (2025)
por: Blair-Stanek, Andrew, et al.
Publicado: (2025)
Tracing Pharmacological Knowledge In Large Language Models
por: Khwaja, Basil Hasan, et al.
Publicado: (2026)
por: Khwaja, Basil Hasan, et al.
Publicado: (2026)
Ejemplares similares
-
"According to ...": Prompting Language Models Improves Quoting from Pre-Training Data
por: Weller, Orion, et al.
Publicado: (2023) -
mmBERT: A Modern Multilingual Encoder with Annealed Language Learning
por: Marone, Marc, et al.
Publicado: (2025) -
NevIR: Negation in Neural Information Retrieval
por: Weller, Orion, et al.
Publicado: (2023) -
Seq vs Seq: An Open Suite of Paired Encoders and Decoders
por: Weller, Orion, et al.
Publicado: (2025) -
Verifiable by Design: Aligning Language Models to Quote from Pre-Training Data
por: Zhang, Jingyu, et al.
Publicado: (2024)