Enregistré dans:
| Auteurs principaux: | Zou, Xinrui, Zhang, Ming, Weir, Nathaniel, Van Durme, Benjamin, Holzenberger, Nils |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2401.06715 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Language Models and Logic Programs for Trustworthy Tax Reasoning
par: Jurayj, William, et autres
Publié: (2025)
par: Jurayj, William, et autres
Publié: (2025)
TV-TREES: Multimodal Entailment Trees for Neuro-Symbolic Video Reasoning
par: Sanders, Kate, et autres
Publié: (2024)
par: Sanders, Kate, et autres
Publié: (2024)
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
par: Blair-Stanek, Andrew, et autres
Publié: (2023)
par: Blair-Stanek, Andrew, et autres
Publié: (2023)
Can LLMs Identify Tax Abuse?
par: Blair-Stanek, Andrew, et autres
Publié: (2025)
par: Blair-Stanek, Andrew, et autres
Publié: (2025)
BLT: Can Large Language Models Handle Basic Legal Text?
par: Blair-Stanek, Andrew, et autres
Publié: (2023)
par: Blair-Stanek, Andrew, et autres
Publié: (2023)
"According to ...": Prompting Language Models Improves Quoting from Pre-Training Data
par: Weller, Orion, et autres
Publié: (2023)
par: Weller, Orion, et autres
Publié: (2023)
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
par: Jiang, Dongwei, et autres
Publié: (2024)
par: Jiang, Dongwei, et autres
Publié: (2024)
NELLIE: A Neuro-Symbolic Inference Engine for Grounded, Compositional, and Explainable Reasoning
par: Weir, Nathaniel, et autres
Publié: (2022)
par: Weir, Nathaniel, et autres
Publié: (2022)
Compactor: Calibrated Query-Agnostic KV Cache Compression with Approximate Leverage Scores
par: Chari, Vivek, et autres
Publié: (2025)
par: Chari, Vivek, et autres
Publié: (2025)
LELA: An End-to-end LLM-based Entity Linking Framework with Zero-shot Domain Adaptation
par: Haffoudhi, Samy, et autres
Publié: (2026)
par: Haffoudhi, Samy, et autres
Publié: (2026)
Learning to Reason via Program Generation, Emulation, and Search
par: Weir, Nathaniel, et autres
Publié: (2024)
par: Weir, Nathaniel, et autres
Publié: (2024)
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning
par: Singh, Vikash, et autres
Publié: (2026)
par: Singh, Vikash, et autres
Publié: (2026)
Enhancing Systematic Decompositional Natural Language Inference Using Informal Logic
par: Weir, Nathaniel, et autres
Publié: (2024)
par: Weir, Nathaniel, et autres
Publié: (2024)
Generating Data-Driven Reasoning Rubrics for Domain-Adaptive Reward Modeling
par: Sanders, Kate, et autres
Publié: (2026)
par: Sanders, Kate, et autres
Publié: (2026)
RE-Adapt: Reverse Engineered Adaptation of Large Language Models
par: Fleshman, William, et autres
Publié: (2024)
par: Fleshman, William, et autres
Publié: (2024)
SEQR: Secure and Efficient QR-based LoRA Routing
par: Fleshman, William, et autres
Publié: (2025)
par: Fleshman, William, et autres
Publié: (2025)
LoRA-Augmented Generation (LAG) for Knowledge-Intensive Language Tasks
par: Fleshman, William, et autres
Publié: (2025)
par: Fleshman, William, et autres
Publié: (2025)
SpectR: Dynamically Composing LM Experts with Spectral Routing
par: Fleshman, William, et autres
Publié: (2025)
par: Fleshman, William, et autres
Publié: (2025)
Bonsai: Interpretable Tree-Adaptive Grounded Reasoning
par: Sanders, Kate, et autres
Publié: (2025)
par: Sanders, Kate, et autres
Publié: (2025)
LM Agents for Coordinating Multi-User Information Gathering
par: Jhamtani, Harsh, et autres
Publié: (2025)
par: Jhamtani, Harsh, et autres
Publié: (2025)
KV-Distill: Nearly Lossless Learnable Context Compression for LLMs
par: Chari, Vivek, et autres
Publié: (2025)
par: Chari, Vivek, et autres
Publié: (2025)
DeonticBench: A Benchmark for Reasoning over Rules
par: Dou, Guangyao, et autres
Publié: (2026)
par: Dou, Guangyao, et autres
Publié: (2026)
RE-AdaptIR: Improving Information Retrieval through Reverse Engineered Adaptation
par: Fleshman, William, et autres
Publié: (2024)
par: Fleshman, William, et autres
Publié: (2024)
RATIONALYST: Mining Implicit Rationales for Process Supervision of Reasoning
par: Jiang, Dongwei, et autres
Publié: (2024)
par: Jiang, Dongwei, et autres
Publié: (2024)
Reasoners or Translators? Contamination-aware Evaluation and Neuro-Symbolic Robustness in Tax Law
par: Kordjamshidi, Parisa, et autres
Publié: (2026)
par: Kordjamshidi, Parisa, et autres
Publié: (2026)
Crystal: Characterizing Relative Impact of Scholarly Publications
par: Collison, Hannah, et autres
Publié: (2026)
par: Collison, Hannah, et autres
Publié: (2026)
Controllable Safety Alignment: Inference-Time Adaptation to Diverse Safety Requirements
par: Zhang, Jingyu, et autres
Publié: (2024)
par: Zhang, Jingyu, et autres
Publié: (2024)
Gaps or Hallucinations? Gazing into Machine-Generated Legal Analysis for Fine-grained Text Evaluations
par: Hou, Abe Bohan, et autres
Publié: (2024)
par: Hou, Abe Bohan, et autres
Publié: (2024)
Retrieval-Constrained Decoding Reveals Underestimated Parametric Knowledge in Language Models
par: Hamdani, Rajaa El, et autres
Publié: (2025)
par: Hamdani, Rajaa El, et autres
Publié: (2025)
AdapterSwap: Continuous Training of LLMs with Data Removal and Access-Control Guarantees
par: Fleshman, William, et autres
Publié: (2024)
par: Fleshman, William, et autres
Publié: (2024)
Always Tell Me The Odds: Fine-grained Conditional Probability Estimation
par: Wang, Liaoyaqi, et autres
Publié: (2025)
par: Wang, Liaoyaqi, et autres
Publié: (2025)
Many-Tier Instruction Hierarchy in LLM Agents
par: Zhang, Jingyu, et autres
Publié: (2026)
par: Zhang, Jingyu, et autres
Publié: (2026)
Are Machines Better at Complex Reasoning? Unveiling Human-Machine Inference Gaps in Entailment Verification
par: Sanyal, Soumya, et autres
Publié: (2024)
par: Sanyal, Soumya, et autres
Publié: (2024)
Defending Against Disinformation Attacks in Open-Domain Question Answering
par: Weller, Orion, et autres
Publié: (2022)
par: Weller, Orion, et autres
Publié: (2022)
The Factuality of Large Language Models in the Legal Domain
par: Hamdani, Rajaa El, et autres
Publié: (2024)
par: Hamdani, Rajaa El, et autres
Publié: (2024)
Do Androids Know They're Only Dreaming of Electric Sheep?
par: CH-Wang, Sky, et autres
Publié: (2023)
par: CH-Wang, Sky, et autres
Publié: (2023)
Sample-Efficient Online Learning in LM Agents via Hindsight Trajectory Rewriting
par: Hu, Michael Y., et autres
Publié: (2025)
par: Hu, Michael Y., et autres
Publié: (2025)
Unlocking LLM Creativity in Science through Analogical Reasoning
par: Shen, Andrew, et autres
Publié: (2026)
par: Shen, Andrew, et autres
Publié: (2026)
Interpreting User Requests in the Context of Natural Language Standing Instructions
par: Moghe, Nikita, et autres
Publié: (2023)
par: Moghe, Nikita, et autres
Publié: (2023)
Adversarial Attacks and Defense for Conversation Entailment Task
par: Yang, Zhenning, et autres
Publié: (2024)
par: Yang, Zhenning, et autres
Publié: (2024)
Documents similaires
-
Language Models and Logic Programs for Trustworthy Tax Reasoning
par: Jurayj, William, et autres
Publié: (2025) -
TV-TREES: Multimodal Entailment Trees for Neuro-Symbolic Video Reasoning
par: Sanders, Kate, et autres
Publié: (2024) -
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
par: Blair-Stanek, Andrew, et autres
Publié: (2023) -
Can LLMs Identify Tax Abuse?
par: Blair-Stanek, Andrew, et autres
Publié: (2025) -
BLT: Can Large Language Models Handle Basic Legal Text?
par: Blair-Stanek, Andrew, et autres
Publié: (2023)