Dodo: Dynamic Contextual Compression for Decoder-only LMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Qin, Guanghui, Rosset, Corby, Chau, Ethan C., Rao, Nikhil, Van Durme, Benjamin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
di: Hashemi, Helia, et al.
Pubblicazione: (2024)
di: Hashemi, Helia, et al.
Pubblicazione: (2024)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025)
di: Fadli, Samih
Pubblicazione: (2025)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
di: Gwak, Jiho, et al.
Pubblicazione: (2025)
di: Gwak, Jiho, et al.
Pubblicazione: (2025)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
di: Liu, Zhongxin, et al.
Pubblicazione: (2025)
di: Liu, Zhongxin, et al.
Pubblicazione: (2025)
On the Influence of Discourse Relations in Persuasive Texts
di: Turk, Nawar, et al.
Pubblicazione: (2025)
di: Turk, Nawar, et al.
Pubblicazione: (2025)
Calibrated Confidence Estimation for Tabular Question Answering
di: Voss, Lukas
Pubblicazione: (2026)
di: Voss, Lukas
Pubblicazione: (2026)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
di: Han, Xudong, et al.
Pubblicazione: (2025)
di: Han, Xudong, et al.
Pubblicazione: (2025)
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
di: Kugler, Kai
Pubblicazione: (2025)
di: Kugler, Kai
Pubblicazione: (2025)
Character-Level Transformer for Tajik-Persian Transliteration with a Parallel Lexical Corpus
di: Arabov, Mullosharaf K.
Pubblicazione: (2026)
di: Arabov, Mullosharaf K.
Pubblicazione: (2026)
Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation
di: Resck, Lucas, et al.
Pubblicazione: (2026)
di: Resck, Lucas, et al.
Pubblicazione: (2026)
Bridging the Gap: An Intermediate Language for Enhanced and Cost-Effective Grapheme-to-Phoneme Conversion with Homographs with Multiple Pronunciations Disambiguation
di: Bertina, Abbas, et al.
Pubblicazione: (2025)
di: Bertina, Abbas, et al.
Pubblicazione: (2025)
Induce, Align, Predict: Zero-Shot Stance Detection via Cognitive Inductive Reasoning
di: Zhang, Bowen, et al.
Pubblicazione: (2025)
di: Zhang, Bowen, et al.
Pubblicazione: (2025)
EmoLoom-2B: Fast Base-Model Screening for Emotion Classification and VAD with Lexicon-Weak Supervision and KV-Off Evaluation
di: Li, Zilin, et al.
Pubblicazione: (2026)
di: Li, Zilin, et al.
Pubblicazione: (2026)
TRiMS: Real-Time Tracking of Minimal Sufficient Length for Efficient Reasoning via RL
di: Bian, Tingcheng, et al.
Pubblicazione: (2026)
di: Bian, Tingcheng, et al.
Pubblicazione: (2026)
MCP: A Control-Theoretic Orchestration Framework for Synergistic Efficiency and Interpretability in Multimodal Large Language Models
di: Zhang, Luyan
Pubblicazione: (2025)
di: Zhang, Luyan
Pubblicazione: (2025)
Mixup Model Merge: Enhancing Model Merging Performance through Randomized Linear Interpolation
di: Zhou, Yue, et al.
Pubblicazione: (2025)
di: Zhou, Yue, et al.
Pubblicazione: (2025)
D-COT: Disciplined Chain-of-Thought Learning for Efficient Reasoning in Small Language Models
di: Ubukata, Shunsuke
Pubblicazione: (2026)
di: Ubukata, Shunsuke
Pubblicazione: (2026)
UrduBench: An Urdu Reasoning Benchmark using Contextually Ensembled Translations with Human-in-the-Loop
di: Shafique, Muhammad Ali, et al.
Pubblicazione: (2026)
di: Shafique, Muhammad Ali, et al.
Pubblicazione: (2026)
LLM Vocabulary Compression for Low-Compute Environments
di: Vennam, Sreeram, et al.
Pubblicazione: (2024)
di: Vennam, Sreeram, et al.
Pubblicazione: (2024)
Truth as a Compression Artifact in Language Model Training
di: Krestnikov, Konstantin
Pubblicazione: (2026)
di: Krestnikov, Konstantin
Pubblicazione: (2026)
Memory Bank Compression for Continual Adaptation of Large Language Models
di: Katraouras, Thomas, et al.
Pubblicazione: (2026)
di: Katraouras, Thomas, et al.
Pubblicazione: (2026)
Fuzzy, Symbolic, and Contextual: Enhancing LLM Instruction via Cognitive Scaffolding
di: Figueiredo, Vanessa
Pubblicazione: (2025)
di: Figueiredo, Vanessa
Pubblicazione: (2025)
Continuous-Depth Transformers with Learned Control Dynamics
di: Jemley, Peter
Pubblicazione: (2026)
di: Jemley, Peter
Pubblicazione: (2026)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
di: Du, Bangde, et al.
Pubblicazione: (2025)
di: Du, Bangde, et al.
Pubblicazione: (2025)
Learning the meanings of function words from grounded language using a visual question answering model
di: Portelance, Eva, et al.
Pubblicazione: (2023)
di: Portelance, Eva, et al.
Pubblicazione: (2023)
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
di: Yang, Shan
Pubblicazione: (2026)
di: Yang, Shan
Pubblicazione: (2026)
SAGE: A Strategy-Aware Graph-Enhanced Generation Framework For Online Counseling
di: Aharon, Eliya Naomi, et al.
Pubblicazione: (2026)
di: Aharon, Eliya Naomi, et al.
Pubblicazione: (2026)
HyDRA: Hybrid Dynamic Routing Architecture for Heterogeneous LLM Pools
di: Garg, Aashna, et al.
Pubblicazione: (2026)
di: Garg, Aashna, et al.
Pubblicazione: (2026)
Annotation Entropy Predicts Per-Example Learning Dynamics in LoRA Fine-Tuning
di: Steele, Brady
Pubblicazione: (2026)
di: Steele, Brady
Pubblicazione: (2026)
Ouroboros: Dynamic Weight Generation for Recursive Transformers via Input-Conditioned LoRA Modulation
di: Jaber, Jaber, et al.
Pubblicazione: (2026)
di: Jaber, Jaber, et al.
Pubblicazione: (2026)
Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales
di: Salfati, Samuel
Pubblicazione: (2026)
di: Salfati, Samuel
Pubblicazione: (2026)
Contextual Integrity in LLMs via Reasoning and Reinforcement Learning
di: Lan, Guangchen, et al.
Pubblicazione: (2025)
di: Lan, Guangchen, et al.
Pubblicazione: (2025)
PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues
di: Kajare, Prajwal Vijay, et al.
Pubblicazione: (2026)
di: Kajare, Prajwal Vijay, et al.
Pubblicazione: (2026)
propella-1: Multi-Property Document Annotation for LLM Data Curation at Scale
di: Idahl, Maximilian, et al.
Pubblicazione: (2026)
di: Idahl, Maximilian, et al.
Pubblicazione: (2026)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
di: Gonzalez, Alberto Andres Valdes
Pubblicazione: (2026)
di: Gonzalez, Alberto Andres Valdes
Pubblicazione: (2026)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
di: Reddy, Sandeep, et al.
Pubblicazione: (2025)
di: Reddy, Sandeep, et al.
Pubblicazione: (2025)
Alternating Reinforcement Learning with Contextual Rubric Rewards: Beyond the Scalarization Strategy
di: Lan, Guangchen, et al.
Pubblicazione: (2026)
di: Lan, Guangchen, et al.
Pubblicazione: (2026)
Evaluation of Hate Speech Detection Using Large Language Models and Geographical Contextualization
di: Zahid, Anwar Hossain, et al.
Pubblicazione: (2025)
di: Zahid, Anwar Hossain, et al.
Pubblicazione: (2025)
BitSkip: An Empirical Analysis of Quantization and Early Exit Composition in Transformers
di: Bhuvaneswaran, Ramshankar, et al.
Pubblicazione: (2025)
di: Bhuvaneswaran, Ramshankar, et al.
Pubblicazione: (2025)
Discovering Transformer Circuits via a Hybrid Attribution and Pruning Framework
di: Gu, Hao, et al.
Pubblicazione: (2025)
di: Gu, Hao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
di: Hashemi, Helia, et al.
Pubblicazione: (2024) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025) -
Layer-Aware Embedding Fusion for LLMs in Text Classifications
di: Gwak, Jiho, et al.
Pubblicazione: (2025) -
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
di: Liu, Zhongxin, et al.
Pubblicazione: (2025) -
On the Influence of Discourse Relations in Persuasive Texts
di: Turk, Nawar, et al.
Pubblicazione: (2025)