Idea-Gated Transformers: Enforcing Semantic Coherence via Differentiable Vocabulary Pruning
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Fofadiya, Darshan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Factual Grounding: The Case for Opinion-Aware Retrieval-Augmented Generation
von: Agrawal, Aditya, et al.
Veröffentlicht: (2026)
von: Agrawal, Aditya, et al.
Veröffentlicht: (2026)
Dynamic Vocabulary Pruning in Early-Exit LLMs
von: Vincenti, Jort, et al.
Veröffentlicht: (2024)
von: Vincenti, Jort, et al.
Veröffentlicht: (2024)
SECA: Semantically Equivalent and Coherent Attacks for Eliciting LLM Hallucinations
von: Liang, Buyun, et al.
Veröffentlicht: (2025)
von: Liang, Buyun, et al.
Veröffentlicht: (2025)
Revisiting Large Language Model Pruning using Neuron Semantic Attribution
von: Ding, Yizhuo, et al.
Veröffentlicht: (2025)
von: Ding, Yizhuo, et al.
Veröffentlicht: (2025)
GoAI: Enhancing AI Students' Learning Paths and Idea Generation via Graph of AI Ideas
von: Gao, Xian, et al.
Veröffentlicht: (2025)
von: Gao, Xian, et al.
Veröffentlicht: (2025)
COCORELI: Enforcing Execution Preconditions for Reliable Collaborative Instruction Following
von: Bhar, Swarnadeep, et al.
Veröffentlicht: (2025)
von: Bhar, Swarnadeep, et al.
Veröffentlicht: (2025)
Chopping Trees: Semantic Similarity Based Dynamic Pruning for Tree-of-Thought Reasoning
von: Kim, Joongho, et al.
Veröffentlicht: (2025)
von: Kim, Joongho, et al.
Veröffentlicht: (2025)
Semantic Exploration with Adaptive Gating for Efficient Problem Solving with Language Models
von: Lee, Sungjae, et al.
Veröffentlicht: (2025)
von: Lee, Sungjae, et al.
Veröffentlicht: (2025)
`Keep it Together': Enforcing Cohesion in Extractive Summaries by Simulating Human Memory
von: Cardenas, Ronald, et al.
Veröffentlicht: (2024)
von: Cardenas, Ronald, et al.
Veröffentlicht: (2024)
Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies
von: Tao, Chaofan, et al.
Veröffentlicht: (2024)
von: Tao, Chaofan, et al.
Veröffentlicht: (2024)
Adaptive Computation Pruning for the Forgetting Transformer
von: Lin, Zhixuan, et al.
Veröffentlicht: (2025)
von: Lin, Zhixuan, et al.
Veröffentlicht: (2025)
ReplaceMe: Network Simplification via Depth Pruning and Transformer Block Linearization
von: Shopkhoev, Dmitriy, et al.
Veröffentlicht: (2025)
von: Shopkhoev, Dmitriy, et al.
Veröffentlicht: (2025)
Overcoming Vocabulary Mismatch: Vocabulary-agnostic Teacher Guided Language Modeling
von: Shin, Haebin, et al.
Veröffentlicht: (2025)
von: Shin, Haebin, et al.
Veröffentlicht: (2025)
Semantic Flow Regularization: Teaching LLMs to Generate Diverse Yet Coherent Responses
von: Peng, Kerui, et al.
Veröffentlicht: (2026)
von: Peng, Kerui, et al.
Veröffentlicht: (2026)
AI Idea Bench 2025: AI Research Idea Generation Benchmark
von: Qiu, Yansheng, et al.
Veröffentlicht: (2025)
von: Qiu, Yansheng, et al.
Veröffentlicht: (2025)
Learning and Enforcing Context-Sensitive Control for LLMs
von: Albinhassan, Mohammad, et al.
Veröffentlicht: (2026)
von: Albinhassan, Mohammad, et al.
Veröffentlicht: (2026)
Chain of Ideas: Revolutionizing Research Via Novel Idea Development with LLM Agents
von: Li, Long, et al.
Veröffentlicht: (2024)
von: Li, Long, et al.
Veröffentlicht: (2024)
EvoP: Robust LLM Inference via Evolutionary Pruning
von: Wu, Shangyu, et al.
Veröffentlicht: (2025)
von: Wu, Shangyu, et al.
Veröffentlicht: (2025)
EEG2TEXT: Open Vocabulary EEG-to-Text Decoding with EEG Pre-Training and Multi-View Transformer
von: Liu, Hanwen, et al.
Veröffentlicht: (2024)
von: Liu, Hanwen, et al.
Veröffentlicht: (2024)
Developing Adaptive Context Compression Techniques for Large Language Models (LLMs) in Long-Running Interactions
von: Fofadiya, Payal, et al.
Veröffentlicht: (2026)
von: Fofadiya, Payal, et al.
Veröffentlicht: (2026)
Novel Memory Forgetting Techniques for Autonomous AI Agents: Balancing Relevance and Efficiency
von: Fofadiya, Payal, et al.
Veröffentlicht: (2026)
von: Fofadiya, Payal, et al.
Veröffentlicht: (2026)
Multi-Layered Memory Architectures for LLM Agents: An Experimental Evaluation of Long-Term Context Retention
von: Tiwari, Sunil, et al.
Veröffentlicht: (2026)
von: Tiwari, Sunil, et al.
Veröffentlicht: (2026)
Reading Between the Lines: Combining Pause Dynamics and Semantic Coherence for Automated Assessment of Thought Disorder
von: Chen, Feng, et al.
Veröffentlicht: (2025)
von: Chen, Feng, et al.
Veröffentlicht: (2025)
S$^4$C: Speculative Sampling with Syntactic and Semantic Coherence for Efficient Inference of Large Language Models
von: He, Tao, et al.
Veröffentlicht: (2025)
von: He, Tao, et al.
Veröffentlicht: (2025)
Shayona@SMM4H23: COVID-19 Self diagnosis classification using BERT and LightGBM models
von: Chavda, Rushi, et al.
Veröffentlicht: (2024)
von: Chavda, Rushi, et al.
Veröffentlicht: (2024)
PruneCD: Contrasting Pruned Self Model to Improve Decoding Factuality
von: Yu, Byeongho, et al.
Veröffentlicht: (2025)
von: Yu, Byeongho, et al.
Veröffentlicht: (2025)
Pruning via Merging: Compressing LLMs via Manifold Alignment Based Layer Merging
von: Liu, Deyuan, et al.
Veröffentlicht: (2024)
von: Liu, Deyuan, et al.
Veröffentlicht: (2024)
PGB: One-Shot Pruning for BERT via Weight Grouping and Permutation
von: Lim, Hyemin, et al.
Veröffentlicht: (2025)
von: Lim, Hyemin, et al.
Veröffentlicht: (2025)
Context Pruning for Coding Agents via Multi-Rubric Latent Reasoning
von: Wang, Jingjing, et al.
Veröffentlicht: (2026)
von: Wang, Jingjing, et al.
Veröffentlicht: (2026)
LaCo: Large Language Model Pruning via Layer Collapse
von: Yang, Yifei, et al.
Veröffentlicht: (2024)
von: Yang, Yifei, et al.
Veröffentlicht: (2024)
Vocabulary Expansion of Large Language Models via Kullback-Leibler-Based Self-Distillation
von: Linder, Max Rehman
Veröffentlicht: (2025)
von: Linder, Max Rehman
Veröffentlicht: (2025)
EvoSpec: Evolving Speculative Decoding via Real-Time Vocabulary and Parameter Adaptation
von: Zhang, Shuyu, et al.
Veröffentlicht: (2026)
von: Zhang, Shuyu, et al.
Veröffentlicht: (2026)
Enforcing Monotonic Progress in Legal Cross-Examination: Preventing Long-Horizon Stagnation in LLM-Based Inquiry
von: Liao, Hsien-Jyh
Veröffentlicht: (2026)
von: Liao, Hsien-Jyh
Veröffentlicht: (2026)
DVAGen: Dynamic Vocabulary Augmented Generation
von: Du, Wei, et al.
Veröffentlicht: (2025)
von: Du, Wei, et al.
Veröffentlicht: (2025)
Towards Robust Pruning: An Adaptive Knowledge-Retention Pruning Strategy for Language Models
von: Li, Jianwei, et al.
Veröffentlicht: (2023)
von: Li, Jianwei, et al.
Veröffentlicht: (2023)
Probe Pruning: Accelerating LLMs through Dynamic Pruning via Model-Probing
von: Le, Qi, et al.
Veröffentlicht: (2025)
von: Le, Qi, et al.
Veröffentlicht: (2025)
Focus on the Core: Efficient Attention via Pruned Token Compression for Document Classification
von: Yun, Jungmin, et al.
Veröffentlicht: (2024)
von: Yun, Jungmin, et al.
Veröffentlicht: (2024)
Scaling LLM Pre-training with Vocabulary Curriculum
von: Yu, Fangyuan
Veröffentlicht: (2025)
von: Yu, Fangyuan
Veröffentlicht: (2025)
Forgetting Transformer: Softmax Attention with a Forget Gate
von: Lin, Zhixuan, et al.
Veröffentlicht: (2025)
von: Lin, Zhixuan, et al.
Veröffentlicht: (2025)
Enforcing Temporal Constraints for LLM Agents
von: Kamath, Adharsh, et al.
Veröffentlicht: (2025)
von: Kamath, Adharsh, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Beyond Factual Grounding: The Case for Opinion-Aware Retrieval-Augmented Generation
von: Agrawal, Aditya, et al.
Veröffentlicht: (2026) -
Dynamic Vocabulary Pruning in Early-Exit LLMs
von: Vincenti, Jort, et al.
Veröffentlicht: (2024) -
SECA: Semantically Equivalent and Coherent Attacks for Eliciting LLM Hallucinations
von: Liang, Buyun, et al.
Veröffentlicht: (2025) -
Revisiting Large Language Model Pruning using Neuron Semantic Attribution
von: Ding, Yizhuo, et al.
Veröffentlicht: (2025) -
GoAI: Enhancing AI Students' Learning Paths and Idea Generation via Graph of AI Ideas
von: Gao, Xian, et al.
Veröffentlicht: (2025)