Continuity and Isolation Lead to Doubts or Dilemmas in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pasten, Hector, Urrutia, Felipe, Jimenez, Hector, Calderon, Cristian B., Rojas, Cristóbal, Kozachinskiy, Alexander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Decoupling Positional and Symbolic Attention Behavior in Transformers
von: Urrutia, Felipe, et al.
Veröffentlicht: (2025)
von: Urrutia, Felipe, et al.
Veröffentlicht: (2025)
Strassen Attention, Split VC Dimension and Compositionality in Transformers
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2025)
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2025)
Lower bounds on transformers with infinite precision
von: Kozachinskiy, Alexander
Veröffentlicht: (2024)
von: Kozachinskiy, Alexander
Veröffentlicht: (2024)
Positional versus Symbolic Attention Heads: Learning Dynamics, RoPE Geometry, and Length Generalization
von: Urrutia, Felipe, et al.
Veröffentlicht: (2026)
von: Urrutia, Felipe, et al.
Veröffentlicht: (2026)
Message Passing on the Edge: Towards Scalable and Expressive GNNs
von: Barceló, Pablo, et al.
Veröffentlicht: (2025)
von: Barceló, Pablo, et al.
Veröffentlicht: (2025)
A completely uniform transformer for parity
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2025)
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2025)
Simple online learning with consistent oracle
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2023)
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2023)
Ehrenfeucht-Haussler Rank and Chain of Thought
von: Barceló, Pablo, et al.
Veröffentlicht: (2025)
von: Barceló, Pablo, et al.
Veröffentlicht: (2025)
Parity, Sensitivity, and Transformers
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2026)
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2026)
On the Limits of Self-Improving in Large Language Models: The Singularity Is Not Near Without Symbolic Model Synthesis
von: Zenil, Hector
Veröffentlicht: (2026)
von: Zenil, Hector
Veröffentlicht: (2026)
On Token's Dilemma: Dynamic MoE with Drift-Aware Token Assignment for Continual Learning of Large Vision Language Models
von: Zhao, Chongyang, et al.
Veröffentlicht: (2026)
von: Zhao, Chongyang, et al.
Veröffentlicht: (2026)
Risk-Sensitive RL for Alleviating Exploration Dilemmas in Large Language Models
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
Optimal bounds for dissatisfaction in perpetual voting
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2024)
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2024)
On dimensionality of feature vectors in MPNNs
von: Bravo, César, et al.
Veröffentlicht: (2024)
von: Bravo, César, et al.
Veröffentlicht: (2024)
Explaining k-Nearest Neighbors: Abductive and Counterfactual Explanations
von: Barceló, Pablo, et al.
Veröffentlicht: (2025)
von: Barceló, Pablo, et al.
Veröffentlicht: (2025)
Shared Doubt: Zero-shot Cross-Lingual Confidence Estimation for Language Models
von: Kyriakou, Athina, et al.
Veröffentlicht: (2026)
von: Kyriakou, Athina, et al.
Veröffentlicht: (2026)
Language Generation: Complexity Barriers and Implications for Learning
von: Arenas, Marcelo, et al.
Veröffentlicht: (2025)
von: Arenas, Marcelo, et al.
Veröffentlicht: (2025)
Meta-Cognitive Reinforcement Learning with Self-Doubt and Recovery
von: Zhang, Zhipeng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhipeng, et al.
Veröffentlicht: (2026)
Concisely Explaining the Doubt: Minimum-Size Abductive Explanations for Linear Models with a Reject Option
von: Fernandes, Gleilson Pedro, et al.
Veröffentlicht: (2026)
von: Fernandes, Gleilson Pedro, et al.
Veröffentlicht: (2026)
Echoes of Socratic Doubt: Embracing Uncertainty in Calibrated Evidential Reinforcement Learning
von: Stutts, Alex Christopher, et al.
Veröffentlicht: (2024)
von: Stutts, Alex Christopher, et al.
Veröffentlicht: (2024)
The Constitutional Controller: Doubt-Calibrated Steering of Compliant Agents
von: Kohaut, Simon, et al.
Veröffentlicht: (2025)
von: Kohaut, Simon, et al.
Veröffentlicht: (2025)
STABLE: Gated Continual Learning for Large Language Models
von: Hoy, William, et al.
Veröffentlicht: (2025)
von: Hoy, William, et al.
Veröffentlicht: (2025)
Routing-Based Continual Learning for Multimodal Large Language Models
von: Mohta, Jay, et al.
Veröffentlicht: (2025)
von: Mohta, Jay, et al.
Veröffentlicht: (2025)
The Data Addition Dilemma
von: Shen, Judy Hanwen, et al.
Veröffentlicht: (2024)
von: Shen, Judy Hanwen, et al.
Veröffentlicht: (2024)
JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models
von: Dragomir, Alexandra, et al.
Veröffentlicht: (2026)
von: Dragomir, Alexandra, et al.
Veröffentlicht: (2026)
Why Keep Your Doubts to Yourself? Trading Visual Uncertainties in Multi-Agent Bandit Systems
von: Zhang, Jusheng, et al.
Veröffentlicht: (2026)
von: Zhang, Jusheng, et al.
Veröffentlicht: (2026)
CAMEL: Continuous Action Masking Enabled by Large Language Models for Reinforcement Learning
von: Zhao, Yanxiao, et al.
Veröffentlicht: (2025)
von: Zhao, Yanxiao, et al.
Veröffentlicht: (2025)
When Continue Learning Meets Multimodal Large Language Model: A Survey
von: Huo, Yukang, et al.
Veröffentlicht: (2025)
von: Huo, Yukang, et al.
Veröffentlicht: (2025)
The Chicken and Egg Dilemma: Co-optimizing Data and Model Configurations for LLMs
von: Chen, Zhiliang, et al.
Veröffentlicht: (2026)
von: Chen, Zhiliang, et al.
Veröffentlicht: (2026)
Learning Generalized Policies for Fully Observable Non-Deterministic Planning Domains
von: Hofmann, Till, et al.
Veröffentlicht: (2024)
von: Hofmann, Till, et al.
Veröffentlicht: (2024)
Are Large-Language Models Graph Algorithmic Reasoners?
von: Taylor, Alexander K, et al.
Veröffentlicht: (2024)
von: Taylor, Alexander K, et al.
Veröffentlicht: (2024)
Functional Component Ablation Reveals Specialization Patterns in Hybrid Language Model Architectures
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
When Less is More: 8-bit Quantization Improves Continual Learning in Large Language Models
von: Zhang, Michael S., et al.
Veröffentlicht: (2025)
von: Zhang, Michael S., et al.
Veröffentlicht: (2025)
Robust Uncertainty Quantification for Self-Evolving Large Language Models via Continual Domain Pretraining
von: Zhou, Xiaofan, et al.
Veröffentlicht: (2025)
von: Zhou, Xiaofan, et al.
Veröffentlicht: (2025)
Multi-Level Safety Continual Projection for Fine-Tuned Large Language Models without Retraining
von: Han, Bing, et al.
Veröffentlicht: (2025)
von: Han, Bing, et al.
Veröffentlicht: (2025)
COPAL: Continual Pruning in Large Language Generative Models
von: Malla, Srikanth, et al.
Veröffentlicht: (2024)
von: Malla, Srikanth, et al.
Veröffentlicht: (2024)
SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks
von: Robey, Alexander, et al.
Veröffentlicht: (2023)
von: Robey, Alexander, et al.
Veröffentlicht: (2023)
Jailbreaking Black Box Large Language Models in Twenty Queries
von: Chao, Patrick, et al.
Veröffentlicht: (2023)
von: Chao, Patrick, et al.
Veröffentlicht: (2023)
Learning Dynamics in Continual Pre-Training for Large Language Models
von: Wang, Xingjin, et al.
Veröffentlicht: (2025)
von: Wang, Xingjin, et al.
Veröffentlicht: (2025)
Continual Learning of Large Language Models: A Comprehensive Survey
von: Shi, Haizhou, et al.
Veröffentlicht: (2024)
von: Shi, Haizhou, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Decoupling Positional and Symbolic Attention Behavior in Transformers
von: Urrutia, Felipe, et al.
Veröffentlicht: (2025) -
Strassen Attention, Split VC Dimension and Compositionality in Transformers
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2025) -
Lower bounds on transformers with infinite precision
von: Kozachinskiy, Alexander
Veröffentlicht: (2024) -
Positional versus Symbolic Attention Heads: Learning Dynamics, RoPE Geometry, and Length Generalization
von: Urrutia, Felipe, et al.
Veröffentlicht: (2026) -
Message Passing on the Edge: Towards Scalable and Expressive GNNs
von: Barceló, Pablo, et al.
Veröffentlicht: (2025)