Entropy-Reservoir Bregman Projection: An Information-Geometric Unification of Model Collapse
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Chen, Jingwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Survey Transfer Learning: Recycling Data with Silicon Responses
von: Amini, Ali
Veröffentlicht: (2025)
von: Amini, Ali
Veröffentlicht: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
Does Editing Provide Evidence for Localization?
von: Wang, Zihao, et al.
Veröffentlicht: (2025)
von: Wang, Zihao, et al.
Veröffentlicht: (2025)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
von: Bachar, Or, et al.
Veröffentlicht: (2026)
von: Bachar, Or, et al.
Veröffentlicht: (2026)
On Semantic Loss Fine-Tuning Approach for Preventing Model Collapse in Causal Reasoning
von: Deshmukh, Pratik, et al.
Veröffentlicht: (2026)
von: Deshmukh, Pratik, et al.
Veröffentlicht: (2026)
Less is More: Learning Graph Tasks with Just LLMs
von: Shirai, Sola, et al.
Veröffentlicht: (2025)
von: Shirai, Sola, et al.
Veröffentlicht: (2025)
Model Collapse as Cultural Evolution
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Measure what Matters: Psychometric Evaluation of AI with Situational Judgment Tests
von: Yost, Alexandra, et al.
Veröffentlicht: (2025)
von: Yost, Alexandra, et al.
Veröffentlicht: (2025)
Theoretical Analysis of Positional Encodings in Transformer Models: Impact on Expressiveness and Generalization
von: Li, Yin
Veröffentlicht: (2025)
von: Li, Yin
Veröffentlicht: (2025)
Latent Cache Flow: Model-to-Model Communication Without Text
von: Rossi, Maximillian, et al.
Veröffentlicht: (2026)
von: Rossi, Maximillian, et al.
Veröffentlicht: (2026)
Alif: Advancing Urdu Large Language Models via Multilingual Synthetic Data Distillation
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2025)
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2025)
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
von: Yocam, Eric, et al.
Veröffentlicht: (2026)
von: Yocam, Eric, et al.
Veröffentlicht: (2026)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
von: Costa, Rimom
Veröffentlicht: (2025)
von: Costa, Rimom
Veröffentlicht: (2025)
Advancing Transformer Architecture in Long-Context Large Language Models: A Comprehensive Survey
von: Huang, Yunpeng, et al.
Veröffentlicht: (2023)
von: Huang, Yunpeng, et al.
Veröffentlicht: (2023)
Reinforcement Learning for Scalable and Trustworthy Intelligent Systems
von: Lan, Guangchen
Veröffentlicht: (2026)
von: Lan, Guangchen
Veröffentlicht: (2026)
Autonomous Deep Agent
von: Yu, Amy, et al.
Veröffentlicht: (2025)
von: Yu, Amy, et al.
Veröffentlicht: (2025)
The Geometry of Thought: How Scale Restructures Reasoning In Large Language Models
von: Anderson, Samuel Cyrenius
Veröffentlicht: (2026)
von: Anderson, Samuel Cyrenius
Veröffentlicht: (2026)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
von: Sun, Yuhui, et al.
Veröffentlicht: (2025)
von: Sun, Yuhui, et al.
Veröffentlicht: (2025)
Reasoning Large Language Model Errors Arise from Hallucinating Critical Problem Features
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels
von: Rath, Plawan Kumar, et al.
Veröffentlicht: (2026)
von: Rath, Plawan Kumar, et al.
Veröffentlicht: (2026)
The Anti-Ouroboros Effect: Emergent Resilience in Large Language Models from Recursive Selective Feedback
von: Adapala, Sai Teja Reddy
Veröffentlicht: (2025)
von: Adapala, Sai Teja Reddy
Veröffentlicht: (2025)
Controlled Territory and Conflict Tracking (CONTACT): (Geo-)Mapping Occupied Territory from Open Source Intelligence
von: Mandal, Paul K., et al.
Veröffentlicht: (2025)
von: Mandal, Paul K., et al.
Veröffentlicht: (2025)
PRPO: Aligning Process Reward with Outcome Reward in Policy Optimization
von: Ding, Ruiyi, et al.
Veröffentlicht: (2026)
von: Ding, Ruiyi, et al.
Veröffentlicht: (2026)
FastForward Pruning: Efficient LLM Pruning via Single-Step Reinforcement Learning
von: Yuan, Xin, et al.
Veröffentlicht: (2025)
von: Yuan, Xin, et al.
Veröffentlicht: (2025)
On the Compatibility of Generative AI and Generative Linguistics
von: Portelance, Eva, et al.
Veröffentlicht: (2024)
von: Portelance, Eva, et al.
Veröffentlicht: (2024)
Generalization of Graph Neural Network Models for Distribution Grid Fault Detection
von: Karabulut, Burak, et al.
Veröffentlicht: (2025)
von: Karabulut, Burak, et al.
Veröffentlicht: (2025)
GraphEval36K: Benchmarking Coding and Reasoning Capabilities of Large Language Models on Graph Datasets
von: Wu, Qiming, et al.
Veröffentlicht: (2024)
von: Wu, Qiming, et al.
Veröffentlicht: (2024)
Social Cooperation in Conversational AI Agents
von: Çelikok, Mustafa Mert, et al.
Veröffentlicht: (2025)
von: Çelikok, Mustafa Mert, et al.
Veröffentlicht: (2025)
OFMU: Optimization-Driven Framework for Machine Unlearning
von: Asif, Sadia, et al.
Veröffentlicht: (2025)
von: Asif, Sadia, et al.
Veröffentlicht: (2025)
Persona Features Control Emergent Misalignment
von: Wang, Miles, et al.
Veröffentlicht: (2025)
von: Wang, Miles, et al.
Veröffentlicht: (2025)
FastGRPO: Accelerating Policy Optimization via Concurrency-aware Speculative Decoding and Online Draft Learning
von: Zhang, Yizhou, et al.
Veröffentlicht: (2025)
von: Zhang, Yizhou, et al.
Veröffentlicht: (2025)
Overclocking LLM Reasoning: Monitoring and Controlling Thinking Path Lengths in LLMs
von: Eisenstadt, Roy, et al.
Veröffentlicht: (2025)
von: Eisenstadt, Roy, et al.
Veröffentlicht: (2025)
Node-Level Uncertainty Estimation in LLM-Generated SQL
von: Hasson, Hilaf, et al.
Veröffentlicht: (2025)
von: Hasson, Hilaf, et al.
Veröffentlicht: (2025)
ReflexGrad: Within-Episode Failure Recovery in LLM Agents via Progress-Gated Dual-Process Routing
von: Kadu, Ankush, et al.
Veröffentlicht: (2025)
von: Kadu, Ankush, et al.
Veröffentlicht: (2025)
How Does Unfaithful Reasoning Emerge from Autoregressive Training? A Study of Synthetic Experiments
von: Wang, Fuxin, et al.
Veröffentlicht: (2026)
von: Wang, Fuxin, et al.
Veröffentlicht: (2026)
Extreme AutoML: Analysis of Classification, Regression, and NLP Performance
von: Ratner, Edward, et al.
Veröffentlicht: (2024)
von: Ratner, Edward, et al.
Veröffentlicht: (2024)
CircuitProbe: Predicting Reasoning Circuits in Transformers via Stability Zone Detection
von: Panuganti, Rajkiran
Veröffentlicht: (2026)
von: Panuganti, Rajkiran
Veröffentlicht: (2026)
BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2026)
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2026)
Task-Conditioned Routing Signatures in Sparse Mixture-of-Experts Transformers
von: Avinash, Mynampati Sri Ranganadha
Veröffentlicht: (2026)
von: Avinash, Mynampati Sri Ranganadha
Veröffentlicht: (2026)
Shattered Compositionality: Counterintuitive Learning Dynamics of Transformers for Arithmetic
von: Zhao, Xingyu, et al.
Veröffentlicht: (2026)
von: Zhao, Xingyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Survey Transfer Learning: Recycling Data with Silicon Responses
von: Amini, Ali
Veröffentlicht: (2025) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025) -
Does Editing Provide Evidence for Localization?
von: Wang, Zihao, et al.
Veröffentlicht: (2025) -
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
von: Bachar, Or, et al.
Veröffentlicht: (2026) -
On Semantic Loss Fine-Tuning Approach for Preventing Model Collapse in Causal Reasoning
von: Deshmukh, Pratik, et al.
Veröffentlicht: (2026)