Towards Effective Theory of LLMs: A Representation Learning Approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ustaomeroglu, Muhammed, Qu, Guannan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2026)
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2026)
A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2025)
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2025)
Internal Planning in Language Models: Characterizing Horizon and Branch Awareness
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2025)
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2025)
Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer
von: Askin, Baris, et al.
Veröffentlicht: (2026)
von: Askin, Baris, et al.
Veröffentlicht: (2026)
Transformer-Based Scalable Multi-Agent Reinforcement Learning for Networked Systems with Long-Range Interactions
von: Sinha, Vidur, et al.
Veröffentlicht: (2025)
von: Sinha, Vidur, et al.
Veröffentlicht: (2025)
Towards a Learning Theory of Representation Alignment
von: Insulla, Francesco, et al.
Veröffentlicht: (2025)
von: Insulla, Francesco, et al.
Veröffentlicht: (2025)
Bayesian Kolmogorov Arnold Networks (Bayesian_KANs): A Probabilistic Approach to Enhance Accuracy and Interpretability
von: Hassan, Masoud Muhammed
Veröffentlicht: (2024)
von: Hassan, Masoud Muhammed
Veröffentlicht: (2024)
Towards a Generic Representation of Combinatorial Problems for Learning-Based Approaches
von: Boisvert, Léo, et al.
Veröffentlicht: (2024)
von: Boisvert, Léo, et al.
Veröffentlicht: (2024)
Thinking Beyond Visibility: A Near-Optimal Policy Framework for Locally Interdependent Multi-Agent MDPs
von: DeWeese, Alex, et al.
Veröffentlicht: (2025)
von: DeWeese, Alex, et al.
Veröffentlicht: (2025)
Toward Effective Digraph Representation Learning: A Magnetic Adaptive Propagation based Approach
von: Li, Xunkai, et al.
Veröffentlicht: (2025)
von: Li, Xunkai, et al.
Veröffentlicht: (2025)
Locally Interdependent Multi-Agent MDP: Theoretical Framework for Decentralized Agents with Dynamic Dependencies
von: DeWeese, Alex, et al.
Veröffentlicht: (2024)
von: DeWeese, Alex, et al.
Veröffentlicht: (2024)
RED: Effective Trajectory Representation Learning with Comprehensive Information
von: Zhou, Silin, et al.
Veröffentlicht: (2024)
von: Zhou, Silin, et al.
Veröffentlicht: (2024)
The Effectiveness of LLMs as Annotators: A Comparative Overview and Empirical Analysis of Direct Representation
von: Pavlovic, Maja, et al.
Veröffentlicht: (2024)
von: Pavlovic, Maja, et al.
Veröffentlicht: (2024)
any4: Learned 4-bit Numeric Representation for LLMs
von: Elhoushi, Mostafa, et al.
Veröffentlicht: (2025)
von: Elhoushi, Mostafa, et al.
Veröffentlicht: (2025)
MolMix: A Simple Yet Effective Baseline for Multimodal Molecular Representation Learning
von: Manolache, Andrei, et al.
Veröffentlicht: (2024)
von: Manolache, Andrei, et al.
Veröffentlicht: (2024)
Toward Temporal Causal Representation Learning with Tensor Decomposition
von: Chen, Jianhong, et al.
Veröffentlicht: (2025)
von: Chen, Jianhong, et al.
Veröffentlicht: (2025)
Exploiting Latent Linearity in LLMs Improves Explainable Molecular Representation Learning
von: Li, Zhuoran, et al.
Veröffentlicht: (2024)
von: Li, Zhuoran, et al.
Veröffentlicht: (2024)
MolReasoner: Toward Effective and Interpretable Reasoning for Molecular LLMs
von: Zhao, Guojiang, et al.
Veröffentlicht: (2025)
von: Zhao, Guojiang, et al.
Veröffentlicht: (2025)
Towards Effective Experiential Learning: Dual Guidance for Utilization and Internalization
von: Bai, Fei, et al.
Veröffentlicht: (2026)
von: Bai, Fei, et al.
Veröffentlicht: (2026)
Merge then Realign: Simple and Effective Modality-Incremental Continual Learning for Multimodal LLMs
von: Zhang, Dingkun, et al.
Veröffentlicht: (2025)
von: Zhang, Dingkun, et al.
Veröffentlicht: (2025)
Predictive Modeling of Homeless Service Assignment: A Representation Learning Approach
von: Rahman, Khandker Sadia, et al.
Veröffentlicht: (2024)
von: Rahman, Khandker Sadia, et al.
Veröffentlicht: (2024)
A Theory of Training Profit-Optimal LLMs
von: Hao, Sophie, et al.
Veröffentlicht: (2026)
von: Hao, Sophie, et al.
Veröffentlicht: (2026)
Toward Enhancing Representation Learning in Federated Multi-Task Settings
von: Setayesh, Mehdi, et al.
Veröffentlicht: (2026)
von: Setayesh, Mehdi, et al.
Veröffentlicht: (2026)
DeepGate3: Towards Scalable Circuit Representation Learning
von: Shi, Zhengyuan, et al.
Veröffentlicht: (2024)
von: Shi, Zhengyuan, et al.
Veröffentlicht: (2024)
Schema-Adaptive Tabular Representation Learning with LLMs for Generalizable Multimodal Clinical Reasoning
von: Mao, Hongxi, et al.
Veröffentlicht: (2026)
von: Mao, Hongxi, et al.
Veröffentlicht: (2026)
Refining Latent Representations: A Generative SSL Approach for Heterogeneous Graph Learning
von: Hu, Yulan, et al.
Veröffentlicht: (2023)
von: Hu, Yulan, et al.
Veröffentlicht: (2023)
TSI: A Multi-View Representation Learning Approach for Time Series Forecasting
von: Gao, Wentao, et al.
Veröffentlicht: (2024)
von: Gao, Wentao, et al.
Veröffentlicht: (2024)
FedSA: A Unified Representation Learning via Semantic Anchors for Prototype-based Federated Learning
von: Zhou, Yanbing, et al.
Veröffentlicht: (2025)
von: Zhou, Yanbing, et al.
Veröffentlicht: (2025)
COSMOS: Predictable and Cost-Effective Adaptation of LLMs
von: Wang, Jiayu, et al.
Veröffentlicht: (2025)
von: Wang, Jiayu, et al.
Veröffentlicht: (2025)
Simple and Effective Specialized Representations for Fair Classifiers
von: Sinigaglia, Alberto, et al.
Veröffentlicht: (2025)
von: Sinigaglia, Alberto, et al.
Veröffentlicht: (2025)
When LLMs Learn to Be Consistently Wrong: A Multi-Model Study of Linear Representations of Synthetic Deception
von: Zolfaghari, Vahideh
Veröffentlicht: (2026)
von: Zolfaghari, Vahideh
Veröffentlicht: (2026)
Towards Robust Trajectory Representations: Isolating Environmental Confounders with Causal Learning
von: Luo, Kang, et al.
Veröffentlicht: (2024)
von: Luo, Kang, et al.
Veröffentlicht: (2024)
Exploring Open-world Continual Learning with Knowns-Unknowns Knowledge Transfer
von: Li, Yujie, et al.
Veröffentlicht: (2025)
von: Li, Yujie, et al.
Veröffentlicht: (2025)
A2PO: Towards Effective Offline Reinforcement Learning from an Advantage-aware Perspective
von: Qing, Yunpeng, et al.
Veröffentlicht: (2024)
von: Qing, Yunpeng, et al.
Veröffentlicht: (2024)
LightDiC: A Simple yet Effective Approach for Large-scale Digraph Representation Learning
von: Li, Xunkai, et al.
Veröffentlicht: (2024)
von: Li, Xunkai, et al.
Veröffentlicht: (2024)
A Self-guided Multimodal Approach to Enhancing Graph Representation Learning for Alzheimer's Diseases
von: Wang, Zhepeng, et al.
Veröffentlicht: (2024)
von: Wang, Zhepeng, et al.
Veröffentlicht: (2024)
Exploring Task Unification in Graph Representation Learning via Generative Approach
von: Hu, Yulan, et al.
Veröffentlicht: (2024)
von: Hu, Yulan, et al.
Veröffentlicht: (2024)
EMP: Effective Multidimensional Persistence for Graph Representation Learning
von: Segovia-Dominguez, Ignacio, et al.
Veröffentlicht: (2024)
von: Segovia-Dominguez, Ignacio, et al.
Veröffentlicht: (2024)
DECRL: A Deep Evolutionary Clustering Jointed Temporal Knowledge Graph Representation Learning Approach
von: Chen, Qian, et al.
Veröffentlicht: (2024)
von: Chen, Qian, et al.
Veröffentlicht: (2024)
Frequency-Masked Embedding Inference: A Non-Contrastive Approach for Time Series Representation Learning
von: Fu, En, et al.
Veröffentlicht: (2024)
von: Fu, En, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2026) -
A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2025) -
Internal Planning in Language Models: Characterizing Horizon and Branch Awareness
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2025) -
Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer
von: Askin, Baris, et al.
Veröffentlicht: (2026) -
Transformer-Based Scalable Multi-Agent Reinforcement Learning for Networked Systems with Long-Range Interactions
von: Sinha, Vidur, et al.
Veröffentlicht: (2025)