Pelican Soup Framework: A Theoretical Framework for Language Model Capabilities
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chiang, Ting-Rui, Yogatama, Dani |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Rotary Position Embedding May Cause Dimension Inefficiency in Attention Heads for Long-Distance Retrieval
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2025)
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2025)
LocateBench: Evaluating the Locating Ability of Vision Language Models
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2024)
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2024)
DeLLMa: Decision Making Under Uncertainty with Large Language Models
von: Liu, Ollie, et al.
Veröffentlicht: (2024)
von: Liu, Ollie, et al.
Veröffentlicht: (2024)
Causal Interventions on Causal Paths: Mapping GPT-2's Reasoning From Syntax to Semantics
von: Lee, Isabelle, et al.
Veröffentlicht: (2024)
von: Lee, Isabelle, et al.
Veröffentlicht: (2024)
Bone Soups: A Seek-and-Soup Model Merging Approach for Controllable Multi-Objective Generation
von: Xie, Guofu, et al.
Veröffentlicht: (2025)
von: Xie, Guofu, et al.
Veröffentlicht: (2025)
IsoBench: Benchmarking Multimodal Foundation Models on Isomorphic Representations
von: Fu, Deqing, et al.
Veröffentlicht: (2024)
von: Fu, Deqing, et al.
Veröffentlicht: (2024)
On Retrieval Augmentation and the Limitations of Language Model Training
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2023)
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2023)
An Information-Theoretic Framework for Robust Large Language Model Editing
von: Chen, Qizhou, et al.
Veröffentlicht: (2025)
von: Chen, Qizhou, et al.
Veröffentlicht: (2025)
Representations as Language: An Information-Theoretic Framework for Interpretability
von: Conklin, Henry, et al.
Veröffentlicht: (2024)
von: Conklin, Henry, et al.
Veröffentlicht: (2024)
LieCraft: A Multi-Agent Framework for Evaluating Deceptive Capabilities in Language Models
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2026)
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2026)
USB-Rec: An Effective Framework for Improving Conversational Recommendation Capability of Large Language Model
von: Wen, Jianyu, et al.
Veröffentlicht: (2025)
von: Wen, Jianyu, et al.
Veröffentlicht: (2025)
Elo-Evolve: A Co-evolutionary Framework for Language Model Alignment
von: Zhao, Jing, et al.
Veröffentlicht: (2026)
von: Zhao, Jing, et al.
Veröffentlicht: (2026)
NewsBench: A Systematic Evaluation Framework for Assessing Editorial Capabilities of Large Language Models in Chinese Journalism
von: Li, Miao, et al.
Veröffentlicht: (2024)
von: Li, Miao, et al.
Veröffentlicht: (2024)
MUG-Eval: A Proxy Evaluation Framework for Multilingual Generation Capabilities in Any Language
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
Emergent Hierarchical Structure in Large Language Models: An Information-Theoretic Framework for Multi-Scale Representation
von: Zhang, Yukin, et al.
Veröffentlicht: (2025)
von: Zhang, Yukin, et al.
Veröffentlicht: (2025)
CogniDual Framework: Self-Training Large Language Models within a Dual-System Theoretical Framework for Improving Cognitive Tasks
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
Cross-Lingual Transfer and Parameter-Efficient Adaptation in the Turkic Language Family: A Theoretical Framework for Low-Resource Language Models
von: Ibrahimzade, O., et al.
Veröffentlicht: (2026)
von: Ibrahimzade, O., et al.
Veröffentlicht: (2026)
Towards Better Multi-task Learning: A Framework for Optimizing Dataset Combinations in Large Language Models
von: Zhan, Zaifu, et al.
Veröffentlicht: (2024)
von: Zhan, Zaifu, et al.
Veröffentlicht: (2024)
Redefining Evaluation Standards: A Unified Framework for Evaluating the Korean Capabilities of Language Models
von: Lee, Hanwool, et al.
Veröffentlicht: (2025)
von: Lee, Hanwool, et al.
Veröffentlicht: (2025)
Cybench: A Framework for Evaluating Cybersecurity Capabilities and Risks of Language Models
von: Zhang, Andy K., et al.
Veröffentlicht: (2024)
von: Zhang, Andy K., et al.
Veröffentlicht: (2024)
Sycophancy in Vision-Language Models: A Systematic Analysis and an Inference-Time Mitigation Framework
von: Zhao, Yunpu, et al.
Veröffentlicht: (2024)
von: Zhao, Yunpu, et al.
Veröffentlicht: (2024)
Self-Review Framework for Enhancing Instruction Following Capability of LLM
von: Park, Sihyun
Veröffentlicht: (2025)
von: Park, Sihyun
Veröffentlicht: (2025)
Toward Verifiable Misinformation Detection: A Multi-Tool LLM Agent Framework
von: Cui, Zikun, et al.
Veröffentlicht: (2025)
von: Cui, Zikun, et al.
Veröffentlicht: (2025)
SQLBench: A Comprehensive Evaluation for Text-to-SQL Capabilities of Large Language Models
von: Zhang, Bin, et al.
Veröffentlicht: (2024)
von: Zhang, Bin, et al.
Veröffentlicht: (2024)
GS-KGC: A Generative Subgraph-based Framework for Knowledge Graph Completion with Large Language Models
von: Yang, Rui, et al.
Veröffentlicht: (2024)
von: Yang, Rui, et al.
Veröffentlicht: (2024)
EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models
von: Zhou, Weikang, et al.
Veröffentlicht: (2024)
von: Zhou, Weikang, et al.
Veröffentlicht: (2024)
Enhancing Emotional Generation Capability of Large Language Models via Emotional Chain-of-Thought
von: Li, Zaijing, et al.
Veröffentlicht: (2024)
von: Li, Zaijing, et al.
Veröffentlicht: (2024)
CoT-Driven Framework for Short Text Classification: Enhancing and Transferring Capabilities from Large to Smaller Model
von: Wu, Hui, et al.
Veröffentlicht: (2024)
von: Wu, Hui, et al.
Veröffentlicht: (2024)
LLMs Judge Themselves: A Game-Theoretic Framework for Human-Aligned Evaluation
von: Yang, Gao, et al.
Veröffentlicht: (2025)
von: Yang, Gao, et al.
Veröffentlicht: (2025)
Unraveling the Capabilities of Language Models in News Summarization
von: Odabaşı, Abdurrahman, et al.
Veröffentlicht: (2025)
von: Odabaşı, Abdurrahman, et al.
Veröffentlicht: (2025)
Forecasting Frontier Language Model Agent Capabilities
von: Pimpale, Govind, et al.
Veröffentlicht: (2025)
von: Pimpale, Govind, et al.
Veröffentlicht: (2025)
A Novel Self-Evolution Framework for Large Language Models
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
CNSL-bench: Benchmarking the Sign Language Understanding Capabilities of MLLMs on Chinese National Sign Language
von: Zhao, Rui, et al.
Veröffentlicht: (2026)
von: Zhao, Rui, et al.
Veröffentlicht: (2026)
Semantic Substrate Theory: An Operator-Theoretic Framework for Geometric Semantic Drift
von: Russell, Stephen
Veröffentlicht: (2026)
von: Russell, Stephen
Veröffentlicht: (2026)
CoT-Space: A Theoretical Framework for Internal Slow-Thinking via Reinforcement Learning
von: Gan, Zeyu, et al.
Veröffentlicht: (2025)
von: Gan, Zeyu, et al.
Veröffentlicht: (2025)
PromptWizard: Task-Aware Prompt Optimization Framework
von: Agarwal, Eshaan, et al.
Veröffentlicht: (2024)
von: Agarwal, Eshaan, et al.
Veröffentlicht: (2024)
Quamba2: A Robust and Scalable Post-training Quantization Framework for Selective State Space Models
von: Chiang, Hung-Yueh, et al.
Veröffentlicht: (2025)
von: Chiang, Hung-Yueh, et al.
Veröffentlicht: (2025)
Evaluating Consistency and Reasoning Capabilities of Large Language Models
von: Saxena, Yash, et al.
Veröffentlicht: (2024)
von: Saxena, Yash, et al.
Veröffentlicht: (2024)
Distilling Mathematical Reasoning Capabilities into Small Language Models
von: Zhu, Xunyu, et al.
Veröffentlicht: (2024)
von: Zhu, Xunyu, et al.
Veröffentlicht: (2024)
On the Use of Large Language Models to Generate Capability Ontologies
von: da Silva, Luis Miguel Vieira, et al.
Veröffentlicht: (2024)
von: da Silva, Luis Miguel Vieira, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The Rotary Position Embedding May Cause Dimension Inefficiency in Attention Heads for Long-Distance Retrieval
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2025) -
LocateBench: Evaluating the Locating Ability of Vision Language Models
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2024) -
DeLLMa: Decision Making Under Uncertainty with Large Language Models
von: Liu, Ollie, et al.
Veröffentlicht: (2024) -
Causal Interventions on Causal Paths: Mapping GPT-2's Reasoning From Syntax to Semantics
von: Lee, Isabelle, et al.
Veröffentlicht: (2024) -
Bone Soups: A Seek-and-Soup Model Merging Approach for Controllable Multi-Objective Generation
von: Xie, Guofu, et al.
Veröffentlicht: (2025)