Learning Compact Representations of LLM Abilities via Item Response Theory
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Jianhao, Wang, Chenxu, Zhang, Gengrui, Ye, Peng, Bai, Lei, Hu, Wei, Qu, Yuzhong, Hu, Shuyue |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ICL-Router: In-Context Learned Model Representations for LLM Routing
by: Wang, Chenxu, et al.
Published: (2025)
by: Wang, Chenxu, et al.
Published: (2025)
Do We Truly Need So Many Samples? Multi-LLM Repeated Sampling Efficiently Scales Test-Time Compute
by: Chen, Jianhao, et al.
Published: (2025)
by: Chen, Jianhao, et al.
Published: (2025)
Large Language Models are Near-Optimal Decision-Makers with a Non-Human Learning Behavior
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Timeline-based Sentence Decomposition with In-Context Learning for Temporal Fact Extraction
by: Chen, Jianhao, et al.
Published: (2024)
by: Chen, Jianhao, et al.
Published: (2024)
Conflict Detection for Temporal Knowledge Graphs:A Fast Constraint Mining Algorithm and New Benchmarks
by: Chen, Jianhao, et al.
Published: (2023)
by: Chen, Jianhao, et al.
Published: (2023)
LLMRouterBench: A Massive Benchmark and Unified Framework for LLM Routing
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
JE-IRT: A Geometric Lens on LLM Abilities through Joint Embedding Item Response Theory
by: Yao, Louie Hong, et al.
Published: (2025)
by: Yao, Louie Hong, et al.
Published: (2025)
Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory
by: Choi, Junhyuk, et al.
Published: (2026)
by: Choi, Junhyuk, et al.
Published: (2026)
IRT-Router: Effective and Interpretable Multi-LLM Routing via Item Response Theory
by: Song, Wei, et al.
Published: (2025)
by: Song, Wei, et al.
Published: (2025)
Beyond Gemini-3-Pro: Revisiting LLM Routing and Aggregation at Scale
by: Tang, Shengji, et al.
Published: (2026)
by: Tang, Shengji, et al.
Published: (2026)
LLM-Empowered Representation Learning for Emerging Item Recommendation
by: Zhang, Ziying, et al.
Published: (2025)
by: Zhang, Ziying, et al.
Published: (2025)
FITRep: Attention-Guided Item Representation via MLLMs
by: Zhang, Guoxiao, et al.
Published: (2025)
by: Zhang, Guoxiao, et al.
Published: (2025)
Adaptive Theory of Mind for LLM-based Multi-Agent Coordination
by: Mu, Chunjiang, et al.
Published: (2026)
by: Mu, Chunjiang, et al.
Published: (2026)
A Scalable Multi-LLM Collaboration System with Retrieval-based Selection and Exploration-Exploitation-Driven Enhancement
by: Tang, Shengji, et al.
Published: (2025)
by: Tang, Shengji, et al.
Published: (2025)
Truly Assessing Fluid Intelligence of Large Language Models through Dynamic Reasoning Evaluation
by: Yang, Yue, et al.
Published: (2025)
by: Yang, Yue, et al.
Published: (2025)
Single-Agent Scaling Fails Multi-Agent Intelligence: Towards Foundation Models with Native Multi-Agent Intelligence
by: Hu, Shuyue, et al.
Published: (2025)
by: Hu, Shuyue, et al.
Published: (2025)
MTRouter: Cost-Aware Multi-Turn LLM Routing with History-Model Joint Embeddings
by: Zhang, Yiqun, et al.
Published: (2026)
by: Zhang, Yiqun, et al.
Published: (2026)
Token-Efficient Item Representation via Images for LLM Recommender Systems
by: Kim, Kibum, et al.
Published: (2025)
by: Kim, Kibum, et al.
Published: (2025)
Skin Tokens: A Learned Compact Representation for Unified Autoregressive Rigging
by: Zhang, Jia-peng, et al.
Published: (2026)
by: Zhang, Jia-peng, et al.
Published: (2026)
The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
by: Schmucker, Robin, et al.
Published: (2025)
by: Schmucker, Robin, et al.
Published: (2025)
Predicting LLM Output Length via Entropy-Guided Representations
by: Xie, Huanyi, et al.
Published: (2026)
by: Xie, Huanyi, et al.
Published: (2026)
PAPO: Stabilizing Rubric Integration Training via Decoupled Advantage Normalization
by: Tan, Zelin, et al.
Published: (2026)
by: Tan, Zelin, et al.
Published: (2026)
Learning Partially Aligned Item Representation for Cross-Domain Sequential Recommendation
by: Yin, Mingjia, et al.
Published: (2024)
by: Yin, Mingjia, et al.
Published: (2024)
Multi-modal Relational Item Representation Learning for Inferring Substitutable and Complementary Items
by: Wang, Junting, et al.
Published: (2025)
by: Wang, Junting, et al.
Published: (2025)
PerPilot: Personalizing VLM-based Mobile Agents via Memory and Exploration
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
Towards Effective Theory of LLMs: A Representation Learning Approach
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
Learning Hierarchical and Geometry-Aware Graph Representations for Text-to-CAD
by: Gong, Shengjie, et al.
Published: (2026)
by: Gong, Shengjie, et al.
Published: (2026)
EmbedLLM: Learning Compact Representations of Large Language Models
by: Zhuang, Richard, et al.
Published: (2024)
by: Zhuang, Richard, et al.
Published: (2024)
InnovatorBench: Evaluating Agents' Ability to Conduct Innovative LLM Research
by: Wu, Yunze, et al.
Published: (2025)
by: Wu, Yunze, et al.
Published: (2025)
LLM-Empowered State Representation for Reinforcement Learning
by: Wang, Boyuan, et al.
Published: (2024)
by: Wang, Boyuan, et al.
Published: (2024)
Automated urban waterlogging assessment and early warning through a mixture of foundation models
by: Zhang, Chenxu, et al.
Published: (2025)
by: Zhang, Chenxu, et al.
Published: (2025)
PhysicsMinions: Winning Gold Medals in the Latest Physics Olympiads with a Coevolutionary Multimodal Multi-Agent System
by: Yu, Fangchen, et al.
Published: (2025)
by: Yu, Fangchen, et al.
Published: (2025)
Diversity-Incentivized Exploration for Versatile Reasoning
by: Hu, Zican, et al.
Published: (2025)
by: Hu, Zican, et al.
Published: (2025)
Beyond Scores: Diagnostic LLM Evaluation via Fine-Grained Abilities
by: Zhang, Xu, et al.
Published: (2026)
by: Zhang, Xu, et al.
Published: (2026)
Leveraging Natural Language and Item Response Theory Models for ESG Scoring
by: Soares, César Pedrosa
Published: (2024)
by: Soares, César Pedrosa
Published: (2024)
Enhancing Essay Cohesion Assessment: A Novel Item Response Theory Approach
by: Rosa, Bruno Alexandre, et al.
Published: (2025)
by: Rosa, Bruno Alexandre, et al.
Published: (2025)
AgentDR: Dynamic Recommendation with Implicit Item-Item Relations via LLM-based Agents
by: Yang, Mingdai, et al.
Published: (2025)
by: Yang, Mingdai, et al.
Published: (2025)
Reputation as a Solution to Cooperation Collapse in LLM-based MASs
by: Ren, Siyue, et al.
Published: (2025)
by: Ren, Siyue, et al.
Published: (2025)
LLM-EvRep: Learning an LLM-Compatible Event Representation Using a Self-Supervised Framework
by: Yu, Zongyou, et al.
Published: (2025)
by: Yu, Zongyou, et al.
Published: (2025)
FCN-LLM: Empower LLM for Brain Functional Connectivity Network Understanding via Graph-level Multi-task Instruction Tuning
by: Hu, Xingcan, et al.
Published: (2026)
by: Hu, Xingcan, et al.
Published: (2026)
Similar Items
-
ICL-Router: In-Context Learned Model Representations for LLM Routing
by: Wang, Chenxu, et al.
Published: (2025) -
Do We Truly Need So Many Samples? Multi-LLM Repeated Sampling Efficiently Scales Test-Time Compute
by: Chen, Jianhao, et al.
Published: (2025) -
Large Language Models are Near-Optimal Decision-Makers with a Non-Human Learning Behavior
by: Li, Hao, et al.
Published: (2025) -
Timeline-based Sentence Decomposition with In-Context Learning for Temporal Fact Extraction
by: Chen, Jianhao, et al.
Published: (2024) -
Conflict Detection for Temporal Knowledge Graphs:A Fast Constraint Mining Algorithm and New Benchmarks
by: Chen, Jianhao, et al.
Published: (2023)