Product of Experts with LLMs: Boosting Performance on ARC Is a Matter of Perspective
Fuente:
arXiv
Salvato in:
| Autori principali: | Franzen, Daniel, Disselhoff, Jan, Hartmann, David |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ARC-TGI: Human-Validated Task Generators with Reasoning Chain Templates for ARC-AGI
di: Lehmann, Jens, et al.
Pubblicazione: (2026)
di: Lehmann, Jens, et al.
Pubblicazione: (2026)
Capturing Sparks of Abstraction for the ARC Challenge
di: Andrews, Martin
Pubblicazione: (2024)
di: Andrews, Martin
Pubblicazione: (2024)
CausalARC: Abstract Reasoning with Causal World Models
di: Maasch, Jacqueline, et al.
Pubblicazione: (2025)
di: Maasch, Jacqueline, et al.
Pubblicazione: (2025)
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
di: Zhuang, Haomin, et al.
Pubblicazione: (2024)
di: Zhuang, Haomin, et al.
Pubblicazione: (2024)
Every Expert Matters: Towards Effective Knowledge Distillation for Mixture-of-Experts Language Models
di: Kim, Gyeongman, et al.
Pubblicazione: (2025)
di: Kim, Gyeongman, et al.
Pubblicazione: (2025)
Towards Efficient Neurally-Guided Program Induction for ARC-AGI
di: Ouellette, Simon
Pubblicazione: (2024)
di: Ouellette, Simon
Pubblicazione: (2024)
Unveiling Downstream Performance Scaling of LLMs: A Clustering-Based Perspective
di: Xu, Chengyin, et al.
Pubblicazione: (2025)
di: Xu, Chengyin, et al.
Pubblicazione: (2025)
Dynamic Experts Search: Enhancing Reasoning in Mixture-of-Experts LLMs at Test Time
di: Han, Yixuan, et al.
Pubblicazione: (2025)
di: Han, Yixuan, et al.
Pubblicazione: (2025)
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
di: Wei, Anjiang, et al.
Pubblicazione: (2025)
di: Wei, Anjiang, et al.
Pubblicazione: (2025)
Evaluating LLMs on Real-World Forecasting Against Expert Forecasters
di: Lu, Janna
Pubblicazione: (2025)
di: Lu, Janna
Pubblicazione: (2025)
AQuA -- Combining Experts' and Non-Experts' Views To Assess Deliberation Quality in Online Discussions Using LLMs
di: Behrendt, Maike, et al.
Pubblicazione: (2024)
di: Behrendt, Maike, et al.
Pubblicazione: (2024)
RAG-Instruct: Boosting LLMs with Diverse Retrieval-Augmented Instructions
di: Liu, Wanlong, et al.
Pubblicazione: (2024)
di: Liu, Wanlong, et al.
Pubblicazione: (2024)
Knowledge Localization in Mixture-of-Experts LLMs Using Cross-Lingual Inconsistency
di: Bandarkar, Lucas, et al.
Pubblicazione: (2026)
di: Bandarkar, Lucas, et al.
Pubblicazione: (2026)
Boosting In-Context Learning in LLMs Through the Lens of Classical Supervised Learning
di: Gundem, Korel, et al.
Pubblicazione: (2025)
di: Gundem, Korel, et al.
Pubblicazione: (2025)
DistiLLM-2: A Contrastive Approach Boosts the Distillation of LLMs
di: Ko, Jongwoo, et al.
Pubblicazione: (2025)
di: Ko, Jongwoo, et al.
Pubblicazione: (2025)
REAM: Merging Improves Pruning of Experts in LLMs
di: Jha, Saurav, et al.
Pubblicazione: (2026)
di: Jha, Saurav, et al.
Pubblicazione: (2026)
Leveraging LLM Inconsistency to Boost Pass@k Performance
di: Dalal, Uri, et al.
Pubblicazione: (2025)
di: Dalal, Uri, et al.
Pubblicazione: (2025)
FlyLoRA: Boosting Task Decoupling and Parameter Efficiency via Implicit Rank-Wise Mixture-of-Experts
di: Zou, Heming, et al.
Pubblicazione: (2025)
di: Zou, Heming, et al.
Pubblicazione: (2025)
On the Spatial Structure of Mixture-of-Experts in Transformers
di: Bershatsky, Daniel, et al.
Pubblicazione: (2025)
di: Bershatsky, Daniel, et al.
Pubblicazione: (2025)
GaLore$+$: Boosting Low-Rank Adaptation for LLMs with Cross-Head Projection
di: Liao, Xutao, et al.
Pubblicazione: (2024)
di: Liao, Xutao, et al.
Pubblicazione: (2024)
Configurable Foundation Models: Building LLMs from a Modular Perspective
di: Xiao, Chaojun, et al.
Pubblicazione: (2024)
di: Xiao, Chaojun, et al.
Pubblicazione: (2024)
Mixture-of-Experts as Soft Clustering: A Dual Jacobian-PCA Spectral Geometry Perspective
di: Liu, Feilong
Pubblicazione: (2026)
di: Liu, Feilong
Pubblicazione: (2026)
A Decomposition Perspective to Long-context Reasoning for LLMs
di: Xiao, Yanling, et al.
Pubblicazione: (2026)
di: Xiao, Yanling, et al.
Pubblicazione: (2026)
Mission Impossible: A Statistical Perspective on Jailbreaking LLMs
di: Su, Jingtong, et al.
Pubblicazione: (2024)
di: Su, Jingtong, et al.
Pubblicazione: (2024)
Less is More: Extreme Gradient Boost Rank-1 Adaption for Efficient Finetuning of LLMs
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
On the Performance of LLMs for Real Estate Appraisal
di: Geerts, Margot, et al.
Pubblicazione: (2025)
di: Geerts, Margot, et al.
Pubblicazione: (2025)
Principled RL for Diffusion LLMs Emerges from a Sequence-Level Perspective
di: Ou, Jingyang, et al.
Pubblicazione: (2025)
di: Ou, Jingyang, et al.
Pubblicazione: (2025)
Profiling News Media for Factuality and Bias Using LLMs and the Fact-Checking Methodology of Human Experts
di: Mujahid, Zain Muhammad, et al.
Pubblicazione: (2025)
di: Mujahid, Zain Muhammad, et al.
Pubblicazione: (2025)
Shifting Perspectives: Steering Vectors for Robust Bias Mitigation in LLMs
di: Siddique, Zara, et al.
Pubblicazione: (2025)
di: Siddique, Zara, et al.
Pubblicazione: (2025)
Are the Values of LLMs Structurally Aligned with Humans? A Causal Perspective
di: Kang, Yipeng, et al.
Pubblicazione: (2024)
di: Kang, Yipeng, et al.
Pubblicazione: (2024)
Scaling Laws for Predicting Downstream Performance in LLMs
di: Chen, Yangyi, et al.
Pubblicazione: (2024)
di: Chen, Yangyi, et al.
Pubblicazione: (2024)
Bridging the Creativity Understanding Gap: Small-Scale Human Alignment Enables Expert-Level Humor Ranking in LLMs
di: Zhou, Kuan Lok, et al.
Pubblicazione: (2025)
di: Zhou, Kuan Lok, et al.
Pubblicazione: (2025)
Joint MoE Scaling Laws: Mixture of Experts Can Be Memory Efficient
di: Ludziejewski, Jan, et al.
Pubblicazione: (2025)
di: Ludziejewski, Jan, et al.
Pubblicazione: (2025)
Assembly of Experts: Linear-time construction of the Chimera LLM variants with emergent and adaptable behaviors
di: Klagges, Henrik, et al.
Pubblicazione: (2025)
di: Klagges, Henrik, et al.
Pubblicazione: (2025)
Not All Experts are Equal: Efficient Expert Pruning and Skipping for Mixture-of-Experts Large Language Models
di: Lu, Xudong, et al.
Pubblicazione: (2024)
di: Lu, Xudong, et al.
Pubblicazione: (2024)
Scaling Laws for Fine-Grained Mixture of Experts
di: Krajewski, Jakub, et al.
Pubblicazione: (2024)
di: Krajewski, Jakub, et al.
Pubblicazione: (2024)
Beyond Confidence: Rethinking Self-Assessments for Performance Prediction in LLMs
di: Bhattacharyya, Sree, et al.
Pubblicazione: (2026)
di: Bhattacharyya, Sree, et al.
Pubblicazione: (2026)
Instruction Learning Paradigms: A Dual Perspective on White-box and Black-box LLMs
di: Ren, Yanwei, et al.
Pubblicazione: (2025)
di: Ren, Yanwei, et al.
Pubblicazione: (2025)
How Data Inter-connectivity Shapes LLMs Unlearning: A Structural Unlearning Perspective
di: Qiu, Xinchi, et al.
Pubblicazione: (2024)
di: Qiu, Xinchi, et al.
Pubblicazione: (2024)
Self-Distillation as a Performance Recovery Mechanism for LLMs: Counteracting Compression and Catastrophic Forgetting
di: Liu, Chi, et al.
Pubblicazione: (2026)
di: Liu, Chi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
ARC-TGI: Human-Validated Task Generators with Reasoning Chain Templates for ARC-AGI
di: Lehmann, Jens, et al.
Pubblicazione: (2026) -
Capturing Sparks of Abstraction for the ARC Challenge
di: Andrews, Martin
Pubblicazione: (2024) -
CausalARC: Abstract Reasoning with Causal World Models
di: Maasch, Jacqueline, et al.
Pubblicazione: (2025) -
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
di: Zhuang, Haomin, et al.
Pubblicazione: (2024) -
Every Expert Matters: Towards Effective Knowledge Distillation for Mixture-of-Experts Language Models
di: Kim, Gyeongman, et al.
Pubblicazione: (2025)