Quantifying Generalization Complexity for Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Qi, Zhenting, Luo, Hongyin, Huang, Xuliang, Zhao, Zhuokai, Jiang, Yibo, Fan, Xiangjun, Lakkaraju, Himabindu, Glass, James |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Manipulating Large Language Models to Increase Product Visibility
di: Kumar, Aounon, et al.
Pubblicazione: (2024)
di: Kumar, Aounon, et al.
Pubblicazione: (2024)
Towards Uncovering How Large Language Model Works: An Explainability Perspective
di: Zhao, Haiyan, et al.
Pubblicazione: (2024)
di: Zhao, Haiyan, et al.
Pubblicazione: (2024)
Follow My Instruction and Spill the Beans: Scalable Data Extraction from Retrieval-Augmented Generation Systems
di: Qi, Zhenting, et al.
Pubblicazione: (2024)
di: Qi, Zhenting, et al.
Pubblicazione: (2024)
Self-Improving Language Models with Bidirectional Evolutionary Search
di: Xu, Guowei, et al.
Pubblicazione: (2026)
di: Xu, Guowei, et al.
Pubblicazione: (2026)
Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models
di: Agarwal, Chirag, et al.
Pubblicazione: (2024)
di: Agarwal, Chirag, et al.
Pubblicazione: (2024)
EvoLM: In Search of Lost Language Model Training Dynamics
di: Qi, Zhenting, et al.
Pubblicazione: (2025)
di: Qi, Zhenting, et al.
Pubblicazione: (2025)
On the Hardness of Faithful Chain-of-Thought Reasoning in Large Language Models
di: Tanneru, Sree Harsha, et al.
Pubblicazione: (2024)
di: Tanneru, Sree Harsha, et al.
Pubblicazione: (2024)
Confronting LLMs with Traditional ML: Rethinking the Fairness of Large Language Models in Tabular Classifications
di: Liu, Yanchen, et al.
Pubblicazione: (2023)
di: Liu, Yanchen, et al.
Pubblicazione: (2023)
Understanding the Effects of Iterative Prompting on Truthfulness
di: Krishna, Satyapriya, et al.
Pubblicazione: (2024)
di: Krishna, Satyapriya, et al.
Pubblicazione: (2024)
On the Impact of Fine-Tuning on Chain-of-Thought Reasoning
di: Lobo, Elita, et al.
Pubblicazione: (2024)
di: Lobo, Elita, et al.
Pubblicazione: (2024)
THREAD: Thinking Deeper with Recursive Spawning
di: Schroeder, Philip, et al.
Pubblicazione: (2024)
di: Schroeder, Philip, et al.
Pubblicazione: (2024)
Measuring the Faithfulness of Thinking Drafts in Large Reasoning Models
di: Xiong, Zidi, et al.
Pubblicazione: (2025)
di: Xiong, Zidi, et al.
Pubblicazione: (2025)
ROVER: Recursive Reasoning Over Videos with Vision-Language Models for Embodied Tasks
di: Schroeder, Philip, et al.
Pubblicazione: (2025)
di: Schroeder, Philip, et al.
Pubblicazione: (2025)
DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models
di: Chuang, Yung-Sung, et al.
Pubblicazione: (2023)
di: Chuang, Yung-Sung, et al.
Pubblicazione: (2023)
Generalizing Trust: Weak-to-Strong Trustworthiness in Language Models
di: Pawelczyk, Martin, et al.
Pubblicazione: (2024)
di: Pawelczyk, Martin, et al.
Pubblicazione: (2024)
More RLHF, More Trust? On The Impact of Preference Alignment On Trustworthiness
di: Li, Aaron J., et al.
Pubblicazione: (2024)
di: Li, Aaron J., et al.
Pubblicazione: (2024)
Decoding on Graphs: Faithful and Sound Reasoning on Knowledge Graphs through Generation of Well-Formed Chains
di: Li, Kun, et al.
Pubblicazione: (2024)
di: Li, Kun, et al.
Pubblicazione: (2024)
Challenges and Responses in the Practice of Large Language Models
di: Zhu, Hongyin
Pubblicazione: (2024)
di: Zhu, Hongyin
Pubblicazione: (2024)
Architectural Foundations for the Large Language Model Infrastructures
di: Zhu, Hongyin
Pubblicazione: (2024)
di: Zhu, Hongyin
Pubblicazione: (2024)
Self-Specialization: Uncovering Latent Expertise within Large Language Models
di: Kang, Junmo, et al.
Pubblicazione: (2023)
di: Kang, Junmo, et al.
Pubblicazione: (2023)
Addition is All You Need for Energy-efficient Language Models
di: Luo, Hongyin, et al.
Pubblicazione: (2024)
di: Luo, Hongyin, et al.
Pubblicazione: (2024)
Climate Change from Large Language Models
di: Zhu, Hongyin, et al.
Pubblicazione: (2023)
di: Zhu, Hongyin, et al.
Pubblicazione: (2023)
Interpretability Needs a New Paradigm
di: Madsen, Andreas, et al.
Pubblicazione: (2024)
di: Madsen, Andreas, et al.
Pubblicazione: (2024)
Self-MoE: Towards Compositional Large Language Models with Self-Specialized Experts
di: Kang, Junmo, et al.
Pubblicazione: (2024)
di: Kang, Junmo, et al.
Pubblicazione: (2024)
Adaptive Query Rewriting: Aligning Rewriters through Marginal Probability of Conversational Answers
di: Zhang, Tianhua, et al.
Pubblicazione: (2024)
di: Zhang, Tianhua, et al.
Pubblicazione: (2024)
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
di: Kroeger, Nicholas, et al.
Pubblicazione: (2023)
di: Kroeger, Nicholas, et al.
Pubblicazione: (2023)
Evaluating Adversarial Robustness of Concept Representations in Sparse Autoencoders
di: Li, Aaron J., et al.
Pubblicazione: (2025)
di: Li, Aaron J., et al.
Pubblicazione: (2025)
Generate, Discriminate, Evolve: Enhancing Context Faithfulness via Fine-Grained Sentence-Level Self-Evolution
di: Li, Kun, et al.
Pubblicazione: (2025)
di: Li, Kun, et al.
Pubblicazione: (2025)
Temporal Sparse Autoencoders: Leveraging the Sequential Nature of Language for Interpretability
di: Bhalla, Usha, et al.
Pubblicazione: (2025)
di: Bhalla, Usha, et al.
Pubblicazione: (2025)
A Single Layer to Explain Them All:Understanding Massive Activations in Large Language Models
di: Shi, Zeru, et al.
Pubblicazione: (2026)
di: Shi, Zeru, et al.
Pubblicazione: (2026)
RAG-Zeval: Towards Robust and Interpretable Evaluation on RAG Responses through End-to-End Rule-Guided Reasoning
di: Li, Kun, et al.
Pubblicazione: (2025)
di: Li, Kun, et al.
Pubblicazione: (2025)
Synthetic Sandbox for Training Machine Learning Engineering Agents
di: Zhou, Yuhang, et al.
Pubblicazione: (2026)
di: Zhou, Yuhang, et al.
Pubblicazione: (2026)
Natural Language Embedded Programs for Hybrid Language Symbolic Reasoning
di: Zhang, Tianhua, et al.
Pubblicazione: (2023)
di: Zhang, Tianhua, et al.
Pubblicazione: (2023)
Learning Recourse Costs from Pairwise Feature Comparisons
di: Rawal, Kaivalya, et al.
Pubblicazione: (2024)
di: Rawal, Kaivalya, et al.
Pubblicazione: (2024)
TreePS-RAG: Tree-based Process Supervision for Reinforcement Learning in Agentic RAG
di: Zhang, Tianhua, et al.
Pubblicazione: (2026)
di: Zhang, Tianhua, et al.
Pubblicazione: (2026)
Unmasking and Quantifying Racial Bias of Large Language Models in Medical Report Generation
di: Yang, Yifan, et al.
Pubblicazione: (2024)
di: Yang, Yifan, et al.
Pubblicazione: (2024)
OmniOPD: Logit-Free On-Policy Distillation via Speculative Verification
di: Zhou, Yuhang, et al.
Pubblicazione: (2026)
di: Zhou, Yuhang, et al.
Pubblicazione: (2026)
CAFe: Unifying Representation and Generation with Contrastive-Autoregressive Finetuning
di: Yu, Hao, et al.
Pubblicazione: (2025)
di: Yu, Hao, et al.
Pubblicazione: (2025)
S'MoRE: Structural Mixture of Residual Experts for Parameter-Efficient LLM Fine-tuning
di: Zeng, Hanqing, et al.
Pubblicazione: (2025)
di: Zeng, Hanqing, et al.
Pubblicazione: (2025)
On the Generalization Gap in Self-Evolving Language Model Reasoning
di: Qi, Zhenting, et al.
Pubblicazione: (2026)
di: Qi, Zhenting, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Manipulating Large Language Models to Increase Product Visibility
di: Kumar, Aounon, et al.
Pubblicazione: (2024) -
Towards Uncovering How Large Language Model Works: An Explainability Perspective
di: Zhao, Haiyan, et al.
Pubblicazione: (2024) -
Follow My Instruction and Spill the Beans: Scalable Data Extraction from Retrieval-Augmented Generation Systems
di: Qi, Zhenting, et al.
Pubblicazione: (2024) -
Self-Improving Language Models with Bidirectional Evolutionary Search
di: Xu, Guowei, et al.
Pubblicazione: (2026) -
Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models
di: Agarwal, Chirag, et al.
Pubblicazione: (2024)