ZeroTuning: Unlocking the Initial Token's Power to Enhance Large Language Models Without Training
Fuente:
arXiv
Guardado en:
| Autores principales: | Han, Feijiang, Yu, Xiaodong, Tang, Jianheng, Rao, Delip, Du, Weihua, Ungar, Lyle |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ThinknCheck: Grounded Claim Verification with Compact, Reasoning-Driven, and Interpretable Models
por: Rao, Delip, et al.
Publicado: (2026)
por: Rao, Delip, et al.
Publicado: (2026)
Closing the Confidence-Faithfulness Gap in Large Language Models
por: Miao, Miranda Muqing, et al.
Publicado: (2026)
por: Miao, Miranda Muqing, et al.
Publicado: (2026)
What Do Claim Verification Datasets Actually Test? A Reasoning Trace Analysis
por: Rao, Delip, et al.
Publicado: (2026)
por: Rao, Delip, et al.
Publicado: (2026)
Autorubric: Unifying Rubric-based LLM Evaluation
por: Rao, Delip, et al.
Publicado: (2026)
por: Rao, Delip, et al.
Publicado: (2026)
BibTeX Citation Hallucinations in Scientific Publishing Agents: Evaluation and Mitigation
por: Rao, Delip, et al.
Publicado: (2026)
por: Rao, Delip, et al.
Publicado: (2026)
Agreement Metrics for LLM-as-Judge Evaluation: What to Report and Why
por: Rao, Delip, et al.
Publicado: (2026)
por: Rao, Delip, et al.
Publicado: (2026)
Detecting and Correcting Reference Hallucinations in Commercial LLMs and Deep Research Agents
por: Rao, Delip, et al.
Publicado: (2026)
por: Rao, Delip, et al.
Publicado: (2026)
Unlocking Full Efficiency of Token Filtering in Large Language Model Training
por: Chai, Di, et al.
Publicado: (2025)
por: Chai, Di, et al.
Publicado: (2025)
WithdrarXiv: A Large-Scale Dataset for Retraction Study
por: Rao, Delip, et al.
Publicado: (2024)
por: Rao, Delip, et al.
Publicado: (2024)
INTERS: Unlocking the Power of Large Language Models in Search with Instruction Tuning
por: Zhu, Yutao, et al.
Publicado: (2024)
por: Zhu, Yutao, et al.
Publicado: (2024)
NSF-SciFy: Mining the NSF Awards Database for Scientific Claims
por: Rao, Delip, et al.
Publicado: (2025)
por: Rao, Delip, et al.
Publicado: (2025)
Comparing Styles across Languages: A Cross-Cultural Exploration of Politeness
por: Havaldar, Shreya, et al.
Publicado: (2023)
por: Havaldar, Shreya, et al.
Publicado: (2023)
How does Misinformation Affect Large Language Model Behaviors and Preferences?
por: Peng, Miao, et al.
Publicado: (2025)
por: Peng, Miao, et al.
Publicado: (2025)
Re-Initialization Token Learning for Tool-Augmented Large Language Models
por: Li, Chenghao, et al.
Publicado: (2025)
por: Li, Chenghao, et al.
Publicado: (2025)
When Verification Fails: How Compositionally Infeasible Claims Escape Rejection
por: Liu, Muxin, et al.
Publicado: (2026)
por: Liu, Muxin, et al.
Publicado: (2026)
Unlocking the Power of Large Language Models for Multi-table Entity Matching
por: Tang, Yingkai, et al.
Publicado: (2026)
por: Tang, Yingkai, et al.
Publicado: (2026)
Enhancing Large Language Model Reasoning via Selective Critical Token Fine-Tuning
por: Ruan, Zhiwen, et al.
Publicado: (2025)
por: Ruan, Zhiwen, et al.
Publicado: (2025)
Inverse-Q*: Token Level Reinforcement Learning for Aligning Large Language Models Without Preference Data
por: Xia, Han, et al.
Publicado: (2024)
por: Xia, Han, et al.
Publicado: (2024)
GraphArena: Evaluating and Exploring Large Language Models on Graph Computation
por: Tang, Jianheng, et al.
Publicado: (2024)
por: Tang, Jianheng, et al.
Publicado: (2024)
Enhancing Large Language Models for Mobility Analytics with Semantic Location Tokenization
por: Chen, Yile, et al.
Publicado: (2025)
por: Chen, Yile, et al.
Publicado: (2025)
Estimating Knowledge in Large Language Models Without Generating a Single Token
por: Gottesman, Daniela, et al.
Publicado: (2024)
por: Gottesman, Daniela, et al.
Publicado: (2024)
Unlocking the Power of Large Language Models for Entity Alignment
por: Jiang, Xuhui, et al.
Publicado: (2024)
por: Jiang, Xuhui, et al.
Publicado: (2024)
Can LLMs Handle WebShell Detection? Overcoming Detection Challenges with Behavioral Function-Aware Framework
por: Han, Feijiang, et al.
Publicado: (2025)
por: Han, Feijiang, et al.
Publicado: (2025)
GraphWiz: An Instruction-Following Language Model for Graph Problems
por: Chen, Nuo, et al.
Publicado: (2024)
por: Chen, Nuo, et al.
Publicado: (2024)
Supervised Fine-Tuning Needs to Unlock the Potential of Token Priority
por: Shen, Zhanming, et al.
Publicado: (2026)
por: Shen, Zhanming, et al.
Publicado: (2026)
DiverseDialogue: A Methodology for Designing Chatbots with Human-Like Diversity
por: Lin, Xiaoyu, et al.
Publicado: (2024)
por: Lin, Xiaoyu, et al.
Publicado: (2024)
PACIT: Unlocking the Power of Examples for Better In-Context Instruction Tuning
por: Xue, Tianci, et al.
Publicado: (2023)
por: Xue, Tianci, et al.
Publicado: (2023)
Zero Token-Driven Deep Thinking in LLMs: Unlocking the Full Potential of Existing Parameters via Cyclic Refinement
por: Li, Guanghao, et al.
Publicado: (2025)
por: Li, Guanghao, et al.
Publicado: (2025)
Towards Style Alignment in Cross-Cultural Translation
por: Havaldar, Shreya, et al.
Publicado: (2025)
por: Havaldar, Shreya, et al.
Publicado: (2025)
GCoder: Improving Large Language Model for Generalized Graph Problem Solving
por: Zhang, Qifan, et al.
Publicado: (2024)
por: Zhang, Qifan, et al.
Publicado: (2024)
SEAP: Training-free Sparse Expert Activation Pruning Unlock the Brainpower of Large Language Models
por: Liang, Xun, et al.
Publicado: (2025)
por: Liang, Xun, et al.
Publicado: (2025)
Trans-Zero: Self-Play Incentivizes Large Language Models for Multilingual Translation Without Parallel Data
por: Zou, Wei, et al.
Publicado: (2025)
por: Zou, Wei, et al.
Publicado: (2025)
A Learning Rate Path Switching Training Paradigm for Version Updates of Large Language Models
por: Wang, Zhihao, et al.
Publicado: (2024)
por: Wang, Zhihao, et al.
Publicado: (2024)
BitLM: Unlocking Multi-Token Language Generation with Bitwise Continuous Diffusion
por: Zhuang, Shaobin, et al.
Publicado: (2026)
por: Zhuang, Shaobin, et al.
Publicado: (2026)
Self-Training Large Language Models for Tool-Use Without Demonstrations
por: Luo, Ne, et al.
Publicado: (2025)
por: Luo, Ne, et al.
Publicado: (2025)
Interactive Concept Learning for Uncovering Latent Themes in Large Text Collections
por: Pacheco, Maria Leonor, et al.
Publicado: (2023)
por: Pacheco, Maria Leonor, et al.
Publicado: (2023)
The Impact of Language Mixing on Bilingual LLM Reasoning
por: Li, Yihao, et al.
Publicado: (2025)
por: Li, Yihao, et al.
Publicado: (2025)
Culturally-Aware Conversations: A Framework & Benchmark for LLMs
por: Havaldar, Shreya, et al.
Publicado: (2025)
por: Havaldar, Shreya, et al.
Publicado: (2025)
Harnessing the Zero-Shot Power of Instruction-Tuned Large Language Model in End-to-End Speech Recognition
por: Higuchi, Yosuke, et al.
Publicado: (2023)
por: Higuchi, Yosuke, et al.
Publicado: (2023)
Unveiling the Generalization Power of Fine-Tuned Large Language Models
por: Yang, Haoran, et al.
Publicado: (2024)
por: Yang, Haoran, et al.
Publicado: (2024)
Ejemplares similares
-
ThinknCheck: Grounded Claim Verification with Compact, Reasoning-Driven, and Interpretable Models
por: Rao, Delip, et al.
Publicado: (2026) -
Closing the Confidence-Faithfulness Gap in Large Language Models
por: Miao, Miranda Muqing, et al.
Publicado: (2026) -
What Do Claim Verification Datasets Actually Test? A Reasoning Trace Analysis
por: Rao, Delip, et al.
Publicado: (2026) -
Autorubric: Unifying Rubric-based LLM Evaluation
por: Rao, Delip, et al.
Publicado: (2026) -
BibTeX Citation Hallucinations in Scientific Publishing Agents: Evaluation and Mitigation
por: Rao, Delip, et al.
Publicado: (2026)