Guardado en:
| Autores principales: | Zhuang, Nan, Wang, Wenshuo, Qian, Lekai, Wang, Yuxiao, Cao, Boyu, Liu, Qi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2512.03082 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
iTAG: Inverse Design for Natural Text Generation with Accurate Causal Graph Annotations
por: Wang, Wenshuo, et al.
Publicado: (2026)
por: Wang, Wenshuo, et al.
Publicado: (2026)
Anchored Cyclic Generation: A Novel Paradigm for Long-Sequence Symbolic Music Generation
por: Cao, Boyu, et al.
Publicado: (2026)
por: Cao, Boyu, et al.
Publicado: (2026)
Does Reasoning Introduce Bias? A Study of Social Bias Evaluation and Mitigation in LLM Reasoning
por: Wu, Xuyang, et al.
Publicado: (2025)
por: Wu, Xuyang, et al.
Publicado: (2025)
Steering LLM Thinking with Budget Guidance
por: Li, Junyan, et al.
Publicado: (2025)
por: Li, Junyan, et al.
Publicado: (2025)
Intermediate Languages Matter: Formal Choice Drives Neurosymbolic LLM Reasoning
por: Beiser, Alexander, et al.
Publicado: (2025)
por: Beiser, Alexander, et al.
Publicado: (2025)
Enhancing Generalization in Chain of Thought Reasoning for Smaller Models
por: Yin, Maxwell J., et al.
Publicado: (2025)
por: Yin, Maxwell J., et al.
Publicado: (2025)
UniBias: Unveiling and Mitigating LLM Bias through Internal Attention and FFN Manipulation
por: Zhou, Hanzhang, et al.
Publicado: (2024)
por: Zhou, Hanzhang, et al.
Publicado: (2024)
LKD-KGC: Domain-Specific KG Construction via LLM-driven Knowledge Dependency Parsing
por: Sun, Jiaqi, et al.
Publicado: (2025)
por: Sun, Jiaqi, et al.
Publicado: (2025)
PRISM: Probing Reasoning, Instruction, and Source Memory in LLM Hallucinations
por: Wu, Yuhe, et al.
Publicado: (2026)
por: Wu, Yuhe, et al.
Publicado: (2026)
Emotional Support with LLM-based Empathetic Dialogue Generation
por: Wang, Shiquan, et al.
Publicado: (2025)
por: Wang, Shiquan, et al.
Publicado: (2025)
CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
por: Ke, Pei, et al.
Publicado: (2023)
por: Ke, Pei, et al.
Publicado: (2023)
TrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them
por: Wang, Yidong, et al.
Publicado: (2025)
por: Wang, Yidong, et al.
Publicado: (2025)
Experiences Build Characters: The Linguistic Origins and Functional Impact of LLM Personality
por: Wang, Xi, et al.
Publicado: (2026)
por: Wang, Xi, et al.
Publicado: (2026)
Assessing Bias in Metric Models for LLM Open-Ended Generation Bias Benchmarks
por: Demchak, Nathaniel, et al.
Publicado: (2024)
por: Demchak, Nathaniel, et al.
Publicado: (2024)
On LLM-Based Scientific Inductive Reasoning Beyond Equations
por: Lin, Brian S., et al.
Publicado: (2025)
por: Lin, Brian S., et al.
Publicado: (2025)
Demystifying Reasoning Dynamics with Mutual Information: Thinking Tokens are Information Peaks in LLM Reasoning
por: Qian, Chen, et al.
Publicado: (2025)
por: Qian, Chen, et al.
Publicado: (2025)
LLM Reasoning Is Latent, Not the Chain of Thought
por: Wang, Wenshuo
Publicado: (2026)
por: Wang, Wenshuo
Publicado: (2026)
Automating the Generation of Prompts for LLM-based Action Choice in PDDL Planning
por: Stein, Katharina, et al.
Publicado: (2023)
por: Stein, Katharina, et al.
Publicado: (2023)
SurveyEval: Towards Comprehensive Evaluation of LLM-Generated Academic Surveys
por: Zhao, Jiahao, et al.
Publicado: (2025)
por: Zhao, Jiahao, et al.
Publicado: (2025)
More Thinking, More Bias: Length-Driven Position Bias in Reasoning Models
por: Wang, Xiao
Publicado: (2026)
por: Wang, Xiao
Publicado: (2026)
Pragmatic Reasoning improves LLM Code Generation
por: Cao, Zhuchen, et al.
Publicado: (2025)
por: Cao, Zhuchen, et al.
Publicado: (2025)
Differentiating Choices via Commonality for Multiple-Choice Question Answering
por: Deng, Wenqing, et al.
Publicado: (2024)
por: Deng, Wenqing, et al.
Publicado: (2024)
LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models
por: Gui, Jiayi, et al.
Publicado: (2024)
por: Gui, Jiayi, et al.
Publicado: (2024)
AgenticMath: Enhancing LLM Reasoning via Agentic-based Math Data Generation
por: Liu, Xianyang, et al.
Publicado: (2025)
por: Liu, Xianyang, et al.
Publicado: (2025)
LLM should think and action as a human
por: Leung, Haun, et al.
Publicado: (2025)
por: Leung, Haun, et al.
Publicado: (2025)
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning
por: Zhuang, Yuchen, et al.
Publicado: (2025)
por: Zhuang, Yuchen, et al.
Publicado: (2025)
LegalChainReasoner: A Legal Chain-guided Framework for Criminal Judicial Opinion Generation
por: Shi, Weizhe, et al.
Publicado: (2025)
por: Shi, Weizhe, et al.
Publicado: (2025)
Harnessing the Power of Large Language Models for Empathetic Response Generation: Empirical Investigations and Improvements
por: Qian, Yushan, et al.
Publicado: (2023)
por: Qian, Yushan, et al.
Publicado: (2023)
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning
por: Zhang, Kongcheng, et al.
Publicado: (2025)
por: Zhang, Kongcheng, et al.
Publicado: (2025)
Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents
por: Shi, Yaorui, et al.
Publicado: (2025)
por: Shi, Yaorui, et al.
Publicado: (2025)
Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to Deliberation
por: Zhang, Zhiwei, et al.
Publicado: (2025)
por: Zhang, Zhiwei, et al.
Publicado: (2025)
Algorithmic Fragility and Persona Bias in LLM-Generated Autistic Communication
por: Rizvi, Naba, et al.
Publicado: (2026)
por: Rizvi, Naba, et al.
Publicado: (2026)
Vero: An Open RL Recipe for General Visual Reasoning
por: Sarch, Gabriel, et al.
Publicado: (2026)
por: Sarch, Gabriel, et al.
Publicado: (2026)
Parallel LLM Reasoning for Bias-Resilient, Robust Conceptual Abstraction
por: Adeseye, Aisvarya, et al.
Publicado: (2026)
por: Adeseye, Aisvarya, et al.
Publicado: (2026)
Balancing Rigor and Utility: Mitigating Cognitive Biases in Large Language Models for Multiple-Choice Questions
por: Zhong, Hanyang, et al.
Publicado: (2024)
por: Zhong, Hanyang, et al.
Publicado: (2024)
Can We Trust a Black-box LLM? LLM Untrustworthy Boundary Detection via Bias-Diffusion and Multi-Agent Reinforcement Learning
por: Zhou, Xiaotian, et al.
Publicado: (2026)
por: Zhou, Xiaotian, et al.
Publicado: (2026)
BiasScope: Towards Automated Detection of Bias in LLM-as-a-Judge Evaluation
por: Lai, Peng, et al.
Publicado: (2026)
por: Lai, Peng, et al.
Publicado: (2026)
Exploring and Evaluating Multimodal Knowledge Reasoning Consistency of Multimodal Large Language Models
por: Jia, Boyu, et al.
Publicado: (2025)
por: Jia, Boyu, et al.
Publicado: (2025)
LLM-MRD: LLM-Guided Multi-View Reasoning Distillation for Fake News Detection
por: Zhou, Weilin, et al.
Publicado: (2026)
por: Zhou, Weilin, et al.
Publicado: (2026)
Developing A Framework to Support Human Evaluation of Bias in Generated Free Response Text
por: Healey, Jennifer, et al.
Publicado: (2025)
por: Healey, Jennifer, et al.
Publicado: (2025)
Ejemplares similares
-
iTAG: Inverse Design for Natural Text Generation with Accurate Causal Graph Annotations
por: Wang, Wenshuo, et al.
Publicado: (2026) -
Anchored Cyclic Generation: A Novel Paradigm for Long-Sequence Symbolic Music Generation
por: Cao, Boyu, et al.
Publicado: (2026) -
Does Reasoning Introduce Bias? A Study of Social Bias Evaluation and Mitigation in LLM Reasoning
por: Wu, Xuyang, et al.
Publicado: (2025) -
Steering LLM Thinking with Budget Guidance
por: Li, Junyan, et al.
Publicado: (2025) -
Intermediate Languages Matter: Formal Choice Drives Neurosymbolic LLM Reasoning
por: Beiser, Alexander, et al.
Publicado: (2025)