Gespeichert in:
| Hauptverfasser: | Zhang, Zizhuo, Zhu, Jianing, Ge, Xinmu, Zhao, Zihua, Zhou, Zhanke, Li, Xuan, Feng, Xiao, Yao, Jiangchao, Han, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2508.00410 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeepInception: Hypnotize Large Language Model to Be Jailbreaker
von: Li, Xuan, et al.
Veröffentlicht: (2023)
von: Li, Xuan, et al.
Veröffentlicht: (2023)
From Passive to Active Reasoning: Can Large Language Models Ask the Right Questions under Incomplete Information?
von: Zhou, Zhanke, et al.
Veröffentlicht: (2025)
von: Zhou, Zhanke, et al.
Veröffentlicht: (2025)
Less is More: One-shot Subgraph Reasoning on Large-scale Knowledge Graphs
von: Zhou, Zhanke, et al.
Veröffentlicht: (2024)
von: Zhou, Zhanke, et al.
Veröffentlicht: (2024)
Towards Understanding Valuable Preference Data for Large Language Model Alignment
von: Zhang, Zizhuo, et al.
Veröffentlicht: (2025)
von: Zhang, Zizhuo, et al.
Veröffentlicht: (2025)
RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models
von: Feng, Xiao, et al.
Veröffentlicht: (2026)
von: Feng, Xiao, et al.
Veröffentlicht: (2026)
Eliciting Causal Abilities in Large Language Models for Reasoning Tasks
von: Wang, Yajing, et al.
Veröffentlicht: (2024)
von: Wang, Yajing, et al.
Veröffentlicht: (2024)
Self-Calibrated Tuning of Vision-Language Models for Out-of-Distribution Detection
von: Yu, Geng, et al.
Veröffentlicht: (2024)
von: Yu, Geng, et al.
Veröffentlicht: (2024)
Reference-guided Policy Optimization for Molecular Optimization via LLM Reasoning
von: Li, Xuan, et al.
Veröffentlicht: (2026)
von: Li, Xuan, et al.
Veröffentlicht: (2026)
Neural Atoms: Propagating Long-range Interaction in Molecular Graphs through Efficient Communication Channel
von: Li, Xuan, et al.
Veröffentlicht: (2023)
von: Li, Xuan, et al.
Veröffentlicht: (2023)
Landscape of Thoughts: Visualizing the Reasoning Process of Large Language Models
von: Zhou, Zhanke, et al.
Veröffentlicht: (2025)
von: Zhou, Zhanke, et al.
Veröffentlicht: (2025)
Can Language Models Perform Robust Reasoning in Chain-of-thought Prompting with Noisy Rationales?
von: Zhou, Zhanke, et al.
Veröffentlicht: (2024)
von: Zhou, Zhanke, et al.
Veröffentlicht: (2024)
Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards
von: Huang, Yu, et al.
Veröffentlicht: (2026)
von: Huang, Yu, et al.
Veröffentlicht: (2026)
Model Inversion Attacks: A Survey of Approaches and Countermeasures
von: Zhou, Zhanke, et al.
Veröffentlicht: (2024)
von: Zhou, Zhanke, et al.
Veröffentlicht: (2024)
Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory
von: Sun, Jingwei, et al.
Veröffentlicht: (2026)
von: Sun, Jingwei, et al.
Veröffentlicht: (2026)
Fast and Accurate Blind Flexible Docking
von: Zhang, Zizhuo, et al.
Veröffentlicht: (2025)
von: Zhang, Zizhuo, et al.
Veröffentlicht: (2025)
Decoupling the Class Label and the Target Concept in Machine Unlearning
von: Zhu, Jianing, et al.
Veröffentlicht: (2024)
von: Zhu, Jianing, et al.
Veröffentlicht: (2024)
Physics Reasoner: Knowledge-Augmented Reasoning for Solving Physics Problems with Large Language Models
von: Pang, Xinyu, et al.
Veröffentlicht: (2024)
von: Pang, Xinyu, et al.
Veröffentlicht: (2024)
Envisioning Outlier Exposure by Large Language Models for Out-of-Distribution Detection
von: Cao, Chentao, et al.
Veröffentlicht: (2024)
von: Cao, Chentao, et al.
Veröffentlicht: (2024)
AlphaApollo: A System for Deep Agentic Reasoning
von: Zhou, Zhanke, et al.
Veröffentlicht: (2025)
von: Zhou, Zhanke, et al.
Veröffentlicht: (2025)
Per-parameter Task Arithmetic for Unlearning in Large Language Models
von: Cai, Chengyi, et al.
Veröffentlicht: (2026)
von: Cai, Chengyi, et al.
Veröffentlicht: (2026)
Eliciting Medical Reasoning with Knowledge-enhanced Data Synthesis: A Semi-Supervised Reinforcement Learning Approach
von: Li, Haolin, et al.
Veröffentlicht: (2026)
von: Li, Haolin, et al.
Veröffentlicht: (2026)
Understanding Fairness Surrogate Functions in Algorithmic Fairness
von: Yao, Wei, et al.
Veröffentlicht: (2023)
von: Yao, Wei, et al.
Veröffentlicht: (2023)
Self-Debias: Self-correcting for Debiasing Large Language Models
von: Feng, Xuan, et al.
Veröffentlicht: (2026)
von: Feng, Xuan, et al.
Veröffentlicht: (2026)
Mitigating Noisy Correspondence by Geometrical Structure Consistency Learning
von: Zhao, Zihua, et al.
Veröffentlicht: (2024)
von: Zhao, Zihua, et al.
Veröffentlicht: (2024)
Self-Training Elicits Concise Reasoning in Large Language Models
von: Munkhbat, Tergel, et al.
Veröffentlicht: (2025)
von: Munkhbat, Tergel, et al.
Veröffentlicht: (2025)
From Debate to Equilibrium: Belief-Driven Multi-Agent LLM Reasoning via Bayesian Nash Equilibrium
von: Yi, Xie, et al.
Veröffentlicht: (2025)
von: Yi, Xie, et al.
Veröffentlicht: (2025)
Dual-granularity Sinkhorn Distillation for Enhanced Learning from Long-tailed Noisy Data
von: Hong, Feng, et al.
Veröffentlicht: (2025)
von: Hong, Feng, et al.
Veröffentlicht: (2025)
Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL
von: He, Haoyang, et al.
Veröffentlicht: (2025)
von: He, Haoyang, et al.
Veröffentlicht: (2025)
SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning
von: Guo, Xiaojun, et al.
Veröffentlicht: (2025)
von: Guo, Xiaojun, et al.
Veröffentlicht: (2025)
Exploring Training on Heterogeneous Data with Mixture of Low-rank Adapters
von: Zhou, Yuhang, et al.
Veröffentlicht: (2024)
von: Zhou, Yuhang, et al.
Veröffentlicht: (2024)
Distributive fairness during the transition to adolescence: The role of peer comparison and social value orientation
von: Siqi Liu, et al.
Veröffentlicht: (2024)
von: Siqi Liu, et al.
Veröffentlicht: (2024)
Enhancing Sequential Recommendation with World Knowledge from Large Language Models
von: Dai, Tianjie, et al.
Veröffentlicht: (2025)
von: Dai, Tianjie, et al.
Veröffentlicht: (2025)
Is There Only One Vacuum?
von: Zong, Xinmu, et al.
Veröffentlicht: (2025)
von: Zong, Xinmu, et al.
Veröffentlicht: (2025)
Mind the Gap Between Prototypes and Images in Cross-domain Finetuning
von: Tian, Hongduan, et al.
Veröffentlicht: (2024)
von: Tian, Hongduan, et al.
Veröffentlicht: (2024)
Differential-informed Sample Selection Accelerates Multimodal Contrastive Learning
von: Zhao, Zihua, et al.
Veröffentlicht: (2025)
von: Zhao, Zihua, et al.
Veröffentlicht: (2025)
Unraveling the Impact of Heterophilic Structures on Graph Positive-Unlabeled Learning
von: Wu, Yuhao, et al.
Veröffentlicht: (2024)
von: Wu, Yuhao, et al.
Veröffentlicht: (2024)
Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models
von: Qiu, Yifu, et al.
Veröffentlicht: (2025)
von: Qiu, Yifu, et al.
Veröffentlicht: (2025)
Noisy Test-Time Adaptation in Vision-Language Models
von: Cao, Chentao, et al.
Veröffentlicht: (2025)
von: Cao, Chentao, et al.
Veröffentlicht: (2025)
Self-supervised network distillation: an effective approach to exploration in sparse reward environments
von: Pecháč, Matej, et al.
Veröffentlicht: (2023)
von: Pecháč, Matej, et al.
Veröffentlicht: (2023)
GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization
von: Liu, Shih-Yang, et al.
Veröffentlicht: (2026)
von: Liu, Shih-Yang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DeepInception: Hypnotize Large Language Model to Be Jailbreaker
von: Li, Xuan, et al.
Veröffentlicht: (2023) -
From Passive to Active Reasoning: Can Large Language Models Ask the Right Questions under Incomplete Information?
von: Zhou, Zhanke, et al.
Veröffentlicht: (2025) -
Less is More: One-shot Subgraph Reasoning on Large-scale Knowledge Graphs
von: Zhou, Zhanke, et al.
Veröffentlicht: (2024) -
Towards Understanding Valuable Preference Data for Large Language Model Alignment
von: Zhang, Zizhuo, et al.
Veröffentlicht: (2025) -
RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models
von: Feng, Xiao, et al.
Veröffentlicht: (2026)