A Survey on Post-training of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tie, Guiyao, Zhao, Zeli, Song, Dingjie, Wei, Fuyang, Zhou, Rong, Dai, Yurou, Yin, Wen, Yang, Zhejian, Yan, Jiangyue, Su, Yao, Dai, Zhenhan, Xie, Yifeng, Cao, Yihan, Sun, Lichao, Zhou, Pan, He, Lifang, Chen, Hechang, Zhang, Yu, Wen, Qingsong, Liu, Tianming, Gong, Neil Zhenqiang, Tang, Jiliang, Xiong, Caiming, Ji, Heng, Yu, Philip S., Gao, Jianfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents
von: Yang, Zhejian, et al.
Veröffentlicht: (2025)
von: Yang, Zhejian, et al.
Veröffentlicht: (2025)
BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
Prompt Injection Attack to Tool Selection in LLM Agents
von: Shi, Jiawen, et al.
Veröffentlicht: (2025)
von: Shi, Jiawen, et al.
Veröffentlicht: (2025)
A Survey of AI Scientists
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
EmbodiedClaw: Conversational Workflow Execution for Embodied AI Development
von: Zhou, Xueyang, et al.
Veröffentlicht: (2026)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2026)
SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
Exploring the Necessity of Reasoning in LLM-based Agent Scenarios
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
From Correctness to Comprehension: AI Agents for Personalized Error Diagnosis in Education
von: Zhang, Yi-Fan, et al.
Veröffentlicht: (2025)
von: Zhang, Yi-Fan, et al.
Veröffentlicht: (2025)
LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
TGSBM: Transformer-Guided Stochastic Block Model for Link Prediction
von: Yang, Zhejian, et al.
Veröffentlicht: (2026)
von: Yang, Zhejian, et al.
Veröffentlicht: (2026)
Advancing Autonomous Driving System Testing: Demands, Challenges, and Future Directions
von: Liao, Yihan, et al.
Veröffentlicht: (2025)
von: Liao, Yihan, et al.
Veröffentlicht: (2025)
BadToken: Token-level Backdoor Attacks to Multi-modal Large Language Models
von: Yuan, Zenghui, et al.
Veröffentlicht: (2025)
von: Yuan, Zenghui, et al.
Veröffentlicht: (2025)
Decision Mamba: Reinforcement Learning via Hybrid Selective Sequence Modeling
von: Huang, Sili, et al.
Veröffentlicht: (2024)
von: Huang, Sili, et al.
Veröffentlicht: (2024)
Continual Diffuser (CoD): Mastering Continual Offline Reinforcement Learning with Experience Rehearsal
von: Hu, Jifeng, et al.
Veröffentlicht: (2024)
von: Hu, Jifeng, et al.
Veröffentlicht: (2024)
Automate Knowledge Concept Tagging on Math Questions with LLMs
von: Li, Hang, et al.
Veröffentlicht: (2024)
von: Li, Hang, et al.
Veröffentlicht: (2024)
Knowledge Tagging System on Math Questions via LLMs with Flexible Demonstration Retriever
von: Li, Hang, et al.
Veröffentlicht: (2024)
von: Li, Hang, et al.
Veröffentlicht: (2024)
Conceptual Steganography
von: Zhou, Zhejian, et al.
Veröffentlicht: (2026)
von: Zhou, Zhejian, et al.
Veröffentlicht: (2026)
Self-Cognition in Large Language Models: An Exploratory Study
von: Chen, Dongping, et al.
Veröffentlicht: (2024)
von: Chen, Dongping, et al.
Veröffentlicht: (2024)
MMLU-Reason: Benchmarking Multi-Task Multi-modal Language Understanding and Reasoning
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
Enabling Extensible Embodied Capabilities with Tools
von: Zhou, Xueyang, et al.
Veröffentlicht: (2026)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2026)
Analytic Energy-Guided Policy Optimization for Offline Reinforcement Learning
von: Hu, Jifeng, et al.
Veröffentlicht: (2025)
von: Hu, Jifeng, et al.
Veröffentlicht: (2025)
A Survey of Mix-based Data Augmentation: Taxonomy, Methods, Applications, and Explainability
von: Cao, Chengtai, et al.
Veröffentlicht: (2022)
von: Cao, Chengtai, et al.
Veröffentlicht: (2022)
Can MLLMs Read Students' Minds? Unpacking Multimodal Error Analysis in Handwritten Math
von: Song, Dingjie, et al.
Veröffentlicht: (2026)
von: Song, Dingjie, et al.
Veröffentlicht: (2026)
Towards a Medical AI Scientist
von: Wu, Hongtao, et al.
Veröffentlicht: (2026)
von: Wu, Hongtao, et al.
Veröffentlicht: (2026)
RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty
von: Zhang, Ziqian, et al.
Veröffentlicht: (2026)
von: Zhang, Ziqian, et al.
Veröffentlicht: (2026)
SAMed-2: Selective Memory Enhanced Medical Segment Anything Model
von: Yan, Zhiling, et al.
Veröffentlicht: (2025)
von: Yan, Zhiling, et al.
Veröffentlicht: (2025)
Comment on “Drivers of Frequent Emergency Department Use in Socioeconomically Disadvantaged Older Adults: A Qualitative Study” by Chary et al.
von: Yu Dai, et al.
Veröffentlicht: (2025)
von: Yu Dai, et al.
Veröffentlicht: (2025)
Optimization-based Prompt Injection Attack to LLM-as-a-Judge
von: Shi, Jiawen, et al.
Veröffentlicht: (2024)
von: Shi, Jiawen, et al.
Veröffentlicht: (2024)
Solving Continual Offline RL through Selective Weights Activation on Aligned Spaces
von: Hu, Jifeng, et al.
Veröffentlicht: (2024)
von: Hu, Jifeng, et al.
Veröffentlicht: (2024)
Visual Hallucinations of Multi-modal Large Language Models
von: Huang, Wen, et al.
Veröffentlicht: (2024)
von: Huang, Wen, et al.
Veröffentlicht: (2024)
Bilateral globus pallidus internus‐deep brain stimulation in a 5‐year‐old boy with SGCE‐related myoclonus dystonia syndrome
von: Xiaojuan Tian, et al.
Veröffentlicht: (2024)
von: Xiaojuan Tian, et al.
Veröffentlicht: (2024)
How Language Models Process Negation
von: Zhou, Zhejian, et al.
Veröffentlicht: (2026)
von: Zhou, Zhejian, et al.
Veröffentlicht: (2026)
Revisiting the Graph Reasoning Ability of Large Language Models: Case Studies in Translation, Connectivity and Shortest Path
von: Dai, Xinnan, et al.
Veröffentlicht: (2024)
von: Dai, Xinnan, et al.
Veröffentlicht: (2024)
Creation of Three‐Scroll Hidden Conservative Lorenz‐Like Chaotic Flows
von: Guiyao Ke
Veröffentlicht: (2024)
von: Guiyao Ke
Veröffentlicht: (2024)
Large Language Models for Education: A Survey and Outlook
von: Wang, Shen, et al.
Veröffentlicht: (2024)
von: Wang, Shen, et al.
Veröffentlicht: (2024)
Recursive structures of molecules and cells in Gelfand $S_n$-graphs
von: Dai, Zhiqiang, et al.
Veröffentlicht: (2026)
von: Dai, Zhiqiang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents
von: Yang, Zhejian, et al.
Veröffentlicht: (2025) -
BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025) -
Prompt Injection Attack to Tool Selection in LLM Agents
von: Shi, Jiawen, et al.
Veröffentlicht: (2025) -
A Survey of AI Scientists
von: Tie, Guiyao, et al.
Veröffentlicht: (2025) -
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)