Think Socially via Cognitive Reasoning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhou, Jinfeng, Chen, Zheyu, Wang, Shuai, Dai, Quanyu, Dong, Zhenhua, Wang, Hongning, Huang, Minlie |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
SocialEval: Evaluating Social Intelligence of Large Language Models
par: Zhou, Jinfeng, et autres
Publié: (2025)
par: Zhou, Jinfeng, et autres
Publié: (2025)
SocialSim: Towards Socialized Simulation of Emotional Support Conversation
par: Chen, Zhuang, et autres
Publié: (2025)
par: Chen, Zhuang, et autres
Publié: (2025)
Unlocking Reasoning Potential in Large Langauge Models by Scaling Code-form Planning
par: Wen, Jiaxin, et autres
Publié: (2024)
par: Wen, Jiaxin, et autres
Publié: (2024)
Language Model Decoding as Direct Metrics Optimization
par: Ji, Haozhe, et autres
Publié: (2023)
par: Ji, Haozhe, et autres
Publié: (2023)
Data Selection via Optimal Control for Language Models
par: Gu, Yuxian, et autres
Publié: (2024)
par: Gu, Yuxian, et autres
Publié: (2024)
Crisp: Cognitive Restructuring of Negative Thoughts through Multi-turn Supportive Dialogues
par: Zhou, Jinfeng, et autres
Publié: (2025)
par: Zhou, Jinfeng, et autres
Publié: (2025)
HPSS: Heuristic Prompting Strategy Search for LLM Evaluators
par: Wen, Bosi, et autres
Publié: (2025)
par: Wen, Bosi, et autres
Publié: (2025)
ShieldVLM: Safeguarding the Multimodal Implicit Toxicity via Deliberative Reasoning with LVLMs
par: Cui, Shiyao, et autres
Publié: (2025)
par: Cui, Shiyao, et autres
Publié: (2025)
Learning Task Decomposition to Assist Humans in Competitive Programming
par: Wen, Jiaxin, et autres
Publié: (2024)
par: Wen, Jiaxin, et autres
Publié: (2024)
Defending Large Language Models Against Jailbreaking Attacks Through Goal Prioritization
par: Zhang, Zhexin, et autres
Publié: (2023)
par: Zhang, Zhexin, et autres
Publié: (2023)
AMOR: A Recipe for Building Adaptable Modular Knowledge Agents Through Process Feedback
par: Guan, Jian, et autres
Publié: (2024)
par: Guan, Jian, et autres
Publié: (2024)
LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models
par: Gui, Jiayi, et autres
Publié: (2024)
par: Gui, Jiayi, et autres
Publié: (2024)
Agent-SafetyBench: Evaluating the Safety of LLM Agents
par: Zhang, Zhexin, et autres
Publié: (2024)
par: Zhang, Zhexin, et autres
Publié: (2024)
RAVEL: Reasoning Agents for Validating and Evaluating LLM Text Synthesis
par: Feng, Andrew Zhuoer, et autres
Publié: (2026)
par: Feng, Andrew Zhuoer, et autres
Publié: (2026)
RLAR: An Agentic Reward System for Multi-task Reinforcement Learning on Large Language Models
par: Feng, Andrew Zhuoer, et autres
Publié: (2026)
par: Feng, Andrew Zhuoer, et autres
Publié: (2026)
MemBench: Towards More Comprehensive Evaluation on the Memory of LLM-based Agents
par: Tan, Haoran, et autres
Publié: (2025)
par: Tan, Haoran, et autres
Publié: (2025)
Black-Box Prompt Optimization: Aligning Large Language Models without Model Training
par: Cheng, Jiale, et autres
Publié: (2023)
par: Cheng, Jiale, et autres
Publié: (2023)
Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!
par: Zhang, Zhexin, et autres
Publié: (2025)
par: Zhang, Zhexin, et autres
Publié: (2025)
Improving Retrospective Language Agents via Joint Policy Gradient Optimization
par: Feng, Xueyang, et autres
Publié: (2025)
par: Feng, Xueyang, et autres
Publié: (2025)
IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation
par: Wen, Bosi, et autres
Publié: (2026)
par: Wen, Bosi, et autres
Publié: (2026)
How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study
par: Zhang, Zhexin, et autres
Publié: (2025)
par: Zhang, Zhexin, et autres
Publié: (2025)
KnowTrace: Bootstrapping Iterative Retrieval-Augmented Generation with Structured Knowledge Tracing
par: Li, Rui, et autres
Publié: (2025)
par: Li, Rui, et autres
Publié: (2025)
Beyond Single-Point Judgment: Distribution Alignment for LLM-as-a-Judge
par: Chen, Luyu, et autres
Publié: (2025)
par: Chen, Luyu, et autres
Publié: (2025)
Towards Efficient Exact Optimization of Language Model Alignment
par: Ji, Haozhe, et autres
Publié: (2024)
par: Ji, Haozhe, et autres
Publié: (2024)
HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing
par: Feng, Andrew Zhuoer, et autres
Publié: (2026)
par: Feng, Andrew Zhuoer, et autres
Publié: (2026)
IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation
par: Wen, Bosi, et autres
Publié: (2025)
par: Wen, Bosi, et autres
Publié: (2025)
When Smiley Turns Hostile: Interpreting How Emojis Trigger LLMs' Toxicity
par: Cui, Shiyao, et autres
Publié: (2025)
par: Cui, Shiyao, et autres
Publié: (2025)
Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework
par: Zhang, Zeyu, et autres
Publié: (2025)
par: Zhang, Zeyu, et autres
Publié: (2025)
CharacterBench: Benchmarking Character Customization of Large Language Models
par: Zhou, Jinfeng, et autres
Publié: (2024)
par: Zhou, Jinfeng, et autres
Publié: (2024)
BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs
par: Yang, Junxiao, et autres
Publié: (2025)
par: Yang, Junxiao, et autres
Publié: (2025)
LongSafety: Evaluating Long-Context Safety of Large Language Models
par: Lu, Yida, et autres
Publié: (2025)
par: Lu, Yida, et autres
Publié: (2025)
COKE: A Cognitive Knowledge Graph for Machine Theory of Mind
par: Wu, Jincenzi, et autres
Publié: (2023)
par: Wu, Jincenzi, et autres
Publié: (2023)
Benchmarking Complex Instruction-Following with Multiple Constraints Composition
par: Wen, Bosi, et autres
Publié: (2024)
par: Wen, Bosi, et autres
Publié: (2024)
Expectation Confirmation Preference Optimization for Multi-Turn Conversational Recommendation Agent
par: Feng, Xueyang, et autres
Publié: (2025)
par: Feng, Xueyang, et autres
Publié: (2025)
Prompt and Parameter Co-Optimization for Large Language Models
par: Bo, Xiaohe, et autres
Publié: (2025)
par: Bo, Xiaohe, et autres
Publié: (2025)
CAM: A Constructivist View of Agentic Memory for LLM-Based Reading Comprehension
par: Li, Rui, et autres
Publié: (2025)
par: Li, Rui, et autres
Publié: (2025)
ChatGLM-RLHF: Practices of Aligning Large Language Models with Human Feedback
par: Hou, Zhenyu, et autres
Publié: (2024)
par: Hou, Zhenyu, et autres
Publié: (2024)
AutoDetect: Towards a Unified Framework for Automated Weakness Detection in Large Language Models
par: Cheng, Jiale, et autres
Publié: (2024)
par: Cheng, Jiale, et autres
Publié: (2024)
Does RLHF Scale? Exploring the Impacts From Data, Model, and Method
par: Hou, Zhenyu, et autres
Publié: (2024)
par: Hou, Zhenyu, et autres
Publié: (2024)
SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
par: Cheng, Jiale, et autres
Publié: (2024)
par: Cheng, Jiale, et autres
Publié: (2024)
Documents similaires
-
SocialEval: Evaluating Social Intelligence of Large Language Models
par: Zhou, Jinfeng, et autres
Publié: (2025) -
SocialSim: Towards Socialized Simulation of Emotional Support Conversation
par: Chen, Zhuang, et autres
Publié: (2025) -
Unlocking Reasoning Potential in Large Langauge Models by Scaling Code-form Planning
par: Wen, Jiaxin, et autres
Publié: (2024) -
Language Model Decoding as Direct Metrics Optimization
par: Ji, Haozhe, et autres
Publié: (2023) -
Data Selection via Optimal Control for Language Models
par: Gu, Yuxian, et autres
Publié: (2024)