From Stimuli to Minds: Enhancing Psychological Reasoning in LLMs via Bilateral Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Feng, Yichao, Luo, Haoran, Feng, Lang, Zhao, Shuai, Luu, Anh Tuan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search
por: Luo, Haoran, et al.
Publicado: (2025)
por: Luo, Haoran, et al.
Publicado: (2025)
Aspect-Based Summarization with Self-Aspect Retrieval Enhanced Generation
por: Feng, Yichao, et al.
Publicado: (2025)
por: Feng, Yichao, et al.
Publicado: (2025)
Not All RDF is Created Equal: Investigating RDF Load Times on Resource-Constrained Devices
por: Sowinski, Piotr, et al.
Publicado: (2024)
por: Sowinski, Piotr, et al.
Publicado: (2024)
LLMIA: An Out-of-the-Box Index Advisor via In-Context Learning with LLMs
por: Zhao, Xinxin, et al.
Publicado: (2025)
por: Zhao, Xinxin, et al.
Publicado: (2025)
Decomposition-Driven Multi-Table Retrieval and Reasoning for Numerical Question Answering
por: Luo, Feng, et al.
Publicado: (2026)
por: Luo, Feng, et al.
Publicado: (2026)
OrchMAS: Orchestrated Reasoning with Multi Collaborative Heterogeneous Scientific Expert Structured Agents
por: Feng, Yichao, et al.
Publicado: (2026)
por: Feng, Yichao, et al.
Publicado: (2026)
MonoM: Enhancing Monotonicity in Learned Cardinality Estimators
por: Yi, Lyu, et al.
Publicado: (2025)
por: Yi, Lyu, et al.
Publicado: (2025)
Unlearning Backdoor Attacks for LLMs with Weak-to-Strong Knowledge Distillation
por: Zhao, Shuai, et al.
Publicado: (2024)
por: Zhao, Shuai, et al.
Publicado: (2024)
Rethinking Reasoning: A Survey on Reasoning-based Backdoors in LLMs
por: Hu, Man, et al.
Publicado: (2025)
por: Hu, Man, et al.
Publicado: (2025)
Enhancing Multimodal Entity Linking with Jaccard Distance-based Conditional Contrastive Learning and Contextual Visual Augmentation
por: Nguyen, Cong-Duy, et al.
Publicado: (2025)
por: Nguyen, Cong-Duy, et al.
Publicado: (2025)
SQL-R1: Training Natural Language to SQL Reasoning Model By Reinforcement Learning
por: Ma, Peixian, et al.
Publicado: (2025)
por: Ma, Peixian, et al.
Publicado: (2025)
FairDAG: Consensus Fairness over Multi-Proposer Causal Design
por: Kang, Dakai, et al.
Publicado: (2025)
por: Kang, Dakai, et al.
Publicado: (2025)
TKG-Thinker: Towards Dynamic Reasoning over Temporal Knowledge Graphs via Agentic Reinforcement Learning
por: Jiang, Zihao, et al.
Publicado: (2026)
por: Jiang, Zihao, et al.
Publicado: (2026)
Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL
por: Pourreza, Mohammadreza, et al.
Publicado: (2025)
por: Pourreza, Mohammadreza, et al.
Publicado: (2025)
VectorMaton: Efficient Vector Search with Pattern Constraints via an Enhanced Suffix Automaton
por: Xie, Haoxuan, et al.
Publicado: (2026)
por: Xie, Haoxuan, et al.
Publicado: (2026)
Mind the Data Gap: Bridging LLMs to Enterprise Data Integration
por: Kayali, Moe, et al.
Publicado: (2024)
por: Kayali, Moe, et al.
Publicado: (2024)
DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving
por: Hou, Xinmeng, et al.
Publicado: (2025)
por: Hou, Xinmeng, et al.
Publicado: (2025)
TabSQLify: Enhancing Reasoning Capabilities of LLMs Through Table Decomposition
por: Nahid, Md Mahadi Hasan, et al.
Publicado: (2024)
por: Nahid, Md Mahadi Hasan, et al.
Publicado: (2024)
AdaNDV: Adaptive Number of Distinct Value Estimation via Learning to Select and Fuse Estimators
por: Xu, Xianghong, et al.
Publicado: (2025)
por: Xu, Xianghong, et al.
Publicado: (2025)
Query Rewriting via LLMs
por: Dharwada, Sriram, et al.
Publicado: (2025)
por: Dharwada, Sriram, et al.
Publicado: (2025)
RLMiner: Finding the Most Frequent k-sized Subgraph via Reinforcement Learning
por: Huang, Wei, et al.
Publicado: (2026)
por: Huang, Wei, et al.
Publicado: (2026)
Learning from Uncertain Data: From Possible Worlds to Possible Models
por: Zhu, Jiongli, et al.
Publicado: (2024)
por: Zhu, Jiongli, et al.
Publicado: (2024)
Advances in LLMs with Focus on Reasoning, Adaptability, Efficiency and Ethics
por: Khan, Asifullah, et al.
Publicado: (2025)
por: Khan, Asifullah, et al.
Publicado: (2025)
Three Minds, One Legend: Jailbreak Large Reasoning Model with Adaptive Stacked Ciphers
por: Nguyen, Viet-Anh, et al.
Publicado: (2025)
por: Nguyen, Viet-Anh, et al.
Publicado: (2025)
Towards Cost-effective LLMs Routing with Batch Prompting
por: Xu, Haotian, et al.
Publicado: (2026)
por: Xu, Haotian, et al.
Publicado: (2026)
KOR-Bench: Benchmarking Language Models on Knowledge-Orthogonal Reasoning Tasks
por: Ma, Kaijing, et al.
Publicado: (2024)
por: Ma, Kaijing, et al.
Publicado: (2024)
AutoIndexer: A Reinforcement Learning-Enhanced Index Advisor Towards Scaling Workloads
por: Wang, Taiyi, et al.
Publicado: (2025)
por: Wang, Taiyi, et al.
Publicado: (2025)
SQL-o1: A Self-Reward Heuristic Dynamic Search Method for Text-to-SQL
por: Lyu, Shuai, et al.
Publicado: (2025)
por: Lyu, Shuai, et al.
Publicado: (2025)
JudgeSQL: Reasoning over SQL Candidates with Weighted Consensus Tournament
por: Bai, Jiayuan, et al.
Publicado: (2025)
por: Bai, Jiayuan, et al.
Publicado: (2025)
DW-Bench: Benchmarking LLMs on Data Warehouse Graph Topology Reasoning
por: Ahmed, Ahmed G. A. H, et al.
Publicado: (2026)
por: Ahmed, Ahmed G. A. H, et al.
Publicado: (2026)
KnobTree: Intelligent Database Parameter Configuration via Explainable Reinforcement Learning
por: Chen, Jiahan, et al.
Publicado: (2024)
por: Chen, Jiahan, et al.
Publicado: (2024)
HAKES: Scalable Vector Database for Embedding Search Service
por: Hu, Guoyu, et al.
Publicado: (2025)
por: Hu, Guoyu, et al.
Publicado: (2025)
ReMatch: Retrieval Enhanced Schema Matching with LLMs
por: Sheetrit, Eitam, et al.
Publicado: (2024)
por: Sheetrit, Eitam, et al.
Publicado: (2024)
Aster: Enhancing LSM-structures for Scalable Graph Database
por: Mo, Dingheng, et al.
Publicado: (2025)
por: Mo, Dingheng, et al.
Publicado: (2025)
A New Paradigm in Tuning Learned Indexes: A Reinforcement Learning Enhanced Approach
por: Wang, Taiyi, et al.
Publicado: (2025)
por: Wang, Taiyi, et al.
Publicado: (2025)
BQSched: A Non-intrusive Scheduler for Batch Concurrent Queries via Reinforcement Learning
por: Xu, Chenhao, et al.
Publicado: (2025)
por: Xu, Chenhao, et al.
Publicado: (2025)
Think2SQL: Reinforce LLM Reasoning Capabilities for Text2SQL
por: Papicchio, Simone, et al.
Publicado: (2025)
por: Papicchio, Simone, et al.
Publicado: (2025)
ACCIO: Table Understanding Enhanced via Contrastive Learning with Aggregations
por: Cho, Whanhee
Publicado: (2024)
por: Cho, Whanhee
Publicado: (2024)
Toward a Cognitive Data Model: Exploring a Mind-Inspired Approach to Database Design
por: Pieris, Dhammika
Publicado: (2025)
por: Pieris, Dhammika
Publicado: (2025)
TabTracer: Monte Carlo Tree Search for Complex Table Reasoning with Large Language Models
por: Luo, Zhizhao, et al.
Publicado: (2026)
por: Luo, Zhizhao, et al.
Publicado: (2026)
Ejemplares similares
-
KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search
por: Luo, Haoran, et al.
Publicado: (2025) -
Aspect-Based Summarization with Self-Aspect Retrieval Enhanced Generation
por: Feng, Yichao, et al.
Publicado: (2025) -
Not All RDF is Created Equal: Investigating RDF Load Times on Resource-Constrained Devices
por: Sowinski, Piotr, et al.
Publicado: (2024) -
LLMIA: An Out-of-the-Box Index Advisor via In-Context Learning with LLMs
por: Zhao, Xinxin, et al.
Publicado: (2025) -
Decomposition-Driven Multi-Table Retrieval and Reasoning for Numerical Question Answering
por: Luo, Feng, et al.
Publicado: (2026)