Supervising the search process produces reliable and generalizable information-seeking agents
Fuente:
arXiv
Saved in:
| Main Authors: | Xiong, Guangzhi, Jin, Qiao, Wang, Xiao, Fang, Yin, Liu, Haolin, Yang, Yifan, Chen, Fangyuan, Song, Zhixing, Wang, Dengyu, Zhang, Minjia, Lu, Zhiyong, Zhang, Aidong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Retrieval-Augmented Generation in Medicine with Iterative Follow-up Questions
by: Xiong, Guangzhi, et al.
Published: (2024)
by: Xiong, Guangzhi, et al.
Published: (2024)
Benchmarking Retrieval-Augmented Generation for Medicine
by: Xiong, Guangzhi, et al.
Published: (2024)
by: Xiong, Guangzhi, et al.
Published: (2024)
MedCite: Can Language Models Generate Verifiable Text for Medicine?
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models
by: Xiong, Guangzhi, et al.
Published: (2026)
by: Xiong, Guangzhi, et al.
Published: (2026)
Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning
by: Fang, Yin, et al.
Published: (2025)
by: Fang, Yin, et al.
Published: (2025)
Structural Causality-based Generalizable Concept Discovery Models
by: Sinha, Sanchit, et al.
Published: (2024)
by: Sinha, Sanchit, et al.
Published: (2024)
ASCENT-ViT: Attention-based Scale-aware Concept Learning Framework for Enhanced Alignment in Vision Transformers
by: Sinha, Sanchit, et al.
Published: (2025)
by: Sinha, Sanchit, et al.
Published: (2025)
COCO-Tree: Compositional Hierarchical Concept Trees for Enhanced Reasoning in Vision Language Models
by: Sinha, Sanchit, et al.
Published: (2025)
by: Sinha, Sanchit, et al.
Published: (2025)
CoLiDR: Concept Learning using Aggregated Disentangled Representations
by: Sinha, Sanchit, et al.
Published: (2024)
by: Sinha, Sanchit, et al.
Published: (2024)
Neural Additive Experts: Context-Gated Experts for Controllable Model Additivity
by: Xiong, Guangzhi, et al.
Published: (2026)
by: Xiong, Guangzhi, et al.
Published: (2026)
A Self-explaining Neural Architecture for Generalizable Concept Learning
by: Sinha, Sanchit, et al.
Published: (2024)
by: Sinha, Sanchit, et al.
Published: (2024)
ProtoNAM: Prototypical Neural Additive Models for Interpretable Deep Tabular Learning
by: Xiong, Guangzhi, et al.
Published: (2024)
by: Xiong, Guangzhi, et al.
Published: (2024)
GCAV: A Global Concept Activation Vector Framework for Cross-Layer Consistency in Interpretability
by: He, Zhenghao, et al.
Published: (2025)
by: He, Zhenghao, et al.
Published: (2025)
Concept-RuleNet: Grounded Multi-Agent Neurosymbolic Reasoning in Vision Language Models
by: Sinha, Sanchit, et al.
Published: (2025)
by: Sinha, Sanchit, et al.
Published: (2025)
Retrieving Counterfactuals Improves Visual In-Context Learning
by: Xiong, Guangzhi, et al.
Published: (2026)
by: Xiong, Guangzhi, et al.
Published: (2026)
CASL: Concept-Aligned Sparse Latents for Interpreting Diffusion Models
by: He, Zhenghao, et al.
Published: (2026)
by: He, Zhenghao, et al.
Published: (2026)
Large Language Models Lack Temporal Awareness of Medical Knowledge
by: Guan, Zihan, et al.
Published: (2026)
by: Guan, Zihan, et al.
Published: (2026)
Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution
by: Jin, Qiao, et al.
Published: (2026)
by: Jin, Qiao, et al.
Published: (2026)
Reasoning Beyond Chain-of-Thought: A Latent Computational Mode in Large Language Models
by: He, Zhenghao, et al.
Published: (2026)
by: He, Zhenghao, et al.
Published: (2026)
Attention-guided Fine-tuning of Multimodal Large Language Models Improves Chain-of-Thought Reasoning
by: Sinha, Sanchit, et al.
Published: (2026)
by: Sinha, Sanchit, et al.
Published: (2026)
Toward Faithful Retrieval-Augmented Generation with Sparse Autoencoders
by: Xiong, Guangzhi, et al.
Published: (2025)
by: Xiong, Guangzhi, et al.
Published: (2025)
Humans and Large Language Models in Clinical Decision Support: A Study with Medical Calculators
by: Wan, Nicholas, et al.
Published: (2024)
by: Wan, Nicholas, et al.
Published: (2024)
Improving Scientific Hypothesis Generation with Knowledge Grounded Large Language Models
by: Xiong, Guangzhi, et al.
Published: (2024)
by: Xiong, Guangzhi, et al.
Published: (2024)
Towards a conversational information seeking process model: Characterizing mixed‐initiative user–agent interaction
by: Shiting Fu, et al.
Published: (2025)
by: Shiting Fu, et al.
Published: (2025)
Gene-R1: Reasoning with Data-Augmented Lightweight LLMs for Gene Set Analysis
by: Wang, Zhizheng, et al.
Published: (2025)
by: Wang, Zhizheng, et al.
Published: (2025)
IdeaBench: Benchmarking Large Language Models for Research Idea Generation
by: Guo, Sikun, et al.
Published: (2024)
by: Guo, Sikun, et al.
Published: (2024)
AMaze: An intuitive benchmark generator for fast prototyping of generalizable agents
by: Godin-Dubois, Kevin, et al.
Published: (2024)
by: Godin-Dubois, Kevin, et al.
Published: (2024)
Toddlers do not preferentially transmit generalizable information to others
by: Didar Karadağ, et al.
Published: (2024)
by: Didar Karadağ, et al.
Published: (2024)
Borrowing from anything: A generalizable framework for reference-guided instance editing
by: Zhou, Shengxiao, et al.
Published: (2025)
by: Zhou, Shengxiao, et al.
Published: (2025)
Model Tells You Where to Merge: Adaptive KV Cache Merging for LLMs on Long-Context Tasks
by: Wang, Zheng, et al.
Published: (2024)
by: Wang, Zheng, et al.
Published: (2024)
GRF-based Predictive Flocking Control with Dynamic Pattern Formation
by: Yu, Chenghao, et al.
Published: (2024)
by: Yu, Chenghao, et al.
Published: (2024)
Toward Reliable Scientific Hypothesis Generation: Evaluating Truthfulness and Hallucination in Large Language Models
by: Xiong, Guangzhi, et al.
Published: (2025)
by: Xiong, Guangzhi, et al.
Published: (2025)
GeneGPT: Augmenting Large Language Models with Domain Tools for Improved Access to Biomedical Information
by: Jin, Qiao, et al.
Published: (2023)
by: Jin, Qiao, et al.
Published: (2023)
Adversarial Attacks on Large Language Models in Medicine
by: Yang, Yifan, et al.
Published: (2024)
by: Yang, Yifan, et al.
Published: (2024)
NAS-PINNv2: Improved neural architecture search framework for physics-informed neural networks in low-temperature plasma simulation
by: Wang, Yifan, et al.
Published: (2025)
by: Wang, Yifan, et al.
Published: (2025)
Energy harvester reliability study by Gaidai reliability method
by: Oleg Gaidai, et al.
Published: (2024)
by: Oleg Gaidai, et al.
Published: (2024)
PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference
by: Zhao, Yushu, et al.
Published: (2025)
by: Zhao, Yushu, et al.
Published: (2025)
MedClarify: An information-seeking AI agent for medical diagnosis with case-specific follow-up questions
by: Wong, Hui Min, et al.
Published: (2026)
by: Wong, Hui Min, et al.
Published: (2026)
Communicating socially acceptable risk judgments: The role of impression information insufficiency in the risk information seeking and processing model
by: Timothy K. F. Fung, et al.
Published: (2024)
by: Timothy K. F. Fung, et al.
Published: (2024)
Entry-level guide to the use of large language models for medical research
by: Jin, Qiao, et al.
Published: (2024)
by: Jin, Qiao, et al.
Published: (2024)
Similar Items
-
Improving Retrieval-Augmented Generation in Medicine with Iterative Follow-up Questions
by: Xiong, Guangzhi, et al.
Published: (2024) -
Benchmarking Retrieval-Augmented Generation for Medicine
by: Xiong, Guangzhi, et al.
Published: (2024) -
MedCite: Can Language Models Generate Verifiable Text for Medicine?
by: Wang, Xiao, et al.
Published: (2025) -
Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models
by: Xiong, Guangzhi, et al.
Published: (2026) -
Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning
by: Fang, Yin, et al.
Published: (2025)