Optimizing Agentic Reasoning with Retrieval via Synthetic Semantic Information Gain Reward
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Hu, Senkang, Dai, Yong, Zhao, Yuzhi, Tao, Yihang, Guo, Yu, Fang, Zhengru, Kwong, Sam Tak Wu, Fang, Yuguang |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Self-Induced Outcome Potential: Turn-Level Credit Assignment for Agents without Verifiers
par: Hu, Senkang, et autres
Publié: (2026)
par: Hu, Senkang, et autres
Publié: (2026)
Directed-CP: Directed Collaborative Perception for Connected and Autonomous Vehicles via Proactive Attention
par: Tao, Yihang, et autres
Publié: (2024)
par: Tao, Yihang, et autres
Publié: (2024)
Learning Mutual View Information Graph for Adaptive Adversarial Collaborative Perception
par: Tao, Yihang, et autres
Publié: (2026)
par: Tao, Yihang, et autres
Publié: (2026)
Task-Aware Parameter-Efficient Fine-Tuning of Large Pre-Trained Models at the Edge
par: Hu, Senkang, et autres
Publié: (2025)
par: Hu, Senkang, et autres
Publié: (2025)
Towards Full-scene Domain Generalization in Multi-agent Collaborative Bird's Eye View Segmentation for Connected and Autonomous Driving
par: Hu, Senkang, et autres
Publié: (2023)
par: Hu, Senkang, et autres
Publié: (2023)
Distribution-Aligned Decoding for Efficient LLM Task Adaptation
par: Hu, Senkang, et autres
Publié: (2025)
par: Hu, Senkang, et autres
Publié: (2025)
Task-Oriented Semantic Compression for Localization at the Network Edge
par: Fang, Zhengru, et autres
Publié: (2025)
par: Fang, Zhengru, et autres
Publié: (2025)
AgentsCoMerge: Large Language Model Empowered Collaborative Decision Making for Ramp Merging
par: Hu, Senkang, et autres
Publié: (2024)
par: Hu, Senkang, et autres
Publié: (2024)
DSP-Reg: Domain-Sensitive Parameter Regularization for Robust Domain Generalization
par: Han, Xudong, et autres
Publié: (2026)
par: Han, Xudong, et autres
Publié: (2026)
V2XCrafter: Learning to Generate Driving Scene Across Agents
par: Tao, Yihang, et autres
Publié: (2026)
par: Tao, Yihang, et autres
Publié: (2026)
CP-Guard+: A New Paradigm for Malicious Agent Detection and Defense in Collaborative Perception
par: Hu, Senkang, et autres
Publié: (2025)
par: Hu, Senkang, et autres
Publié: (2025)
CP-Guard: Malicious Agent Detection and Defense in Collaborative Bird's Eye View Perception
par: Hu, Senkang, et autres
Publié: (2024)
par: Hu, Senkang, et autres
Publié: (2024)
Agent-Centric Observation Adaptation for Robust Visual Control under Dynamic Perturbations
par: Fang, Zhengru, et autres
Publié: (2026)
par: Fang, Zhengru, et autres
Publié: (2026)
CP-uniGuard: A Unified, Probability-Agnostic, and Adaptive Framework for Malicious Agent Detection and Defense in Multi-Agent Embodied Perception Systems
par: Hu, Senkang, et autres
Publié: (2025)
par: Hu, Senkang, et autres
Publié: (2025)
Collaborative Perception for Connected and Autonomous Driving: Challenges, Possible Solutions and Opportunities
par: Hu, Senkang, et autres
Publié: (2024)
par: Hu, Senkang, et autres
Publié: (2024)
FRUC: Feedforward Dynamic Scene Reconstruction from Uncalibrated Collaborative Driving Views
par: Tao, Yihang, et autres
Publié: (2026)
par: Tao, Yihang, et autres
Publié: (2026)
PIB: Prioritized Information Bottleneck Framework for Collaborative Edge Video Analytics
par: Fang, Zhengru, et autres
Publié: (2024)
par: Fang, Zhengru, et autres
Publié: (2024)
Prioritized Information Bottleneck Theoretic Framework with Distributed Online Learning for Edge Video Analytics
par: Fang, Zhengru, et autres
Publié: (2024)
par: Fang, Zhengru, et autres
Publié: (2024)
AgentsCoDriver: Large Language Model Empowered Collaborative Driving with Lifelong Learning
par: Hu, Senkang, et autres
Publié: (2024)
par: Hu, Senkang, et autres
Publié: (2024)
UAV-enabled Computing Power Networks: Task Completion Probability Analysis
par: Deng, Yiqin, et autres
Publié: (2025)
par: Deng, Yiqin, et autres
Publié: (2025)
Sense4FL: Vehicular Crowdsensing Enhanced Federated Learning for Object Detection in Autonomous Driving
par: Ma, Yanan, et autres
Publié: (2025)
par: Ma, Yanan, et autres
Publié: (2025)
Inference-Time Budget Control for LLM Search Agents
par: Fang, Zhengru, et autres
Publié: (2026)
par: Fang, Zhengru, et autres
Publié: (2026)
UAV-enabled Computing Power Networks: Design and Performance Analysis under Energy Constraints
par: Deng, Yiqin, et autres
Publié: (2026)
par: Deng, Yiqin, et autres
Publié: (2026)
GCP: Guarded Collaborative Perception with Spatial-Temporal Aware Malicious Agent Detection
par: Tao, Yihang, et autres
Publié: (2025)
par: Tao, Yihang, et autres
Publié: (2025)
Adaptive Communications in Collaborative Perception with Domain Alignment for Autonomous Driving
par: Hu, Senkang, et autres
Publié: (2023)
par: Hu, Senkang, et autres
Publié: (2023)
Channel-Aware Throughput Maximization for Cooperative Data Fusion in CAV
par: An, Haonan, et autres
Publié: (2024)
par: An, Haonan, et autres
Publié: (2024)
CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs
par: Fang, Zhengru, et autres
Publié: (2026)
par: Fang, Zhengru, et autres
Publié: (2026)
Dynamic Uncertainty-aware Multimodal Fusion for Outdoor Health Monitoring
par: Fang, Zihan, et autres
Publié: (2025)
par: Fang, Zihan, et autres
Publié: (2025)
PACP: Priority-Aware Collaborative Perception for Connected and Autonomous Vehicles
par: Fang, Zhengru, et autres
Publié: (2024)
par: Fang, Zhengru, et autres
Publié: (2024)
Birdcast: Interest-aware BEV Multicasting for Infrastructure-assisted Collaborative Perception
par: Ma, Yanan, et autres
Publié: (2026)
par: Ma, Yanan, et autres
Publié: (2026)
Skill-Conditioned Gated Self-Distillation for LLM Reasoning
par: Huang, Jiazhen, et autres
Publié: (2026)
par: Huang, Jiazhen, et autres
Publié: (2026)
Unified Context Evolution for LLM Agents
par: Zhu, Zixuan, et autres
Publié: (2026)
par: Zhu, Zixuan, et autres
Publié: (2026)
R-ACP: Real-Time Adaptive Collaborative Perception Leveraging Robust Task-Oriented Communications
par: Fang, Zhengru, et autres
Publié: (2024)
par: Fang, Zhengru, et autres
Publié: (2024)
HFedMoE: Resource-aware Heterogeneous Federated Learning with Mixture-of-Experts
par: Fang, Zihan, et autres
Publié: (2026)
par: Fang, Zihan, et autres
Publié: (2026)
Shared Spatial Memory Through Predictive Coding
par: Fang, Zhengru, et autres
Publié: (2025)
par: Fang, Zhengru, et autres
Publié: (2025)
RAISE: Optimizing RIS Placement to Maximize Task Throughput in Multi-Server Vehicular Edge Computing
par: Ma, Yanan, et autres
Publié: (2025)
par: Ma, Yanan, et autres
Publié: (2025)
Decoder Gradient Shield: Provable and High-Fidelity Prevention of Gradient-Based Box-Free Watermark Removal
par: An, Haonan, et autres
Publié: (2025)
par: An, Haonan, et autres
Publié: (2025)
Removing Box-Free Watermarks for Image-to-Image Models via Query-Based Reverse Engineering
par: An, Haonan, et autres
Publié: (2025)
par: An, Haonan, et autres
Publié: (2025)
IC3M: In-Car Multimodal Multi-object Monitoring for Abnormal Status of Both Driver and Passengers
par: Fang, Zihan, et autres
Publié: (2024)
par: Fang, Zihan, et autres
Publié: (2024)
Secure Traffic Sign Recognition: An Attention-Enabled Universal Image Inpainting Mechanism against Light Patch Attacks
par: Cao, Hangcheng, et autres
Publié: (2024)
par: Cao, Hangcheng, et autres
Publié: (2024)
Documents similaires
-
Self-Induced Outcome Potential: Turn-Level Credit Assignment for Agents without Verifiers
par: Hu, Senkang, et autres
Publié: (2026) -
Directed-CP: Directed Collaborative Perception for Connected and Autonomous Vehicles via Proactive Attention
par: Tao, Yihang, et autres
Publié: (2024) -
Learning Mutual View Information Graph for Adaptive Adversarial Collaborative Perception
par: Tao, Yihang, et autres
Publié: (2026) -
Task-Aware Parameter-Efficient Fine-Tuning of Large Pre-Trained Models at the Edge
par: Hu, Senkang, et autres
Publié: (2025) -
Towards Full-scene Domain Generalization in Multi-agent Collaborative Bird's Eye View Segmentation for Connected and Autonomous Driving
par: Hu, Senkang, et autres
Publié: (2023)