Self-Induced Outcome Potential: Turn-Level Credit Assignment for Agents without Verifiers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Senkang, Dai, Yong, Han, Xudong, Fang, Zhengru, Zhao, Yuzhi, Kwong, Sam Tak Wu, Fang, Yuguang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimizing Agentic Reasoning with Retrieval via Synthetic Semantic Information Gain Reward
von: Hu, Senkang, et al.
Veröffentlicht: (2026)
von: Hu, Senkang, et al.
Veröffentlicht: (2026)
AgentsCoMerge: Large Language Model Empowered Collaborative Decision Making for Ramp Merging
von: Hu, Senkang, et al.
Veröffentlicht: (2024)
von: Hu, Senkang, et al.
Veröffentlicht: (2024)
Towards Full-scene Domain Generalization in Multi-agent Collaborative Bird's Eye View Segmentation for Connected and Autonomous Driving
von: Hu, Senkang, et al.
Veröffentlicht: (2023)
von: Hu, Senkang, et al.
Veröffentlicht: (2023)
Distribution-Aligned Decoding for Efficient LLM Task Adaptation
von: Hu, Senkang, et al.
Veröffentlicht: (2025)
von: Hu, Senkang, et al.
Veröffentlicht: (2025)
Directed-CP: Directed Collaborative Perception for Connected and Autonomous Vehicles via Proactive Attention
von: Tao, Yihang, et al.
Veröffentlicht: (2024)
von: Tao, Yihang, et al.
Veröffentlicht: (2024)
Task-Aware Parameter-Efficient Fine-Tuning of Large Pre-Trained Models at the Edge
von: Hu, Senkang, et al.
Veröffentlicht: (2025)
von: Hu, Senkang, et al.
Veröffentlicht: (2025)
DSP-Reg: Domain-Sensitive Parameter Regularization for Robust Domain Generalization
von: Han, Xudong, et al.
Veröffentlicht: (2026)
von: Han, Xudong, et al.
Veröffentlicht: (2026)
AgentsCoDriver: Large Language Model Empowered Collaborative Driving with Lifelong Learning
von: Hu, Senkang, et al.
Veröffentlicht: (2024)
von: Hu, Senkang, et al.
Veröffentlicht: (2024)
Collaborative Perception for Connected and Autonomous Driving: Challenges, Possible Solutions and Opportunities
von: Hu, Senkang, et al.
Veröffentlicht: (2024)
von: Hu, Senkang, et al.
Veröffentlicht: (2024)
Task-Oriented Semantic Compression for Localization at the Network Edge
von: Fang, Zhengru, et al.
Veröffentlicht: (2025)
von: Fang, Zhengru, et al.
Veröffentlicht: (2025)
V2XCrafter: Learning to Generate Driving Scene Across Agents
von: Tao, Yihang, et al.
Veröffentlicht: (2026)
von: Tao, Yihang, et al.
Veröffentlicht: (2026)
CP-Guard+: A New Paradigm for Malicious Agent Detection and Defense in Collaborative Perception
von: Hu, Senkang, et al.
Veröffentlicht: (2025)
von: Hu, Senkang, et al.
Veröffentlicht: (2025)
Learning Mutual View Information Graph for Adaptive Adversarial Collaborative Perception
von: Tao, Yihang, et al.
Veröffentlicht: (2026)
von: Tao, Yihang, et al.
Veröffentlicht: (2026)
CP-uniGuard: A Unified, Probability-Agnostic, and Adaptive Framework for Malicious Agent Detection and Defense in Multi-Agent Embodied Perception Systems
von: Hu, Senkang, et al.
Veröffentlicht: (2025)
von: Hu, Senkang, et al.
Veröffentlicht: (2025)
Prioritized Information Bottleneck Theoretic Framework with Distributed Online Learning for Edge Video Analytics
von: Fang, Zhengru, et al.
Veröffentlicht: (2024)
von: Fang, Zhengru, et al.
Veröffentlicht: (2024)
PIB: Prioritized Information Bottleneck Framework for Collaborative Edge Video Analytics
von: Fang, Zhengru, et al.
Veröffentlicht: (2024)
von: Fang, Zhengru, et al.
Veröffentlicht: (2024)
UAV-enabled Computing Power Networks: Task Completion Probability Analysis
von: Deng, Yiqin, et al.
Veröffentlicht: (2025)
von: Deng, Yiqin, et al.
Veröffentlicht: (2025)
Sense4FL: Vehicular Crowdsensing Enhanced Federated Learning for Object Detection in Autonomous Driving
von: Ma, Yanan, et al.
Veröffentlicht: (2025)
von: Ma, Yanan, et al.
Veröffentlicht: (2025)
CP-Guard: Malicious Agent Detection and Defense in Collaborative Bird's Eye View Perception
von: Hu, Senkang, et al.
Veröffentlicht: (2024)
von: Hu, Senkang, et al.
Veröffentlicht: (2024)
Adaptive Communications in Collaborative Perception with Domain Alignment for Autonomous Driving
von: Hu, Senkang, et al.
Veröffentlicht: (2023)
von: Hu, Senkang, et al.
Veröffentlicht: (2023)
Channel-Aware Throughput Maximization for Cooperative Data Fusion in CAV
von: An, Haonan, et al.
Veröffentlicht: (2024)
von: An, Haonan, et al.
Veröffentlicht: (2024)
Agent-Centric Observation Adaptation for Robust Visual Control under Dynamic Perturbations
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
Unified Context Evolution for LLM Agents
von: Zhu, Zixuan, et al.
Veröffentlicht: (2026)
von: Zhu, Zixuan, et al.
Veröffentlicht: (2026)
UAV-enabled Computing Power Networks: Design and Performance Analysis under Energy Constraints
von: Deng, Yiqin, et al.
Veröffentlicht: (2026)
von: Deng, Yiqin, et al.
Veröffentlicht: (2026)
PACP: Priority-Aware Collaborative Perception for Connected and Autonomous Vehicles
von: Fang, Zhengru, et al.
Veröffentlicht: (2024)
von: Fang, Zhengru, et al.
Veröffentlicht: (2024)
Inference-Time Budget Control for LLM Search Agents
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
Proximity-Based Multi-Turn Optimization: Practical Credit Assignment for LLM Agent Training
von: Fang, Yangyi, et al.
Veröffentlicht: (2026)
von: Fang, Yangyi, et al.
Veröffentlicht: (2026)
Skill-Conditioned Gated Self-Distillation for LLM Reasoning
von: Huang, Jiazhen, et al.
Veröffentlicht: (2026)
von: Huang, Jiazhen, et al.
Veröffentlicht: (2026)
GCP: Guarded Collaborative Perception with Spatial-Temporal Aware Malicious Agent Detection
von: Tao, Yihang, et al.
Veröffentlicht: (2025)
von: Tao, Yihang, et al.
Veröffentlicht: (2025)
Not All Turns Matter: Credit Assignment for Multi-Turn Jailbreaking
von: He, Zhida, et al.
Veröffentlicht: (2026)
von: He, Zhida, et al.
Veröffentlicht: (2026)
FRUC: Feedforward Dynamic Scene Reconstruction from Uncalibrated Collaborative Driving Views
von: Tao, Yihang, et al.
Veröffentlicht: (2026)
von: Tao, Yihang, et al.
Veröffentlicht: (2026)
Decoder Gradient Shield: Provable and High-Fidelity Prevention of Gradient-Based Box-Free Watermark Removal
von: An, Haonan, et al.
Veröffentlicht: (2025)
von: An, Haonan, et al.
Veröffentlicht: (2025)
Shared Spatial Memory Through Predictive Coding
von: Fang, Zhengru, et al.
Veröffentlicht: (2025)
von: Fang, Zhengru, et al.
Veröffentlicht: (2025)
IC3M: In-Car Multimodal Multi-object Monitoring for Abnormal Status of Both Driver and Passengers
von: Fang, Zihan, et al.
Veröffentlicht: (2024)
von: Fang, Zihan, et al.
Veröffentlicht: (2024)
Dynamic Uncertainty-aware Multimodal Fusion for Outdoor Health Monitoring
von: Fang, Zihan, et al.
Veröffentlicht: (2025)
von: Fang, Zihan, et al.
Veröffentlicht: (2025)
RAISE: Optimizing RIS Placement to Maximize Task Throughput in Multi-Server Vehicular Edge Computing
von: Ma, Yanan, et al.
Veröffentlicht: (2025)
von: Ma, Yanan, et al.
Veröffentlicht: (2025)
Secure Traffic Sign Recognition: An Attention-Enabled Universal Image Inpainting Mechanism against Light Patch Attacks
von: Cao, Hangcheng, et al.
Veröffentlicht: (2024)
von: Cao, Hangcheng, et al.
Veröffentlicht: (2024)
AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment
von: Wei, Zhenlin, et al.
Veröffentlicht: (2026)
von: Wei, Zhenlin, et al.
Veröffentlicht: (2026)
SmartCooper: Vehicular Collaborative Perception with Adaptive Fusion and Judger Mechanism
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Optimizing Agentic Reasoning with Retrieval via Synthetic Semantic Information Gain Reward
von: Hu, Senkang, et al.
Veröffentlicht: (2026) -
AgentsCoMerge: Large Language Model Empowered Collaborative Decision Making for Ramp Merging
von: Hu, Senkang, et al.
Veröffentlicht: (2024) -
Towards Full-scene Domain Generalization in Multi-agent Collaborative Bird's Eye View Segmentation for Connected and Autonomous Driving
von: Hu, Senkang, et al.
Veröffentlicht: (2023) -
Distribution-Aligned Decoding for Efficient LLM Task Adaptation
von: Hu, Senkang, et al.
Veröffentlicht: (2025) -
Directed-CP: Directed Collaborative Perception for Connected and Autonomous Vehicles via Proactive Attention
von: Tao, Yihang, et al.
Veröffentlicht: (2024)