Uncovering Systemic and Environment Errors in Autonomous Systems Using Differential Testing
Fuente:
arXiv
Saved in:
| Main Authors: | Anand, Yashwanthi, Mehta, Rahil P, Motwani, Manish, Saisubramanian, Sandhya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Objective Planning with Contextual Lexicographic Reward Preferences
by: Rustagi, Pulkit, et al.
Published: (2025)
by: Rustagi, Pulkit, et al.
Published: (2025)
Adaptive Querying for Reward Learning from Human Feedback
by: Anand, Yashwanthi, et al.
Published: (2024)
by: Anand, Yashwanthi, et al.
Published: (2024)
Learning Transferable Latent User Preferences for Human-Aligned Decision Making
by: Hyk, Alina, et al.
Published: (2026)
by: Hyk, Alina, et al.
Published: (2026)
Learning with Expert Abstractions for Efficient Multi-Task Continuous Control
by: Jewett, Jeff, et al.
Published: (2025)
by: Jewett, Jeff, et al.
Published: (2025)
Calibrating Biophysical Models for Grape Phenology Prediction via Multi-Task Learning
by: Solow, William, et al.
Published: (2025)
by: Solow, William, et al.
Published: (2025)
WOFOSTGym: A Crop Simulator for Learning Annual and Perennial Crop Management Strategies
by: Solow, William, et al.
Published: (2025)
by: Solow, William, et al.
Published: (2025)
Multi-Objective Multi-Agent Path Finding with Lexicographic Cost Preferences
by: Rustagi, Pulkit, et al.
Published: (2025)
by: Rustagi, Pulkit, et al.
Published: (2025)
Mitigating Side Effects in Multi-Agent Systems Using Blame Assignment
by: Rustagi, Pulkit, et al.
Published: (2024)
by: Rustagi, Pulkit, et al.
Published: (2024)
A Hybrid Modeling Framework for Crop Prediction Tasks via Dynamic Parameter Calibration and Multi-Task Learning
by: Solow, William, et al.
Published: (2026)
by: Solow, William, et al.
Published: (2026)
FedSCAM (Federated Sharpness-Aware Minimization with Clustered Aggregation and Modulation): Scam-resistant SAM for Robust Federated Optimization in Heterogeneous Environments
by: Rahil, Sameer, et al.
Published: (2025)
by: Rahil, Sameer, et al.
Published: (2025)
A Digital Twin Framework for Metamorphic Testing of Autonomous Driving Systems Using Generative Model
by: Zhang, Tony, et al.
Published: (2025)
by: Zhang, Tony, et al.
Published: (2025)
An LLM-Driven Closed-Loop Autonomous Learning Framework for Robots Facing Uncovered Tasks in Open Environments
by: Su, Hong
Published: (2026)
by: Su, Hong
Published: (2026)
Beyond Accuracy: A Multi-Dimensional Framework for Evaluating Enterprise Agentic AI Systems
by: Mehta, Sushant
Published: (2025)
by: Mehta, Sushant
Published: (2025)
An Intent Modeling and Inference Framework for Autonomous and Remotely Piloted Aerial Systems
by: Kaza, Kesav, et al.
Published: (2024)
by: Kaza, Kesav, et al.
Published: (2024)
Make Full Use of Testing Information: An Integrated Accelerated Testing and Evaluation Method for Autonomous Driving Systems
by: Wu, Xinzheng, et al.
Published: (2025)
by: Wu, Xinzheng, et al.
Published: (2025)
Environment-Grounded Multi-Agent Workflow for Autonomous Penetration Testing
by: Somma, Michael, et al.
Published: (2026)
by: Somma, Michael, et al.
Published: (2026)
STCLocker: Deadlock Avoidance Testing for Autonomous Driving Systems
by: Cheng, Mingfei, et al.
Published: (2025)
by: Cheng, Mingfei, et al.
Published: (2025)
Generative AI for Testing of Autonomous Driving Systems: A Survey
by: Song, Qunying, et al.
Published: (2025)
by: Song, Qunying, et al.
Published: (2025)
Autonomous Robot for Disaster Mapping and Victim Localization
by: Potter, Michael, et al.
Published: (2024)
by: Potter, Michael, et al.
Published: (2024)
Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
by: Putta, Pranav, et al.
Published: (2024)
by: Putta, Pranav, et al.
Published: (2024)
Advancing Autonomous Driving System Testing: Demands, Challenges, and Future Directions
by: Liao, Yihan, et al.
Published: (2025)
by: Liao, Yihan, et al.
Published: (2025)
Boundary State Generation for Testing and Improvement of Autonomous Driving Systems
by: Biagiola, Matteo, et al.
Published: (2023)
by: Biagiola, Matteo, et al.
Published: (2023)
Adaptive Monitoring and Real-World Evaluation of Agentic AI Systems
by: Shukla, Manish
Published: (2025)
by: Shukla, Manish
Published: (2025)
Git for Sketches: An Intelligent Tracking System for Capturing Design Evolution
by: B, Sankar, et al.
Published: (2026)
by: B, Sankar, et al.
Published: (2026)
HELM: A Human-Centered Evaluation Framework for LLM-Powered Recommender Systems
by: Mehta, Sushant
Published: (2026)
by: Mehta, Sushant
Published: (2026)
MEMTRACK: Evaluating Long-Term Memory and State Tracking in Multi-Platform Dynamic Agent Environments
by: Deshpande, Darshan, et al.
Published: (2025)
by: Deshpande, Darshan, et al.
Published: (2025)
A Proposed Biomedical Data Policy Framework to Reduce Fragmentation, Improve Quality, and Incentivize Sharing in Indian Healthcare in the era of Artificial Intelligence and Digital Health
by: Mehta, Nikhil, et al.
Published: (2026)
by: Mehta, Nikhil, et al.
Published: (2026)
Autonomous Building Cyber-Physical Systems Using Decentralized Autonomous Organizations, Digital Twins, and Large Language Model
by: Ly, Reachsak, et al.
Published: (2024)
by: Ly, Reachsak, et al.
Published: (2024)
Mission-Aligned Learning-Informed Control of Autonomous Systems: Formulation and Foundations
by: Kungurtsev, Vyacheslav, et al.
Published: (2025)
by: Kungurtsev, Vyacheslav, et al.
Published: (2025)
ProMAS: Proactive Error Forecasting for Multi-Agent Systems Using Markov Transition Dynamics
by: Zhao, Xinkui, et al.
Published: (2026)
by: Zhao, Xinkui, et al.
Published: (2026)
Measuring Error Alignment for Decision-Making Systems
by: Xu, Binxia, et al.
Published: (2024)
by: Xu, Binxia, et al.
Published: (2024)
ASIA: an Autonomous System Identification Agent
by: Piga, Dario, et al.
Published: (2026)
by: Piga, Dario, et al.
Published: (2026)
Bridging the Gap between Real-world and Synthetic Images for Testing Autonomous Driving Systems
by: Amini, Mohammad Hossein, et al.
Published: (2024)
by: Amini, Mohammad Hossein, et al.
Published: (2024)
CanaryBench: Stress Testing Privacy Leakage in Cluster-Level Conversation Summaries
by: Mehta, Deep
Published: (2026)
by: Mehta, Deep
Published: (2026)
Layered Chain-of-Thought Prompting for Multi-Agent LLM Systems: A Comprehensive Approach to Explainable Large Language Models
by: Sanwal, Manish
Published: (2025)
by: Sanwal, Manish
Published: (2025)
Commonsense Reasoning-Aided Autonomous Vehicle Systems
by: Kimbrell, Keegan
Published: (2025)
by: Kimbrell, Keegan
Published: (2025)
Unified Prediction Model for Employability in Indian Higher Education System
by: Thakar, Pooja, et al.
Published: (2024)
by: Thakar, Pooja, et al.
Published: (2024)
PEM: Perception Error Model for Virtual Testing of Autonomous Vehicles
by: Piazzoni, Andrea, et al.
Published: (2023)
by: Piazzoni, Andrea, et al.
Published: (2023)
Investigation of the Privacy Concerns in AI Systems for Young Digital Citizens: A Comparative Stakeholder Analysis
by: Campbell, Molly, et al.
Published: (2025)
by: Campbell, Molly, et al.
Published: (2025)
Verification and Validation of Autonomous Systems
by: Shetiya, Sneha Sudhir, et al.
Published: (2024)
by: Shetiya, Sneha Sudhir, et al.
Published: (2024)
Similar Items
-
Multi-Objective Planning with Contextual Lexicographic Reward Preferences
by: Rustagi, Pulkit, et al.
Published: (2025) -
Adaptive Querying for Reward Learning from Human Feedback
by: Anand, Yashwanthi, et al.
Published: (2024) -
Learning Transferable Latent User Preferences for Human-Aligned Decision Making
by: Hyk, Alina, et al.
Published: (2026) -
Learning with Expert Abstractions for Efficient Multi-Task Continuous Control
by: Jewett, Jeff, et al.
Published: (2025) -
Calibrating Biophysical Models for Grape Phenology Prediction via Multi-Task Learning
by: Solow, William, et al.
Published: (2025)