Can AI Perceive Physical Danger and Intervene?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jindal, Abhishek, Kalashnikov, Dmitry, Hofer, R. Alex, Chang, Oscar, Garikapati, Divya, Majumdar, Anirudha, Sermanet, Pierre, Sindhwani, Vikas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generating Robot Constitutions & Benchmarks for Semantic Safety
von: Sermanet, Pierre, et al.
Veröffentlicht: (2025)
von: Sermanet, Pierre, et al.
Veröffentlicht: (2025)
SciFi-Benchmark: Leveraging Science Fiction To Improve Robot Behavior
von: Sermanet, Pierre, et al.
Veröffentlicht: (2025)
von: Sermanet, Pierre, et al.
Veröffentlicht: (2025)
Predictive Red Teaming: Breaking Policies Without Breaking Robots
von: Majumdar, Anirudha, et al.
Veröffentlicht: (2025)
von: Majumdar, Anirudha, et al.
Veröffentlicht: (2025)
Autonomous Vehicles: Evolution of Artificial Intelligence and Learning Algorithms
von: Garikapati, Divya, et al.
Veröffentlicht: (2024)
von: Garikapati, Divya, et al.
Veröffentlicht: (2024)
Deceptive Risk Minimization: Out-of-Distribution Generalization by Deceiving Distribution Shift Detectors
von: Majumdar, Anirudha
Veröffentlicht: (2025)
von: Majumdar, Anirudha
Veröffentlicht: (2025)
FTA generation using GenAI with an Autonomy sensor Usecase
von: Shetiya, Sneha Sudhir, et al.
Veröffentlicht: (2024)
von: Shetiya, Sneha Sudhir, et al.
Veröffentlicht: (2024)
Thinking Forward and Backward: Effective Backward Planning with Large Language Models
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
How Confident are Video Models? Empowering Video Models to Express their Uncertainty
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
Red Teaming AI Red Teaming
von: Majumdar, Subhabrata, et al.
Veröffentlicht: (2025)
von: Majumdar, Subhabrata, et al.
Veröffentlicht: (2025)
Learning to Intervene on Concept Bottlenecks
von: Steinmann, David, et al.
Veröffentlicht: (2023)
von: Steinmann, David, et al.
Veröffentlicht: (2023)
The Hidden Dangers of Browsing AI Agents
von: Mudryi, Mykyta, et al.
Veröffentlicht: (2025)
von: Mudryi, Mykyta, et al.
Veröffentlicht: (2025)
Gen AI in Automotive: Applications, Challenges, and Opportunities with a Case study on In-Vehicle Experience
von: Shinde, Chaitanya, et al.
Veröffentlicht: (2025)
von: Shinde, Chaitanya, et al.
Veröffentlicht: (2025)
STEER: Flexible Robotic Manipulation via Dense Language Grounding
von: Smith, Laura, et al.
Veröffentlicht: (2024)
von: Smith, Laura, et al.
Veröffentlicht: (2024)
A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions
von: Shorinwa, Ola, et al.
Veröffentlicht: (2024)
von: Shorinwa, Ola, et al.
Veröffentlicht: (2024)
Technical Requirements for Halting Dangerous AI Activities
von: Barnett, Peter, et al.
Veröffentlicht: (2025)
von: Barnett, Peter, et al.
Veröffentlicht: (2025)
Evaluating Gemini Robotics Policies in a Veo World Simulator
von: Gemini Robotics Team, et al.
Veröffentlicht: (2025)
von: Gemini Robotics Team, et al.
Veröffentlicht: (2025)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
A Short Note on Evaluating RepNet for Temporal Repetition Counting in Videos
von: Dwibedi, Debidatta, et al.
Veröffentlicht: (2024)
von: Dwibedi, Debidatta, et al.
Veröffentlicht: (2024)
Flash STU: Fast Spectral Transform Units
von: Liu, Y. Isabel, et al.
Veröffentlicht: (2024)
von: Liu, Y. Isabel, et al.
Veröffentlicht: (2024)
The AI Risk Spectrum: From Dangerous Capabilities to Existential Threats
von: Grey, Markov, et al.
Veröffentlicht: (2025)
von: Grey, Markov, et al.
Veröffentlicht: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
Fighting AI with AI: Leveraging Foundation Models for Assuring AI-Enabled Safety-Critical Systems
von: Mavridou, Anastasia, et al.
Veröffentlicht: (2025)
von: Mavridou, Anastasia, et al.
Veröffentlicht: (2025)
Clawed and Dangerous: Can We Trust Open Agentic Systems?
von: Chen, Shiping, et al.
Veröffentlicht: (2026)
von: Chen, Shiping, et al.
Veröffentlicht: (2026)
V-CEM: Bridging Performance and Intervenability in Concept-based Models
von: De Santis, Francesco, et al.
Veröffentlicht: (2025)
von: De Santis, Francesco, et al.
Veröffentlicht: (2025)
Can LLMs Perceive Time? An Empirical Investigation
von: Garikaparthi, Aniketh
Veröffentlicht: (2026)
von: Garikaparthi, Aniketh
Veröffentlicht: (2026)
Can Reinforcement Learning Unlock the Hidden Dangers in Aligned Large Language Models?
von: Karkevandi, Mohammad Bahrami, et al.
Veröffentlicht: (2024)
von: Karkevandi, Mohammad Bahrami, et al.
Veröffentlicht: (2024)
ResearStudio: A Human-Intervenable Framework for Building Controllable Deep-Research Agents
von: Yang, Linyi, et al.
Veröffentlicht: (2025)
von: Yang, Linyi, et al.
Veröffentlicht: (2025)
Temporal Reasoning in AI systems
von: Sharma, Abhishek
Veröffentlicht: (2025)
von: Sharma, Abhishek
Veröffentlicht: (2025)
YETI (YET to Intervene) Proactive Interventions by Multimodal AI Agents in Augmented Reality Tasks
von: Bandyopadhyay, Saptarashmi, et al.
Veröffentlicht: (2025)
von: Bandyopadhyay, Saptarashmi, et al.
Veröffentlicht: (2025)
Identifying Intervenable and Interpretable Features via Orthogonality Regularization
von: Miller, Moritz, et al.
Veröffentlicht: (2026)
von: Miller, Moritz, et al.
Veröffentlicht: (2026)
Judging by Appearances? Auditing and Intervening Vision-Language Models for Bail Prediction
von: Basu, Sagnik, et al.
Veröffentlicht: (2025)
von: Basu, Sagnik, et al.
Veröffentlicht: (2025)
From Ethical Declarations to Provable Independence: An Ontology-Driven Optimal-Transport Framework for Certifiably Fair AI Systems
von: Bhattacharya, Sukriti, et al.
Veröffentlicht: (2025)
von: Bhattacharya, Sukriti, et al.
Veröffentlicht: (2025)
RT-H: Action Hierarchies Using Language
von: Belkhale, Suneel, et al.
Veröffentlicht: (2024)
von: Belkhale, Suneel, et al.
Veröffentlicht: (2024)
IntervenGen: Interventional Data Generation for Robust and Data-Efficient Robot Imitation Learning
von: Hoque, Ryan, et al.
Veröffentlicht: (2024)
von: Hoque, Ryan, et al.
Veröffentlicht: (2024)
Reliable and Scalable Robot Policy Evaluation with Imperfect Simulators
von: Badithela, Apurva, et al.
Veröffentlicht: (2025)
von: Badithela, Apurva, et al.
Veröffentlicht: (2025)
Structured Unrestricted-Rank Matrices for Parameter Efficient Fine-tuning
von: Sehanobish, Arijit, et al.
Veröffentlicht: (2024)
von: Sehanobish, Arijit, et al.
Veröffentlicht: (2024)
Key Safety Design Overview in AI-driven Autonomous Vehicles
von: Vyas, Vikas, et al.
Veröffentlicht: (2024)
von: Vyas, Vikas, et al.
Veröffentlicht: (2024)
Concept Layers: Enhancing Interpretability and Intervenability via LLM Conceptualization
von: Bidusa, Or Raphael, et al.
Veröffentlicht: (2025)
von: Bidusa, Or Raphael, et al.
Veröffentlicht: (2025)
LAP: Language-Action Pre-Training Enables Zero-shot Cross-Embodiment Transfer
von: Zha, Lihan, et al.
Veröffentlicht: (2026)
von: Zha, Lihan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Generating Robot Constitutions & Benchmarks for Semantic Safety
von: Sermanet, Pierre, et al.
Veröffentlicht: (2025) -
SciFi-Benchmark: Leveraging Science Fiction To Improve Robot Behavior
von: Sermanet, Pierre, et al.
Veröffentlicht: (2025) -
Predictive Red Teaming: Breaking Policies Without Breaking Robots
von: Majumdar, Anirudha, et al.
Veröffentlicht: (2025) -
Autonomous Vehicles: Evolution of Artificial Intelligence and Learning Algorithms
von: Garikapati, Divya, et al.
Veröffentlicht: (2024) -
Deceptive Risk Minimization: Out-of-Distribution Generalization by Deceiving Distribution Shift Detectors
von: Majumdar, Anirudha
Veröffentlicht: (2025)