Revisiting Out-of-Distribution Detection in Real-time Object Detection: From Benchmark Pitfalls to a New Mitigation Paradigm
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Changshun, He, Weicheng, Cheng, Chih-Hong, Huang, Xiaowei, Bensalem, Saddek |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BAM: Box Abstraction Monitors for Real-time OoD Detection in Object Detection
von: Wu, Changshun, et al.
Veröffentlicht: (2024)
von: Wu, Changshun, et al.
Veröffentlicht: (2024)
What, Indeed, is an Achievable Provable Guarantee for Learning-Enabled Safety Critical Systems
von: Bensalem, Saddek, et al.
Veröffentlicht: (2023)
von: Bensalem, Saddek, et al.
Veröffentlicht: (2023)
Runtime Monitoring and Enforcement of Conditional Fairness in Generative AIs
von: Cheng, Chih-Hong, et al.
Veröffentlicht: (2024)
von: Cheng, Chih-Hong, et al.
Veröffentlicht: (2024)
From Bias To Improved Prompts: A Case Study of Bias Mitigation of Clone Detection Models
von: Chen, QiHong, et al.
Veröffentlicht: (2025)
von: Chen, QiHong, et al.
Veröffentlicht: (2025)
Towards More Trustworthy Deep Code Models by Enabling Out-of-Distribution Detection
von: Yan, Yanfu, et al.
Veröffentlicht: (2025)
von: Yan, Yanfu, et al.
Veröffentlicht: (2025)
LoRA-BAM: Input Filtering for Fine-tuned LLMs via Boxed Abstraction Monitors over LoRA Layers
von: Wu, Changshun, et al.
Veröffentlicht: (2025)
von: Wu, Changshun, et al.
Veröffentlicht: (2025)
Taxonomy of the Retrieval System Framework: Pitfalls and Paradigms
von: Shah, Deep, et al.
Veröffentlicht: (2026)
von: Shah, Deep, et al.
Veröffentlicht: (2026)
Revisiting Code Search in a Two-Stage Paradigm
von: Hu, Fan, et al.
Veröffentlicht: (2022)
von: Hu, Fan, et al.
Veröffentlicht: (2022)
SecVulEval: Benchmarking LLMs for Real-World C/C++ Vulnerability Detection
von: Ahmed, Md Basim Uddin, et al.
Veröffentlicht: (2025)
von: Ahmed, Md Basim Uddin, et al.
Veröffentlicht: (2025)
The Promise and Pitfalls of WebAssembly: Perspectives from the Industry
von: He, Ningyu, et al.
Veröffentlicht: (2025)
von: He, Ningyu, et al.
Veröffentlicht: (2025)
Detecting and Mitigating Flakiness in REST API Fuzzing
von: Zhang, Man, et al.
Veröffentlicht: (2026)
von: Zhang, Man, et al.
Veröffentlicht: (2026)
Out of Distribution, Out of Luck: How Well Can LLMs Trained on Vulnerability Datasets Detect Top 25 CWE Weaknesses?
von: Li, Yikun, et al.
Veröffentlicht: (2025)
von: Li, Yikun, et al.
Veröffentlicht: (2025)
Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs
von: Feng, Dylan, et al.
Veröffentlicht: (2026)
von: Feng, Dylan, et al.
Veröffentlicht: (2026)
Beyond Blind Spots: Analytic Hints for Mitigating LLM-Based Evaluation Pitfalls
von: Fandina, Ora Nova, et al.
Veröffentlicht: (2025)
von: Fandina, Ora Nova, et al.
Veröffentlicht: (2025)
Benchmarking Large Language Models for Multi-Language Software Vulnerability Detection
von: Zhang, Ting, et al.
Veröffentlicht: (2025)
von: Zhang, Ting, et al.
Veröffentlicht: (2025)
Out of Distribution Detection in Self-adaptive Robots with AI-powered Digital Twins
von: Isaku, Erblin, et al.
Veröffentlicht: (2025)
von: Isaku, Erblin, et al.
Veröffentlicht: (2025)
Boosting Vulnerability Detection with Inter-function Multilateral Association Insights
von: Qiu, Shaojian, et al.
Veröffentlicht: (2025)
von: Qiu, Shaojian, et al.
Veröffentlicht: (2025)
Automated Detection and Mitigation of Dependability Failures in Healthcare Scenarios through Digital Twins
von: Guindani, Bruno, et al.
Veröffentlicht: (2026)
von: Guindani, Bruno, et al.
Veröffentlicht: (2026)
Towards Benchmarking Design Pattern Detection Under Obfuscation: Reproducing and Evaluating Attention-Based Detection Method
von: Shenoy, Manthan, et al.
Veröffentlicht: (2025)
von: Shenoy, Manthan, et al.
Veröffentlicht: (2025)
Digital Twin-based Out-of-Distribution Detection in Autonomous Vessels
von: Isaku, Erblin, et al.
Veröffentlicht: (2025)
von: Isaku, Erblin, et al.
Veröffentlicht: (2025)
From UI to Code: Mobile Ads Detection via LLM-Unified Static-Dynamic Analysis
von: Ma, Shang, et al.
Veröffentlicht: (2026)
von: Ma, Shang, et al.
Veröffentlicht: (2026)
How to Save My Gas Fees: Understanding and Detecting Real-world Gas Issues in Solidity Programs
von: He, Mengting, et al.
Veröffentlicht: (2024)
von: He, Mengting, et al.
Veröffentlicht: (2024)
Multivariate Log-based Anomaly Detection for Distributed Database
von: Zhang, Lingzhe, et al.
Veröffentlicht: (2024)
von: Zhang, Lingzhe, et al.
Veröffentlicht: (2024)
Advanced Vulnerability Scanning for Open Source Software: Detection and Mitigation of Log4j Vulnerabilities
von: Wen, Victor, et al.
Veröffentlicht: (2026)
von: Wen, Victor, et al.
Veröffentlicht: (2026)
Leveraging Self-Paced Learning for Software Vulnerability Detection
von: Cheng, Zeru, et al.
Veröffentlicht: (2025)
von: Cheng, Zeru, et al.
Veröffentlicht: (2025)
Characterizing Real-World Bugs in Tile Programs for Automated Bug Detection
von: Rathnasuriya, Ravishka, et al.
Veröffentlicht: (2026)
von: Rathnasuriya, Ravishka, et al.
Veröffentlicht: (2026)
Benchmarking and Revisiting Code Generation Assessment: A Mutation-Based Approach
von: Wang, Longtian, et al.
Veröffentlicht: (2025)
von: Wang, Longtian, et al.
Veröffentlicht: (2025)
Randomized Smoothing Meets Vision-Language Models
von: Seferis, Emmanouil, et al.
Veröffentlicht: (2025)
von: Seferis, Emmanouil, et al.
Veröffentlicht: (2025)
Detect Repair Verify for Securing LLM Generated Code: A Multi-Language Empirical Study
von: Cheng, Cheng
Veröffentlicht: (2026)
von: Cheng, Cheng
Veröffentlicht: (2026)
PATCHEVAL: A New Benchmark for Evaluating LLMs on Patching Real-World Vulnerabilities
von: Wei, Zichao, et al.
Veröffentlicht: (2025)
von: Wei, Zichao, et al.
Veröffentlicht: (2025)
Detect--Repair--Verify for LLM-Generated Code: A Multi-Language, Multi-Granularity Empirical Study
von: Cheng, Cheng
Veröffentlicht: (2026)
von: Cheng, Cheng
Veröffentlicht: (2026)
Defects4Log: Benchmarking LLMs for Logging Code Defect Detection and Reasoning
von: Wang, Xin, et al.
Veröffentlicht: (2025)
von: Wang, Xin, et al.
Veröffentlicht: (2025)
Understanding NPM Malicious Package Detection: A Benchmark-Driven Empirical Analysis
von: Guo, Wenbo, et al.
Veröffentlicht: (2026)
von: Guo, Wenbo, et al.
Veröffentlicht: (2026)
ZeroLog: Zero-Label Generalizable Cross-System Log-based Anomaly Detection
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
A Benchmark for Language Models in Real-World System Building
von: Jin, Weilin, et al.
Veröffentlicht: (2026)
von: Jin, Weilin, et al.
Veröffentlicht: (2026)
"Should I Give Up Now?" Investigating LLM Pitfalls in Software Engineering
von: Tie, Jiessie, et al.
Veröffentlicht: (2024)
von: Tie, Jiessie, et al.
Veröffentlicht: (2024)
From Prompts to Templates: A Systematic Prompt Template Analysis for Real-world LLMapps
von: Mao, Yuetian, et al.
Veröffentlicht: (2025)
von: Mao, Yuetian, et al.
Veröffentlicht: (2025)
GDPR-Bench-Android: A Benchmark for Evaluating Automated GDPR Compliance Detection in Android
von: Ran, Huaijin, et al.
Veröffentlicht: (2025)
von: Ran, Huaijin, et al.
Veröffentlicht: (2025)
From Knowledge to Noise: CTIM-Rover and the Pitfalls of Episodic Memory in Software Engineering Agents
von: Lindenbauer, Tobias, et al.
Veröffentlicht: (2025)
von: Lindenbauer, Tobias, et al.
Veröffentlicht: (2025)
HACMony: Automatically Detecting Hopping-related Audio-stream Conflict Issues on HarmonyOS
von: He, Jinlong, et al.
Veröffentlicht: (2025)
von: He, Jinlong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
BAM: Box Abstraction Monitors for Real-time OoD Detection in Object Detection
von: Wu, Changshun, et al.
Veröffentlicht: (2024) -
What, Indeed, is an Achievable Provable Guarantee for Learning-Enabled Safety Critical Systems
von: Bensalem, Saddek, et al.
Veröffentlicht: (2023) -
Runtime Monitoring and Enforcement of Conditional Fairness in Generative AIs
von: Cheng, Chih-Hong, et al.
Veröffentlicht: (2024) -
From Bias To Improved Prompts: A Case Study of Bias Mitigation of Clone Detection Models
von: Chen, QiHong, et al.
Veröffentlicht: (2025) -
Towards More Trustworthy Deep Code Models by Enabling Out-of-Distribution Detection
von: Yan, Yanfu, et al.
Veröffentlicht: (2025)