A3Rank: Augmentation Alignment Analysis for Prioritizing Overconfident Failing Samples for Deep Learning Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Zhengyuan, Wang, Haipeng, Zhou, Qilin, Chan, W. K. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Toward Patch Robustness Certification and Detection for Deep Learning Systems Beyond Consistent Samples
by: Zhou, Qilin, et al.
Published: (2025)
by: Zhou, Qilin, et al.
Published: (2025)
CrossCert: A Cross-Checking Detection Approach to Patch Robustness Certification for Deep Learning Models
by: Zhou, Qilin, et al.
Published: (2024)
by: Zhou, Qilin, et al.
Published: (2024)
Scalable and Precise Patch Robustness Certification for Deep Learning Models with Top-k Predictions
by: Zhou, Qilin, et al.
Published: (2025)
by: Zhou, Qilin, et al.
Published: (2025)
Context-Aware Fuzzing for Robustness Enhancement of Deep Learning Models
by: Wang, Haipeng, et al.
Published: (2024)
by: Wang, Haipeng, et al.
Published: (2024)
Automated Snippet-Alignment Data Augmentation for Code Translation
by: Zhang, Zhiming, et al.
Published: (2025)
by: Zhang, Zhiming, et al.
Published: (2025)
LLMs for Qualitative Data Analysis Fail on Security-specificComments in Human Experiments
by: Camporese, Maria, et al.
Published: (2026)
by: Camporese, Maria, et al.
Published: (2026)
Reasoning over Precedents Alongside Statutes: Case-Augmented Deliberative Alignment for LLM Safety
by: Jin, Can, et al.
Published: (2026)
by: Jin, Can, et al.
Published: (2026)
Where Do AI Coding Agents Fail? An Empirical Study of Failed Agentic Pull Requests in GitHub
by: Ehsani, Ramtin, et al.
Published: (2026)
by: Ehsani, Ramtin, et al.
Published: (2026)
Butterfly Effects in Toolchains: A Comprehensive Analysis of Failed Parameter Filling in LLM Tool-Agent Systems
by: Xiong, Qian, et al.
Published: (2025)
by: Xiong, Qian, et al.
Published: (2025)
Fuzzy Inference System for Test Case Prioritization in Software Testing
by: Karatayev, Aron, et al.
Published: (2024)
by: Karatayev, Aron, et al.
Published: (2024)
Automated Bug Report Prioritization in Large Open-Source Projects
by: Pierson, Riley, et al.
Published: (2025)
by: Pierson, Riley, et al.
Published: (2025)
AutoP2C: An LLM-Based Agent Framework for Code Repository Generation from Multimodal Content in Academic Papers
by: Lin, Zijie, et al.
Published: (2025)
by: Lin, Zijie, et al.
Published: (2025)
Faster Configuration Performance Bug Testing with Neural Dual-level Prioritization
by: Ma, Youpeng, et al.
Published: (2025)
by: Ma, Youpeng, et al.
Published: (2025)
WebSuite: Systematically Evaluating Why Web Agents Fail
by: Li, Eric, et al.
Published: (2024)
by: Li, Eric, et al.
Published: (2024)
GenCode: A Generic Data Augmentation Framework for Boosting Deep Learning-Based Code Understanding
by: Dong, Zeming, et al.
Published: (2024)
by: Dong, Zeming, et al.
Published: (2024)
How Do LLMs Fail In Agentic Scenarios? A Qualitative Analysis of Success and Failure Scenarios of Various LLMs in Agentic Simulations
by: Roig, JV
Published: (2025)
by: Roig, JV
Published: (2025)
Coherence Collapse: Diagnosing Why Code Agents Fail After Reaching the Right Code
by: Kim, Myeongsoo, et al.
Published: (2026)
by: Kim, Myeongsoo, et al.
Published: (2026)
CODE-DITING: A Reasoning-Based Metric for Functional Alignment in Code Evaluation
by: Yang, Guang, et al.
Published: (2025)
by: Yang, Guang, et al.
Published: (2025)
SweRank+: Multilingual, Multi-Turn Code Ranking for Software Issue Localization
by: Reddy, Revanth Gangi, et al.
Published: (2025)
by: Reddy, Revanth Gangi, et al.
Published: (2025)
Why Attention Fails: A Taxonomy of Faults in Attention-Based Neural Networks
by: Jahan, Sigma, et al.
Published: (2025)
by: Jahan, Sigma, et al.
Published: (2025)
Five Fatal Assumptions: Why T-Shirt Sizing Systematically Fails for AI Projects
by: Soundaramourty, Raja, et al.
Published: (2026)
by: Soundaramourty, Raja, et al.
Published: (2026)
Enhancing Cross-Language Code Translation via Task-Specific Embedding Alignment in Retrieval-Augmented Generation
by: Bhattarai, Manish, et al.
Published: (2024)
by: Bhattarai, Manish, et al.
Published: (2024)
An Empirical Investigation of Pre-Trained Deep Learning Model Reuse in the Scientific Process
by: Synovic, Nicholas M., et al.
Published: (2026)
by: Synovic, Nicholas M., et al.
Published: (2026)
Augmenting Large Language Models with Static Code Analysis for Automated Code Quality Improvements
by: Abtahi, Seyed Moein, et al.
Published: (2025)
by: Abtahi, Seyed Moein, et al.
Published: (2025)
Exploring Autonomous Agents: A Closer Look at Why They Fail When Completing Tasks
by: Lu, Ruofan, et al.
Published: (2025)
by: Lu, Ruofan, et al.
Published: (2025)
DeepSample: DNN sampling-based testing for operational accuracy assessment
by: Guerriero, Antonio, et al.
Published: (2024)
by: Guerriero, Antonio, et al.
Published: (2024)
Enhancing Code Translation in Language Models with Few-Shot Learning via Retrieval-Augmented Generation
by: Bhattarai, Manish, et al.
Published: (2024)
by: Bhattarai, Manish, et al.
Published: (2024)
A Contemporary Survey of Large Language Model Assisted Program Analysis
by: Wang, Jiayimei, et al.
Published: (2025)
by: Wang, Jiayimei, et al.
Published: (2025)
RAILS: Retrieval-Augmented Intelligence for Learning Software Development
by: Abdullah, Wali Mohammad, et al.
Published: (2025)
by: Abdullah, Wali Mohammad, et al.
Published: (2025)
Reducing Events to Augment Log-based Anomaly Detection Models: An Empirical Study
by: Zhang, Lingzhe, et al.
Published: (2024)
by: Zhang, Lingzhe, et al.
Published: (2024)
KADEL: Knowledge-Aware Denoising Learning for Commit Message Generation
by: Tao, Wei, et al.
Published: (2024)
by: Tao, Wei, et al.
Published: (2024)
Deep Learning Library Testing: Definition, Methods and Challenges
by: Zhang, Xiaoyu, et al.
Published: (2024)
by: Zhang, Xiaoyu, et al.
Published: (2024)
Deep Learning and Data Augmentation for Detecting Self-Admitted Technical Debt
by: Sutoyo, Edi, et al.
Published: (2024)
by: Sutoyo, Edi, et al.
Published: (2024)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
by: Giramata, Suavis, et al.
Published: (2025)
by: Giramata, Suavis, et al.
Published: (2025)
Keeping Deep Learning Models in Check: A History-Based Approach to Mitigate Overfitting
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Towards Fair Machine Learning Software: Understanding and Addressing Model Bias Through Counterfactual Thinking
by: Wang, Zichong, et al.
Published: (2023)
by: Wang, Zichong, et al.
Published: (2023)
Boosting Source Code Learning with Text-Oriented Data Augmentation: An Empirical Study
by: Dong, Zeming, et al.
Published: (2023)
by: Dong, Zeming, et al.
Published: (2023)
BabelCoder: Agentic Code Translation with Specification Alignment
by: Rabbi, Fazle, et al.
Published: (2025)
by: Rabbi, Fazle, et al.
Published: (2025)
Deep Learning-Based Identification of Inconsistent Method Names: How Far Are We?
by: Wang, Taiming, et al.
Published: (2025)
by: Wang, Taiming, et al.
Published: (2025)
On Security Weaknesses and Vulnerabilities in Deep Learning Systems
by: Lai, Zhongzheng, et al.
Published: (2024)
by: Lai, Zhongzheng, et al.
Published: (2024)
Similar Items
-
Toward Patch Robustness Certification and Detection for Deep Learning Systems Beyond Consistent Samples
by: Zhou, Qilin, et al.
Published: (2025) -
CrossCert: A Cross-Checking Detection Approach to Patch Robustness Certification for Deep Learning Models
by: Zhou, Qilin, et al.
Published: (2024) -
Scalable and Precise Patch Robustness Certification for Deep Learning Models with Top-k Predictions
by: Zhou, Qilin, et al.
Published: (2025) -
Context-Aware Fuzzing for Robustness Enhancement of Deep Learning Models
by: Wang, Haipeng, et al.
Published: (2024) -
Automated Snippet-Alignment Data Augmentation for Code Translation
by: Zhang, Zhiming, et al.
Published: (2025)