Human-aligned AI Model Cards with Weighted Hierarchy Architecture
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Pengyue, Jin, Haolin, Zeng, Qingwen, Wen, Jiawen, Rao, Harry, Chen, Huaming |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Uncovering Systematic Failures of LLMs in Verifying Code Against Natural Language Specifications
by: Jin, Haolin, et al.
Published: (2025)
by: Jin, Haolin, et al.
Published: (2025)
Are LLMs Reliable Code Reviewers? Systematic Overcorrection in Requirement Conformance Judgement
by: Jin, Haolin, et al.
Published: (2026)
by: Jin, Haolin, et al.
Published: (2026)
Towards Advancing Code Generation with Large Language Models: A Research Roadmap
by: Jin, Haolin, et al.
Published: (2025)
by: Jin, Haolin, et al.
Published: (2025)
RGD: Multi-LLM Based Agent Debugger via Refinement and Generation Guidance
by: Jin, Haolin, et al.
Published: (2024)
by: Jin, Haolin, et al.
Published: (2024)
What You See Is Not Always What You Get: Evaluating GPT's Comprehension of Source Code
by: Wen, Jiawen, et al.
Published: (2024)
by: Wen, Jiawen, et al.
Published: (2024)
Fairpriori: Improving Biased Subgroup Discovery for Deep Neural Network Fairness
by: Zhou, Kacy, et al.
Published: (2024)
by: Zhou, Kacy, et al.
Published: (2024)
From LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and Future
by: Jin, Haolin, et al.
Published: (2024)
by: Jin, Haolin, et al.
Published: (2024)
Model Cards Revisited: Bridging the Gap Between Theory and Practice for Ethical AI Requirements
by: Puhlfürß, Tim, et al.
Published: (2025)
by: Puhlfürß, Tim, et al.
Published: (2025)
What's documented in AI? Systematic Analysis of 32K AI Model Cards
by: Liang, Weixin, et al.
Published: (2024)
by: Liang, Weixin, et al.
Published: (2024)
LLMs are All You Need? Improving Fuzz Testing for MOJO with Large Language Models
by: Huang, Linghan, et al.
Published: (2025)
by: Huang, Linghan, et al.
Published: (2025)
On the Challenges of Fuzzing Techniques via Large Language Models
by: Huang, Linghan, et al.
Published: (2024)
by: Huang, Linghan, et al.
Published: (2024)
DesignCoder: Hierarchy-Aware and Self-Correcting UI Code Generation with Large Language Models
by: Chen, Yunnong, et al.
Published: (2025)
by: Chen, Yunnong, et al.
Published: (2025)
AI Transparency Atlas: Framework, Scoring, and Real-Time Model Card Evaluation Pipeline
by: Mamirov, Akhmadillo, et al.
Published: (2025)
by: Mamirov, Akhmadillo, et al.
Published: (2025)
A Grounded Theory of Debugging in Professional Software Engineering Practice
by: Li, Haolin, et al.
Published: (2026)
by: Li, Haolin, et al.
Published: (2026)
SWE-TRACE: Optimizing Long-Horizon SWE Agents Through Rubric Process Reward Models and Heuristic Test-Time Scaling
by: Han, Hao, et al.
Published: (2026)
by: Han, Hao, et al.
Published: (2026)
Formal Architecture Descriptors as Navigation Primitives for AI Coding Agents
by: Jin, Ruoqi
Published: (2026)
by: Jin, Ruoqi
Published: (2026)
Agentic Model Checking
by: Sun, Youcheng, et al.
Published: (2026)
by: Sun, Youcheng, et al.
Published: (2026)
Quantum Algorithm Cards: Streamlining the development of hybrid classical-quantum applications
by: Stirbu, Vlad, et al.
Published: (2023)
by: Stirbu, Vlad, et al.
Published: (2023)
Adversarial Feature Map Pruning for Backdoor
by: Huang, Dong, et al.
Published: (2023)
by: Huang, Dong, et al.
Published: (2023)
Supporting Software Maintenance with Dynamically Generated Document Hierarchies
by: Dearstyne, Katherine R., et al.
Published: (2024)
by: Dearstyne, Katherine R., et al.
Published: (2024)
Agile System Development Lifecycle for AI Systems: Decision Architecture
by: Gill, Asif Q.
Published: (2025)
by: Gill, Asif Q.
Published: (2025)
ArchBench: Benchmarking Generative-AI for Software Architecture Tasks
by: Adnan, Bassam, et al.
Published: (2026)
by: Adnan, Bassam, et al.
Published: (2026)
Synergy-Guided Compiler Auto-Tuning of Nested LLVM Pass Pipelines
by: Pan, Haolin, et al.
Published: (2025)
by: Pan, Haolin, et al.
Published: (2025)
A Hybrid, Knowledge-Guided Evolutionary Framework for Personalized Compiler Auto-Tuning
by: Pan, Haolin, et al.
Published: (2025)
by: Pan, Haolin, et al.
Published: (2025)
Human to Document, AI to Code: Comparing GenAI for Notebook Competitions
by: Settewong, Tasha, et al.
Published: (2025)
by: Settewong, Tasha, et al.
Published: (2025)
Understanding Open Source Contributor Profiles in Popular Machine Learning Libraries
by: Liu, Jiawen, et al.
Published: (2024)
by: Liu, Jiawen, et al.
Published: (2024)
Human-AI Synergy in Agentic Code Review
by: Zhong, Suzhen, et al.
Published: (2026)
by: Zhong, Suzhen, et al.
Published: (2026)
A Systematic Mapping Study on Software Architecture for AI-based Mobility Systems
by: Ramic, Amra, et al.
Published: (2025)
by: Ramic, Amra, et al.
Published: (2025)
Mining Architectural Information: A Systematic Mapping Study
by: de Dieu, Musengamana Jean, et al.
Published: (2022)
by: de Dieu, Musengamana Jean, et al.
Published: (2022)
AI Policy, Disclosure, and Human in the Loop: How Are Contribution Guidelines Adapting to GenAI?
by: Hora, Andre, et al.
Published: (2026)
by: Hora, Andre, et al.
Published: (2026)
From Prompt-Response to Goal-Directed Systems: The Evolution of Agentic AI Software Architecture
by: Alenezi, Mamdouh
Published: (2026)
by: Alenezi, Mamdouh
Published: (2026)
Agentic AI in the Software Development Lifecycle: Architecture, Empirical Evidence, and the Reshaping of Software Engineering
by: Bhati, Happy
Published: (2026)
by: Bhati, Happy
Published: (2026)
GRACE: Globally-Seeded Representation-Aware Cluster-Specific Evolution for Compiler Auto-Tuning
by: Pan, Haolin, et al.
Published: (2025)
by: Pan, Haolin, et al.
Published: (2025)
Enhancing COBOL Code Explanations: A Multi-Agents Approach Using Large Language Models
by: Lei, Fangjian, et al.
Published: (2025)
by: Lei, Fangjian, et al.
Published: (2025)
Think Like Human Developers: Harnessing Community Knowledge for Structured Code Reasoning
by: Yang, Chengran, et al.
Published: (2025)
by: Yang, Chengran, et al.
Published: (2025)
A11YN: aligning LLMs for accessible web UI code generation
by: Yoon, Janghan, et al.
Published: (2025)
by: Yoon, Janghan, et al.
Published: (2025)
Unveiling Assumptions: Exploring the Decisions of AI Chatbots and Human Testers
by: Neto, Francisco Gomes de Oliveira
Published: (2024)
by: Neto, Francisco Gomes de Oliveira
Published: (2024)
A Reference Architecture for Gamified Cultural Heritage Applications Leveraging Generative AI and Augmented Reality
by: Martusciello, Federico, et al.
Published: (2025)
by: Martusciello, Federico, et al.
Published: (2025)
The Impact of AI-Generated Solutions on Software Architecture and Productivity: Results from a Survey Study
by: Amasanti, Giorgio, et al.
Published: (2025)
by: Amasanti, Giorgio, et al.
Published: (2025)
Reducing Labeling Effort in Architecture Technical Debt Detection through Active Learning and Explainable AI
by: Sutoyo, Edi, et al.
Published: (2026)
by: Sutoyo, Edi, et al.
Published: (2026)
Similar Items
-
Uncovering Systematic Failures of LLMs in Verifying Code Against Natural Language Specifications
by: Jin, Haolin, et al.
Published: (2025) -
Are LLMs Reliable Code Reviewers? Systematic Overcorrection in Requirement Conformance Judgement
by: Jin, Haolin, et al.
Published: (2026) -
Towards Advancing Code Generation with Large Language Models: A Research Roadmap
by: Jin, Haolin, et al.
Published: (2025) -
RGD: Multi-LLM Based Agent Debugger via Refinement and Generation Guidance
by: Jin, Haolin, et al.
Published: (2024) -
What You See Is Not Always What You Get: Evaluating GPT's Comprehension of Source Code
by: Wen, Jiawen, et al.
Published: (2024)