Human-aligned AI Model Cards with Weighted Hierarchy Architecture
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Pengyue, Jin, Haolin, Zeng, Qingwen, Wen, Jiawen, Rao, Harry, Chen, Huaming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Uncovering Systematic Failures of LLMs in Verifying Code Against Natural Language Specifications
von: Jin, Haolin, et al.
Veröffentlicht: (2025)
von: Jin, Haolin, et al.
Veröffentlicht: (2025)
Are LLMs Reliable Code Reviewers? Systematic Overcorrection in Requirement Conformance Judgement
von: Jin, Haolin, et al.
Veröffentlicht: (2026)
von: Jin, Haolin, et al.
Veröffentlicht: (2026)
Towards Advancing Code Generation with Large Language Models: A Research Roadmap
von: Jin, Haolin, et al.
Veröffentlicht: (2025)
von: Jin, Haolin, et al.
Veröffentlicht: (2025)
RGD: Multi-LLM Based Agent Debugger via Refinement and Generation Guidance
von: Jin, Haolin, et al.
Veröffentlicht: (2024)
von: Jin, Haolin, et al.
Veröffentlicht: (2024)
What You See Is Not Always What You Get: Evaluating GPT's Comprehension of Source Code
von: Wen, Jiawen, et al.
Veröffentlicht: (2024)
von: Wen, Jiawen, et al.
Veröffentlicht: (2024)
Fairpriori: Improving Biased Subgroup Discovery for Deep Neural Network Fairness
von: Zhou, Kacy, et al.
Veröffentlicht: (2024)
von: Zhou, Kacy, et al.
Veröffentlicht: (2024)
From LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and Future
von: Jin, Haolin, et al.
Veröffentlicht: (2024)
von: Jin, Haolin, et al.
Veröffentlicht: (2024)
Model Cards Revisited: Bridging the Gap Between Theory and Practice for Ethical AI Requirements
von: Puhlfürß, Tim, et al.
Veröffentlicht: (2025)
von: Puhlfürß, Tim, et al.
Veröffentlicht: (2025)
What's documented in AI? Systematic Analysis of 32K AI Model Cards
von: Liang, Weixin, et al.
Veröffentlicht: (2024)
von: Liang, Weixin, et al.
Veröffentlicht: (2024)
LLMs are All You Need? Improving Fuzz Testing for MOJO with Large Language Models
von: Huang, Linghan, et al.
Veröffentlicht: (2025)
von: Huang, Linghan, et al.
Veröffentlicht: (2025)
On the Challenges of Fuzzing Techniques via Large Language Models
von: Huang, Linghan, et al.
Veröffentlicht: (2024)
von: Huang, Linghan, et al.
Veröffentlicht: (2024)
DesignCoder: Hierarchy-Aware and Self-Correcting UI Code Generation with Large Language Models
von: Chen, Yunnong, et al.
Veröffentlicht: (2025)
von: Chen, Yunnong, et al.
Veröffentlicht: (2025)
AI Transparency Atlas: Framework, Scoring, and Real-Time Model Card Evaluation Pipeline
von: Mamirov, Akhmadillo, et al.
Veröffentlicht: (2025)
von: Mamirov, Akhmadillo, et al.
Veröffentlicht: (2025)
A Grounded Theory of Debugging in Professional Software Engineering Practice
von: Li, Haolin, et al.
Veröffentlicht: (2026)
von: Li, Haolin, et al.
Veröffentlicht: (2026)
SWE-TRACE: Optimizing Long-Horizon SWE Agents Through Rubric Process Reward Models and Heuristic Test-Time Scaling
von: Han, Hao, et al.
Veröffentlicht: (2026)
von: Han, Hao, et al.
Veröffentlicht: (2026)
Formal Architecture Descriptors as Navigation Primitives for AI Coding Agents
von: Jin, Ruoqi
Veröffentlicht: (2026)
von: Jin, Ruoqi
Veröffentlicht: (2026)
Agentic Model Checking
von: Sun, Youcheng, et al.
Veröffentlicht: (2026)
von: Sun, Youcheng, et al.
Veröffentlicht: (2026)
Quantum Algorithm Cards: Streamlining the development of hybrid classical-quantum applications
von: Stirbu, Vlad, et al.
Veröffentlicht: (2023)
von: Stirbu, Vlad, et al.
Veröffentlicht: (2023)
Adversarial Feature Map Pruning for Backdoor
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
Supporting Software Maintenance with Dynamically Generated Document Hierarchies
von: Dearstyne, Katherine R., et al.
Veröffentlicht: (2024)
von: Dearstyne, Katherine R., et al.
Veröffentlicht: (2024)
Agile System Development Lifecycle for AI Systems: Decision Architecture
von: Gill, Asif Q.
Veröffentlicht: (2025)
von: Gill, Asif Q.
Veröffentlicht: (2025)
ArchBench: Benchmarking Generative-AI for Software Architecture Tasks
von: Adnan, Bassam, et al.
Veröffentlicht: (2026)
von: Adnan, Bassam, et al.
Veröffentlicht: (2026)
Synergy-Guided Compiler Auto-Tuning of Nested LLVM Pass Pipelines
von: Pan, Haolin, et al.
Veröffentlicht: (2025)
von: Pan, Haolin, et al.
Veröffentlicht: (2025)
A Hybrid, Knowledge-Guided Evolutionary Framework for Personalized Compiler Auto-Tuning
von: Pan, Haolin, et al.
Veröffentlicht: (2025)
von: Pan, Haolin, et al.
Veröffentlicht: (2025)
Human to Document, AI to Code: Comparing GenAI for Notebook Competitions
von: Settewong, Tasha, et al.
Veröffentlicht: (2025)
von: Settewong, Tasha, et al.
Veröffentlicht: (2025)
Understanding Open Source Contributor Profiles in Popular Machine Learning Libraries
von: Liu, Jiawen, et al.
Veröffentlicht: (2024)
von: Liu, Jiawen, et al.
Veröffentlicht: (2024)
Human-AI Synergy in Agentic Code Review
von: Zhong, Suzhen, et al.
Veröffentlicht: (2026)
von: Zhong, Suzhen, et al.
Veröffentlicht: (2026)
A Systematic Mapping Study on Software Architecture for AI-based Mobility Systems
von: Ramic, Amra, et al.
Veröffentlicht: (2025)
von: Ramic, Amra, et al.
Veröffentlicht: (2025)
Mining Architectural Information: A Systematic Mapping Study
von: de Dieu, Musengamana Jean, et al.
Veröffentlicht: (2022)
von: de Dieu, Musengamana Jean, et al.
Veröffentlicht: (2022)
AI Policy, Disclosure, and Human in the Loop: How Are Contribution Guidelines Adapting to GenAI?
von: Hora, Andre, et al.
Veröffentlicht: (2026)
von: Hora, Andre, et al.
Veröffentlicht: (2026)
From Prompt-Response to Goal-Directed Systems: The Evolution of Agentic AI Software Architecture
von: Alenezi, Mamdouh
Veröffentlicht: (2026)
von: Alenezi, Mamdouh
Veröffentlicht: (2026)
Agentic AI in the Software Development Lifecycle: Architecture, Empirical Evidence, and the Reshaping of Software Engineering
von: Bhati, Happy
Veröffentlicht: (2026)
von: Bhati, Happy
Veröffentlicht: (2026)
GRACE: Globally-Seeded Representation-Aware Cluster-Specific Evolution for Compiler Auto-Tuning
von: Pan, Haolin, et al.
Veröffentlicht: (2025)
von: Pan, Haolin, et al.
Veröffentlicht: (2025)
Enhancing COBOL Code Explanations: A Multi-Agents Approach Using Large Language Models
von: Lei, Fangjian, et al.
Veröffentlicht: (2025)
von: Lei, Fangjian, et al.
Veröffentlicht: (2025)
Think Like Human Developers: Harnessing Community Knowledge for Structured Code Reasoning
von: Yang, Chengran, et al.
Veröffentlicht: (2025)
von: Yang, Chengran, et al.
Veröffentlicht: (2025)
A11YN: aligning LLMs for accessible web UI code generation
von: Yoon, Janghan, et al.
Veröffentlicht: (2025)
von: Yoon, Janghan, et al.
Veröffentlicht: (2025)
Unveiling Assumptions: Exploring the Decisions of AI Chatbots and Human Testers
von: Neto, Francisco Gomes de Oliveira
Veröffentlicht: (2024)
von: Neto, Francisco Gomes de Oliveira
Veröffentlicht: (2024)
A Reference Architecture for Gamified Cultural Heritage Applications Leveraging Generative AI and Augmented Reality
von: Martusciello, Federico, et al.
Veröffentlicht: (2025)
von: Martusciello, Federico, et al.
Veröffentlicht: (2025)
The Impact of AI-Generated Solutions on Software Architecture and Productivity: Results from a Survey Study
von: Amasanti, Giorgio, et al.
Veröffentlicht: (2025)
von: Amasanti, Giorgio, et al.
Veröffentlicht: (2025)
Reducing Labeling Effort in Architecture Technical Debt Detection through Active Learning and Explainable AI
von: Sutoyo, Edi, et al.
Veröffentlicht: (2026)
von: Sutoyo, Edi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Uncovering Systematic Failures of LLMs in Verifying Code Against Natural Language Specifications
von: Jin, Haolin, et al.
Veröffentlicht: (2025) -
Are LLMs Reliable Code Reviewers? Systematic Overcorrection in Requirement Conformance Judgement
von: Jin, Haolin, et al.
Veröffentlicht: (2026) -
Towards Advancing Code Generation with Large Language Models: A Research Roadmap
von: Jin, Haolin, et al.
Veröffentlicht: (2025) -
RGD: Multi-LLM Based Agent Debugger via Refinement and Generation Guidance
von: Jin, Haolin, et al.
Veröffentlicht: (2024) -
What You See Is Not Always What You Get: Evaluating GPT's Comprehension of Source Code
von: Wen, Jiawen, et al.
Veröffentlicht: (2024)