Can AI Models Direct Each Other? Organizational Structure as a Probe into Training Limitations
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Liu, Rui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Investigating Training Data Detection in AI Coders
von: Li, Tianlin, et al.
Veröffentlicht: (2025)
von: Li, Tianlin, et al.
Veröffentlicht: (2025)
A Framework for the Adoption and Integration of Generative AI in Midsize Organizations and Enterprises (FAIGMOE)
von: Weinberg, Abraham Itzhak
Veröffentlicht: (2025)
von: Weinberg, Abraham Itzhak
Veröffentlicht: (2025)
Let the Barbarians In: How AI Can Accelerate Systems Performance Research
von: Cheng, Audrey, et al.
Veröffentlicht: (2025)
von: Cheng, Audrey, et al.
Veröffentlicht: (2025)
ProgramBench: Can Language Models Rebuild Programs From Scratch?
von: Yang, John, et al.
Veröffentlicht: (2026)
von: Yang, John, et al.
Veröffentlicht: (2026)
Structure-Aware Corpus Construction and User-Perception-Aligned Metrics for Large-Language-Model Code Completion
von: Liu, Dengfeng, et al.
Veröffentlicht: (2025)
von: Liu, Dengfeng, et al.
Veröffentlicht: (2025)
Unveiling Code Pre-Trained Models: Investigating Syntax and Semantics Capacities
von: Ma, Wei, et al.
Veröffentlicht: (2022)
von: Ma, Wei, et al.
Veröffentlicht: (2022)
Repository Structure-Aware Training Makes SLMs Better Issue Resolver
von: Ma, Zexiong, et al.
Veröffentlicht: (2024)
von: Ma, Zexiong, et al.
Veröffentlicht: (2024)
GitHub's Copilot Code Review: Can AI Spot Security Flaws Before You Commit?
von: Amro, Amena, et al.
Veröffentlicht: (2025)
von: Amro, Amena, et al.
Veröffentlicht: (2025)
Verification Limits Code LLM Training
von: Gureja, Srishti, et al.
Veröffentlicht: (2025)
von: Gureja, Srishti, et al.
Veröffentlicht: (2025)
Can Github issues be solved with Tree Of Thoughts?
von: La Rosa, Ricardo, et al.
Veröffentlicht: (2024)
von: La Rosa, Ricardo, et al.
Veröffentlicht: (2024)
Can Agents Fix Agent Issues?
von: Rahardja, Alfin Wijaya, et al.
Veröffentlicht: (2025)
von: Rahardja, Alfin Wijaya, et al.
Veröffentlicht: (2025)
CodeFort: Robust Training for Code Generation Models
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
Attention Distance: A Novel Metric for Directed Fuzzing with Large Language Models
von: Bin, Wang, et al.
Veröffentlicht: (2025)
von: Bin, Wang, et al.
Veröffentlicht: (2025)
AI-Driven Tools in Modern Software Quality Assurance: An Assessment of Benefits, Challenges, and Future Directions
von: Pysmennyi, Ihor, et al.
Veröffentlicht: (2025)
von: Pysmennyi, Ihor, et al.
Veröffentlicht: (2025)
Accuracy Can Lie: On the Impact of Surrogate Model in Configuration Tuning
von: Chen, Pengzhou, et al.
Veröffentlicht: (2025)
von: Chen, Pengzhou, et al.
Veröffentlicht: (2025)
Breaking the Myth: Can Small Models Infer Postconditions Too?
von: Zhang, Gehao, et al.
Veröffentlicht: (2025)
von: Zhang, Gehao, et al.
Veröffentlicht: (2025)
Can LLM Generate Regression Tests for Software Commits?
von: Liu, Jing, et al.
Veröffentlicht: (2025)
von: Liu, Jing, et al.
Veröffentlicht: (2025)
Large Language Models for In-File Vulnerability Localization Can Be "Lost in the End"
von: Sovrano, Francesco, et al.
Veröffentlicht: (2025)
von: Sovrano, Francesco, et al.
Veröffentlicht: (2025)
PostTrainBench: Can LLM Agents Automate LLM Post-Training?
von: Rank, Ben, et al.
Veröffentlicht: (2026)
von: Rank, Ben, et al.
Veröffentlicht: (2026)
Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment
von: Yu, Shasha, et al.
Veröffentlicht: (2026)
von: Yu, Shasha, et al.
Veröffentlicht: (2026)
LLM Company Policies and Policy Implications in Software Organizations
von: Khojah, Ranim, et al.
Veröffentlicht: (2025)
von: Khojah, Ranim, et al.
Veröffentlicht: (2025)
Structural Enforcement of Statistical Rigor in AI-Driven Discovery: A Functional Architecture
von: Sargsyan, Karen
Veröffentlicht: (2025)
von: Sargsyan, Karen
Veröffentlicht: (2025)
VeriStruct: AI-assisted Automated Verification of Data-Structure Modules in Verus
von: Sun, Chuyue, et al.
Veröffentlicht: (2025)
von: Sun, Chuyue, et al.
Veröffentlicht: (2025)
How Mature is Requirements Engineering for AI-based Systems? A Systematic Mapping Study on Practices, Challenges, and Future Research Directions
von: Habiba, Umm-e-, et al.
Veröffentlicht: (2024)
von: Habiba, Umm-e-, et al.
Veröffentlicht: (2024)
Generative Language Models Potential for Requirement Engineering Applications: Insights into Current Strengths and Limitations
von: Saleem, Summra, et al.
Veröffentlicht: (2024)
von: Saleem, Summra, et al.
Veröffentlicht: (2024)
DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs
von: Liu, Zhihan, et al.
Veröffentlicht: (2024)
von: Liu, Zhihan, et al.
Veröffentlicht: (2024)
Digital Twins & ZeroConf AI: Structuring Automated Intelligent Pipelines for Industrial Applications
von: Picone, Marco, et al.
Veröffentlicht: (2026)
von: Picone, Marco, et al.
Veröffentlicht: (2026)
Querying Large Automotive Software Models: Agentic vs. Direct LLM Approaches
von: Mazur, Lukasz, et al.
Veröffentlicht: (2025)
von: Mazur, Lukasz, et al.
Veröffentlicht: (2025)
I Can Find You in Seconds! Leveraging Large Language Models for Code Authorship Attribution
von: Choi, Soohyeon, et al.
Veröffentlicht: (2025)
von: Choi, Soohyeon, et al.
Veröffentlicht: (2025)
GenAI for Simulation Model in Model-Based Systems Engineering
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
Your Code Agent Can Grow Alongside You with Structured Memory
von: Deng, Yi-Xuan, et al.
Veröffentlicht: (2026)
von: Deng, Yi-Xuan, et al.
Veröffentlicht: (2026)
Automating Structural Analysis Across Multiple Software Platforms Using Large Language Models
von: Geng, Ziheng, et al.
Veröffentlicht: (2026)
von: Geng, Ziheng, et al.
Veröffentlicht: (2026)
Decentralised Governance-Driven Architecture for Designing Foundation Model based Systems: Exploring the Role of Blockchain in Responsible AI
von: Liu, Yue, et al.
Veröffentlicht: (2023)
von: Liu, Yue, et al.
Veröffentlicht: (2023)
An Empirical Investigation of Pre-Trained Deep Learning Model Reuse in the Scientific Process
von: Synovic, Nicholas M., et al.
Veröffentlicht: (2026)
von: Synovic, Nicholas M., et al.
Veröffentlicht: (2026)
Learning to Debug: LLM-Organized Knowledge Trees for Solving RTL Assertion Failures
von: Bai, Yunsheng, et al.
Veröffentlicht: (2025)
von: Bai, Yunsheng, et al.
Veröffentlicht: (2025)
Less is More: Towards Green Code Large Language Models via Unified Structural Pruning
von: Yang, Guang, et al.
Veröffentlicht: (2024)
von: Yang, Guang, et al.
Veröffentlicht: (2024)
RepoMirage: Probing Repository Context Reasoning in Code Agents with Perturbations
von: Li, Hanyu, et al.
Veröffentlicht: (2026)
von: Li, Hanyu, et al.
Veröffentlicht: (2026)
Contractual Skills: A GovernSpec Design Framework for Enterprise AI Agents
von: Liu, Ting
Veröffentlicht: (2026)
von: Liu, Ting
Veröffentlicht: (2026)
Can an LLM Detect Instances of Microservice Infrastructure Patterns?
von: Duarte, Carlos Eduardo, et al.
Veröffentlicht: (2026)
von: Duarte, Carlos Eduardo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Investigating Training Data Detection in AI Coders
von: Li, Tianlin, et al.
Veröffentlicht: (2025) -
A Framework for the Adoption and Integration of Generative AI in Midsize Organizations and Enterprises (FAIGMOE)
von: Weinberg, Abraham Itzhak
Veröffentlicht: (2025) -
Let the Barbarians In: How AI Can Accelerate Systems Performance Research
von: Cheng, Audrey, et al.
Veröffentlicht: (2025) -
ProgramBench: Can Language Models Rebuild Programs From Scratch?
von: Yang, John, et al.
Veröffentlicht: (2026) -
Structure-Aware Corpus Construction and User-Perception-Aligned Metrics for Large-Language-Model Code Completion
von: Liu, Dengfeng, et al.
Veröffentlicht: (2025)