Beyond Quantity: Trajectory Diversity Scaling for Code Agents
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Guhong, Sun, Chenghao, Fu, Cheng, Wang, Qiyao, Huang, Zhihong, Wei, Chaopeng, Chen, Guangxu, Fang, Feiteng, Argha, Ahmadreza, Zhao, Bing, Xu, Xander, Han, Qi, Alinejad-Rokny, Hamid, Qu, Qiang, Li, Binhua, Ni, Shiwen, Yang, Min, Wei, Hu, Li, Yongbin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Lower Layers Matter: Alleviating Hallucination via Multi-Layer Fusion Contrastive Decoding with Truthfulness Refocused
di: Chen, Dingwei, et al.
Pubblicazione: (2024)
di: Chen, Dingwei, et al.
Pubblicazione: (2024)
Expanding before Inferring: Enhancing Factuality in Large Language Models through Premature Layers Interpolation
di: Chen, Dingwei, et al.
Pubblicazione: (2025)
di: Chen, Dingwei, et al.
Pubblicazione: (2025)
AutoPatent: A Multi-Agent Framework for Automatic Patent Generation
di: Wang, Qiyao, et al.
Pubblicazione: (2024)
di: Wang, Qiyao, et al.
Pubblicazione: (2024)
xJailbreak: Representation Space Guided Reinforcement Learning for Interpretable LLM Jailbreaking
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
ETAGE: Enhanced Test Time Adaptation with Integrated Entropy and Gradient Norms for Robust Model Performance
di: Shamsi, Afshar, et al.
Pubblicazione: (2024)
di: Shamsi, Afshar, et al.
Pubblicazione: (2024)
Interpretable graph-based models on multimodal biomedical data integration: A technical review and benchmarking
di: Sadeghi, Alireza, et al.
Pubblicazione: (2025)
di: Sadeghi, Alireza, et al.
Pubblicazione: (2025)
FlowPIE: Test-Time Scientific Idea Evolution with Flow-Guided Literature Exploration
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
AgentCourt: Simulating Court with Adversarial Evolvable Lawyer Agents
di: Chen, Guhong, et al.
Pubblicazione: (2024)
di: Chen, Guhong, et al.
Pubblicazione: (2024)
RxSafeBench: Identifying Medication Safety Issues of Large Language Models in Simulated Consultation
di: Zhao, Jiahao, et al.
Pubblicazione: (2025)
di: Zhao, Jiahao, et al.
Pubblicazione: (2025)
CLinNET: An Interpretable and Uncertainty‐Aware Deep Learning Framework for Multi‐Modal Clinical Genomics
di: Ivan Bakhshayeshi, et al.
Pubblicazione: (2026)
di: Ivan Bakhshayeshi, et al.
Pubblicazione: (2026)
Small Language Model as Data Prospector for Large Language Model
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
SemanticST: Spatially Informed Semantic Graph Learning for Clustering, Integration, and Scalable Analysis of Spatial Transcriptomics
di: Zahedi, Roxana, et al.
Pubblicazione: (2025)
di: Zahedi, Roxana, et al.
Pubblicazione: (2025)
Automatic Paper Reviewing with Heterogeneous Graph Reasoning over LLM-Simulated Reviewer-Author Debates
di: Li, Shuaimin, et al.
Pubblicazione: (2025)
di: Li, Shuaimin, et al.
Pubblicazione: (2025)
PersonaMath: Boosting Mathematical Reasoning via Persona-Driven Data Augmentation
di: Luo, Jing, et al.
Pubblicazione: (2024)
di: Luo, Jing, et al.
Pubblicazione: (2024)
STORYTELLER: An Enhanced Plot-Planning Framework for Coherent and Cohesive Story Generation
di: Li, Jiaming, et al.
Pubblicazione: (2025)
di: Li, Jiaming, et al.
Pubblicazione: (2025)
PatRe: A Full-Stage Office Action and Rebuttal Generation Benchmark for Patent Examination
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
di: Wang, Qiyao, et al.
Pubblicazione: (2026)
Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability
di: Liang, Yuheng, et al.
Pubblicazione: (2026)
di: Liang, Yuheng, et al.
Pubblicazione: (2026)
How chromatin interactions shed light on interpreting non-coding genomic variants: opportunities and future direc-tions
di: Liang, Yuheng, et al.
Pubblicazione: (2024)
di: Liang, Yuheng, et al.
Pubblicazione: (2024)
Structuring Reasoning for Complex Rules Beyond Flat Representations
di: Yang, Zhihao, et al.
Pubblicazione: (2025)
di: Yang, Zhihao, et al.
Pubblicazione: (2025)
CollectiveSFT: Scaling Large Language Models for Chinese Medical Benchmark with Collective Instructions in Healthcare
di: Zhu, Jingwei, et al.
Pubblicazione: (2024)
di: Zhu, Jingwei, et al.
Pubblicazione: (2024)
Act-Adaptive Margin: Dynamically Calibrating Reward Models for Subjective Ambiguity
di: Fang, Feiteng, et al.
Pubblicazione: (2025)
di: Fang, Feiteng, et al.
Pubblicazione: (2025)
IPBench: Benchmarking the Knowledge of Large Language Models in Intellectual Property
di: Wang, Qiyao, et al.
Pubblicazione: (2025)
di: Wang, Qiyao, et al.
Pubblicazione: (2025)
PLOT: Enhancing Preference Learning via Optimal Transport
di: Zhu, Liang, et al.
Pubblicazione: (2026)
di: Zhu, Liang, et al.
Pubblicazione: (2026)
To Diff or Not to Diff? Structure-Aware and Adaptive Output Formats for Efficient LLM-based Code Editing
di: Cheng, Wei, et al.
Pubblicazione: (2026)
di: Cheng, Wei, et al.
Pubblicazione: (2026)
Enhancing Monte Carlo Dropout Performance for Uncertainty Quantification
di: Asgharnezhad, Hamzeh, et al.
Pubblicazione: (2025)
di: Asgharnezhad, Hamzeh, et al.
Pubblicazione: (2025)
CoTJudger: A Graph-Driven Framework for Automatic Evaluation of Chain-of-Thought Efficiency and Redundancy in LRMs
di: Li, Siyi, et al.
Pubblicazione: (2026)
di: Li, Siyi, et al.
Pubblicazione: (2026)
A Survey on Large Language Model Benchmarks
di: Ni, Shiwen, et al.
Pubblicazione: (2025)
di: Ni, Shiwen, et al.
Pubblicazione: (2025)
Enhancing Noise Robustness of Retrieval-Augmented Language Models with Adaptive Adversarial Training
di: Fang, Feiteng, et al.
Pubblicazione: (2024)
di: Fang, Feiteng, et al.
Pubblicazione: (2024)
Training Superior Sparse Autoencoders for Instruct Models
di: Li, Jiaming, et al.
Pubblicazione: (2025)
di: Li, Jiaming, et al.
Pubblicazione: (2025)
HiST: Histological Images Reconstruct Tumor Spatial Transcriptomics via MultiScale Fusion Deep Learning
di: Wei Li, et al.
Pubblicazione: (2026)
di: Wei Li, et al.
Pubblicazione: (2026)
Advancing Medical Image Segmentation with Mini-Net: A Lightweight Solution Tailored for Efficient Segmentation of Medical Images
di: Javed, Syed, et al.
Pubblicazione: (2024)
di: Javed, Syed, et al.
Pubblicazione: (2024)
OpenOmni: Advancing Open-Source Omnimodal Large Language Models with Progressive Multimodal Alignment and Real-Time Self-Aware Emotional Speech Synthesis
di: Luo, Run, et al.
Pubblicazione: (2025)
di: Luo, Run, et al.
Pubblicazione: (2025)
Quantification of Large Language Model Distillation
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
RuCL: Stratified Rubric-Based Curriculum Learning for Multimodal Large Language Model Reasoning
di: Chen, Yukun, et al.
Pubblicazione: (2026)
di: Chen, Yukun, et al.
Pubblicazione: (2026)
A Diagnostic Model for Acute Lymphoblastic Leukemia Using Metaheuristics and Deep Learning Methods
di: Rahmani, Amir Masoud, et al.
Pubblicazione: (2024)
di: Rahmani, Amir Masoud, et al.
Pubblicazione: (2024)
Focal Modulation and Bidirectional Feature Fusion Network for Medical Image Segmentation
di: Safdar, Moin, et al.
Pubblicazione: (2025)
di: Safdar, Moin, et al.
Pubblicazione: (2025)
Quantity leadership in dual‐channel supply chains with production economies or diseconomies
di: Jinru Feng, et al.
Pubblicazione: (2026)
di: Jinru Feng, et al.
Pubblicazione: (2026)
SWE-CI: Evaluating Agent Capabilities in Maintaining Codebases via Continuous Integration
di: Chen, Jialong, et al.
Pubblicazione: (2026)
di: Chen, Jialong, et al.
Pubblicazione: (2026)
CLaSp: In-Context Layer Skip for Self-Speculative Decoding
di: Chen, Longze, et al.
Pubblicazione: (2025)
di: Chen, Longze, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Lower Layers Matter: Alleviating Hallucination via Multi-Layer Fusion Contrastive Decoding with Truthfulness Refocused
di: Chen, Dingwei, et al.
Pubblicazione: (2024) -
Expanding before Inferring: Enhancing Factuality in Large Language Models through Premature Layers Interpolation
di: Chen, Dingwei, et al.
Pubblicazione: (2025) -
AutoPatent: A Multi-Agent Framework for Automatic Patent Generation
di: Wang, Qiyao, et al.
Pubblicazione: (2024) -
xJailbreak: Representation Space Guided Reinforcement Learning for Interpretable LLM Jailbreaking
di: Lee, Sunbowen, et al.
Pubblicazione: (2025) -
ETAGE: Enhanced Test Time Adaptation with Integrated Entropy and Gradient Norms for Robust Model Performance
di: Shamsi, Afshar, et al.
Pubblicazione: (2024)