RECODE-H: A Benchmark for Research Code Development with Interactive Human Feedback
Fuente:
arXiv
Salvato in:
| Autori principali: | Miao, Chunyu, Zou, Henry Peng, Li, Yangning, Chen, Yankai, Wang, Yibo, Wang, Fangxin, Li, Yifan, Yang, Wooseong, He, Bowei, Zhang, Xinni, Yu, Dianzhi, Yang, Hanchen, Nguyen, Hoang H, Zhou, Yue, Yang, Jie, Guo, Jizhou, Fan, Wenzhe, Yeh, Chin-Yuan, Meng, Panpan, Fang, Liancheng, Qi, Jinhu, Huang, Wei-Chieh, Gu, Zhengyao, Han, Yuwei, He, Langzhou, Yang, Yuyao, Li, Yinghui, Zheng, Hai-Tao, Liu, Xue, King, Irwin, Yu, Philip S. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Deep Research with Open-Domain Evaluation and Multi-Stage Guardrails for Safety
di: Huang, Wei-Chieh, et al.
Pubblicazione: (2025)
di: Huang, Wei-Chieh, et al.
Pubblicazione: (2025)
Recent Advances of Multimodal Continual Learning: A Comprehensive Survey
di: Yu, Dianzhi, et al.
Pubblicazione: (2024)
di: Yu, Dianzhi, et al.
Pubblicazione: (2024)
LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey
di: Zou, Henry Peng, et al.
Pubblicazione: (2025)
di: Zou, Henry Peng, et al.
Pubblicazione: (2025)
When Users Change Their Mind: Evaluating Interruptible Agents in Long-Horizon Web Navigation
di: Zou, Henry Peng, et al.
Pubblicazione: (2026)
di: Zou, Henry Peng, et al.
Pubblicazione: (2026)
Embracing Trustworthy Brain-Agent Collaboration as Paradigm Extension for Intelligent Assistive Technologies
di: Chen, Yankai, et al.
Pubblicazione: (2025)
di: Chen, Yankai, et al.
Pubblicazione: (2025)
TestNUC: Enhancing Test-Time Computing Approaches and Scaling through Neighboring Unlabeled Data Consistency
di: Zou, Henry Peng, et al.
Pubblicazione: (2025)
di: Zou, Henry Peng, et al.
Pubblicazione: (2025)
Mining Intrinsic Rewards from LLM Hidden States for Efficient Best-of-N Sampling
di: Guo, Jizhou, et al.
Pubblicazione: (2025)
di: Guo, Jizhou, et al.
Pubblicazione: (2025)
ConSurv: Multimodal Continual Learning for Survival Analysis
di: Yu, Dianzhi, et al.
Pubblicazione: (2025)
di: Yu, Dianzhi, et al.
Pubblicazione: (2025)
Sensitivity of the RECODE Reactor CEvNS Experiment to the Dark Axion Portal
di: Shen, Yingji, et al.
Pubblicazione: (2026)
di: Shen, Yingji, et al.
Pubblicazione: (2026)
OKG-LLM: Aligning Ocean Knowledge Graph with Observation Data via LLMs for Global Sea Surface Temperature Prediction
di: Yang, Hanchen, et al.
Pubblicazione: (2025)
di: Yang, Hanchen, et al.
Pubblicazione: (2025)
Filter-then-Weight: Online Data Selection and Reweighting for LLM Fine-Tuning
di: Wang, Fangxin, et al.
Pubblicazione: (2026)
di: Wang, Fangxin, et al.
Pubblicazione: (2026)
An Entropy-based Text Watermarking Detection Method
di: Lu, Yijian, et al.
Pubblicazione: (2024)
di: Lu, Yijian, et al.
Pubblicazione: (2024)
Learning Binarized Representations with Pseudo-positive Sample Enhancement for Efficient Graph Collaborative Filtering
di: Chen, Yankai, et al.
Pubblicazione: (2025)
di: Chen, Yankai, et al.
Pubblicazione: (2025)
AdmTree: Compressing Lengthy Context with Adaptive Semantic Trees
di: Li, Yangning, et al.
Pubblicazione: (2025)
di: Li, Yangning, et al.
Pubblicazione: (2025)
1700 groups of frequently used chinese synonyms / Yang Jizhou; Jia Yongfen bian zhu
di: Yang, Jizhou
di: Yang, Jizhou
LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition
di: Chen, Yanyu, et al.
Pubblicazione: (2026)
di: Chen, Yanyu, et al.
Pubblicazione: (2026)
Multi-Agent Autonomous Driving Systems with Large Language Models: A Survey of Recent Advances
di: Wu, Yaozu, et al.
Pubblicazione: (2025)
di: Wu, Yaozu, et al.
Pubblicazione: (2025)
A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy
di: Zou, Henry Peng, et al.
Pubblicazione: (2025)
di: Zou, Henry Peng, et al.
Pubblicazione: (2025)
RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation
di: Li, Shijun, et al.
Pubblicazione: (2026)
di: Li, Shijun, et al.
Pubblicazione: (2026)
Adaptive Candidate Retrieval with Dynamic Knowledge Graph Construction for Cold-Start Recommendation
di: Yang, Wooseong, et al.
Pubblicazione: (2025)
di: Yang, Wooseong, et al.
Pubblicazione: (2025)
Can Molecular Foundation Models Know What They Don't Know? A Simple Remedy with Preference Optimization
di: He, Langzhou, et al.
Pubblicazione: (2025)
di: He, Langzhou, et al.
Pubblicazione: (2025)
Distributionally Robust Set Representation Learning Under Inference-Time Element Corruption
di: Chen, Yankai, et al.
Pubblicazione: (2026)
di: Chen, Yankai, et al.
Pubblicazione: (2026)
Teaching According to Talents! Instruction Tuning LLMs with Competence-Aware Curriculum Learning
di: Li, Yangning, et al.
Pubblicazione: (2025)
di: Li, Yangning, et al.
Pubblicazione: (2025)
Embracing Ambiguity: Improving Similarity-oriented Tasks with Contextual Synonym Knowledge
di: Li, Yangning, et al.
Pubblicazione: (2022)
di: Li, Yangning, et al.
Pubblicazione: (2022)
Entropy, Thermodynamics and the Geometrization of the Language Model
di: Yang, Wenzhe
Pubblicazione: (2024)
di: Yang, Wenzhe
Pubblicazione: (2024)
MDIT: A Model-free Data Interpolation Method for Diverse Instruction Tuning
di: Li, Yangning, et al.
Pubblicazione: (2025)
di: Li, Yangning, et al.
Pubblicazione: (2025)
From Retrieval to Generation: Efficient and Effective Entity Set Expansion
di: Huang, Shulin, et al.
Pubblicazione: (2023)
di: Huang, Shulin, et al.
Pubblicazione: (2023)
Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security
di: Qi, Jinhu, et al.
Pubblicazione: (2026)
di: Qi, Jinhu, et al.
Pubblicazione: (2026)
Targeted Screening of Cyclooxygenase‐2 Inhibitors From Dendropanax dentiger Root Using Affinity Ultrafiltration Coupled With UHPLC‐MS
di: Qihui Wang, et al.
Pubblicazione: (2025)
di: Qihui Wang, et al.
Pubblicazione: (2025)
Co-RaL: Complementary Radar-Leg Odometry with 4-DoF Optimization and Rolling Contact
di: Jung, Sangwoo, et al.
Pubblicazione: (2024)
di: Jung, Sangwoo, et al.
Pubblicazione: (2024)
Ground-Optimized 4D Radar-Inertial Odometry via Continuous Velocity Integration using Gaussian Process
di: Yang, Wooseong, et al.
Pubblicazione: (2025)
di: Yang, Wooseong, et al.
Pubblicazione: (2025)
Subgraph-level Universal Prompt Tuning
di: Lee, Junhyun, et al.
Pubblicazione: (2024)
di: Lee, Junhyun, et al.
Pubblicazione: (2024)
DP-DGAD: A Generalist Dynamic Graph Anomaly Detector with Dynamic Prototypes
di: Zheng, Jialun, et al.
Pubblicazione: (2025)
di: Zheng, Jialun, et al.
Pubblicazione: (2025)
PSG-Agent: Personality-Aware Safety Guardrail for LLM-based Agents
di: Wu, Yaozu, et al.
Pubblicazione: (2025)
di: Wu, Yaozu, et al.
Pubblicazione: (2025)
Item Cluster-aware Prompt Learning for Session-based Recommendation
di: Yang, Wooseong, et al.
Pubblicazione: (2024)
di: Yang, Wooseong, et al.
Pubblicazione: (2024)
Actor-Curator: Co-adaptive Curriculum Learning via Policy-Improvement Bandits for RL Post-Training
di: Gu, Zhengyao, et al.
Pubblicazione: (2026)
di: Gu, Zhengyao, et al.
Pubblicazione: (2026)
Plausible Colloidal Methods to Synthesize Semiconductor Nanowires: Deep Study From ZnSe Nanorods
di: Chunyu Yu, et al.
Pubblicazione: (2024)
di: Chunyu Yu, et al.
Pubblicazione: (2024)
Magnesium Chalcogenide Nanocrystals and Their Core/Shell Nanostructures
di: Chunyu Yu, et al.
Pubblicazione: (2024)
di: Chunyu Yu, et al.
Pubblicazione: (2024)
Process-Level Trajectory Evaluation for Environment Configuration in Software Engineering Agents
di: Kuang, Jiayi, et al.
Pubblicazione: (2025)
di: Kuang, Jiayi, et al.
Pubblicazione: (2025)
Mitigating Large Language Model Hallucination with Faithful Finetuning
di: Hu, Minda, et al.
Pubblicazione: (2024)
di: Hu, Minda, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Deep Research with Open-Domain Evaluation and Multi-Stage Guardrails for Safety
di: Huang, Wei-Chieh, et al.
Pubblicazione: (2025) -
Recent Advances of Multimodal Continual Learning: A Comprehensive Survey
di: Yu, Dianzhi, et al.
Pubblicazione: (2024) -
LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey
di: Zou, Henry Peng, et al.
Pubblicazione: (2025) -
When Users Change Their Mind: Evaluating Interruptible Agents in Long-Horizon Web Navigation
di: Zou, Henry Peng, et al.
Pubblicazione: (2026) -
Embracing Trustworthy Brain-Agent Collaboration as Paradigm Extension for Intelligent Assistive Technologies
di: Chen, Yankai, et al.
Pubblicazione: (2025)