RECODE-H: A Benchmark for Research Code Development with Interactive Human Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Miao, Chunyu, Zou, Henry Peng, Li, Yangning, Chen, Yankai, Wang, Yibo, Wang, Fangxin, Li, Yifan, Yang, Wooseong, He, Bowei, Zhang, Xinni, Yu, Dianzhi, Yang, Hanchen, Nguyen, Hoang H, Zhou, Yue, Yang, Jie, Guo, Jizhou, Fan, Wenzhe, Yeh, Chin-Yuan, Meng, Panpan, Fang, Liancheng, Qi, Jinhu, Huang, Wei-Chieh, Gu, Zhengyao, Han, Yuwei, He, Langzhou, Yang, Yuyao, Li, Yinghui, Zheng, Hai-Tao, Liu, Xue, King, Irwin, Yu, Philip S. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deep Research with Open-Domain Evaluation and Multi-Stage Guardrails for Safety
by: Huang, Wei-Chieh, et al.
Published: (2025)
by: Huang, Wei-Chieh, et al.
Published: (2025)
Recent Advances of Multimodal Continual Learning: A Comprehensive Survey
by: Yu, Dianzhi, et al.
Published: (2024)
by: Yu, Dianzhi, et al.
Published: (2024)
LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey
by: Zou, Henry Peng, et al.
Published: (2025)
by: Zou, Henry Peng, et al.
Published: (2025)
When Users Change Their Mind: Evaluating Interruptible Agents in Long-Horizon Web Navigation
by: Zou, Henry Peng, et al.
Published: (2026)
by: Zou, Henry Peng, et al.
Published: (2026)
Embracing Trustworthy Brain-Agent Collaboration as Paradigm Extension for Intelligent Assistive Technologies
by: Chen, Yankai, et al.
Published: (2025)
by: Chen, Yankai, et al.
Published: (2025)
TestNUC: Enhancing Test-Time Computing Approaches and Scaling through Neighboring Unlabeled Data Consistency
by: Zou, Henry Peng, et al.
Published: (2025)
by: Zou, Henry Peng, et al.
Published: (2025)
Mining Intrinsic Rewards from LLM Hidden States for Efficient Best-of-N Sampling
by: Guo, Jizhou, et al.
Published: (2025)
by: Guo, Jizhou, et al.
Published: (2025)
ConSurv: Multimodal Continual Learning for Survival Analysis
by: Yu, Dianzhi, et al.
Published: (2025)
by: Yu, Dianzhi, et al.
Published: (2025)
Sensitivity of the RECODE Reactor CEvNS Experiment to the Dark Axion Portal
by: Shen, Yingji, et al.
Published: (2026)
by: Shen, Yingji, et al.
Published: (2026)
OKG-LLM: Aligning Ocean Knowledge Graph with Observation Data via LLMs for Global Sea Surface Temperature Prediction
by: Yang, Hanchen, et al.
Published: (2025)
by: Yang, Hanchen, et al.
Published: (2025)
Filter-then-Weight: Online Data Selection and Reweighting for LLM Fine-Tuning
by: Wang, Fangxin, et al.
Published: (2026)
by: Wang, Fangxin, et al.
Published: (2026)
An Entropy-based Text Watermarking Detection Method
by: Lu, Yijian, et al.
Published: (2024)
by: Lu, Yijian, et al.
Published: (2024)
Learning Binarized Representations with Pseudo-positive Sample Enhancement for Efficient Graph Collaborative Filtering
by: Chen, Yankai, et al.
Published: (2025)
by: Chen, Yankai, et al.
Published: (2025)
AdmTree: Compressing Lengthy Context with Adaptive Semantic Trees
by: Li, Yangning, et al.
Published: (2025)
by: Li, Yangning, et al.
Published: (2025)
1700 groups of frequently used chinese synonyms / Yang Jizhou; Jia Yongfen bian zhu
by: Yang, Jizhou
by: Yang, Jizhou
LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition
by: Chen, Yanyu, et al.
Published: (2026)
by: Chen, Yanyu, et al.
Published: (2026)
Multi-Agent Autonomous Driving Systems with Large Language Models: A Survey of Recent Advances
by: Wu, Yaozu, et al.
Published: (2025)
by: Wu, Yaozu, et al.
Published: (2025)
A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy
by: Zou, Henry Peng, et al.
Published: (2025)
by: Zou, Henry Peng, et al.
Published: (2025)
RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation
by: Li, Shijun, et al.
Published: (2026)
by: Li, Shijun, et al.
Published: (2026)
Adaptive Candidate Retrieval with Dynamic Knowledge Graph Construction for Cold-Start Recommendation
by: Yang, Wooseong, et al.
Published: (2025)
by: Yang, Wooseong, et al.
Published: (2025)
Can Molecular Foundation Models Know What They Don't Know? A Simple Remedy with Preference Optimization
by: He, Langzhou, et al.
Published: (2025)
by: He, Langzhou, et al.
Published: (2025)
Distributionally Robust Set Representation Learning Under Inference-Time Element Corruption
by: Chen, Yankai, et al.
Published: (2026)
by: Chen, Yankai, et al.
Published: (2026)
Teaching According to Talents! Instruction Tuning LLMs with Competence-Aware Curriculum Learning
by: Li, Yangning, et al.
Published: (2025)
by: Li, Yangning, et al.
Published: (2025)
Embracing Ambiguity: Improving Similarity-oriented Tasks with Contextual Synonym Knowledge
by: Li, Yangning, et al.
Published: (2022)
by: Li, Yangning, et al.
Published: (2022)
Entropy, Thermodynamics and the Geometrization of the Language Model
by: Yang, Wenzhe
Published: (2024)
by: Yang, Wenzhe
Published: (2024)
MDIT: A Model-free Data Interpolation Method for Diverse Instruction Tuning
by: Li, Yangning, et al.
Published: (2025)
by: Li, Yangning, et al.
Published: (2025)
From Retrieval to Generation: Efficient and Effective Entity Set Expansion
by: Huang, Shulin, et al.
Published: (2023)
by: Huang, Shulin, et al.
Published: (2023)
Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security
by: Qi, Jinhu, et al.
Published: (2026)
by: Qi, Jinhu, et al.
Published: (2026)
Targeted Screening of Cyclooxygenase‐2 Inhibitors From Dendropanax dentiger Root Using Affinity Ultrafiltration Coupled With UHPLC‐MS
by: Qihui Wang, et al.
Published: (2025)
by: Qihui Wang, et al.
Published: (2025)
Co-RaL: Complementary Radar-Leg Odometry with 4-DoF Optimization and Rolling Contact
by: Jung, Sangwoo, et al.
Published: (2024)
by: Jung, Sangwoo, et al.
Published: (2024)
Ground-Optimized 4D Radar-Inertial Odometry via Continuous Velocity Integration using Gaussian Process
by: Yang, Wooseong, et al.
Published: (2025)
by: Yang, Wooseong, et al.
Published: (2025)
Subgraph-level Universal Prompt Tuning
by: Lee, Junhyun, et al.
Published: (2024)
by: Lee, Junhyun, et al.
Published: (2024)
DP-DGAD: A Generalist Dynamic Graph Anomaly Detector with Dynamic Prototypes
by: Zheng, Jialun, et al.
Published: (2025)
by: Zheng, Jialun, et al.
Published: (2025)
PSG-Agent: Personality-Aware Safety Guardrail for LLM-based Agents
by: Wu, Yaozu, et al.
Published: (2025)
by: Wu, Yaozu, et al.
Published: (2025)
Item Cluster-aware Prompt Learning for Session-based Recommendation
by: Yang, Wooseong, et al.
Published: (2024)
by: Yang, Wooseong, et al.
Published: (2024)
Actor-Curator: Co-adaptive Curriculum Learning via Policy-Improvement Bandits for RL Post-Training
by: Gu, Zhengyao, et al.
Published: (2026)
by: Gu, Zhengyao, et al.
Published: (2026)
Plausible Colloidal Methods to Synthesize Semiconductor Nanowires: Deep Study From ZnSe Nanorods
by: Chunyu Yu, et al.
Published: (2024)
by: Chunyu Yu, et al.
Published: (2024)
Magnesium Chalcogenide Nanocrystals and Their Core/Shell Nanostructures
by: Chunyu Yu, et al.
Published: (2024)
by: Chunyu Yu, et al.
Published: (2024)
Process-Level Trajectory Evaluation for Environment Configuration in Software Engineering Agents
by: Kuang, Jiayi, et al.
Published: (2025)
by: Kuang, Jiayi, et al.
Published: (2025)
Mitigating Large Language Model Hallucination with Faithful Finetuning
by: Hu, Minda, et al.
Published: (2024)
by: Hu, Minda, et al.
Published: (2024)
Similar Items
-
Deep Research with Open-Domain Evaluation and Multi-Stage Guardrails for Safety
by: Huang, Wei-Chieh, et al.
Published: (2025) -
Recent Advances of Multimodal Continual Learning: A Comprehensive Survey
by: Yu, Dianzhi, et al.
Published: (2024) -
LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey
by: Zou, Henry Peng, et al.
Published: (2025) -
When Users Change Their Mind: Evaluating Interruptible Agents in Long-Horizon Web Navigation
by: Zou, Henry Peng, et al.
Published: (2026) -
Embracing Trustworthy Brain-Agent Collaboration as Paradigm Extension for Intelligent Assistive Technologies
by: Chen, Yankai, et al.
Published: (2025)