EvoIdeator: Evolving Scientific Ideas through Checklist-Grounded Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sauter, Andreas, Zhao, Yuyue, Urbani, Jacopo, Hu, Wenxiang, Meng, Zaiqiao, Zhou, Lun, Yan, Xiaohui, Lyu, Yougang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evo-DKD: Dual-Knowledge Decoding for Autonomous Ontology Evolution in Large Language Models
von: Raman, Vishal, et al.
Veröffentlicht: (2025)
von: Raman, Vishal, et al.
Veröffentlicht: (2025)
CoupleEvo: Evolving Heuristics for Coupled Optimization Problems Using Large Language Models
von: Bömer, Thomas, et al.
Veröffentlicht: (2026)
von: Bömer, Thomas, et al.
Veröffentlicht: (2026)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
Survey Transfer Learning: Recycling Data with Silicon Responses
von: Amini, Ali
Veröffentlicht: (2025)
von: Amini, Ali
Veröffentlicht: (2025)
Learned Relay Representations for Forward-Thinking Discrete Diffusion Models
von: Rozonoyer, Benjamin, et al.
Veröffentlicht: (2026)
von: Rozonoyer, Benjamin, et al.
Veröffentlicht: (2026)
Adapting While Learning: Grounding LLMs for Scientific Problems with Intelligent Tool Usage Adaptation
von: Lyu, Bohan, et al.
Veröffentlicht: (2024)
von: Lyu, Bohan, et al.
Veröffentlicht: (2024)
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
von: Luijkx, Jelle, et al.
Veröffentlicht: (2025)
von: Luijkx, Jelle, et al.
Veröffentlicht: (2025)
PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues
von: Kajare, Prajwal Vijay, et al.
Veröffentlicht: (2026)
von: Kajare, Prajwal Vijay, et al.
Veröffentlicht: (2026)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
Learning Natural Language Constraints for Safe Reinforcement Learning of Language Agents
von: Chua, Jaymari, et al.
Veröffentlicht: (2025)
von: Chua, Jaymari, et al.
Veröffentlicht: (2025)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
von: Du, Bangde, et al.
Veröffentlicht: (2025)
von: Du, Bangde, et al.
Veröffentlicht: (2025)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
von: Costa, Rimom
Veröffentlicht: (2025)
von: Costa, Rimom
Veröffentlicht: (2025)
Learning the meanings of function words from grounded language using a visual question answering model
von: Portelance, Eva, et al.
Veröffentlicht: (2023)
von: Portelance, Eva, et al.
Veröffentlicht: (2023)
OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
von: Yang, Shan
Veröffentlicht: (2026)
von: Yang, Shan
Veröffentlicht: (2026)
HIP Network: Historical Information Passing Network for Extrapolation Reasoning on Temporal Knowledge Graph
von: He, Yongquan, et al.
Veröffentlicht: (2024)
von: He, Yongquan, et al.
Veröffentlicht: (2024)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
von: Hashemi, Helia, et al.
Veröffentlicht: (2024)
von: Hashemi, Helia, et al.
Veröffentlicht: (2024)
Latent Cache Flow: Model-to-Model Communication Without Text
von: Rossi, Maximillian, et al.
Veröffentlicht: (2026)
von: Rossi, Maximillian, et al.
Veröffentlicht: (2026)
SAGE: A Strategy-Aware Graph-Enhanced Generation Framework For Online Counseling
von: Aharon, Eliya Naomi, et al.
Veröffentlicht: (2026)
von: Aharon, Eliya Naomi, et al.
Veröffentlicht: (2026)
A Domain-Agnostic Neurosymbolic Approach for Big Social Data Analysis: Evaluating Mental Health Sentiment on Social Media during COVID-19
von: Khandelwal, Vedant, et al.
Veröffentlicht: (2024)
von: Khandelwal, Vedant, et al.
Veröffentlicht: (2024)
Automated Circuit Interpretation via Probe Prompting
von: Birardi, Giuseppe
Veröffentlicht: (2025)
von: Birardi, Giuseppe
Veröffentlicht: (2025)
Unleashing LLMs in Bayesian Optimization: Preference-Guided Framework for Scientific Discovery
von: Yuan, Xinzhe, et al.
Veröffentlicht: (2026)
von: Yuan, Xinzhe, et al.
Veröffentlicht: (2026)
Graph Your Way to Inspiration: Integrating Co-Author Graphs with Retrieval-Augmented Generation for Large Language Model Based Scientific Idea Generation
von: Xie, Pengzhen, et al.
Veröffentlicht: (2025)
von: Xie, Pengzhen, et al.
Veröffentlicht: (2025)
Automated Bug Triaging using Instruction-Tuned Large Language Models
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
Merge-Bench: Resolve Merge Conflicts with Large Language Models
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
ACE: Exploring Activation Cosine Similarity and Variance for Accurate and Calibration-Efficient LLM Pruning
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
von: Gwak, Jiho, et al.
Veröffentlicht: (2025)
von: Gwak, Jiho, et al.
Veröffentlicht: (2025)
On the Influence of Discourse Relations in Persuasive Texts
von: Turk, Nawar, et al.
Veröffentlicht: (2025)
von: Turk, Nawar, et al.
Veröffentlicht: (2025)
Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs
von: Easley, Eric, et al.
Veröffentlicht: (2026)
von: Easley, Eric, et al.
Veröffentlicht: (2026)
Calibrated Confidence Estimation for Tabular Question Answering
von: Voss, Lukas
Veröffentlicht: (2026)
von: Voss, Lukas
Veröffentlicht: (2026)
Automated CAD Modeling Sequence Generation from Text Descriptions via Transformer-Based Large Language Models
von: Liao, Jianxing, et al.
Veröffentlicht: (2025)
von: Liao, Jianxing, et al.
Veröffentlicht: (2025)
Scalable GPU-Accelerated Euler Characteristic Curves: Optimization and Differentiable Learning for PyTorch
von: Saxena, Udit
Veröffentlicht: (2025)
von: Saxena, Udit
Veröffentlicht: (2025)
JURY-RL: Votes Propose, Proofs Dispose for Label-Free RLVR
von: Chen, Xinjie, et al.
Veröffentlicht: (2026)
von: Chen, Xinjie, et al.
Veröffentlicht: (2026)
Descriptive Collision in Sparse Autoencoder Auto-Interpretability: When One Explanation Describes Many Features
von: McCann, Jordan F.
Veröffentlicht: (2026)
von: McCann, Jordan F.
Veröffentlicht: (2026)
Super Apriel: One Checkpoint, Many Speeds
von: Labs, SLAM, et al.
Veröffentlicht: (2026)
von: Labs, SLAM, et al.
Veröffentlicht: (2026)
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind
von: Srishty, Sharmin Sultana, et al.
Veröffentlicht: (2026)
von: Srishty, Sharmin Sultana, et al.
Veröffentlicht: (2026)
Synergy over Discrepancy: A Partition-Based Approach to Multi-Domain LLM Fine-Tuning
von: Ye, Hua, et al.
Veröffentlicht: (2025)
von: Ye, Hua, et al.
Veröffentlicht: (2025)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
von: Han, Xudong, et al.
Veröffentlicht: (2025)
von: Han, Xudong, et al.
Veröffentlicht: (2025)
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
von: Kugler, Kai
Veröffentlicht: (2025)
von: Kugler, Kai
Veröffentlicht: (2025)
Ähnliche Einträge
-
Evo-DKD: Dual-Knowledge Decoding for Autonomous Ontology Evolution in Large Language Models
von: Raman, Vishal, et al.
Veröffentlicht: (2025) -
CoupleEvo: Evolving Heuristics for Coupled Optimization Problems Using Large Language Models
von: Bömer, Thomas, et al.
Veröffentlicht: (2026) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025) -
Survey Transfer Learning: Recycling Data with Silicon Responses
von: Amini, Ali
Veröffentlicht: (2025) -
Learned Relay Representations for Forward-Thinking Discrete Diffusion Models
von: Rozonoyer, Benjamin, et al.
Veröffentlicht: (2026)