Research Re: search & Re-search
Fuente:
arXiv
Saved in:
| Main Author: | Plaat, Aske |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CausalPlayground: Addressing Data-Generation Requirements in Cutting-Edge Causality Research
by: Sauter, Andreas W M, et al.
Published: (2024)
by: Sauter, Andreas W M, et al.
Published: (2024)
Analysis of Bluffing by DQN and CFR in Leduc Hold'em Poker
by: Zaciragic, Tarik, et al.
Published: (2025)
by: Zaciragic, Tarik, et al.
Published: (2025)
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
Baba is LLM: Reasoning in a Game with Dynamic Rules
by: van Wetten, Fien, et al.
Published: (2025)
by: van Wetten, Fien, et al.
Published: (2025)
ACTIVA: Amortized Causal Effect Estimation via Transformer-based Variational Autoencoder
by: Sauter, Andreas, et al.
Published: (2025)
by: Sauter, Andreas, et al.
Published: (2025)
CoComposer: LLM Multi-agent Collaborative Music Composition
by: Xing, Peiwen, et al.
Published: (2025)
by: Xing, Peiwen, et al.
Published: (2025)
Reset-free Reinforcement Learning with World Models
by: Yang, Zhao, et al.
Published: (2024)
by: Yang, Zhao, et al.
Published: (2024)
Chargax: A JAX Accelerated EV Charging Simulator
by: Ponse, Koen, et al.
Published: (2025)
by: Ponse, Koen, et al.
Published: (2025)
Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learning
by: Spoor, Lindsay, et al.
Published: (2025)
by: Spoor, Lindsay, et al.
Published: (2025)
Explicitly Disentangled Representations in Object-Centric Learning
by: Majellaro, Riccardo, et al.
Published: (2024)
by: Majellaro, Riccardo, et al.
Published: (2024)
Mirror Mode in Fire Emblem: Beating Players at their own Game with Imitation and Reinforcement Learning
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
Solving Deep Reinforcement Learning Tasks with Evolution Strategies and Linear Policy Networks
by: Wong, Annie, et al.
Published: (2024)
by: Wong, Annie, et al.
Published: (2024)
Reasoning Capabilities of Large Language Models on Dynamic Tasks
by: Wong, Annie, et al.
Published: (2025)
by: Wong, Annie, et al.
Published: (2025)
State Design Matters: How Representations Shape Dynamic Reasoning in Large Language Models
by: Wong, Annie, et al.
Published: (2026)
by: Wong, Annie, et al.
Published: (2026)
Guiding Skill Discovery with Foundation Models
by: Yang, Zhao, et al.
Published: (2025)
by: Yang, Zhao, et al.
Published: (2025)
A Benchmark Study of Deep Reinforcement Learning Algorithms for the Container Stowage Planning Problem
by: Huang, Yunqi, et al.
Published: (2025)
by: Huang, Yunqi, et al.
Published: (2025)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
by: Sauter, Andreas W. M., et al.
Published: (2024)
by: Sauter, Andreas W. M., et al.
Published: (2024)
Assessing Reproducibility in Evolutionary Computation: A Case Study using Human- and LLM-based Assessment
by: Da Ros, Francesca, et al.
Published: (2026)
by: Da Ros, Francesca, et al.
Published: (2026)
Multi-Step Reasoning with Large Language Models, a Survey
by: Plaat, Aske, et al.
Published: (2024)
by: Plaat, Aske, et al.
Published: (2024)
Reinforcement Learning for Sustainable Energy: A Survey
by: Ponse, Koen, et al.
Published: (2024)
by: Ponse, Koen, et al.
Published: (2024)
Advanced For-Loop for QML algorithm search
by: Wong, FuTe
Published: (2025)
by: Wong, FuTe
Published: (2025)
Prime the search: Using large language models for guiding geometric task and motion planning by warm-starting tree search
by: Lee, Dongryung, et al.
Published: (2025)
by: Lee, Dongryung, et al.
Published: (2025)
A Hybrid Intelligence Method for Argument Mining
by: van der Meer, Michiel, et al.
Published: (2024)
by: van der Meer, Michiel, et al.
Published: (2024)
Agentic Large Language Models, a survey
by: Plaat, Aske, et al.
Published: (2025)
by: Plaat, Aske, et al.
Published: (2025)
ReAttn: Improving Attention-based Re-ranking via Attention Re-weighting
by: Tian, Yuxing, et al.
Published: (2026)
by: Tian, Yuxing, et al.
Published: (2026)
Focused ReAct: Improving ReAct through Reiterate and Early Stop
by: Li, Shuoqiu, et al.
Published: (2024)
by: Li, Shuoqiu, et al.
Published: (2024)
Co-ReAct: Rubrics as Step-Level Collaborators for ReAct Agents
by: Kang, Jiazheng, et al.
Published: (2026)
by: Kang, Jiazheng, et al.
Published: (2026)
Re$^2$Math: Benchmarking Theorem Retrieval in Research-Level Mathematics
by: Lyu, Zicheng, et al.
Published: (2026)
by: Lyu, Zicheng, et al.
Published: (2026)
$\textbf{Re}^{2}$: Unlocking LLM Reasoning via Reinforcement Learning with Re-solving
by: Wang, Pinzheng, et al.
Published: (2026)
by: Wang, Pinzheng, et al.
Published: (2026)
Topological quantification of ambiguity in semantic search
by: Barillot, Thomas Roland, et al.
Published: (2024)
by: Barillot, Thomas Roland, et al.
Published: (2024)
Generative structure search for efficient and diverse discovery of molecular and crystal structures
by: Qin, Yifang, et al.
Published: (2026)
by: Qin, Yifang, et al.
Published: (2026)
Task tree retrieval from FOON using search algorithms
by: Attapu, Amitha
Published: (2023)
by: Attapu, Amitha
Published: (2023)
Proposing and solving olympiad geometry with guided tree search
by: Zhang, Chi, et al.
Published: (2024)
by: Zhang, Chi, et al.
Published: (2024)
Hierarchical quantum circuit representations for neural architecture search
by: Lourens, Matt, et al.
Published: (2022)
by: Lourens, Matt, et al.
Published: (2022)
Evaluation Pipeline for systematically searching for Anomaly Detection Systems
by: Rokohl, Florian, et al.
Published: (2025)
by: Rokohl, Florian, et al.
Published: (2025)
N-ReLU: Zero-Mean Stochastic Extension of ReLU
by: Manik, Md Motaleb Hossen, et al.
Published: (2025)
by: Manik, Md Motaleb Hossen, et al.
Published: (2025)
The Resurrection of the ReLU
by: Horuz, Coşku Can, et al.
Published: (2025)
by: Horuz, Coşku Can, et al.
Published: (2025)
ReIFE: Re-evaluating Instruction-Following Evaluation
by: Liu, Yixin, et al.
Published: (2024)
by: Liu, Yixin, et al.
Published: (2024)
RePo: Language Models with Context Re-Positioning
by: Li, Huayang, et al.
Published: (2025)
by: Li, Huayang, et al.
Published: (2025)
RePO: Understanding Preference Learning Through ReLU-Based Optimization
by: Wu, Junkang, et al.
Published: (2025)
by: Wu, Junkang, et al.
Published: (2025)
Similar Items
-
CausalPlayground: Addressing Data-Generation Requirements in Cutting-Edge Causality Research
by: Sauter, Andreas W M, et al.
Published: (2024) -
Analysis of Bluffing by DQN and CFR in Leduc Hold'em Poker
by: Zaciragic, Tarik, et al.
Published: (2025) -
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
by: Chen, Xi, et al.
Published: (2025) -
Baba is LLM: Reasoning in a Game with Dynamic Rules
by: van Wetten, Fien, et al.
Published: (2025) -
ACTIVA: Amortized Causal Effect Estimation via Transformer-based Variational Autoencoder
by: Sauter, Andreas, et al.
Published: (2025)