What Matters in Hierarchical Search for Combinatorial Reasoning Problems?
Fuente:
arXiv
Saved in:
| Main Authors: | Zawalski, Michał, Góral, Gracjan, Tyrolski, Michał, Wiśnios, Emilia, Budrowski, Franciszek, Cygan, Marek, Kuciński, Łukasz, Miłoś, Piotr |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast and Precise: Adjusting Planning Horizon with Adaptive Subgoal Search
by: Zawalski, Michał, et al.
Published: (2022)
by: Zawalski, Michał, et al.
Published: (2022)
Subgoal Search For Complex Reasoning Tasks
by: Czechowski, Konrad, et al.
Published: (2021)
by: Czechowski, Konrad, et al.
Published: (2021)
OpenGVL -- Benchmarking Visual Temporal Progress for Data Curation
by: Budzianowski, Paweł, et al.
Published: (2025)
by: Budzianowski, Paweł, et al.
Published: (2025)
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
by: Góral, Gracjan, et al.
Published: (2024)
by: Góral, Gracjan, et al.
Published: (2024)
Contrastive Representations for Temporal Reasoning
by: Ziarko, Alicja, et al.
Published: (2025)
by: Ziarko, Alicja, et al.
Published: (2025)
Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
by: Nauman, Michal, et al.
Published: (2024)
by: Nauman, Michal, et al.
Published: (2024)
Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem
by: Wołczyk, Maciej, et al.
Published: (2024)
by: Wołczyk, Maciej, et al.
Published: (2024)
Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control
by: Nauman, Michal, et al.
Published: (2024)
by: Nauman, Michal, et al.
Published: (2024)
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2024)
by: Góral, Gracjan, et al.
Published: (2024)
On the Theory of Risk-Aware Agents: Bridging Actor-Critic and Economics
by: Nauman, Michal, et al.
Published: (2023)
by: Nauman, Michal, et al.
Published: (2023)
Off-Policy Correction For Multi-Agent Reinforcement Learning
by: Zawalski, Michał, et al.
Published: (2021)
by: Zawalski, Michał, et al.
Published: (2021)
Depth-Wise Activation Steering for Honest Language Models
by: Góral, Gracjan, et al.
Published: (2025)
by: Góral, Gracjan, et al.
Published: (2025)
tsGT: Stochastic Time Series Modeling With Transformer
by: Kuciński, Łukasz, et al.
Published: (2024)
by: Kuciński, Łukasz, et al.
Published: (2024)
A Case for Validation Buffer in Pessimistic Actor-Critic
by: Nauman, Michal, et al.
Published: (2024)
by: Nauman, Michal, et al.
Published: (2024)
Reward-Conditioned Reinforcement Learning
by: Nauman, Michal, et al.
Published: (2026)
by: Nauman, Michal, et al.
Published: (2026)
RoboMorph: Evolving Robot Morphology using Large Language Models
by: Qiu, Kevin, et al.
Published: (2024)
by: Qiu, Kevin, et al.
Published: (2024)
Catalytic Role Of Noise And Necessity Of Inductive Biases In The Emergence Of Compositional Communication
by: Kuciński, Łukasz, et al.
Published: (2021)
by: Kuciński, Łukasz, et al.
Published: (2021)
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2025)
by: Góral, Gracjan, et al.
Published: (2025)
Trust Your $\nabla$: Gradient-based Intervention Targeting for Causal Discovery
by: Olko, Mateusz, et al.
Published: (2022)
by: Olko, Mateusz, et al.
Published: (2022)
When Does Non-Uniform Replay Matter in Reinforcement Learning?
by: Korniak, Michal, et al.
Published: (2026)
by: Korniak, Michal, et al.
Published: (2026)
Bigger, Regularized, Categorical: High-Capacity Value Functions are Efficient Multi-Task Learners
by: Nauman, Michal, et al.
Published: (2025)
by: Nauman, Michal, et al.
Published: (2025)
Robotic Control via Embodied Chain-of-Thought Reasoning
by: Zawalski, Michał, et al.
Published: (2024)
by: Zawalski, Michał, et al.
Published: (2024)
Towards consistency of rule-based explainer and black box model -- fusion of rule induction and XAI-based feature importance
by: Kozielski, Michał, et al.
Published: (2024)
by: Kozielski, Michał, et al.
Published: (2024)
Joint MoE Scaling Laws: Mixture of Experts Can Be Memory Efficient
by: Ludziejewski, Jan, et al.
Published: (2025)
by: Ludziejewski, Jan, et al.
Published: (2025)
MoE-Mamba: Efficient Selective State Space Models with Mixture of Experts
by: Pióro, Maciej, et al.
Published: (2024)
by: Pióro, Maciej, et al.
Published: (2024)
Attend or Perish: Benchmarking Attention in Algorithmic Reasoning
by: Spiegel, Michal, et al.
Published: (2025)
by: Spiegel, Michal, et al.
Published: (2025)
Debate2Create: Robot Co-design via Multi-Agent LLM Debate
by: Qiu, Kevin, et al.
Published: (2025)
by: Qiu, Kevin, et al.
Published: (2025)
Feature importance analysis for patient management decisions
by: Valko, Michal, et al.
Published: (2026)
by: Valko, Michal, et al.
Published: (2026)
Distance metric learning for conditional anomaly detection
by: Valko, Michal, et al.
Published: (2026)
by: Valko, Michal, et al.
Published: (2026)
Accelerating Goal-Conditioned RL Algorithms and Research
by: Bortkiewicz, Michał, et al.
Published: (2024)
by: Bortkiewicz, Michał, et al.
Published: (2024)
RapidDock: Unlocking Proteome-scale Molecular Docking
by: Powalski, Rafał, et al.
Published: (2024)
by: Powalski, Rafał, et al.
Published: (2024)
EXALT: EXplainable ALgorithmic Tools for Optimization Problems
by: Bączek, Zuzanna, et al.
Published: (2025)
by: Bączek, Zuzanna, et al.
Published: (2025)
Decoupled Relative Learning Rate Schedules
by: Ludziejewski, Jan, et al.
Published: (2025)
by: Ludziejewski, Jan, et al.
Published: (2025)
VectorEdits: A Dataset and Benchmark for Instruction-Based Editing of Vector Graphics
by: Kuchař, Josef, et al.
Published: (2025)
by: Kuchař, Josef, et al.
Published: (2025)
Learning predictive models for combinations of heterogeneous proteomic data sources
by: Valko, Michal, et al.
Published: (2026)
by: Valko, Michal, et al.
Published: (2026)
GUIDE: Guidance-based Incremental Learning with Diffusion Models
by: Cywiński, Bartosz, et al.
Published: (2024)
by: Cywiński, Bartosz, et al.
Published: (2024)
Differentiation of Blackbox Combinatorial Solvers
by: Vlastelica, Marin, et al.
Published: (2019)
by: Vlastelica, Marin, et al.
Published: (2019)
Vid2Sid: Videos Can Help Close the Sim2Real Gap
by: Qiu, Kevin, et al.
Published: (2026)
by: Qiu, Kevin, et al.
Published: (2026)
Graph Property Inference in Small Language Models: Effects of Representation and Reasoning Strategy
by: Podstawski, Michal
Published: (2026)
by: Podstawski, Michal
Published: (2026)
FlySearch: Exploring how vision-language models explore
by: Pardyl, Adam, et al.
Published: (2025)
by: Pardyl, Adam, et al.
Published: (2025)
Similar Items
-
Fast and Precise: Adjusting Planning Horizon with Adaptive Subgoal Search
by: Zawalski, Michał, et al.
Published: (2022) -
Subgoal Search For Complex Reasoning Tasks
by: Czechowski, Konrad, et al.
Published: (2021) -
OpenGVL -- Benchmarking Visual Temporal Progress for Data Curation
by: Budzianowski, Paweł, et al.
Published: (2025) -
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
by: Góral, Gracjan, et al.
Published: (2024) -
Contrastive Representations for Temporal Reasoning
by: Ziarko, Alicja, et al.
Published: (2025)