Solving Rubik's Cube Without Tricky Sampling
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Yicheng, Liang, Siyu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Wormhole Memory: A Rubik's Cube for Cross-Dialogue Retrieval
by: Wang, Libo
Published: (2025)
by: Wang, Libo
Published: (2025)
Solving a Rubik's Cube Using its Local Graph Structure
by: Yao, Shunyu, et al.
Published: (2024)
by: Yao, Shunyu, et al.
Published: (2024)
Fairness Without Harm: An Influence-Guided Active Sampling Approach
by: Pang, Jinlong, et al.
Published: (2024)
by: Pang, Jinlong, et al.
Published: (2024)
Solving Diffusion Inverse Problems with Restart Posterior Sampling
by: Ahmed, Bilal, et al.
Published: (2025)
by: Ahmed, Bilal, et al.
Published: (2025)
Universality in Collective Intelligence on the Rubik's Cube
by: Krakauer, David, et al.
Published: (2025)
by: Krakauer, David, et al.
Published: (2025)
TCR-GPT: Integrating Autoregressive Model and Reinforcement Learning for T-Cell Receptor Repertoires Generation
by: Lin, Yicheng, et al.
Published: (2024)
by: Lin, Yicheng, et al.
Published: (2024)
Recursive Speculative Decoding: Accelerating LLM Inference via Sampling Without Replacement
by: Jeon, Wonseok, et al.
Published: (2024)
by: Jeon, Wonseok, et al.
Published: (2024)
Node Classification and Search on the Rubik's Cube Graph with GNNs
by: Barro, Alessandro
Published: (2025)
by: Barro, Alessandro
Published: (2025)
CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
by: Wang, Feiyang, et al.
Published: (2025)
by: Wang, Feiyang, et al.
Published: (2025)
Improved Sample Complexity For Diffusion Model Training Without Empirical Risk Minimizer Access
by: Gaur, Mudit, et al.
Published: (2025)
by: Gaur, Mudit, et al.
Published: (2025)
AttNS: Attention-Inspired Numerical Solving For Limited Data Scenarios
by: Huang, Zhongzhan, et al.
Published: (2023)
by: Huang, Zhongzhan, et al.
Published: (2023)
GASE: Graph Attention Sampling with Edges Fusion for Solving Vehicle Routing Problems
by: Wang, Zhenwei, et al.
Published: (2024)
by: Wang, Zhenwei, et al.
Published: (2024)
Inconsistencies In Consistency Models: Better ODE Solving Does Not Imply Better Samples
by: Vouitsis, Noël, et al.
Published: (2024)
by: Vouitsis, Noël, et al.
Published: (2024)
Planning Under Observation Mismatch for Traffic Signal Control via Adaptive Modular World Models
by: Huang, Zherui, et al.
Published: (2025)
by: Huang, Zherui, et al.
Published: (2025)
Mastering Chinese Chess AI (Xiangqi) Without Search
by: Chen, Yu, et al.
Published: (2024)
by: Chen, Yu, et al.
Published: (2024)
A Rubik's Cube inspired approach to Clifford synthesis
by: Bao, Ning, et al.
Published: (2023)
by: Bao, Ning, et al.
Published: (2023)
SplitQuantV2: Enhancing Low-Bit Quantization of LLMs Without GPUs
by: Song, Jaewoo, et al.
Published: (2025)
by: Song, Jaewoo, et al.
Published: (2025)
Weak-Form Evolutionary Kolmogorov-Arnold Networks for Solving Partial Differential Equations
by: Kim, Bongseok, et al.
Published: (2026)
by: Kim, Bongseok, et al.
Published: (2026)
Student Engagement in AI Assisted Complex Problem Solving: A Pilot Study of Human AI Rubik's Cube Collaboration
by: Vanacore, Kirk, et al.
Published: (2025)
by: Vanacore, Kirk, et al.
Published: (2025)
Counterfactual experience augmented off-policy reinforcement learning
by: Lee, Sunbowen, et al.
Published: (2025)
by: Lee, Sunbowen, et al.
Published: (2025)
Offline Multi-Agent Reinforcement Learning via In-Sample Sequential Policy Optimization
by: Liu, Zongkai, et al.
Published: (2024)
by: Liu, Zongkai, et al.
Published: (2024)
Beyond Output Faithfulness: Learning Attributions that Preserve Computational Pathways
by: Zhang, Siyu, et al.
Published: (2025)
by: Zhang, Siyu, et al.
Published: (2025)
Maximum Entropy Exploration Without the Rollouts
by: Adamczyk, Jacob, et al.
Published: (2026)
by: Adamczyk, Jacob, et al.
Published: (2026)
Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates
by: Li, Yibo, et al.
Published: (2026)
by: Li, Yibo, et al.
Published: (2026)
Rethinking Transformers in Solving POMDPs
by: Lu, Chenhao, et al.
Published: (2024)
by: Lu, Chenhao, et al.
Published: (2024)
An End-to-End Deep Reinforcement Learning Approach for Solving the Traveling Salesman Problem with Drones
by: Zeng, Taihelong, et al.
Published: (2025)
by: Zeng, Taihelong, et al.
Published: (2025)
Efficient Bias Mitigation Without Privileged Information
by: Zarlenga, Mateo Espinosa, et al.
Published: (2024)
by: Zarlenga, Mateo Espinosa, et al.
Published: (2024)
GLOP: Learning Global Partition and Local Construction for Solving Large-scale Routing Problems in Real-time
by: Ye, Haoran, et al.
Published: (2023)
by: Ye, Haoran, et al.
Published: (2023)
BioCube: A Multimodal Dataset for Biodiversity Research
by: Stasinos, Stylianos, et al.
Published: (2025)
by: Stasinos, Stylianos, et al.
Published: (2025)
Detecting Scarce and Sparse Anomalous: Solving Dual Imbalance in Multi-Instance Learning
by: Jia, Lin-Han, et al.
Published: (2025)
by: Jia, Lin-Han, et al.
Published: (2025)
A Middle Path for On-Premises LLM Deployment: Preserving Privacy Without Sacrificing Model Confidentiality
by: Huang, Hanbo, et al.
Published: (2024)
by: Huang, Hanbo, et al.
Published: (2024)
TW-CRL: Time-Weighted Contrastive Reward Learning for Efficient Inverse Reinforcement Learning
by: Li, Yuxuan, et al.
Published: (2025)
by: Li, Yuxuan, et al.
Published: (2025)
Analytic Continual Test-Time Adaptation for Multi-Modality Corruption
by: Zhang, Yufei, et al.
Published: (2024)
by: Zhang, Yufei, et al.
Published: (2024)
Semantic Prototypes: Enhancing Transparency Without Black Boxes
by: Menis-Mastromichalakis, Orfeas, et al.
Published: (2024)
by: Menis-Mastromichalakis, Orfeas, et al.
Published: (2024)
Learning Quantifiable Visual Explanations Without Ground-Truth
by: Singh, Amritpal, et al.
Published: (2026)
by: Singh, Amritpal, et al.
Published: (2026)
Prototypical Self-Explainable Models Without Re-training
by: Gautam, Srishti, et al.
Published: (2023)
by: Gautam, Srishti, et al.
Published: (2023)
Impatient Bandits: Optimizing for the Long-Term Without Delay
by: Zhang, Kelly W., et al.
Published: (2025)
by: Zhang, Kelly W., et al.
Published: (2025)
Eliciting Numerical Predictive Distributions of LLMs Without Autoregression
by: Piskorz, Julianna, et al.
Published: (2026)
by: Piskorz, Julianna, et al.
Published: (2026)
Coupled Data and Measurement Space Dynamics for Enhanced Diffusion Posterior Sampling
by: Hamidi, Shayan Mohajer, et al.
Published: (2025)
by: Hamidi, Shayan Mohajer, et al.
Published: (2025)
Equally Critical: Samples, Targets, and Their Mappings in Datasets
by: Yang, Runkang, et al.
Published: (2025)
by: Yang, Runkang, et al.
Published: (2025)
Similar Items
-
Wormhole Memory: A Rubik's Cube for Cross-Dialogue Retrieval
by: Wang, Libo
Published: (2025) -
Solving a Rubik's Cube Using its Local Graph Structure
by: Yao, Shunyu, et al.
Published: (2024) -
Fairness Without Harm: An Influence-Guided Active Sampling Approach
by: Pang, Jinlong, et al.
Published: (2024) -
Solving Diffusion Inverse Problems with Restart Posterior Sampling
by: Ahmed, Bilal, et al.
Published: (2025) -
Universality in Collective Intelligence on the Rubik's Cube
by: Krakauer, David, et al.
Published: (2025)