Saved in:
| Main Authors: | Odonchimed, Sodtavilan, Matsushima, Tatsuya, Holk, Simon, Iwasawa, Yusuke, Matsuo, Yutaka |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.21452 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GenORM: Generalizable One-shot Rope Manipulation with Parameter-Aware Policy
by: Kuroki, So, et al.
Published: (2023)
by: Kuroki, So, et al.
Published: (2023)
SPARK: Graph-Based Online Semantic Integration System for Robot Task Planning
by: Shirasaka, Mimo, et al.
Published: (2025)
by: Shirasaka, Mimo, et al.
Published: (2025)
A Comprehensive Survey on Physical Risk Control in the Era of Foundation Model-enabled Robotics
by: Kojima, Takeshi, et al.
Published: (2025)
by: Kojima, Takeshi, et al.
Published: (2025)
Leave No Observation Behind: Real-time Correction for VLA Action Chunks
by: Sendai, Kohei, et al.
Published: (2025)
by: Sendai, Kohei, et al.
Published: (2025)
Conditional Diffusion Models for Semantic 3D Brain MRI Synthesis
by: Dorjsembe, Zolnamar, et al.
Published: (2023)
by: Dorjsembe, Zolnamar, et al.
Published: (2023)
Collective Intelligence for 2D Push Manipulations with Mobile Robots
by: Kuroki, So, et al.
Published: (2022)
by: Kuroki, So, et al.
Published: (2022)
PREDILECT: Preferences Delineated with Zero-Shot Language-based Reasoning in Reinforcement Learning
by: Holk, Simon, et al.
Published: (2024)
by: Holk, Simon, et al.
Published: (2024)
Bridging Lottery Ticket and Grokking: Understanding Grokking from Inner Structure of Networks
by: Minegishi, Gouki, et al.
Published: (2023)
by: Minegishi, Gouki, et al.
Published: (2023)
GenDOM: Generalizable One-shot Deformable Object Manipulation with Parameter-Aware Policy
by: Kuroki, So, et al.
Published: (2023)
by: Kuroki, So, et al.
Published: (2023)
Does "Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients?
by: Onoda, Ku, et al.
Published: (2026)
by: Onoda, Ku, et al.
Published: (2026)
Language Models Do Hard Arithmetic Tasks Easily and Hardly Do Easy Arithmetic Tasks
by: Gambardella, Andrew, et al.
Published: (2024)
by: Gambardella, Andrew, et al.
Published: (2024)
Continuous Reasoning for Vision-Language-Action
by: Wu, Yueh-Hua, et al.
Published: (2026)
by: Wu, Yueh-Hua, et al.
Published: (2026)
Zipping the Thought: When and How Compressed Reasoning Data Works in LLM Post-Training
by: Matsutani, Kohsei, et al.
Published: (2026)
by: Matsutani, Kohsei, et al.
Published: (2026)
Towards Empirical Interpretation of Internal Circuits and Properties in Grokked Transformers on Modular Polynomials
by: Furuta, Hiroki, et al.
Published: (2024)
by: Furuta, Hiroki, et al.
Published: (2024)
Residual Koopman Spectral Profiling for Predicting and Preventing Transformer Training Instability
by: Kim, Bum Jun, et al.
Published: (2026)
by: Kim, Bum Jun, et al.
Published: (2026)
C-voting: Confidence-Based Test-Time Voting without Explicit Energy Functions
by: Kubo, Kenji, et al.
Published: (2026)
by: Kubo, Kenji, et al.
Published: (2026)
$\infty$-MoE: Generalizing Mixture of Experts to Infinite Experts
by: Takashiro, Shota, et al.
Published: (2026)
by: Takashiro, Shota, et al.
Published: (2026)
Inconsistent Tokenizations Cause Language Models to be Perplexed by Japanese Grammar
by: Gambardella, Andrew, et al.
Published: (2025)
by: Gambardella, Andrew, et al.
Published: (2025)
Rethinking Evaluation of Sparse Autoencoders through the Representation of Polysemous Words
by: Minegishi, Gouki, et al.
Published: (2025)
by: Minegishi, Gouki, et al.
Published: (2025)
R2-Dreamer: Redundancy-Reduced World Models without Decoders or Augmentation
by: Morihira, Naoki, et al.
Published: (2026)
by: Morihira, Naoki, et al.
Published: (2026)
FLoRA: Sample-Efficient Preference-based RL via Low-Rank Style Adaptation of Reward Functions
by: Marta, Daniel, et al.
Published: (2025)
by: Marta, Daniel, et al.
Published: (2025)
Unlocking Noise-Resistant Vision: Key Architectural Secrets for Robust Models
by: Kim, Bum Jun, et al.
Published: (2025)
by: Kim, Bum Jun, et al.
Published: (2025)
Understanding Emergent Misalignment via Feature Superposition Geometry
by: Minegishi, Gouki, et al.
Published: (2026)
by: Minegishi, Gouki, et al.
Published: (2026)
QuadNorm: Resolution-Robust Normalization for Neural Operators
by: Kim, Bum Jun, et al.
Published: (2026)
by: Kim, Bum Jun, et al.
Published: (2026)
Robustifying a Policy in Multi-Agent RL with Diverse Cooperative Behaviors and Adversarial Style Sampling for Assistive Tasks
by: Osa, Takayuki, et al.
Published: (2024)
by: Osa, Takayuki, et al.
Published: (2024)
Real-Time Operator Takeover for Visuomotor Diffusion Policy Training
by: Moletta, Marco, et al.
Published: (2025)
by: Moletta, Marco, et al.
Published: (2025)
Training-Free Imitation Learning with Closed-Form Diffusion Policies
by: Mishra, Raghav, et al.
Published: (2026)
by: Mishra, Raghav, et al.
Published: (2026)
COLLAGE: Adaptive Fusion-based Retrieval for Augmented Policy Learning
by: Kumar, Sateesh, et al.
Published: (2025)
by: Kumar, Sateesh, et al.
Published: (2025)
STRAP: Robot Sub-Trajectory Retrieval for Augmented Policy Learning
by: Memmel, Marius, et al.
Published: (2024)
by: Memmel, Marius, et al.
Published: (2024)
WorldPack: Compressed Memory Improves Spatial Consistency in Video World Modeling
by: Oshima, Yuta, et al.
Published: (2025)
by: Oshima, Yuta, et al.
Published: (2025)
Semantic Token Clustering for Efficient Uncertainty Quantification in Large Language Models
by: Cao, Qi, et al.
Published: (2026)
by: Cao, Qi, et al.
Published: (2026)
Speeding up 6-DoF Grasp Sampling with Quality-Diversity
by: Huber, Johann, et al.
Published: (2024)
by: Huber, Johann, et al.
Published: (2024)
Mini Diffuser: Fast Multi-task Diffusion Policy Training Using Two-level Mini-batches
by: Hu, Yutong, et al.
Published: (2025)
by: Hu, Yutong, et al.
Published: (2025)
The Role of Domain Randomization in Training Diffusion Policies for Whole-Body Humanoid Control
by: Kaidanov, Oleg, et al.
Published: (2024)
by: Kaidanov, Oleg, et al.
Published: (2024)
Diffusion Policy Policy Optimization
by: Ren, Allen Z., et al.
Published: (2024)
by: Ren, Allen Z., et al.
Published: (2024)
Variable-Speed Teaching-Playback as Real-World Data Augmentation for Imitation Learning
by: Masuya, Nozomu, et al.
Published: (2024)
by: Masuya, Nozomu, et al.
Published: (2024)
Rethinking Policy Diversity in Ensemble Policy Gradient in Large-Scale Reinforcement Learning
by: Shitanda, Naoki, et al.
Published: (2026)
by: Shitanda, Naoki, et al.
Published: (2026)
Thinking While Listening: Fast-Slow Recurrence for Long-Horizon Sequential Modeling
by: Takashiro, Shota, et al.
Published: (2026)
by: Takashiro, Shota, et al.
Published: (2026)
Safe Transformer: An Explicit Safety Bit For Interpretable And Controllable Alignment
by: Feng, Jingyuan, et al.
Published: (2026)
by: Feng, Jingyuan, et al.
Published: (2026)
Real-World Robot Applications of Foundation Models: A Review
by: Kawaharazuka, Kento, et al.
Published: (2024)
by: Kawaharazuka, Kento, et al.
Published: (2024)
Similar Items
-
GenORM: Generalizable One-shot Rope Manipulation with Parameter-Aware Policy
by: Kuroki, So, et al.
Published: (2023) -
SPARK: Graph-Based Online Semantic Integration System for Robot Task Planning
by: Shirasaka, Mimo, et al.
Published: (2025) -
A Comprehensive Survey on Physical Risk Control in the Era of Foundation Model-enabled Robotics
by: Kojima, Takeshi, et al.
Published: (2025) -
Leave No Observation Behind: Real-time Correction for VLA Action Chunks
by: Sendai, Kohei, et al.
Published: (2025) -
Conditional Diffusion Models for Semantic 3D Brain MRI Synthesis
by: Dorjsembe, Zolnamar, et al.
Published: (2023)