Making VLMs More Robot-Friendly: Self-Critical Distillation of Low-Level Procedural Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Park, Chan Young, Fisher, Jillian, Memmel, Marius, Khullar, Dipika, Yun, Seoho, Gupta, Abhishek, Choi, Yejin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
STRAP: Robot Sub-Trajectory Retrieval for Augmented Policy Learning
by: Memmel, Marius, et al.
Published: (2024)
by: Memmel, Marius, et al.
Published: (2024)
ASID: Active Exploration for System Identification in Robotic Manipulation
by: Memmel, Marius, et al.
Published: (2024)
by: Memmel, Marius, et al.
Published: (2024)
Modular Pluralism: Pluralistic Alignment via Multi-LLM Collaboration
by: Feng, Shangbin, et al.
Published: (2024)
by: Feng, Shangbin, et al.
Published: (2024)
Response-Aware User Memory Selection for LLM Personalization
by: Fisher, Jillian, et al.
Published: (2026)
by: Fisher, Jillian, et al.
Published: (2026)
StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
by: Fisher, Jillian, et al.
Published: (2024)
by: Fisher, Jillian, et al.
Published: (2024)
Improved Few-Shot Image Classification Through Multiple-Choice Questions
by: Khullar, Dipika, et al.
Published: (2024)
by: Khullar, Dipika, et al.
Published: (2024)
FDARxBench: Benchmarking Regulatory and Clinical Reasoning on FDA Generic Drug Assessment
by: Xiong, Betty, et al.
Published: (2026)
by: Xiong, Betty, et al.
Published: (2026)
Biased AI can Influence Political Decision-Making
by: Fisher, Jillian, et al.
Published: (2024)
by: Fisher, Jillian, et al.
Published: (2024)
Self-Attribution Bias: When AI Monitors Go Easy on Themselves
by: Khullar, Dipika, et al.
Published: (2026)
by: Khullar, Dipika, et al.
Published: (2026)
Spectrum Tuning: Post-Training for Distributional Coverage and In-Context Steerability
by: Sorensen, Taylor, et al.
Published: (2025)
by: Sorensen, Taylor, et al.
Published: (2025)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
by: Zhang, Jesse, et al.
Published: (2025)
by: Zhang, Jesse, et al.
Published: (2025)
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing
by: Jung, Jaehun, et al.
Published: (2023)
by: Jung, Jaehun, et al.
Published: (2023)
Generalizing from SIMPLE to HARD Visual Reasoning: Can We Mitigate Modality Imbalance in VLMs?
by: Park, Simon, et al.
Published: (2025)
by: Park, Simon, et al.
Published: (2025)
Visual CoT Makes VLMs Smarter but More Fragile
by: Xu, Chunxue, et al.
Published: (2025)
by: Xu, Chunxue, et al.
Published: (2025)
Configurable 3D‐Printed Microstructured Stamp as a User‐Friendly Tool for Versatile Patterning of Low‐Viscosity Bioinks
by: Yejin Choi, et al.
Published: (2025)
by: Yejin Choi, et al.
Published: (2025)
JAMDEC: Unsupervised Authorship Obfuscation using Constrained Decoding over Small Language Models
by: Fisher, Jillian, et al.
Published: (2024)
by: Fisher, Jillian, et al.
Published: (2024)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
by: Li, Yi, et al.
Published: (2025)
by: Li, Yi, et al.
Published: (2025)
URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images
by: Chen, Zoey, et al.
Published: (2024)
by: Chen, Zoey, et al.
Published: (2024)
Thinking-Free Policy Initialization Makes Distilled Reasoning Models More Effective and Efficient Reasoners
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
VLM-SubtleBench: How Far Are VLMs from Human-Level Subtle Comparative Reasoning?
by: Kim, Minkyu, et al.
Published: (2026)
by: Kim, Minkyu, et al.
Published: (2026)
Automatic Structure Identification for Highly Nonlinear MIMO Volterra Tensor Networks
by: Memmel, Eva, et al.
Published: (2025)
by: Memmel, Eva, et al.
Published: (2025)
Product Line Profitability and Margin Performance Analysis on Nassau Candy Distributor
by: Yadav, Dipika
Published: (2026)
by: Yadav, Dipika
Published: (2026)
Legal challenges in expanding the provider base for abortion in Asia
by: Dipika Jain
Published: (2024)
by: Dipika Jain
Published: (2024)
Think Twice to See More: Iterative Visual Reasoning in Medical VLMs
by: Chen, Kaitao, et al.
Published: (2025)
by: Chen, Kaitao, et al.
Published: (2025)
Political Neutrality in AI Is Impossible- But Here Is How to Approximate It
by: Fisher, Jillian, et al.
Published: (2025)
by: Fisher, Jillian, et al.
Published: (2025)
MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes
by: Chiu, Yu Ying, et al.
Published: (2025)
by: Chiu, Yu Ying, et al.
Published: (2025)
Relevance to Utility: Process-Supervised Rewrite for RAG
by: Kim, Jaeyoung, et al.
Published: (2025)
by: Kim, Jaeyoung, et al.
Published: (2025)
Co-Evolving Agents: Learning from Failures as Hard Negatives
by: Jung, Yeonsung, et al.
Published: (2025)
by: Jung, Yeonsung, et al.
Published: (2025)
LongPerceptualThoughts: Distilling System-2 Reasoning for System-1 Perception
by: Liao, Yuan-Hong, et al.
Published: (2025)
by: Liao, Yuan-Hong, et al.
Published: (2025)
PlaSma: Making Small Language Models Better Procedural Knowledge Models for (Counterfactual) Planning
by: Brahman, Faeze, et al.
Published: (2023)
by: Brahman, Faeze, et al.
Published: (2023)
Similarity-Guided Diffusion for Contrastive Sequential Recommendation
by: Choi, Jinkyeong, et al.
Published: (2025)
by: Choi, Jinkyeong, et al.
Published: (2025)
DRAWER: Digital Reconstruction and Articulation With Environment Realism
by: Xia, Hongchi, et al.
Published: (2025)
by: Xia, Hongchi, et al.
Published: (2025)
World2VLM: Distilling World Model Imagination into VLMs for Dynamic Spatial Reasoning
by: Zhang, Wanyue, et al.
Published: (2026)
by: Zhang, Wanyue, et al.
Published: (2026)
Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
by: Ravichander, Abhilasha, et al.
Published: (2025)
by: Ravichander, Abhilasha, et al.
Published: (2025)
More User-Friendly Metrics
by: Erwin Krauskopf
Published: (2019)
by: Erwin Krauskopf
Published: (2019)
Subtle Risks, Critical Failures: A Framework for Diagnosing Physical Safety of LLMs for Embodied Decision Making
by: Son, Yejin, et al.
Published: (2025)
by: Son, Yejin, et al.
Published: (2025)
Tailoring Self-Rationalizers with Multi-Reward Distillation
by: Ramnath, Sahana, et al.
Published: (2023)
by: Ramnath, Sahana, et al.
Published: (2023)
Adaptive Travel Behaviors During Crisis Waves: A Co‐Occurrence Network of Social Media Discourse
by: Yejin Lee, et al.
Published: (2026)
by: Yejin Lee, et al.
Published: (2026)
Making Friends with Print.
by: Robinson, Dorothy W.
Published: (1982)
by: Robinson, Dorothy W.
Published: (1982)
Are VLMs Really Blind
by: Singh, Ayush, et al.
Published: (2024)
by: Singh, Ayush, et al.
Published: (2024)
Similar Items
-
STRAP: Robot Sub-Trajectory Retrieval for Augmented Policy Learning
by: Memmel, Marius, et al.
Published: (2024) -
ASID: Active Exploration for System Identification in Robotic Manipulation
by: Memmel, Marius, et al.
Published: (2024) -
Modular Pluralism: Pluralistic Alignment via Multi-LLM Collaboration
by: Feng, Shangbin, et al.
Published: (2024) -
Response-Aware User Memory Selection for LLM Personalization
by: Fisher, Jillian, et al.
Published: (2026) -
StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
by: Fisher, Jillian, et al.
Published: (2024)