Learning the RoPEs: Better 2D and 3D Position Encodings with STRING
Fuente:
arXiv
Saved in:
| Main Authors: | Schenck, Connor, Reid, Isaac, Jacob, Mithun George, Bewley, Alex, Ainslie, Joshua, Rendleman, David, Jain, Deepali, Sharma, Mohit, Dubey, Avinava, Wahid, Ayzaan, Singh, Sumeet, Wagner, René, Ding, Tianli, Fu, Chuyuan, Byravan, Arunkumar, Varley, Jake, Gritsenko, Alexey, Minderer, Matthias, Kalashnikov, Dmitry, Tompson, Jonathan, Sindhwani, Vikas, Choromanski, Krzysztof |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Embodied AI with Two Arms: Zero-shot Learning, Safety and Modularity
by: Varley, Jake, et al.
Published: (2024)
by: Varley, Jake, et al.
Published: (2024)
Modeling the Real World with High-Density Visual Particle Dynamics
by: Whitney, William F., et al.
Published: (2024)
by: Whitney, William F., et al.
Published: (2024)
Linear Transformer Topological Masking with Graph Random Features
by: Reid, Isaac, et al.
Published: (2024)
by: Reid, Isaac, et al.
Published: (2024)
Structured Unrestricted-Rank Matrices for Parameter Efficient Fine-tuning
by: Sehanobish, Arijit, et al.
Published: (2024)
by: Sehanobish, Arijit, et al.
Published: (2024)
Scaling Open-Vocabulary Object Detection
by: Minderer, Matthias, et al.
Published: (2023)
by: Minderer, Matthias, et al.
Published: (2023)
Self-Improving Embodied Foundation Models
by: Ghasemipour, Seyed Kamyar Seyed, et al.
Published: (2025)
by: Ghasemipour, Seyed Kamyar Seyed, et al.
Published: (2025)
Predictive Red Teaming: Breaking Policies Without Breaking Robots
by: Majumdar, Anirudha, et al.
Published: (2025)
by: Majumdar, Anirudha, et al.
Published: (2025)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
by: Salzmann, Tim, et al.
Published: (2024)
by: Salzmann, Tim, et al.
Published: (2024)
ALOHA Unleashed: A Simple Recipe for Robot Dexterity
by: Zhao, Tony Z., et al.
Published: (2024)
by: Zhao, Tony Z., et al.
Published: (2024)
RelFlexformer: Efficient Attention 3D-Transformers for Integrable Relative Positional Encodings
by: Kim, Byeongchan, et al.
Published: (2026)
by: Kim, Byeongchan, et al.
Published: (2026)
CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining
by: Liu, I-Chun Arthur, et al.
Published: (2026)
by: Liu, I-Chun Arthur, et al.
Published: (2026)
Optimal Time Complexity Algorithms for Computing General Random Walk Graph Kernels on Sparse Graphs
by: Choromanski, Krzysztof, et al.
Published: (2024)
by: Choromanski, Krzysztof, et al.
Published: (2024)
Computationally-efficient Graph Modeling with Refined Graph Random Features
by: Choromanski, Krzysztof, et al.
Published: (2025)
by: Choromanski, Krzysztof, et al.
Published: (2025)
Diffusion Augmented Agents: A Framework for Efficient Exploration and Transfer Learning
by: Di Palo, Norman, et al.
Published: (2024)
by: Di Palo, Norman, et al.
Published: (2024)
Scalable Neural Network Kernels
by: Sehanobish, Arijit, et al.
Published: (2023)
by: Sehanobish, Arijit, et al.
Published: (2023)
SWING: Unlocking Implicit Graph Representations for Graph Random Features
by: Manenti, Alessandro, et al.
Published: (2026)
by: Manenti, Alessandro, et al.
Published: (2026)
Robot Data Curation with Mutual Information Estimators
by: Hejna, Joey, et al.
Published: (2025)
by: Hejna, Joey, et al.
Published: (2025)
THE OPERANT CONDITIONING OF LETTER STRING PROBLEM SOLVING
by: Marco A. Pulido
Published: (2010)
by: Marco A. Pulido
Published: (2010)
Generating Robot Constitutions & Benchmarks for Semantic Safety
by: Sermanet, Pierre, et al.
Published: (2025)
by: Sermanet, Pierre, et al.
Published: (2025)
Towards Scalable Exact Machine Unlearning Using Parameter-Efficient Fine-Tuning
by: Chowdhury, Somnath Basu Roy, et al.
Published: (2024)
by: Chowdhury, Somnath Basu Roy, et al.
Published: (2024)
Vision Language Models are In-Context Value Learners
by: Ma, Yecheng Jason, et al.
Published: (2024)
by: Ma, Yecheng Jason, et al.
Published: (2024)
Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs
by: Chiang, Hao-Tien Lewis, et al.
Published: (2024)
by: Chiang, Hao-Tien Lewis, et al.
Published: (2024)
Splatting Physical Scenes: End-to-End Real-to-Sim from Imperfect Robot Data
by: Moran, Ben, et al.
Published: (2025)
by: Moran, Ben, et al.
Published: (2025)
Modular differential equations of minimal orders of the elliptic genus of Calabi--Yau varieties
by: Adler, Dmitrii, et al.
Published: (2025)
by: Adler, Dmitrii, et al.
Published: (2025)
Fast Tree-Field Integrators: From Low Displacement Rank to Topological Transformers
by: Choromanski, Krzysztof, et al.
Published: (2024)
by: Choromanski, Krzysztof, et al.
Published: (2024)
ViTaPEs: Visuotactile Position Encodings for Cross-Modal Alignment in Multimodal Transformers
by: Lygerakis, Fotios, et al.
Published: (2025)
by: Lygerakis, Fotios, et al.
Published: (2025)
G$^{2}$D: Boosting Multimodal Learning with Gradient-Guided Distillation
by: Rakib, Mohammed, et al.
Published: (2025)
by: Rakib, Mohammed, et al.
Published: (2025)
kylieainslie/mitey: mitey 0.3.1
by: Kylie Ainslie
Published: (2026)
by: Kylie Ainslie
Published: (2026)
India : historia del subcontinente desde las culturas del indo hasta el comienzo del dominio inglés / Ainslie T. Embree y Friedrich Wilhelm; traductores Antón Dieterich, María Isabel Carrillo
by: Embree, Ainslie
by: Embree, Ainslie
Inadequate housing is not neglect: How the family regulation system punishes parents for a housing crisis out of their control
by: Ainslie Martin
Published: (2025)
by: Ainslie Martin
Published: (2025)
Communicating Competencies and Collaboration.
by: Zipperer, Lorri, et al.
Published: (2002)
by: Zipperer, Lorri, et al.
Published: (2002)
A toric degeneration of Kronecker moduli spaces
by: Kalashnikov, Elana
Published: (2024)
by: Kalashnikov, Elana
Published: (2024)
Laurent polynomial mirrors for quiver flag zero loci
by: Kalashnikov, Elana
Published: (2019)
by: Kalashnikov, Elana
Published: (2019)
Structure of Filled Functions: Why Gaussian and Cauchy Templates Are Most Efficient
by: Vyacheslav Kalashnikov
Published: (2016)
by: Vyacheslav Kalashnikov
Published: (2016)
Un modelo de migración humana: experimentos numéricos basados sobre los datos de las tres ciudades laguneras
by: Vyacheslav Kalashnikov
Published: (2007)
by: Vyacheslav Kalashnikov
Published: (2007)
Commentary on Epidemiology of mental health comorbidity in patients with atopic dermatitis: An analysis of global trends from 1998 to 2022
by: Anthony Bewley
Published: (2024)
by: Anthony Bewley
Published: (2024)
Magnetic Tape Care, Storage, and Error Recovery.
by: Schenck, Thomas
Published: (1984)
by: Schenck, Thomas
Published: (1984)
Sobre el poder del conocimiento, o el conocimiento como poder: Reflexiones sobre la política de la Ciencia Política
by: Marcela Schenck
Published: (2019)
by: Marcela Schenck
Published: (2019)
Incorporación de la diversidad genérico-sexual en salud: claves teóricas para un modelo analítico
by: Marcela Schenck
Published: (2018)
by: Marcela Schenck
Published: (2018)
TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment
by: Cao, Bingyi, et al.
Published: (2026)
by: Cao, Bingyi, et al.
Published: (2026)
Similar Items
-
Embodied AI with Two Arms: Zero-shot Learning, Safety and Modularity
by: Varley, Jake, et al.
Published: (2024) -
Modeling the Real World with High-Density Visual Particle Dynamics
by: Whitney, William F., et al.
Published: (2024) -
Linear Transformer Topological Masking with Graph Random Features
by: Reid, Isaac, et al.
Published: (2024) -
Structured Unrestricted-Rank Matrices for Parameter Efficient Fine-tuning
by: Sehanobish, Arijit, et al.
Published: (2024) -
Scaling Open-Vocabulary Object Detection
by: Minderer, Matthias, et al.
Published: (2023)