C^2ROPE: Causal Continuous Rotary Positional Encoding for 3D Large Multimodal-Models Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Guanting, Zhao, Qiyan, Yu, Wenhao, Zhang, Xiaofeng, Ji, Jianmin, Zhang, Yanyong, Yuen, Ka-Veng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SoPE: Spherical Coordinate-Based Positional Embedding for Enhancing Spatial Perception of 3D LVLMs
von: Ye, Guanting, et al.
Veröffentlicht: (2026)
von: Ye, Guanting, et al.
Veröffentlicht: (2026)
Bayesian Online Joint Identification on Axle Positions, Axle Loads, and Bridge Structural Parameters Using Evolving Virtual Axle Configurations Without Axle Detector
von: Hou-Zuo Guo, et al.
Veröffentlicht: (2026)
von: Hou-Zuo Guo, et al.
Veröffentlicht: (2026)
Transmissibility Physics‐Guided Deep Learning Network for Seismic Response Prediction Under Limited Domain Knowledge, Sparse Measurements, and Unknown Excitations
von: Yuntai Zhang, et al.
Veröffentlicht: (2026)
von: Yuntai Zhang, et al.
Veröffentlicht: (2026)
Rotary Position Encodings for Graphs
von: Reid, Isaac, et al.
Veröffentlicht: (2025)
von: Reid, Isaac, et al.
Veröffentlicht: (2025)
MHRC: Closed-loop Decentralized Multi-Heterogeneous Robot Collaboration with Large Language Models
von: Yu, Wenhao, et al.
Veröffentlicht: (2024)
von: Yu, Wenhao, et al.
Veröffentlicht: (2024)
3D-RPE: Enhancing Long-Context Modeling Through 3D Rotary Position Encoding
von: Ma, Xindian, et al.
Veröffentlicht: (2024)
von: Ma, Xindian, et al.
Veröffentlicht: (2024)
OA-DET3D: Embedding Object Awareness as a General Plug-in for Multi-Camera 3D Object Detection
von: Chu, Xiaomeng, et al.
Veröffentlicht: (2023)
von: Chu, Xiaomeng, et al.
Veröffentlicht: (2023)
LDP: A Local Diffusion Planner for Efficient Robot Navigation and Collision Avoidance
von: Yu, Wenhao, et al.
Veröffentlicht: (2024)
von: Yu, Wenhao, et al.
Veröffentlicht: (2024)
CLMASP: Coupling Large Language Models with Answer Set Programming for Robotic Task Planning
von: Lin, Xinrui, et al.
Veröffentlicht: (2024)
von: Lin, Xinrui, et al.
Veröffentlicht: (2024)
Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Models
von: Zhang, Chenyu, et al.
Veröffentlicht: (2026)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2026)
MM-Gaussian: 3D Gaussian-based Multi-modal Fusion for Localization and Reconstruction in Unbounded Scenes
von: Wu, Chenyang, et al.
Veröffentlicht: (2024)
von: Wu, Chenyang, et al.
Veröffentlicht: (2024)
Navigating Uncertainties in Machine Learning for Structural Dynamics: A Comprehensive Survey of Probabilistic and Non-Probabilistic Approaches in Forward and Inverse Problems
von: Yan, Wang-Ji, et al.
Veröffentlicht: (2024)
von: Yan, Wang-Ji, et al.
Veröffentlicht: (2024)
SpatialSplat: Efficient Semantic 3D from Sparse Unposed Images
von: Sheng, Yu, et al.
Veröffentlicht: (2025)
von: Sheng, Yu, et al.
Veröffentlicht: (2025)
STDArm: Transferring Visuomotor Policies From Static Data Training to Dynamic Robot Manipulation
von: Duan, Yifan, et al.
Veröffentlicht: (2025)
von: Duan, Yifan, et al.
Veröffentlicht: (2025)
OCC-VO: Dense Mapping via 3D Occupancy-Based Visual Odometry for Autonomous Driving
von: Li, Heng, et al.
Veröffentlicht: (2023)
von: Li, Heng, et al.
Veröffentlicht: (2023)
AdaToken-3D: Dynamic Spatial Gating for Efficient 3D Large Multimodal-Models Reasoning
von: Zhang, Kai, et al.
Veröffentlicht: (2025)
von: Zhang, Kai, et al.
Veröffentlicht: (2025)
MrRoPE: Mixed-radix Rotary Position Embedding
von: Tian, Qingyuan, et al.
Veröffentlicht: (2026)
von: Tian, Qingyuan, et al.
Veröffentlicht: (2026)
Head-wise Adaptive Rotary Positional Encoding for Fine-Grained Image Generation
von: Li, Jiaye, et al.
Veröffentlicht: (2025)
von: Li, Jiaye, et al.
Veröffentlicht: (2025)
Round and Round We Go! What makes Rotary Positional Encodings useful?
von: Barbero, Federico, et al.
Veröffentlicht: (2024)
von: Barbero, Federico, et al.
Veröffentlicht: (2024)
Progressive Multimodal Reasoning via Active Retrieval
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
Length Generalization of Causal Transformers without Position Encoding
von: Wang, Jie, et al.
Veröffentlicht: (2024)
von: Wang, Jie, et al.
Veröffentlicht: (2024)
MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models
von: Zhao, Qiyan, et al.
Veröffentlicht: (2025)
von: Zhao, Qiyan, et al.
Veröffentlicht: (2025)
HoPE: Hyperbolic Rotary Positional Encoding for Stable Long-Range Dependency Modeling in Large Language Models
von: Dai, Chang, et al.
Veröffentlicht: (2025)
von: Dai, Chang, et al.
Veröffentlicht: (2025)
IT IS ALLOWED TO PROHIBIT: “WOKE”, “DEI” AND THE ROPE OF METAPHOR.
von: Custódio, Sérgio José
Veröffentlicht: (2025)
von: Custódio, Sérgio José
Veröffentlicht: (2025)
Bayesian Learning in Structural Dynamics: A Comprehensive Review and Emerging Trends
von: Yan, Wang-Ji, et al.
Veröffentlicht: (2025)
von: Yan, Wang-Ji, et al.
Veröffentlicht: (2025)
Rendering-Enhanced Automatic Image-to-Point Cloud Registration for Roadside Scenes
von: Sheng, Yu, et al.
Veröffentlicht: (2024)
von: Sheng, Yu, et al.
Veröffentlicht: (2024)
Efficient Matrix Implementation for Rotary Position Embedding
von: Minqi, Chen, et al.
Veröffentlicht: (2026)
von: Minqi, Chen, et al.
Veröffentlicht: (2026)
GraspCoT: Integrating Physical Property Reasoning for 6-DoF Grasping under Flexible Language Instructions
von: Chu, Xiaomeng, et al.
Veröffentlicht: (2025)
von: Chu, Xiaomeng, et al.
Veröffentlicht: (2025)
CAFE-AD: Cross-Scenario Adaptive Feature Enhancement for Trajectory Planning in Autonomous Driving
von: Zhang, Junrui, et al.
Veröffentlicht: (2025)
von: Zhang, Junrui, et al.
Veröffentlicht: (2025)
MT-PCR: Leveraging Modality Transformation for Large-Scale Point Cloud Registration with Limited Overlap
von: Wu, Yilong, et al.
Veröffentlicht: (2025)
von: Wu, Yilong, et al.
Veröffentlicht: (2025)
Selective Rotary Position Embedding
von: Movahedi, Sajad, et al.
Veröffentlicht: (2025)
von: Movahedi, Sajad, et al.
Veröffentlicht: (2025)
ReDirector: Creating Any-Length Video Retakes with Rotary Camera Encoding
von: Park, Byeongjun, et al.
Veröffentlicht: (2025)
von: Park, Byeongjun, et al.
Veröffentlicht: (2025)
Benchmarking Rotary Position Embeddings for Automatic Speech Recognition
von: Zhang, Shucong, et al.
Veröffentlicht: (2025)
von: Zhang, Shucong, et al.
Veröffentlicht: (2025)
On quadratic Novikov algebras
von: Dong, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Dong, Xiaofeng, et al.
Veröffentlicht: (2025)
RomanTex: Decoupling 3D-aware Rotary Positional Embedded Multi-Attention Network for Texture Synthesis
von: Feng, Yifei, et al.
Veröffentlicht: (2025)
von: Feng, Yifei, et al.
Veröffentlicht: (2025)
Cross-Axis Transformer with 3D Rotary Positional Embeddings
von: Erickson, Lily
Veröffentlicht: (2023)
von: Erickson, Lily
Veröffentlicht: (2023)
UrgenGo: Urgency-Aware Transparent GPU Kernel Launching for Autonomous Driving
von: Zhu, Hanqi, et al.
Veröffentlicht: (2025)
von: Zhu, Hanqi, et al.
Veröffentlicht: (2025)
Map++: Towards User-Participatory Visual SLAM Systems with Efficient Map Expansion and Sharing
von: Zhang, Xinran, et al.
Veröffentlicht: (2024)
von: Zhang, Xinran, et al.
Veröffentlicht: (2024)
Rotary Position Embedding for Vision Transformer
von: Heo, Byeongho, et al.
Veröffentlicht: (2024)
von: Heo, Byeongho, et al.
Veröffentlicht: (2024)
Context-aware Rotary Position Embedding
von: Veisi, Ali, et al.
Veröffentlicht: (2025)
von: Veisi, Ali, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SoPE: Spherical Coordinate-Based Positional Embedding for Enhancing Spatial Perception of 3D LVLMs
von: Ye, Guanting, et al.
Veröffentlicht: (2026) -
Bayesian Online Joint Identification on Axle Positions, Axle Loads, and Bridge Structural Parameters Using Evolving Virtual Axle Configurations Without Axle Detector
von: Hou-Zuo Guo, et al.
Veröffentlicht: (2026) -
Transmissibility Physics‐Guided Deep Learning Network for Seismic Response Prediction Under Limited Domain Knowledge, Sparse Measurements, and Unknown Excitations
von: Yuntai Zhang, et al.
Veröffentlicht: (2026) -
Rotary Position Encodings for Graphs
von: Reid, Isaac, et al.
Veröffentlicht: (2025) -
MHRC: Closed-loop Decentralized Multi-Heterogeneous Robot Collaboration with Large Language Models
von: Yu, Wenhao, et al.
Veröffentlicht: (2024)