Parallelized Spatiotemporal Binding
Fuente:
arXiv
Guardado en:
| Autores principales: | Singh, Gautam, Wang, Yue, Yang, Jiawei, Ivanovic, Boris, Ahn, Sungjin, Pavone, Marco, Che, Tong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation Models
por: Lu, Ziqi, et al.
Publicado: (2024)
por: Lu, Ziqi, et al.
Publicado: (2024)
Producing and Leveraging Online Map Uncertainty in Trajectory Prediction
por: Gu, Xunjiang, et al.
Publicado: (2024)
por: Gu, Xunjiang, et al.
Publicado: (2024)
Accelerating Online Mapping and Behavior Prediction via Direct BEV Feature Attention
por: Gu, Xunjiang, et al.
Publicado: (2024)
por: Gu, Xunjiang, et al.
Publicado: (2024)
Efficient Multi-Camera Tokenization with Triplanes for End-to-End Driving
por: Ivanovic, Boris, et al.
Publicado: (2025)
por: Ivanovic, Boris, et al.
Publicado: (2025)
Dreamweaver: Learning Compositional World Models from Pixels
por: Baek, Junyeob, et al.
Publicado: (2025)
por: Baek, Junyeob, et al.
Publicado: (2025)
STORM: Spatio-Temporal Reconstruction Model for Large-Scale Outdoor Scenes
por: Yang, Jiawei, et al.
Publicado: (2024)
por: Yang, Jiawei, et al.
Publicado: (2024)
Neural Language of Thought Models
por: Wu, Yi-Fu, et al.
Publicado: (2024)
por: Wu, Yi-Fu, et al.
Publicado: (2024)
Learning to Compose: Improving Object Centric Learning by Injecting Compositionality
por: Jung, Whie, et al.
Publicado: (2024)
por: Jung, Whie, et al.
Publicado: (2024)
Distributed NeRF Learning for Collaborative Multi-Robot Perception
por: Zhao, Hongrui, et al.
Publicado: (2024)
por: Zhao, Hongrui, et al.
Publicado: (2024)
Data Scaling Laws for End-to-End Autonomous Driving
por: Naumann, Alexander, et al.
Publicado: (2025)
por: Naumann, Alexander, et al.
Publicado: (2025)
Extrapolated Urban View Synthesis Benchmark
por: Han, Xiangyu, et al.
Publicado: (2024)
por: Han, Xiangyu, et al.
Publicado: (2024)
RoaD: Rollouts as Demonstrations for Closed-Loop Supervised Fine-Tuning of Autonomous Driving Policies
por: Garcia-Cobo, Guillermo, et al.
Publicado: (2025)
por: Garcia-Cobo, Guillermo, et al.
Publicado: (2025)
Language-Image Models with 3D Understanding
por: Cho, Jang Hyun, et al.
Publicado: (2024)
por: Cho, Jang Hyun, et al.
Publicado: (2024)
An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning
por: Yoon, Jaesik, et al.
Publicado: (2023)
por: Yoon, Jaesik, et al.
Publicado: (2023)
Multimodal Transformer With a Low-Computational-Cost Guarantee
por: Park, Sungjin, et al.
Publicado: (2024)
por: Park, Sungjin, et al.
Publicado: (2024)
Towards Efficient and Effective Multi-Camera Encoding for End-to-End Driving
por: Yang, Jiawei, et al.
Publicado: (2025)
por: Yang, Jiawei, et al.
Publicado: (2025)
NAVSIM: Data-Driven Non-Reactive Autonomous Vehicle Simulation and Benchmarking
por: Dauner, Daniel, et al.
Publicado: (2024)
por: Dauner, Daniel, et al.
Publicado: (2024)
Scan, Materialize, Simulate: A Generalizable Framework for Physically Grounded Robot Planning
por: Elhafsi, Amine, et al.
Publicado: (2025)
por: Elhafsi, Amine, et al.
Publicado: (2025)
Pseudo-Simulation for Autonomous Driving
por: Cao, Wei, et al.
Publicado: (2025)
por: Cao, Wei, et al.
Publicado: (2025)
LOTUS: A Leaderboard for Detailed Image Captioning from Quality to Societal Bias and User Preferences
por: Hirota, Yusuke, et al.
Publicado: (2025)
por: Hirota, Yusuke, et al.
Publicado: (2025)
Bisecle: Binding and Separation in Continual Learning for Video Language Understanding
por: Tan, Yue, et al.
Publicado: (2025)
por: Tan, Yue, et al.
Publicado: (2025)
Incremental Multi-Scene Modeling via Continual Neural Graphics Primitives
por: Singh, Prajwal, et al.
Publicado: (2024)
por: Singh, Prajwal, et al.
Publicado: (2024)
Learning Traffic Crashes as Language: Datasets, Benchmarks, and What-if Causal Analyses
por: Fan, Zhiwen, et al.
Publicado: (2024)
por: Fan, Zhiwen, et al.
Publicado: (2024)
Cocoon: Robust Multi-Modal Perception with Uncertainty-Aware Sensor Fusion
por: Cho, Minkyoung, et al.
Publicado: (2024)
por: Cho, Minkyoung, et al.
Publicado: (2024)
Vision Foundation Model Embedding-Based Semantic Anomaly Detection
por: Ronecker, Max Peter, et al.
Publicado: (2025)
por: Ronecker, Max Peter, et al.
Publicado: (2025)
Generative Spatiotemporal Data Augmentation
por: Zhou, Jinfan, et al.
Publicado: (2025)
por: Zhou, Jinfan, et al.
Publicado: (2025)
Concept Pinpoint Eraser for Text-to-image Diffusion Models via Residual Attention Gate
por: Lee, Byung Hyun, et al.
Publicado: (2025)
por: Lee, Byung Hyun, et al.
Publicado: (2025)
Spatiotemporal Attention Learning Framework for Event-Driven Object Recognition
por: Xie, Tiantian, et al.
Publicado: (2025)
por: Xie, Tiantian, et al.
Publicado: (2025)
DeepDamageNet: A two-step deep-learning model for multi-disaster building damage segmentation and classification using satellite imagery
por: Alisjahbana, Irene, et al.
Publicado: (2024)
por: Alisjahbana, Irene, et al.
Publicado: (2024)
OmniRe: Omni Urban Scene Reconstruction
por: Chen, Ziyu, et al.
Publicado: (2024)
por: Chen, Ziyu, et al.
Publicado: (2024)
Spatiotemporal Satellite Image Downscaling with Transfer Encoders and Autoregressive Generative Models
por: Xiang, Yang, et al.
Publicado: (2025)
por: Xiang, Yang, et al.
Publicado: (2025)
Physics-Driven Spatiotemporal Modeling for AI-Generated Video Detection
por: Zhang, Shuhai, et al.
Publicado: (2025)
por: Zhang, Shuhai, et al.
Publicado: (2025)
Representation Learning for Spatiotemporal Physical Systems
por: Qu, Helen, et al.
Publicado: (2026)
por: Qu, Helen, et al.
Publicado: (2026)
DCL-SE: Dynamic Curriculum Learning for Spatiotemporal Encoding of Brain Imaging
por: Zhou, Meihua, et al.
Publicado: (2025)
por: Zhou, Meihua, et al.
Publicado: (2025)
Multi-modal Data Binding for Survival Analysis Modeling with Incomplete Data and Annotations
por: Qu, Linhao, et al.
Publicado: (2024)
por: Qu, Linhao, et al.
Publicado: (2024)
SEE-DPO: Self Entropy Enhanced Direct Preference Optimization
por: Shekhar, Shivanshu, et al.
Publicado: (2024)
por: Shekhar, Shivanshu, et al.
Publicado: (2024)
UCTB: An Urban Computing Tool Box for Building Spatiotemporal Prediction Services
por: Fang, Jiangyi, et al.
Publicado: (2023)
por: Fang, Jiangyi, et al.
Publicado: (2023)
Wolf: Dense Video Captioning with a World Summarization Framework
por: Li, Boyi, et al.
Publicado: (2024)
por: Li, Boyi, et al.
Publicado: (2024)
Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression
por: Sarkar, Sreetama, et al.
Publicado: (2025)
por: Sarkar, Sreetama, et al.
Publicado: (2025)
SemBind: Binding Diffusion Watermarks to Semantics Against Black-Box Forgery Attacks
por: Zhang, Xin, et al.
Publicado: (2026)
por: Zhang, Xin, et al.
Publicado: (2026)
Ejemplares similares
-
LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation Models
por: Lu, Ziqi, et al.
Publicado: (2024) -
Producing and Leveraging Online Map Uncertainty in Trajectory Prediction
por: Gu, Xunjiang, et al.
Publicado: (2024) -
Accelerating Online Mapping and Behavior Prediction via Direct BEV Feature Attention
por: Gu, Xunjiang, et al.
Publicado: (2024) -
Efficient Multi-Camera Tokenization with Triplanes for End-to-End Driving
por: Ivanovic, Boris, et al.
Publicado: (2025) -
Dreamweaver: Learning Compositional World Models from Pixels
por: Baek, Junyeob, et al.
Publicado: (2025)