Motor Focus: Fast Ego-Motion Prediction for Assistive Visual Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Hao, Qin, Jiayou, Chen, Xiwen, Bastola, Ashish, Suchanek, John, Gong, Zihao, Razi, Abolfazl |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VisionGPT: LLM-Assisted Real-Time Anomaly Detection for Safe Visual Navigation
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Diffusion Prism: Enhancing Diversity and Morphology Consistency in Mask-to-Image Diffusion
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
Fast 2DGS: Efficient Image Representation with Deep Gaussian Prior
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
Motion Focus Recognition in Fast-Moving Egocentric Video
by: Hong, Si-En, et al.
Published: (2026)
by: Hong, Si-En, et al.
Published: (2026)
FLAME Diffuser: Wildfire Image Synthesis using Mask Guided Diffusion
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
RobustFormer: Noise-Robust Pre-training for images and videos
by: Bastola, Ashish, et al.
Published: (2024)
by: Bastola, Ashish, et al.
Published: (2024)
RBAD: A Dataset and Benchmark for Retinal Vessels Branching Angle Detection
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Spatial-Conditioned Reasoning in Long-Egocentric Videos
by: Tribble, James, et al.
Published: (2026)
by: Tribble, James, et al.
Published: (2026)
Imaging Signal Recovery Using Neural Network Priors Under Uncertain Forward Model Parameters
by: Chen, Xiwen, et al.
Published: (2024)
by: Chen, Xiwen, et al.
Published: (2024)
FedMIL: Federated-Multiple Instance Learning for Video Analysis with Optimized DPP Scheduling
by: Bastola, Ashish, et al.
Published: (2024)
by: Bastola, Ashish, et al.
Published: (2024)
Enhancing Digital Hologram Reconstruction Using Reverse-Attention Loss for Untrained Physics-Driven Deep Learning Models with Uncertain Distance
by: Chen, Xiwen, et al.
Published: (2024)
by: Chen, Xiwen, et al.
Published: (2024)
DGR-MIL: Exploring Diverse Global Representation in Multiple Instance Learning for Whole Slide Image Classification
by: Zhu, Wenhui, et al.
Published: (2024)
by: Zhu, Wenhui, et al.
Published: (2024)
AtomDiffuser: Time-Aware Degradation Modeling for Drift and Beam Damage in STEM Imaging
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
OTPrune: Distribution-Aligned Visual Token Pruning via Optimal Transport
by: Chen, Xiwen, et al.
Published: (2026)
by: Chen, Xiwen, et al.
Published: (2026)
SelfReg-UNet: Self-Regularized UNet for Medical Image Segmentation
by: Zhu, Wenhui, et al.
Published: (2024)
by: Zhu, Wenhui, et al.
Published: (2024)
EgoAVU: Egocentric Audio-Visual Understanding
by: Seth, Ashish, et al.
Published: (2026)
by: Seth, Ashish, et al.
Published: (2026)
Anomaly Detection in Cooperative Vehicle Perception Systems under Imperfect Communication
by: Bastola, Ashish, et al.
Published: (2025)
by: Bastola, Ashish, et al.
Published: (2025)
LoFA: Learning to Predict Personalized Priors for Fast Adaptation of Visual Generative Models
by: Hao, Yiming, et al.
Published: (2025)
by: Hao, Yiming, et al.
Published: (2025)
EgoMotion: Hierarchical Reasoning and Diffusion for Egocentric Vision-Language Motion Generation
by: Hou, Ruibing, et al.
Published: (2026)
by: Hou, Ruibing, et al.
Published: (2026)
Many-MobileNet: Multi-Model Augmentation for Robust Retinal Disease Classification
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Improving Keystep Recognition in Ego-Video via Dexterous Focus
by: Chavis, Zachary, et al.
Published: (2025)
by: Chavis, Zachary, et al.
Published: (2025)
Prompt-OT: An Optimal Transport Regularization Paradigm for Knowledge Preservation in Vision-Language Model Adaptation
by: Chen, Xiwen, et al.
Published: (2025)
by: Chen, Xiwen, et al.
Published: (2025)
How Effective Can Dropout Be in Multiple Instance Learning ?
by: Zhu, Wenhui, et al.
Published: (2025)
by: Zhu, Wenhui, et al.
Published: (2025)
Efficient Multi-domain Text Recognition Deep Neural Network Parameterization with Residual Adapters
by: Chao, Jiayou, et al.
Published: (2024)
by: Chao, Jiayou, et al.
Published: (2024)
EgoExo-Gen: Ego-centric Video Prediction by Watching Exo-centric Videos
by: Xu, Jilan, et al.
Published: (2025)
by: Xu, Jilan, et al.
Published: (2025)
EgoFlowNet: Non-Rigid Scene Flow from Point Clouds with Ego-Motion Support
by: Battrawy, Ramy, et al.
Published: (2024)
by: Battrawy, Ramy, et al.
Published: (2024)
Cracking Instance Jigsaw Puzzles: An Alternative to Multiple Instance Learning for Whole Slide Image Analysis
by: Chen, Xiwen, et al.
Published: (2025)
by: Chen, Xiwen, et al.
Published: (2025)
UniEgoMotion: A Unified Model for Egocentric Motion Reconstruction, Forecasting, and Generation
by: Patel, Chaitanya, et al.
Published: (2025)
by: Patel, Chaitanya, et al.
Published: (2025)
Multimodal Variational Autoencoder: a Barycentric View
by: Qiu, Peijie, et al.
Published: (2024)
by: Qiu, Peijie, et al.
Published: (2024)
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models
by: Zhu, Wenhui, et al.
Published: (2025)
by: Zhu, Wenhui, et al.
Published: (2025)
EgoLM: Multi-Modal Language Model of Egocentric Motions
by: Hong, Fangzhou, et al.
Published: (2024)
by: Hong, Fangzhou, et al.
Published: (2024)
Human Pose-Constrained UV Map Estimation
by: Suchanek, Matej, et al.
Published: (2025)
by: Suchanek, Matej, et al.
Published: (2025)
Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2026)
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2026)
All You Need for Object Detection: From Pixels, Points, and Prompts to Next-Gen Fusion and Multimodal LLMs/VLMs in Autonomous Vehicles
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2025)
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2025)
DINO-Explorer: Active Underwater Discovery via Ego-Motion Compensated Semantic Predictive Coding
by: Jin, Yuhan, et al.
Published: (2026)
by: Jin, Yuhan, et al.
Published: (2026)
GEM: A Generalizable Ego-Vision Multimodal World Model for Fine-Grained Ego-Motion, Object Dynamics, and Scene Composition Control
by: Hassan, Mariam, et al.
Published: (2024)
by: Hassan, Mariam, et al.
Published: (2024)
SMF-VO: Direct Ego-Motion Estimation via Sparse Motion Fields
by: Yang, Sangheon, et al.
Published: (2025)
by: Yang, Sangheon, et al.
Published: (2025)
Discriminately Treating Motion Components Evolves Joint Depth and Ego-Motion Learning
by: Zhang, Mengtan, et al.
Published: (2025)
by: Zhang, Mengtan, et al.
Published: (2025)
EgoMusic-driven Human Dance Motion Estimation with Skeleton Mamba
by: Nguyen, Quang, et al.
Published: (2025)
by: Nguyen, Quang, et al.
Published: (2025)
Towards Open Environments and Instructions: General Vision-Language Navigation via Fast-Slow Interactive Reasoning
by: Li, Yang, et al.
Published: (2026)
by: Li, Yang, et al.
Published: (2026)
Similar Items
-
VisionGPT: LLM-Assisted Real-Time Anomaly Detection for Safe Visual Navigation
by: Wang, Hao, et al.
Published: (2024) -
Diffusion Prism: Enhancing Diversity and Morphology Consistency in Mask-to-Image Diffusion
by: Wang, Hao, et al.
Published: (2025) -
Fast 2DGS: Efficient Image Representation with Deep Gaussian Prior
by: Wang, Hao, et al.
Published: (2025) -
Motion Focus Recognition in Fast-Moving Egocentric Video
by: Hong, Si-En, et al.
Published: (2026) -
FLAME Diffuser: Wildfire Image Synthesis using Mask Guided Diffusion
by: Wang, Hao, et al.
Published: (2024)