Can Foundation Models Revolutionize Mobile AR Sparse Sensing?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yiqin, Guo, Tian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PI-HMR: Towards Robust In-bed Temporal Human Shape Reconstruction with Contact Pressure Sensing
von: Wu, Ziyu, et al.
Veröffentlicht: (2025)
von: Wu, Ziyu, et al.
Veröffentlicht: (2025)
Spatio-Temporal Foundation Models: Vision, Challenges, and Opportunities
von: Goodge, Adam, et al.
Veröffentlicht: (2025)
von: Goodge, Adam, et al.
Veröffentlicht: (2025)
AR Glulam: Accurate Augmented Reality Using Multiple Fiducial Markers for Glulam Fabrication
von: Kyaw, Alexander Htet, et al.
Veröffentlicht: (2025)
von: Kyaw, Alexander Htet, et al.
Veröffentlicht: (2025)
VOLMO: Versatile and Open Large Models for Ophthalmology
von: Qin, Zhenyue, et al.
Veröffentlicht: (2026)
von: Qin, Zhenyue, et al.
Veröffentlicht: (2026)
From Pixels to Nucleotides: End-to-End Token-Based Video Compression for DNA Storage
von: Ruan, Cihan, et al.
Veröffentlicht: (2026)
von: Ruan, Cihan, et al.
Veröffentlicht: (2026)
Scaling Ultrasound Volumetric Reconstruction via Mobile Augmented Reality
von: Ng, Kian Wei, et al.
Veröffentlicht: (2026)
von: Ng, Kian Wei, et al.
Veröffentlicht: (2026)
Learned Display Radiance Fields with Lensless Cameras
von: Chen, Ziyang, et al.
Veröffentlicht: (2025)
von: Chen, Ziyang, et al.
Veröffentlicht: (2025)
Introducing Nylon Face Mask Attacks: A Dataset for Evaluating Generalised Face Presentation Attack Detection
von: Manasa, et al.
Veröffentlicht: (2025)
von: Manasa, et al.
Veröffentlicht: (2025)
SynSpill: Improved Industrial Spill Detection With Synthetic Data
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025)
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025)
Towards Railway Domain Adaptation for LiDAR-based 3D Detection: Road-to-Rail and Sim-to-Real via SynDRA-BBox
von: Diaz, Xavier, et al.
Veröffentlicht: (2025)
von: Diaz, Xavier, et al.
Veröffentlicht: (2025)
DashCam Video: A complementary low-cost data stream for on-demand forest-infrastructure system monitoring
von: Joshi, Durga, et al.
Veröffentlicht: (2025)
von: Joshi, Durga, et al.
Veröffentlicht: (2025)
Attention-based Generative Latent Replay: A Continual Learning Approach for WSI Analysis
von: Kumari, Pratibha, et al.
Veröffentlicht: (2025)
von: Kumari, Pratibha, et al.
Veröffentlicht: (2025)
Diff-GNSS: Diffusion-based Pseudorange Error Estimation
von: Zhu, Jiaqi, et al.
Veröffentlicht: (2025)
von: Zhu, Jiaqi, et al.
Veröffentlicht: (2025)
Probabilistic Online Event Downsampling
von: Girbau-Xalabarder, Andreu, et al.
Veröffentlicht: (2025)
von: Girbau-Xalabarder, Andreu, et al.
Veröffentlicht: (2025)
A Manually Annotated Image-Caption Dataset for Detecting Children in the Wild
von: Kireev, Klim, et al.
Veröffentlicht: (2025)
von: Kireev, Klim, et al.
Veröffentlicht: (2025)
Hyperspectral Sensors and Autonomous Driving: Technologies, Limitations, and Opportunities
von: Shah, Imad Ali, et al.
Veröffentlicht: (2025)
von: Shah, Imad Ali, et al.
Veröffentlicht: (2025)
Unmanned Aerial Vehicle (UAV)-Based Mapping of Iris Pseudacorus L. Invasion in Laguna del Sauce (Uruguay) Coast
von: Silvarrey, Alejo, et al.
Veröffentlicht: (2025)
von: Silvarrey, Alejo, et al.
Veröffentlicht: (2025)
Multi-Image Super Resolution Framework for Detection and Analysis of Plant Roots
von: Agarwal, Shubham, et al.
Veröffentlicht: (2026)
von: Agarwal, Shubham, et al.
Veröffentlicht: (2026)
Evaluating and Enhancing Trustworthiness of LLMs in Perception Tasks
von: Dona, Malsha Ashani Mahawatta, et al.
Veröffentlicht: (2024)
von: Dona, Malsha Ashani Mahawatta, et al.
Veröffentlicht: (2024)
Scrutinizing Data from Sky: An Examination of Its Veracity in Area Based Traffic Contexts
von: Ali, Yawar, et al.
Veröffentlicht: (2024)
von: Ali, Yawar, et al.
Veröffentlicht: (2024)
Fourier-based Action Recognition for Wildlife Behavior Quantification with Event Cameras
von: Hamann, Friedhelm, et al.
Veröffentlicht: (2024)
von: Hamann, Friedhelm, et al.
Veröffentlicht: (2024)
x-RAGE: eXtended Reality -- Action & Gesture Events Dataset
von: Parmar, Vivek, et al.
Veröffentlicht: (2024)
von: Parmar, Vivek, et al.
Veröffentlicht: (2024)
Fast Quantum Convolutional Neural Networks for Low-Complexity Object Detection in Autonomous Driving Applications
von: Baek, Hankyul, et al.
Veröffentlicht: (2023)
von: Baek, Hankyul, et al.
Veröffentlicht: (2023)
Enhancing Autism Spectrum Disorder Early Detection with the Parent-Child Dyads Block-Play Protocol and an Attention-enhanced GCN-xLSTM Hybrid Deep Learning Framework
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Assistive Image Annotation Systems with Deep Learning and Natural Language Capabilities: A Review
von: Mots'oehli, Moseli
Veröffentlicht: (2024)
von: Mots'oehli, Moseli
Veröffentlicht: (2024)
Self-evolving Embodied AI
von: Feng, Tongtong, et al.
Veröffentlicht: (2026)
von: Feng, Tongtong, et al.
Veröffentlicht: (2026)
Extracting Object Heights From LiDAR & Aerial Imagery
von: Guerrero, Jesus
Veröffentlicht: (2024)
von: Guerrero, Jesus
Veröffentlicht: (2024)
Unlocking Comics: The AI4VA Dataset for Visual Understanding
von: Grönquist, Peter, et al.
Veröffentlicht: (2024)
von: Grönquist, Peter, et al.
Veröffentlicht: (2024)
SpecTrack: Learned Multi-Rotation Tracking via Speckle Imaging
von: Chen, Ziyang, et al.
Veröffentlicht: (2024)
von: Chen, Ziyang, et al.
Veröffentlicht: (2024)
EATFormer: Improving Vision Transformer Inspired by Evolutionary Algorithm
von: Zhang, Jiangning, et al.
Veröffentlicht: (2022)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2022)
INSIGHT: Indoor Scene Intelligence from Geometric-Semantic Hierarchy Transfer for Public~Safety
von: Dimopoulos, Alexander Nikitas, et al.
Veröffentlicht: (2026)
von: Dimopoulos, Alexander Nikitas, et al.
Veröffentlicht: (2026)
Neuromorphic Face Analysis: a Survey
von: Becattini, Federico, et al.
Veröffentlicht: (2024)
von: Becattini, Federico, et al.
Veröffentlicht: (2024)
Detecting sexually explicit content in the context of the child sexual abuse materials (CSAM): end-to-end classifiers and region-based networks
von: Gutfeter, Weronika, et al.
Veröffentlicht: (2024)
von: Gutfeter, Weronika, et al.
Veröffentlicht: (2024)
All-Optical Segmentation via Diffractive Neural Networks for Autonomous Driving
von: Li, Yingjie, et al.
Veröffentlicht: (2026)
von: Li, Yingjie, et al.
Veröffentlicht: (2026)
V-RoAst: Visual Road Assessment. Can VLM be a Road Safety Assessor Using the iRAP Standard?
von: Jongwiriyanurak, Natchapon, et al.
Veröffentlicht: (2024)
von: Jongwiriyanurak, Natchapon, et al.
Veröffentlicht: (2024)
From Flat to Spatial: Comparison of 4 methods constructing 3D, 2 and 1/2D Models from 2D Plans with neural networks
von: Sam, Jacob, et al.
Veröffentlicht: (2024)
von: Sam, Jacob, et al.
Veröffentlicht: (2024)
Built Environment Reasoning from Remote Sensing Imagery Using Large Vision--Language Models
von: Wang, Dongdong, et al.
Veröffentlicht: (2026)
von: Wang, Dongdong, et al.
Veröffentlicht: (2026)
Human-in-the-Loop: Quantitative Evaluation of 3D Models Generation by Large Language Models
von: Sadik, Ahmed R., et al.
Veröffentlicht: (2025)
von: Sadik, Ahmed R., et al.
Veröffentlicht: (2025)
Complex-Valued Holographic Radiance Fields
von: Zhan, Yicheng, et al.
Veröffentlicht: (2025)
von: Zhan, Yicheng, et al.
Veröffentlicht: (2025)
EgoPoseVR: Spatiotemporal Multi-Modal Reasoning for Egocentric Full-Body Pose in Virtual Reality
von: Cheng, Haojie, et al.
Veröffentlicht: (2026)
von: Cheng, Haojie, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PI-HMR: Towards Robust In-bed Temporal Human Shape Reconstruction with Contact Pressure Sensing
von: Wu, Ziyu, et al.
Veröffentlicht: (2025) -
Spatio-Temporal Foundation Models: Vision, Challenges, and Opportunities
von: Goodge, Adam, et al.
Veröffentlicht: (2025) -
AR Glulam: Accurate Augmented Reality Using Multiple Fiducial Markers for Glulam Fabrication
von: Kyaw, Alexander Htet, et al.
Veröffentlicht: (2025) -
VOLMO: Versatile and Open Large Models for Ophthalmology
von: Qin, Zhenyue, et al.
Veröffentlicht: (2026) -
From Pixels to Nucleotides: End-to-End Token-Based Video Compression for DNA Storage
von: Ruan, Cihan, et al.
Veröffentlicht: (2026)