Multi-Step Guided Diffusion for Image Restoration on Edge Devices: Toward Lightweight Perception in Embodied AI
Fuente:
arXiv
Guardado en:
| Autor principal: | Chakravarty, Aditya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AI-driven Dispensing of Coral Reseeding Devices for Broad-scale Restoration of the Great Barrier Reef
por: Raine, Scarlett, et al.
Publicado: (2025)
por: Raine, Scarlett, et al.
Publicado: (2025)
CrackESS: A Self-Prompting Crack Segmentation System for Edge Devices
por: Wang, Yingchu, et al.
Publicado: (2024)
por: Wang, Yingchu, et al.
Publicado: (2024)
ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop
por: Hong, Yining, et al.
Publicado: (2026)
por: Hong, Yining, et al.
Publicado: (2026)
From Detection to Action Recognition: An Edge-Based Pipeline for Robot Human Perception
por: Toupas, Petros, et al.
Publicado: (2023)
por: Toupas, Petros, et al.
Publicado: (2023)
DGFusion: Depth-Guided Sensor Fusion for Robust Semantic Perception
por: Broedermannn, Tim, et al.
Publicado: (2025)
por: Broedermannn, Tim, et al.
Publicado: (2025)
Diff-2-in-1: Bridging Generation and Dense Perception with Diffusion Models
por: Zheng, Shuhong, et al.
Publicado: (2024)
por: Zheng, Shuhong, et al.
Publicado: (2024)
Vidar: Embodied Video Diffusion Model for Generalist Manipulation
por: Feng, Yao, et al.
Publicado: (2025)
por: Feng, Yao, et al.
Publicado: (2025)
ECBench: Can Multi-modal Foundation Models Understand the Egocentric World? A Holistic Embodied Cognition Benchmark
por: Dang, Ronghao, et al.
Publicado: (2025)
por: Dang, Ronghao, et al.
Publicado: (2025)
NeurAll: Towards a Unified Visual Perception Model for Automated Driving
por: Sistu, Ganesh, et al.
Publicado: (2019)
por: Sistu, Ganesh, et al.
Publicado: (2019)
Distributed NeRF Learning for Collaborative Multi-Robot Perception
por: Zhao, Hongrui, et al.
Publicado: (2024)
por: Zhao, Hongrui, et al.
Publicado: (2024)
AI-Driven Marine Robotics: Emerging Trends in Underwater Perception and Ecosystem Monitoring
por: Raine, Scarlett, et al.
Publicado: (2025)
por: Raine, Scarlett, et al.
Publicado: (2025)
EnerVerse: Envisioning Embodied Future Space for Robotics Manipulation
por: Huang, Siyuan, et al.
Publicado: (2025)
por: Huang, Siyuan, et al.
Publicado: (2025)
Planning-Guided Diffusion Policy Learning for Generalizable Contact-Rich Bimanual Manipulation
por: Li, Xuanlin, et al.
Publicado: (2024)
por: Li, Xuanlin, et al.
Publicado: (2024)
Deep Learning-Based Multi-Modal Fusion for Robust Robot Perception and Navigation
por: Lai, Delun, et al.
Publicado: (2025)
por: Lai, Delun, et al.
Publicado: (2025)
EfficientFlow: Efficient Equivariant Flow Policy Learning for Embodied AI
por: Chang, Jianlei, et al.
Publicado: (2025)
por: Chang, Jianlei, et al.
Publicado: (2025)
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
por: Zhang, Junjie, et al.
Publicado: (2024)
por: Zhang, Junjie, et al.
Publicado: (2024)
ESPIRE: A Diagnostic Benchmark for Embodied Spatial Reasoning of Vision-Language Models
por: Zhao, Yanpeng, et al.
Publicado: (2026)
por: Zhao, Yanpeng, et al.
Publicado: (2026)
Adver-City: Open-Source Multi-Modal Dataset for Collaborative Perception Under Adverse Weather Conditions
por: Karvat, Mateus, et al.
Publicado: (2024)
por: Karvat, Mateus, et al.
Publicado: (2024)
PhyScene: Physically Interactable 3D Scene Synthesis for Embodied AI
por: Yang, Yandan, et al.
Publicado: (2024)
por: Yang, Yandan, et al.
Publicado: (2024)
Deep Learning Models in Speech Recognition: Measuring GPU Energy Consumption, Impact of Noise and Model Quantization for Edge Deployment
por: Chakravarty, Aditya
Publicado: (2024)
por: Chakravarty, Aditya
Publicado: (2024)
Multi-Space Alignments Towards Universal LiDAR Segmentation
por: Liu, Youquan, et al.
Publicado: (2024)
por: Liu, Youquan, et al.
Publicado: (2024)
AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems
por: AgiBot-World-Contributors, et al.
Publicado: (2025)
por: AgiBot-World-Contributors, et al.
Publicado: (2025)
Seeing Farther and Smarter: Value-Guided Multi-Path Reflection for VLM Policy Optimization
por: Yang, Yanting, et al.
Publicado: (2026)
por: Yang, Yanting, et al.
Publicado: (2026)
M2H: Multi-Task Learning with Efficient Window-Based Cross-Task Attention for Monocular Spatial Perception
por: Udugama, U. V. B. L, et al.
Publicado: (2025)
por: Udugama, U. V. B. L, et al.
Publicado: (2025)
Towards Multi-Source Domain Generalization for Sleep Staging with Noisy Labels
por: Wang, Kening, et al.
Publicado: (2026)
por: Wang, Kening, et al.
Publicado: (2026)
The GOOSE Dataset for Perception in Unstructured Environments
por: Mortimer, Peter, et al.
Publicado: (2023)
por: Mortimer, Peter, et al.
Publicado: (2023)
ZeD-MAP: Bundle Adjustment Guided Zero-Shot Depth Maps for Real-Time Aerial Imaging
por: Iz, Selim Ahmet, et al.
Publicado: (2026)
por: Iz, Selim Ahmet, et al.
Publicado: (2026)
SLNet: A Super-Lightweight Geometry-Adaptive Network for 3D Point Cloud Recognition
por: Saeid, Mohammad, et al.
Publicado: (2026)
por: Saeid, Mohammad, et al.
Publicado: (2026)
Panoptic Perception for Autonomous Driving: A Survey
por: Li, Yunge, et al.
Publicado: (2024)
por: Li, Yunge, et al.
Publicado: (2024)
Lightweight Multimodal Artificial Intelligence Framework for Maritime Multi-Scene Recognition
por: Xi, Xinyu, et al.
Publicado: (2025)
por: Xi, Xinyu, et al.
Publicado: (2025)
TransDiffuser: Diverse Trajectory Generation with Decorrelated Multi-modal Representation for End-to-end Autonomous Driving
por: Jiang, Xuefeng, et al.
Publicado: (2025)
por: Jiang, Xuefeng, et al.
Publicado: (2025)
Intelligent Sensing-to-Action for Robust Autonomy at the Edge: Opportunities and Challenges
por: Trivedi, Amit Ranjan, et al.
Publicado: (2025)
por: Trivedi, Amit Ranjan, et al.
Publicado: (2025)
Image Quality Assessment for Embodied AI
por: Li, Chunyi, et al.
Publicado: (2025)
por: Li, Chunyi, et al.
Publicado: (2025)
GAC-KAN: An Ultra-Lightweight GNSS Interference Classifier for GenAI-Powered Consumer Edge Devices
por: Zeng, Zhihan, et al.
Publicado: (2026)
por: Zeng, Zhihan, et al.
Publicado: (2026)
AnyThermal: Towards Learning Universal Representations for Thermal Perception
por: Maheshwari, Parv, et al.
Publicado: (2026)
por: Maheshwari, Parv, et al.
Publicado: (2026)
LiDAR Based Semantic Perception for Forklifts in Outdoor Environments
por: Serfling, Benjamin, et al.
Publicado: (2025)
por: Serfling, Benjamin, et al.
Publicado: (2025)
Diffusion Reward: Learning Rewards via Conditional Video Diffusion
por: Huang, Tao, et al.
Publicado: (2023)
por: Huang, Tao, et al.
Publicado: (2023)
Depth Matters: Multimodal RGB-D Perception for Robust Autonomous Agents
por: Clement, Mihaela-Larisa, et al.
Publicado: (2025)
por: Clement, Mihaela-Larisa, et al.
Publicado: (2025)
Is Intermediate Fusion All You Need for UAV-based Collaborative Perception?
por: Hao, Jiuwu, et al.
Publicado: (2025)
por: Hao, Jiuwu, et al.
Publicado: (2025)
Embodied AI with Foundation Models for Mobile Service Robots: A Systematic Review
por: Lisondra, Matthew, et al.
Publicado: (2025)
por: Lisondra, Matthew, et al.
Publicado: (2025)
Ejemplares similares
-
AI-driven Dispensing of Coral Reseeding Devices for Broad-scale Restoration of the Great Barrier Reef
por: Raine, Scarlett, et al.
Publicado: (2025) -
CrackESS: A Self-Prompting Crack Segmentation System for Edge Devices
por: Wang, Yingchu, et al.
Publicado: (2024) -
ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop
por: Hong, Yining, et al.
Publicado: (2026) -
From Detection to Action Recognition: An Edge-Based Pipeline for Robot Human Perception
por: Toupas, Petros, et al.
Publicado: (2023) -
DGFusion: Depth-Guided Sensor Fusion for Robust Semantic Perception
por: Broedermannn, Tim, et al.
Publicado: (2025)