Dual-Projection Fusion for Accurate Upright Panorama Generation in Robotic Vision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shan, Yuhao, Yuan, Qianyi, Liu, Jingguo, Li, Shigang, Li, Jianfeng, Chen, Tong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
360-Degree Full-view Image Segmentation by Spherical Convolution compatible with Large-scale Planar Pre-trained Models
von: Liu, Jingguo, et al.
Veröffentlicht: (2025)
von: Liu, Jingguo, et al.
Veröffentlicht: (2025)
Estimating Depth of Monocular Panoramic Image with Teacher-Student Model Fusing Equirectangular and Spherical Representations
von: Liu, Jingguo, et al.
Veröffentlicht: (2024)
von: Liu, Jingguo, et al.
Veröffentlicht: (2024)
Cascaded Dual Vision Transformer for Accurate Facial Landmark Detection
von: Dang, Ziqiang, et al.
Veröffentlicht: (2024)
von: Dang, Ziqiang, et al.
Veröffentlicht: (2024)
Upright adjustment with graph convolutional networks
von: Jung, Raehyuk, et al.
Veröffentlicht: (2024)
von: Jung, Raehyuk, et al.
Veröffentlicht: (2024)
SphereFusion: Efficient Panorama Depth Estimation via Gated Fusion
von: Yan, Qingsong, et al.
Veröffentlicht: (2025)
von: Yan, Qingsong, et al.
Veröffentlicht: (2025)
Taming Stable Diffusion for Text to 360° Panorama Image Generation
von: Zhang, Cheng, et al.
Veröffentlicht: (2024)
von: Zhang, Cheng, et al.
Veröffentlicht: (2024)
Florence-VL: Enhancing Vision-Language Models with Generative Vision Encoder and Depth-Breadth Fusion
von: Chen, Jiuhai, et al.
Veröffentlicht: (2024)
von: Chen, Jiuhai, et al.
Veröffentlicht: (2024)
LiftProj: Space Lifting and Projection-Based Panorama Stitching
von: Jia, Yuan, et al.
Veröffentlicht: (2025)
von: Jia, Yuan, et al.
Veröffentlicht: (2025)
RCR: Robust Crowd Reconstruction with Upright Space from a Single Large-scene Image
von: Huang, Jing, et al.
Veröffentlicht: (2024)
von: Huang, Jing, et al.
Veröffentlicht: (2024)
BiomedAP: A Vision-Informed Dual-Anchor Framework with Gated Cross-Modal Fusion for Robust Medical Vision-Language Adaptation
von: Tong, Huanyang, et al.
Veröffentlicht: (2026)
von: Tong, Huanyang, et al.
Veröffentlicht: (2026)
Multimodal Fusion and Vision-Language Models: A Survey for Robot Vision
von: Han, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Han, Xiaofeng, et al.
Veröffentlicht: (2025)
Dual-Task Vision Transformer for Rapid and Accurate Intracerebral Hemorrhage CT Image Classification
von: Fan, Jialiang, et al.
Veröffentlicht: (2024)
von: Fan, Jialiang, et al.
Veröffentlicht: (2024)
PanSplat: 4K Panorama Synthesis with Feed-Forward Gaussian Splatting
von: Zhang, Cheng, et al.
Veröffentlicht: (2024)
von: Zhang, Cheng, et al.
Veröffentlicht: (2024)
LayerPano3D: Layered 3D Panorama for Hyper-Immersive Scene Generation
von: Yang, Shuai, et al.
Veröffentlicht: (2024)
von: Yang, Shuai, et al.
Veröffentlicht: (2024)
CamFreeDiff: Camera-free Image to Panorama Generation with Diffusion Model
von: Yuan, Xiaoding, et al.
Veröffentlicht: (2024)
von: Yuan, Xiaoding, et al.
Veröffentlicht: (2024)
DesignEdit: Multi-Layered Latent Decomposition and Fusion for Unified & Accurate Image Editing
von: Jia, Yueru, et al.
Veröffentlicht: (2024)
von: Jia, Yueru, et al.
Veröffentlicht: (2024)
Dual-Level Precision Edges Guided Multi-View Stereo with Accurate Planarization
von: Chen, Kehua, et al.
Veröffentlicht: (2024)
von: Chen, Kehua, et al.
Veröffentlicht: (2024)
EVLF: Early Vision-Language Fusion for Generative Dataset Distillation
von: Cai, Wenqi, et al.
Veröffentlicht: (2026)
von: Cai, Wenqi, et al.
Veröffentlicht: (2026)
Towards an Accurate and Effective Robot Vision (The Problem of Topological Localization for Mobile Robots)
von: Boros, Emanuela
Veröffentlicht: (2025)
von: Boros, Emanuela
Veröffentlicht: (2025)
DSPFusion: Image Fusion via Degradation and Semantic Dual-Prior Guidance
von: Tang, Linfeng, et al.
Veröffentlicht: (2025)
von: Tang, Linfeng, et al.
Veröffentlicht: (2025)
PSGS: Text-driven Panorama Sliding Scene Generation via Gaussian Splatting
von: Zhang, Xin, et al.
Veröffentlicht: (2026)
von: Zhang, Xin, et al.
Veröffentlicht: (2026)
IDOL: Unified Dual-Modal Latent Diffusion for Human-Centric Joint Video-Depth Generation
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
Adapting Vision-Language Foundation Model for Next Generation Medical Ultrasound Image Analysis
von: Qu, Jingguo, et al.
Veröffentlicht: (2025)
von: Qu, Jingguo, et al.
Veröffentlicht: (2025)
JoPano: Unified Panorama Generation via Joint Modeling
von: Feng, Wancheng, et al.
Veröffentlicht: (2025)
von: Feng, Wancheng, et al.
Veröffentlicht: (2025)
ODM: A Text-Image Further Alignment Pre-training Approach for Scene Text Detection and Spotting
von: Duan, Chen, et al.
Veröffentlicht: (2024)
von: Duan, Chen, et al.
Veröffentlicht: (2024)
SP-Det: Self-Prompted Dual-Text Fusion for Generalized Multi-Label Lesion Detection
von: Xu, Qing, et al.
Veröffentlicht: (2025)
von: Xu, Qing, et al.
Veröffentlicht: (2025)
D2-Mamba: Dual-Scale Fusion and Dual-Path Scanning with SSMs for Shadow Removal
von: Li, Linhao, et al.
Veröffentlicht: (2025)
von: Li, Linhao, et al.
Veröffentlicht: (2025)
DualDiff: Dual-branch Diffusion Model for Autonomous Driving with Semantic Fusion
von: Li, Haoteng, et al.
Veröffentlicht: (2025)
von: Li, Haoteng, et al.
Veröffentlicht: (2025)
RayFusion: Ray Fusion Enhanced Collaborative Visual Perception
von: Wang, Shaohong, et al.
Veröffentlicht: (2025)
von: Wang, Shaohong, et al.
Veröffentlicht: (2025)
OOTDiffusion: Outfitting Fusion based Latent Diffusion for Controllable Virtual Try-on
von: Xu, Yuhao, et al.
Veröffentlicht: (2024)
von: Xu, Yuhao, et al.
Veröffentlicht: (2024)
Towards Accurate Camouflaged Object Detection with Mixture Convolution and Interactive Fusion
von: Chen, Geng, et al.
Veröffentlicht: (2021)
von: Chen, Geng, et al.
Veröffentlicht: (2021)
MoGe: Unlocking Accurate Monocular Geometry Estimation for Open-Domain Images with Optimal Training Supervision
von: Wang, Ruicheng, et al.
Veröffentlicht: (2024)
von: Wang, Ruicheng, et al.
Veröffentlicht: (2024)
HyperVision: A Channel-Adaptive Ground-Based Hyperspectral Vision Pre-trained Backbone
von: Fu, Guanyiman, et al.
Veröffentlicht: (2026)
von: Fu, Guanyiman, et al.
Veröffentlicht: (2026)
ReWeaver: Towards Simulation-Ready and Topology-Accurate Garment Reconstruction
von: Li, Ming, et al.
Veröffentlicht: (2026)
von: Li, Ming, et al.
Veröffentlicht: (2026)
Defurnishing with X-Ray Vision: Joint Removal of Furniture from Panoramas and Mesh
von: Dolhasz, Alan, et al.
Veröffentlicht: (2025)
von: Dolhasz, Alan, et al.
Veröffentlicht: (2025)
Mitigate Replication and Copying in Diffusion Models with Generalized Caption and Dual Fusion Enhancement
von: Li, Chenghao, et al.
Veröffentlicht: (2023)
von: Li, Chenghao, et al.
Veröffentlicht: (2023)
VisionLLM-based Multimodal Fusion Network for Glottic Carcinoma Early Detection
von: Jin, Zhaohui, et al.
Veröffentlicht: (2024)
von: Jin, Zhaohui, et al.
Veröffentlicht: (2024)
Generalized Robot 3D Vision-Language Model with Fast Rendering and Pre-Training Vision-Language Alignment
von: Liu, Kangcheng, et al.
Veröffentlicht: (2023)
von: Liu, Kangcheng, et al.
Veröffentlicht: (2023)
VisionCreator: A Native Visual-Generation Agentic Model with Understanding, Thinking, Planning and Creation
von: Lai, Jinxiang, et al.
Veröffentlicht: (2026)
von: Lai, Jinxiang, et al.
Veröffentlicht: (2026)
MoGe-2: Accurate Monocular Geometry with Metric Scale and Sharp Details
von: Wang, Ruicheng, et al.
Veröffentlicht: (2025)
von: Wang, Ruicheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
360-Degree Full-view Image Segmentation by Spherical Convolution compatible with Large-scale Planar Pre-trained Models
von: Liu, Jingguo, et al.
Veröffentlicht: (2025) -
Estimating Depth of Monocular Panoramic Image with Teacher-Student Model Fusing Equirectangular and Spherical Representations
von: Liu, Jingguo, et al.
Veröffentlicht: (2024) -
Cascaded Dual Vision Transformer for Accurate Facial Landmark Detection
von: Dang, Ziqiang, et al.
Veröffentlicht: (2024) -
Upright adjustment with graph convolutional networks
von: Jung, Raehyuk, et al.
Veröffentlicht: (2024) -
SphereFusion: Efficient Panorama Depth Estimation via Gated Fusion
von: Yan, Qingsong, et al.
Veröffentlicht: (2025)