Dual-Projection Fusion for Accurate Upright Panorama Generation in Robotic Vision
Fuente:
arXiv
Saved in:
| Main Authors: | Shan, Yuhao, Yuan, Qianyi, Liu, Jingguo, Li, Shigang, Li, Jianfeng, Chen, Tong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
360-Degree Full-view Image Segmentation by Spherical Convolution compatible with Large-scale Planar Pre-trained Models
by: Liu, Jingguo, et al.
Published: (2025)
by: Liu, Jingguo, et al.
Published: (2025)
Estimating Depth of Monocular Panoramic Image with Teacher-Student Model Fusing Equirectangular and Spherical Representations
by: Liu, Jingguo, et al.
Published: (2024)
by: Liu, Jingguo, et al.
Published: (2024)
Cascaded Dual Vision Transformer for Accurate Facial Landmark Detection
by: Dang, Ziqiang, et al.
Published: (2024)
by: Dang, Ziqiang, et al.
Published: (2024)
Upright adjustment with graph convolutional networks
by: Jung, Raehyuk, et al.
Published: (2024)
by: Jung, Raehyuk, et al.
Published: (2024)
SphereFusion: Efficient Panorama Depth Estimation via Gated Fusion
by: Yan, Qingsong, et al.
Published: (2025)
by: Yan, Qingsong, et al.
Published: (2025)
Taming Stable Diffusion for Text to 360° Panorama Image Generation
by: Zhang, Cheng, et al.
Published: (2024)
by: Zhang, Cheng, et al.
Published: (2024)
Florence-VL: Enhancing Vision-Language Models with Generative Vision Encoder and Depth-Breadth Fusion
by: Chen, Jiuhai, et al.
Published: (2024)
by: Chen, Jiuhai, et al.
Published: (2024)
LiftProj: Space Lifting and Projection-Based Panorama Stitching
by: Jia, Yuan, et al.
Published: (2025)
by: Jia, Yuan, et al.
Published: (2025)
RCR: Robust Crowd Reconstruction with Upright Space from a Single Large-scene Image
by: Huang, Jing, et al.
Published: (2024)
by: Huang, Jing, et al.
Published: (2024)
BiomedAP: A Vision-Informed Dual-Anchor Framework with Gated Cross-Modal Fusion for Robust Medical Vision-Language Adaptation
by: Tong, Huanyang, et al.
Published: (2026)
by: Tong, Huanyang, et al.
Published: (2026)
Multimodal Fusion and Vision-Language Models: A Survey for Robot Vision
by: Han, Xiaofeng, et al.
Published: (2025)
by: Han, Xiaofeng, et al.
Published: (2025)
Dual-Task Vision Transformer for Rapid and Accurate Intracerebral Hemorrhage CT Image Classification
by: Fan, Jialiang, et al.
Published: (2024)
by: Fan, Jialiang, et al.
Published: (2024)
PanSplat: 4K Panorama Synthesis with Feed-Forward Gaussian Splatting
by: Zhang, Cheng, et al.
Published: (2024)
by: Zhang, Cheng, et al.
Published: (2024)
LayerPano3D: Layered 3D Panorama for Hyper-Immersive Scene Generation
by: Yang, Shuai, et al.
Published: (2024)
by: Yang, Shuai, et al.
Published: (2024)
CamFreeDiff: Camera-free Image to Panorama Generation with Diffusion Model
by: Yuan, Xiaoding, et al.
Published: (2024)
by: Yuan, Xiaoding, et al.
Published: (2024)
DesignEdit: Multi-Layered Latent Decomposition and Fusion for Unified & Accurate Image Editing
by: Jia, Yueru, et al.
Published: (2024)
by: Jia, Yueru, et al.
Published: (2024)
Dual-Level Precision Edges Guided Multi-View Stereo with Accurate Planarization
by: Chen, Kehua, et al.
Published: (2024)
by: Chen, Kehua, et al.
Published: (2024)
EVLF: Early Vision-Language Fusion for Generative Dataset Distillation
by: Cai, Wenqi, et al.
Published: (2026)
by: Cai, Wenqi, et al.
Published: (2026)
Towards an Accurate and Effective Robot Vision (The Problem of Topological Localization for Mobile Robots)
by: Boros, Emanuela
Published: (2025)
by: Boros, Emanuela
Published: (2025)
DSPFusion: Image Fusion via Degradation and Semantic Dual-Prior Guidance
by: Tang, Linfeng, et al.
Published: (2025)
by: Tang, Linfeng, et al.
Published: (2025)
PSGS: Text-driven Panorama Sliding Scene Generation via Gaussian Splatting
by: Zhang, Xin, et al.
Published: (2026)
by: Zhang, Xin, et al.
Published: (2026)
IDOL: Unified Dual-Modal Latent Diffusion for Human-Centric Joint Video-Depth Generation
by: Zhai, Yuanhao, et al.
Published: (2024)
by: Zhai, Yuanhao, et al.
Published: (2024)
Adapting Vision-Language Foundation Model for Next Generation Medical Ultrasound Image Analysis
by: Qu, Jingguo, et al.
Published: (2025)
by: Qu, Jingguo, et al.
Published: (2025)
JoPano: Unified Panorama Generation via Joint Modeling
by: Feng, Wancheng, et al.
Published: (2025)
by: Feng, Wancheng, et al.
Published: (2025)
ODM: A Text-Image Further Alignment Pre-training Approach for Scene Text Detection and Spotting
by: Duan, Chen, et al.
Published: (2024)
by: Duan, Chen, et al.
Published: (2024)
SP-Det: Self-Prompted Dual-Text Fusion for Generalized Multi-Label Lesion Detection
by: Xu, Qing, et al.
Published: (2025)
by: Xu, Qing, et al.
Published: (2025)
D2-Mamba: Dual-Scale Fusion and Dual-Path Scanning with SSMs for Shadow Removal
by: Li, Linhao, et al.
Published: (2025)
by: Li, Linhao, et al.
Published: (2025)
DualDiff: Dual-branch Diffusion Model for Autonomous Driving with Semantic Fusion
by: Li, Haoteng, et al.
Published: (2025)
by: Li, Haoteng, et al.
Published: (2025)
RayFusion: Ray Fusion Enhanced Collaborative Visual Perception
by: Wang, Shaohong, et al.
Published: (2025)
by: Wang, Shaohong, et al.
Published: (2025)
OOTDiffusion: Outfitting Fusion based Latent Diffusion for Controllable Virtual Try-on
by: Xu, Yuhao, et al.
Published: (2024)
by: Xu, Yuhao, et al.
Published: (2024)
Towards Accurate Camouflaged Object Detection with Mixture Convolution and Interactive Fusion
by: Chen, Geng, et al.
Published: (2021)
by: Chen, Geng, et al.
Published: (2021)
MoGe: Unlocking Accurate Monocular Geometry Estimation for Open-Domain Images with Optimal Training Supervision
by: Wang, Ruicheng, et al.
Published: (2024)
by: Wang, Ruicheng, et al.
Published: (2024)
HyperVision: A Channel-Adaptive Ground-Based Hyperspectral Vision Pre-trained Backbone
by: Fu, Guanyiman, et al.
Published: (2026)
by: Fu, Guanyiman, et al.
Published: (2026)
ReWeaver: Towards Simulation-Ready and Topology-Accurate Garment Reconstruction
by: Li, Ming, et al.
Published: (2026)
by: Li, Ming, et al.
Published: (2026)
Defurnishing with X-Ray Vision: Joint Removal of Furniture from Panoramas and Mesh
by: Dolhasz, Alan, et al.
Published: (2025)
by: Dolhasz, Alan, et al.
Published: (2025)
Mitigate Replication and Copying in Diffusion Models with Generalized Caption and Dual Fusion Enhancement
by: Li, Chenghao, et al.
Published: (2023)
by: Li, Chenghao, et al.
Published: (2023)
VisionLLM-based Multimodal Fusion Network for Glottic Carcinoma Early Detection
by: Jin, Zhaohui, et al.
Published: (2024)
by: Jin, Zhaohui, et al.
Published: (2024)
Generalized Robot 3D Vision-Language Model with Fast Rendering and Pre-Training Vision-Language Alignment
by: Liu, Kangcheng, et al.
Published: (2023)
by: Liu, Kangcheng, et al.
Published: (2023)
VisionCreator: A Native Visual-Generation Agentic Model with Understanding, Thinking, Planning and Creation
by: Lai, Jinxiang, et al.
Published: (2026)
by: Lai, Jinxiang, et al.
Published: (2026)
MoGe-2: Accurate Monocular Geometry with Metric Scale and Sharp Details
by: Wang, Ruicheng, et al.
Published: (2025)
by: Wang, Ruicheng, et al.
Published: (2025)
Similar Items
-
360-Degree Full-view Image Segmentation by Spherical Convolution compatible with Large-scale Planar Pre-trained Models
by: Liu, Jingguo, et al.
Published: (2025) -
Estimating Depth of Monocular Panoramic Image with Teacher-Student Model Fusing Equirectangular and Spherical Representations
by: Liu, Jingguo, et al.
Published: (2024) -
Cascaded Dual Vision Transformer for Accurate Facial Landmark Detection
by: Dang, Ziqiang, et al.
Published: (2024) -
Upright adjustment with graph convolutional networks
by: Jung, Raehyuk, et al.
Published: (2024) -
SphereFusion: Efficient Panorama Depth Estimation via Gated Fusion
by: Yan, Qingsong, et al.
Published: (2025)