Saved in:
| Main Authors: | Wakai, Nobuhiko, Sato, Satoshi, Ishii, Yasunori, Yamashita, Takayoshi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2503.14228 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deep Single Image Camera Calibration by Heatmap Regression to Recover Fisheye Images Under Manhattan World Assumption
by: Wakai, Nobuhiko, et al.
Published: (2023)
by: Wakai, Nobuhiko, et al.
Published: (2023)
Bidirectional Action Sequence Learning for Long-term Action Anticipation with Large Language Models
by: Sato, Yuji, et al.
Published: (2025)
by: Sato, Yuji, et al.
Published: (2025)
Location-guided Head Pose Estimation for Fisheye Image
by: Li, Bing, et al.
Published: (2024)
by: Li, Bing, et al.
Published: (2024)
Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tokens
by: Gangopadhyay, Suchisrit, et al.
Published: (2025)
by: Gangopadhyay, Suchisrit, et al.
Published: (2025)
PFDepth: Heterogeneous Pinhole-Fisheye Joint Depth Estimation via Distortion-aware Gaussian-Splatted Volumetric Fusion
by: Zhang, Zhiwei, et al.
Published: (2025)
by: Zhang, Zhiwei, et al.
Published: (2025)
Impact of Video Compression Artifacts on Fisheye Camera Visual Perception Tasks
by: Sakthi, Madhumitha, et al.
Published: (2024)
by: Sakthi, Madhumitha, et al.
Published: (2024)
Enhanced Parking Perception by Multi-Task Fisheye Cross-view Transformers
by: Musabini, Antonyo, et al.
Published: (2024)
by: Musabini, Antonyo, et al.
Published: (2024)
Edge-case Synthesis for Fisheye Object Detection: A Data-centric Perspective
by: Kim, Seunghyeon, et al.
Published: (2025)
by: Kim, Seunghyeon, et al.
Published: (2025)
DAOVI: Distortion-Aware Omnidirectional Video Inpainting
by: Seshimo, Ryosuke, et al.
Published: (2025)
by: Seshimo, Ryosuke, et al.
Published: (2025)
Local Representative Token Guided Merging for Text-to-Image Generation
by: Lee, Min-Jeong, et al.
Published: (2025)
by: Lee, Min-Jeong, et al.
Published: (2025)
What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token Merging
by: Kang, Inha, et al.
Published: (2025)
by: Kang, Inha, et al.
Published: (2025)
Mining Your Own Secrets: Diffusion Classifier Scores for Continual Personalization of Text-to-Image Diffusion Models
by: Jha, Saurav, et al.
Published: (2024)
by: Jha, Saurav, et al.
Published: (2024)
PanoWorld: Geometry-Consistent Panoramic Video World Modeling
by: Jiang, Le, et al.
Published: (2026)
by: Jiang, Le, et al.
Published: (2026)
MPT-PAR:Mix-Parameters Transformer for Panoramic Activity Recognition
by: Gan, Wenqing, et al.
Published: (2024)
by: Gan, Wenqing, et al.
Published: (2024)
OPGAgent: An Agent for Auditable Dental Panoramic X-ray Interpretation
by: Yu, Zhaolin, et al.
Published: (2026)
by: Yu, Zhaolin, et al.
Published: (2026)
TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
by: Qu, Liao, et al.
Published: (2024)
by: Qu, Liao, et al.
Published: (2024)
Unified Panoramic Geometry Estimation via Multi-View Foundation Models
by: Bozic, Vukasin, et al.
Published: (2026)
by: Bozic, Vukasin, et al.
Published: (2026)
GaussianToken: An Effective Image Tokenizer with 2D Gaussian Splatting
by: Dong, Jiajun, et al.
Published: (2025)
by: Dong, Jiajun, et al.
Published: (2025)
Context-Aware Weakly Supervised Image Manipulation Localization with SAM Refinement
by: Wang, Xinghao, et al.
Published: (2025)
by: Wang, Xinghao, et al.
Published: (2025)
Spatial-Frequency Aware for Object Detection in RAW Image
by: Ye, Zhuohua, et al.
Published: (2025)
by: Ye, Zhuohua, et al.
Published: (2025)
Frequency-Aware Token Reduction for Efficient Vision Transformer
by: Lee, Dong-Jae, et al.
Published: (2025)
by: Lee, Dong-Jae, et al.
Published: (2025)
QueryCDR: Query-Based Controllable Distortion Rectification Network for Fisheye Images
by: Guo, Pengbo, et al.
Published: (2024)
by: Guo, Pengbo, et al.
Published: (2024)
StableAnimator++: Overcoming Pose Misalignment and Face Distortion for Human Image Animation
by: Tu, Shuyuan, et al.
Published: (2025)
by: Tu, Shuyuan, et al.
Published: (2025)
PanoLora: Bridging Perspective and Panoramic Video Generation with LoRA Adaptation
by: Dong, Zeyu, et al.
Published: (2025)
by: Dong, Zeyu, et al.
Published: (2025)
TexLiDAR: Automated Text Understanding for Panoramic LiDAR Data
by: Cohen, Naor, et al.
Published: (2025)
by: Cohen, Naor, et al.
Published: (2025)
On the Adversarial Robustness of Discrete Image Tokenizers
by: Bhagwatkar, Rishika, et al.
Published: (2026)
by: Bhagwatkar, Rishika, et al.
Published: (2026)
Homogeneous Tokenizer Matters: Homogeneous Visual Tokenizer for Remote Sensing Image Understanding
by: Shao, Run, et al.
Published: (2024)
by: Shao, Run, et al.
Published: (2024)
Real-Time AI-Driven People Tracking and Counting Using Overhead Cameras
by: Ahamed, Ishrath, et al.
Published: (2024)
by: Ahamed, Ishrath, et al.
Published: (2024)
DistortBench: Benchmarking Vision Language Models on Image Distortion Identification
by: Goyal, Divyanshu, et al.
Published: (2026)
by: Goyal, Divyanshu, et al.
Published: (2026)
VideoPanda: Video Panoramic Diffusion with Multi-view Attention
by: Xie, Kevin, et al.
Published: (2025)
by: Xie, Kevin, et al.
Published: (2025)
From Narrow to Panoramic Vision: Attention-Guided Cold-Start Reshapes Multimodal Reasoning
by: Luo, Ruilin, et al.
Published: (2026)
by: Luo, Ruilin, et al.
Published: (2026)
Frequency Autoregressive Image Generation with Continuous Tokens
by: Yu, Hu, et al.
Published: (2025)
by: Yu, Hu, et al.
Published: (2025)
Hita: Holistic Tokenizer for Autoregressive Image Generation
by: Zheng, Anlin, et al.
Published: (2025)
by: Zheng, Anlin, et al.
Published: (2025)
Scaling Image Tokenizers with Grouped Spherical Quantization
by: Wang, Jiangtao, et al.
Published: (2024)
by: Wang, Jiangtao, et al.
Published: (2024)
Scalable Image Tokenization with Index Backpropagation Quantization
by: Shi, Fengyuan, et al.
Published: (2024)
by: Shi, Fengyuan, et al.
Published: (2024)
VISTA: A Panoramic View of Neural Representations
by: White, Tom
Published: (2024)
by: White, Tom
Published: (2024)
Beyond Pedestrians: Caption-Guided CLIP Framework for High-Difficulty Video-based Person Re-Identification
by: Hamano, Shogo, et al.
Published: (2026)
by: Hamano, Shogo, et al.
Published: (2026)
Fair Generation without Unfair Distortions: Debiasing Text-to-Image Generation with Entanglement-Free Attention
by: Park, Jeonghoon, et al.
Published: (2025)
by: Park, Jeonghoon, et al.
Published: (2025)
Token Pruning using a Lightweight Background Aware Vision Transformer
by: Sah, Sudhakar, et al.
Published: (2024)
by: Sah, Sudhakar, et al.
Published: (2024)
PanoTPS-Net: Panoramic Room Layout Estimation via Thin Plate Spline Transformation
by: Ibrahem, Hatem, et al.
Published: (2025)
by: Ibrahem, Hatem, et al.
Published: (2025)
Similar Items
-
Deep Single Image Camera Calibration by Heatmap Regression to Recover Fisheye Images Under Manhattan World Assumption
by: Wakai, Nobuhiko, et al.
Published: (2023) -
Bidirectional Action Sequence Learning for Long-term Action Anticipation with Large Language Models
by: Sato, Yuji, et al.
Published: (2025) -
Location-guided Head Pose Estimation for Fisheye Image
by: Li, Bing, et al.
Published: (2024) -
Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tokens
by: Gangopadhyay, Suchisrit, et al.
Published: (2025) -
PFDepth: Heterogeneous Pinhole-Fisheye Joint Depth Estimation via Distortion-aware Gaussian-Splatted Volumetric Fusion
by: Zhang, Zhiwei, et al.
Published: (2025)