Probing CLIP's Comprehension of 360-Degree Textual and Visual Semantics
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Hai, Yang, Xiaochen, Dong, Mingzhi, Xue, Jing-Hao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
360PanT: Training-Free Text-Driven 360-Degree Panorama-to-Panorama Translation
por: Wang, Hai, et al.
Publicado: (2024)
por: Wang, Hai, et al.
Publicado: (2024)
A Survey on Text-Driven 360-Degree Panorama Generation
por: Wang, Hai, et al.
Publicado: (2025)
por: Wang, Hai, et al.
Publicado: (2025)
360DVO: Deep Visual Odometry for Monocular 360-Degree Camera
por: Guo, Xiaopeng, et al.
Publicado: (2026)
por: Guo, Xiaopeng, et al.
Publicado: (2026)
Grad-ECLIP: Gradient-based Visual and Textual Explanations for CLIP
por: Zhao, Chenyang, et al.
Publicado: (2025)
por: Zhao, Chenyang, et al.
Publicado: (2025)
Harnessing Textual Semantic Priors for Knowledge Transfer and Refinement in CLIP-Driven Continual Learning
por: He, Lingfeng, et al.
Publicado: (2025)
por: He, Lingfeng, et al.
Publicado: (2025)
FG-CLIP: Fine-Grained Visual and Textual Alignment
por: Xie, Chunyu, et al.
Publicado: (2025)
por: Xie, Chunyu, et al.
Publicado: (2025)
Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning
por: Wang, Yifan, et al.
Publicado: (2026)
por: Wang, Yifan, et al.
Publicado: (2026)
Continual Learning on CLIP via Incremental Prompt Tuning with Intrinsic Textual Anchors
por: Lu, Haodong, et al.
Publicado: (2025)
por: Lu, Haodong, et al.
Publicado: (2025)
CLIPErase: Efficient Unlearning of Visual-Textual Associations in CLIP
por: Yang, Tianyu, et al.
Publicado: (2024)
por: Yang, Tianyu, et al.
Publicado: (2024)
Spherical Vision Transformers for Audio-Visual Saliency Prediction in 360-Degree Videos
por: Cokelek, Mert, et al.
Publicado: (2025)
por: Cokelek, Mert, et al.
Publicado: (2025)
Anomaly Detection for People with Visual Impairments Using an Egocentric 360-Degree Camera
por: Song, Inpyo, et al.
Publicado: (2024)
por: Song, Inpyo, et al.
Publicado: (2024)
360DVD: Controllable Panorama Video Generation with 360-Degree Video Diffusion Model
por: Wang, Qian, et al.
Publicado: (2024)
por: Wang, Qian, et al.
Publicado: (2024)
Elite360D: Towards Efficient 360 Depth Estimation via Semantic- and Distance-Aware Bi-Projection Fusion
por: Ai, Hao, et al.
Publicado: (2024)
por: Ai, Hao, et al.
Publicado: (2024)
Video Question Answering for People with Visual Impairments Using an Egocentric 360-Degree Camera
por: Song, Inpyo, et al.
Publicado: (2024)
por: Song, Inpyo, et al.
Publicado: (2024)
Adaptive Score Alignment Learning for Continual Perceptual Quality Assessment of 360-Degree Videos in Virtual Reality
por: Zhou, Kanglei, et al.
Publicado: (2025)
por: Zhou, Kanglei, et al.
Publicado: (2025)
CLIP-AGIQA: Boosting the Performance of AI-Generated Image Quality Assessment with CLIP
por: Tang, Zhenchen, et al.
Publicado: (2024)
por: Tang, Zhenchen, et al.
Publicado: (2024)
Exploring Textual Semantics Diversity for Image Transmission in Semantic Communication Systems using Visual Language Model
por: Huang, Peishan, et al.
Publicado: (2025)
por: Huang, Peishan, et al.
Publicado: (2025)
CLIP-Driven Semantic Discovery Network for Visible-Infrared Person Re-Identification
por: Yu, Xiaoyan, et al.
Publicado: (2024)
por: Yu, Xiaoyan, et al.
Publicado: (2024)
TSalV360: A Method and Dataset for Text-driven Saliency Detection in 360-Degrees Videos
por: Kontostathis, Ioannis, et al.
Publicado: (2025)
por: Kontostathis, Ioannis, et al.
Publicado: (2025)
PathoSCOPE: Few-Shot Pathology Detection via Self-Supervised Contrastive Learning and Pathology-Informed Synthetic Embeddings
por: Chin, Sinchee, et al.
Publicado: (2025)
por: Chin, Sinchee, et al.
Publicado: (2025)
CLIP-VG: Self-paced Curriculum Adapting of CLIP for Visual Grounding
por: Xiao, Linhui, et al.
Publicado: (2023)
por: Xiao, Linhui, et al.
Publicado: (2023)
PanoDreamer: Consistent Text to 360-Degree Scene Generation
por: Xiong, Zhexiao, et al.
Publicado: (2025)
por: Xiong, Zhexiao, et al.
Publicado: (2025)
CLIP-MUSED: CLIP-Guided Multi-Subject Visual Neural Information Semantic Decoding
por: Zhou, Qiongyi, et al.
Publicado: (2024)
por: Zhou, Qiongyi, et al.
Publicado: (2024)
ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation
por: Wang, Jingyun, et al.
Publicado: (2024)
por: Wang, Jingyun, et al.
Publicado: (2024)
Rethinking Visual Content Refinement in Low-Shot CLIP Adaptation
por: Lu, Jinda, et al.
Publicado: (2024)
por: Lu, Jinda, et al.
Publicado: (2024)
MIGA: Mutual Information-Guided Attack on Denoising Models for Semantic Manipulation
por: Li, Guanghao, et al.
Publicado: (2025)
por: Li, Guanghao, et al.
Publicado: (2025)
DiCLIP: Diffusion Model Enhances CLIP's Dense Knowledge for Weakly Supervised Semantic Segmentation
por: Yang, Zhiwei, et al.
Publicado: (2026)
por: Yang, Zhiwei, et al.
Publicado: (2026)
Free360: Layered Gaussian Splatting for Unbounded 360-Degree View Synthesis from Extremely Sparse and Unposed Views
por: Bao, Chong, et al.
Publicado: (2025)
por: Bao, Chong, et al.
Publicado: (2025)
Show or Tell? A Benchmark To Evaluate Visual and Textual Prompts in Semantic Segmentation
por: Rosi, Gabriele, et al.
Publicado: (2025)
por: Rosi, Gabriele, et al.
Publicado: (2025)
Thinking in 360°: Humanoid Visual Search in the Wild
por: Yu, Heyang, et al.
Publicado: (2025)
por: Yu, Heyang, et al.
Publicado: (2025)
GazeTarget360: Towards Gaze Target Estimation in 360-Degree for Robot Perception
por: Dai, Zhuangzhuang, et al.
Publicado: (2025)
por: Dai, Zhuangzhuang, et al.
Publicado: (2025)
Pose-Free Omnidirectional Gaussian Splatting for 360-Degree Videos with Consistent Depth Priors
por: Zhuang, Chuanqing, et al.
Publicado: (2026)
por: Zhuang, Chuanqing, et al.
Publicado: (2026)
ACT360: An Efficient 360-Degree Action Detection and Summarization Framework for Mission-Critical Training and Debriefing
por: Tiwari, Aditi, et al.
Publicado: (2025)
por: Tiwari, Aditi, et al.
Publicado: (2025)
CLIP-SENet: CLIP-based Semantic Enhancement Network for Vehicle Re-identification
por: Lu, Liping, et al.
Publicado: (2025)
por: Lu, Liping, et al.
Publicado: (2025)
WP-CLIP: Leveraging CLIP to Predict Wölfflin's Principles in Visual Art
por: Ghildyal, Abhijay, et al.
Publicado: (2025)
por: Ghildyal, Abhijay, et al.
Publicado: (2025)
Improving Visual Discriminability of CLIP for Training-Free Open-Vocabulary Semantic Segmentation
por: Zhou, Jinxin, et al.
Publicado: (2025)
por: Zhou, Jinxin, et al.
Publicado: (2025)
EdgeRelight360: Text-Conditioned 360-Degree HDR Image Generation for Real-Time On-Device Video Portrait Relighting
por: Lin, Min-Hui, et al.
Publicado: (2024)
por: Lin, Min-Hui, et al.
Publicado: (2024)
Improving Geometric Consistency for 360-Degree Neural Radiance Fields in Indoor Scenarios
por: Repinetska, Iryna, et al.
Publicado: (2025)
por: Repinetska, Iryna, et al.
Publicado: (2025)
SE360: Semantic Edit in 360$^\circ$ Panoramas via Hierarchical Data Construction
por: Zhong, Haoyi, et al.
Publicado: (2025)
por: Zhong, Haoyi, et al.
Publicado: (2025)
Imagine360: Immersive 360 Video Generation from Perspective Anchor
por: Tan, Jing, et al.
Publicado: (2024)
por: Tan, Jing, et al.
Publicado: (2024)
Ejemplares similares
-
360PanT: Training-Free Text-Driven 360-Degree Panorama-to-Panorama Translation
por: Wang, Hai, et al.
Publicado: (2024) -
A Survey on Text-Driven 360-Degree Panorama Generation
por: Wang, Hai, et al.
Publicado: (2025) -
360DVO: Deep Visual Odometry for Monocular 360-Degree Camera
por: Guo, Xiaopeng, et al.
Publicado: (2026) -
Grad-ECLIP: Gradient-based Visual and Textual Explanations for CLIP
por: Zhao, Chenyang, et al.
Publicado: (2025) -
Harnessing Textual Semantic Priors for Knowledge Transfer and Refinement in CLIP-Driven Continual Learning
por: He, Lingfeng, et al.
Publicado: (2025)