The Shape of Sight: A Homological Framework for Unifying Visual Perception
Fuente:
arXiv
Salvato in:
| Autore principale: | Li, Xin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2018
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CauSight: Learning to Supersense for Visual Causal Discovery
di: Zhang, Yize, et al.
Pubblicazione: (2025)
di: Zhang, Yize, et al.
Pubblicazione: (2025)
AV-Unified: A Unified Framework for Audio-visual Scene Understanding
di: Li, Guangyao, et al.
Pubblicazione: (2026)
di: Li, Guangyao, et al.
Pubblicazione: (2026)
MITO: A Millimeter-Wave Dataset and Simulator for Non-Line-of-Sight Perception
di: Dodds, Laura, et al.
Pubblicazione: (2025)
di: Dodds, Laura, et al.
Pubblicazione: (2025)
What is Memory? A Homological Perspective
di: Li, Xin
Pubblicazione: (2023)
di: Li, Xin
Pubblicazione: (2023)
Panther: Illuminate the Sight of Multimodal LLMs with Instruction-Guided Visual Prompts
di: Li, Honglin, et al.
Pubblicazione: (2024)
di: Li, Honglin, et al.
Pubblicazione: (2024)
Adaptive Perception for Unified Visual Multi-modal Object Tracking
di: Hu, Xiantao, et al.
Pubblicazione: (2025)
di: Hu, Xiantao, et al.
Pubblicazione: (2025)
Hidden in Plain Sight: Evaluating Abstract Shape Recognition in Vision-Language Models
di: Hemmat, Arshia, et al.
Pubblicazione: (2024)
di: Hemmat, Arshia, et al.
Pubblicazione: (2024)
EchoSight: Advancing Visual-Language Models with Wiki Knowledge
di: Yan, Yibin, et al.
Pubblicazione: (2024)
di: Yan, Yibin, et al.
Pubblicazione: (2024)
Bridging Generative and Discriminative Models for Unified Visual Perception with Diffusion Priors
di: Dong, Shiyin, et al.
Pubblicazione: (2024)
di: Dong, Shiyin, et al.
Pubblicazione: (2024)
PreSight: Enhancing Autonomous Vehicle Perception with City-Scale NeRF Priors
di: Yuan, Tianyuan, et al.
Pubblicazione: (2024)
di: Yuan, Tianyuan, et al.
Pubblicazione: (2024)
Antidote: A Unified Framework for Mitigating LVLM Hallucinations in Counterfactual Presupposition and Object Perception
di: Wu, Yuanchen, et al.
Pubblicazione: (2025)
di: Wu, Yuanchen, et al.
Pubblicazione: (2025)
A Unified and Controllable Framework for Layered Image Generation with Visual Effects
di: Yang, Jinrui, et al.
Pubblicazione: (2026)
di: Yang, Jinrui, et al.
Pubblicazione: (2026)
UFO: A Unified Approach to Fine-grained Visual Perception via Open-ended Language Interface
di: Tang, Hao, et al.
Pubblicazione: (2025)
di: Tang, Hao, et al.
Pubblicazione: (2025)
UniModel: A Visual-Only Framework for Unified Multimodal Understanding and Generation
di: Zhang, Chi, et al.
Pubblicazione: (2025)
di: Zhang, Chi, et al.
Pubblicazione: (2025)
Knowledge to Sight: Reasoning over Visual Attributes via Knowledge Decomposition for Abnormality Grounding
di: Li, Jun, et al.
Pubblicazione: (2025)
di: Li, Jun, et al.
Pubblicazione: (2025)
UniVision: A Unified Framework for Vision-Centric 3D Perception
di: Hong, Yu, et al.
Pubblicazione: (2024)
di: Hong, Yu, et al.
Pubblicazione: (2024)
VisionReasoner: Unified Reasoning-Integrated Visual Perception via Reinforcement Learning
di: Liu, Yuqi, et al.
Pubblicazione: (2025)
di: Liu, Yuqi, et al.
Pubblicazione: (2025)
Hidden Meanings in Plain Sight: RebusBench for Evaluating Cognitive Visual Reasoning
di: Kasaei, Seyed Amir, et al.
Pubblicazione: (2026)
di: Kasaei, Seyed Amir, et al.
Pubblicazione: (2026)
FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving
di: Zeng, Shuang, et al.
Pubblicazione: (2025)
di: Zeng, Shuang, et al.
Pubblicazione: (2025)
Sight Over Site: Perception-Aware Reinforcement Learning for Efficient Robotic Inspection
di: Kuhlmann, Richard, et al.
Pubblicazione: (2025)
di: Kuhlmann, Richard, et al.
Pubblicazione: (2025)
MMTL-UniAD: A Unified Framework for Multimodal and Multi-Task Learning in Assistive Driving Perception
di: Liu, Wenzhuo, et al.
Pubblicazione: (2025)
di: Liu, Wenzhuo, et al.
Pubblicazione: (2025)
MergeMix: A Unified Augmentation Paradigm for Visual and Multi-Modal Understanding
di: Jin, Xin, et al.
Pubblicazione: (2025)
di: Jin, Xin, et al.
Pubblicazione: (2025)
ImagiDrive: A Unified Imagination-and-Planning Framework for Autonomous Driving
di: Li, Jingyu, et al.
Pubblicazione: (2025)
di: Li, Jingyu, et al.
Pubblicazione: (2025)
A Unified 3D Object Perception Framework for Real-Time Outside-In Multi-Camera Systems
di: Wang, Yizhou, et al.
Pubblicazione: (2026)
di: Wang, Yizhou, et al.
Pubblicazione: (2026)
SuperEx: Enhancing Indoor Mapping and Exploration using Non-Line-of-Sight Perception
di: Garg, Kush, et al.
Pubblicazione: (2025)
di: Garg, Kush, et al.
Pubblicazione: (2025)
A Unified Framework for Semi-Supervised Image Segmentation and Registration
di: Li, Ruizhe, et al.
Pubblicazione: (2025)
di: Li, Ruizhe, et al.
Pubblicazione: (2025)
UniDGF: A Unified Detection-to-Generation Framework for Hierarchical Object Visual Recognition
di: Nan, Xinyu, et al.
Pubblicazione: (2025)
di: Nan, Xinyu, et al.
Pubblicazione: (2025)
Visual Bridge: Universal Visual Perception Representations Generating
di: Gao, Yilin, et al.
Pubblicazione: (2025)
di: Gao, Yilin, et al.
Pubblicazione: (2025)
Active Visual Perception: Opportunities and Challenges
di: Li, Yian, et al.
Pubblicazione: (2025)
di: Li, Yian, et al.
Pubblicazione: (2025)
VL-UniTrack: A Unified Framework with Visual-Language Prompts for UAV-Ground Visual Tracking
di: Xu, Boyue, et al.
Pubblicazione: (2026)
di: Xu, Boyue, et al.
Pubblicazione: (2026)
OwlSight: A Robust Illumination Adaptation Framework for Dark Video Human Action Recognition
di: Cheng, Shihao, et al.
Pubblicazione: (2025)
di: Cheng, Shihao, et al.
Pubblicazione: (2025)
COOPER: A Unified Model for Cooperative Perception and Reasoning in Spatial Intelligence
di: Zhang, Zefeng, et al.
Pubblicazione: (2025)
di: Zhang, Zefeng, et al.
Pubblicazione: (2025)
StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion
di: Tao, Ming, et al.
Pubblicazione: (2024)
di: Tao, Ming, et al.
Pubblicazione: (2024)
Aligning Generative Denoising with Discriminative Objectives Unleashes Diffusion for Visual Perception
di: Pang, Ziqi, et al.
Pubblicazione: (2025)
di: Pang, Ziqi, et al.
Pubblicazione: (2025)
CoopDETR: A Unified Cooperative Perception Framework for 3D Detection via Object Query
di: Wang, Zhe, et al.
Pubblicazione: (2025)
di: Wang, Zhe, et al.
Pubblicazione: (2025)
StableIdentity: Inserting Anybody into Anywhere at First Sight
di: Wang, Qinghe, et al.
Pubblicazione: (2024)
di: Wang, Qinghe, et al.
Pubblicazione: (2024)
A Unified Framework for 3D Scene Understanding
di: Xu, Wei, et al.
Pubblicazione: (2024)
di: Xu, Wei, et al.
Pubblicazione: (2024)
UV-M3TL: A Unified and Versatile Multimodal Multi-Task Learning Framework for Assistive Driving Perception
di: Liu, Wenzhuo, et al.
Pubblicazione: (2026)
di: Liu, Wenzhuo, et al.
Pubblicazione: (2026)
PatchEAD: Unifying Industrial Visual Prompting Frameworks for Patch-Exclusive Anomaly Detection
di: Huang, Po-Han, et al.
Pubblicazione: (2025)
di: Huang, Po-Han, et al.
Pubblicazione: (2025)
Open-Text Aerial Detection: A Unified Framework For Aerial Visual Grounding And Detection
di: Wei, Guoting, et al.
Pubblicazione: (2026)
di: Wei, Guoting, et al.
Pubblicazione: (2026)
Documenti analoghi
-
CauSight: Learning to Supersense for Visual Causal Discovery
di: Zhang, Yize, et al.
Pubblicazione: (2025) -
AV-Unified: A Unified Framework for Audio-visual Scene Understanding
di: Li, Guangyao, et al.
Pubblicazione: (2026) -
MITO: A Millimeter-Wave Dataset and Simulator for Non-Line-of-Sight Perception
di: Dodds, Laura, et al.
Pubblicazione: (2025) -
What is Memory? A Homological Perspective
di: Li, Xin
Pubblicazione: (2023) -
Panther: Illuminate the Sight of Multimodal LLMs with Instruction-Guided Visual Prompts
di: Li, Honglin, et al.
Pubblicazione: (2024)