Salvato in:
| Autori principali: | Li, Chenhao, Ono, Taishi, Uemori, Takeshi, Moriuchi, Yusuke |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2603.04817 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
NeISF++: Neural Incident Stokes Field for Polarized Inverse Rendering of Conductors and Dielectrics
di: Li, Chenhao, et al.
Pubblicazione: (2024)
di: Li, Chenhao, et al.
Pubblicazione: (2024)
Revisiting Active Learning in the Era of Vision Foundation Models
di: Gupte, Sanket Rajan, et al.
Pubblicazione: (2024)
di: Gupte, Sanket Rajan, et al.
Pubblicazione: (2024)
Revisiting Disparity from Dual-Pixel Images: Physics-Informed Lightweight Depth Estimation
di: Kurita, Teppei, et al.
Pubblicazione: (2024)
di: Kurita, Teppei, et al.
Pubblicazione: (2024)
Deep Polarization Cues for Single-shot Shape and Subsurface Scattering Estimation
di: Li, Chenhao, et al.
Pubblicazione: (2024)
di: Li, Chenhao, et al.
Pubblicazione: (2024)
Revisiting Model Stitching In the Foundation Model Era
di: Mai, Zheda, et al.
Pubblicazione: (2026)
di: Mai, Zheda, et al.
Pubblicazione: (2026)
Vision-and-Language Navigation Today and Tomorrow: A Survey in the Era of Foundation Models
di: Zhang, Yue, et al.
Pubblicazione: (2024)
di: Zhang, Yue, et al.
Pubblicazione: (2024)
Towards Unifying Understanding and Generation in the Era of Vision Foundation Models: A Survey from the Autoregression Perspective
di: Xie, Shenghao, et al.
Pubblicazione: (2024)
di: Xie, Shenghao, et al.
Pubblicazione: (2024)
Revisiting Automatic Data Curation for Vision Foundation Models in Digital Pathology
di: Chen, Boqi, et al.
Pubblicazione: (2025)
di: Chen, Boqi, et al.
Pubblicazione: (2025)
OmniCD: A Foundational Framework for Remote Sensing Image Change Detection Guided by Multimodal Semantics
di: Sun, Chenhao
Pubblicazione: (2026)
di: Sun, Chenhao
Pubblicazione: (2026)
Mixed-precision Supernet Training from Vision Foundation Models using Low Rank Adapter
di: Sakuma, Yuiko, et al.
Pubblicazione: (2024)
di: Sakuma, Yuiko, et al.
Pubblicazione: (2024)
Revisiting Vision Language Foundations for No-Reference Image Quality Assessment
di: Yadav, Ankit, et al.
Pubblicazione: (2025)
di: Yadav, Ankit, et al.
Pubblicazione: (2025)
HitoMi-Cam: A Shape-Agnostic Person Detection Method Using the Spectral Characteristics of Clothing
di: Ono, Shuji
Pubblicazione: (2025)
di: Ono, Shuji
Pubblicazione: (2025)
Image Segmentation in Foundation Model Era: A Survey
di: Zhou, Tianfei, et al.
Pubblicazione: (2024)
di: Zhou, Tianfei, et al.
Pubblicazione: (2024)
Zero-shot Shape Classification of Nanoparticles in SEM Images using Vision Foundation Models
di: Barnatan, Freida, et al.
Pubblicazione: (2025)
di: Barnatan, Freida, et al.
Pubblicazione: (2025)
Revisiting Referring Expression Comprehension Evaluation in the Era of Large Multimodal Models
di: Chen, Jierun, et al.
Pubblicazione: (2024)
di: Chen, Jierun, et al.
Pubblicazione: (2024)
In the Era of Prompt Learning with Vision-Language Models
di: Jha, Ankit
Pubblicazione: (2024)
di: Jha, Ankit
Pubblicazione: (2024)
Revisiting Prompt Pretraining of Vision-Language Models
di: Chen, Zhenyuan, et al.
Pubblicazione: (2024)
di: Chen, Zhenyuan, et al.
Pubblicazione: (2024)
Iterated Learning Improves Compositionality in Large Vision-Language Models
di: Zheng, Chenhao, et al.
Pubblicazione: (2024)
di: Zheng, Chenhao, et al.
Pubblicazione: (2024)
Online Long-term Point Tracking in the Foundation Model Era
di: Aydemir, Görkay
Pubblicazione: (2025)
di: Aydemir, Görkay
Pubblicazione: (2025)
ViTamin: Designing Scalable Vision Models in the Vision-Language Era
di: Chen, Jieneng, et al.
Pubblicazione: (2024)
di: Chen, Jieneng, et al.
Pubblicazione: (2024)
Reflection Removal Using Recurrent Polarization-to-Polarization Network
di: Bian, Wenjiao, et al.
Pubblicazione: (2024)
di: Bian, Wenjiao, et al.
Pubblicazione: (2024)
HSViT: Horizontally Scalable Vision Transformer
di: Xu, Chenhao, et al.
Pubblicazione: (2024)
di: Xu, Chenhao, et al.
Pubblicazione: (2024)
Revisiting Tampered Scene Text Detection in the Era of Generative AI
di: Qu, Chenfan, et al.
Pubblicazione: (2024)
di: Qu, Chenfan, et al.
Pubblicazione: (2024)
PLUG: Revisiting Amodal Segmentation with Foundation Model and Hierarchical Focus
di: Liu, Zhaochen, et al.
Pubblicazione: (2024)
di: Liu, Zhaochen, et al.
Pubblicazione: (2024)
A Novel Benchmark for Few-Shot Semantic Segmentation in the Era of Foundation Models
di: Bensaid, Reda, et al.
Pubblicazione: (2024)
di: Bensaid, Reda, et al.
Pubblicazione: (2024)
Surgical Scene Understanding in the Era of Foundation AI Models: A Comprehensive Review
di: Khan, Ufaq, et al.
Pubblicazione: (2025)
di: Khan, Ufaq, et al.
Pubblicazione: (2025)
Segmentation-Driven Monocular Shape from Polarization based on Physical Model
di: Zhang, Jinyu, et al.
Pubblicazione: (2026)
di: Zhang, Jinyu, et al.
Pubblicazione: (2026)
SuperPlace: The Renaissance of Classical Feature Aggregation for Visual Place Recognition in the Era of Foundation Models
di: Liu, Bingxi, et al.
Pubblicazione: (2025)
di: Liu, Bingxi, et al.
Pubblicazione: (2025)
ZeroSlide: Is Zero-Shot Classification Adequate for Lifelong Learning in Whole-Slide Image Analysis in the Era of Pathology Vision-Language Foundation Models?
di: Bui, Doanh C., et al.
Pubblicazione: (2025)
di: Bui, Doanh C., et al.
Pubblicazione: (2025)
Diversity Covariance-Aware Prompt Learning for Vision-Language Models
di: Dong, Songlin, et al.
Pubblicazione: (2025)
di: Dong, Songlin, et al.
Pubblicazione: (2025)
Revisiting Continual Semantic Segmentation with Pre-trained Vision Models
di: Zhang, Duzhen, et al.
Pubblicazione: (2025)
di: Zhang, Duzhen, et al.
Pubblicazione: (2025)
Are Vision Foundation Models Foundational for Electron Microscopy Image Segmentation?
di: Fuster-Barceló, Caterina, et al.
Pubblicazione: (2026)
di: Fuster-Barceló, Caterina, et al.
Pubblicazione: (2026)
Shape from Polarization of Thermal Emission and Reflection
di: Kitazawa, Kazuma, et al.
Pubblicazione: (2025)
di: Kitazawa, Kazuma, et al.
Pubblicazione: (2025)
Sapiens: Foundation for Human Vision Models
di: Khirodkar, Rawal, et al.
Pubblicazione: (2024)
di: Khirodkar, Rawal, et al.
Pubblicazione: (2024)
Revisiting Multimodal Positional Encoding in Vision-Language Models
di: Huang, Jie, et al.
Pubblicazione: (2025)
di: Huang, Jie, et al.
Pubblicazione: (2025)
Bootstrapping SparseFormers from Vision Foundation Models
di: Gao, Ziteng, et al.
Pubblicazione: (2023)
di: Gao, Ziteng, et al.
Pubblicazione: (2023)
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models
di: Guo, Jianyuan, et al.
Pubblicazione: (2024)
di: Guo, Jianyuan, et al.
Pubblicazione: (2024)
Revisiting Shadow Detection from a Vision-Language Perspective
di: Wang, Yonghui, et al.
Pubblicazione: (2026)
di: Wang, Yonghui, et al.
Pubblicazione: (2026)
PANORAMA: The Rise of Omnidirectional Vision in the Embodied AI Era
di: Zheng, Xu, et al.
Pubblicazione: (2025)
di: Zheng, Xu, et al.
Pubblicazione: (2025)
Fooling Polarization-based Vision using Locally Controllable Polarizing Projection
di: Li, Zhuoxiao, et al.
Pubblicazione: (2023)
di: Li, Zhuoxiao, et al.
Pubblicazione: (2023)
Documenti analoghi
-
NeISF++: Neural Incident Stokes Field for Polarized Inverse Rendering of Conductors and Dielectrics
di: Li, Chenhao, et al.
Pubblicazione: (2024) -
Revisiting Active Learning in the Era of Vision Foundation Models
di: Gupte, Sanket Rajan, et al.
Pubblicazione: (2024) -
Revisiting Disparity from Dual-Pixel Images: Physics-Informed Lightweight Depth Estimation
di: Kurita, Teppei, et al.
Pubblicazione: (2024) -
Deep Polarization Cues for Single-shot Shape and Subsurface Scattering Estimation
di: Li, Chenhao, et al.
Pubblicazione: (2024) -
Revisiting Model Stitching In the Foundation Model Era
di: Mai, Zheda, et al.
Pubblicazione: (2026)