Scaling up Multi-domain Semantic Segmentation with Sentence Embeddings
Fuente:
arXiv
Salvato in:
| Autori principali: | Yin, Wei, Liu, Yifan, Shen, Chunhua, Sun, Baichuan, Hengel, Anton van den |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ViewFusion: Towards Multi-View Consistency via Interpolated Denoising
di: Yang, Xianghui, et al.
Pubblicazione: (2024)
di: Yang, Xianghui, et al.
Pubblicazione: (2024)
Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors
di: Xia, Jiatong, et al.
Pubblicazione: (2026)
di: Xia, Jiatong, et al.
Pubblicazione: (2026)
Frame-wise Conditioning Adaptation for Fine-Tuning Diffusion Models in Text-to-Video Prediction
di: Liu, Zheyuan, et al.
Pubblicazione: (2025)
di: Liu, Zheyuan, et al.
Pubblicazione: (2025)
Let Your Video Listen to Your Music!
di: Zhang, Xinyu, et al.
Pubblicazione: (2025)
di: Zhang, Xinyu, et al.
Pubblicazione: (2025)
Can You Learn to See Without Images? Procedural Warm-Up for Vision Transformers
di: Shinnick, Zachary, et al.
Pubblicazione: (2025)
di: Shinnick, Zachary, et al.
Pubblicazione: (2025)
Math Blind: Failures in Diagram Understanding Undermine Reasoning in MLLMs
di: Sun, Yanpeng, et al.
Pubblicazione: (2025)
di: Sun, Yanpeng, et al.
Pubblicazione: (2025)
Hierarchical Process Reward Models are Symbolic Vision Learners
di: Zhang, Shan, et al.
Pubblicazione: (2025)
di: Zhang, Shan, et al.
Pubblicazione: (2025)
Knowledge Composition using Task Vectors with Learned Anisotropic Scaling
di: Zhang, Frederic Z., et al.
Pubblicazione: (2024)
di: Zhang, Frederic Z., et al.
Pubblicazione: (2024)
Source-Free Unsupervised Domain Adaptation with Hypothesis Consolidation of Prediction Rationale
di: Shu, Yangyang, et al.
Pubblicazione: (2024)
di: Shu, Yangyang, et al.
Pubblicazione: (2024)
Premonition: Using Generative Models to Preempt Future Data Changes in Continual Learning
di: McDonnell, Mark D., et al.
Pubblicazione: (2024)
di: McDonnell, Mark D., et al.
Pubblicazione: (2024)
Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation
di: Zhu, Muzhi, et al.
Pubblicazione: (2024)
di: Zhu, Muzhi, et al.
Pubblicazione: (2024)
DiffuMask: Synthesizing Images with Pixel-level Annotations for Semantic Segmentation Using Diffusion Models
di: Wu, Weijia, et al.
Pubblicazione: (2023)
di: Wu, Weijia, et al.
Pubblicazione: (2023)
Continual Learning on CLIP via Incremental Prompt Tuning with Intrinsic Textual Anchors
di: Lu, Haodong, et al.
Pubblicazione: (2025)
di: Lu, Haodong, et al.
Pubblicazione: (2025)
Unified Open-World Segmentation with Multi-Modal Prompts
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
Length Matters: Length-Aware Transformer for Temporal Sentence Grounding
di: Wang, Yifan, et al.
Pubblicazione: (2025)
di: Wang, Yifan, et al.
Pubblicazione: (2025)
The Devil is in the Distributions: Explicit Modeling of Scene Content is Key in Zero-Shot Video Captioning
di: Tian, Mingkai, et al.
Pubblicazione: (2025)
di: Tian, Mingkai, et al.
Pubblicazione: (2025)
Towards Higher Effective Rank in Parameter-efficient Fine-tuning using Khatri--Rao Product
di: Albert, Paul, et al.
Pubblicazione: (2025)
di: Albert, Paul, et al.
Pubblicazione: (2025)
Multi-Scale Semantic Segmentation with Modified MBConv Blocks
di: Chen, Xi, et al.
Pubblicazione: (2024)
di: Chen, Xi, et al.
Pubblicazione: (2024)
RanPAC: Random Projections and Pre-trained Models for Continual Learning
di: McDonnell, Mark D., et al.
Pubblicazione: (2023)
di: McDonnell, Mark D., et al.
Pubblicazione: (2023)
Probabilistic Modeling of Multi-rater Medical Image Segmentation for Diversity and Personalization
di: Liu, Ke, et al.
Pubblicazione: (2025)
di: Liu, Ke, et al.
Pubblicazione: (2025)
HERO: Hierarchical Embedding-Refinement for Open-Vocabulary Temporal Sentence Grounding in Videos
di: Han, Tingting, et al.
Pubblicazione: (2026)
di: Han, Tingting, et al.
Pubblicazione: (2026)
Open Eyes, Then Reason: Fine-grained Visual Mathematical Understanding in MLLMs
di: Zhang, Shan, et al.
Pubblicazione: (2025)
di: Zhang, Shan, et al.
Pubblicazione: (2025)
Multi-Scale Representations by Varying Window Attention for Semantic Segmentation
di: Yan, Haotian, et al.
Pubblicazione: (2024)
di: Yan, Haotian, et al.
Pubblicazione: (2024)
Optimized Unet with Attention Mechanism for Multi-Scale Semantic Segmentation
di: Li, Xuan, et al.
Pubblicazione: (2025)
di: Li, Xuan, et al.
Pubblicazione: (2025)
Embedding Generalized Semantic Knowledge into Few-Shot Remote Sensing Segmentation
di: Jia, Yuyu, et al.
Pubblicazione: (2024)
di: Jia, Yuyu, et al.
Pubblicazione: (2024)
Prosody-Enhanced Acoustic Pre-training and Acoustic-Disentangled Prosody Adapting for Movie Dubbing
di: Zhang, Zhedong, et al.
Pubblicazione: (2025)
di: Zhang, Zhedong, et al.
Pubblicazione: (2025)
ProgRoCC: A Progressive Approach to Rough Crowd Counting
di: Jiang, Shengqin, et al.
Pubblicazione: (2025)
di: Jiang, Shengqin, et al.
Pubblicazione: (2025)
On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
di: Cai, Yichao, et al.
Pubblicazione: (2025)
di: Cai, Yichao, et al.
Pubblicazione: (2025)
Bridging Granularity Gaps: Hierarchical Semantic Learning for Cross-domain Few-shot Segmentation
di: Sun, Sujun, et al.
Pubblicazione: (2025)
di: Sun, Sujun, et al.
Pubblicazione: (2025)
Robust Visual Localization via Semantic-Guided Multi-Scale Transformer
di: Tian, Zhongtao, et al.
Pubblicazione: (2025)
di: Tian, Zhongtao, et al.
Pubblicazione: (2025)
Open-Vocabulary Semantic Segmentation with Image Embedding Balancing
di: Shan, Xiangheng, et al.
Pubblicazione: (2024)
di: Shan, Xiangheng, et al.
Pubblicazione: (2024)
FlowDubber: Movie Dubbing with LLM-based Semantic-aware Learning and Flow Matching based Voice Enhancing
di: Cong, Gaoxiang, et al.
Pubblicazione: (2025)
di: Cong, Gaoxiang, et al.
Pubblicazione: (2025)
USE: Universal Segment Embeddings for Open-Vocabulary Image Segmentation
di: Wang, Xiaoqi, et al.
Pubblicazione: (2024)
di: Wang, Xiaoqi, et al.
Pubblicazione: (2024)
SSP-SAM: SAM with Semantic-Spatial Prompt for Referring Expression Segmentation
di: Tang, Wei, et al.
Pubblicazione: (2026)
di: Tang, Wei, et al.
Pubblicazione: (2026)
Reducing Unimodal Bias in Multi-Modal Semantic Segmentation with Multi-Scale Functional Entropy Regularization
di: Zheng, Xu, et al.
Pubblicazione: (2025)
di: Zheng, Xu, et al.
Pubblicazione: (2025)
Hard-aware Instance Adaptive Self-training for Unsupervised Cross-domain Semantic Segmentation
di: Zhu, Chuang, et al.
Pubblicazione: (2023)
di: Zhu, Chuang, et al.
Pubblicazione: (2023)
RandLoRA: Full-rank parameter-efficient fine-tuning of large models
di: Albert, Paul, et al.
Pubblicazione: (2025)
di: Albert, Paul, et al.
Pubblicazione: (2025)
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
di: Liu, Yang, et al.
Pubblicazione: (2023)
di: Liu, Yang, et al.
Pubblicazione: (2023)
Domain-invariant Prototypes for Semantic Segmentation
di: Yang, Zhengeng, et al.
Pubblicazione: (2022)
di: Yang, Zhengeng, et al.
Pubblicazione: (2022)
Multi-Scale Grouped Prototypes for Interpretable Semantic Segmentation
di: Porta, Hugo, et al.
Pubblicazione: (2024)
di: Porta, Hugo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ViewFusion: Towards Multi-View Consistency via Interpolated Denoising
di: Yang, Xianghui, et al.
Pubblicazione: (2024) -
Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors
di: Xia, Jiatong, et al.
Pubblicazione: (2026) -
Frame-wise Conditioning Adaptation for Fine-Tuning Diffusion Models in Text-to-Video Prediction
di: Liu, Zheyuan, et al.
Pubblicazione: (2025) -
Let Your Video Listen to Your Music!
di: Zhang, Xinyu, et al.
Pubblicazione: (2025) -
Can You Learn to See Without Images? Procedural Warm-Up for Vision Transformers
di: Shinnick, Zachary, et al.
Pubblicazione: (2025)