DiffCut: Catalyzing Zero-Shot Semantic Segmentation with Diffusion Features and Recursive Normalized Cut
Fuente:
arXiv
Salvato in:
| Autori principali: | Couairon, Paul, Shukor, Mustafa, Haugeard, Jean-Emmanuel, Cord, Matthieu, Thome, Nicolas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing
di: Couairon, Paul, et al.
Pubblicazione: (2023)
di: Couairon, Paul, et al.
Pubblicazione: (2023)
FreeSeg-Diff: Training-Free Open-Vocabulary Segmentation with Diffusion Models
di: Corradini, Barbara Toniella, et al.
Pubblicazione: (2024)
di: Corradini, Barbara Toniella, et al.
Pubblicazione: (2024)
JAFAR: Jack up Any Feature at Any Resolution
di: Couairon, Paul, et al.
Pubblicazione: (2025)
di: Couairon, Paul, et al.
Pubblicazione: (2025)
NAF: Zero-Shot Feature Upsampling via Neighborhood Attention Filtering
di: Chambon, Loick, et al.
Pubblicazione: (2025)
di: Chambon, Loick, et al.
Pubblicazione: (2025)
Skipping Computations in Multimodal LLMs
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
Improved Baselines for Data-efficient Perceptual Augmentation of LLMs
di: Vallaeys, Théophane, et al.
Pubblicazione: (2024)
di: Vallaeys, Théophane, et al.
Pubblicazione: (2024)
Beyond Task Performance: Evaluating and Reducing the Flaws of Large Multimodal Models with In-Context Learning
di: Shukor, Mustafa, et al.
Pubblicazione: (2023)
di: Shukor, Mustafa, et al.
Pubblicazione: (2023)
Unsupervised Segmentation by Diffusing, Walking and Cutting
di: Ivanova, Daniela, et al.
Pubblicazione: (2024)
di: Ivanova, Daniela, et al.
Pubblicazione: (2024)
Analyzing Finetuning Representation Shift for Multimodal LLMs Steering
di: Khayatan, Pegah, et al.
Pubblicazione: (2025)
di: Khayatan, Pegah, et al.
Pubblicazione: (2025)
What Makes Multimodal In-Context Learning Work?
di: Baldassini, Folco Bertini, et al.
Pubblicazione: (2024)
di: Baldassini, Folco Bertini, et al.
Pubblicazione: (2024)
Zero-Shot Refinement of Buildings' Segmentation Models using SAM
di: Mayladan, Ali, et al.
Pubblicazione: (2023)
di: Mayladan, Ali, et al.
Pubblicazione: (2023)
A Concept-Based Explainability Framework for Large Multimodal Models
di: Parekh, Jayneel, et al.
Pubblicazione: (2024)
di: Parekh, Jayneel, et al.
Pubblicazione: (2024)
CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation
di: Sick, Leon, et al.
Pubblicazione: (2024)
di: Sick, Leon, et al.
Pubblicazione: (2024)
Learning to Steer: Input-dependent Steering for Multimodal LLMs
di: Parekh, Jayneel, et al.
Pubblicazione: (2025)
di: Parekh, Jayneel, et al.
Pubblicazione: (2025)
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
di: Khayatan, Pegah, et al.
Pubblicazione: (2026)
di: Khayatan, Pegah, et al.
Pubblicazione: (2026)
Scaling Laws for Native Multimodal Models
di: Shukor, Mustafa, et al.
Pubblicazione: (2025)
di: Shukor, Mustafa, et al.
Pubblicazione: (2025)
ZeroDiff: Solidified Visual-Semantic Correlation in Zero-Shot Learning
di: Ye, Zihan, et al.
Pubblicazione: (2024)
di: Ye, Zihan, et al.
Pubblicazione: (2024)
Studying Image Diffusion Features for Zero-Shot Video Object Segmentation
di: Delatolas, Thanos, et al.
Pubblicazione: (2025)
di: Delatolas, Thanos, et al.
Pubblicazione: (2025)
ECAP: Extensive Cut-and-Paste Augmentation for Unsupervised Domain Adaptive Semantic Segmentation
di: Brorsson, Erik, et al.
Pubblicazione: (2024)
di: Brorsson, Erik, et al.
Pubblicazione: (2024)
Zero-Shot Video Semantic Segmentation based on Pre-Trained Diffusion Models
di: Wang, Qian, et al.
Pubblicazione: (2024)
di: Wang, Qian, et al.
Pubblicazione: (2024)
RefCut: Interactive Segmentation with Reference Guidance
di: Lin, Zheng, et al.
Pubblicazione: (2025)
di: Lin, Zheng, et al.
Pubblicazione: (2025)
Reliability in Semantic Segmentation: Can We Use Synthetic Data?
di: Loiseau, Thibaut, et al.
Pubblicazione: (2023)
di: Loiseau, Thibaut, et al.
Pubblicazione: (2023)
OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer
di: Wang, Boyang, et al.
Pubblicazione: (2026)
di: Wang, Boyang, et al.
Pubblicazione: (2026)
SSDD: Single-Step Diffusion Decoder for Efficient Image Tokenization
di: Vallaeys, Théophane, et al.
Pubblicazione: (2025)
di: Vallaeys, Théophane, et al.
Pubblicazione: (2025)
VoxelDiffusionCut: Non-destructive Internal-part Extraction via Iterative Cutting and Structure Estimation
di: Hachimine, Takumi, et al.
Pubblicazione: (2026)
di: Hachimine, Takumi, et al.
Pubblicazione: (2026)
Diff-SBSR: Learning Multimodal Feature-Enhanced Diffusion Models for Zero-Shot Sketch-Based 3D Shape Retrieval
di: Cheng, Hang, et al.
Pubblicazione: (2026)
di: Cheng, Hang, et al.
Pubblicazione: (2026)
Diffuse, Attend, and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion
di: Tian, Junjiao, et al.
Pubblicazione: (2023)
di: Tian, Junjiao, et al.
Pubblicazione: (2023)
DiffPortrait3D: Controllable Diffusion for Zero-Shot Portrait View Synthesis
di: Gu, Yuming, et al.
Pubblicazione: (2023)
di: Gu, Yuming, et al.
Pubblicazione: (2023)
DiffIR2VR-Zero: Zero-Shot Video Restoration with Diffusion-based Image Restoration Models
di: Yeh, Chang-Han, et al.
Pubblicazione: (2024)
di: Yeh, Chang-Han, et al.
Pubblicazione: (2024)
MeshSegmenter: Zero-Shot Mesh Semantic Segmentation via Texture Synthesis
di: Zhong, Ziming, et al.
Pubblicazione: (2024)
di: Zhong, Ziming, et al.
Pubblicazione: (2024)
SAM-guided Graph Cut for 3D Instance Segmentation
di: Guo, Haoyu, et al.
Pubblicazione: (2023)
di: Guo, Haoyu, et al.
Pubblicazione: (2023)
Falcon: Fractional Alternating Cut with Overcoming Minima in Unsupervised Segmentation
di: Zhang, Xiao, et al.
Pubblicazione: (2025)
di: Zhang, Xiao, et al.
Pubblicazione: (2025)
Evaluating the Efficacy of Cut-and-Paste Data Augmentation in Semantic Segmentation for Satellite Imagery
di: Motoi, Ionut M., et al.
Pubblicazione: (2024)
di: Motoi, Ionut M., et al.
Pubblicazione: (2024)
Conditional Latent Diffusion Models for Zero-Shot Instance Segmentation
di: Ulmer, Maximilian, et al.
Pubblicazione: (2025)
di: Ulmer, Maximilian, et al.
Pubblicazione: (2025)
MaskDiff: Modeling Mask Distribution with Diffusion Probabilistic Model for Few-Shot Instance Segmentation
di: Le, Minh-Quan, et al.
Pubblicazione: (2023)
di: Le, Minh-Quan, et al.
Pubblicazione: (2023)
Conterfactual Generative Zero-Shot Semantic Segmentation
di: Shen, Feihong, et al.
Pubblicazione: (2021)
di: Shen, Feihong, et al.
Pubblicazione: (2021)
VVitCutLER: Towards Unsupervised Object Detection and Segmentation in Videos
di: Lu, Zhijing, et al.
Pubblicazione: (2026)
di: Lu, Zhijing, et al.
Pubblicazione: (2026)
Language-Driven Visual Consensus for Zero-Shot Semantic Segmentation
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
Cut2Next: Generating Next Shot via In-Context Tuning
di: He, Jingwen, et al.
Pubblicazione: (2025)
di: He, Jingwen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing
di: Couairon, Paul, et al.
Pubblicazione: (2023) -
FreeSeg-Diff: Training-Free Open-Vocabulary Segmentation with Diffusion Models
di: Corradini, Barbara Toniella, et al.
Pubblicazione: (2024) -
JAFAR: Jack up Any Feature at Any Resolution
di: Couairon, Paul, et al.
Pubblicazione: (2025) -
NAF: Zero-Shot Feature Upsampling via Neighborhood Attention Filtering
di: Chambon, Loick, et al.
Pubblicazione: (2025) -
Skipping Computations in Multimodal LLMs
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)