Interpretable Text-Guided Image Clustering via Iterative Search
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Bingchen, Mac Aodha, Oisin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Labeled Data Selection for Category Discovery
by: Zhao, Bingchen, et al.
Published: (2024)
by: Zhao, Bingchen, et al.
Published: (2024)
Vision Learners Meet Web Image-Text Pairs
by: Zhao, Bingchen, et al.
Published: (2023)
by: Zhao, Bingchen, et al.
Published: (2023)
Improving Semantic Correspondence with Viewpoint-Guided Spherical Maps
by: Mariotti, Octave, et al.
Published: (2023)
by: Mariotti, Octave, et al.
Published: (2023)
Representational Similarity via Interpretable Visual Concepts
by: Kondapaneni, Neehar, et al.
Published: (2025)
by: Kondapaneni, Neehar, et al.
Published: (2025)
SAOR: Single-View Articulated Object Reconstruction
by: Aygün, Mehmet, et al.
Published: (2023)
by: Aygün, Mehmet, et al.
Published: (2023)
MotionPhysics: Learnable Motion Distillation for Text-Guided Simulation
by: Wang, Miaowei, et al.
Published: (2026)
by: Wang, Miaowei, et al.
Published: (2026)
Less is More: Discovering Concise Network Explanations
by: Kondapaneni, Neehar, et al.
Published: (2024)
by: Kondapaneni, Neehar, et al.
Published: (2024)
Jamais Vu: Exposing the Generalization Gap in Supervised Semantic Correspondence
by: Mariotti, Octave, et al.
Published: (2025)
by: Mariotti, Octave, et al.
Published: (2025)
DepthCues: Evaluating Monocular Depth Perception in Large Vision Models
by: Danier, Duolikun, et al.
Published: (2024)
by: Danier, Duolikun, et al.
Published: (2024)
Enhancing 2D Representation Learning with a 3D Prior
by: Aygün, Mehmet, et al.
Published: (2024)
by: Aygün, Mehmet, et al.
Published: (2024)
Representational Difference Explanations
by: Kondapaneni, Neehar, et al.
Published: (2025)
by: Kondapaneni, Neehar, et al.
Published: (2025)
CleverBirds: A Multiple-Choice Benchmark for Fine-grained Human Knowledge Tracing
by: Bossemeyer, Leonie, et al.
Published: (2025)
by: Bossemeyer, Leonie, et al.
Published: (2025)
Generating Binary Species Range Maps
by: Dorm, Filip, et al.
Published: (2024)
by: Dorm, Filip, et al.
Published: (2024)
Click to Grasp: Zero-Shot Precise Manipulation via Visual Diffusion Descriptors
by: Tsagkas, Nikolaos, et al.
Published: (2024)
by: Tsagkas, Nikolaos, et al.
Published: (2024)
WildSAT: Learning Satellite Image Representations from Wildlife Observations
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
View-Consistent Diffusion Representations for 3D-Consistent Video Generation
by: Danier, Duolikun, et al.
Published: (2025)
by: Danier, Duolikun, et al.
Published: (2025)
BiMotion: B-spline Motion for Text-guided Dynamic 3D Character Generation
by: Wang, Miaowei, et al.
Published: (2026)
by: Wang, Miaowei, et al.
Published: (2026)
AirPlanes: Accurate Plane Estimation via 3D-Consistent Embeddings
by: Watson, Jamie, et al.
Published: (2024)
by: Watson, Jamie, et al.
Published: (2024)
VesselSDF: Distance Field Priors for Vascular Network Reconstruction
by: Esposito, Salvatore, et al.
Published: (2025)
by: Esposito, Salvatore, et al.
Published: (2025)
INQUIRE: A Natural World Text-to-Image Retrieval Benchmark
by: Vendrow, Edward, et al.
Published: (2024)
by: Vendrow, Edward, et al.
Published: (2024)
CrossSDF: 3D Reconstruction of Thin Structures From Cross-Sections
by: Walker, Thomas, et al.
Published: (2024)
by: Walker, Thomas, et al.
Published: (2024)
Sample-efficient Integration of New Modalities into Large Language Models
by: İnce, Osman Batur, et al.
Published: (2025)
by: İnce, Osman Batur, et al.
Published: (2025)
Attentive Feature Aggregation or: How Policies Learn to Stop Worrying about Robustness and Attend to Task-Relevant Visual Cues
by: Tsagkas, Nikolaos, et al.
Published: (2025)
by: Tsagkas, Nikolaos, et al.
Published: (2025)
INQUIRE-Search: Interactive Discovery in Large-Scale Biodiversity Databases
by: Vendrow, Edward, et al.
Published: (2025)
by: Vendrow, Edward, et al.
Published: (2025)
The Temporal Trap: Entanglement in Pre-Trained Visual Representations for Visuomotor Policy Learning
by: Tsagkas, Nikolaos, et al.
Published: (2025)
by: Tsagkas, Nikolaos, et al.
Published: (2025)
Text-Guided Image Clustering
by: Stephan, Andreas, et al.
Published: (2024)
by: Stephan, Andreas, et al.
Published: (2024)
GeoGen: Geometry-Aware Generative Modeling via Signed Distance Functions
by: Esposito, Salvatore, et al.
Published: (2024)
by: Esposito, Salvatore, et al.
Published: (2024)
MVSAnywhere: Zero-Shot Multi-View Stereo
by: Izquierdo, Sergio, et al.
Published: (2025)
by: Izquierdo, Sergio, et al.
Published: (2025)
Text-Guided Alternative Image Clustering
by: Stephan, Andreas, et al.
Published: (2024)
by: Stephan, Andreas, et al.
Published: (2024)
Generalized Category Discovery under the Long-Tailed Distribution
by: Zhao, Bingchen, et al.
Published: (2025)
by: Zhao, Bingchen, et al.
Published: (2025)
Text Image Inpainting via Global Structure-Guided Diffusion Models
by: Zhu, Shipeng, et al.
Published: (2024)
by: Zhu, Shipeng, et al.
Published: (2024)
Feedforward Few-shot Species Range Estimation
by: Lange, Christian, et al.
Published: (2025)
by: Lange, Christian, et al.
Published: (2025)
Feature Aligning Few shot Learning Method Using Local Descriptors Weighted Rules
by: Yan, Bingchen
Published: (2024)
by: Yan, Bingchen
Published: (2024)
Iterative Prompt Refinement for Safer Text-to-Image Generation
by: Jeon, Jinwoo, et al.
Published: (2025)
by: Jeon, Jinwoo, et al.
Published: (2025)
"Principal Components" Enable A New Language of Images
by: Wen, Xin, et al.
Published: (2025)
by: Wen, Xin, et al.
Published: (2025)
Forgedit: Text Guided Image Editing via Learning and Forgetting
by: Zhang, Shiwen, et al.
Published: (2023)
by: Zhang, Shiwen, et al.
Published: (2023)
LiftVSR: Lifting Image Diffusion to Video Super-Resolution via Hybrid Temporal Modeling with Only 4$\times$RTX 4090s
by: Wang, Xijun, et al.
Published: (2025)
by: Wang, Xijun, et al.
Published: (2025)
DIR-TIR: Dialog-Iterative Refinement for Text-to-Image Retrieval
by: Zhen, Zongwei, et al.
Published: (2025)
by: Zhen, Zongwei, et al.
Published: (2025)
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
by: Gao, Xiang, et al.
Published: (2024)
by: Gao, Xiang, et al.
Published: (2024)
Describe-to-Score: Text-Guided Efficient Image Complexity Assessment
by: Liu, Shipeng, et al.
Published: (2025)
by: Liu, Shipeng, et al.
Published: (2025)
Similar Items
-
Labeled Data Selection for Category Discovery
by: Zhao, Bingchen, et al.
Published: (2024) -
Vision Learners Meet Web Image-Text Pairs
by: Zhao, Bingchen, et al.
Published: (2023) -
Improving Semantic Correspondence with Viewpoint-Guided Spherical Maps
by: Mariotti, Octave, et al.
Published: (2023) -
Representational Similarity via Interpretable Visual Concepts
by: Kondapaneni, Neehar, et al.
Published: (2025) -
SAOR: Single-View Articulated Object Reconstruction
by: Aygün, Mehmet, et al.
Published: (2023)