VISTA: A Visual Analytics Framework to Enhance Foundation Model-Generated Data Labels
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xuan, Xiwei, Wang, Xiaoqi, He, Wenbin, Ono, Jorge Piazentin, Gou, Liang, Ma, Kwan-Liu, Ren, Liu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AttributionScanner: A Visual Analytics System for Model Validation with Metadata-Free Slice Finding
von: Xuan, Xiwei, et al.
Veröffentlicht: (2024)
von: Xuan, Xiwei, et al.
Veröffentlicht: (2024)
USE: Universal Segment Embeddings for Open-Vocabulary Image Segmentation
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2024)
ReME: A Data-Centric Framework for Training-Free Open-Vocabulary Segmentation
von: Xuan, Xiwei, et al.
Veröffentlicht: (2025)
von: Xuan, Xiwei, et al.
Veröffentlicht: (2025)
VISLIX: An XAI Framework for Validating Vision Models with Slice Discovery and Analysis
von: Yan, Xinyuan, et al.
Veröffentlicht: (2025)
von: Yan, Xinyuan, et al.
Veröffentlicht: (2025)
ProSAM: Enhancing the Robustness of SAM-based Visual Reference Segmentation with Probabilistic Prompts
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2025)
SUNY: A Visual Interpretation Framework for Convolutional Neural Networks from a Necessary and Sufficient Perspective
von: Xuan, Xiwei, et al.
Veröffentlicht: (2023)
von: Xuan, Xiwei, et al.
Veröffentlicht: (2023)
SLIM: Spuriousness Mitigation with Minimal Human Annotations
von: Xuan, Xiwei, et al.
Veröffentlicht: (2024)
von: Xuan, Xiwei, et al.
Veröffentlicht: (2024)
Hyp-OW: Exploiting Hierarchical Structure Learning with Hyperbolic Distance Enhances Open World Object Detection
von: Doan, Thang, et al.
Veröffentlicht: (2023)
von: Doan, Thang, et al.
Veröffentlicht: (2023)
InterChat: Enhancing Generative Visual Analytics using Multimodal Interactions
von: Chen, Juntong, et al.
Veröffentlicht: (2025)
von: Chen, Juntong, et al.
Veröffentlicht: (2025)
InterChat: Enhancing Generative Visual Analytics using Multimodal Interactions
von: Juntong Chen, et al.
Veröffentlicht: (2025)
von: Juntong Chen, et al.
Veröffentlicht: (2025)
A streamlined Approach to Multimodal Few-Shot Class Incremental Learning for Fine-Grained Datasets
von: Doan, Thang, et al.
Veröffentlicht: (2024)
von: Doan, Thang, et al.
Veröffentlicht: (2024)
InterVLS: Interactive Model Understanding and Improvement with Vision-Language Surrogates
von: Huang, Jinbin, et al.
Veröffentlicht: (2023)
von: Huang, Jinbin, et al.
Veröffentlicht: (2023)
DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models
von: Pan, Chenbin, et al.
Veröffentlicht: (2025)
von: Pan, Chenbin, et al.
Veröffentlicht: (2025)
VISLIX: An XAI Framework for Validating Vision Models with Slice Discovery and Analysis
von: Xinyuan Yan, et al.
Veröffentlicht: (2025)
von: Xinyuan Yan, et al.
Veröffentlicht: (2025)
ViT-Split: Unleashing the Power of Vision Foundation Models via Efficient Splitting Heads
von: Li, Yifan, et al.
Veröffentlicht: (2025)
von: Li, Yifan, et al.
Veröffentlicht: (2025)
VISTA: A Test-Time Self-Improving Video Generation Agent
von: Long, Do Xuan, et al.
Veröffentlicht: (2025)
von: Long, Do Xuan, et al.
Veröffentlicht: (2025)
Investigating Interaction Modes and User Agency in Human-LLM Collaboration for Domain-Specific Data Analysis
von: Guo, Jiajing, et al.
Veröffentlicht: (2024)
von: Guo, Jiajing, et al.
Veröffentlicht: (2024)
VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by Video Spatiotemporal Augmentation
von: Ren, Weiming, et al.
Veröffentlicht: (2024)
von: Ren, Weiming, et al.
Veröffentlicht: (2024)
Rethinking Agentic Workflows: Evaluating Inference-Based Test-Time Scaling Strategies in Text2SQL Tasks
von: Guo, Jiajing, et al.
Veröffentlicht: (2025)
von: Guo, Jiajing, et al.
Veröffentlicht: (2025)
ELIP: Enhanced Visual-Language Foundation Models for Image Retrieval
von: Zhan, Guanqi, et al.
Veröffentlicht: (2025)
von: Zhan, Guanqi, et al.
Veröffentlicht: (2025)
Human Uncertainty-Aware Data Selection and Automatic Labeling in Visual Question Answering
von: Lan, Jian, et al.
Veröffentlicht: (2025)
von: Lan, Jian, et al.
Veröffentlicht: (2025)
3DWG: 3D Weakly Supervised Visual Grounding via Category and Instance-Level Alignment
von: Li, Xiaoqi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoqi, et al.
Veröffentlicht: (2025)
A Sensor Agnostic Domain Generalization Framework for Leveraging Geospatial Foundation Models: Enhancing Semantic Segmentation viaSynergistic Pseudo-Labeling and Generative Learning
von: Yaghmour, Anan, et al.
Veröffentlicht: (2025)
von: Yaghmour, Anan, et al.
Veröffentlicht: (2025)
VISTA: Visualized Text Embedding For Universal Multi-Modal Retrieval
von: Zhou, Junjie, et al.
Veröffentlicht: (2024)
von: Zhou, Junjie, et al.
Veröffentlicht: (2024)
VISTA3D: A Unified Segmentation Foundation Model For 3D Medical Imaging
von: He, Yufan, et al.
Veröffentlicht: (2024)
von: He, Yufan, et al.
Veröffentlicht: (2024)
Sparse Generation: Making Pseudo Labels Sparse for Point Weakly Supervised Object Detection on Low Data Volume
von: Shang, Chuyang, et al.
Veröffentlicht: (2024)
von: Shang, Chuyang, et al.
Veröffentlicht: (2024)
FIAS: Feature Imbalance-Aware Medical Image Segmentation with Dynamic Fusion and Mixing Attention
von: Liu, Xiwei, et al.
Veröffentlicht: (2024)
von: Liu, Xiwei, et al.
Veröffentlicht: (2024)
Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning
von: Gong, Xuan, et al.
Veröffentlicht: (2026)
von: Gong, Xuan, et al.
Veröffentlicht: (2026)
Satellite-Free Training for Drone-View Geo-Localization
von: Liu, Tao, et al.
Veröffentlicht: (2026)
von: Liu, Tao, et al.
Veröffentlicht: (2026)
MD-Face: MoE-Enhanced Label-Free Disentangled Representation for Interactive Facial Attribute Editing
von: Cui, Xuan, et al.
Veröffentlicht: (2026)
von: Cui, Xuan, et al.
Veröffentlicht: (2026)
Online Descriptor Enhancement via Self-Labelling Triplets for Visual Data Association
von: Shaoul, Yorai, et al.
Veröffentlicht: (2020)
von: Shaoul, Yorai, et al.
Veröffentlicht: (2020)
Incomplete Multi-Label Image Recognition by Co-learning Semantic-Aware Features and Label Recovery
von: He, Zhi-Fen, et al.
Veröffentlicht: (2025)
von: He, Zhi-Fen, et al.
Veröffentlicht: (2025)
Native Audio-Visual Alignment for Generation
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
A Unified and Controllable Framework for Layered Image Generation with Visual Effects
von: Yang, Jinrui, et al.
Veröffentlicht: (2026)
von: Yang, Jinrui, et al.
Veröffentlicht: (2026)
Integrating Reinforcement Learning with Visual Generative Models: Foundations and Advances
von: Liang, Yuanzhi, et al.
Veröffentlicht: (2025)
von: Liang, Yuanzhi, et al.
Veröffentlicht: (2025)
VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?
von: Liu, Qing'an, et al.
Veröffentlicht: (2026)
von: Liu, Qing'an, et al.
Veröffentlicht: (2026)
VISTA: A Visual and Textual Attention Dataset for Interpreting Multimodal Models
von: Harshit, et al.
Veröffentlicht: (2024)
von: Harshit, et al.
Veröffentlicht: (2024)
Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction
von: He, Jing, et al.
Veröffentlicht: (2024)
von: He, Jing, et al.
Veröffentlicht: (2024)
Towards Robust Visual Continual Learning with Multi-Prototype Supervision
von: Liu, Xiwei, et al.
Veröffentlicht: (2025)
von: Liu, Xiwei, et al.
Veröffentlicht: (2025)
Data Adaptive Few-shot Multi Label Segmentation with Foundation Model
von: Reddy, Gurunath, et al.
Veröffentlicht: (2024)
von: Reddy, Gurunath, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AttributionScanner: A Visual Analytics System for Model Validation with Metadata-Free Slice Finding
von: Xuan, Xiwei, et al.
Veröffentlicht: (2024) -
USE: Universal Segment Embeddings for Open-Vocabulary Image Segmentation
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2024) -
ReME: A Data-Centric Framework for Training-Free Open-Vocabulary Segmentation
von: Xuan, Xiwei, et al.
Veröffentlicht: (2025) -
VISLIX: An XAI Framework for Validating Vision Models with Slice Discovery and Analysis
von: Yan, Xinyuan, et al.
Veröffentlicht: (2025) -
ProSAM: Enhancing the Robustness of SAM-based Visual Reference Segmentation with Probabilistic Prompts
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2025)