ModalChorus: Visual Probing and Alignment of Multi-modal Embeddings via Modal Fusion Map
Fuente:
arXiv
Saved in:
| Main Authors: | Ye, Yilin, Xiao, Shishi, Zeng, Xingchen, Zeng, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AKRMap: Adaptive Kernel Regression for Trustworthy Visualization of Cross-Modal Embeddings
by: Ye, Yilin, et al.
Published: (2025)
by: Ye, Yilin, et al.
Published: (2025)
Three Modalities, Two Design Probes, One Prototype, and No Vision: Experience-Based Co-Design of a Multi-modal 3D Data Visualization Tool
by: Kamath, Sanchita S., et al.
Published: (2026)
by: Kamath, Sanchita S., et al.
Published: (2026)
Evaluating VisualRAG: Quantifying Cross-Modal Performance in Enterprise Document Understanding
by: Mannam, Varun, et al.
Published: (2025)
by: Mannam, Varun, et al.
Published: (2025)
PoultryTalk: A Multi-modal Retrieval-Augmented Generation (RAG) System for Intelligent Poultry Management and Decision Support
by: Khanal, Kapalik, et al.
Published: (2025)
by: Khanal, Kapalik, et al.
Published: (2025)
Using a negative spatial auto-correlation index to evaluate and improve intrinsic TagMap's multi-scale visualization capabilities
by: Wei, Zhiwei, et al.
Published: (2024)
by: Wei, Zhiwei, et al.
Published: (2024)
Orbit: A Framework for Designing and Evaluating Multi-objective Rankers
by: Yang, Chenyang, et al.
Published: (2024)
by: Yang, Chenyang, et al.
Published: (2024)
CLAS: A Machine Learning Enhanced Framework for Exploring Large 3D Design Datasets
by: Zhang, XiuYu, et al.
Published: (2024)
by: Zhang, XiuYu, et al.
Published: (2024)
Toward Safe and Human-Aligned Game Conversational Recommendation via Multi-Agent Decomposition
by: Hui, Zheng, et al.
Published: (2025)
by: Hui, Zheng, et al.
Published: (2025)
Search Timelines: Visualizing Search History to Enable Cross-Session Exploratory Search
by: Hoeber, Orland, et al.
Published: (2025)
by: Hoeber, Orland, et al.
Published: (2025)
Multi-TAP: Multi-criteria Target Adaptive Persona Modeling for Cross-Domain Recommendation
by: Kang, Daehee, et al.
Published: (2026)
by: Kang, Daehee, et al.
Published: (2026)
Artful Path to Healing: Using Machine Learning for Visual Art Recommendation to Prevent and Reduce Post-Intensive Care
by: Yilma, Bereket A., et al.
Published: (2024)
by: Yilma, Bereket A., et al.
Published: (2024)
IntentTuner: An Interactive Framework for Integrating Human Intents in Fine-tuning Text-to-Image Generative Models
by: Zeng, Xingchen, et al.
Published: (2024)
by: Zeng, Xingchen, et al.
Published: (2024)
Decoy Effect in Search Interaction: A Pilot Study
by: Chen, Nuo, et al.
Published: (2023)
by: Chen, Nuo, et al.
Published: (2023)
Improving Collaborative Filtering Recommendation via Graph Learning
by: Wang, Yongyu
Published: (2023)
by: Wang, Yongyu
Published: (2023)
ChartifyText: Automated Chart Generation from Data-Involved Texts via LLM
by: Zhang, Songheng, et al.
Published: (2024)
by: Zhang, Songheng, et al.
Published: (2024)
Network-based Topic Structure Visualization
by: Jeon, Yeseul, et al.
Published: (2024)
by: Jeon, Yeseul, et al.
Published: (2024)
InterDeepResearch: Enabling Human-Agent Collaborative Information Seeking through Interactive Deep Research
by: Pan, Bo, et al.
Published: (2026)
by: Pan, Bo, et al.
Published: (2026)
Positive-First Most Ambiguous: A Simple Active Learning Criterion for Interactive Retrieval of Rare Categories
by: Zaher, Kawtar, et al.
Published: (2026)
by: Zaher, Kawtar, et al.
Published: (2026)
Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers
by: Zaher, Kawtar, et al.
Published: (2026)
by: Zaher, Kawtar, et al.
Published: (2026)
ERPA: Efficient RPA Model Integrating OCR and LLMs for Intelligent Document Processing
by: Abdellaif, Osama, et al.
Published: (2024)
by: Abdellaif, Osama, et al.
Published: (2024)
Spacewalker: Traversing Representation Spaces for Fast Interactive Exploration and Annotation of Unstructured Data
by: Heine, Lukas, et al.
Published: (2024)
by: Heine, Lukas, et al.
Published: (2024)
A Versatile Dataset of Mouse and Eye Movements on Search Engine Results Pages
by: Latifzadeh, Kayhan, et al.
Published: (2025)
by: Latifzadeh, Kayhan, et al.
Published: (2025)
SymbioticRAG: Enhancing Document Intelligence Through Human-LLM Symbiotic Collaboration
by: Sun, Qiang, et al.
Published: (2025)
by: Sun, Qiang, et al.
Published: (2025)
Memento: Towards Proactive Visualization of Everyday Memories with Personal Wearable AR Assistant
by: Kim, Yoonsang, et al.
Published: (2026)
by: Kim, Yoonsang, et al.
Published: (2026)
Towards Coarse-grained Visual Language Navigation Task Planning Enhanced by Event Knowledge Graph
by: Kaichen, Zhao, et al.
Published: (2024)
by: Kaichen, Zhao, et al.
Published: (2024)
Learning to Rank for Maps at Airbnb
by: Haldar, Malay, et al.
Published: (2024)
by: Haldar, Malay, et al.
Published: (2024)
ArtCognition: A Multimodal AI Framework for Affective State Sensing from Visual and Kinematic Drawing Cues
by: Binaei-Haghighi, Behrad, et al.
Published: (2026)
by: Binaei-Haghighi, Behrad, et al.
Published: (2026)
Emotion-Driven Personalized Recommendation for AI-Generated Content Using Multi-Modal Sentiment and Intent Analysis
by: Hu, Zheqi, et al.
Published: (2025)
by: Hu, Zheqi, et al.
Published: (2025)
Contrastive Learning Method for Sequential Recommendation based on Multi-Intention Disentanglement
by: Hu, Zeyu, et al.
Published: (2024)
by: Hu, Zeyu, et al.
Published: (2024)
VISTA: Visualized Text Embedding For Universal Multi-Modal Retrieval
by: Zhou, Junjie, et al.
Published: (2024)
by: Zhou, Junjie, et al.
Published: (2024)
Argumentative Experience: Reducing Confirmation Bias on Controversial Issues through LLM-Generated Multi-Persona Debates
by: Shi, Li, et al.
Published: (2024)
by: Shi, Li, et al.
Published: (2024)
AFPM: Alignment-based Frame Patch Modeling for Cross-Dataset EEG Decoding
by: Chen, Xiaoqing, et al.
Published: (2025)
by: Chen, Xiaoqing, et al.
Published: (2025)
Colour Contrast on the Web: A WCAG 2.1 Level AA Compliance Audit of Common Crawl's Top 500 Domains
by: Vaughan, Thom, et al.
Published: (2026)
by: Vaughan, Thom, et al.
Published: (2026)
RecGaze: The First Eye Tracking and User Interaction Dataset for Carousel Interfaces
by: de Leon-Martinez, Santiago, et al.
Published: (2025)
by: de Leon-Martinez, Santiago, et al.
Published: (2025)
Blending Queries and Conversations: Understanding Tactics, Trust, Verification, and System Choice in Web Search and Chat Interactions
by: Mayerhofer, Kerstin, et al.
Published: (2025)
by: Mayerhofer, Kerstin, et al.
Published: (2025)
How to Make Museums More Interactive? Case Study of Artistic Chatbot
by: Kucia, Filip J., et al.
Published: (2025)
by: Kucia, Filip J., et al.
Published: (2025)
From SERPs to Agents: A Platform for Comparative Studies of Information Interaction
by: Zerhoudi, Saber, et al.
Published: (2026)
by: Zerhoudi, Saber, et al.
Published: (2026)
Task Supportive and Personalized Human-Large Language Model Interaction: A User Study
by: Wang, Ben, et al.
Published: (2024)
by: Wang, Ben, et al.
Published: (2024)
What's in People's Digital File Collections?
by: Dinneen, Jesse David, et al.
Published: (2024)
by: Dinneen, Jesse David, et al.
Published: (2024)
Enhancing EmoBot: An In-Depth Analysis of User Satisfaction and Faults in an Emotion-Aware Chatbot
by: Mubassira, Taseen, et al.
Published: (2024)
by: Mubassira, Taseen, et al.
Published: (2024)
Similar Items
-
AKRMap: Adaptive Kernel Regression for Trustworthy Visualization of Cross-Modal Embeddings
by: Ye, Yilin, et al.
Published: (2025) -
Three Modalities, Two Design Probes, One Prototype, and No Vision: Experience-Based Co-Design of a Multi-modal 3D Data Visualization Tool
by: Kamath, Sanchita S., et al.
Published: (2026) -
Evaluating VisualRAG: Quantifying Cross-Modal Performance in Enterprise Document Understanding
by: Mannam, Varun, et al.
Published: (2025) -
PoultryTalk: A Multi-modal Retrieval-Augmented Generation (RAG) System for Intelligent Poultry Management and Decision Support
by: Khanal, Kapalik, et al.
Published: (2025) -
Using a negative spatial auto-correlation index to evaluate and improve intrinsic TagMap's multi-scale visualization capabilities
by: Wei, Zhiwei, et al.
Published: (2024)