TextME: Bridging Unseen Modalities Through Text Descriptions
Fuente:
arXiv
Saved in:
| Main Authors: | Hong, Soyeon, Kim, Jinchan, You, Jaegook, Choi, Seungtaek, Kwak, Suha, Cho, Hyunsouk |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FLEX: Expert-level False-Less EXecution Metric for Reliable Text-to-SQL Benchmark
by: Kim, Heegyu, et al.
Published: (2024)
by: Kim, Heegyu, et al.
Published: (2024)
Learning Unified Distance Metric Across Diverse Data Distributions with Parameter-Efficient Transfer Learning
by: Kim, Sungyeon, et al.
Published: (2023)
by: Kim, Sungyeon, et al.
Published: (2023)
Meta-Controller: Few-Shot Imitation of Unseen Embodiments and Tasks in Continuous Control
by: Cho, Seongwoong, et al.
Published: (2024)
by: Cho, Seongwoong, et al.
Published: (2024)
CORN: Contact-based Object Representation for Nonprehensile Manipulation of General Unseen Objects
by: Cho, Yoonyoung, et al.
Published: (2024)
by: Cho, Yoonyoung, et al.
Published: (2024)
Training-Free Safe Text Embedding Guidance for Text-to-Image Diffusion Models
by: Na, Byeonghu, et al.
Published: (2025)
by: Na, Byeonghu, et al.
Published: (2025)
Text2Robot: Evolutionary Robot Design from Text Descriptions
by: Ringel, Ryan P., et al.
Published: (2024)
by: Ringel, Ryan P., et al.
Published: (2024)
Diffusion-Link: Diffusion Probabilistic Model for Bridging the Audio-Text Modality Gap
by: Nam, KiHyun, et al.
Published: (2025)
by: Nam, KiHyun, et al.
Published: (2025)
MoST: Mixing Speech and Text with Modality-Aware Mixture of Experts
by: Lou, Yuxuan, et al.
Published: (2026)
by: Lou, Yuxuan, et al.
Published: (2026)
PhysioME: A Robust Multimodal Self-Supervised Framework for Physiological Signals with Missing Modalities
by: Lee, Cheol-Hui, et al.
Published: (2025)
by: Lee, Cheol-Hui, et al.
Published: (2025)
Learning to Generalize Unseen Domains via Multi-Source Meta Learning for Text Classification
by: Hu, Yuxuan, et al.
Published: (2024)
by: Hu, Yuxuan, et al.
Published: (2024)
Multimodal Forecasting for Commodity Prices Using Spectrogram-Based and Time Series Representations
by: Park, Soyeon, et al.
Published: (2026)
by: Park, Soyeon, et al.
Published: (2026)
Toward Structural Multimodal Representations: Specialization, Selection, and Sparsification via Mixture-of-Experts
by: Choi, Hahyeon, et al.
Published: (2026)
by: Choi, Hahyeon, et al.
Published: (2026)
Trillion 7B Technical Report
by: Han, Sungjun, et al.
Published: (2025)
by: Han, Sungjun, et al.
Published: (2025)
TourSynbio: A Multi-Modal Large Model and Agent Framework to Bridge Text and Protein Sequences for Protein Engineering
by: Shen, Yiqing, et al.
Published: (2024)
by: Shen, Yiqing, et al.
Published: (2024)
Semantic Nutrition Estimation: Predicting Food Healthfulness from Text Descriptions
by: Freudenberg, Dayne R., et al.
Published: (2025)
by: Freudenberg, Dayne R., et al.
Published: (2025)
Alert-ME: An Explainability-Driven Defense Against Adversarial Examples in Transformer-Based Text Classification
by: Sabir, Bushra, et al.
Published: (2023)
by: Sabir, Bushra, et al.
Published: (2023)
Bridging the Gap Between Molecule and Textual Descriptions via Substructure-aware Alignment
by: Park, Hyuntae, et al.
Published: (2025)
by: Park, Hyuntae, et al.
Published: (2025)
HARMONY: Bridging the Personalization-Generalization Gap by Mitigating Representation Skew in Heterogeneous Split Federated Learning
by: Youn, Jiseok, et al.
Published: (2026)
by: Youn, Jiseok, et al.
Published: (2026)
Differentially Private Federated Clustering with Random Rebalancing
by: Yang, Xiyuan, et al.
Published: (2025)
by: Yang, Xiyuan, et al.
Published: (2025)
Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning
by: Kim, Minung, et al.
Published: (2026)
by: Kim, Minung, et al.
Published: (2026)
Hollowed Net for On-Device Personalization of Text-to-Image Diffusion Models
by: Cho, Wonguk, et al.
Published: (2024)
by: Cho, Wonguk, et al.
Published: (2024)
Text2Chart31: Instruction Tuning for Chart Generation with Automatic Feedback
by: Zadeh, Fatemeh Pesaran, et al.
Published: (2024)
by: Zadeh, Fatemeh Pesaran, et al.
Published: (2024)
PFGuard: A Generative Framework with Privacy and Fairness Safeguards
by: Kim, Soyeon, et al.
Published: (2024)
by: Kim, Soyeon, et al.
Published: (2024)
Diffusion Adaptive Text Embedding for Text-to-Image Diffusion Models
by: Na, Byeonghu, et al.
Published: (2025)
by: Na, Byeonghu, et al.
Published: (2025)
FakeInversion: Learning to Detect Images from Unseen Text-to-Image Models by Inverting Stable Diffusion
by: Cazenavette, George, et al.
Published: (2024)
by: Cazenavette, George, et al.
Published: (2024)
Text-to-Image GAN with Pretrained Representations
by: You, Xiaozhou, et al.
Published: (2024)
by: You, Xiaozhou, et al.
Published: (2024)
Text-Aware Image Restoration with Diffusion Models
by: Min, Jaewon, et al.
Published: (2025)
by: Min, Jaewon, et al.
Published: (2025)
GENIUS: A Generative Framework for Universal Multimodal Search
by: Kim, Sungyeon, et al.
Published: (2025)
by: Kim, Sungyeon, et al.
Published: (2025)
Addressing Negative Transfer in Diffusion Models
by: Go, Hyojun, et al.
Published: (2023)
by: Go, Hyojun, et al.
Published: (2023)
Federated Learning for Face Recognition via Intra-subject Self-supervised Learning
by: Kim, Hansol, et al.
Published: (2024)
by: Kim, Hansol, et al.
Published: (2024)
Sparsely-Supervised Data Assimilation via Physics-Informed Schrödinger Bridge
by: Bu, Dohyun, et al.
Published: (2026)
by: Bu, Dohyun, et al.
Published: (2026)
Improving Text-based Person Search via Part-level Cross-modal Correspondence
by: Park, Jicheol, et al.
Published: (2024)
by: Park, Jicheol, et al.
Published: (2024)
Fact-Consistency Evaluation of Text-to-SQL Generation for Business Intelligence Using Exaone 3.5
by: Choi, Jeho
Published: (2025)
by: Choi, Jeho
Published: (2025)
Integrating Unstructured Text into Causal Inference: Empirical Evidence from Real Data
by: Zhou, Boning, et al.
Published: (2026)
by: Zhou, Boning, et al.
Published: (2026)
Text-to-Level Diffusion Models With Various Text Encoders for Super Mario Bros
by: Schrum, Jacob, et al.
Published: (2025)
by: Schrum, Jacob, et al.
Published: (2025)
Automated Filtering of Human Feedback Data for Aligning Text-to-Image Diffusion Models
by: Yang, Yongjin, et al.
Published: (2024)
by: Yang, Yongjin, et al.
Published: (2024)
PnPXAI: A Universal XAI Framework Providing Automatic Explanations Across Diverse Modalities and Models
by: Kim, Seongun, et al.
Published: (2025)
by: Kim, Seongun, et al.
Published: (2025)
EdgeFusion: On-Device Text-to-Image Generation
by: Castells, Thibault, et al.
Published: (2024)
by: Castells, Thibault, et al.
Published: (2024)
Sample-Efficient Diffusion for Text-To-Speech Synthesis
by: Lovelace, Justin, et al.
Published: (2024)
by: Lovelace, Justin, et al.
Published: (2024)
TARDiS : Text Augmentation for Refining Diversity and Separability
by: Kim, Kyungmin, et al.
Published: (2025)
by: Kim, Kyungmin, et al.
Published: (2025)
Similar Items
-
FLEX: Expert-level False-Less EXecution Metric for Reliable Text-to-SQL Benchmark
by: Kim, Heegyu, et al.
Published: (2024) -
Learning Unified Distance Metric Across Diverse Data Distributions with Parameter-Efficient Transfer Learning
by: Kim, Sungyeon, et al.
Published: (2023) -
Meta-Controller: Few-Shot Imitation of Unseen Embodiments and Tasks in Continuous Control
by: Cho, Seongwoong, et al.
Published: (2024) -
CORN: Contact-based Object Representation for Nonprehensile Manipulation of General Unseen Objects
by: Cho, Yoonyoung, et al.
Published: (2024) -
Training-Free Safe Text Embedding Guidance for Text-to-Image Diffusion Models
by: Na, Byeonghu, et al.
Published: (2025)