ReactXT: Understanding Molecular "Reaction-ship" via Reaction-Contextualized Molecule-Text Pretraining
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Zhiyuan, Shi, Yaorui, Zhang, An, Li, Sihang, Zhang, Enzhi, Wang, Xiang, Kawaguchi, Kenji, Chua, Tat-Seng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ProtT3: Protein-to-Text Generation for Text-based Protein Understanding
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2024)
NExT-Mol: 3D Diffusion Meets 1D Language Modeling for 3D Molecule Generation
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
MolCA: Molecular Graph-Language Modeling with Cross-Modal Projector and Uni-Modal Adapter
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2023)
Rethinking Tokenizer and Decoder in Masked Graph Modeling for Molecules
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2023)
Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
Efficient Constraining of Transcoding in DNA-Based Image Storage
von: Sayyed, Sara Al, et al.
Veröffentlicht: (2025)
von: Sayyed, Sara Al, et al.
Veröffentlicht: (2025)
MUDI: A Multimodal Biomedical Dataset for Understanding Pharmacodynamic Drug-Drug Interactions
von: Ngo, Tung-Lam, et al.
Veröffentlicht: (2025)
von: Ngo, Tung-Lam, et al.
Veröffentlicht: (2025)
ReactDiff: Latent Diffusion for Facial Reaction Generation
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
Towards Natural Language-Guided Drones: GeoText-1652 Benchmark with Spatial Relation Matching
von: Chu, Meng, et al.
Veröffentlicht: (2023)
von: Chu, Meng, et al.
Veröffentlicht: (2023)
Simple but Effective Raw-Data Level Multimodal Fusion for Composed Image Retrieval
von: Wen, Haokun, et al.
Veröffentlicht: (2024)
von: Wen, Haokun, et al.
Veröffentlicht: (2024)
Towards 3D Molecule-Text Interpretation in Language Models
von: Li, Sihang, et al.
Veröffentlicht: (2024)
von: Li, Sihang, et al.
Veröffentlicht: (2024)
Discriminative Probing and Tuning for Text-to-Image Generation
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
GestaltMML: Enhancing Rare Genetic Disease Diagnosis through Multimodal Machine Learning Combining Facial Images and Clinical Texts
von: Wu, Da, et al.
Veröffentlicht: (2023)
von: Wu, Da, et al.
Veröffentlicht: (2023)
Extending Visual Dynamics for Video-to-Music Generation
von: Liu, Xiaohao, et al.
Veröffentlicht: (2025)
von: Liu, Xiaohao, et al.
Veröffentlicht: (2025)
Scene-Text Grounding for Text-Based Video Question Answering
von: Zhou, Sheng, et al.
Veröffentlicht: (2024)
von: Zhou, Sheng, et al.
Veröffentlicht: (2024)
Smart Fitting Room: A One-stop Framework for Matching-aware Virtual Try-on
von: Yu, Mingzhe, et al.
Veröffentlicht: (2024)
von: Yu, Mingzhe, et al.
Veröffentlicht: (2024)
ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion Model
von: Cheng, Luo, et al.
Veröffentlicht: (2025)
von: Cheng, Luo, et al.
Veröffentlicht: (2025)
VA3: Virtually Assured Amplification Attack on Probabilistic Copyright Protection for Text-to-Image Generative Models
von: Li, Xiang, et al.
Veröffentlicht: (2023)
von: Li, Xiang, et al.
Veröffentlicht: (2023)
EgoTextVQA: Towards Egocentric Scene-Text Aware Video Question Answering
von: Zhou, Sheng, et al.
Veröffentlicht: (2025)
von: Zhou, Sheng, et al.
Veröffentlicht: (2025)
ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions
von: Luo, Cheng, et al.
Veröffentlicht: (2023)
von: Luo, Cheng, et al.
Veröffentlicht: (2023)
On Copyright Risks of Text-to-Image Diffusion Models
von: Zhang, Yang, et al.
Veröffentlicht: (2023)
von: Zhang, Yang, et al.
Veröffentlicht: (2023)
ReactZyme: A Benchmark for Enzyme-Reaction Prediction
von: Hua, Chenqing, et al.
Veröffentlicht: (2024)
von: Hua, Chenqing, et al.
Veröffentlicht: (2024)
Evaluating Molecule Synthesizability via Retrosynthetic Planning and Reaction Prediction
von: Liu, Songtao, et al.
Veröffentlicht: (2024)
von: Liu, Songtao, et al.
Veröffentlicht: (2024)
CIRP: Cross-Item Relational Pre-training for Multimodal Product Bundling
von: Ma, Yunshan, et al.
Veröffentlicht: (2024)
von: Ma, Yunshan, et al.
Veröffentlicht: (2024)
Can I Trust Your Answer? Visually Grounded Video Question Answering
von: Xiao, Junbin, et al.
Veröffentlicht: (2023)
von: Xiao, Junbin, et al.
Veröffentlicht: (2023)
Principled Multimodal Representation Learning
von: Liu, Xiaohao, et al.
Veröffentlicht: (2025)
von: Liu, Xiaohao, et al.
Veröffentlicht: (2025)
SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
Revolutionizing Text-to-Image Retrieval as Autoregressive Token-to-Voken Generation
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
FashionReGen: LLM-Empowered Fashion Report Generation
von: Ding, Yujuan, et al.
Veröffentlicht: (2024)
von: Ding, Yujuan, et al.
Veröffentlicht: (2024)
ExpLLM: Towards Chain of Thought for Facial Expression Recognition
von: Lan, Xing, et al.
Veröffentlicht: (2024)
von: Lan, Xing, et al.
Veröffentlicht: (2024)
InstructVid2Vid: Controllable Video Editing with Natural Language Instructions
von: Qin, Bosheng, et al.
Veröffentlicht: (2023)
von: Qin, Bosheng, et al.
Veröffentlicht: (2023)
A Picture Is Worth a Graph: A Blueprint Debate Paradigm for Multimodal Reasoning
von: Zheng, Changmeng, et al.
Veröffentlicht: (2024)
von: Zheng, Changmeng, et al.
Veröffentlicht: (2024)
Atomas: Hierarchical Alignment on Molecule-Text for Unified Molecule Understanding and Generation
von: Zhang, Yikun, et al.
Veröffentlicht: (2024)
von: Zhang, Yikun, et al.
Veröffentlicht: (2024)
MolTC: Towards Molecular Relational Modeling In Language Models
von: Fang, Junfeng, et al.
Veröffentlicht: (2024)
von: Fang, Junfeng, et al.
Veröffentlicht: (2024)
Contextual Wireless Video Semantic Communication in MIMO-OFDM Systems
von: Xie, Bingyan, et al.
Veröffentlicht: (2026)
von: Xie, Bingyan, et al.
Veröffentlicht: (2026)
Dynamic Multimodal Fusion via Meta-Learning Towards Micro-Video Recommendation
von: Liu, Han, et al.
Veröffentlicht: (2025)
von: Liu, Han, et al.
Veröffentlicht: (2025)
TIGeR: Unifying Text-to-Image Generation and Retrieval with Large Multimodal Models
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
SemCORE: A Semantic-Enhanced Generative Cross-Modal Retrieval Framework with MLLMs
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
Pretraining a Foundation Model for Small-Molecule Natural Products
von: Ding, Yuheng, et al.
Veröffentlicht: (2025)
von: Ding, Yuheng, et al.
Veröffentlicht: (2025)
Towards Temporal-Aware Multi-Modal Retrieval Augmented Generation in Finance
von: Zhu, Fengbin, et al.
Veröffentlicht: (2025)
von: Zhu, Fengbin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ProtT3: Protein-to-Text Generation for Text-based Protein Understanding
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2024) -
NExT-Mol: 3D Diffusion Meets 1D Language Modeling for 3D Molecule Generation
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025) -
MolCA: Molecular Graph-Language Modeling with Cross-Modal Projector and Uni-Modal Adapter
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2023) -
Rethinking Tokenizer and Decoder in Masked Graph Modeling for Molecules
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2023) -
Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval
von: Fang, Xiang, et al.
Veröffentlicht: (2026)