Uni-SMART: Universal Science Multimodal Analysis and Research Transformer
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Hengxing, Cai, Xiaochen, Yang, Shuwen, Wang, Jiankun, Yao, Lin, Gao, Zhifeng, Chang, Junhan, Li, Sihang, Xu, Mingjun, Wang, Changxin, Wang, Hongshuai, Li, Yongge, Lin, Mujie, Li, Yaqi, Yin, Yuqi, Zhang, Linfeng, Ke, Guolin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis
by: Cai, Hengxing, et al.
Published: (2024)
by: Cai, Hengxing, et al.
Published: (2024)
SciLitLLM: How to Adapt LLMs for Scientific Literature Understanding
by: Li, Sihang, et al.
Published: (2024)
by: Li, Sihang, et al.
Published: (2024)
MolParser: End-to-end Visual Recognition of Molecule Structures in the Wild
by: Fang, Xi, et al.
Published: (2024)
by: Fang, Xi, et al.
Published: (2024)
MolReasoner: Toward Effective and Interpretable Reasoning for Molecular LLMs
by: Zhao, Guojiang, et al.
Published: (2025)
by: Zhao, Guojiang, et al.
Published: (2025)
MM-R5: MultiModal Reasoning-Enhanced ReRanker via Reinforcement Learning for Document Retrieval
by: Xu, Mingjun, et al.
Published: (2025)
by: Xu, Mingjun, et al.
Published: (2025)
Uni-AIMS: AI-Powered Microscopy Image Analysis
by: Hong, Yanhui, et al.
Published: (2025)
by: Hong, Yanhui, et al.
Published: (2025)
Uni-Parser Technical Report
by: Fang, Xi, et al.
Published: (2025)
by: Fang, Xi, et al.
Published: (2025)
Doc2SAR: A Synergistic Framework for High-Fidelity Extraction of Structure-Activity Relationships from Scientific Documents
by: Zhuang, Jiaxi, et al.
Published: (2025)
by: Zhuang, Jiaxi, et al.
Published: (2025)
A Multi-Granularity Retrieval Framework for Visually-Rich Documents
by: Xu, Mingjun, et al.
Published: (2025)
by: Xu, Mingjun, et al.
Published: (2025)
UniEM-3M: A Universal Electron Micrograph Dataset for Microstructural Segmentation and Generation
by: wang, Nan, et al.
Published: (2025)
by: wang, Nan, et al.
Published: (2025)
Draft-Refine-Optimize: Self-Evolved Learning for Natural Language to MongoDB Query Generation
by: Ye, Mingwei, et al.
Published: (2026)
by: Ye, Mingwei, et al.
Published: (2026)
FlightGPT: Towards Generalizable and Interpretable UAV Vision-and-Language Navigation with Vision-Language Models
by: Cai, Hengxing, et al.
Published: (2025)
by: Cai, Hengxing, et al.
Published: (2025)
Masked-and-Reordered Self-Supervision for Reinforcement Learning from Verifiable Rewards
by: Wang, Zhen, et al.
Published: (2025)
by: Wang, Zhen, et al.
Published: (2025)
Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents
by: Shi, Yaorui, et al.
Published: (2025)
by: Shi, Yaorui, et al.
Published: (2025)
Uni-Mol Docking V2: Towards Realistic and Accurate Binding Pose Prediction
by: Alcaide, Eric, et al.
Published: (2024)
by: Alcaide, Eric, et al.
Published: (2024)
Search and Refine During Think: Facilitating Knowledge Refinement for Improved Retrieval-Augmented Reasoning
by: Shi, Yaorui, et al.
Published: (2025)
by: Shi, Yaorui, et al.
Published: (2025)
Intelligent System for Automated Molecular Patent Infringement Assessment
by: Shi, Yaorui, et al.
Published: (2024)
by: Shi, Yaorui, et al.
Published: (2024)
NOSE: Neural Olfactory-Semantic Embedding with Tri-Modal Orthogonal Contrastive Learning
by: Su, Yanyi, et al.
Published: (2026)
by: Su, Yanyi, et al.
Published: (2026)
OmniScience: A Large-scale Multi-modal Dataset for Scientific Image Understanding
by: Tao, Haoyi, et al.
Published: (2026)
by: Tao, Haoyi, et al.
Published: (2026)
Uni-Mol2: Exploring Molecular Pretraining Model at Scale
by: Ji, Xiaohong, et al.
Published: (2024)
by: Ji, Xiaohong, et al.
Published: (2024)
Demystifying Numerosity in Diffusion Models -- Limitations and Remedies
by: Zhao, Yaqi, et al.
Published: (2025)
by: Zhao, Yaqi, et al.
Published: (2025)
System-2 Mathematical Reasoning via Enriched Instruction Tuning
by: Cai, Huanqia, et al.
Published: (2024)
by: Cai, Huanqia, et al.
Published: (2024)
Multi-Task Fine-Tuning Enables Robust Out-of-Distribution Generalization in Atomistic Models
by: Zhang, Chengqian, et al.
Published: (2026)
by: Zhang, Chengqian, et al.
Published: (2026)
RxnBench: A Multimodal Benchmark for Evaluating Large Language Models on Chemical Reaction Understanding from Scientific Literature
by: Li, Hanzheng, et al.
Published: (2025)
by: Li, Hanzheng, et al.
Published: (2025)
Uni-Mol3: A Multi-Molecular Foundation Model for Advancing Organic Reaction Modeling
by: Wu, Lirong, et al.
Published: (2025)
by: Wu, Lirong, et al.
Published: (2025)
UniHDA: A Unified and Versatile Framework for Multi-Modal Hybrid Domain Adaptation
by: Li, Hengjia, et al.
Published: (2024)
by: Li, Hengjia, et al.
Published: (2024)
UniTabE: A Universal Pretraining Protocol for Tabular Foundation Model in Data Science
by: Yang, Yazheng, et al.
Published: (2023)
by: Yang, Yazheng, et al.
Published: (2023)
UniLabOS: An AI-Native Operating System for Autonomous Laboratories
by: Gao, Jing, et al.
Published: (2025)
by: Gao, Jing, et al.
Published: (2025)
MolCA: Molecular Graph-Language Modeling with Cross-Modal Projector and Uni-Modal Adapter
by: Liu, Zhiyuan, et al.
Published: (2023)
by: Liu, Zhiyuan, et al.
Published: (2023)
UniCom: Unified Multimodal Modeling via Compressed Continuous Semantic Representations
by: Zhao, Yaqi, et al.
Published: (2026)
by: Zhao, Yaqi, et al.
Published: (2026)
Representative Attention For Vision Transformers
by: Li, Yuntong, et al.
Published: (2026)
by: Li, Yuntong, et al.
Published: (2026)
Towards Dynamic and Small Objects Refinement for Unsupervised Domain Adaptative Nighttime Semantic Segmentation
by: Pan, Jingyi, et al.
Published: (2023)
by: Pan, Jingyi, et al.
Published: (2023)
UniAP: Unifying Inter- and Intra-Layer Automatic Parallelism by Mixed Integer Quadratic Programming
by: Lin, Hao, et al.
Published: (2023)
by: Lin, Hao, et al.
Published: (2023)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
by: Wang, Xiang, et al.
Published: (2025)
by: Wang, Xiang, et al.
Published: (2025)
Automated Market Makers for Decentralized Finance (DeFi)
by: Wang, Yongge
Published: (2020)
by: Wang, Yongge
Published: (2020)
Forecast the Principal, Stabilize the Residual: Subspace-Aware Feature Caching for Efficient Diffusion Transformers
by: Chen, Guantao, et al.
Published: (2026)
by: Chen, Guantao, et al.
Published: (2026)
Towards a Unified Benchmark and Framework for Deep Learning-Based Prediction of Nuclear Magnetic Resonance Chemical Shifts
by: Xu, Fanjie, et al.
Published: (2024)
by: Xu, Fanjie, et al.
Published: (2024)
Realizing Immersive Volumetric Video: A Multimodal Framework for 6-DoF VR Engagement
by: Yang, Zhengxian, et al.
Published: (2026)
by: Yang, Zhengxian, et al.
Published: (2026)
Interpretable Reward Model via Sparse Autoencoder
by: Zhang, Shuyi, et al.
Published: (2025)
by: Zhang, Shuyi, et al.
Published: (2025)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
by: Wang, Xiang, et al.
Published: (2024)
by: Wang, Xiang, et al.
Published: (2024)
Similar Items
-
SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis
by: Cai, Hengxing, et al.
Published: (2024) -
SciLitLLM: How to Adapt LLMs for Scientific Literature Understanding
by: Li, Sihang, et al.
Published: (2024) -
MolParser: End-to-end Visual Recognition of Molecule Structures in the Wild
by: Fang, Xi, et al.
Published: (2024) -
MolReasoner: Toward Effective and Interpretable Reasoning for Molecular LLMs
by: Zhao, Guojiang, et al.
Published: (2025) -
MM-R5: MultiModal Reasoning-Enhanced ReRanker via Reinforcement Learning for Document Retrieval
by: Xu, Mingjun, et al.
Published: (2025)