Tokensome: Towards a Genetic Vision-Language GPT for Explainable and Cognitive Karyotyping
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Haoxi, Zhang, Xinxu, Lin, Yuanxin, Wang, Maiqi, Lai, Yi, Wang, Yu, Yu, Linfeng, Xu, Yufeng, Cheng, Ran, Szczerbicki, Edward |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Realistic Scene Generation with LiDAR Diffusion Models
by: Ran, Haoxi, et al.
Published: (2024)
by: Ran, Haoxi, et al.
Published: (2024)
X-Driver: Explainable Autonomous Driving with Vision-Language Models
by: Liu, Wei, et al.
Published: (2025)
by: Liu, Wei, et al.
Published: (2025)
Invisible Languages of the LLM Universe
by: Khanna, Saurabh, et al.
Published: (2025)
by: Khanna, Saurabh, et al.
Published: (2025)
Towards Transparent AI: A Survey on Explainable Large Language Models
by: Palikhe, Avash, et al.
Published: (2025)
by: Palikhe, Avash, et al.
Published: (2025)
Instant4D: 4D Gaussian Splatting in Minutes
by: Luo, Zhanpeng, et al.
Published: (2025)
by: Luo, Zhanpeng, et al.
Published: (2025)
Diffusion-Guided Pretraining for Brain Graph Foundation Models
by: Wei, Xinxu, et al.
Published: (2026)
by: Wei, Xinxu, et al.
Published: (2026)
Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models
by: Liu, Yuqing, et al.
Published: (2024)
by: Liu, Yuqing, et al.
Published: (2024)
Explainable Few-shot Knowledge Tracing
by: Li, Haoxuan, et al.
Published: (2024)
by: Li, Haoxuan, et al.
Published: (2024)
Resurgence number of matroid configuration
by: Hu, Haoxi
Published: (2025)
by: Hu, Haoxi
Published: (2025)
RegionGPT: Towards Region Understanding Vision Language Model
by: Guo, Qiushan, et al.
Published: (2024)
by: Guo, Qiushan, et al.
Published: (2024)
AgriGPT-VL: Agricultural Vision-Language Understanding Suite
by: Yang, Bo, et al.
Published: (2025)
by: Yang, Bo, et al.
Published: (2025)
Is Meta-Path Attention an Explanation? Evidence of Alignment and Decoupling in Heterogeneous GNNs
by: Jiang, Maiqi, et al.
Published: (2026)
by: Jiang, Maiqi, et al.
Published: (2026)
Think How to Think: Mitigating Overthinking with Autonomous Difficulty Cognition in Large Reasoning Models
by: Liu, Yongjiang, et al.
Published: (2025)
by: Liu, Yongjiang, et al.
Published: (2025)
Organic Gradient Homojunction via D‐A Engineering Enables Photoelectric/Photothermal Dual‐Assisted Catalysis Toward Full Spectrum Light‐Coupled Low‐Temperature Seawater Batteries
by: Yi Lin, et al.
Published: (2025)
by: Yi Lin, et al.
Published: (2025)
SurgXBench: Explainable Vision-Language Model Benchmark for Surgery
by: Cheng, Jiajun, et al.
Published: (2025)
by: Cheng, Jiajun, et al.
Published: (2025)
Evaluating Chromosomal Mosaicism in Prenatal Diagnosis: The Complementary Roles of Chromosomal Microarray Analysis and Karyotyping
by: Chenxia Xu, et al.
Published: (2025)
by: Chenxia Xu, et al.
Published: (2025)
OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models
by: Jia, Mengdi, et al.
Published: (2025)
by: Jia, Mengdi, et al.
Published: (2025)
Almost sure convergence of differentially positive systems on a globally orderable manifold
by: Niu, Lin, et al.
Published: (2024)
by: Niu, Lin, et al.
Published: (2024)
FlightGPT: Towards Generalizable and Interpretable UAV Vision-and-Language Navigation with Vision-Language Models
by: Cai, Hengxing, et al.
Published: (2025)
by: Cai, Hengxing, et al.
Published: (2025)
Simulation and Deformation Analysis of Construction Process of Large Eccentric Frame‐Core Tube Structure
by: Guanghua Yin, et al.
Published: (2024)
by: Guanghua Yin, et al.
Published: (2024)
A Brain Graph Foundation Model: Pre-Training and Prompt-Tuning across Broad Atlases and Disorders
by: Wei, Xinxu, et al.
Published: (2025)
by: Wei, Xinxu, et al.
Published: (2025)
VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction
by: Fu, Chaoyou, et al.
Published: (2025)
by: Fu, Chaoyou, et al.
Published: (2025)
Dynamic Retriever for In-Context Knowledge Editing via Policy Optimization
by: Nafee, Mahmud Wasif, et al.
Published: (2025)
by: Nafee, Mahmud Wasif, et al.
Published: (2025)
Toward Ultra-Long-Horizon Agentic Science: Cognitive Accumulation for Machine Learning Engineering
by: Zhu, Xinyu, et al.
Published: (2026)
by: Zhu, Xinyu, et al.
Published: (2026)
Seeing Sarcasm Through Different Eyes: Analyzing Multimodal Sarcasm Perception in Large Vision-Language Models
by: Chen, Junjie, et al.
Published: (2025)
by: Chen, Junjie, et al.
Published: (2025)
HarmoCLIP: Harmonizing Global and Regional Representations in Contrastive Vision-Language Models
by: Zeng, Haoxi, et al.
Published: (2025)
by: Zeng, Haoxi, et al.
Published: (2025)
Botulinum Toxin Type A as a Therapeutic Agent in Epilepsy: Attenuation of Neuronal Ferroptosis and Cognitive Dysfunction
by: Shuang Li, et al.
Published: (2025)
by: Shuang Li, et al.
Published: (2025)
Dissecting Role Cognition in Medical LLMs via Neuronal Ablation
by: Liang, Xun, et al.
Published: (2025)
by: Liang, Xun, et al.
Published: (2025)
GPT4AIGChip: Towards Next-Generation AI Accelerator Design Automation via Large Language Models
by: Fu, Yonggan, et al.
Published: (2023)
by: Fu, Yonggan, et al.
Published: (2023)
HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale
by: Chen, Junying, et al.
Published: (2024)
by: Chen, Junying, et al.
Published: (2024)
Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model
by: Liu, Ting, et al.
Published: (2024)
by: Liu, Ting, et al.
Published: (2024)
GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization
by: Pan, Yu, et al.
Published: (2026)
by: Pan, Yu, et al.
Published: (2026)
MC-GPT: Empowering Vision-and-Language Navigation with Memory Map and Reasoning Chains
by: Zhan, Zhaohuan, et al.
Published: (2024)
by: Zhan, Zhaohuan, et al.
Published: (2024)
RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics
by: Zhou, Enshen, et al.
Published: (2025)
by: Zhou, Enshen, et al.
Published: (2025)
GestureGPT: Toward Zero-Shot Free-Form Hand Gesture Understanding with Large Language Model Agents
by: Zeng, Xin, et al.
Published: (2023)
by: Zeng, Xin, et al.
Published: (2023)
ProtChatGPT: Towards Understanding Proteins with Large Language Models
by: Wang, Chao, et al.
Published: (2024)
by: Wang, Chao, et al.
Published: (2024)
Towards Compatible Fine-tuning for Vision-Language Model Updates
by: Wang, Zhengbo, et al.
Published: (2024)
by: Wang, Zhengbo, et al.
Published: (2024)
Comment on “Diagnostic Performance of ChatGPT ‐4o and DeepSeek ‐3 Differential Diagnosis of Complex Oral Lesions: A Multimodal Imaging and Case Difficulty Analysis”
by: Yuxuan Zhang, et al.
Published: (2025)
by: Yuxuan Zhang, et al.
Published: (2025)
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
by: Xu, Junhao, et al.
Published: (2025)
by: Xu, Junhao, et al.
Published: (2025)
Towards Distribution Matching between Collaborative and Language Spaces for Generative Recommendation
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
Similar Items
-
Towards Realistic Scene Generation with LiDAR Diffusion Models
by: Ran, Haoxi, et al.
Published: (2024) -
X-Driver: Explainable Autonomous Driving with Vision-Language Models
by: Liu, Wei, et al.
Published: (2025) -
Invisible Languages of the LLM Universe
by: Khanna, Saurabh, et al.
Published: (2025) -
Towards Transparent AI: A Survey on Explainable Large Language Models
by: Palikhe, Avash, et al.
Published: (2025) -
Instant4D: 4D Gaussian Splatting in Minutes
by: Luo, Zhanpeng, et al.
Published: (2025)