Singpath-VL Technical Report
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qiu, Zhen, Xiao, Kaiwen, Lu, Zhengwei, Liu, Xiangyu, Zhao, Lei, Zhang, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Kimi-VL Technical Report
von: Kimi Team, et al.
Veröffentlicht: (2025)
von: Kimi Team, et al.
Veröffentlicht: (2025)
Kwai Keye-VL Technical Report
von: Kwai Keye Team, et al.
Veröffentlicht: (2025)
von: Kwai Keye Team, et al.
Veröffentlicht: (2025)
SAIL-VL2 Technical Report
von: Yin, Weijie, et al.
Veröffentlicht: (2025)
von: Yin, Weijie, et al.
Veröffentlicht: (2025)
Kwai Keye-VL 1.5 Technical Report
von: Yang, Biao, et al.
Veröffentlicht: (2025)
von: Yang, Biao, et al.
Veröffentlicht: (2025)
STEP3-VL-10B Technical Report
von: Huang, Ailin, et al.
Veröffentlicht: (2026)
von: Huang, Ailin, et al.
Veröffentlicht: (2026)
Seed1.5-VL Technical Report
von: Guo, Dong, et al.
Veröffentlicht: (2025)
von: Guo, Dong, et al.
Veröffentlicht: (2025)
Qwen3-VL Technical Report
von: Bai, Shuai, et al.
Veröffentlicht: (2025)
von: Bai, Shuai, et al.
Veröffentlicht: (2025)
Qwen2.5-VL Technical Report
von: Bai, Shuai, et al.
Veröffentlicht: (2025)
von: Bai, Shuai, et al.
Veröffentlicht: (2025)
Xiaomi MiMo-VL-Miloco Technical Report
von: Li, Jiaze, et al.
Veröffentlicht: (2025)
von: Li, Jiaze, et al.
Veröffentlicht: (2025)
ZAYA1-VL-8B Technical Report
von: Shapourian, Hassan, et al.
Veröffentlicht: (2026)
von: Shapourian, Hassan, et al.
Veröffentlicht: (2026)
PLaMo 2.1-VL Technical Report
von: Kerola, Tommi, et al.
Veröffentlicht: (2026)
von: Kerola, Tommi, et al.
Veröffentlicht: (2026)
Phoenix-VL 1.5 Medium Technical Report
von: Phoenix, Team, et al.
Veröffentlicht: (2026)
von: Phoenix, Team, et al.
Veröffentlicht: (2026)
AndesVL Technical Report: An Efficient Mobile-side Multimodal Large Language Model
von: Jin, Zhiwei, et al.
Veröffentlicht: (2025)
von: Jin, Zhiwei, et al.
Veröffentlicht: (2025)
SophiaVL-R1: Reinforcing MLLMs Reasoning with Thinking Reward
von: Fan, Kaixuan, et al.
Veröffentlicht: (2025)
von: Fan, Kaixuan, et al.
Veröffentlicht: (2025)
Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning
von: Wang, Xiaokun, et al.
Veröffentlicht: (2025)
von: Wang, Xiaokun, et al.
Veröffentlicht: (2025)
Step-GUI Technical Report
von: Yan, Haolong, et al.
Veröffentlicht: (2025)
von: Yan, Haolong, et al.
Veröffentlicht: (2025)
Kelix Technical Report
von: Ding, Boyang, et al.
Veröffentlicht: (2026)
von: Ding, Boyang, et al.
Veröffentlicht: (2026)
Agent0-VL: Exploring Self-Evolving Agent for Tool-Integrated Vision-Language Reasoning
von: Liu, Jiaqi, et al.
Veröffentlicht: (2025)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2025)
Kling-Omni Technical Report
von: Kling Team, et al.
Veröffentlicht: (2025)
von: Kling Team, et al.
Veröffentlicht: (2025)
HunyuanImage 3.0 Technical Report
von: Cao, Siyu, et al.
Veröffentlicht: (2025)
von: Cao, Siyu, et al.
Veröffentlicht: (2025)
Qwen-Image Technical Report
von: Wu, Chenfei, et al.
Veröffentlicht: (2025)
von: Wu, Chenfei, et al.
Veröffentlicht: (2025)
STAvatar: Soft Binding and Temporal Density Control for Monocular 3D Head Avatars Reconstruction
von: Zhao, Jiankuo, et al.
Veröffentlicht: (2025)
von: Zhao, Jiankuo, et al.
Veröffentlicht: (2025)
HunyuanVideo 1.5 Technical Report
von: Wu, Bing, et al.
Veröffentlicht: (2025)
von: Wu, Bing, et al.
Veröffentlicht: (2025)
OpenSearch-VL: An Open Recipe for Frontier Multimodal Search Agents
von: Chen, Shuang, et al.
Veröffentlicht: (2026)
von: Chen, Shuang, et al.
Veröffentlicht: (2026)
Seedream 3.0 Technical Report
von: Gao, Yu, et al.
Veröffentlicht: (2025)
von: Gao, Yu, et al.
Veröffentlicht: (2025)
InternVL-X: Advancing and Accelerating InternVL Series with Efficient Visual Token Compression
von: Lu, Dongchen, et al.
Veröffentlicht: (2025)
von: Lu, Dongchen, et al.
Veröffentlicht: (2025)
Uni-Parser Technical Report
von: Fang, Xi, et al.
Veröffentlicht: (2025)
von: Fang, Xi, et al.
Veröffentlicht: (2025)
Modeling Spoof Noise by De-spoofing Diffusion and its Application in Face Anti-spoofing
von: Zhang, Bin, et al.
Veröffentlicht: (2024)
von: Zhang, Bin, et al.
Veröffentlicht: (2024)
Kling-MotionControl Technical Report
von: Kling Team, et al.
Veröffentlicht: (2026)
von: Kling Team, et al.
Veröffentlicht: (2026)
VL-Mamba: Exploring State Space Models for Multimodal Learning
von: Qiao, Yanyuan, et al.
Veröffentlicht: (2024)
von: Qiao, Yanyuan, et al.
Veröffentlicht: (2024)
PaddleOCR 3.0 Technical Report
von: Cui, Cheng, et al.
Veröffentlicht: (2025)
von: Cui, Cheng, et al.
Veröffentlicht: (2025)
VL-UR: Vision-Language-guided Universal Restoration of Images Degraded by Adverse Weather Conditions
von: Liu, Ziyan, et al.
Veröffentlicht: (2025)
von: Liu, Ziyan, et al.
Veröffentlicht: (2025)
Qwen-Image-2.0 Technical Report
von: Zhao, Bing, et al.
Veröffentlicht: (2026)
von: Zhao, Bing, et al.
Veröffentlicht: (2026)
Penguin-VL: Exploring the Efficiency Limits of VLM with LLM-based Vision Encoders
von: Zhang, Boqiang, et al.
Veröffentlicht: (2026)
von: Zhang, Boqiang, et al.
Veröffentlicht: (2026)
Top-Down Guidance for Learning Object-Centric Representations
von: Zou, Junhong, et al.
Veröffentlicht: (2024)
von: Zou, Junhong, et al.
Veröffentlicht: (2024)
HaploVL: A Single-Transformer Baseline for Multi-Modal Understanding
von: Yang, Rui, et al.
Veröffentlicht: (2025)
von: Yang, Rui, et al.
Veröffentlicht: (2025)
MVBoost: Boost 3D Reconstruction with Multi-View Refinement
von: Liu, Xiangyu, et al.
Veröffentlicht: (2024)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2024)
Dolphin v1.0 Technical Report
von: Weng, Taohan, et al.
Veröffentlicht: (2025)
von: Weng, Taohan, et al.
Veröffentlicht: (2025)
KlingAvatar 2.0 Technical Report
von: Kling Team, et al.
Veröffentlicht: (2025)
von: Kling Team, et al.
Veröffentlicht: (2025)
Direct Discrepancy Replay: Distribution-Discrepancy Condensation and Manifold-Consistent Replay for Continual Face Forgery Detection
von: Zhang, Tianshuo, et al.
Veröffentlicht: (2026)
von: Zhang, Tianshuo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Kimi-VL Technical Report
von: Kimi Team, et al.
Veröffentlicht: (2025) -
Kwai Keye-VL Technical Report
von: Kwai Keye Team, et al.
Veröffentlicht: (2025) -
SAIL-VL2 Technical Report
von: Yin, Weijie, et al.
Veröffentlicht: (2025) -
Kwai Keye-VL 1.5 Technical Report
von: Yang, Biao, et al.
Veröffentlicht: (2025) -
STEP3-VL-10B Technical Report
von: Huang, Ailin, et al.
Veröffentlicht: (2026)