FireRed-OCR Technical Report
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Hao, Lou, Haoran, Li, Xinyue, Zhong, Zuodong, Sun, Zhaojun, Chen, Phellon, Zhou, Xuanhe, Zuo, Kai, Chen, Yibo, Tang, Xu, Hu, Yao, Zhou, Boxiang, Wu, Jian, Wu, Yongji, Yu, Wenxin, Liu, Yingmiao, Huang, Yuhao, Xu, Manjie, Liu, Gang, Ma, Yidong, Sun, Zhichao, Qiao, Changhao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FireRed-Image-Edit-1.0 Technical Report
by: Super Intelligence Team, et al.
Published: (2026)
by: Super Intelligence Team, et al.
Published: (2026)
IVC-Prune: Revealing the Implicit Visual Coordinates in LVLMs for Vision Token Pruning
by: Sun, Zhichao, et al.
Published: (2026)
by: Sun, Zhichao, et al.
Published: (2026)
Cross-Scenario Unified Modeling of User Interests at Billion Scale
by: Xu, Manjie, et al.
Published: (2025)
by: Xu, Manjie, et al.
Published: (2025)
CQ-DINO: Mitigating Gradient Dilution via Category Queries for Vast Vocabulary Object Detection
by: Sun, Zhichao, et al.
Published: (2025)
by: Sun, Zhichao, et al.
Published: (2025)
SurveyBench: Can LLM(-Agents) Write Academic Surveys that Align with Reader Needs?
by: Sun, Zhaojun, et al.
Published: (2025)
by: Sun, Zhaojun, et al.
Published: (2025)
HyMiRec: A Hybrid Multi-interest Learning Framework for LLM-based Sequential Recommendation
by: Zhou, Jingyi, et al.
Published: (2025)
by: Zhou, Jingyi, et al.
Published: (2025)
GaussMaster: An LLM-based Database Copilot System
by: Zhou, Wei, et al.
Published: (2025)
by: Zhou, Wei, et al.
Published: (2025)
DBAIOps: A Reasoning LLM-Enhanced Database Operation and Maintenance System using Knowledge Graphs
by: Zhou, Wei, et al.
Published: (2025)
by: Zhou, Wei, et al.
Published: (2025)
P‐156: Efficiency Improvement Mechanism Analysis of Sidewall Passivation GaN based Micro‐LEDs by Atomic Layer Deposition
by: Mengyuan Zhanghu, et al.
Published: (2024)
by: Mengyuan Zhanghu, et al.
Published: (2024)
Mamba-VMR: Multimodal Query Augmentation via Generated Videos for Precise Temporal Grounding
by: Sun, Yunzhuo, et al.
Published: (2026)
by: Sun, Yunzhuo, et al.
Published: (2026)
Trisimplicial vertices in (fork, odd parachute)-free graphs
by: Lan, Kaiyang, et al.
Published: (2025)
by: Lan, Kaiyang, et al.
Published: (2025)
STS: Efficient Sparse Attention with Speculative Token Sparsity
by: Xu, Ceyu, et al.
Published: (2026)
by: Xu, Ceyu, et al.
Published: (2026)
A Survey of LLM $\times$ DATA
by: Zhou, Xuanhe, et al.
Published: (2025)
by: Zhou, Xuanhe, et al.
Published: (2025)
Mott insulating states of the anisotropic SU(4) Dirac fermions
by: Xu, Han, et al.
Published: (2019)
by: Xu, Han, et al.
Published: (2019)
Automating Database-Native Function Code Synthesis with LLMs
by: Zhou, Wei, et al.
Published: (2026)
by: Zhou, Wei, et al.
Published: (2026)
SRC-gAudio: Sampling-Rate-Controlled Audio Generation
by: Li, Chenxing, et al.
Published: (2024)
by: Li, Chenxing, et al.
Published: (2024)
Dial: A Knowledge-Grounded Dialect-Specific NL2SQL System
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs
by: Lou, Haoran, et al.
Published: (2026)
by: Lou, Haoran, et al.
Published: (2026)
Global coordination and trade‐off of grassland species traits and climatic drivers
by: Kuo Sun, et al.
Published: (2025)
by: Kuo Sun, et al.
Published: (2025)
PARROT: A Benchmark for Evaluating LLMs in Cross-System SQL Translation
by: Zhou, Wei, et al.
Published: (2025)
by: Zhou, Wei, et al.
Published: (2025)
Single‐cell transcriptomics reveal metastatic CLDN4+ cancer cells underlying the recurrence of malignant pleural effusion in patients with advanced non‐small‐cell lung cancer
by: Xiaoshen Zhang, et al.
Published: (2024)
by: Xiaoshen Zhang, et al.
Published: (2024)
PanoNav: Mapless Zero-Shot Object Navigation with Panoramic Scene Parsing and Dynamic Memory
by: Jin, Qunchao, et al.
Published: (2025)
by: Jin, Qunchao, et al.
Published: (2025)
Specific multi-emitter identification via multi-label learning
by: Chen, Yuhao, et al.
Published: (2025)
by: Chen, Yuhao, et al.
Published: (2025)
GAP-LA: GPU-Accelerated Performance-Driven Layer Assignment
by: Zhao, Chunyuan, et al.
Published: (2025)
by: Zhao, Chunyuan, et al.
Published: (2025)
Graphs with girth $2\ell+1$ and without longer odd holes that contain an odd $K_4$-subdivision
by: Chen, Rong, et al.
Published: (2022)
by: Chen, Rong, et al.
Published: (2022)
Optimizing treatment sequence for inoperable locally advanced breast cancer: Long‐term outcomes of surgery first versus neoadjuvant chemotherapy in a real‐world setting
by: Bowen Liu, et al.
Published: (2024)
by: Bowen Liu, et al.
Published: (2024)
Local Behavior of Fractional Equations in Grushin-type Spaces
by: Xu, Boxiang, et al.
Published: (2025)
by: Xu, Boxiang, et al.
Published: (2025)
VcLLM: Video Codecs are Secretly Tensor Codecs
by: Xu, Ceyu, et al.
Published: (2024)
by: Xu, Ceyu, et al.
Published: (2024)
Vector Search for the Future: From Memory-Resident, Static Heterogeneous Storage, to Cloud-Native Architectures
by: Song, Yitong, et al.
Published: (2026)
by: Song, Yitong, et al.
Published: (2026)
Exiting National Anti-Poverty Campaign, Social Support, and Improved Mental Health
by: Liu, Zhengwen, et al.
Published: (2025)
by: Liu, Zhengwen, et al.
Published: (2025)
Training-Free Multi-Style Fusion Through Reference-Based Adaptive Modulation
by: Liu, Xu, et al.
Published: (2025)
by: Liu, Xu, et al.
Published: (2025)
General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model
by: Wei, Haoran, et al.
Published: (2024)
by: Wei, Haoran, et al.
Published: (2024)
Event Data Association via Robust Model Fitting for Event-based Object Tracking
by: Chen, Haosheng, et al.
Published: (2021)
by: Chen, Haosheng, et al.
Published: (2021)
Synthesis imaging with a lunar orbit array: II. Impacts of instrument-induced phase errors
by: Zhou, Meng, et al.
Published: (2026)
by: Zhou, Meng, et al.
Published: (2026)
Nonlinear spiked covariance matrices and signal propagation in deep neural networks
by: Wang, Zhichao, et al.
Published: (2024)
by: Wang, Zhichao, et al.
Published: (2024)
FruitNinja: 3D Object Interior Texture Generation with Gaussian Splatting
by: Wu, Fangyu, et al.
Published: (2024)
by: Wu, Fangyu, et al.
Published: (2024)
“Fight Fire with Fire”: Trace‐Water‐Induced Controllable Synthesis of Hydrophobic Armors to Stabilize Metal‐Organic Cage‐Based Crystalline Frameworks
by: Ze‐Jiu Diao, et al.
Published: (2025)
by: Ze‐Jiu Diao, et al.
Published: (2025)
SimpleOCR: Rendering Visualized Questions to Teach MLLMs to Read
by: Peng, Yibo, et al.
Published: (2026)
by: Peng, Yibo, et al.
Published: (2026)
Optimal Planning and Control under Signal Temporal Logic Specifications
by: Pan, Zuodong, et al.
Published: (2026)
by: Pan, Zuodong, et al.
Published: (2026)
AutoSCORE: Enhancing Automated Scoring with Multi-Agent Large Language Models via Structured Component Recognition
by: Wang, Yun, et al.
Published: (2025)
by: Wang, Yun, et al.
Published: (2025)
Similar Items
-
FireRed-Image-Edit-1.0 Technical Report
by: Super Intelligence Team, et al.
Published: (2026) -
IVC-Prune: Revealing the Implicit Visual Coordinates in LVLMs for Vision Token Pruning
by: Sun, Zhichao, et al.
Published: (2026) -
Cross-Scenario Unified Modeling of User Interests at Billion Scale
by: Xu, Manjie, et al.
Published: (2025) -
CQ-DINO: Mitigating Gradient Dilution via Category Queries for Vast Vocabulary Object Detection
by: Sun, Zhichao, et al.
Published: (2025) -
SurveyBench: Can LLM(-Agents) Write Academic Surveys that Align with Reader Needs?
by: Sun, Zhaojun, et al.
Published: (2025)