Texo: Formula Recognition within 20M Parameters
Fuente:
arXiv
Saved in:
| Main Author: | Mao, Sicheng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PP-FormulaNet: Bridging Accuracy and Efficiency in Advanced Formula Recognition
by: Liu, Hongen, et al.
Published: (2025)
by: Liu, Hongen, et al.
Published: (2025)
AutoMR: A Universal Time Series Motion Recognition Pipeline
by: Zhang, Likun, et al.
Published: (2025)
by: Zhang, Likun, et al.
Published: (2025)
Triple-domain Feature Learning with Frequency-aware Memory Enhancement for Moving Infrared Small Target Detection
by: Duan, Weiwei, et al.
Published: (2024)
by: Duan, Weiwei, et al.
Published: (2024)
MPT-PAR:Mix-Parameters Transformer for Panoramic Activity Recognition
by: Gan, Wenqing, et al.
Published: (2024)
by: Gan, Wenqing, et al.
Published: (2024)
Bridging Coarse and Fine Recognition: A Hybrid Approach for Open-Ended Multi-Granularity Object Recognition in Interactive Educational Games
by: Yi, Hanling, et al.
Published: (2026)
by: Yi, Hanling, et al.
Published: (2026)
Enhancing Table Recognition with Vision LLMs: A Benchmark and Neighbor-Guided Toolchain Reasoner
by: Zhou, Yitong, et al.
Published: (2024)
by: Zhou, Yitong, et al.
Published: (2024)
Towards A Robust Group-level Emotion Recognition via Uncertainty-Aware Learning
by: Zhu, Qing, et al.
Published: (2023)
by: Zhu, Qing, et al.
Published: (2023)
Feature Fusion for Human Activity Recognition using Parameter-Optimized Multi-Stage Graph Convolutional Network and Transformer Models
by: Belal, Mohammad, et al.
Published: (2024)
by: Belal, Mohammad, et al.
Published: (2024)
Toward Optimal Sampling Rate Selection and Unbiased Classification for Precise Animal Activity Recognition
by: Mao, Axiu, et al.
Published: (2026)
by: Mao, Axiu, et al.
Published: (2026)
SmolRGPT: Efficient Spatial Reasoning for Warehouse Environments with 600M Parameters
by: Traore, Abdarahmane, et al.
Published: (2025)
by: Traore, Abdarahmane, et al.
Published: (2025)
SSMamba: A Self-Supervised Hybrid State Space Model for Pathological Image Classification
by: Chai, Enhui, et al.
Published: (2026)
by: Chai, Enhui, et al.
Published: (2026)
MoireMix: A Formula-Based Data Augmentation for Improving Image Classification Robustness
by: Matsuo, Yuto, et al.
Published: (2026)
by: Matsuo, Yuto, et al.
Published: (2026)
TCFormer: A 5M-Parameter Transformer with Density-Guided Aggregation for Weakly-Supervised Crowd Counting
by: Guo, Qiang, et al.
Published: (2025)
by: Guo, Qiang, et al.
Published: (2025)
MambaBack: Bridging Local Features and Global Contexts in Whole Slide Image Analysis
by: Chen, Sicheng, et al.
Published: (2026)
by: Chen, Sicheng, et al.
Published: (2026)
SegMix:Shuffle-based Feedback Learning for Semantic Segmentation of Pathology Images
by: Yan, Zhiling, et al.
Published: (2026)
by: Yan, Zhiling, et al.
Published: (2026)
RewardMap: Tackling Sparse Rewards in Fine-grained Visual Reasoning via Multi-Stage Reinforcement Learning
by: Feng, Sicheng, et al.
Published: (2025)
by: Feng, Sicheng, et al.
Published: (2025)
AdaTP: Attention-Debiased Token Pruning for Video Large Language Models
by: Sun, Fengyuan, et al.
Published: (2025)
by: Sun, Fengyuan, et al.
Published: (2025)
A$^2$M$^2$-Net: Adaptively Aligned Multi-Scale Moment for Few-Shot Action Recognition
by: Gao, Zilin, et al.
Published: (2025)
by: Gao, Zilin, et al.
Published: (2025)
Geometry-Aware State Space Model: A New Paradigm for Whole-Slide Image Representation
by: Chai, Enhui, et al.
Published: (2026)
by: Chai, Enhui, et al.
Published: (2026)
WebChain: A Large-Scale Human-Annotated Dataset of Real-World Web Interaction Traces
by: Fan, Sicheng, et al.
Published: (2026)
by: Fan, Sicheng, et al.
Published: (2026)
BALR-SAM: Boundary-Aware Low-Rank Adaptation of SAM for Resource-Efficient Medical Image Segmentation
by: Liu, Zelin, et al.
Published: (2025)
by: Liu, Zelin, et al.
Published: (2025)
LLMI3D: MLLM-based 3D Perception from a Single 2D Image
by: Yang, Fan, et al.
Published: (2024)
by: Yang, Fan, et al.
Published: (2024)
VASA-3D: Lifelike Audio-Driven Gaussian Head Avatars from a Single Image
by: Xu, Sicheng, et al.
Published: (2025)
by: Xu, Sicheng, et al.
Published: (2025)
Generalizable Facial Expression Recognition
by: Zhang, Yuhang, et al.
Published: (2024)
by: Zhang, Yuhang, et al.
Published: (2024)
Adversarial Watermarking for Face Recognition
by: Yao, Yuguang, et al.
Published: (2024)
by: Yao, Yuguang, et al.
Published: (2024)
M3GCLR: Multi-View Mini-Max Infinite Skeleton-Data Game Contrastive Learning For Skeleton-Based Action Recognition
by: Li, Yanshan, et al.
Published: (2026)
by: Li, Yanshan, et al.
Published: (2026)
OVFoodSeg: Elevating Open-Vocabulary Food Image Segmentation via Image-Informed Textual Representation
by: Wu, Xiongwei, et al.
Published: (2024)
by: Wu, Xiongwei, et al.
Published: (2024)
Study of detecting behavioral signatures within DeepFake videos
by: Miao, Qiaomu, et al.
Published: (2022)
by: Miao, Qiaomu, et al.
Published: (2022)
e5-omni: Explicit Cross-modal Alignment for Omni-modal Embeddings
by: Chen, Haonan, et al.
Published: (2026)
by: Chen, Haonan, et al.
Published: (2026)
Parameter-Efficient Active Learning for Foundational models
by: Narayanan, Athmanarayanan Lakshmi, et al.
Published: (2024)
by: Narayanan, Athmanarayanan Lakshmi, et al.
Published: (2024)
Towards Understanding Deep Learning Model in Image Recognition via Coverage Test
by: Li, Wenkai, et al.
Published: (2025)
by: Li, Wenkai, et al.
Published: (2025)
A Survey of Deep Learning for Group-level Emotion Recognition
by: Huang, Xiaohua, et al.
Published: (2024)
by: Huang, Xiaohua, et al.
Published: (2024)
Bridge then Begin Anew: Generating Target-relevant Intermediate Model for Source-free Visual Emotion Adaptation
by: Zhu, Jiankun, et al.
Published: (2024)
by: Zhu, Jiankun, et al.
Published: (2024)
HEIE: MLLM-Based Hierarchical Explainable AIGC Image Implausibility Evaluator
by: Yang, Fan, et al.
Published: (2024)
by: Yang, Fan, et al.
Published: (2024)
Weaknesses of Facial Emotion Recognition Systems
by: Jamróz, Aleksandra, et al.
Published: (2026)
by: Jamróz, Aleksandra, et al.
Published: (2026)
Deep Ensemble Art Style Recognition
by: Menis-Mastromichalakis, Orfeas, et al.
Published: (2024)
by: Menis-Mastromichalakis, Orfeas, et al.
Published: (2024)
Exploring Explainability in Video Action Recognition
by: Saha, Avinab, et al.
Published: (2024)
by: Saha, Avinab, et al.
Published: (2024)
Handwritten Text Recognition: A Survey
by: Garrido-Munoz, Carlos, et al.
Published: (2025)
by: Garrido-Munoz, Carlos, et al.
Published: (2025)
On Denoising Walking Videos for Gait Recognition
by: Jin, Dongyang, et al.
Published: (2025)
by: Jin, Dongyang, et al.
Published: (2025)
Robust Dynamic Facial Expression Recognition
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
Similar Items
-
PP-FormulaNet: Bridging Accuracy and Efficiency in Advanced Formula Recognition
by: Liu, Hongen, et al.
Published: (2025) -
AutoMR: A Universal Time Series Motion Recognition Pipeline
by: Zhang, Likun, et al.
Published: (2025) -
Triple-domain Feature Learning with Frequency-aware Memory Enhancement for Moving Infrared Small Target Detection
by: Duan, Weiwei, et al.
Published: (2024) -
MPT-PAR:Mix-Parameters Transformer for Panoramic Activity Recognition
by: Gan, Wenqing, et al.
Published: (2024) -
Bridging Coarse and Fine Recognition: A Hybrid Approach for Open-Ended Multi-Granularity Object Recognition in Interactive Educational Games
by: Yi, Hanling, et al.
Published: (2026)