Efficient Reasoning via Thought Compression for Language Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Qing, Zhang, Shiyu, Jia, Yuyu, Gao, Junyu, Ni, Weiping, Wu, Junzheng, Wang, Qi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Benchmark for Multi-Lingual Vision-Language Learning in Remote Sensing Image Captioning
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
Embedding Generalized Semantic Knowledge into Few-Shot Remote Sensing Segmentation
von: Jia, Yuyu, et al.
Veröffentlicht: (2024)
von: Jia, Yuyu, et al.
Veröffentlicht: (2024)
Like Humans to Few-Shot Learning through Knowledge Permeation of Vision and Text
von: Jia, Yuyu, et al.
Veröffentlicht: (2024)
von: Jia, Yuyu, et al.
Veröffentlicht: (2024)
Iterative Camera-LiDAR Extrinsic Optimization via Surrogate Diffusion
von: Ou, Ni, et al.
Veröffentlicht: (2025)
von: Ou, Ni, et al.
Veröffentlicht: (2025)
Iterative Camera-LiDAR Extrinsic Optimization via Surrogate Diffusion
von: Ou, Ni, et al.
Veröffentlicht: (2024)
von: Ou, Ni, et al.
Veröffentlicht: (2024)
Scale Efficient Training for Large Datasets
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
STAR-IOD: Scale-decoupled Topology Alignment with Pseudo-label Refinement for Remote Sensing Incremental Object Detection
von: Zhang, Yaoteng, et al.
Veröffentlicht: (2026)
von: Zhang, Yaoteng, et al.
Veröffentlicht: (2026)
Discriminative Perception via Anchored Description for Reasoning Segmentation
von: Yang, Tao, et al.
Veröffentlicht: (2026)
von: Yang, Tao, et al.
Veröffentlicht: (2026)
Native-Domain Cross-Attention for Camera-LiDAR Extrinsic Calibration Under Large Initial Perturbations
von: Ou, Ni, et al.
Veröffentlicht: (2026)
von: Ou, Ni, et al.
Veröffentlicht: (2026)
Beyond Prompt Degradation: Prototype-guided Dual-pool Prompting for Incremental Object Detection
von: Zhang, Yaoteng, et al.
Veröffentlicht: (2026)
von: Zhang, Yaoteng, et al.
Veröffentlicht: (2026)
ImgCoT: Compressing Long Chain of Thought into Compact Visual Tokens for Efficient Reasoning of Large Language Model
von: Chen, Xiaoshu, et al.
Veröffentlicht: (2026)
von: Chen, Xiaoshu, et al.
Veröffentlicht: (2026)
SamLP: A Customized Segment Anything Model for License Plate Detection
von: Ding, Haoxuan, et al.
Veröffentlicht: (2024)
von: Ding, Haoxuan, et al.
Veröffentlicht: (2024)
Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning
von: Jiang, Qing, et al.
Veröffentlicht: (2025)
von: Jiang, Qing, et al.
Veröffentlicht: (2025)
Video-based Sign Language Recognition without Temporal Segmentation
von: Huang, Jie, et al.
Veröffentlicht: (2018)
von: Huang, Jie, et al.
Veröffentlicht: (2018)
Chain-of-Thought Compression Should Not Be Blind: V-Skip for Efficient Multimodal Reasoning via Dual-Path Anchoring
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
ArgusCogito: Chain-of-Thought for Cross-Modal Synergy and Omnidirectional Reasoning in Camouflaged Object Segmentation
von: Tan, Jianwen, et al.
Veröffentlicht: (2025)
von: Tan, Jianwen, et al.
Veröffentlicht: (2025)
Prototype-Based Low Altitude UAV Semantic Segmentation
von: Zhang, Da, et al.
Veröffentlicht: (2026)
von: Zhang, Da, et al.
Veröffentlicht: (2026)
Explainable Action Form Assessment by Exploiting Multimodal Chain-of-Thoughts Reasoning
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
RSVP: Reasoning Segmentation via Visual Prompting and Multi-modal Chain-of-Thought
von: Lu, Yi, et al.
Veröffentlicht: (2025)
von: Lu, Yi, et al.
Veröffentlicht: (2025)
LISA: Reasoning Segmentation via Large Language Model
von: Lai, Xin, et al.
Veröffentlicht: (2023)
von: Lai, Xin, et al.
Veröffentlicht: (2023)
Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs
von: Zhang, Xintong, et al.
Veröffentlicht: (2025)
von: Zhang, Xintong, et al.
Veröffentlicht: (2025)
Video Evidence to Reasoning Efficient Video Understanding via Explicit Evidence Grounding
von: Huang, Yanxiang, et al.
Veröffentlicht: (2026)
von: Huang, Yanxiang, et al.
Veröffentlicht: (2026)
MSF-Net: Multi-Stage Feature Extraction and Fusion for Robust Photometric Stereo
von: Qin, Shiyu, et al.
Veröffentlicht: (2025)
von: Qin, Shiyu, et al.
Veröffentlicht: (2025)
VideoTIR: Accurate Understanding for Long Videos with Efficient Tool-Integrated Reasoning
von: Gao, Zhe, et al.
Veröffentlicht: (2026)
von: Gao, Zhe, et al.
Veröffentlicht: (2026)
Exploring Cross-Domain Few-Shot Classification via Frequency-Aware Prompting
von: Zhang, Tiange, et al.
Veröffentlicht: (2024)
von: Zhang, Tiange, et al.
Veröffentlicht: (2024)
HRGR: Enhancing Image Manipulation Detection via Hierarchical Region-aware Graph Reasoning
von: Wang, Xudong, et al.
Veröffentlicht: (2024)
von: Wang, Xudong, et al.
Veröffentlicht: (2024)
Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2024)
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2024)
PGR-Net: Prior-Guided ROI Reasoning Network for Brain Tumor MRI Segmentation
von: Lu, Jiacheng, et al.
Veröffentlicht: (2026)
von: Lu, Jiacheng, et al.
Veröffentlicht: (2026)
Image-of-Thought Prompting for Visual Reasoning Refinement in Multimodal Large Language Models
von: Zhou, Qiji, et al.
Veröffentlicht: (2024)
von: Zhou, Qiji, et al.
Veröffentlicht: (2024)
RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation
von: Wen, Junwei, et al.
Veröffentlicht: (2026)
von: Wen, Junwei, et al.
Veröffentlicht: (2026)
CoT-Segmenter: Enhancing OOD Detection in Dense Road Scenes via Chain-of-Thought Reasoning
von: Song, Jeonghyo, et al.
Veröffentlicht: (2025)
von: Song, Jeonghyo, et al.
Veröffentlicht: (2025)
Think Before You Segment: High-Quality Reasoning Segmentation with GPT Chain of Thoughts
von: Kao, Shiu-hong, et al.
Veröffentlicht: (2025)
von: Kao, Shiu-hong, et al.
Veröffentlicht: (2025)
Dynamic Proxy Domain Generalizes the Crowd Localization by Better Binary Segmentation
von: Gao, Junyu, et al.
Veröffentlicht: (2024)
von: Gao, Junyu, et al.
Veröffentlicht: (2024)
Batch Loss Score for Dynamic Data Pruning
von: Zhou, Qing, et al.
Veröffentlicht: (2026)
von: Zhou, Qing, et al.
Veröffentlicht: (2026)
TokenSeg: Efficient 3D Medical Image Segmentation via Hierarchical Visual Token Compression
von: Zeng, Sen, et al.
Veröffentlicht: (2026)
von: Zeng, Sen, et al.
Veröffentlicht: (2026)
ViLLa: Video Reasoning Segmentation with Large Language Model
von: Zheng, Rongkun, et al.
Veröffentlicht: (2024)
von: Zheng, Rongkun, et al.
Veröffentlicht: (2024)
LL-ICM: Image Compression for Low-level Machine Vision via Large Vision-Language Model
von: Xue, Yuan, et al.
Veröffentlicht: (2024)
von: Xue, Yuan, et al.
Veröffentlicht: (2024)
SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
von: Xu, Guoan, et al.
Veröffentlicht: (2024)
CoT-RVS: Zero-Shot Chain-of-Thought Reasoning Segmentation for Videos
von: Kao, Shiu-hong, et al.
Veröffentlicht: (2025)
von: Kao, Shiu-hong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Benchmark for Multi-Lingual Vision-Language Learning in Remote Sensing Image Captioning
von: Zhou, Qing, et al.
Veröffentlicht: (2025) -
Embedding Generalized Semantic Knowledge into Few-Shot Remote Sensing Segmentation
von: Jia, Yuyu, et al.
Veröffentlicht: (2024) -
Like Humans to Few-Shot Learning through Knowledge Permeation of Vision and Text
von: Jia, Yuyu, et al.
Veröffentlicht: (2024) -
Iterative Camera-LiDAR Extrinsic Optimization via Surrogate Diffusion
von: Ou, Ni, et al.
Veröffentlicht: (2025) -
Iterative Camera-LiDAR Extrinsic Optimization via Surrogate Diffusion
von: Ou, Ni, et al.
Veröffentlicht: (2024)