Guardado en:
| Autores principales: | Zeng, Boyi, Hao, Yiqin, Li, He, Song, Shixiang, Song, Feichen, Wang, Zitong, Huang, Siyuan, Xu, Yi, He, ZiWei, Wang, Xinbing, Lin, Zhouhan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.08220 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PonderLM-2: Pretraining LLM with Latent Thoughts in Continuous Space
por: Zeng, Boyi, et al.
Publicado: (2025)
por: Zeng, Boyi, et al.
Publicado: (2025)
AdaPonderLM: Gated Pondering Language Models with Token-Wise Adaptive Depth
por: Song, Shixiang, et al.
Publicado: (2026)
por: Song, Shixiang, et al.
Publicado: (2026)
PonderLM-3: Adaptive Token-Wise Pondering with Differentiable Masking
por: Li, He, et al.
Publicado: (2026)
por: Li, He, et al.
Publicado: (2026)
PonderLM: Pretraining Language Models to Ponder in Continuous Space
por: Zeng, Boyi, et al.
Publicado: (2025)
por: Zeng, Boyi, et al.
Publicado: (2025)
AWM: Accurate Weight-Matrix Fingerprint for Large Language Models
por: Zeng, Boyi, et al.
Publicado: (2025)
por: Zeng, Boyi, et al.
Publicado: (2025)
Graph Parsing Networks
por: Song, Yunchong, et al.
Publicado: (2024)
por: Song, Yunchong, et al.
Publicado: (2024)
HuRef: HUman-REadable Fingerprint for Large Language Models
por: Zeng, Boyi, et al.
Publicado: (2023)
por: Zeng, Boyi, et al.
Publicado: (2023)
Cluster-wise Graph Transformer with Dual-granularity Kernelized Attention
por: Huang, Siyuan, et al.
Publicado: (2024)
por: Huang, Siyuan, et al.
Publicado: (2024)
FreqKV: Key-Value Compression in Frequency Domain for Context Window Extension
por: Kai, Jushi, et al.
Publicado: (2025)
por: Kai, Jushi, et al.
Publicado: (2025)
Diagnosing Memorization in Chain-of-Thought Reasoning, One Token at a Time
por: Li, Huihan, et al.
Publicado: (2025)
por: Li, Huihan, et al.
Publicado: (2025)
Fourier Transformer: Fast Long Range Modeling by Removing Sequence Redundancy with FFT Operator
por: He, Ziwei, et al.
Publicado: (2023)
por: He, Ziwei, et al.
Publicado: (2023)
Effects of Perceived Control in the Relationship Between Psychological Distress and Posttraumatic Growth After Glioma Diagnosis: A Longitudinal Mediation Analysis
por: Xu ZiWei, et al.
Publicado: (2025)
por: Xu ZiWei, et al.
Publicado: (2025)
Flow of Spans: Generalizing Language Models to Dynamic Span-Vocabulary via GFlowNets
por: Xue, Bo, et al.
Publicado: (2026)
por: Xue, Bo, et al.
Publicado: (2026)
One Size Does Not Fit All: Token-Wise Adaptive Compression for KV Cache
por: Lu, Liming, et al.
Publicado: (2026)
por: Lu, Liming, et al.
Publicado: (2026)
CoT-X: An Adaptive Framework for Cross-Model Chain-of-Thought Transfer and Optimization
por: Bi, Ziqian, et al.
Publicado: (2025)
por: Bi, Ziqian, et al.
Publicado: (2025)
LLM Reasoning Is Latent, Not the Chain of Thought
por: Wang, Wenshuo
Publicado: (2026)
por: Wang, Wenshuo
Publicado: (2026)
Context-level Language Modeling by Learning Predictive Context Embeddings
por: Dai, Beiya, et al.
Publicado: (2025)
por: Dai, Beiya, et al.
Publicado: (2025)
Latent Chain-of-Thought as Planning: Decoupling Reasoning from Verbalization
por: Wang, Jiecong, et al.
Publicado: (2026)
por: Wang, Jiecong, et al.
Publicado: (2026)
Towards Controlled Table-to-Text Generation with Scientific Reasoning
por: Guo, Zhixin, et al.
Publicado: (2023)
por: Guo, Zhixin, et al.
Publicado: (2023)
Fourier Compressor: Frequency-Domain Visual Token Compression for Vision-Language Models
por: Wang, Huanyu, et al.
Publicado: (2025)
por: Wang, Huanyu, et al.
Publicado: (2025)
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN
por: Xu, Yao, et al.
Publicado: (2025)
por: Xu, Yao, et al.
Publicado: (2025)
GeoGalactica: A Scientific Large Language Model in Geoscience
por: Lin, Zhouhan, et al.
Publicado: (2023)
por: Lin, Zhouhan, et al.
Publicado: (2023)
Accelerating Structured Chain-of-Thought in Autonomous Vehicles
por: Gu, Yi, et al.
Publicado: (2026)
por: Gu, Yi, et al.
Publicado: (2026)
Learning Modal-Mixed Chain-of-Thought Reasoning with Latent Embeddings
por: Shao, Yifei, et al.
Publicado: (2026)
por: Shao, Yifei, et al.
Publicado: (2026)
Prehabilitation enhances functional and structural recovery following anterior cruciate ligament reconstruction: A randomized controlled trial
por: Yuping Fu, et al.
Publicado: (2025)
por: Yuping Fu, et al.
Publicado: (2025)
Rethinking Token-Level Policy Optimization for Multimodal Chain-of-Thought
por: Li, Yunheng, et al.
Publicado: (2026)
por: Li, Yunheng, et al.
Publicado: (2026)
Toward the $p$-adic Hodge parameters in the potentially crystalline representations of $\mathrm{GL}_n$
por: He, Yiqin
Publicado: (2026)
por: He, Yiqin
Publicado: (2026)
Companion points and locally analytic socle conjecture for Steinberg case
por: He, Yiqin
Publicado: (2024)
por: He, Yiqin
Publicado: (2024)
Towards the $p$-adic Hodge parameters in semistable representations of $\mathrm{GL}_n(\mathrm{Q}_p)$
por: He, Yiqin
Publicado: (2026)
por: He, Yiqin
Publicado: (2026)
Chain-of-Thought Tokens are Computer Program Variables
por: Zhu, Fangwei, et al.
Publicado: (2025)
por: Zhu, Fangwei, et al.
Publicado: (2025)
Latent Chain-of-Thought for Visual Reasoning
por: Sun, Guohao, et al.
Publicado: (2025)
por: Sun, Guohao, et al.
Publicado: (2025)
FaithCoT-Bench: Benchmarking Instance-Level Faithfulness of Chain-of-Thought Reasoning
por: Shen, Xu, et al.
Publicado: (2025)
por: Shen, Xu, et al.
Publicado: (2025)
Mirror-Consistency: Harnessing Inconsistency in Majority Voting
por: Huang, Siyuan, et al.
Publicado: (2024)
por: Huang, Siyuan, et al.
Publicado: (2024)
Investigating the Fundamental Limit: A Feasibility Study of Hybrid-Neural Archival
por: Armstrong, Marcus, et al.
Publicado: (2026)
por: Armstrong, Marcus, et al.
Publicado: (2026)
Towards Enhanced Image Generation Via Multi-modal Chain of Thought in Unified Generative Models
por: Wang, Yi, et al.
Publicado: (2025)
por: Wang, Yi, et al.
Publicado: (2025)
Imbalanced Graph-Level Anomaly Detection via Counterfactual Augmentation and Feature Learning
por: Wang, Zitong, et al.
Publicado: (2024)
por: Wang, Zitong, et al.
Publicado: (2024)
Mitigating Premature Discretization with Progressive Quantization for Robust Vector Tokenization
por: Zhao, Wenhao, et al.
Publicado: (2026)
por: Zhao, Wenhao, et al.
Publicado: (2026)
Rethinking Chain-of-Thought Reasoning for Videos
por: Zhong, Yiwu, et al.
Publicado: (2025)
por: Zhong, Yiwu, et al.
Publicado: (2025)
Inference-Time Chain-of-Thought Pruning with Latent Informativeness Signals
por: Li, Sophie, et al.
Publicado: (2025)
por: Li, Sophie, et al.
Publicado: (2025)
SemCoT: Accelerating Chain-of-Thought Reasoning through Semantically-Aligned Implicit Tokens
por: He, Yinhan, et al.
Publicado: (2025)
por: He, Yinhan, et al.
Publicado: (2025)
Ejemplares similares
-
PonderLM-2: Pretraining LLM with Latent Thoughts in Continuous Space
por: Zeng, Boyi, et al.
Publicado: (2025) -
AdaPonderLM: Gated Pondering Language Models with Token-Wise Adaptive Depth
por: Song, Shixiang, et al.
Publicado: (2026) -
PonderLM-3: Adaptive Token-Wise Pondering with Differentiable Masking
por: Li, He, et al.
Publicado: (2026) -
PonderLM: Pretraining Language Models to Ponder in Continuous Space
por: Zeng, Boyi, et al.
Publicado: (2025) -
AWM: Accurate Weight-Matrix Fingerprint for Large Language Models
por: Zeng, Boyi, et al.
Publicado: (2025)