FBS: Modeling Native Parallel Reading inside a Transformer
Fuente:
arXiv
Salvato in:
| Autore principale: | Wang, Tongxi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ChiEngMixBench: Evaluating Large Language Models on Spontaneous and Natural Chinese-English Code-Mixed Generation
di: Yang, Qingyan, et al.
Pubblicazione: (2026)
di: Yang, Qingyan, et al.
Pubblicazione: (2026)
ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute
di: Wen, Hao, et al.
Pubblicazione: (2025)
di: Wen, Hao, et al.
Pubblicazione: (2025)
Automated Extraction and Creation of FBS Design Reasoning Knowledge Graphs from Structured Data in Product Catalogues Lacking Contextual Information
di: Sahadevan, Vijayalaxmi, et al.
Pubblicazione: (2024)
di: Sahadevan, Vijayalaxmi, et al.
Pubblicazione: (2024)
Parallel Decoder Transformer: Planner-Seeded Latent Coordination for Synchronized Parallel Decoding
di: Robbins, Logan
Pubblicazione: (2025)
di: Robbins, Logan
Pubblicazione: (2025)
Applications of the Transformer Architecture in AI-Assisted English Reading Comprehension
di: Li, Ping
Pubblicazione: (2026)
di: Li, Ping
Pubblicazione: (2026)
NAG: A Unified Native Architecture for Encoder-free Text-Graph Modeling in Language Models
di: Gong, Haisong, et al.
Pubblicazione: (2026)
di: Gong, Haisong, et al.
Pubblicazione: (2026)
TransformLLM: Adapting Large Language Models via LLM-Transformed Reading Comprehension Text
di: Arbel, Iftach, et al.
Pubblicazione: (2024)
di: Arbel, Iftach, et al.
Pubblicazione: (2024)
Sharp Spectral Thresholds for Logit Fixed Points
di: Wang, Tongxi
Pubblicazione: (2026)
di: Wang, Tongxi
Pubblicazione: (2026)
Improving Context Fidelity via Native Retrieval-Augmented Reasoning
di: Wang, Suyuchen, et al.
Pubblicazione: (2025)
di: Wang, Suyuchen, et al.
Pubblicazione: (2025)
Characterizing Model-Native Skills
di: Kang, Feiyang, et al.
Pubblicazione: (2026)
di: Kang, Feiyang, et al.
Pubblicazione: (2026)
Learning Adaptive Parallel Reasoning with Language Models
di: Pan, Jiayi, et al.
Pubblicazione: (2025)
di: Pan, Jiayi, et al.
Pubblicazione: (2025)
Multilingual Multi-Aspect Explainability Analyses on Machine Reading Comprehension Models
di: Cui, Yiming, et al.
Pubblicazione: (2021)
di: Cui, Yiming, et al.
Pubblicazione: (2021)
Read Before You Think: Mitigating LLM Comprehension Failures with Step-by-Step Reading
di: Han, Feijiang, et al.
Pubblicazione: (2025)
di: Han, Feijiang, et al.
Pubblicazione: (2025)
ViRanker: A BGE-M3 & Blockwise Parallel Transformer Cross-Encoder for Vietnamese Reranking
di: Dang, Phuong-Nam, et al.
Pubblicazione: (2025)
di: Dang, Phuong-Nam, et al.
Pubblicazione: (2025)
Accelerating Transformer Inference for Translation via Parallel Decoding
di: Santilli, Andrea, et al.
Pubblicazione: (2023)
di: Santilli, Andrea, et al.
Pubblicazione: (2023)
ParallelMuse: Agentic Parallel Thinking for Deep Information Seeking
di: Li, Baixuan, et al.
Pubblicazione: (2025)
di: Li, Baixuan, et al.
Pubblicazione: (2025)
Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
di: Yuan, Jingyang, et al.
Pubblicazione: (2025)
di: Yuan, Jingyang, et al.
Pubblicazione: (2025)
Automatically Planning Optimal Parallel Strategy for Large Language Models
di: Li, Zongbiao, et al.
Pubblicazione: (2024)
di: Li, Zongbiao, et al.
Pubblicazione: (2024)
CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit
di: Wang, Kangyu, et al.
Pubblicazione: (2025)
di: Wang, Kangyu, et al.
Pubblicazione: (2025)
Parallel Universes, Parallel Languages: A Comprehensive Study on LLM-based Multilingual Counterfactual Example Generation
di: Wang, Qianli, et al.
Pubblicazione: (2026)
di: Wang, Qianli, et al.
Pubblicazione: (2026)
Skeleton-of-Thought: Prompting LLMs for Efficient Parallel Generation
di: Ning, Xuefei, et al.
Pubblicazione: (2023)
di: Ning, Xuefei, et al.
Pubblicazione: (2023)
CSCD-NS: a Chinese Spelling Check Dataset for Native Speakers
di: Hu, Yong, et al.
Pubblicazione: (2022)
di: Hu, Yong, et al.
Pubblicazione: (2022)
PaDeLLM-NER: Parallel Decoding in Large Language Models for Named Entity Recognition
di: Lu, Jinghui, et al.
Pubblicazione: (2024)
di: Lu, Jinghui, et al.
Pubblicazione: (2024)
Native Hybrid Attention for Efficient Sequence Modeling
di: Du, Jusen, et al.
Pubblicazione: (2025)
di: Du, Jusen, et al.
Pubblicazione: (2025)
ASPD: Unlocking Adaptive Serial-Parallel Decoding by Exploring Intrinsic Parallelism in LLMs
di: Chen, Keyu, et al.
Pubblicazione: (2025)
di: Chen, Keyu, et al.
Pubblicazione: (2025)
Just Go Parallel: Improving the Multilingual Capabilities of Large Language Models
di: Qorib, Muhammad Reza, et al.
Pubblicazione: (2025)
di: Qorib, Muhammad Reza, et al.
Pubblicazione: (2025)
Improving Multilingual Capabilities with Cultural and Local Knowledge in Large Language Models While Enhancing Native Performance
di: Kadiyala, Ram Mohan Rao, et al.
Pubblicazione: (2025)
di: Kadiyala, Ram Mohan Rao, et al.
Pubblicazione: (2025)
Beyond Facts: Benchmarking Distributional Reading Comprehension in Large Language Models
di: Guo, Pei-Fu, et al.
Pubblicazione: (2026)
di: Guo, Pei-Fu, et al.
Pubblicazione: (2026)
Read Quietly, Think Aloud: Decoupling Comprehension and Reasoning in LLMs
di: Wang, Yuanxin, et al.
Pubblicazione: (2025)
di: Wang, Yuanxin, et al.
Pubblicazione: (2025)
Generating Reading Comprehension Exercises with Large Language Models for Educational Applications
di: Huang, Xingyu, et al.
Pubblicazione: (2025)
di: Huang, Xingyu, et al.
Pubblicazione: (2025)
Why Diffusion Language Models Struggle with Truly Parallel (Non-Autoregressive) Decoding?
di: Li, Pengxiang, et al.
Pubblicazione: (2026)
di: Li, Pengxiang, et al.
Pubblicazione: (2026)
EPIC: Efficient and Parallel Inference under CFG Constraints for Diffusion Language Models
di: Jin, Hyundong, et al.
Pubblicazione: (2026)
di: Jin, Hyundong, et al.
Pubblicazione: (2026)
Multilingual and Explainable Text Detoxification with Parallel Corpora
di: Dementieva, Daryna, et al.
Pubblicazione: (2024)
di: Dementieva, Daryna, et al.
Pubblicazione: (2024)
What Should Baby Models Read? Exploring Sample-Efficient Data Composition on Model Performance
di: Yam, Hong Meng, et al.
Pubblicazione: (2024)
di: Yam, Hong Meng, et al.
Pubblicazione: (2024)
NOSA: Native and Offloadable Sparse Attention
di: Huang, Yuxiang, et al.
Pubblicazione: (2025)
di: Huang, Yuxiang, et al.
Pubblicazione: (2025)
Automatic Generation of Inference Making Questions for Reading Comprehension Assessments
di: Ma, Wanjing Anya, et al.
Pubblicazione: (2025)
di: Ma, Wanjing Anya, et al.
Pubblicazione: (2025)
Reading Comprehension using Entity-based Memory Network
di: Wang, Xun, et al.
Pubblicazione: (2016)
di: Wang, Xun, et al.
Pubblicazione: (2016)
ArabicNumBench: Evaluating Arabic Number Reading in Large Language Models
di: Alhumud, Anas, et al.
Pubblicazione: (2026)
di: Alhumud, Anas, et al.
Pubblicazione: (2026)
PARD-2: Target-Aligned Parallel Draft Model for Dual-Mode Speculative Decoding
di: An, Zihao, et al.
Pubblicazione: (2026)
di: An, Zihao, et al.
Pubblicazione: (2026)
Parallel Test-Time Scaling for Latent Reasoning Models
di: You, Runyang, et al.
Pubblicazione: (2025)
di: You, Runyang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
ChiEngMixBench: Evaluating Large Language Models on Spontaneous and Natural Chinese-English Code-Mixed Generation
di: Yang, Qingyan, et al.
Pubblicazione: (2026) -
ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute
di: Wen, Hao, et al.
Pubblicazione: (2025) -
Automated Extraction and Creation of FBS Design Reasoning Knowledge Graphs from Structured Data in Product Catalogues Lacking Contextual Information
di: Sahadevan, Vijayalaxmi, et al.
Pubblicazione: (2024) -
Parallel Decoder Transformer: Planner-Seeded Latent Coordination for Synchronized Parallel Decoding
di: Robbins, Logan
Pubblicazione: (2025) -
Applications of the Transformer Architecture in AI-Assisted English Reading Comprehension
di: Li, Ping
Pubblicazione: (2026)