Gespeichert in:
| 1. Verfasser: | Zhou, Ruiyang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2402.12249 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Certified Robustness Under Bounded Levenshtein Distance
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025)
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025)
A Comprehensive Approach to Misspelling Correction with BERT and Levenshtein Distance
von: Naziri, Amirreza, et al.
Veröffentlicht: (2024)
von: Naziri, Amirreza, et al.
Veröffentlicht: (2024)
Efficient Encoder-Decoder Transformer Decoding for Decomposable Tasks
von: Lu, Bo-Ru, et al.
Veröffentlicht: (2024)
von: Lu, Bo-Ru, et al.
Veröffentlicht: (2024)
DecoderLens: Layerwise Interpretation of Encoder-Decoder Transformers
von: Langedijk, Anna, et al.
Veröffentlicht: (2023)
von: Langedijk, Anna, et al.
Veröffentlicht: (2023)
Beyond Levenshtein: Leveraging Multiple Algorithms for Robust Word Error Rate Computations And Granular Error Classifications
von: Kuhn, Korbinian, et al.
Veröffentlicht: (2024)
von: Kuhn, Korbinian, et al.
Veröffentlicht: (2024)
Guided Decoding and Its Critical Role in Retrieval-Augmented Generation
von: Uğur, Özgür, et al.
Veröffentlicht: (2025)
von: Uğur, Özgür, et al.
Veröffentlicht: (2025)
ExPO: Unlocking Hard Reasoning with Self-Explanation-Guided Reinforcement Learning
von: Zhou, Ruiyang, et al.
Veröffentlicht: (2025)
von: Zhou, Ruiyang, et al.
Veröffentlicht: (2025)
A Decoding Algorithm for Length-Control Summarization Based on Directed Acyclic Transformers
von: Huang, Chenyang, et al.
Veröffentlicht: (2025)
von: Huang, Chenyang, et al.
Veröffentlicht: (2025)
Decoder-only Streaming Transformer for Simultaneous Translation
von: Guo, Shoutao, et al.
Veröffentlicht: (2024)
von: Guo, Shoutao, et al.
Veröffentlicht: (2024)
Linearity of Relation Decoding in Transformer Language Models
von: Hernandez, Evan, et al.
Veröffentlicht: (2023)
von: Hernandez, Evan, et al.
Veröffentlicht: (2023)
On The Adaptation of Unlimiformer for Decoder-Only Transformers
von: Ahrabian, Kian, et al.
Veröffentlicht: (2024)
von: Ahrabian, Kian, et al.
Veröffentlicht: (2024)
Top 10 Open Challenges Steering the Future of Diffusion Language Model and Its Variants
von: Wang, Yunhe, et al.
Veröffentlicht: (2026)
von: Wang, Yunhe, et al.
Veröffentlicht: (2026)
NDT: Non-Differential Transformer and Its Application to Sentiment Analysis
von: Ghoshal, Soudeep, et al.
Veröffentlicht: (2026)
von: Ghoshal, Soudeep, et al.
Veröffentlicht: (2026)
SMRC: Aligning Large Language Models with Student Reasoning for Mathematical Error Correction
von: Zeng, Biaojie, et al.
Veröffentlicht: (2025)
von: Zeng, Biaojie, et al.
Veröffentlicht: (2025)
Parallel Decoder Transformer: Planner-Seeded Latent Coordination for Synchronized Parallel Decoding
von: Robbins, Logan
Veröffentlicht: (2025)
von: Robbins, Logan
Veröffentlicht: (2025)
Combining Constrained and Unconstrained Decoding via Boosting: BoostCD and Its Application to Information Extraction
von: Šakota, Marija, et al.
Veröffentlicht: (2025)
von: Šakota, Marija, et al.
Veröffentlicht: (2025)
Personalized Language Modeling from Personalized Human Feedback
von: Li, Xinyu, et al.
Veröffentlicht: (2024)
von: Li, Xinyu, et al.
Veröffentlicht: (2024)
Speculative Contrastive Decoding
von: Yuan, Hongyi, et al.
Veröffentlicht: (2023)
von: Yuan, Hongyi, et al.
Veröffentlicht: (2023)
LayerNorm Induces Recency Bias in Transformer Decoders
von: Kim, Junu, et al.
Veröffentlicht: (2025)
von: Kim, Junu, et al.
Veröffentlicht: (2025)
Dissociating Decodability and Causal Use in Bracket-Sequence Transformers
von: Sharma, Aryan, et al.
Veröffentlicht: (2026)
von: Sharma, Aryan, et al.
Veröffentlicht: (2026)
How Powerful are Decoder-Only Transformer Neural Models?
von: Roberts, Jesse
Veröffentlicht: (2023)
von: Roberts, Jesse
Veröffentlicht: (2023)
Generative Sentiment Analysis via Latent Category Distribution and Constrained Decoding
von: Zhou, Jun, et al.
Veröffentlicht: (2024)
von: Zhou, Jun, et al.
Veröffentlicht: (2024)
StableMask: Refining Causal Masking in Decoder-only Transformer
von: Yin, Qingyu, et al.
Veröffentlicht: (2024)
von: Yin, Qingyu, et al.
Veröffentlicht: (2024)
Faster Transformer Decoding: N-gram Masked Self-Attention
von: Chelba, Ciprian, et al.
Veröffentlicht: (2020)
von: Chelba, Ciprian, et al.
Veröffentlicht: (2020)
From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection
von: Zhou, Hongxu
Veröffentlicht: (2026)
von: Zhou, Hongxu
Veröffentlicht: (2026)
Enabling On-Device Large Language Model Personalization with Self-Supervised Data Selection and Synthesis
von: Qin, Ruiyang, et al.
Veröffentlicht: (2023)
von: Qin, Ruiyang, et al.
Veröffentlicht: (2023)
Bridging Logic and Learning: Decoding Temporal Logic Embeddings via Transformers
von: Candussio, Sara, et al.
Veröffentlicht: (2025)
von: Candussio, Sara, et al.
Veröffentlicht: (2025)
Encoder-Decoder or Decoder-Only? Revisiting Encoder-Decoder Large Language Model
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
Accelerating Transformer Inference for Translation via Parallel Decoding
von: Santilli, Andrea, et al.
Veröffentlicht: (2023)
von: Santilli, Andrea, et al.
Veröffentlicht: (2023)
Latent Chain-of-Thought? Decoding the Depth-Recurrent Transformer
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
Decoding Student Minds: Leveraging Conversational Agents for Psychological and Learning Analysis
von: Chaabene, Nour El Houda Ben, et al.
Veröffentlicht: (2025)
von: Chaabene, Nour El Houda Ben, et al.
Veröffentlicht: (2025)
TraceMem: Weaving Narrative Memory Schemata from User Conversational Traces
von: Shu, Yiming, et al.
Veröffentlicht: (2026)
von: Shu, Yiming, et al.
Veröffentlicht: (2026)
Flexible and Efficient Grammar-Constrained Decoding
von: Park, Kanghee, et al.
Veröffentlicht: (2025)
von: Park, Kanghee, et al.
Veröffentlicht: (2025)
A One-Layer Decoder-Only Transformer is a Two-Layer RNN: With an Application to Certified Robustness
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
Decoding Speculative Decoding
von: Yan, Minghao, et al.
Veröffentlicht: (2024)
von: Yan, Minghao, et al.
Veröffentlicht: (2024)
Quality-Aware Decoding: Unifying Quality Estimation and Decoding
von: Koneru, Sai, et al.
Veröffentlicht: (2025)
von: Koneru, Sai, et al.
Veröffentlicht: (2025)
Decoding the Flow: CauseMotion for Emotional Causality Analysis in Long-form Conversations
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2025)
Parameter-Efficient Detoxification with Contrastive Decoding
von: Niu, Tong, et al.
Veröffentlicht: (2024)
von: Niu, Tong, et al.
Veröffentlicht: (2024)
Empirical Analysis of Decoding Biases in Masked Diffusion Models
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Huang, Pengcheng, et al.
Veröffentlicht: (2025)
Improving Generative Cross-lingual Aspect-Based Sentiment Analysis with Constrained Decoding
von: Šmíd, Jakub, et al.
Veröffentlicht: (2025)
von: Šmíd, Jakub, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Certified Robustness Under Bounded Levenshtein Distance
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025) -
A Comprehensive Approach to Misspelling Correction with BERT and Levenshtein Distance
von: Naziri, Amirreza, et al.
Veröffentlicht: (2024) -
Efficient Encoder-Decoder Transformer Decoding for Decomposable Tasks
von: Lu, Bo-Ru, et al.
Veröffentlicht: (2024) -
DecoderLens: Layerwise Interpretation of Encoder-Decoder Transformers
von: Langedijk, Anna, et al.
Veröffentlicht: (2023) -
Beyond Levenshtein: Leveraging Multiple Algorithms for Robust Word Error Rate Computations And Granular Error Classifications
von: Kuhn, Korbinian, et al.
Veröffentlicht: (2024)