Adversarial Tokenization
Fuente:
arXiv
Salvato in:
| Autori principali: | Geh, Renato Lui, Shao, Zilei, Broeck, Guy Van den |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enabling Autoregressive Models to Fill In Masked Tokens
di: Israel, Daniel, et al.
Pubblicazione: (2025)
di: Israel, Daniel, et al.
Pubblicazione: (2025)
Where is the signal in tokenization space?
di: Geh, Renato Lui, et al.
Pubblicazione: (2024)
di: Geh, Renato Lui, et al.
Pubblicazione: (2024)
Probabilistic Programs of Thought
di: Garg, Poorva, et al.
Pubblicazione: (2026)
di: Garg, Poorva, et al.
Pubblicazione: (2026)
The Pitfalls of KV Cache Compression
di: Chen, Alex, et al.
Pubblicazione: (2025)
di: Chen, Alex, et al.
Pubblicazione: (2025)
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
di: Zhao, Heng, et al.
Pubblicazione: (2026)
di: Zhao, Heng, et al.
Pubblicazione: (2026)
A Pseudo-Semantic Loss for Autoregressive Models with Logical Constraints
di: Ahmed, Kareem, et al.
Pubblicazione: (2023)
di: Ahmed, Kareem, et al.
Pubblicazione: (2023)
Accelerating Diffusion LLMs via Adaptive Parallel Decoding
di: Israel, Daniel, et al.
Pubblicazione: (2025)
di: Israel, Daniel, et al.
Pubblicazione: (2025)
Prepacking: A Simple Method for Fast Prefilling and Increased Throughput in Large Language Models
di: Zhao, Siyan, et al.
Pubblicazione: (2024)
di: Zhao, Siyan, et al.
Pubblicazione: (2024)
Breaking the Factorization Barrier in Diffusion Language Models
di: Li, Ian, et al.
Pubblicazione: (2026)
di: Li, Ian, et al.
Pubblicazione: (2026)
Collapsed Inference for Bayesian Deep Learning
di: Zeng, Zhe, et al.
Pubblicazione: (2023)
di: Zeng, Zhe, et al.
Pubblicazione: (2023)
On the Relationship Between Monotone and Squared Probabilistic Circuits
di: Wang, Benjie, et al.
Pubblicazione: (2024)
di: Wang, Benjie, et al.
Pubblicazione: (2024)
Zero-Variance Gradients for Variational Autoencoders
di: Shao, Zilei, et al.
Pubblicazione: (2025)
di: Shao, Zilei, et al.
Pubblicazione: (2025)
Rethinking Probabilistic Circuit Parameter Learning
di: Liu, Anji, et al.
Pubblicazione: (2025)
di: Liu, Anji, et al.
Pubblicazione: (2025)
How to Marginalize in Causal Structure Learning?
di: Zhao, William, et al.
Pubblicazione: (2025)
di: Zhao, William, et al.
Pubblicazione: (2025)
Scaling Up Probabilistic Circuits by Latent Variable Distillation
di: Liu, Anji, et al.
Pubblicazione: (2022)
di: Liu, Anji, et al.
Pubblicazione: (2022)
LoRA-Pro: Are Low-Rank Adapters Properly Optimized?
di: Wang, Zhengbo, et al.
Pubblicazione: (2024)
di: Wang, Zhengbo, et al.
Pubblicazione: (2024)
Taming Momentum: Rethinking Optimizer States Through Low-Rank Approximation
di: Wang, Zhengbo, et al.
Pubblicazione: (2026)
di: Wang, Zhengbo, et al.
Pubblicazione: (2026)
Token-Efficient Leverage Learning in Large Language Models
di: Zeng, Yuanhao, et al.
Pubblicazione: (2024)
di: Zeng, Yuanhao, et al.
Pubblicazione: (2024)
Beyond Next Token Prediction: Patch-Level Training for Large Language Models
di: Shao, Chenze, et al.
Pubblicazione: (2024)
di: Shao, Chenze, et al.
Pubblicazione: (2024)
TRACE Back from the Future: A Probabilistic Reasoning Approach to Controllable Language Generation
di: Weng, Gwen Yidou, et al.
Pubblicazione: (2025)
di: Weng, Gwen Yidou, et al.
Pubblicazione: (2025)
Controllable Generation via Locally Constrained Resampling
di: Ahmed, Kareem, et al.
Pubblicazione: (2024)
di: Ahmed, Kareem, et al.
Pubblicazione: (2024)
A Tractable Inference Perspective of Offline RL
di: Liu, Xuejie, et al.
Pubblicazione: (2023)
di: Liu, Xuejie, et al.
Pubblicazione: (2023)
Restructuring Tractable Probabilistic Circuits
di: Zhang, Honghua, et al.
Pubblicazione: (2024)
di: Zhang, Honghua, et al.
Pubblicazione: (2024)
SIMPLE: A Gradient Estimator for $k$-Subset Sampling
di: Ahmed, Kareem, et al.
Pubblicazione: (2022)
di: Ahmed, Kareem, et al.
Pubblicazione: (2022)
A Compositional Atlas for Algebraic Circuits
di: Wang, Benjie, et al.
Pubblicazione: (2024)
di: Wang, Benjie, et al.
Pubblicazione: (2024)
Adversarial Attacks on AI-Generated Text Detection Models: A Token Probability-Based Approach Using Embeddings
di: Kadhim, Ahmed K., et al.
Pubblicazione: (2025)
di: Kadhim, Ahmed K., et al.
Pubblicazione: (2025)
Luna-2: Scalable Single-Token Evaluation with Small Language Models
di: Goel, Vatsal, et al.
Pubblicazione: (2026)
di: Goel, Vatsal, et al.
Pubblicazione: (2026)
TokenButler: Token Importance is Predictable
di: Akhauri, Yash, et al.
Pubblicazione: (2025)
di: Akhauri, Yash, et al.
Pubblicazione: (2025)
Lossless Token Sequence Compression via Meta-Tokens
di: Harvill, John, et al.
Pubblicazione: (2025)
di: Harvill, John, et al.
Pubblicazione: (2025)
mini-vec2vec: Scaling Universal Geometry Alignment with Linear Transformations
di: Dar, Guy
Pubblicazione: (2025)
di: Dar, Guy
Pubblicazione: (2025)
Tokenization Matters: Navigating Data-Scarce Tokenization for Gender Inclusive Language Technologies
di: Ovalle, Anaelia, et al.
Pubblicazione: (2023)
di: Ovalle, Anaelia, et al.
Pubblicazione: (2023)
Probabilistic Circuits for Cumulative Distribution Functions
di: Broadrick, Oliver, et al.
Pubblicazione: (2024)
di: Broadrick, Oliver, et al.
Pubblicazione: (2024)
OrthoRank: Token Selection via Sink Token Orthogonality for Efficient LLM inference
di: Shin, Seungjun, et al.
Pubblicazione: (2025)
di: Shin, Seungjun, et al.
Pubblicazione: (2025)
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
di: Lin, Zicheng, et al.
Pubblicazione: (2024)
di: Lin, Zicheng, et al.
Pubblicazione: (2024)
Multimodal Medical Code Tokenizer
di: Su, Xiaorui, et al.
Pubblicazione: (2025)
di: Su, Xiaorui, et al.
Pubblicazione: (2025)
Soft Tokens, Hard Truths
di: Butt, Natasha, et al.
Pubblicazione: (2025)
di: Butt, Natasha, et al.
Pubblicazione: (2025)
Learning to Reason with Mixture of Tokens
di: Jain, Adit, et al.
Pubblicazione: (2025)
di: Jain, Adit, et al.
Pubblicazione: (2025)
Cautious Next Token Prediction
di: Wang, Yizhou, et al.
Pubblicazione: (2025)
di: Wang, Yizhou, et al.
Pubblicazione: (2025)
Revisiting Graph-Tokenizing Large Language Models: A Systematic Evaluation of Graph Token Understanding
di: Zhang, Zhongjian, et al.
Pubblicazione: (2026)
di: Zhang, Zhongjian, et al.
Pubblicazione: (2026)
Generative Adversarial Reasoner: Enhancing LLM Reasoning with Adversarial Reinforcement Learning
di: Liu, Qihao, et al.
Pubblicazione: (2025)
di: Liu, Qihao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Enabling Autoregressive Models to Fill In Masked Tokens
di: Israel, Daniel, et al.
Pubblicazione: (2025) -
Where is the signal in tokenization space?
di: Geh, Renato Lui, et al.
Pubblicazione: (2024) -
Probabilistic Programs of Thought
di: Garg, Poorva, et al.
Pubblicazione: (2026) -
The Pitfalls of KV Cache Compression
di: Chen, Alex, et al.
Pubblicazione: (2025) -
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
di: Zhao, Heng, et al.
Pubblicazione: (2026)