Gespeichert in:
| Hauptverfasser: | Yu, Yao-Ching, Kuo, Chun-Chih, Ye, Ziqi, Chang, Yu-Cheng, Li, Yueh-Se |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2406.12585 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Breaking the Capability Ceiling of LLM Post-Training by Reintroducing Markov States
von: Yuan, Yurun, et al.
Veröffentlicht: (2026)
von: Yuan, Yurun, et al.
Veröffentlicht: (2026)
Breaking the Ceiling: Exploring the Potential of Jailbreak Attacks through Expanding Strategy Space
von: Huang, Yao, et al.
Veröffentlicht: (2025)
von: Huang, Yao, et al.
Veröffentlicht: (2025)
MAD-OPD: Breaking the Ceiling in On-Policy Distillation via Multi-Agent Debate
von: Wang, Jianze, et al.
Veröffentlicht: (2026)
von: Wang, Jianze, et al.
Veröffentlicht: (2026)
When to Ensemble: Identifying Token-Level Points for Stable and Fast LLM Ensembling
von: Yun, Heecheol, et al.
Veröffentlicht: (2025)
von: Yun, Heecheol, et al.
Veröffentlicht: (2025)
Text2Freq: Learning Series Patterns from Text via Frequency Domain
von: Lo, Ming-Chih, et al.
Veröffentlicht: (2024)
von: Lo, Ming-Chih, et al.
Veröffentlicht: (2024)
On the Fallacy of Global Token Perplexity in Spoken Language Model Evaluation
von: Hsu, Chan-Jan, et al.
Veröffentlicht: (2026)
von: Hsu, Chan-Jan, et al.
Veröffentlicht: (2026)
Style-News: Incorporating Stylized News Generation and Adversarial Verification for Neural Fake News Detection
von: Wang, Wei-Yao, et al.
Veröffentlicht: (2024)
von: Wang, Wei-Yao, et al.
Veröffentlicht: (2024)
Next Token Perception Score: Analytical Assessment of your LLM Perception Skills
von: Cheng, Yu-Ang, et al.
Veröffentlicht: (2025)
von: Cheng, Yu-Ang, et al.
Veröffentlicht: (2025)
Benchmarking Cognitive Domains for LLMs: Insights from Taiwanese Hakka Culture
von: Chang, Chen-Chi, et al.
Veröffentlicht: (2024)
von: Chang, Chen-Chi, et al.
Veröffentlicht: (2024)
Reinforcement Learning with Token-level Feedback for Controllable Text Generation
von: Li, Wendi, et al.
Veröffentlicht: (2024)
von: Li, Wendi, et al.
Veröffentlicht: (2024)
Enhancing Robustness of LLM-Synthetic Text Detectors for Academic Writing: A Comprehensive Analysis
von: Dou, Zhicheng, et al.
Veröffentlicht: (2024)
von: Dou, Zhicheng, et al.
Veröffentlicht: (2024)
FLRC: Fine-grained Low-Rank Compressor for Efficient LLM Inference
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
PromptEmbedder:: Efficient and Transferable Text Embedding via Dual-LLM Soft Prompting
von: Tsai, Yu-Che, et al.
Veröffentlicht: (2026)
von: Tsai, Yu-Che, et al.
Veröffentlicht: (2026)
Breaking the Reviewer: Assessing the Vulnerability of Large Language Models in Automated Peer Review Under Textual Adversarial Attacks
von: Lin, Tzu-Ling, et al.
Veröffentlicht: (2025)
von: Lin, Tzu-Ling, et al.
Veröffentlicht: (2025)
Low-Resource Fine-Tuning for Multi-Task Structured Information Extraction with a Billion-Parameter Instruction-Tuned Model
von: Chih, Yu Cheng, et al.
Veröffentlicht: (2025)
von: Chih, Yu Cheng, et al.
Veröffentlicht: (2025)
Primus: A Pioneering Collection of Open-Source Datasets for Cybersecurity LLM Training
von: Yu, Yao-Ching, et al.
Veröffentlicht: (2025)
von: Yu, Yao-Ching, et al.
Veröffentlicht: (2025)
TokenButler: Token Importance is Predictable
von: Akhauri, Yash, et al.
Veröffentlicht: (2025)
von: Akhauri, Yash, et al.
Veröffentlicht: (2025)
Generate-on-Graph: Treat LLM as both Agent and KG in Incomplete Knowledge Graph Question Answering
von: Xu, Yao, et al.
Veröffentlicht: (2024)
von: Xu, Yao, et al.
Veröffentlicht: (2024)
TokenRec: Learning to Tokenize ID for LLM-based Generative Recommendation
von: Qu, Haohao, et al.
Veröffentlicht: (2024)
von: Qu, Haohao, et al.
Veröffentlicht: (2024)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
The Art of Breaking Words: Rethinking Multilingual Tokenizer Design
von: Thakur, Aamod, et al.
Veröffentlicht: (2025)
von: Thakur, Aamod, et al.
Veröffentlicht: (2025)
Harnessing Consistency for Robust Test-Time LLM Ensemble
von: Zeng, Zhichen, et al.
Veröffentlicht: (2025)
von: Zeng, Zhichen, et al.
Veröffentlicht: (2025)
VocalNet: Speech LLM with Multi-Token Prediction for Faster and High-Quality Generation
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
Generative Adversarial Reasoner: Enhancing LLM Reasoning with Adversarial Reinforcement Learning
von: Liu, Qihao, et al.
Veröffentlicht: (2025)
von: Liu, Qihao, et al.
Veröffentlicht: (2025)
A Simple Ensemble Strategy for LLM Inference: Towards More Stable Text Classification
von: Niimi, Junichiro
Veröffentlicht: (2025)
von: Niimi, Junichiro
Veröffentlicht: (2025)
Measuring Human Involvement in AI-Generated Text: A Case Study on Academic Writing
von: Guo, Yuchen, et al.
Veröffentlicht: (2025)
von: Guo, Yuchen, et al.
Veröffentlicht: (2025)
Explainability-Based Token Replacement on LLM-Generated Text
von: Mohammadi, Hadi, et al.
Veröffentlicht: (2025)
von: Mohammadi, Hadi, et al.
Veröffentlicht: (2025)
ToBlend: Token-Level Blending With an Ensemble of LLMs to Attack AI-Generated Text Detection
von: Huang, Fan, et al.
Veröffentlicht: (2024)
von: Huang, Fan, et al.
Veröffentlicht: (2024)
Generation Space Size: Understanding and Calibrating Open-Endedness of LLM Generations
von: Yu, Sunny, et al.
Veröffentlicht: (2025)
von: Yu, Sunny, et al.
Veröffentlicht: (2025)
S$^2$-MAD: Breaking the Token Barrier to Enhance Multi-Agent Debate Efficiency
von: Zeng, Yuting, et al.
Veröffentlicht: (2025)
von: Zeng, Yuting, et al.
Veröffentlicht: (2025)
Does Using Counterfactual Help LLMs Explain Textual Importance in Classification?
von: Tan, Nelvin, et al.
Veröffentlicht: (2025)
von: Tan, Nelvin, et al.
Veröffentlicht: (2025)
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM
von: Yu, Chengyue, et al.
Veröffentlicht: (2024)
von: Yu, Chengyue, et al.
Veröffentlicht: (2024)
Next-Token Prediction Task Assumes Optimal Data Ordering for LLM Training in Proof Generation
von: An, Chenyang, et al.
Veröffentlicht: (2024)
von: An, Chenyang, et al.
Veröffentlicht: (2024)
Multi-Programming Language Ensemble for Code Generation in Large Language Model
von: Xue, Tengfei, et al.
Veröffentlicht: (2024)
von: Xue, Tengfei, et al.
Veröffentlicht: (2024)
Enhancing Annotated Bibliography Generation with LLM Ensembles
von: Bermejo, Sergio
Veröffentlicht: (2024)
von: Bermejo, Sergio
Veröffentlicht: (2024)
SemToken: Semantic-Aware Tokenization for Efficient Long-Context Language Modeling
von: Liu, Dong, et al.
Veröffentlicht: (2025)
von: Liu, Dong, et al.
Veröffentlicht: (2025)
An Exploration of Mamba for Speech Self-Supervised Models
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2025)
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2025)
KunlunBaize: LLM with Multi-Scale Convolution and Multi-Token Prediction Under TransformerX Framework
von: Li, Cheng, et al.
Veröffentlicht: (2025)
von: Li, Cheng, et al.
Veröffentlicht: (2025)
Selecting Between BERT and GPT for Text Classification in Political Science Research
von: Wang, Yu, et al.
Veröffentlicht: (2024)
von: Wang, Yu, et al.
Veröffentlicht: (2024)
Dynamic Generation of Multi-LLM Agents Communication Topologies with Graph Diffusion Models
von: Jiang, Eric Hanchen, et al.
Veröffentlicht: (2025)
von: Jiang, Eric Hanchen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Breaking the Capability Ceiling of LLM Post-Training by Reintroducing Markov States
von: Yuan, Yurun, et al.
Veröffentlicht: (2026) -
Breaking the Ceiling: Exploring the Potential of Jailbreak Attacks through Expanding Strategy Space
von: Huang, Yao, et al.
Veröffentlicht: (2025) -
MAD-OPD: Breaking the Ceiling in On-Policy Distillation via Multi-Agent Debate
von: Wang, Jianze, et al.
Veröffentlicht: (2026) -
When to Ensemble: Identifying Token-Level Points for Stable and Fast LLM Ensembling
von: Yun, Heecheol, et al.
Veröffentlicht: (2025) -
Text2Freq: Learning Series Patterns from Text via Frequency Domain
von: Lo, Ming-Chih, et al.
Veröffentlicht: (2024)