GlitchMiner: Mining Glitch Tokens in Large Language Models via Gradient-based Discrete Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Zihui, Gao, Haichang, Wang, Ping, Zhang, Shudong, Liu, Zhaoxiang, Lian, Shiguo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GlitchProber: Advancing Effective Detection and Mitigation of Glitch Tokens in Large Language Models
von: Zhang, Zhibo, et al.
Veröffentlicht: (2024)
von: Zhang, Zhibo, et al.
Veröffentlicht: (2024)
The Dark Side of Function Calling: Pathways to Jailbreaking Large Language Models
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
TempGlitch: Evaluating Vision-Language Models for Temporal Glitch Detection in Gameplay Videos
von: Yu, Yakun, et al.
Veröffentlicht: (2026)
von: Yu, Yakun, et al.
Veröffentlicht: (2026)
Piculet: Specialized Models-Guided Hallucination Decrease for MultiModal Large Language Models
von: Wang, Kohou, et al.
Veröffentlicht: (2024)
von: Wang, Kohou, et al.
Veröffentlicht: (2024)
A Systematic Security Evaluation of OpenClaw and Its Variants
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
A Large Vision-Language Model based Environment Perception System for Visually Impaired People
von: Chen, Zezhou, et al.
Veröffentlicht: (2025)
von: Chen, Zezhou, et al.
Veröffentlicht: (2025)
A Glitch in the Matrix? Locating and Detecting Language Model Grounding with Fakepedia
von: Monea, Giovanni, et al.
Veröffentlicht: (2023)
von: Monea, Giovanni, et al.
Veröffentlicht: (2023)
From Assistant to Double Agent: Formalizing and Benchmarking Attacks on OpenClaw for Personalized Local AI Agent
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
Patch-wise Auto-Encoder for Visual Anomaly Detection
von: Cui, Yajie, et al.
Veröffentlicht: (2023)
von: Cui, Yajie, et al.
Veröffentlicht: (2023)
Semantic Glitch: Agency and Artistry in an Autonomous Pixel Cloud
von: Zhang, Qing, et al.
Veröffentlicht: (2025)
von: Zhang, Qing, et al.
Veröffentlicht: (2025)
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
What is the best model? Application-driven Evaluation for Large Language Models
von: Lian, Shiguo, et al.
Veröffentlicht: (2024)
von: Lian, Shiguo, et al.
Veröffentlicht: (2024)
CHiSafetyBench: A Chinese Hierarchical Safety Benchmark for Large Language Models
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
Using Deep Convolutional Neural Networks to Detect Rendered Glitches in Video Games
von: Ling, Carlos Garcia, et al.
Veröffentlicht: (2024)
von: Ling, Carlos Garcia, et al.
Veröffentlicht: (2024)
Methodology of Adapting Large English Language Models for Specific Cultural Contexts
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
Sparser Block-Sparse Attention via Token Permutation
von: Wang, Xinghao, et al.
Veröffentlicht: (2025)
von: Wang, Xinghao, et al.
Veröffentlicht: (2025)
GlitchBench: Can large multimodal models detect video game glitches?
von: Taesiri, Mohammad Reza, et al.
Veröffentlicht: (2023)
von: Taesiri, Mohammad Reza, et al.
Veröffentlicht: (2023)
Mixture of Heterogeneous Grouped Experts for Language Modeling
von: Ma, Zhicheng, et al.
Veröffentlicht: (2026)
von: Ma, Zhicheng, et al.
Veröffentlicht: (2026)
Hierarchical Deep Fusion Framework for Multi-dimensional Facial Forgery Detection -- The 2024 Global Deepfake Image Detection Challenge
von: Wang, Kohou, et al.
Veröffentlicht: (2025)
von: Wang, Kohou, et al.
Veröffentlicht: (2025)
HumorReject: Decoupling LLM Safety from Refusal Prefix via A Little Humor
von: Wu, Zihui, et al.
Veröffentlicht: (2025)
von: Wu, Zihui, et al.
Veröffentlicht: (2025)
LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation
von: Gao, Huanlin, et al.
Veröffentlicht: (2025)
von: Gao, Huanlin, et al.
Veröffentlicht: (2025)
Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
ICU-Bench:Benchmarking Continual Unlearning in Multimodal Large Language Models
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
LFTR: Learning-Free Token Reduction for Multimodal Large Language Models
von: Zhao, Zihui, et al.
Veröffentlicht: (2025)
von: Zhao, Zihui, et al.
Veröffentlicht: (2025)
A Multimodal Benchmark Dataset and Model for Crop Disease Diagnosis
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
Normal Distribution of Crab Pulsar Glitch Activity from a Glitch Cluster Perspective
von: Zhu, Pei-Xin, et al.
Veröffentlicht: (2025)
von: Zhu, Pei-Xin, et al.
Veröffentlicht: (2025)
DAST: Difficulty-Adaptive Slow-Thinking for Large Reasoning Models
von: Shen, Yi, et al.
Veröffentlicht: (2025)
von: Shen, Yi, et al.
Veröffentlicht: (2025)
Image Tokens Matter: Mitigating Hallucination in Discrete Tokenizer-based Large Vision-Language Models via Latent Editing
von: Wang, Weixing, et al.
Veröffentlicht: (2025)
von: Wang, Weixing, et al.
Veröffentlicht: (2025)
Entropy-Gated Selective Policy Optimization:Token-Level Gradient Allocation for Hybrid Training of Large Language Models
von: Hu, Yuelin, et al.
Veröffentlicht: (2026)
von: Hu, Yuelin, et al.
Veröffentlicht: (2026)
Fuzzy Reasoning Chain (FRC): An Innovative Reasoning Framework from Fuzziness to Clarity
von: Chen, Ping, et al.
Veröffentlicht: (2025)
von: Chen, Ping, et al.
Veröffentlicht: (2025)
MineAgent: Towards Remote-Sensing Mineral Exploration with Multimodal Large Language Models
von: Yu, Beibei, et al.
Veröffentlicht: (2024)
von: Yu, Beibei, et al.
Veröffentlicht: (2024)
HEAL: Hindsight Entropy-Assisted Learning for Reasoning Distillation
von: Zhang, Wenjing, et al.
Veröffentlicht: (2026)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2026)
VQ-Map: Bird's-Eye-View Map Layout Estimation in Tokenized Discrete Space via Vector Quantization
von: Zhang, Yiwei, et al.
Veröffentlicht: (2024)
von: Zhang, Yiwei, et al.
Veröffentlicht: (2024)
Exploring Superfluid Angular Momentum Reservoir Effect on Pulsar Glitches and Forecasting Next Glitches of the Crab Pulsar
von: Zhu, Pei-Xin, et al.
Veröffentlicht: (2026)
von: Zhu, Pei-Xin, et al.
Veröffentlicht: (2026)
Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
MITS: A Large-Scale Multimodal Benchmark Dataset for Intelligent Traffic Surveillance
von: Zhao, Kaikai, et al.
Veröffentlicht: (2025)
von: Zhao, Kaikai, et al.
Veröffentlicht: (2025)
Glitches and glitching clusters in rotation-powered pulsars
von: Zhu, Pei-Xin, et al.
Veröffentlicht: (2025)
von: Zhu, Pei-Xin, et al.
Veröffentlicht: (2025)
EA4LLM: A Gradient-Free Approach to Large Language Model Optimization via Evolutionary Algorithms
von: Liu, WenTao, et al.
Veröffentlicht: (2025)
von: Liu, WenTao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GlitchProber: Advancing Effective Detection and Mitigation of Glitch Tokens in Large Language Models
von: Zhang, Zhibo, et al.
Veröffentlicht: (2024) -
The Dark Side of Function Calling: Pathways to Jailbreaking Large Language Models
von: Wu, Zihui, et al.
Veröffentlicht: (2024) -
TempGlitch: Evaluating Vision-Language Models for Temporal Glitch Detection in Gameplay Videos
von: Yu, Yakun, et al.
Veröffentlicht: (2026) -
Piculet: Specialized Models-Guided Hallucination Decrease for MultiModal Large Language Models
von: Wang, Kohou, et al.
Veröffentlicht: (2024) -
A Systematic Security Evaluation of OpenClaw and Its Variants
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)