Non-Halting Queries: Exploiting Fixed Points in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hammouri, Ghaith, Derya, Kemal, Sunar, Berk |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compression in LLMs
von: Okutomi, Akira
Veröffentlicht: (2025)
von: Okutomi, Akira
Veröffentlicht: (2025)
QueryBandits for Hallucination Mitigation: Exploiting Semantic Features for No-Regret Rewriting
von: Cho, Nicole, et al.
Veröffentlicht: (2025)
von: Cho, Nicole, et al.
Veröffentlicht: (2025)
ThermoQA: A Three-Tier Benchmark for Evaluating Thermodynamic Reasoning in Large Language Models
von: Düzkar, Kemal
Veröffentlicht: (2026)
von: Düzkar, Kemal
Veröffentlicht: (2026)
μRL: Discovering Transient Execution Vulnerabilities Using Reinforcement Learning
von: Tol, M. Caner, et al.
Veröffentlicht: (2025)
von: Tol, M. Caner, et al.
Veröffentlicht: (2025)
Exploiting Synergistic Cognitive Biases to Bypass Safety in LLMs
von: Yang, Xikang, et al.
Veröffentlicht: (2025)
von: Yang, Xikang, et al.
Veröffentlicht: (2025)
Super Suffixes: Bypassing Text Generation Alignment and Guard Models Simultaneously
von: Adiletta, Andrew, et al.
Veröffentlicht: (2025)
von: Adiletta, Andrew, et al.
Veröffentlicht: (2025)
Can We Count on LLMs? The Fixed-Effect Fallacy and Claims of GPT-4 Capabilities
von: Ball, Thomas, et al.
Veröffentlicht: (2024)
von: Ball, Thomas, et al.
Veröffentlicht: (2024)
Aligning Tree-Search Policies with Fixed Token Budgets in Test-Time Scaling of LLMs
von: Miyamoto, Sora, et al.
Veröffentlicht: (2026)
von: Miyamoto, Sora, et al.
Veröffentlicht: (2026)
Prompt Repetition Improves Non-Reasoning LLMs
von: Leviathan, Yaniv, et al.
Veröffentlicht: (2025)
von: Leviathan, Yaniv, et al.
Veröffentlicht: (2025)
VerAs: Verify then Assess STEM Lab Reports
von: Atil, Berk, et al.
Veröffentlicht: (2024)
von: Atil, Berk, et al.
Veröffentlicht: (2024)
Prompting with Phonemes: Enhancing LLMs' Multilinguality for Non-Latin Script Languages
von: Nguyen, Hoang H, et al.
Veröffentlicht: (2024)
von: Nguyen, Hoang H, et al.
Veröffentlicht: (2024)
Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
Should You Use Your Large Language Model to Explore or Exploit?
von: Harris, Keegan, et al.
Veröffentlicht: (2025)
von: Harris, Keegan, et al.
Veröffentlicht: (2025)
Balanced Aggregation: Understanding and Fixing Aggregation Bias in GRPO
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2026)
Dual Encoder: Exploiting the Potential of Syntactic and Semantic for Aspect Sentiment Triplet Extraction
von: Zhao, Xiaowei, et al.
Veröffentlicht: (2024)
von: Zhao, Xiaowei, et al.
Veröffentlicht: (2024)
Revisiting JBShield: Breaking and Rebuilding Representation-Level Jailbreak Defenses
von: Derya, Kemal, et al.
Veröffentlicht: (2026)
von: Derya, Kemal, et al.
Veröffentlicht: (2026)
LLMs versus the Halting Problem: Characterizing Program Termination Reasoning
von: Sultan, Oren, et al.
Veröffentlicht: (2026)
von: Sultan, Oren, et al.
Veröffentlicht: (2026)
Smaug: Fixing Failure Modes of Preference Optimisation with DPO-Positive
von: Pal, Arka, et al.
Veröffentlicht: (2024)
von: Pal, Arka, et al.
Veröffentlicht: (2024)
Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes
von: Fu, Yuqian, et al.
Veröffentlicht: (2026)
von: Fu, Yuqian, et al.
Veröffentlicht: (2026)
ConfPO: Exploiting Policy Model Confidence for Critical Token Selection in Preference Optimization
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2025)
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2025)
Query-Guided Self-Supervised Summarization of Nursing Notes
von: Gao, Ya, et al.
Veröffentlicht: (2024)
von: Gao, Ya, et al.
Veröffentlicht: (2024)
Structured Query Construction via Knowledge Graph Embedding
von: Wang, Ruijie, et al.
Veröffentlicht: (2019)
von: Wang, Ruijie, et al.
Veröffentlicht: (2019)
Hypothesis-Conditioned Query Rewriting for Decision-Useful Retrieval
von: Chang, Hangeol, et al.
Veröffentlicht: (2026)
von: Chang, Hangeol, et al.
Veröffentlicht: (2026)
MoMQ: Mixture-of-Experts Enhances Multi-Dialect Query Generation across Relational and Non-Relational Databases
von: Lin, Zhisheng, et al.
Veröffentlicht: (2024)
von: Lin, Zhisheng, et al.
Veröffentlicht: (2024)
Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2024)
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2024)
SAC-KG: Exploiting Large Language Models as Skilled Automatic Constructors for Domain Knowledge Graphs
von: Chen, Hanzhu, et al.
Veröffentlicht: (2024)
von: Chen, Hanzhu, et al.
Veröffentlicht: (2024)
Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases
von: Hahm, Dongyoon, et al.
Veröffentlicht: (2026)
von: Hahm, Dongyoon, et al.
Veröffentlicht: (2026)
Exploiting Transliterated Words for Finding Similarity in Inter-Language News Articles using Machine Learning
von: Naeem, Sameea, et al.
Veröffentlicht: (2022)
von: Naeem, Sameea, et al.
Veröffentlicht: (2022)
AQuA -- Combining Experts' and Non-Experts' Views To Assess Deliberation Quality in Online Discussions Using LLMs
von: Behrendt, Maike, et al.
Veröffentlicht: (2024)
von: Behrendt, Maike, et al.
Veröffentlicht: (2024)
Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing
von: Ding, Dujian, et al.
Veröffentlicht: (2024)
von: Ding, Dujian, et al.
Veröffentlicht: (2024)
No One Size Fits All: QueryBandits for Hallucination Mitigation
von: Cho, Nicole, et al.
Veröffentlicht: (2026)
von: Cho, Nicole, et al.
Veröffentlicht: (2026)
Query-Dependent Prompt Evaluation and Optimization with Offline Inverse RL
von: Sun, Hao, et al.
Veröffentlicht: (2023)
von: Sun, Hao, et al.
Veröffentlicht: (2023)
Cost-Optimal Grouped-Query Attention for Long-Context Modeling
von: Chen, Yingfa, et al.
Veröffentlicht: (2025)
von: Chen, Yingfa, et al.
Veröffentlicht: (2025)
PDC & DM-SFT: A Road for LLM SQL Bug-Fix Enhancing
von: Duan, Yiwen, et al.
Veröffentlicht: (2024)
von: Duan, Yiwen, et al.
Veröffentlicht: (2024)
Query-Conditioned Test-Time Self-Training for Large Language Models
von: Song, Chaehee, et al.
Veröffentlicht: (2026)
von: Song, Chaehee, et al.
Veröffentlicht: (2026)
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
von: Li, Chengpeng, et al.
Veröffentlicht: (2023)
von: Li, Chengpeng, et al.
Veröffentlicht: (2023)
Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory
von: Zhang, Haozhen, et al.
Veröffentlicht: (2026)
von: Zhang, Haozhen, et al.
Veröffentlicht: (2026)
Diversity Measures: Domain-Independent Proxies for Failure in Language Model Queries
von: Ngu, Noel, et al.
Veröffentlicht: (2023)
von: Ngu, Noel, et al.
Veröffentlicht: (2023)
No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping
von: Le, Thanh-Long V., et al.
Veröffentlicht: (2025)
von: Le, Thanh-Long V., et al.
Veröffentlicht: (2025)
Investigating Recurrent Transformers with Dynamic Halt
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2024)
von: Chowdhury, Jishnu Ray, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compression in LLMs
von: Okutomi, Akira
Veröffentlicht: (2025) -
QueryBandits for Hallucination Mitigation: Exploiting Semantic Features for No-Regret Rewriting
von: Cho, Nicole, et al.
Veröffentlicht: (2025) -
ThermoQA: A Three-Tier Benchmark for Evaluating Thermodynamic Reasoning in Large Language Models
von: Düzkar, Kemal
Veröffentlicht: (2026) -
μRL: Discovering Transient Execution Vulnerabilities Using Reinforcement Learning
von: Tol, M. Caner, et al.
Veröffentlicht: (2025) -
Exploiting Synergistic Cognitive Biases to Bypass Safety in LLMs
von: Yang, Xikang, et al.
Veröffentlicht: (2025)