Cats Confuse Reasoning LLM: Query Agnostic Adversarial Triggers for Reasoning Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Rajeev, Meghana, Ramamurthy, Rajkumar, Trivedi, Prapti, Yadav, Vikas, Bamgbose, Oluwanifemi, Madhusudan, Sathwik Tejaswi, Zou, James, Rajani, Nazneen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DNR Bench: Benchmarking Over-Reasoning in Reasoning LLMs
di: Hashemi, Masoud, et al.
Pubblicazione: (2025)
di: Hashemi, Masoud, et al.
Pubblicazione: (2025)
VERITAS: A Unified Approach to Reliability Evaluation
di: Ramamurthy, Rajkumar, et al.
Pubblicazione: (2024)
di: Ramamurthy, Rajkumar, et al.
Pubblicazione: (2024)
Self-rationalization improves LLM as a fine-grained judge
di: Trivedi, Prapti, et al.
Pubblicazione: (2024)
di: Trivedi, Prapti, et al.
Pubblicazione: (2024)
Impatient Users Confuse AI Agents: High-fidelity Simulations of Human Traits for Testing Agents
di: He, Muyu, et al.
Pubblicazione: (2025)
di: He, Muyu, et al.
Pubblicazione: (2025)
Augmenting LLM Reasoning with Dynamic Notes Writing for Complex QA
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2025)
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2025)
AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs
di: Nguyen, Hoang, et al.
Pubblicazione: (2025)
di: Nguyen, Hoang, et al.
Pubblicazione: (2025)
Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models
di: Madhusudhan, Nishanth, et al.
Pubblicazione: (2024)
di: Madhusudhan, Nishanth, et al.
Pubblicazione: (2024)
Auto-Cypher: Improving LLMs on Cypher generation via LLM-supervised generation-verification framework
di: Tiwari, Aman, et al.
Pubblicazione: (2024)
di: Tiwari, Aman, et al.
Pubblicazione: (2024)
Curry-DPO: Enhancing Alignment using Curriculum Learning & Ranked Preferences
di: Pattnaik, Pulkit, et al.
Pubblicazione: (2024)
di: Pattnaik, Pulkit, et al.
Pubblicazione: (2024)
M2Lingual: Enhancing Multilingual, Multi-Turn Instruction Alignment in Large Language Models
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2024)
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2024)
The Valley of Code Reasoning: Scaling Knowledge Distillation of Large Language Models
di: He, Muyu, et al.
Pubblicazione: (2025)
di: He, Muyu, et al.
Pubblicazione: (2025)
Practical Guide for Causal Pathways and Sub-group Disparity Analysis
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
DeepSRGM -- Sequence Classification and Ranking in Indian Classical Music with Deep Learning
di: Madhusudhan, Sathwik Tejaswi, et al.
Pubblicazione: (2024)
di: Madhusudhan, Sathwik Tejaswi, et al.
Pubblicazione: (2024)
Evaluating Robustness of Large Language Models in Enterprise Applications: Benchmarks for Perturbation Consistency Across Formats and Languages
di: Bogavelli, Tara, et al.
Pubblicazione: (2026)
di: Bogavelli, Tara, et al.
Pubblicazione: (2026)
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2024)
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2024)
Revitalizing Saturated Benchmarks: A Weighted Metric Approach for Differentiating Large Language Model Performance
di: Etzine, Bryan, et al.
Pubblicazione: (2025)
di: Etzine, Bryan, et al.
Pubblicazione: (2025)
Grammar Search for Multi-Agent Systems
di: Singh, Mayank, et al.
Pubblicazione: (2025)
di: Singh, Mayank, et al.
Pubblicazione: (2025)
FakeWatch: A Framework for Detecting Fake News to Ensure Credible Elections
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis
di: Raza, Shaina, et al.
Pubblicazione: (2023)
di: Raza, Shaina, et al.
Pubblicazione: (2023)
A Toolbox, Not a Hammer -- Multi-TAG: Scaling Math Reasoning with Multi-Tool Aggregation
di: Yao, Bohan, et al.
Pubblicazione: (2025)
di: Yao, Bohan, et al.
Pubblicazione: (2025)
Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems
di: Madhusudhan, Nishanth, et al.
Pubblicazione: (2026)
di: Madhusudhan, Nishanth, et al.
Pubblicazione: (2026)
Developing Safe and Responsible Large Language Model : Can We Balance Bias Reduction and Language Understanding in Large Language Models?
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
ReasonOps: Operator Segmentation for LLM Reasoning Traces
di: Lee, Daniel, et al.
Pubblicazione: (2026)
di: Lee, Daniel, et al.
Pubblicazione: (2026)
BigCharts-R1: Enhanced Chart Reasoning with Visual Reinforcement Finetuning
di: Masry, Ahmed, et al.
Pubblicazione: (2025)
di: Masry, Ahmed, et al.
Pubblicazione: (2025)
Stage-wise Fine-tuning for Graph-to-Text Generation
di: Wang, Qingyun, et al.
Pubblicazione: (2021)
di: Wang, Qingyun, et al.
Pubblicazione: (2021)
Pragmatic and Discourse Functions in Jenifa’s Diary
di: Ganiu Bamgbose
Pubblicazione: (2021)
di: Ganiu Bamgbose
Pubblicazione: (2021)
Generative Adversarial Reasoner: Enhancing LLM Reasoning with Adversarial Reinforcement Learning
di: Liu, Qihao, et al.
Pubblicazione: (2025)
di: Liu, Qihao, et al.
Pubblicazione: (2025)
Modal Logic for Reasoning About Uncertainty and Confusion
di: Bílková, Marta, et al.
Pubblicazione: (2025)
di: Bílková, Marta, et al.
Pubblicazione: (2025)
Quantifying Cross-Query Contradictions in Multi-Query LLM Reasoning
di: Salla, Rohit Kumar, et al.
Pubblicazione: (2026)
di: Salla, Rohit Kumar, et al.
Pubblicazione: (2026)
ARM: Discovering Agentic Reasoning Modules for Generalizable Multi-Agent Systems
di: Yao, Bohan, et al.
Pubblicazione: (2025)
di: Yao, Bohan, et al.
Pubblicazione: (2025)
What's documented in AI? Systematic Analysis of 32K AI Model Cards
di: Liang, Weixin, et al.
Pubblicazione: (2024)
di: Liang, Weixin, et al.
Pubblicazione: (2024)
iReasoner: Trajectory-Aware Intrinsic Reasoning Supervision for Self-Evolving Large Multimodal Models
di: Sunil, Meghana, et al.
Pubblicazione: (2026)
di: Sunil, Meghana, et al.
Pubblicazione: (2026)
Reverse
di: Rahman, Prapti
Pubblicazione: (2024)
di: Rahman, Prapti
Pubblicazione: (2024)
Sago‐grain nodules in tubercular pleural effusion
di: Manoj Madhusudan, et al.
Pubblicazione: (2024)
di: Manoj Madhusudan, et al.
Pubblicazione: (2024)
ConciseRL: Conciseness-Guided Reinforcement Learning for Efficient Reasoning Models
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2025)
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2025)
Unlocking LLM Creativity in Science through Analogical Reasoning
di: Shen, Andrew, et al.
Pubblicazione: (2026)
di: Shen, Andrew, et al.
Pubblicazione: (2026)
Forecasting LLM Inference Performance via Hardware-Agnostic Analytical Modeling
di: Patwari, Rajeev, et al.
Pubblicazione: (2025)
di: Patwari, Rajeev, et al.
Pubblicazione: (2025)
The Dog the Cat Chased Stumped the Model: Measuring When Language Models Abandon Structure for Shortcuts
di: Madhusudan, Sangmitra, et al.
Pubblicazione: (2025)
di: Madhusudan, Sangmitra, et al.
Pubblicazione: (2025)
QTrack: Query-Driven Reasoning for Multi-modal MOT
di: Ashraf, Tajamul, et al.
Pubblicazione: (2026)
di: Ashraf, Tajamul, et al.
Pubblicazione: (2026)
Apriel-1.5-15b-Thinker
di: Radhakrishna, Shruthan, et al.
Pubblicazione: (2025)
di: Radhakrishna, Shruthan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
DNR Bench: Benchmarking Over-Reasoning in Reasoning LLMs
di: Hashemi, Masoud, et al.
Pubblicazione: (2025) -
VERITAS: A Unified Approach to Reliability Evaluation
di: Ramamurthy, Rajkumar, et al.
Pubblicazione: (2024) -
Self-rationalization improves LLM as a fine-grained judge
di: Trivedi, Prapti, et al.
Pubblicazione: (2024) -
Impatient Users Confuse AI Agents: High-fidelity Simulations of Human Traits for Testing Agents
di: He, Muyu, et al.
Pubblicazione: (2025) -
Augmenting LLM Reasoning with Dynamic Notes Writing for Complex QA
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2025)