ODAR: Principled Adaptive Routing for LLM Reasoning via Active Inference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Siyuan, Gao, Bo, Jia, Xiaojun, Qin, Simeng, Li, Tianlin, Ma, Ke, Jia, Xiaoshuang, Ren, Wenqi, Liu, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reasoning as an Attack Surface: Adaptive Evolutionary CoT Jailbreaks for LLMs
von: Li, Jianan, et al.
Veröffentlicht: (2026)
von: Li, Jianan, et al.
Veröffentlicht: (2026)
Inverse Reinforcement Learning with Dynamic Reward Scaling for LLM Alignment
von: Cheng, Ruoxi, et al.
Veröffentlicht: (2025)
von: Cheng, Ruoxi, et al.
Veröffentlicht: (2025)
Obscure but Effective: Classical Chinese Jailbreak Prompt Optimization via Bio-Inspired Search
von: Huang, Xun, et al.
Veröffentlicht: (2026)
von: Huang, Xun, et al.
Veröffentlicht: (2026)
Heuristic-Induced Multimodal Risk Distribution Jailbreak Attack for Multimodal Large Language Models
von: Teng, Ma, et al.
Veröffentlicht: (2024)
von: Teng, Ma, et al.
Veröffentlicht: (2024)
SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents
von: Jia, Xiaojun, et al.
Veröffentlicht: (2026)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2026)
GeoShield: Safeguarding Geolocation Privacy from Vision-Language Models via Adversarial Perturbations
von: Liu, Xinwei, et al.
Veröffentlicht: (2025)
von: Liu, Xinwei, et al.
Veröffentlicht: (2025)
PBI-Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization
von: Cheng, Ruoxi, et al.
Veröffentlicht: (2024)
von: Cheng, Ruoxi, et al.
Veröffentlicht: (2024)
MMR-Bench: A Comprehensive Benchmark for Multimodal LLM Routing
von: Ma, Haoxuan, et al.
Veröffentlicht: (2026)
von: Ma, Haoxuan, et al.
Veröffentlicht: (2026)
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
von: Li, Zheng, et al.
Veröffentlicht: (2025)
von: Li, Zheng, et al.
Veröffentlicht: (2025)
Cannot See the Forest for the Trees: Invoking Heuristics and Biases to Elicit Irrational Choices of LLMs
von: Yang, Haoming, et al.
Veröffentlicht: (2025)
von: Yang, Haoming, et al.
Veröffentlicht: (2025)
Boosting LLM Reasoning via Human-Inspired Reward Shaping
von: Lin, Wenze, et al.
Veröffentlicht: (2026)
von: Lin, Wenze, et al.
Veröffentlicht: (2026)
When Routing Collapses: On the Degenerate Convergence of LLM Routers
von: Lai, Guannan, et al.
Veröffentlicht: (2026)
von: Lai, Guannan, et al.
Veröffentlicht: (2026)
Martingale Foresight Sampling: A Principled Approach to Inference-Time LLM Decoding
von: Li, Huayu, et al.
Veröffentlicht: (2026)
von: Li, Huayu, et al.
Veröffentlicht: (2026)
Robust Multi-bit Text Watermark with LLM-based Paraphrasers
von: Xu, Xiaojun, et al.
Veröffentlicht: (2024)
von: Xu, Xiaojun, et al.
Veröffentlicht: (2024)
SeCon-RAG: A Two-Stage Semantic Filtering and Conflict-Free Framework for Trustworthy RAG
von: Si, Xiaonan, et al.
Veröffentlicht: (2025)
von: Si, Xiaonan, et al.
Veröffentlicht: (2025)
Evolution-based Region Adversarial Prompt Learning for Robustness Enhancement in Vision-Language Models
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
Boosting LLM via Learning from Data Iteratively and Selectively
von: Jia, Qi, et al.
Veröffentlicht: (2024)
von: Jia, Qi, et al.
Veröffentlicht: (2024)
OmniSafeBench-MM: A Unified Benchmark and Toolbox for Multimodal Jailbreak Attack-Defense Evaluation
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
UTMath: Math Evaluation with Unit Test via Reasoning-to-Coding Thoughts
von: Yang, Bo, et al.
Veröffentlicht: (2024)
von: Yang, Bo, et al.
Veröffentlicht: (2024)
Revisiting LLM Reasoning via Information Bottleneck
von: Lei, Shiye, et al.
Veröffentlicht: (2025)
von: Lei, Shiye, et al.
Veröffentlicht: (2025)
Large Language Models for Predictive Analysis: How Far Are They?
von: Chen, Qin, et al.
Veröffentlicht: (2025)
von: Chen, Qin, et al.
Veröffentlicht: (2025)
Meta-Reasoner: Dynamic Guidance for Optimized Inference-time Reasoning in Large Language Models
von: Sui, Yuan, et al.
Veröffentlicht: (2025)
von: Sui, Yuan, et al.
Veröffentlicht: (2025)
The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans?
von: Xun, Yuan, et al.
Veröffentlicht: (2025)
von: Xun, Yuan, et al.
Veröffentlicht: (2025)
MASteer: Multi-Agent Adaptive Steer Strategy for End-to-End LLM Trustworthiness Repair
von: Li, Changqing, et al.
Veröffentlicht: (2025)
von: Li, Changqing, et al.
Veröffentlicht: (2025)
CoT2-Meta: Budgeted Metacognitive Control for Test-Time Reasoning
von: Ma, Siyuan, et al.
Veröffentlicht: (2026)
von: Ma, Siyuan, et al.
Veröffentlicht: (2026)
Semantic-Aligned Adversarial Evolution Triangle for High-Transferability Vision-Language Attack
von: Jia, Xiaojun, et al.
Veröffentlicht: (2024)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2024)
Adaptive Stopping for Multi-Turn LLM Reasoning
von: Zhou, Xiaofan, et al.
Veröffentlicht: (2026)
von: Zhou, Xiaofan, et al.
Veröffentlicht: (2026)
Multi-modal Situated Reasoning in 3D Scenes
von: Linghu, Xiongkun, et al.
Veröffentlicht: (2024)
von: Linghu, Xiongkun, et al.
Veröffentlicht: (2024)
Environmental Matching Attack Against Unmanned Aerial Vehicles Object Detection
von: Kong, Dehong, et al.
Veröffentlicht: (2024)
von: Kong, Dehong, et al.
Veröffentlicht: (2024)
GAR: Carbon-Aware Routing for LLM Inference via Constrained Optimization
von: Sheshanarayana, Disha, et al.
Veröffentlicht: (2026)
von: Sheshanarayana, Disha, et al.
Veröffentlicht: (2026)
LLM Routing as Reasoning: A MaxSAT View
von: Nguyen, Son, et al.
Veröffentlicht: (2026)
von: Nguyen, Son, et al.
Veröffentlicht: (2026)
CleanerCLIP: Fine-grained Counterfactual Semantic Augmentation for Backdoor Defense in Contrastive Learning
von: Xun, Yuan, et al.
Veröffentlicht: (2024)
von: Xun, Yuan, et al.
Veröffentlicht: (2024)
PRACT: Optimizing Principled Reasoning and Acting of LLM Agent
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
Semantic Scheduling for LLM Inference
von: Hua, Wenyue, et al.
Veröffentlicht: (2025)
von: Hua, Wenyue, et al.
Veröffentlicht: (2025)
One Request, Multiple Experts: LLM Orchestrates Domain Specific Models via Adaptive Task Routing
von: Yang, Xu, et al.
Veröffentlicht: (2025)
von: Yang, Xu, et al.
Veröffentlicht: (2025)
Effective Learning for Small Reasoning Models: An Empirical Study on 0.5B Reasoning LLMs
von: Zhuang, Xialie, et al.
Veröffentlicht: (2025)
von: Zhuang, Xialie, et al.
Veröffentlicht: (2025)
SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning
von: Pan, Rui, et al.
Veröffentlicht: (2025)
von: Pan, Rui, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Trip Purpose Inference from GPS Trajectories via POI Semantic Zones and Pareto Calibration
von: Yang, Bo, et al.
Veröffentlicht: (2026)
von: Yang, Bo, et al.
Veröffentlicht: (2026)
SheetDesigner: MLLM-Powered Spreadsheet Layout Generation with Rule-Based and Vision-Based Reflection
von: Chen, Qin, et al.
Veröffentlicht: (2025)
von: Chen, Qin, et al.
Veröffentlicht: (2025)
Active Inference in Discrete State Spaces from First Principles
von: Kenny, Patrick
Veröffentlicht: (2025)
von: Kenny, Patrick
Veröffentlicht: (2025)
Ähnliche Einträge
-
Reasoning as an Attack Surface: Adaptive Evolutionary CoT Jailbreaks for LLMs
von: Li, Jianan, et al.
Veröffentlicht: (2026) -
Inverse Reinforcement Learning with Dynamic Reward Scaling for LLM Alignment
von: Cheng, Ruoxi, et al.
Veröffentlicht: (2025) -
Obscure but Effective: Classical Chinese Jailbreak Prompt Optimization via Bio-Inspired Search
von: Huang, Xun, et al.
Veröffentlicht: (2026) -
Heuristic-Induced Multimodal Risk Distribution Jailbreak Attack for Multimodal Large Language Models
von: Teng, Ma, et al.
Veröffentlicht: (2024) -
SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents
von: Jia, Xiaojun, et al.
Veröffentlicht: (2026)