Saved in:
| Main Authors: | Cai, Rui, Mo, Weijie Jacky, Wen, Xiaofei, Ma, Qiyao, Zhu, Wenhui, Chen, Xiwen, Chen, Muhao, Zhao, Zhe |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.07075 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diagnosing and Mitigating Modality Interference in Multimodal Large Language Models
by: Cai, Rui, et al.
Published: (2025)
by: Cai, Rui, et al.
Published: (2025)
ThinkGuard: Deliberative Slow Thinking Leads to Cautious Guardrails
by: Wen, Xiaofei, et al.
Published: (2025)
by: Wen, Xiaofei, et al.
Published: (2025)
OmniGuard: Unified Omni-Modal Guardrails with Deliberate Reasoning
by: Zhu, Boyu, et al.
Published: (2025)
by: Zhu, Boyu, et al.
Published: (2025)
DebugLM: Learning Traceable Training Data Provenance for LLMs
by: Mo, Wenjie Jacky, et al.
Published: (2026)
by: Mo, Wenjie Jacky, et al.
Published: (2026)
Triaging Threats to Specialized Guardrails
by: Mo, Wenjie Jacky, et al.
Published: (2026)
by: Mo, Wenjie Jacky, et al.
Published: (2026)
Learning Efficient Guardrails for Compliance
by: Wen, Xiaofei, et al.
Published: (2025)
by: Wen, Xiaofei, et al.
Published: (2025)
Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization
by: Ma, Qiyao, et al.
Published: (2026)
by: Ma, Qiyao, et al.
Published: (2026)
RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
by: Mo, Wenjie Jacky, et al.
Published: (2025)
by: Mo, Wenjie Jacky, et al.
Published: (2025)
When Vision Speaks for Sound
by: Wen, Xiaofei, et al.
Published: (2026)
by: Wen, Xiaofei, et al.
Published: (2026)
Form-Finding and Physical Property Predictions of Tensegrity Structures Using Deep Neural Networks
by: Chen, Muhao, et al.
Published: (2024)
by: Chen, Muhao, et al.
Published: (2024)
You Are the Best Reviewer of Your Own Papers: The Isotonic Mechanism
by: Su, Weijie
Published: (2022)
by: Su, Weijie
Published: (2022)
AI Pangaea: Unifying Intelligence Islands for Adapting Myriad Tasks
by: Chang, Jianlong, et al.
Published: (2025)
by: Chang, Jianlong, et al.
Published: (2025)
Red Teaming Language Models for Processing Contradictory Dialogues
by: Wen, Xiaofei, et al.
Published: (2024)
by: Wen, Xiaofei, et al.
Published: (2024)
DRA-GRPO: Your GRPO Needs to Know Diverse Reasoning Paths for Mathematical Reasoning
by: Chen, Xiwen, et al.
Published: (2025)
by: Chen, Xiwen, et al.
Published: (2025)
SHARP: A Self-Evolving Human-Auditable Rubric Policy for Financial Trading Agents
by: Chen, Xiwen, et al.
Published: (2026)
by: Chen, Xiwen, et al.
Published: (2026)
From Introspection to Best Practices: Principled Analysis of Demonstrations in Multimodal In-Context Learning
by: Xu, Nan, et al.
Published: (2024)
by: Xu, Nan, et al.
Published: (2024)
SODA: Semi On-Policy Black-Box Distillation for Large Language Models
by: Chen, Xiwen, et al.
Published: (2026)
by: Chen, Xiwen, et al.
Published: (2026)
Towards a Mechanistic Understanding of Propositional Logical Reasoning in Large Language Models
by: Chen, Danchun, et al.
Published: (2026)
by: Chen, Danchun, et al.
Published: (2026)
Symmetry-Reduced Physics-Informed Learning of Tensegrity Dynamics
by: Qin, Jing, et al.
Published: (2026)
by: Qin, Jing, et al.
Published: (2026)
Active Layer-Contrastive Decoding Reduces Hallucination in Large Language Model Generation
by: Zhang, Hongxiang, et al.
Published: (2025)
by: Zhang, Hongxiang, et al.
Published: (2025)
Modeling Time Series Dynamics with Fourier Ordinary Differential Equations
by: Guo, Muhao, et al.
Published: (2025)
by: Guo, Muhao, et al.
Published: (2025)
Solar Photovoltaic Assessment with Large Language Model
by: Guo, Muhao, et al.
Published: (2025)
by: Guo, Muhao, et al.
Published: (2025)
K-Myriad: Jump-starting reinforcement learning with unsupervised parallel agents
by: De Paola, Vincenzo, et al.
Published: (2026)
by: De Paola, Vincenzo, et al.
Published: (2026)
Imaging Signal Recovery Using Neural Network Priors Under Uncertain Forward Model Parameters
by: Chen, Xiwen, et al.
Published: (2024)
by: Chen, Xiwen, et al.
Published: (2024)
Mitigating Backdoor Threats to Large Language Models: Advancement and Challenges
by: Liu, Qin, et al.
Published: (2024)
by: Liu, Qin, et al.
Published: (2024)
FlexCare: Leveraging Cross-Task Synergy for Flexible Multimodal Healthcare Prediction
by: Xu, Muhao, et al.
Published: (2024)
by: Xu, Muhao, et al.
Published: (2024)
EVTP-IVS: Effective Visual Token Pruning For Unifying Instruction Visual Segmentation In Multi-Modal Large Language Models
by: Zhu, Wenhui, et al.
Published: (2025)
by: Zhu, Wenhui, et al.
Published: (2025)
Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
by: Xu, Jiashu, et al.
Published: (2023)
by: Xu, Jiashu, et al.
Published: (2023)
Efficient Prompt Optimization Through the Lens of Best Arm Identification
by: Shi, Chengshuai, et al.
Published: (2024)
by: Shi, Chengshuai, et al.
Published: (2024)
What Makes Quantization for Large Language Models Hard? An Empirical Study from the Lens of Perturbation
by: Gong, Zhuocheng, et al.
Published: (2024)
by: Gong, Zhuocheng, et al.
Published: (2024)
Veni, Vidi, Vici: Solving the Myriad of Challenges before Knowledge Graph Learning
by: Sardina, Jeffrey, et al.
Published: (2024)
by: Sardina, Jeffrey, et al.
Published: (2024)
Securing Multi-turn Conversational Language Models From Distributed Backdoor Triggers
by: Tong, Terry, et al.
Published: (2024)
by: Tong, Terry, et al.
Published: (2024)
TimeMIL: Advancing Multivariate Time Series Classification via a Time-aware Multiple Instance Learning
by: Chen, Xiwen, et al.
Published: (2024)
by: Chen, Xiwen, et al.
Published: (2024)
Sequence Complementor: Complementing Transformers For Time Series Forecasting with Learnable Sequences
by: Chen, Xiwen, et al.
Published: (2025)
by: Chen, Xiwen, et al.
Published: (2025)
SiGNN: A Spike-induced Graph Neural Network for Dynamic Graph Representation Learning
by: Chen, Dong, et al.
Published: (2024)
by: Chen, Dong, et al.
Published: (2024)
DeepScientist: Advancing Frontier-Pushing Scientific Findings Progressively
by: Weng, Yixuan, et al.
Published: (2025)
by: Weng, Yixuan, et al.
Published: (2025)
Instructional Fingerprinting of Large Language Models
by: Xu, Jiashu, et al.
Published: (2024)
by: Xu, Jiashu, et al.
Published: (2024)
Your Agent is More Brittle Than You Think: Uncovering Indirect Injection Vulnerabilities in Agentic LLMs
by: Zhu, Wenhui, et al.
Published: (2026)
by: Zhu, Wenhui, et al.
Published: (2026)
How Effective Can Dropout Be in Multiple Instance Learning ?
by: Zhu, Wenhui, et al.
Published: (2025)
by: Zhu, Wenhui, et al.
Published: (2025)
Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models
by: Nahin, Shahriar Kabir, et al.
Published: (2025)
by: Nahin, Shahriar Kabir, et al.
Published: (2025)
Similar Items
-
Diagnosing and Mitigating Modality Interference in Multimodal Large Language Models
by: Cai, Rui, et al.
Published: (2025) -
ThinkGuard: Deliberative Slow Thinking Leads to Cautious Guardrails
by: Wen, Xiaofei, et al.
Published: (2025) -
OmniGuard: Unified Omni-Modal Guardrails with Deliberate Reasoning
by: Zhu, Boyu, et al.
Published: (2025) -
DebugLM: Learning Traceable Training Data Provenance for LLMs
by: Mo, Wenjie Jacky, et al.
Published: (2026) -
Triaging Threats to Specialized Guardrails
by: Mo, Wenjie Jacky, et al.
Published: (2026)