LLMScan: Causal Scan for LLM Misbehavior Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Mengdi, Goh, Kai Kiat, Zhang, Peixin, Sun, Jun, Xin, Rose Lin, Zhang, Hongyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Planning with Multi-Constraints via Collaborative Language Agents
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
Generalization of RLVR Using Causal Reasoning as a Testbed
von: Lu, Brian, et al.
Veröffentlicht: (2025)
von: Lu, Brian, et al.
Veröffentlicht: (2025)
LLM4Causal: Democratized Causal Tools for Everyone via Large Language Model
von: Jiang, Haitao, et al.
Veröffentlicht: (2023)
von: Jiang, Haitao, et al.
Veröffentlicht: (2023)
Higher-order Linear Attention
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates
von: Yang, Ling, et al.
Veröffentlicht: (2025)
von: Yang, Ling, et al.
Veröffentlicht: (2025)
Causal Abstraction in Model Interpretability: A Compact Survey
von: Zhang, Yihao
Veröffentlicht: (2024)
von: Zhang, Yihao
Veröffentlicht: (2024)
Code Comprehension then Auditing for Unsupervised LLM Evaluation
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
Absorber LLM: Harnessing Causal Synchronization for Test-Time Training
von: Zhang, Zhixin, et al.
Veröffentlicht: (2026)
von: Zhang, Zhixin, et al.
Veröffentlicht: (2026)
SABER: Switchable and Balanced Training for Efficient LLM Reasoning
von: Zhao, Kai, et al.
Veröffentlicht: (2025)
von: Zhao, Kai, et al.
Veröffentlicht: (2025)
StructEval: Deepen and Broaden Large Language Model Assessment via Structured Evaluation
von: Cao, Boxi, et al.
Veröffentlicht: (2024)
von: Cao, Boxi, et al.
Veröffentlicht: (2024)
InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory
von: Xiao, Chaojun, et al.
Veröffentlicht: (2024)
von: Xiao, Chaojun, et al.
Veröffentlicht: (2024)
Adversarial Preference Optimization: Enhancing Your Alignment via RM-LLM Game
von: Cheng, Pengyu, et al.
Veröffentlicht: (2023)
von: Cheng, Pengyu, et al.
Veröffentlicht: (2023)
LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals
von: Sun, Lihao, et al.
Veröffentlicht: (2026)
von: Sun, Lihao, et al.
Veröffentlicht: (2026)
Efficient Detection of LLM-generated Texts with a Bayesian Surrogate Model
von: Miao, Yibo, et al.
Veröffentlicht: (2023)
von: Miao, Yibo, et al.
Veröffentlicht: (2023)
Capability Instruction Tuning: A New Paradigm for Dynamic LLM Routing
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2025)
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2025)
Temporal Consistency for LLM Reasoning Process Error Identification
von: Guo, Jiacheng, et al.
Veröffentlicht: (2025)
von: Guo, Jiacheng, et al.
Veröffentlicht: (2025)
Causal Inference for Human-Language Model Collaboration
von: Zhang, Bohan, et al.
Veröffentlicht: (2024)
von: Zhang, Bohan, et al.
Veröffentlicht: (2024)
AIOS Compiler: LLM as Interpreter for Natural Language Programming and Flow Programming of AI Agents
von: Xu, Shuyuan, et al.
Veröffentlicht: (2024)
von: Xu, Shuyuan, et al.
Veröffentlicht: (2024)
Non-Markovian Discrete Diffusion with Causal Language Models
von: Zhang, Yangtian, et al.
Veröffentlicht: (2025)
von: Zhang, Yangtian, et al.
Veröffentlicht: (2025)
Evaluation is All You Need: Strategic Overclaiming of LLM Reasoning Capabilities Through Evaluation Design
von: Sun, Lin, et al.
Veröffentlicht: (2025)
von: Sun, Lin, et al.
Veröffentlicht: (2025)
FlashSampling: Fast and Memory-Efficient Exact Sampling
von: Ruiz, Tomas, et al.
Veröffentlicht: (2026)
von: Ruiz, Tomas, et al.
Veröffentlicht: (2026)
Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text
von: Zhou, Hongyi, et al.
Veröffentlicht: (2026)
von: Zhou, Hongyi, et al.
Veröffentlicht: (2026)
Layerwise Convergence Fingerprints for Runtime Misbehavior Detection in Large Language Models
von: Min, Nay Myat, et al.
Veröffentlicht: (2026)
von: Min, Nay Myat, et al.
Veröffentlicht: (2026)
Causality Extraction from Nuclear Licensee Event Reports Using a Hybrid Framework
von: Sohag, Shahidur Rahoman, et al.
Veröffentlicht: (2024)
von: Sohag, Shahidur Rahoman, et al.
Veröffentlicht: (2024)
AdaDetectGPT: Adaptive Detection of LLM-Generated Text with Statistical Guarantees
von: Zhou, Hongyi, et al.
Veröffentlicht: (2025)
von: Zhou, Hongyi, et al.
Veröffentlicht: (2025)
Sparse-RL: Breaking the Memory Wall in LLM Reinforcement Learning via Stable Sparse Rollouts
von: Luo, Sijia, et al.
Veröffentlicht: (2026)
von: Luo, Sijia, et al.
Veröffentlicht: (2026)
LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
Interactive Benchmarks
von: Yue, Baoqing, et al.
Veröffentlicht: (2026)
von: Yue, Baoqing, et al.
Veröffentlicht: (2026)
Causally-Enhanced Reinforcement Policy Optimization
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
Text Rationalization for Robust Causal Effect Estimation
von: Zhang, Lijinghua, et al.
Veröffentlicht: (2025)
von: Zhang, Lijinghua, et al.
Veröffentlicht: (2025)
Enhancing LLM Agent Safety via Causal Influence Prompting
von: Hahm, Dongyoon, et al.
Veröffentlicht: (2025)
von: Hahm, Dongyoon, et al.
Veröffentlicht: (2025)
SpecDec++: Boosting Speculative Decoding via Adaptive Candidate Lengths
von: Huang, Kaixuan, et al.
Veröffentlicht: (2024)
von: Huang, Kaixuan, et al.
Veröffentlicht: (2024)
Complete Chess Games Enable LLM Become A Chess Master
von: Zhang, Yinqi, et al.
Veröffentlicht: (2025)
von: Zhang, Yinqi, et al.
Veröffentlicht: (2025)
DPIC: Decoupling Prompt and Intrinsic Characteristics for LLM Generated Text Detection
von: Yu, Xiao, et al.
Veröffentlicht: (2023)
von: Yu, Xiao, et al.
Veröffentlicht: (2023)
FlowRL: Matching Reward Distributions for LLM Reasoning
von: Zhu, Xuekai, et al.
Veröffentlicht: (2025)
von: Zhu, Xuekai, et al.
Veröffentlicht: (2025)
GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation
von: Li, Sijia, et al.
Veröffentlicht: (2026)
von: Li, Sijia, et al.
Veröffentlicht: (2026)
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2023)
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2023)
Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge
von: Zhang, Wenbo, et al.
Veröffentlicht: (2026)
von: Zhang, Wenbo, et al.
Veröffentlicht: (2026)
Discovering Hierarchical Latent Capabilities of Language Models via Causal Representation Learning
von: Jin, Jikai, et al.
Veröffentlicht: (2025)
von: Jin, Jikai, et al.
Veröffentlicht: (2025)
Causality for Large Language Models
von: Wu, Anpeng, et al.
Veröffentlicht: (2024)
von: Wu, Anpeng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Planning with Multi-Constraints via Collaborative Language Agents
von: Zhang, Cong, et al.
Veröffentlicht: (2024) -
Generalization of RLVR Using Causal Reasoning as a Testbed
von: Lu, Brian, et al.
Veröffentlicht: (2025) -
LLM4Causal: Democratized Causal Tools for Everyone via Large Language Model
von: Jiang, Haitao, et al.
Veröffentlicht: (2023) -
Higher-order Linear Attention
von: Zhang, Yifan, et al.
Veröffentlicht: (2025) -
ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates
von: Yang, Ling, et al.
Veröffentlicht: (2025)