CoCoTen: Detecting Adversarial Inputs to Large Language Models through Latent Space Features of Contextual Co-occurrence Tensors
Fuente:
arXiv
Saved in:
| Main Authors: | Kadali, Sri Durga Sai Sowmya, Papalexakis, Evangelos E. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Jailbreaking Leaves a Trace: Understanding and Detecting Jailbreak Attacks from Internal Representations of Large Language Models
by: Kadali, Sri Durga Sai Sowmya, et al.
Published: (2026)
by: Kadali, Sri Durga Sai Sowmya, et al.
Published: (2026)
Do Internal Layers of LLMs Reveal Patterns for Jailbreak Detection?
by: Kadali, Sri Durga Sai Sowmya, et al.
Published: (2025)
by: Kadali, Sri Durga Sai Sowmya, et al.
Published: (2025)
GPT-generated Text Detection: Benchmark Dataset and Tensor-based Detection Method
by: Qazi, Zubair, et al.
Published: (2024)
by: Qazi, Zubair, et al.
Published: (2024)
Beyond ROUGE: N-Gram Subspace Features for LLM Hallucination Detection
by: Li, Jerry, et al.
Published: (2025)
by: Li, Jerry, et al.
Published: (2025)
Robust Vision-Language Models via Tensor Decomposition: A Defense Against Adversarial Attacks
by: Patel, Het, et al.
Published: (2025)
by: Patel, Het, et al.
Published: (2025)
Gradient Co-occurrence Analysis for Detecting Unsafe Prompts in Large Language Models
by: Yang, Jingyuan, et al.
Published: (2025)
by: Yang, Jingyuan, et al.
Published: (2025)
Co-occurrence is not Factual Association in Language Models
by: Zhang, Xiao, et al.
Published: (2024)
by: Zhang, Xiao, et al.
Published: (2024)
TRAWL: Tensor Reduced and Approximated Weights for Large Language Models
by: Luo, Yiran, et al.
Published: (2024)
by: Luo, Yiran, et al.
Published: (2024)
SamBaTen: Sampling-based Batch Incremental Tensor Decomposition
by: Gujral, Ekta, et al.
Published: (2017)
by: Gujral, Ekta, et al.
Published: (2017)
Every Response Counts: Quantifying Uncertainty of LLM-based Multi-Agent Systems through Tensor Decomposition
by: Chen, Tiejin, et al.
Published: (2026)
by: Chen, Tiejin, et al.
Published: (2026)
On the Distinctive Co-occurrence Characteristics of Antonymy
by: Cao, Zhihan, et al.
Published: (2025)
by: Cao, Zhihan, et al.
Published: (2025)
Hierarchical Concept Geometry in Language Models Emerges from Word Co-occurrence
by: Nava, Andres, et al.
Published: (2026)
by: Nava, Andres, et al.
Published: (2026)
Gradient-Regularized Latent Space Modulation in Large Language Models for Structured Contextual Synthesis
by: Yotheringhay, Derek, et al.
Published: (2025)
by: Yotheringhay, Derek, et al.
Published: (2025)
FRAPPE: $\underline{\text{F}}$ast $\underline{\text{Ra}}$nk $\underline{\text{App}}$roximation with $\underline{\text{E}}$xplainable Features for Tensors
by: Shiao, William, et al.
Published: (2022)
by: Shiao, William, et al.
Published: (2022)
Contextual Gradient Flow Modeling for Large Language Model Generalization in Multi-Scale Feature Spaces
by: Quillington, Daphne, et al.
Published: (2025)
by: Quillington, Daphne, et al.
Published: (2025)
Can a Large Language Model Learn Matrix Functions In Context?
by: Goulart, Paimon, et al.
Published: (2024)
by: Goulart, Paimon, et al.
Published: (2024)
CoViPAL: Layer-wise Contextualized Visual Token Pruning for Large Vision-Language Models
by: Tang, Zicong, et al.
Published: (2025)
by: Tang, Zicong, et al.
Published: (2025)
CoCoA: Confidence and Context-Aware Adaptive Decoding for Resolving Knowledge Conflicts in Large Language Models
by: Khandelwal, Anant, et al.
Published: (2025)
by: Khandelwal, Anant, et al.
Published: (2025)
CoCo-CoLa: Evaluating and Improving Language Adherence in Multilingual LLMs
by: Rahmati, Elnaz, et al.
Published: (2025)
by: Rahmati, Elnaz, et al.
Published: (2025)
Contextual Categorization Enhancement through LLMs Latent-Space
by: Bettouche, Zineddine, et al.
Published: (2024)
by: Bettouche, Zineddine, et al.
Published: (2024)
RoCoIns: Enhancing Robustness of Large Language Models through Code-Style Instructions
by: Zhang, Yuansen, et al.
Published: (2024)
by: Zhang, Yuansen, et al.
Published: (2024)
MATA: Mindful Assessment of the Telugu Abilities of Large Language Models
by: Kranti, Chalamalasetti, et al.
Published: (2025)
by: Kranti, Chalamalasetti, et al.
Published: (2025)
Prompt and Parameter Co-Optimization for Large Language Models
by: Bo, Xiaohe, et al.
Published: (2025)
by: Bo, Xiaohe, et al.
Published: (2025)
Incorporating Co‐occurrence Into the Operationalization of Speech Disfluency for Second Language Pronunciation and Oral Proficiency Assessment
by: Xun Yan, et al.
Published: (2025)
by: Xun Yan, et al.
Published: (2025)
Global and Local Structure Learning for Sparse Tensor Completion
by: Ahn, Dawon, et al.
Published: (2025)
by: Ahn, Dawon, et al.
Published: (2025)
LatentBreak: Jailbreaking Large Language Models through Latent Space Feedback
by: Mura, Raffaele, et al.
Published: (2025)
by: Mura, Raffaele, et al.
Published: (2025)
ProCoT: Stimulating Critical Thinking and Writing of Students through Engagement with Large Language Models (LLMs)
by: Adewumi, Tosin, et al.
Published: (2023)
by: Adewumi, Tosin, et al.
Published: (2023)
Instruction Tuning and CoT Prompting for Contextual Medical QA with LLMs
by: Le, Chenqian, et al.
Published: (2025)
by: Le, Chenqian, et al.
Published: (2025)
Investigating Co-Constructive Behavior of Large Language Models in Explanation Dialogues
by: Fichtel, Leandra, et al.
Published: (2025)
by: Fichtel, Leandra, et al.
Published: (2025)
Federated Co-tuning Framework for Large and Small Language Models
by: Fan, Tao, et al.
Published: (2024)
by: Fan, Tao, et al.
Published: (2024)
CoLLEGe: Concept Embedding Generation for Large Language Models
by: Teehan, Ryan, et al.
Published: (2024)
by: Teehan, Ryan, et al.
Published: (2024)
Uncertainty Quantification of Large Language Models through Multi-Dimensional Responses
by: Chen, Tiejin, et al.
Published: (2025)
by: Chen, Tiejin, et al.
Published: (2025)
CoLT: Reasoning with Chain of Latent Tool Calls
by: Zhu, Fangwei, et al.
Published: (2026)
by: Zhu, Fangwei, et al.
Published: (2026)
Investigating CoT Monitorability in Large Reasoning Models
by: Yang, Shu, et al.
Published: (2025)
by: Yang, Shu, et al.
Published: (2025)
DeCoVec: Building Decoding Space based Task Vector for Large Language Models via In-Context Learning
by: Li, Feiyang, et al.
Published: (2026)
by: Li, Feiyang, et al.
Published: (2026)
VidCoM: Fast Video Comprehension through Large Language Models with Multimodal Tools
by: Qi, Ji, et al.
Published: (2023)
by: Qi, Ji, et al.
Published: (2023)
Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders
by: Patel, Het, et al.
Published: (2026)
by: Patel, Het, et al.
Published: (2026)
Hierarchical Contextual Manifold Alignment for Structuring Latent Representations in Large Language Models
by: Dong, Meiquan, et al.
Published: (2025)
by: Dong, Meiquan, et al.
Published: (2025)
CoLa: Learning to Interactively Collaborate with Large Language Models
by: Sharma, Abhishek, et al.
Published: (2025)
by: Sharma, Abhishek, et al.
Published: (2025)
Making Metadata More FAIR Using Large Language Models
by: Sundaram, Sowmya S., et al.
Published: (2023)
by: Sundaram, Sowmya S., et al.
Published: (2023)
Similar Items
-
Jailbreaking Leaves a Trace: Understanding and Detecting Jailbreak Attacks from Internal Representations of Large Language Models
by: Kadali, Sri Durga Sai Sowmya, et al.
Published: (2026) -
Do Internal Layers of LLMs Reveal Patterns for Jailbreak Detection?
by: Kadali, Sri Durga Sai Sowmya, et al.
Published: (2025) -
GPT-generated Text Detection: Benchmark Dataset and Tensor-based Detection Method
by: Qazi, Zubair, et al.
Published: (2024) -
Beyond ROUGE: N-Gram Subspace Features for LLM Hallucination Detection
by: Li, Jerry, et al.
Published: (2025) -
Robust Vision-Language Models via Tensor Decomposition: A Defense Against Adversarial Attacks
by: Patel, Het, et al.
Published: (2025)