Gespeichert in:
| 1. Verfasser: | Parekh, Swapnil |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2604.00770 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
CIRCUS: Circuit Consensus under Uncertainty via Stability Ensembles
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
CaptionFool: Universal Image Captioning Model Attacks
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
FLAT: Latent-Driven Arbitrary-Target Backdoor Attacks in Federated Learning
von: Nguyen, Tuan, et al.
Veröffentlicht: (2025)
von: Nguyen, Tuan, et al.
Veröffentlicht: (2025)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
von: Xu, Xin, et al.
Veröffentlicht: (2026)
von: Xu, Xin, et al.
Veröffentlicht: (2026)
Thinking in Latents: Adaptive Anchor Refinement for Implicit Reasoning in LLMs
von: Sheshanarayana, Disha, et al.
Veröffentlicht: (2026)
von: Sheshanarayana, Disha, et al.
Veröffentlicht: (2026)
Heterogeneous Graph Backdoor Attack
von: Chen, Jiawei, et al.
Veröffentlicht: (2025)
von: Chen, Jiawei, et al.
Veröffentlicht: (2025)
LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning
von: Ye, Xinwu, et al.
Veröffentlicht: (2026)
von: Ye, Xinwu, et al.
Veröffentlicht: (2026)
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning
von: Guo, Weiran, et al.
Veröffentlicht: (2025)
von: Guo, Weiran, et al.
Veröffentlicht: (2025)
Defending Deep Regression Models against Backdoor Attacks
von: Du, Lingyu, et al.
Veröffentlicht: (2024)
von: Du, Lingyu, et al.
Veröffentlicht: (2024)
Backdoor Vectors: a Task Arithmetic View on Backdoor Attacks and Defenses
von: Pawlak, Stanisław, et al.
Veröffentlicht: (2025)
von: Pawlak, Stanisław, et al.
Veröffentlicht: (2025)
Leveraging Knowledge Graphs and LLM Reasoning to Identify Operational Bottlenecks for Warehouse Planning Assistance
von: Parekh, Rishi, et al.
Veröffentlicht: (2025)
von: Parekh, Rishi, et al.
Veröffentlicht: (2025)
Do Latent-CoT Models Think Step-by-Step? A Mechanistic Study on Sequential Reasoning Tasks
von: Liang, Jia, et al.
Veröffentlicht: (2026)
von: Liang, Jia, et al.
Veröffentlicht: (2026)
Cooperative Backdoor Attack in Decentralized Reinforcement Learning with Theoretical Guarantee
von: Gao, Mengtong, et al.
Veröffentlicht: (2024)
von: Gao, Mengtong, et al.
Veröffentlicht: (2024)
Adversarial Robustness Unhardening via Backdoor Attacks in Federated Learning
von: Kim, Taejin, et al.
Veröffentlicht: (2023)
von: Kim, Taejin, et al.
Veröffentlicht: (2023)
Geometry-Aware Backdoor Attacks: Leveraging Curvature in Hyperbolic Embeddings
von: Baheri, Ali
Veröffentlicht: (2025)
von: Baheri, Ali
Veröffentlicht: (2025)
Unlearn to Relearn Backdoors: Deferred Backdoor Functionality Attacks on Deep Learning Models
von: Shin, Jeongjin, et al.
Veröffentlicht: (2024)
von: Shin, Jeongjin, et al.
Veröffentlicht: (2024)
Backdoor for Debias: Mitigating Model Bias with Backdoor Attack-based Artificial Bias
von: Wu, Shangxi, et al.
Veröffentlicht: (2023)
von: Wu, Shangxi, et al.
Veröffentlicht: (2023)
CANARY: Zero-Label Detection of Fine-Tuning Contamination in Language Models
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
PeerGuard: Defending Multi-Agent Systems Against Backdoor Attacks Through Mutual Reasoning
von: Fan, Falong, et al.
Veröffentlicht: (2025)
von: Fan, Falong, et al.
Veröffentlicht: (2025)
Latent Space Data Fusion Outperforms Early Fusion in Multimodal Mental Health Digital Phenotyping Data
von: Barkat, Youcef, et al.
Veröffentlicht: (2025)
von: Barkat, Youcef, et al.
Veröffentlicht: (2025)
Compromising Embodied Agents with Contextual Backdoor Attacks
von: Liu, Aishan, et al.
Veröffentlicht: (2024)
von: Liu, Aishan, et al.
Veröffentlicht: (2024)
The Eminence in Shadow: Exploiting Feature Boundary Ambiguity for Robust Backdoor Attacks
von: Feng, Zhou, et al.
Veröffentlicht: (2025)
von: Feng, Zhou, et al.
Veröffentlicht: (2025)
HeteroBA: A Structure-Manipulating Backdoor Attack on Heterogeneous Graphs
von: Gao, Honglin, et al.
Veröffentlicht: (2025)
von: Gao, Honglin, et al.
Veröffentlicht: (2025)
ACES: Accent Subspaces for Coupling, Explanations, and Stress-Testing in Automatic Speech Recognition
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
Structure-Aware Distributed Backdoor Attacks in Federated Learning
von: Jian, Wang, et al.
Veröffentlicht: (2026)
von: Jian, Wang, et al.
Veröffentlicht: (2026)
ICLShield: Exploring and Mitigating In-Context Learning Backdoor Attacks
von: Ren, Zhiyao, et al.
Veröffentlicht: (2025)
von: Ren, Zhiyao, et al.
Veröffentlicht: (2025)
BACKTIME: Backdoor Attacks on Multivariate Time Series Forecasting
von: Lin, Xiao, et al.
Veröffentlicht: (2024)
von: Lin, Xiao, et al.
Veröffentlicht: (2024)
Invisible Backdoor Attack Through Singular Value Decomposition
von: Chen, Wenmin, et al.
Veröffentlicht: (2024)
von: Chen, Wenmin, et al.
Veröffentlicht: (2024)
Poisoning the Inner Prediction Logic of Graph Neural Networks for Clean-Label Backdoor Attacks
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2026)
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2026)
Exploiting Layer-Specific Vulnerabilities to Backdoor Attack in Federated Learning
von: Foroughi, Mohammad Hadi, et al.
Veröffentlicht: (2026)
von: Foroughi, Mohammad Hadi, et al.
Veröffentlicht: (2026)
Backdoor Attacks on Fault Detection and Localization in Cyber-Physical Systems
von: Jean, Abile, et al.
Veröffentlicht: (2026)
von: Jean, Abile, et al.
Veröffentlicht: (2026)
Your Agent Can Defend Itself against Backdoor Attacks
von: Changjiang, Li, et al.
Veröffentlicht: (2025)
von: Changjiang, Li, et al.
Veröffentlicht: (2025)
Backdoor Attack on Vertical Federated Graph Neural Network Learning
von: Yang, Jirui, et al.
Veröffentlicht: (2024)
von: Yang, Jirui, et al.
Veröffentlicht: (2024)
Revisiting Backdoor Attacks on Time Series Classification in the Frequency Domain
von: Huang, Yuanmin, et al.
Veröffentlicht: (2025)
von: Huang, Yuanmin, et al.
Veröffentlicht: (2025)
Client-Side Patching against Backdoor Attacks in Federated Learning
von: Molina-Coronado, Borja
Veröffentlicht: (2024)
von: Molina-Coronado, Borja
Veröffentlicht: (2024)
All AI Models are Wrong, but Some are Optimal
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
Think in Blocks: Adaptive Reasoning from Direct Response to Deep Reasoning
von: Zhu, Yekun, et al.
Veröffentlicht: (2025)
von: Zhu, Yekun, et al.
Veröffentlicht: (2025)
ART: Adaptive Reasoning Trees for Explainable Claim Verification
von: Wadhwa, Sahil, et al.
Veröffentlicht: (2026)
von: Wadhwa, Sahil, et al.
Veröffentlicht: (2026)
Certifying Language Model Robustness with Fuzzed Randomized Smoothing: An Efficient Defense Against Backdoor Attacks
von: He, Bowei, et al.
Veröffentlicht: (2025)
von: He, Bowei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning
von: Parekh, Swapnil
Veröffentlicht: (2026) -
CIRCUS: Circuit Consensus under Uncertainty via Stability Ensembles
von: Parekh, Swapnil
Veröffentlicht: (2026) -
CaptionFool: Universal Image Captioning Model Attacks
von: Parekh, Swapnil
Veröffentlicht: (2026) -
FLAT: Latent-Driven Arbitrary-Target Backdoor Attacks in Federated Learning
von: Nguyen, Tuan, et al.
Veröffentlicht: (2025) -
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
von: Xu, Xin, et al.
Veröffentlicht: (2026)