Selective Adversarial Attacks on LLM Benchmarks
Fuente:
arXiv
Saved in:
| Main Authors: | Dubrovsky, Ivan, Orlova, Anastasia, Iov, Illarion, Gubina, Nina, Gureeva, Irena, Zaytsev, Alexey |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SLIME: Stabilized Likelihood Implicit Margin Enforcement for Preference Optimization
by: Afanasyev, Maksim, et al.
Published: (2026)
by: Afanasyev, Maksim, et al.
Published: (2026)
Concealed Adversarial attacks on neural networks for sequential data
by: Sokerin, Petr, et al.
Published: (2025)
by: Sokerin, Petr, et al.
Published: (2025)
Uniting contrastive and generative learning for event sequences models
by: Yugay, Aleksandr, et al.
Published: (2024)
by: Yugay, Aleksandr, et al.
Published: (2024)
Foundation for unbiased cross-validation of spatio-temporal models for species distribution modeling
by: Koldasbayeva, Diana, et al.
Published: (2025)
by: Koldasbayeva, Diana, et al.
Published: (2025)
Hiding Backdoors within Event Sequence Data via Poisoning Attacks
by: Ermilova, Alina, et al.
Published: (2023)
by: Ermilova, Alina, et al.
Published: (2023)
Parameter-Efficient Neural CDEs via Implicit Function Jacobians
by: Kuleshov, Ilya, et al.
Published: (2025)
by: Kuleshov, Ilya, et al.
Published: (2025)
Beyond Simple Averaging: Improving NLP Ensemble Performance with Topological-Data-Analysis-Based Weighting
by: Proskura, Polina, et al.
Published: (2024)
by: Proskura, Polina, et al.
Published: (2024)
Efficient Neural Controlled Differential Equations via Attentive Kernel Smoothing
by: Serov, Egor, et al.
Published: (2026)
by: Serov, Egor, et al.
Published: (2026)
Uncertainty Estimation of Transformers' Predictions via Topological Analysis of the Attention Matrices
by: Kostenok, Elizaveta, et al.
Published: (2023)
by: Kostenok, Elizaveta, et al.
Published: (2023)
U-Former ODE: Fast Probabilistic Forecasting of Irregular Time Series
by: Kuleshov, Ilya, et al.
Published: (2026)
by: Kuleshov, Ilya, et al.
Published: (2026)
Holistic Uncertainty Estimation For Open-Set Recognition
by: Erlygin, Leonid, et al.
Published: (2024)
by: Erlygin, Leonid, et al.
Published: (2024)
A theoretical framework for self-supervised contrastive learning for continuous dependent data
by: Marusov, Alexander, et al.
Published: (2025)
by: Marusov, Alexander, et al.
Published: (2025)
WWAggr: A Window Wasserstein-based Aggregation for Ensemble Change Point Detection
by: Stepikin, Alexander, et al.
Published: (2025)
by: Stepikin, Alexander, et al.
Published: (2025)
Language steering in latent space to mitigate unintended code-switching
by: Goncharov, Andrey, et al.
Published: (2025)
by: Goncharov, Andrey, et al.
Published: (2025)
Normalizing self-supervised learning for provably reliable Change Point Detection
by: Bazarova, Alexandra, et al.
Published: (2024)
by: Bazarova, Alexandra, et al.
Published: (2024)
When an LLM is apprehensive about its answers -- and when its uncertainty is justified
by: Sychev, Petr, et al.
Published: (2025)
by: Sychev, Petr, et al.
Published: (2025)
Strong Linear Baselines Strike Back: Closed-Form Linear Models as Gaussian Process Conditional Density Estimators for TSAD
by: Yugay, Aleksandr, et al.
Published: (2026)
by: Yugay, Aleksandr, et al.
Published: (2026)
Label Attention Network for Temporal Sets Prediction: You Were Looking at a Wrong Self-Attention
by: Kovtun, Elizaveta, et al.
Published: (2023)
by: Kovtun, Elizaveta, et al.
Published: (2023)
InDiD: Instant Disorder Detection via Representation Learning
by: Romanenkova, Evgenia, et al.
Published: (2021)
by: Romanenkova, Evgenia, et al.
Published: (2021)
Benchmarking Transferable Adversarial Attacks
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
Unveiling the Potential of AI for Nanomaterial Morphology Prediction
by: Dubrovsky, Ivan, et al.
Published: (2024)
by: Dubrovsky, Ivan, et al.
Published: (2024)
TabAttackBench: A Benchmark for Adversarial Attacks on Tabular Data
by: He, Zhipeng, et al.
Published: (2025)
by: He, Zhipeng, et al.
Published: (2025)
Never Skip a Batch: Continuous Training of Temporal GNNs via Adaptive Pseudo-Supervision
by: Panyshev, Alexander, et al.
Published: (2025)
by: Panyshev, Alexander, et al.
Published: (2025)
PINE: Pipeline for Important Node Exploration in Attributed Networks
by: Kovtun, Elizaveta, et al.
Published: (2025)
by: Kovtun, Elizaveta, et al.
Published: (2025)
From Variability to Stability: Advancing RecSys Benchmarking Practices
by: Shevchenko, Valeriy, et al.
Published: (2024)
by: Shevchenko, Valeriy, et al.
Published: (2024)
Adversarial Contrastive Learning for LLM Quantization Attacks
by: Song, Dinghong, et al.
Published: (2026)
by: Song, Dinghong, et al.
Published: (2026)
Collusion Detection with Graph Neural Networks
by: Gomes, Lucas, et al.
Published: (2024)
by: Gomes, Lucas, et al.
Published: (2024)
Complexity-aware fine-tuning
by: Goncharov, Andrey, et al.
Published: (2025)
by: Goncharov, Andrey, et al.
Published: (2025)
Surrogate uncertainty estimation for your time series forecasting black-box: learn when to trust
by: Erlygin, Leonid, et al.
Published: (2023)
by: Erlygin, Leonid, et al.
Published: (2023)
Paraphrasing Adversarial Attack on LLM-as-a-Reviewer
by: Kaneko, Masahiro
Published: (2026)
by: Kaneko, Masahiro
Published: (2026)
Looking around you: external information enhances representations for event sequences
by: Sokerin, Petr, et al.
Published: (2025)
by: Sokerin, Petr, et al.
Published: (2025)
Long-term drought prediction using deep neural networks based on geospatial weather data
by: Marusov, Alexander, et al.
Published: (2023)
by: Marusov, Alexander, et al.
Published: (2023)
Moshi Moshi? A Model Selection Hijacking Adversarial Attack
by: Petrucci, Riccardo, et al.
Published: (2025)
by: Petrucci, Riccardo, et al.
Published: (2025)
Enhancing Adversarial Attacks via Parameter Adaptive Adversarial Attack
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
Adversarial Attacks for Drift Detection
by: Hinder, Fabian, et al.
Published: (2024)
by: Hinder, Fabian, et al.
Published: (2024)
Adversarial Attacks on Data Attribution
by: Wang, Xinhe, et al.
Published: (2024)
by: Wang, Xinhe, et al.
Published: (2024)
Benchmarking Adversarial Patch Selection and Location
by: Kimhi, Shai, et al.
Published: (2025)
by: Kimhi, Shai, et al.
Published: (2025)
DARD: Dice Adversarial Robustness Distillation against Adversarial Attacks
by: Zou, Jing, et al.
Published: (2025)
by: Zou, Jing, et al.
Published: (2025)
Enhancing the Reliability of Medical AI through Expert-guided Uncertainty Modeling
by: Khalin, Aleksei, et al.
Published: (2026)
by: Khalin, Aleksei, et al.
Published: (2026)
DeNOTS: Stable Deep Neural ODEs for Time Series
by: Kuleshov, Ilya, et al.
Published: (2024)
by: Kuleshov, Ilya, et al.
Published: (2024)
Similar Items
-
SLIME: Stabilized Likelihood Implicit Margin Enforcement for Preference Optimization
by: Afanasyev, Maksim, et al.
Published: (2026) -
Concealed Adversarial attacks on neural networks for sequential data
by: Sokerin, Petr, et al.
Published: (2025) -
Uniting contrastive and generative learning for event sequences models
by: Yugay, Aleksandr, et al.
Published: (2024) -
Foundation for unbiased cross-validation of spatio-temporal models for species distribution modeling
by: Koldasbayeva, Diana, et al.
Published: (2025) -
Hiding Backdoors within Event Sequence Data via Poisoning Attacks
by: Ermilova, Alina, et al.
Published: (2023)