Label Forensics: Interpreting Hard Labels in Black-Box Text Classifier
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Du, Mengyao, Yang, Gang, Fang, Han, Yin, Quanjun, Chang, Ee-chien |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Accelerating Targeted Hard-Label Adversarial Attacks in Low-Query Black-Box Settings
von: Swaminathan, Arjhun, et al.
Veröffentlicht: (2025)
von: Swaminathan, Arjhun, et al.
Veröffentlicht: (2025)
Learning from Ambiguous Data with Hard Labels
von: Xie, Zeke, et al.
Veröffentlicht: (2025)
von: Xie, Zeke, et al.
Veröffentlicht: (2025)
Interpretable Discriminative Text Representations via Agreement and Label Disentanglement
von: Wang, Tong, et al.
Veröffentlicht: (2026)
von: Wang, Tong, et al.
Veröffentlicht: (2026)
Least-Ambiguous Multi-Label Classifier
von: Hagos, Misgina Tsighe, et al.
Veröffentlicht: (2025)
von: Hagos, Misgina Tsighe, et al.
Veröffentlicht: (2025)
Classifying Long-tailed and Label-noise Data via Disentangling and Unlearning
von: Shu, Chen, et al.
Veröffentlicht: (2025)
von: Shu, Chen, et al.
Veröffentlicht: (2025)
Learning from Hard Labels with Additional Supervision on Non-Hard-Labeled Classes
von: Sugiyama, Kosuke, et al.
Veröffentlicht: (2025)
von: Sugiyama, Kosuke, et al.
Veröffentlicht: (2025)
The Risk of Federated Learning to Skew Fine-Tuning Features and Underperform Out-of-Distribution Robustness
von: Du, Mengyao, et al.
Veröffentlicht: (2024)
von: Du, Mengyao, et al.
Veröffentlicht: (2024)
Classifier Chain Networks for Multi-Label Classification
von: Touw, Daniel J. W., et al.
Veröffentlicht: (2024)
von: Touw, Daniel J. W., et al.
Veröffentlicht: (2024)
Dual-Label Learning With Irregularly Present Labels
von: Li, Mingqian, et al.
Veröffentlicht: (2024)
von: Li, Mingqian, et al.
Veröffentlicht: (2024)
Multi-Label Bayesian Active Learning with Inter-Label Relationships
von: Qi, Yuanyuan, et al.
Veröffentlicht: (2024)
von: Qi, Yuanyuan, et al.
Veröffentlicht: (2024)
TrapSuffix: Proactive Defense Against Adversarial Suffixes in Jailbreaking
von: Du, Mengyao, et al.
Veröffentlicht: (2026)
von: Du, Mengyao, et al.
Veröffentlicht: (2026)
Domain Bridge: Generative model-based domain forensic for black-box models
von: Zhang, Jiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Jiyi, et al.
Veröffentlicht: (2024)
Calibratable Disambiguation Loss for Multi-Instance Partial-Label Learning
von: Tang, Wei, et al.
Veröffentlicht: (2025)
von: Tang, Wei, et al.
Veröffentlicht: (2025)
Uncertainty Calibration of Multi-Label Bird Sound Classifiers
von: Schwinger, Raphael, et al.
Veröffentlicht: (2025)
von: Schwinger, Raphael, et al.
Veröffentlicht: (2025)
Multi-Instance Partial-Label Learning with Margin Adjustment
von: Tang, Wei, et al.
Veröffentlicht: (2025)
von: Tang, Wei, et al.
Veröffentlicht: (2025)
Model Evaluation in the Dark: Robust Classifier Metrics with Missing Labels
von: Dervovic, Danial, et al.
Veröffentlicht: (2025)
von: Dervovic, Danial, et al.
Veröffentlicht: (2025)
Graphite: A Graph-based Extreme Multi-Label Short Text Classifier for Keyphrase Recommendation
von: Mishra, Ashirbad, et al.
Veröffentlicht: (2024)
von: Mishra, Ashirbad, et al.
Veröffentlicht: (2024)
Learning with Confidence: Training Better Classifiers from Soft Labels
von: de Vries, Sjoerd, et al.
Veröffentlicht: (2024)
von: de Vries, Sjoerd, et al.
Veröffentlicht: (2024)
LICORICE: Label-Efficient Concept-Based Interpretable Reinforcement Learning
von: Ye, Zhuorui, et al.
Veröffentlicht: (2024)
von: Ye, Zhuorui, et al.
Veröffentlicht: (2024)
Image-based Novel Fault Detection with Deep Learning Classifiers using Hierarchical Labels
von: Sergin, Nurettin, et al.
Veröffentlicht: (2024)
von: Sergin, Nurettin, et al.
Veröffentlicht: (2024)
Federated Learning with Only Positive Labels by Exploring Label Correlations
von: An, Xuming, et al.
Veröffentlicht: (2024)
von: An, Xuming, et al.
Veröffentlicht: (2024)
Not All Data are Good Labels: On the Self-supervised Labeling for Time Series Forecasting
von: Yang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Yang, Yuxuan, et al.
Veröffentlicht: (2025)
Is the Hard-Label Cryptanalytic Model Extraction Really Polynomial?
von: Ito, Akira, et al.
Veröffentlicht: (2025)
von: Ito, Akira, et al.
Veröffentlicht: (2025)
Hard-Label Cryptanalytic Extraction of Neural Network Models
von: Chen, Yi, et al.
Veröffentlicht: (2024)
von: Chen, Yi, et al.
Veröffentlicht: (2024)
A Theoretical Analysis of Soft-Label vs Hard-Label Training in Neural Networks
von: Mandal, Saptarshi, et al.
Veröffentlicht: (2024)
von: Mandal, Saptarshi, et al.
Veröffentlicht: (2024)
Labels in Extremes: How Well Calibrated are Extreme Multi-label Classifiers?
von: Ullah, Nasib, et al.
Veröffentlicht: (2024)
von: Ullah, Nasib, et al.
Veröffentlicht: (2024)
Multi-Label Adaptive Batch Selection by Highlighting Hard and Imbalanced Samples
von: Zhou, Ao, et al.
Veröffentlicht: (2024)
von: Zhou, Ao, et al.
Veröffentlicht: (2024)
Decoupling the Class Label and the Target Concept in Machine Unlearning
von: Zhu, Jianing, et al.
Veröffentlicht: (2024)
von: Zhu, Jianing, et al.
Veröffentlicht: (2024)
EchoAlign: Bridging Generative and Discriminative Learning under Noisy Labels
von: Zheng, Yuxiang, et al.
Veröffentlicht: (2024)
von: Zheng, Yuxiang, et al.
Veröffentlicht: (2024)
The Bridge-Garden Dilemma in LLM Distillation: Why Mixing Hard and Soft Labels Works
von: Wang, Guanghui, et al.
Veröffentlicht: (2026)
von: Wang, Guanghui, et al.
Veröffentlicht: (2026)
Conformal Prediction of Classifiers with Many Classes based on Noisy Labels
von: Penso, Coby, et al.
Veröffentlicht: (2025)
von: Penso, Coby, et al.
Veröffentlicht: (2025)
On the Role of Label Noise in the Feature Learning Process
von: Han, Andi, et al.
Veröffentlicht: (2025)
von: Han, Andi, et al.
Veröffentlicht: (2025)
From Biased Selective Labels to Pseudo-Labels: An Expectation-Maximization Framework for Learning from Biased Decisions
von: Chang, Trenton, et al.
Veröffentlicht: (2024)
von: Chang, Trenton, et al.
Veröffentlicht: (2024)
Enhanced Labeling Technique for Reddit Text and Fine-Tuned Longformer Models for Classifying Depression Severity in English and Luganda
von: Kimera, Richard, et al.
Veröffentlicht: (2024)
von: Kimera, Richard, et al.
Veröffentlicht: (2024)
Remining Hard Negatives for Generative Pseudo Labeled Domain Adaptation
von: Yuksel, Goksenin, et al.
Veröffentlicht: (2025)
von: Yuksel, Goksenin, et al.
Veröffentlicht: (2025)
Low-Cost Hard-Label Adversarial Attack with Theoretical Foundations
von: Liu, Jun, et al.
Veröffentlicht: (2026)
von: Liu, Jun, et al.
Veröffentlicht: (2026)
Retraining with Predicted Hard Labels Provably Increases Model Accuracy
von: Das, Rudrajit, et al.
Veröffentlicht: (2024)
von: Das, Rudrajit, et al.
Veröffentlicht: (2024)
Unlocking the Black Box of Latent Reasoning: An Interpretability-Guided Approach to Intervention
von: Chang, Shuochen, et al.
Veröffentlicht: (2026)
von: Chang, Shuochen, et al.
Veröffentlicht: (2026)
Feature-Label Modal Alignment for Robust Partial Multi-Label Learning
von: Chen, Yu, et al.
Veröffentlicht: (2026)
von: Chen, Yu, et al.
Veröffentlicht: (2026)
Same Target, Different Basins: Hard vs. Soft Labels for Annotator Distributions
von: Gheibi, Mirerfan, et al.
Veröffentlicht: (2026)
von: Gheibi, Mirerfan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Accelerating Targeted Hard-Label Adversarial Attacks in Low-Query Black-Box Settings
von: Swaminathan, Arjhun, et al.
Veröffentlicht: (2025) -
Learning from Ambiguous Data with Hard Labels
von: Xie, Zeke, et al.
Veröffentlicht: (2025) -
Interpretable Discriminative Text Representations via Agreement and Label Disentanglement
von: Wang, Tong, et al.
Veröffentlicht: (2026) -
Least-Ambiguous Multi-Label Classifier
von: Hagos, Misgina Tsighe, et al.
Veröffentlicht: (2025) -
Classifying Long-tailed and Label-noise Data via Disentangling and Unlearning
von: Shu, Chen, et al.
Veröffentlicht: (2025)