Phrase-Level Adversarial Training for Mitigating Bias in Neural Network-based Automatic Essay Scoring
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Philip, Haddad, Tashu, Tsegaye Misikir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mapping Transformer Leveraged Embeddings for Cross-Lingual Document Representation
von: Tashu, Tsegaye Misikir, et al.
Veröffentlicht: (2024)
von: Tashu, Tsegaye Misikir, et al.
Veröffentlicht: (2024)
Transformer-based Joint Modelling for Automatic Essay Scoring and Off-Topic Detection
von: Das, Sourya Dipta, et al.
Veröffentlicht: (2024)
von: Das, Sourya Dipta, et al.
Veröffentlicht: (2024)
Confidence Calibration in Large Language Model-Based Entity Matching
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
Pair2Score: Pairwise-to-Absolute Transfer for LLM-Based Essay Scoring
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026)
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026)
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
von: Li, Miaomiao, et al.
Veröffentlicht: (2025)
von: Li, Miaomiao, et al.
Veröffentlicht: (2025)
Understanding and Mitigating Tokenization Bias in Language Models
von: Phan, Buu, et al.
Veröffentlicht: (2024)
von: Phan, Buu, et al.
Veröffentlicht: (2024)
Mitigating Extrinsic Gender Bias for Bangla Classification Tasks
von: Joy, Sajib Kumar Saha, et al.
Veröffentlicht: (2024)
von: Joy, Sajib Kumar Saha, et al.
Veröffentlicht: (2024)
Quantifying and Mitigating Self-Preference Bias of LLM Judges
von: Yang, Jinming, et al.
Veröffentlicht: (2026)
von: Yang, Jinming, et al.
Veröffentlicht: (2026)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
A Rhythm-Aware Phrase Insertion for Classical Arabic Poetry Composition
von: Elzohbi, Mohamad, et al.
Veröffentlicht: (2025)
von: Elzohbi, Mohamad, et al.
Veröffentlicht: (2025)
Improve LLM-based Automatic Essay Scoring with Linguistic Features
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
Synthetic Prefixes to Mitigate Bias in Real-Time Neural Query Autocomplete
von: Rajan, Adithya, et al.
Veröffentlicht: (2025)
von: Rajan, Adithya, et al.
Veröffentlicht: (2025)
AGR: Age Group fairness Reward for Bias Mitigation in LLMs
von: Cao, Shuirong, et al.
Veröffentlicht: (2024)
von: Cao, Shuirong, et al.
Veröffentlicht: (2024)
Attention Speaks Volumes: Localizing and Mitigating Bias in Language Models
von: Adiga, Rishabh, et al.
Veröffentlicht: (2024)
von: Adiga, Rishabh, et al.
Veröffentlicht: (2024)
Shifting Perspectives: Steering Vectors for Robust Bias Mitigation in LLMs
von: Siddique, Zara, et al.
Veröffentlicht: (2025)
von: Siddique, Zara, et al.
Veröffentlicht: (2025)
MoESD: Mixture of Experts Stable Diffusion to Mitigate Gender Bias
von: Wang, Guorun, et al.
Veröffentlicht: (2024)
von: Wang, Guorun, et al.
Veröffentlicht: (2024)
Interactive Training: Feedback-Driven Neural Network Optimization
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Mitigating Degree Bias Adaptively with Hard-to-Learn Nodes in Graph Contrastive Learning
von: Hu, Jingyu, et al.
Veröffentlicht: (2025)
von: Hu, Jingyu, et al.
Veröffentlicht: (2025)
Mitigating Selection Bias in Large Language Models via Permutation-Aware GRPO
von: Zheng, Jinquan, et al.
Veröffentlicht: (2026)
von: Zheng, Jinquan, et al.
Veröffentlicht: (2026)
Mitigating Biases for Instruction-following Language Models via Bias Neurons Elimination
von: Yang, Nakyeong, et al.
Veröffentlicht: (2023)
von: Yang, Nakyeong, et al.
Veröffentlicht: (2023)
What is in a name? Mitigating Name Bias in Text Embeddings via Anonymization
von: Manchanda, Sahil, et al.
Veröffentlicht: (2025)
von: Manchanda, Sahil, et al.
Veröffentlicht: (2025)
Training the Untrainable: Introducing Inductive Bias via Representational Alignment
von: Subramaniam, Vighnesh, et al.
Veröffentlicht: (2024)
von: Subramaniam, Vighnesh, et al.
Veröffentlicht: (2024)
Reasoning Towards Fairness: Mitigating Bias in Language Models through Reasoning-Guided Fine-Tuning
von: Kabra, Sanchit, et al.
Veröffentlicht: (2025)
von: Kabra, Sanchit, et al.
Veröffentlicht: (2025)
Teaching LLMs for Step-Level Automatic Math Correction via Reinforcement Learning
von: Li, Junsong, et al.
Veröffentlicht: (2025)
von: Li, Junsong, et al.
Veröffentlicht: (2025)
Deceiving Question-Answering Models: A Hybrid Word-Level Adversarial Approach
von: Li, Jiyao, et al.
Veröffentlicht: (2024)
von: Li, Jiyao, et al.
Veröffentlicht: (2024)
When Greedy Wins: Emergent Exploitation Bias in Meta-Bandit LLM Training
von: Chen, Sanxing, et al.
Veröffentlicht: (2025)
von: Chen, Sanxing, et al.
Veröffentlicht: (2025)
Latent Adversarial Training Improves Robustness to Persistent Harmful Behaviors in LLMs
von: Sheshadri, Abhay, et al.
Veröffentlicht: (2024)
von: Sheshadri, Abhay, et al.
Veröffentlicht: (2024)
ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks
von: Yu, Xiaodong, et al.
Veröffentlicht: (2023)
von: Yu, Xiaodong, et al.
Veröffentlicht: (2023)
Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training
von: Lai, Song, et al.
Veröffentlicht: (2025)
von: Lai, Song, et al.
Veröffentlicht: (2025)
Identifying Features Associated with Bias Against 93 Stigmatized Groups in Language Models and Guardrail Model Safety Mitigation
von: Gueorguieva, Anna-Maria, et al.
Veröffentlicht: (2025)
von: Gueorguieva, Anna-Maria, et al.
Veröffentlicht: (2025)
Automatic Essay Scoring and Feedback Generation in Basque Language Learning
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
TLPO: Token-Level Policy Optimization for Mitigating Language Confusion in Large Language Models
von: Choo, Jinho, et al.
Veröffentlicht: (2026)
von: Choo, Jinho, et al.
Veröffentlicht: (2026)
Lightweight Safety Guardrails via Synthetic Data and RL-guided Adversarial Training
von: Ilin, Aleksei, et al.
Veröffentlicht: (2025)
von: Ilin, Aleksei, et al.
Veröffentlicht: (2025)
Mitigating Training Imbalance in LLM Fine-Tuning via Selective Parameter Merging
von: Ju, Yiming, et al.
Veröffentlicht: (2024)
von: Ju, Yiming, et al.
Veröffentlicht: (2024)
Emergence of Episodic Memory in Transformers: Characterizing Changes in Temporal Structure of Attention Scores During Training
von: Mistry, Deven Mahesh, et al.
Veröffentlicht: (2025)
von: Mistry, Deven Mahesh, et al.
Veröffentlicht: (2025)
Foundational Automatic Evaluators: Scaling Multi-Task Generative Evaluator Training for Reasoning-Centric Domains
von: Xu, Austin, et al.
Veröffentlicht: (2025)
von: Xu, Austin, et al.
Veröffentlicht: (2025)
Correcting Mean Bias in Text Embeddings: A Refined Renormalization with Training-Free Improvements on MMTEB
von: Ren, Xingyu, et al.
Veröffentlicht: (2025)
von: Ren, Xingyu, et al.
Veröffentlicht: (2025)
Impacts of Racial Bias in Historical Training Data for News AI
von: Bhargava, Rahul, et al.
Veröffentlicht: (2025)
von: Bhargava, Rahul, et al.
Veröffentlicht: (2025)
Mitigating Reversal Curse in Large Language Models via Semantic-aware Permutation Training
von: Guo, Qingyan, et al.
Veröffentlicht: (2024)
von: Guo, Qingyan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Mapping Transformer Leveraged Embeddings for Cross-Lingual Document Representation
von: Tashu, Tsegaye Misikir, et al.
Veröffentlicht: (2024) -
Transformer-based Joint Modelling for Automatic Essay Scoring and Off-Topic Detection
von: Das, Sourya Dipta, et al.
Veröffentlicht: (2024) -
Confidence Calibration in Large Language Model-Based Entity Matching
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025) -
Pair2Score: Pairwise-to-Absolute Transfer for LLM-Based Essay Scoring
von: Hallaç, İbrahim Rıza, et al.
Veröffentlicht: (2026) -
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
von: Li, Miaomiao, et al.
Veröffentlicht: (2025)