Gespeichert in:
| Hauptverfasser: | Kaneko, Masahiro, Bollegala, Danushka, Baldwin, Timothy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2401.08511 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
Eagle: Ethical Dataset Given from Real Interactions
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
In-Contextual Gender Bias Suppression for Large Language Models
von: Oba, Daisuke, et al.
Veröffentlicht: (2023)
von: Oba, Daisuke, et al.
Veröffentlicht: (2023)
Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
Evaluating the Evaluation of Diversity in Commonsense Generation
von: Zhang, Tianhui, et al.
Veröffentlicht: (2025)
von: Zhang, Tianhui, et al.
Veröffentlicht: (2025)
A Semantic Distance Metric Learning approach for Lexical Semantic Change Detection
von: Aida, Taichi, et al.
Veröffentlicht: (2024)
von: Aida, Taichi, et al.
Veröffentlicht: (2024)
Investigating the Contextualised Word Embedding Dimensions Specified for Contextual and Temporal Semantic Changes
von: Aida, Taichi, et al.
Veröffentlicht: (2024)
von: Aida, Taichi, et al.
Veröffentlicht: (2024)
SCDTour: Embedding Axis Ordering and Merging for Interpretable Semantic Change Detection
von: Aida, Taichi, et al.
Veröffentlicht: (2025)
von: Aida, Taichi, et al.
Veröffentlicht: (2025)
Map of Encoders -- Mapping Sentence Encoders using Quantum Relative Entropy
von: Zhang, Gaifan, et al.
Veröffentlicht: (2026)
von: Zhang, Gaifan, et al.
Veröffentlicht: (2026)
Evaluating the Effect of Retrieval Augmentation on Social Biases
von: Zhang, Tianhui, et al.
Veröffentlicht: (2025)
von: Zhang, Tianhui, et al.
Veröffentlicht: (2025)
Evaluating Unsupervised Dimensionality Reduction Methods for Pretrained Sentence Embeddings
von: Zhang, Gaifan, et al.
Veröffentlicht: (2024)
von: Zhang, Gaifan, et al.
Veröffentlicht: (2024)
Evaluating Short-Term Temporal Fluctuations of Social Biases in Social Media Data and Masked Language Models
von: Zhou, Yi, et al.
Veröffentlicht: (2024)
von: Zhou, Yi, et al.
Veröffentlicht: (2024)
A Little Leak Will Sink a Great Ship: Survey of Transparency for Large Language Models from Start to Finish
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
Improving Diversity of Commonsense Generation by Large Language Models via In-Context Learning
von: Zhang, Tianhui, et al.
Veröffentlicht: (2024)
von: Zhang, Tianhui, et al.
Veröffentlicht: (2024)
Synthetic Data Generation for Training Diversified Commonsense Reasoning Models
von: Zhang, Tianhui, et al.
Veröffentlicht: (2026)
von: Zhang, Tianhui, et al.
Veröffentlicht: (2026)
Annotating Training Data for Conditional Semantic Textual Similarity Measurement using Large Language Models
von: Zhang, Gaifan, et al.
Veröffentlicht: (2025)
von: Zhang, Gaifan, et al.
Veröffentlicht: (2025)
CASE -- Condition-Aware Sentence Embeddings for Conditional Semantic Textual Similarity Measurement
von: Zhang, Gaifan, et al.
Veröffentlicht: (2025)
von: Zhang, Gaifan, et al.
Veröffentlicht: (2025)
Evaluating Gender Bias of Pre-trained Language Models in Natural Language Inference by Considering All Labels
von: Anantaprayoon, Panatchakorn, et al.
Veröffentlicht: (2023)
von: Anantaprayoon, Panatchakorn, et al.
Veröffentlicht: (2023)
Bits Leaked per Query: Information-Theoretic Bounds on Adversarial Attacks against LLMs
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
Improving Pre-trained Language Model Sensitivity via Mask Specific losses: A case study on Biomedical NER
von: Abaho, Micheal, et al.
Veröffentlicht: (2024)
von: Abaho, Micheal, et al.
Veröffentlicht: (2024)
Online Learning Defense against Iterative Jailbreak Attacks via Prompt Optimization
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
Beyond the Resumé: A Rubric-Aware Automatic Interview System for Information Elicitation
von: Stuart, Harry, et al.
Veröffentlicht: (2026)
von: Stuart, Harry, et al.
Veröffentlicht: (2026)
Improving Unsupervised Constituency Parsing via Maximizing Semantic Information
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
Unsupervised Parsing by Searching for Frequent Word Sequences among Sentences with Equivalent Predicate-Argument Structures
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
Neuron-Level Analysis of Cultural Understanding in Large Language Models
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
JailNewsBench: Multi-Lingual and Regional Benchmark for Fake News Generation under Jailbreak Attacks
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2026)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2026)
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
Social Bias Evaluation for Large Language Models Requires Prompt Variations
von: Hida, Rem, et al.
Veröffentlicht: (2024)
von: Hida, Rem, et al.
Veröffentlicht: (2024)
A Japanese Benchmark for Evaluating Social Bias in Reasoning Based on Attribution Theory
von: Shiotani, Taihei, et al.
Veröffentlicht: (2026)
von: Shiotani, Taihei, et al.
Veröffentlicht: (2026)
Bias Beyond English: Evaluating Social Bias and Debiasing Methods in a Low-Resource Setting
von: Zhou, Ej, et al.
Veröffentlicht: (2025)
von: Zhou, Ej, et al.
Veröffentlicht: (2025)
Likelihood-based Mitigation of Evaluation Bias in Large Language Models
von: Oi, Masanari, et al.
Veröffentlicht: (2024)
von: Oi, Masanari, et al.
Veröffentlicht: (2024)
Inference-Time Selective Debiasing to Enhance Fairness in Text Classification Models
von: Kuzmin, Gleb, et al.
Veröffentlicht: (2024)
von: Kuzmin, Gleb, et al.
Veröffentlicht: (2024)
OffsetBias: Leveraging Debiased Data for Tuning Evaluators
von: Park, Junsoo, et al.
Veröffentlicht: (2024)
von: Park, Junsoo, et al.
Veröffentlicht: (2024)
Applying Intrinsic Debiasing on Downstream Tasks: Challenges and Considerations for Machine Translation
von: Iluz, Bar, et al.
Veröffentlicht: (2024)
von: Iluz, Bar, et al.
Veröffentlicht: (2024)
Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models
von: Mackraz, Natalie, et al.
Veröffentlicht: (2024)
von: Mackraz, Natalie, et al.
Veröffentlicht: (2024)
Connecting the Dots in News Analysis: Bridging the Cross-Disciplinary Disparities in Media Bias and Framing
von: Vallejo, Gisela, et al.
Veröffentlicht: (2023)
von: Vallejo, Gisela, et al.
Veröffentlicht: (2023)
Paraphrasing Adversarial Attack on LLM-as-a-Reviewer
von: Kaneko, Masahiro
Veröffentlicht: (2026)
von: Kaneko, Masahiro
Veröffentlicht: (2026)
Benchmarking Gender and Political Bias in Large Language Models
von: Yang, Jinrui, et al.
Veröffentlicht: (2025)
von: Yang, Jinrui, et al.
Veröffentlicht: (2025)
Language Bias in Information Retrieval: The Nature of the Beast and Mitigation Methods
von: Yang, Jinrui, et al.
Veröffentlicht: (2025)
von: Yang, Jinrui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024) -
Eagle: Ethical Dataset Given from Real Interactions
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024) -
In-Contextual Gender Bias Suppression for Large Language Models
von: Oba, Daisuke, et al.
Veröffentlicht: (2023) -
Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding
von: Oba, Daisuke, et al.
Veröffentlicht: (2026) -
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)