A Note on Bias to Complete
Fuente:
arXiv
Guardado en:
| Autores principales: | Xu, Jia, Diab, Mona |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evaluating Large Language Model Biases in Persona-Steered Generation
por: Liu, Andy, et al.
Publicado: (2024)
por: Liu, Andy, et al.
Publicado: (2024)
Combining Discrete Wavelet and Cosine Transforms for Efficient Sentence Embedding
por: Salama, Rana, et al.
Publicado: (2025)
por: Salama, Rana, et al.
Publicado: (2025)
Semantic Compression for Word and Sentence Embeddings using Discrete Wavelet Transform
por: Salama, Rana Aref, et al.
Publicado: (2025)
por: Salama, Rana Aref, et al.
Publicado: (2025)
Automatic Generation of Model and Data Cards: A Step Towards Responsible AI
por: Liu, Jiarui, et al.
Publicado: (2024)
por: Liu, Jiarui, et al.
Publicado: (2024)
StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection
por: Alqahtani, Amal, et al.
Publicado: (2025)
por: Alqahtani, Amal, et al.
Publicado: (2025)
DWTSumm: Discrete Wavelet Transform for Document Summarization
por: Salama, Rana, et al.
Publicado: (2026)
por: Salama, Rana, et al.
Publicado: (2026)
Biases Propagate in Encoder-based Vision-Language Models: A Systematic Analysis From Intrinsic Measures to Zero-shot Retrieval Outcomes
por: Ghate, Kshitish, et al.
Publicado: (2025)
por: Ghate, Kshitish, et al.
Publicado: (2025)
LLM Microscope: What Model Internals Reveal About Answer Correctness and Context Utilization
por: Liu, Jiarui, et al.
Publicado: (2025)
por: Liu, Jiarui, et al.
Publicado: (2025)
Decoding Dark Matter: Specialized Sparse Autoencoders for Interpreting Rare Concepts in Foundation Models
por: Muhamed, Aashiq, et al.
Publicado: (2024)
por: Muhamed, Aashiq, et al.
Publicado: (2024)
Emotion Classification in Low and Moderate Resource Languages
por: Tafreshi, Shabnam, et al.
Publicado: (2024)
por: Tafreshi, Shabnam, et al.
Publicado: (2024)
SimBA: Simplifying Benchmark Analysis Using Performance Matrices Alone
por: Subramani, Nishant, et al.
Publicado: (2025)
por: Subramani, Nishant, et al.
Publicado: (2025)
Depth-Wise Attention (DWAtt): A Layer Fusion Method for Data-Efficient Classification
por: ElNokrashy, Muhammad, et al.
Publicado: (2022)
por: ElNokrashy, Muhammad, et al.
Publicado: (2022)
Taming Object Hallucinations with Verified Atomic Confidence Estimation
por: Liu, Jiarui, et al.
Publicado: (2025)
por: Liu, Jiarui, et al.
Publicado: (2025)
CoRAG: Collaborative Retrieval-Augmented Generation
por: Muhamed, Aashiq, et al.
Publicado: (2025)
por: Muhamed, Aashiq, et al.
Publicado: (2025)
Personal Information Parroting in Language Models
por: Subramani, Nishant, et al.
Publicado: (2026)
por: Subramani, Nishant, et al.
Publicado: (2026)
Beyond Understanding: Evaluating the Pragmatic Gap in LLMs' Cultural Processing of Figurative Language
por: Attia, Mena, et al.
Publicado: (2025)
por: Attia, Mena, et al.
Publicado: (2025)
Humanizing Machines: Rethinking LLM Anthropomorphism Through a Multi-Level Framework of Design
por: Xiao, Yunze, et al.
Publicado: (2025)
por: Xiao, Yunze, et al.
Publicado: (2025)
Hire Your Anthropologist! Rethinking Culture Benchmarks Through an Anthropological Lens
por: AlKhamissi, Mai, et al.
Publicado: (2025)
por: AlKhamissi, Mai, et al.
Publicado: (2025)
DSPA: Dynamic SAE Steering for Data-Efficient Preference Alignment
por: Wedgwood, James, et al.
Publicado: (2026)
por: Wedgwood, James, et al.
Publicado: (2026)
Investigating Cultural Alignment of Large Language Models
por: AlKhamissi, Badr, et al.
Publicado: (2024)
por: AlKhamissi, Badr, et al.
Publicado: (2024)
BIG5-CHAT: Shaping LLM Personalities Through Training on Human-Grounded Data
por: Li, Wenkai, et al.
Publicado: (2024)
por: Li, Wenkai, et al.
Publicado: (2024)
SAEs $\textit{Can}$ Improve Unlearning: Dynamic Sparse Autoencoder Guardrails for Precision Unlearning in LLMs
por: Muhamed, Aashiq, et al.
Publicado: (2025)
por: Muhamed, Aashiq, et al.
Publicado: (2025)
Sentipolis: Emotion-Aware Agents for Social Simulations
por: Fu, Chiyuan, et al.
Publicado: (2026)
por: Fu, Chiyuan, et al.
Publicado: (2026)
Towards Global AI Inclusivity: A Large-Scale Multilingual Terminology Dataset (GIST)
por: Liu, Jiarui, et al.
Publicado: (2024)
por: Liu, Jiarui, et al.
Publicado: (2024)
Towards Valid Student Simulation with Large Language Models
por: Yuan, Zhihao, et al.
Publicado: (2026)
por: Yuan, Zhihao, et al.
Publicado: (2026)
JiraiBench: A Bilingual Benchmark for Evaluating Large Language Models' Detection of Human Self-Destructive Behavior Content in Jirai Community
por: Xiao, Yunze, et al.
Publicado: (2025)
por: Xiao, Yunze, et al.
Publicado: (2025)
MixSD: Mixed Contextual Self-Distillation for Knowledge Injection
por: Liu, Jiarui, et al.
Publicado: (2026)
por: Liu, Jiarui, et al.
Publicado: (2026)
EVALUESTEER: Measuring Reward Model Steerability Towards Values and Preferences
por: Ghate, Kshitish, et al.
Publicado: (2025)
por: Ghate, Kshitish, et al.
Publicado: (2025)
RefusalBench: Generative Evaluation of Selective Refusal in Grounded Language Models
por: Muhamed, Aashiq, et al.
Publicado: (2025)
por: Muhamed, Aashiq, et al.
Publicado: (2025)
Synthetic Socratic Debates: Examining Persona Effects on Moral Decision and Persuasion Dynamics
por: Liu, Jiarui, et al.
Publicado: (2025)
por: Liu, Jiarui, et al.
Publicado: (2025)
Generative Value Conflicts Reveal LLM Priorities
por: Liu, Andy, et al.
Publicado: (2025)
por: Liu, Andy, et al.
Publicado: (2025)
The FIGNEWS Shared Task on News Media Narratives
por: Zaghouani, Wajdi, et al.
Publicado: (2024)
por: Zaghouani, Wajdi, et al.
Publicado: (2024)
BiasAlert: A Plug-and-play Tool for Social Bias Detection in LLMs
por: Fan, Zhiting, et al.
Publicado: (2024)
por: Fan, Zhiting, et al.
Publicado: (2024)
NoteChat: A Dataset of Synthetic Doctor-Patient Conversations Conditioned on Clinical Notes
por: Wang, Junda, et al.
Publicado: (2023)
por: Wang, Junda, et al.
Publicado: (2023)
DeepNote: Note-Centric Deep Retrieval-Augmented Generation
por: Wang, Ruobing, et al.
Publicado: (2024)
por: Wang, Ruobing, et al.
Publicado: (2024)
Analyzing the Role of Semantic Representations in the Era of Large Language Models
por: Jin, Zhijing, et al.
Publicado: (2024)
por: Jin, Zhijing, et al.
Publicado: (2024)
Can Large Language Models Infer Causation from Correlation?
por: Jin, Zhijing, et al.
Publicado: (2023)
por: Jin, Zhijing, et al.
Publicado: (2023)
Position: Mechanistic Interpretability Should Prioritize Feature Consistency in SAEs
por: Song, Xiangchen, et al.
Publicado: (2025)
por: Song, Xiangchen, et al.
Publicado: (2025)
ICA-RAG: Information Completeness Guided Adaptive Retrieval-Augmented Generation for Disease Diagnosis
por: He, Jiawei, et al.
Publicado: (2025)
por: He, Jiawei, et al.
Publicado: (2025)
The Point of View of a Sentiment: Towards Clinician Bias Detection in Psychiatric Notes
por: Valentine, Alissa A., et al.
Publicado: (2024)
por: Valentine, Alissa A., et al.
Publicado: (2024)
Ejemplares similares
-
Evaluating Large Language Model Biases in Persona-Steered Generation
por: Liu, Andy, et al.
Publicado: (2024) -
Combining Discrete Wavelet and Cosine Transforms for Efficient Sentence Embedding
por: Salama, Rana, et al.
Publicado: (2025) -
Semantic Compression for Word and Sentence Embeddings using Discrete Wavelet Transform
por: Salama, Rana Aref, et al.
Publicado: (2025) -
Automatic Generation of Model and Data Cards: A Step Towards Responsible AI
por: Liu, Jiarui, et al.
Publicado: (2024) -
StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection
por: Alqahtani, Amal, et al.
Publicado: (2025)