Investigating the Effects of Fairness Interventions Using Pointwise Representational Similarity
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Kolling, Camila, Speicher, Till, Nanda, Vedant, Toneva, Mariya, Gummadi, Krishna P. |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Understanding the Role of Invariance in Transfer Learning
par: Speicher, Till, et autres
Publié: (2024)
par: Speicher, Till, et autres
Publié: (2024)
Towards Reliable Latent Knowledge Estimation in LLMs: Zero-Prompt Many-Shot Based Factual Knowledge Extraction
par: Wu, Qinyuan, et autres
Publié: (2024)
par: Wu, Qinyuan, et autres
Publié: (2024)
Understanding Memorisation in LLMs: Dynamics, Influencing Factors, and Implications
par: Speicher, Till, et autres
Publié: (2024)
par: Speicher, Till, et autres
Publié: (2024)
Large Language Models as Model Organisms for Human Associative Learning
par: Kolling, Camila, et autres
Publié: (2025)
par: Kolling, Camila, et autres
Publié: (2025)
Revisiting Privacy, Utility, and Efficiency Trade-offs when Fine-Tuning Large Language Models
par: Das, Soumi, et autres
Publié: (2025)
par: Das, Soumi, et autres
Publié: (2025)
Perturbed examples reveal invariances shared by language models
par: Rawal, Ruchit, et autres
Publié: (2023)
par: Rawal, Ruchit, et autres
Publié: (2023)
Tracking Equivalent Mechanistic Interpretations Across Neural Networks
par: Sun, Alan, et autres
Publié: (2026)
par: Sun, Alan, et autres
Publié: (2026)
Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective
par: Ghosh, Bishwamittra, et autres
Publié: (2026)
par: Ghosh, Bishwamittra, et autres
Publié: (2026)
Fractional Rotation, Full Potential? Investigating Performance and Convergence of Partial RoPE
par: Khan, Mohammad Aflah, et autres
Publié: (2026)
par: Khan, Mohammad Aflah, et autres
Publié: (2026)
Lawma: The Power of Specialization for Legal Annotation
par: Dominguez-Olmedo, Ricardo, et autres
Publié: (2024)
par: Dominguez-Olmedo, Ricardo, et autres
Publié: (2024)
From Set Convergence to Pointwise Convergence: Finite-Time Guarantees for Average-Reward Q-Learning with Adaptive Stepsizes
par: Chen, Zaiwei, et autres
Publié: (2025)
par: Chen, Zaiwei, et autres
Publié: (2025)
Speech language models lack important brain-relevant semantics
par: Oota, Subba Reddy, et autres
Publié: (2023)
par: Oota, Subba Reddy, et autres
Publié: (2023)
Investigating Data Interventions for Subgroup Fairness: An ICU Case Study
par: Tan, Erin, et autres
Publié: (2026)
par: Tan, Erin, et autres
Publié: (2026)
Improving LLM Final Representations with Inter-Layer Geometry
par: Ulanovski, Tom, et autres
Publié: (2026)
par: Ulanovski, Tom, et autres
Publié: (2026)
LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging
par: Lee, Seungeon, et autres
Publié: (2025)
par: Lee, Seungeon, et autres
Publié: (2025)
FairExpand: Individual Fairness on Graphs with Partial Similarity Information
par: Salganik, Rebecca, et autres
Publié: (2025)
par: Salganik, Rebecca, et autres
Publié: (2025)
A General Recipe for Contractive Graph Neural Networks -- Technical Report
par: Bechler-Speicher, Maya, et autres
Publié: (2024)
par: Bechler-Speicher, Maya, et autres
Publié: (2024)
General Causal Imputation via Synthetic Interventions
par: Jiralerspong, Marco, et autres
Publié: (2024)
par: Jiralerspong, Marco, et autres
Publié: (2024)
Perturbation Effects on Accuracy and Fairness among Similar Individuals
par: Li, Xuran, et autres
Publié: (2024)
par: Li, Xuran, et autres
Publié: (2024)
The Impact of Inference Acceleration on Bias of LLMs
par: Kirsten, Elisabeth, et autres
Publié: (2024)
par: Kirsten, Elisabeth, et autres
Publié: (2024)
Rethinking Memorization Measures and their Implications in Large Language Models
par: Ghosh, Bishwamittra, et autres
Publié: (2025)
par: Ghosh, Bishwamittra, et autres
Publié: (2025)
CLIP is All You Need for Human-like Semantic Representations in Stable Diffusion
par: Braunstein, Cameron, et autres
Publié: (2025)
par: Braunstein, Cameron, et autres
Publié: (2025)
Fair Classification by Direct Intervention on Operating Characteristics
par: Jiang, Kevin, et autres
Publié: (2025)
par: Jiang, Kevin, et autres
Publié: (2025)
Entrywise application of non-linear functions on orthogonally invariant matrices
par: Speicher, Roland, et autres
Publié: (2024)
par: Speicher, Roland, et autres
Publié: (2024)
Reasoning-Finetuning Repurposes Latent Representations in Base Models
par: Ward, Jake, et autres
Publié: (2025)
par: Ward, Jake, et autres
Publié: (2025)
Multimodal Visual-Tactile Representation Learning through Self-Supervised Contrastive Pre-Training
par: Dave, Vedant, et autres
Publié: (2024)
par: Dave, Vedant, et autres
Publié: (2024)
Brain-tuned Speech Models Better Reflect Speech Processing Stages in the Brain
par: Moussa, Omer, et autres
Publié: (2025)
par: Moussa, Omer, et autres
Publié: (2025)
Language models and brains align due to more than next-word prediction and word-level information
par: Merlin, Gabriele, et autres
Publié: (2022)
par: Merlin, Gabriele, et autres
Publié: (2022)
Brain-tuning Improves Generalizability and Efficiency of Brain Alignment in Speech Models
par: Moussa, Omer, et autres
Publié: (2025)
par: Moussa, Omer, et autres
Publié: (2025)
When Language Models Lose Their Mind: The Consequences of Brain Misalignment
par: Merlin, Gabriele, et autres
Publié: (2026)
par: Merlin, Gabriele, et autres
Publié: (2026)
Towards Poisoning Fair Representations
par: Liu, Tianci, et autres
Publié: (2023)
par: Liu, Tianci, et autres
Publié: (2023)
Adaptive Federated Learning Defences via Trust-Aware Deep Q-Networks
par: Palit, Vedant
Publié: (2025)
par: Palit, Vedant
Publié: (2025)
Enhancing Model Fairness and Accuracy with Similarity Networks: A Methodological Approach
par: Maghool, Samira, et autres
Publié: (2024)
par: Maghool, Samira, et autres
Publié: (2024)
Multiscale Dubuc: A New Similarity Measure for Time Series
par: Khazaei, Mahsa, et autres
Publié: (2024)
par: Khazaei, Mahsa, et autres
Publié: (2024)
On the Pointwise Behavior of Recursive Partitioning and Its Implications for Heterogeneous Causal Effect Estimation
par: Cattaneo, Matias D., et autres
Publié: (2022)
par: Cattaneo, Matias D., et autres
Publié: (2022)
Kernel Representation and Similarity Measure for Incomplete Data
par: Cao, Yang, et autres
Publié: (2025)
par: Cao, Yang, et autres
Publié: (2025)
Interventional Causal Representation Learning
par: Ahuja, Kartik, et autres
Publié: (2022)
par: Ahuja, Kartik, et autres
Publié: (2022)
Brain-Computer Interfaces for Emotional Regulation in Patients with Various Disorders
par: Mehta, Vedant
Publié: (2024)
par: Mehta, Vedant
Publié: (2024)
On the Power of Randomization in Fair Classification and Representation
par: Agarwal, Sushant, et autres
Publié: (2024)
par: Agarwal, Sushant, et autres
Publié: (2024)
Is It Still Fair? Investigating Gender Fairness in Cross-Corpus Speech Emotion Recognition
par: Upadhyay, Shreya G., et autres
Publié: (2025)
par: Upadhyay, Shreya G., et autres
Publié: (2025)
Documents similaires
-
Understanding the Role of Invariance in Transfer Learning
par: Speicher, Till, et autres
Publié: (2024) -
Towards Reliable Latent Knowledge Estimation in LLMs: Zero-Prompt Many-Shot Based Factual Knowledge Extraction
par: Wu, Qinyuan, et autres
Publié: (2024) -
Understanding Memorisation in LLMs: Dynamics, Influencing Factors, and Implications
par: Speicher, Till, et autres
Publié: (2024) -
Large Language Models as Model Organisms for Human Associative Learning
par: Kolling, Camila, et autres
Publié: (2025) -
Revisiting Privacy, Utility, and Efficiency Trade-offs when Fine-Tuning Large Language Models
par: Das, Soumi, et autres
Publié: (2025)