ChatGPT Based Data Augmentation for Improved Parameter-Efficient Debiasing of LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Han, Pengrui, Kocielnik, Rafal, Saravanan, Adhithya, Jiang, Roy, Sharir, Or, Anandkumar, Anima |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Generating Natural-Language Surgical Feedback: From Structured Representation to Domain-Grounded Evaluation
por: Nasriddinov, Firdavs, et al.
Publicado: (2025)
por: Nasriddinov, Firdavs, et al.
Publicado: (2025)
Automating Feedback Analysis in Surgical Training: Detection, Categorization, and Assessment
por: Nasriddinov, Firdavs, et al.
Publicado: (2024)
por: Nasriddinov, Firdavs, et al.
Publicado: (2024)
On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance
por: Casanova, Etienne, et al.
Publicado: (2026)
por: Casanova, Etienne, et al.
Publicado: (2026)
Examining Linguistic Shifts in Academic Writing Before and After the Launch of ChatGPT: A Study on Preprint Papers
por: Bao, Tong, et al.
Publicado: (2025)
por: Bao, Tong, et al.
Publicado: (2025)
An Iterative Optimizing Framework for Radiology Report Summarization with ChatGPT
por: Ma, Chong, et al.
Publicado: (2023)
por: Ma, Chong, et al.
Publicado: (2023)
The AI Fiction Paradox
por: Elkins, Katherine
Publicado: (2026)
por: Elkins, Katherine
Publicado: (2026)
Chatbots put to the test in math and logic problems: A preliminary comparison and assessment of ChatGPT-3.5, ChatGPT-4, and Google Bard
por: Plevris, Vagelis, et al.
Publicado: (2023)
por: Plevris, Vagelis, et al.
Publicado: (2023)
How Few-shot Demonstrations Affect Prompt-based Defenses Against LLM Jailbreak Attacks
por: Wang, Yanshu, et al.
Publicado: (2026)
por: Wang, Yanshu, et al.
Publicado: (2026)
On-Device Generative AI for GDPR-Compliant Visual Monitoring: Natural Language Alerts from Local Object Detection
por: Schappacher-Tilp, Gudrun, et al.
Publicado: (2026)
por: Schappacher-Tilp, Gudrun, et al.
Publicado: (2026)
ArGen: Auto-Regulation of Generative AI via GRPO and Policy-as-Code
por: Madan, Kapil
Publicado: (2025)
por: Madan, Kapil
Publicado: (2025)
The Ethics Engine: A Modular Pipeline for Accessible Psychometric Assessment of Large Language Models
por: Van Clief, Jake, et al.
Publicado: (2025)
por: Van Clief, Jake, et al.
Publicado: (2025)
Self-Anchored Attention Model for Sample-Efficient Classification of Prosocial Text Chat
por: Li, Zhuofang, et al.
Publicado: (2025)
por: Li, Zhuofang, et al.
Publicado: (2025)
Seeing Hate Differently: Hate Subspace Modeling for Culture-Aware Hate Speech Detection
por: Cai, Weibin, et al.
Publicado: (2025)
por: Cai, Weibin, et al.
Publicado: (2025)
AVEC: Bootstrapping Privacy for Local LLMs
por: Gaikwad, Madhava
Publicado: (2025)
por: Gaikwad, Madhava
Publicado: (2025)
EQUITRIAGE: A Fairness Audit of Gender Bias in LLM-Based Emergency Department Triage
por: Young, Richard J., et al.
Publicado: (2026)
por: Young, Richard J., et al.
Publicado: (2026)
On the Validity of Traditional Vulnerability Scoring Systems for Adversarial Attacks against LLMs
por: Bahar, Atmane Ayoub Mansour, et al.
Publicado: (2024)
por: Bahar, Atmane Ayoub Mansour, et al.
Publicado: (2024)
When Retrieval Succeeds and Fails: Rethinking Retrieval-Augmented Generation for LLMs
por: Wang, Yongjie, et al.
Publicado: (2025)
por: Wang, Yongjie, et al.
Publicado: (2025)
When Names Change Verdicts: Intervention Consistency Reveals Systematic Bias in LLM Decision-Making
por: Basu, Abhinaba, et al.
Publicado: (2026)
por: Basu, Abhinaba, et al.
Publicado: (2026)
Prosocial Behavior Detection in Player Game Chat: From Aligning Human-AI Definitions to Efficient Annotation at Scale
por: Kocielnik, Rafal, et al.
Publicado: (2025)
por: Kocielnik, Rafal, et al.
Publicado: (2025)
Multi-Modal Self-Supervised Learning for Surgical Feedback Effectiveness Assessment
por: Gupta, Arushi, et al.
Publicado: (2024)
por: Gupta, Arushi, et al.
Publicado: (2024)
When No Benchmark Exists: Validating Comparative LLM Safety Scoring Without Ground-Truth Labels
por: Gautam, Sushant, et al.
Publicado: (2026)
por: Gautam, Sushant, et al.
Publicado: (2026)
Recipient Profiling: Predicting Characteristics from Messages
por: Borquez, Martin, et al.
Publicado: (2024)
por: Borquez, Martin, et al.
Publicado: (2024)
AI-Powered Citation Auditing: A Zero-Assumption Protocol for Systematic Reference Verification in Academic Research
por: van Rensburg, L. J. Janse
Publicado: (2025)
por: van Rensburg, L. J. Janse
Publicado: (2025)
MELoRA: Mini-Ensemble Low-Rank Adapters for Parameter-Efficient Fine-Tuning
por: Ren, Pengjie, et al.
Publicado: (2024)
por: Ren, Pengjie, et al.
Publicado: (2024)
Beyond Human Judgment: A Bayesian Evaluation of LLMs' Moral Values Understanding
por: Skorski, Maciej, et al.
Publicado: (2025)
por: Skorski, Maciej, et al.
Publicado: (2025)
Estimating Exam Item Difficulty with LLMs: A Benchmark on Brazil's ENEM Corpus
por: Brant, Thiago, et al.
Publicado: (2026)
por: Brant, Thiago, et al.
Publicado: (2026)
IndianBailJudgments-1200: A Multi-Attribute Dataset for Legal NLP on Indian Bail Orders
por: Deshmukh, Sneha, et al.
Publicado: (2025)
por: Deshmukh, Sneha, et al.
Publicado: (2025)
Can LLMs Compute with Reasons?
por: Sandilya, Harshit, et al.
Publicado: (2024)
por: Sandilya, Harshit, et al.
Publicado: (2024)
Formal Proofs as Structured Explanations: Proposing Several Tasks on Explainable Natural Language Inference
por: Abzianidze, Lasha
Publicado: (2023)
por: Abzianidze, Lasha
Publicado: (2023)
Grandes modelos de lenguaje: de la predicción de palabras a la comprensión?
por: Gómez-Rodríguez, Carlos
Publicado: (2025)
por: Gómez-Rodríguez, Carlos
Publicado: (2025)
Who's Asking? Investigating Bias Through the Lens of Disability Framed Queries in LLMs
por: Hari, Vishnu, et al.
Publicado: (2025)
por: Hari, Vishnu, et al.
Publicado: (2025)
LLMs for Legal Subsumption in German Employment Contracts
por: Wardas, Oliver, et al.
Publicado: (2025)
por: Wardas, Oliver, et al.
Publicado: (2025)
Reviewriter: AI-Generated Instructions For Peer Review Writing
por: Su, Xiaotian, et al.
Publicado: (2025)
por: Su, Xiaotian, et al.
Publicado: (2025)
Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs
por: Balter, Samuel G., et al.
Publicado: (2026)
por: Balter, Samuel G., et al.
Publicado: (2026)
The Unlikely Duel: Evaluating Creative Writing in LLMs through a Unique Scenario
por: Gómez-Rodríguez, Carlos, et al.
Publicado: (2024)
por: Gómez-Rodríguez, Carlos, et al.
Publicado: (2024)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
por: Fang, Xi, et al.
Publicado: (2024)
por: Fang, Xi, et al.
Publicado: (2024)
ScoreRAG: A Retrieval-Augmented Generation Framework with Consistency-Relevance Scoring and Structured Summarization for News Generation
por: Lin, Pei-Yun, et al.
Publicado: (2025)
por: Lin, Pei-Yun, et al.
Publicado: (2025)
Democratizing LLMs: An Exploration of Cost-Performance Trade-offs in Self-Refined Open-Source Models
por: Shashidhar, Sumuk, et al.
Publicado: (2023)
por: Shashidhar, Sumuk, et al.
Publicado: (2023)
Distilling Large Language Models for Efficient Clinical Information Extraction
por: Vedula, Karthik S., et al.
Publicado: (2024)
por: Vedula, Karthik S., et al.
Publicado: (2024)
Towards Effective and Efficient Continual Pre-training of Large Language Models
por: Chen, Jie, et al.
Publicado: (2024)
por: Chen, Jie, et al.
Publicado: (2024)
Ejemplares similares
-
Generating Natural-Language Surgical Feedback: From Structured Representation to Domain-Grounded Evaluation
por: Nasriddinov, Firdavs, et al.
Publicado: (2025) -
Automating Feedback Analysis in Surgical Training: Detection, Categorization, and Assessment
por: Nasriddinov, Firdavs, et al.
Publicado: (2024) -
On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance
por: Casanova, Etienne, et al.
Publicado: (2026) -
Examining Linguistic Shifts in Academic Writing Before and After the Launch of ChatGPT: A Study on Preprint Papers
por: Bao, Tong, et al.
Publicado: (2025) -
An Iterative Optimizing Framework for Radiology Report Summarization with ChatGPT
por: Ma, Chong, et al.
Publicado: (2023)