Salvato in:
| Autori principali: | Liu, Tianci, Wang, Haoyu, Wang, Shiyang, Cheng, Yu, Gao, Jing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2406.00548 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Universal Debiasing for Language Models-based Tabular Data Generation
di: Li, Tianchun, et al.
Pubblicazione: (2025)
di: Li, Tianchun, et al.
Pubblicazione: (2025)
RUBRIC-ARROW: Alternating Pointwise Rubric Reward Modeling for LLM Post-training in Non-verifiable Domains
di: Jiang, Haoxiang, et al.
Pubblicazione: (2026)
di: Jiang, Haoxiang, et al.
Pubblicazione: (2026)
Mitigating Heterogeneous Token Overfitting in LLM Knowledge Editing
di: Liu, Tianci, et al.
Pubblicazione: (2025)
di: Liu, Tianci, et al.
Pubblicazione: (2025)
DBR: Divergence-Based Regularization for Debiasing Natural Language Understanding Models
di: Li, Zihao, et al.
Pubblicazione: (2025)
di: Li, Zihao, et al.
Pubblicazione: (2025)
Unlocking Efficient, Scalable, and Continual Knowledge Editing with Basis-Level Representation Fine-Tuning
di: Liu, Tianci, et al.
Pubblicazione: (2025)
di: Liu, Tianci, et al.
Pubblicazione: (2025)
RoseRAG: Robust Retrieval-augmented Generation with Small-scale LLMs via Margin-aware Preference Optimization
di: Liu, Tianci, et al.
Pubblicazione: (2025)
di: Liu, Tianci, et al.
Pubblicazione: (2025)
Alternating Reinforcement Learning for Rubric-Based Reward Modeling in Non-Verifiable LLM Post-Training
di: Xu, Ran, et al.
Pubblicazione: (2026)
di: Xu, Ran, et al.
Pubblicazione: (2026)
Towards Federated RLHF with Aggregated Client Preference for LLMs
di: Wu, Feijie, et al.
Pubblicazione: (2024)
di: Wu, Feijie, et al.
Pubblicazione: (2024)
Efficient Temporal Tokenization for Mobility Prediction with Large Language Models
di: He, Haoyu, et al.
Pubblicazione: (2025)
di: He, Haoyu, et al.
Pubblicazione: (2025)
Self-Supervised Position Debiasing for Large Language Models
di: Liu, Zhongkun, et al.
Pubblicazione: (2024)
di: Liu, Zhongkun, et al.
Pubblicazione: (2024)
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
di: Gallegos, Isabel O., et al.
Pubblicazione: (2024)
di: Gallegos, Isabel O., et al.
Pubblicazione: (2024)
When the Same Coefficients Reach Different Places: Asymmetric Realizability in Transplanting Tokenizers across Large Language Models
di: Liu, Xiaoze, et al.
Pubblicazione: (2025)
di: Liu, Xiaoze, et al.
Pubblicazione: (2025)
PEANuT: Parameter-Efficient Adaptation with Weight-aware Neural Tweakers
di: Zhong, Yibo, et al.
Pubblicazione: (2024)
di: Zhong, Yibo, et al.
Pubblicazione: (2024)
Debiasing Watermarks for Large Language Models via Maximal Coupling
di: Xie, Yangxinyu, et al.
Pubblicazione: (2024)
di: Xie, Yangxinyu, et al.
Pubblicazione: (2024)
Composable Interventions for Language Models
di: Kolbeinsson, Arinbjorn, et al.
Pubblicazione: (2024)
di: Kolbeinsson, Arinbjorn, et al.
Pubblicazione: (2024)
Debiasing Large Vision-Language Models by Ablating Protected Attribute Representations
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
AXOLOTL: Fairness through Assisted Self-Debiasing of Large Language Model Outputs
di: Ebrahimi, Sana, et al.
Pubblicazione: (2024)
di: Ebrahimi, Sana, et al.
Pubblicazione: (2024)
SUV: Scalable Large Language Model Copyright Compliance with Regularized Selective Unlearning
di: Xu, Tianyang, et al.
Pubblicazione: (2025)
di: Xu, Tianyang, et al.
Pubblicazione: (2025)
Evaluating the Factuality of Large Language Models using Large-Scale Knowledge Graphs
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
Backdooring Instruction-Tuned Large Language Models with Virtual Prompt Injection
di: Yan, Jun, et al.
Pubblicazione: (2023)
di: Yan, Jun, et al.
Pubblicazione: (2023)
Large Language Model Selection with Limited Annotations
di: Durmazkeser, Yavuz, et al.
Pubblicazione: (2026)
di: Durmazkeser, Yavuz, et al.
Pubblicazione: (2026)
Model Hemorrhage and the Robustness Limits of Large Language Models
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
API Is Enough: Conformal Prediction for Large Language Models Without Logit-Access
di: Su, Jiayuan, et al.
Pubblicazione: (2024)
di: Su, Jiayuan, et al.
Pubblicazione: (2024)
RoseLoRA: Row and Column-wise Sparse Low-rank Adaptation of Pre-trained Language Model for Knowledge Editing and Fine-tuning
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
DrugAgent: Multi-Agent Large Language Model-Based Reasoning for Drug-Target Interaction Prediction
di: Inoue, Yoshitaka, et al.
Pubblicazione: (2024)
di: Inoue, Yoshitaka, et al.
Pubblicazione: (2024)
Dynamic Attention-Guided Context Decoding for Mitigating Context Faithfulness Hallucinations in Large Language Models
di: Huang, Yanwen, et al.
Pubblicazione: (2025)
di: Huang, Yanwen, et al.
Pubblicazione: (2025)
On the Thinking-Language Modeling Gap in Large Language Models
di: Liu, Chenxi, et al.
Pubblicazione: (2025)
di: Liu, Chenxi, et al.
Pubblicazione: (2025)
The Flexibility Trap: Why Arbitrary Order Limits Reasoning Potential in Diffusion Language Models
di: Ni, Zanlin, et al.
Pubblicazione: (2026)
di: Ni, Zanlin, et al.
Pubblicazione: (2026)
Beyond the Limits: A Survey of Techniques to Extend the Context Length in Large Language Models
di: Wang, Xindi, et al.
Pubblicazione: (2024)
di: Wang, Xindi, et al.
Pubblicazione: (2024)
The Role of Diversity in In-Context Learning for Large Language Models
di: Xiao, Wenyang, et al.
Pubblicazione: (2025)
di: Xiao, Wenyang, et al.
Pubblicazione: (2025)
DoTA: Weight-Decomposed Tensor Adaptation for Large Language Models
di: Hu, Xiaolin, et al.
Pubblicazione: (2024)
di: Hu, Xiaolin, et al.
Pubblicazione: (2024)
Recurrent Drafter for Fast Speculative Decoding in Large Language Models
di: Cheng, Yunfei, et al.
Pubblicazione: (2024)
di: Cheng, Yunfei, et al.
Pubblicazione: (2024)
Bias Is a Subspace, Not a Coordinate: A Geometric Rethinking of Post-hoc Debiasing in Vision-Language Models
di: Zhao, Dachuan, et al.
Pubblicazione: (2025)
di: Zhao, Dachuan, et al.
Pubblicazione: (2025)
Non-linear Interventions on Large Language Models
di: Kim, Sangwoo
Pubblicazione: (2026)
di: Kim, Sangwoo
Pubblicazione: (2026)
Digger: Detecting Copyright Content Mis-usage in Large Language Model Training
di: Li, Haodong, et al.
Pubblicazione: (2024)
di: Li, Haodong, et al.
Pubblicazione: (2024)
Erasing Without Remembering: Implicit Knowledge Forgetting in Large Language Models
di: Wang, Huazheng, et al.
Pubblicazione: (2025)
di: Wang, Huazheng, et al.
Pubblicazione: (2025)
Towards Reasoning-Preserving Unlearning in Multimodal Large Language Models
di: Li, Hongji, et al.
Pubblicazione: (2025)
di: Li, Hongji, et al.
Pubblicazione: (2025)
Vulnerability Mitigation for Safety-Aligned Language Models via Debiasing
di: Tran, Thien Q., et al.
Pubblicazione: (2025)
di: Tran, Thien Q., et al.
Pubblicazione: (2025)
A Survey on Mixture of Experts in Large Language Models
di: Cai, Weilin, et al.
Pubblicazione: (2024)
di: Cai, Weilin, et al.
Pubblicazione: (2024)
How Can Large Language Models Understand Spatial-Temporal Data?
di: Liu, Lei, et al.
Pubblicazione: (2024)
di: Liu, Lei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Towards Universal Debiasing for Language Models-based Tabular Data Generation
di: Li, Tianchun, et al.
Pubblicazione: (2025) -
RUBRIC-ARROW: Alternating Pointwise Rubric Reward Modeling for LLM Post-training in Non-verifiable Domains
di: Jiang, Haoxiang, et al.
Pubblicazione: (2026) -
Mitigating Heterogeneous Token Overfitting in LLM Knowledge Editing
di: Liu, Tianci, et al.
Pubblicazione: (2025) -
DBR: Divergence-Based Regularization for Debiasing Natural Language Understanding Models
di: Li, Zihao, et al.
Pubblicazione: (2025) -
Unlocking Efficient, Scalable, and Continual Knowledge Editing with Basis-Level Representation Fine-Tuning
di: Liu, Tianci, et al.
Pubblicazione: (2025)