Salvato in:
| Autori principali: | Wang, Qian, Zhao, Xuandong, Zhang, Zirui, Lou, Zhanzhi, Chen, Nuo, Song, Dawn, He, Bingsheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.01528 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Assessing Judging Bias in Large Reasoning Models: An Empirical Study
di: Wang, Qian, et al.
Pubblicazione: (2025)
di: Wang, Qian, et al.
Pubblicazione: (2025)
Towards Evaluting Fake Reasoning Bias in Language Models
di: Wang, Qian, et al.
Pubblicazione: (2025)
di: Wang, Qian, et al.
Pubblicazione: (2025)
Self-Sovereign Agent
di: Qu, Wenjie, et al.
Pubblicazione: (2026)
di: Qu, Wenjie, et al.
Pubblicazione: (2026)
Learning to Reason without External Rewards
di: Zhao, Xuandong, et al.
Pubblicazione: (2025)
di: Zhao, Xuandong, et al.
Pubblicazione: (2025)
Scalable Best-of-N Selection for Large Language Models via Self-Certainty
di: Kang, Zhewei, et al.
Pubblicazione: (2025)
di: Kang, Zhewei, et al.
Pubblicazione: (2025)
Are You Getting What You Pay For? Auditing Model Substitution in LLM APIs
di: Cai, Will, et al.
Pubblicazione: (2025)
di: Cai, Will, et al.
Pubblicazione: (2025)
MemFail: Stress-Testing Failure Modes of LLM Memory Systems
di: Garg, Ishir, et al.
Pubblicazione: (2026)
di: Garg, Ishir, et al.
Pubblicazione: (2026)
Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies
di: Lou, Zhanzhi, et al.
Pubblicazione: (2026)
di: Lou, Zhanzhi, et al.
Pubblicazione: (2026)
"They've Stolen My GPL-Licensed Model!": Toward Standardized and Transparent Model Licensing
di: Duan, Moming, et al.
Pubblicazione: (2024)
di: Duan, Moming, et al.
Pubblicazione: (2024)
Integrating Reason-Based Moral Decision-Making in the Reinforcement Learning Architecture
di: Dargasz, Lisa
Pubblicazione: (2025)
di: Dargasz, Lisa
Pubblicazione: (2025)
An Undetectable Watermark for Generative Image Models
di: Gunn, Sam, et al.
Pubblicazione: (2024)
di: Gunn, Sam, et al.
Pubblicazione: (2024)
Position: The Current AI Conference Model is Unsustainable! Diagnosing the Crisis of Centralized AI Conference
di: Chen, Nuo, et al.
Pubblicazione: (2025)
di: Chen, Nuo, et al.
Pubblicazione: (2025)
Training Fair Models in Federated Learning without Data Privacy Infringement
di: Che, Xin, et al.
Pubblicazione: (2021)
di: Che, Xin, et al.
Pubblicazione: (2021)
The Hidden Risks of Large Reasoning Models: A Safety Assessment of R1
di: Zhou, Kaiwen, et al.
Pubblicazione: (2025)
di: Zhou, Kaiwen, et al.
Pubblicazione: (2025)
From ChatGPT to DeepSeek: Can LLMs Simulate Humanity?
di: Wang, Qian, et al.
Pubblicazione: (2025)
di: Wang, Qian, et al.
Pubblicazione: (2025)
Improving LLM Safety Alignment with Dual-Objective Optimization
di: Zhao, Xuandong, et al.
Pubblicazione: (2025)
di: Zhao, Xuandong, et al.
Pubblicazione: (2025)
Improving Adversarial Robust Fairness via Anti-Bias Soft Label Distillation
di: Zhao, Shiji, et al.
Pubblicazione: (2023)
di: Zhao, Shiji, et al.
Pubblicazione: (2023)
The Landscape of Memorization in LLMs: Mechanisms, Measurement, and Mitigation
di: Xiong, Alexander, et al.
Pubblicazione: (2025)
di: Xiong, Alexander, et al.
Pubblicazione: (2025)
DCAST: Diverse Class-Aware Self-Training Mitigates Selection Bias for Fairer Learning
di: Tepeli, Yasin I., et al.
Pubblicazione: (2024)
di: Tepeli, Yasin I., et al.
Pubblicazione: (2024)
Machine Learning-Driven Student Performance Prediction for Enhancing Tiered Instruction
di: Chen, Yawen, et al.
Pubblicazione: (2025)
di: Chen, Yawen, et al.
Pubblicazione: (2025)
Adapting Static Fairness to Sequential Decision-Making: Bias Mitigation Strategies towards Equal Long-term Benefit Rate
di: Xu, Yuancheng, et al.
Pubblicazione: (2023)
di: Xu, Yuancheng, et al.
Pubblicazione: (2023)
Beyond Brainstorming: What Drives High-Quality Scientific Ideas? Lessons from Multi-Agent Collaboration
di: Chen, Nuo, et al.
Pubblicazione: (2025)
di: Chen, Nuo, et al.
Pubblicazione: (2025)
Integrating Social Determinants of Health into Knowledge Graphs: Evaluating Prediction Bias and Fairness in Healthcare
di: Shang, Tianqi, et al.
Pubblicazione: (2024)
di: Shang, Tianqi, et al.
Pubblicazione: (2024)
Towards Responsible AI in Banking: Addressing Bias for Fair Decision-Making
di: Castelnovo, Alessandro
Pubblicazione: (2024)
di: Castelnovo, Alessandro
Pubblicazione: (2024)
Exploring LLM Cryptocurrency Trading Through Fact-Subjectivity Aware Reasoning
di: Wang, Qian, et al.
Pubblicazione: (2024)
di: Wang, Qian, et al.
Pubblicazione: (2024)
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code
di: Bao, Keqin, et al.
Pubblicazione: (2025)
di: Bao, Keqin, et al.
Pubblicazione: (2025)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
di: Xu, Xin, et al.
Pubblicazione: (2025)
di: Xu, Xin, et al.
Pubblicazione: (2025)
Robust Meta-Model for Predicting the Need for Blood Transfusion in Non-traumatic ICU Patients
di: Rafiei, Alireza, et al.
Pubblicazione: (2024)
di: Rafiei, Alireza, et al.
Pubblicazione: (2024)
LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage
di: Nie, Yuzhou, et al.
Pubblicazione: (2024)
di: Nie, Yuzhou, et al.
Pubblicazione: (2024)
Backdoor for Debias: Mitigating Model Bias with Backdoor Attack-based Artificial Bias
di: Wu, Shangxi, et al.
Pubblicazione: (2023)
di: Wu, Shangxi, et al.
Pubblicazione: (2023)
LLM DNA: Tracing Model Evolution via Functional Representations
di: Wu, Zhaomin, et al.
Pubblicazione: (2025)
di: Wu, Zhaomin, et al.
Pubblicazione: (2025)
Bias Begins with Data: The FairGround Corpus for Robust and Reproducible Research on Algorithmic Fairness
di: Simson, Jan, et al.
Pubblicazione: (2025)
di: Simson, Jan, et al.
Pubblicazione: (2025)
Predicting Human Mobility during Extreme Events via LLM-Enhanced Cross-City Learning
di: Tang, Yinzhou, et al.
Pubblicazione: (2025)
di: Tang, Yinzhou, et al.
Pubblicazione: (2025)
Reinforced Sequential Decision-Making for Sepsis Treatment: The POSNEGDM Framework with Mortality Classifier and Transformer
di: Tamboli, Dipesh, et al.
Pubblicazione: (2024)
di: Tamboli, Dipesh, et al.
Pubblicazione: (2024)
Addressing Discretization-Induced Bias in Demographic Prediction
di: Dong, Evan, et al.
Pubblicazione: (2024)
di: Dong, Evan, et al.
Pubblicazione: (2024)
BadFair: Backdoored Fairness Attacks with Group-conditioned Triggers
di: Xue, Jiaqi, et al.
Pubblicazione: (2024)
di: Xue, Jiaqi, et al.
Pubblicazione: (2024)
The Relative Value of Prediction in Algorithmic Decision Making
di: Perdomo, Juan Carlos
Pubblicazione: (2023)
di: Perdomo, Juan Carlos
Pubblicazione: (2023)
Mitigating Gender Bias in Depression Detection via Counterfactual Inference
di: Hu, Mingxuan, et al.
Pubblicazione: (2025)
di: Hu, Mingxuan, et al.
Pubblicazione: (2025)
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
di: Alamdari, Parand A., et al.
Pubblicazione: (2023)
di: Alamdari, Parand A., et al.
Pubblicazione: (2023)
Counterfactually Fair Reinforcement Learning via Sequential Data Preprocessing
di: Wang, Jitao, et al.
Pubblicazione: (2025)
di: Wang, Jitao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Assessing Judging Bias in Large Reasoning Models: An Empirical Study
di: Wang, Qian, et al.
Pubblicazione: (2025) -
Towards Evaluting Fake Reasoning Bias in Language Models
di: Wang, Qian, et al.
Pubblicazione: (2025) -
Self-Sovereign Agent
di: Qu, Wenjie, et al.
Pubblicazione: (2026) -
Learning to Reason without External Rewards
di: Zhao, Xuandong, et al.
Pubblicazione: (2025) -
Scalable Best-of-N Selection for Large Language Models via Self-Certainty
di: Kang, Zhewei, et al.
Pubblicazione: (2025)