Salvato in:
| Autori principali: | Yang, Haoming, Ma, Ke, Jia, Xiaojun, Sun, Yingfei, Xu, Qianqian, Huang, Qingming |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2505.02862 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Sequential Manipulation Against Rank Aggregation: Theory and Algorithm
di: Ma, Ke, et al.
Pubblicazione: (2024)
di: Ma, Ke, et al.
Pubblicazione: (2024)
Divide and Conquer: Heterogeneous Noise Integration for Diffusion-based Adversarial Purification
di: Pei, Gaozheng, et al.
Pubblicazione: (2025)
di: Pei, Gaozheng, et al.
Pubblicazione: (2025)
Diffusion-based Adversarial Purification from the Perspective of the Frequency Domain
di: Pei, Gaozheng, et al.
Pubblicazione: (2025)
di: Pei, Gaozheng, et al.
Pubblicazione: (2025)
Can't See the Forest for the Trees: Benchmarking Multimodal Safety Awareness for Multimodal LLMs
di: Wang, Wenxuan, et al.
Pubblicazione: (2025)
di: Wang, Wenxuan, et al.
Pubblicazione: (2025)
Systematic Analysis of LLM Contributions to Planning: Solver, Verifier, Heuristic
di: Li, Haoming, et al.
Pubblicazione: (2024)
di: Li, Haoming, et al.
Pubblicazione: (2024)
Logic-Regularized Verifier Elicits Reasoning from LLMs
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
Computer Environments Elicit General Agentic Intelligence in LLMs
di: Cheng, Daixuan, et al.
Pubblicazione: (2026)
di: Cheng, Daixuan, et al.
Pubblicazione: (2026)
Heuristics and Biases in AI Decision-Making: Implications for Responsible AGI
di: Saeedi, Payam, et al.
Pubblicazione: (2024)
di: Saeedi, Payam, et al.
Pubblicazione: (2024)
Balancing Rigor and Utility: Mitigating Cognitive Biases in Large Language Models for Multiple-Choice Questions
di: Zhong, Hanyang, et al.
Pubblicazione: (2024)
di: Zhong, Hanyang, et al.
Pubblicazione: (2024)
To See is Not to Master: Teaching LLMs to Use Private Libraries for Code Generation
di: Zhang, Yitong, et al.
Pubblicazione: (2026)
di: Zhang, Yitong, et al.
Pubblicazione: (2026)
Exploring Query Efficient Data Generation towards Data-free Model Stealing in Hard Label Setting
di: Pei, Gaozheng, et al.
Pubblicazione: (2024)
di: Pei, Gaozheng, et al.
Pubblicazione: (2024)
Empathy and the Right to Be an Exception: What LLMs Can and Cannot Do
di: Kidder, William, et al.
Pubblicazione: (2024)
di: Kidder, William, et al.
Pubblicazione: (2024)
See What LLMs Cannot Answer: A Self-Challenge Framework for Uncovering LLM Weaknesses
di: Chen, Yulong, et al.
Pubblicazione: (2024)
di: Chen, Yulong, et al.
Pubblicazione: (2024)
BiasGym: A Simple and Generalizable Framework for Analyzing and Removing Biases through Elicitation
di: Islam, Sekh Mainul, et al.
Pubblicazione: (2025)
di: Islam, Sekh Mainul, et al.
Pubblicazione: (2025)
What External Knowledge is Preferred by LLMs? Characterizing and Exploring Chain of Evidence in Imperfect Context for Multi-Hop QA
di: Chang, Zhiyuan, et al.
Pubblicazione: (2024)
di: Chang, Zhiyuan, et al.
Pubblicazione: (2024)
From Biased Chatbots to Biased Agents: Examining Role Assignment Effects on LLM Agent Robustness
di: Cao, Linbo, et al.
Pubblicazione: (2026)
di: Cao, Linbo, et al.
Pubblicazione: (2026)
Semantic Anchors in In-Context Learning: Why Small LLMs Cannot Flip Their Labels
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
Can LLMs Evaluate What They Cannot Annotate? Revisiting LLM Reliability in Hate Speech Detection
di: Piot, Paloma, et al.
Pubblicazione: (2025)
di: Piot, Paloma, et al.
Pubblicazione: (2025)
A Comprehensive Evaluation of Cognitive Biases in LLMs
di: Malberg, Simon, et al.
Pubblicazione: (2024)
di: Malberg, Simon, et al.
Pubblicazione: (2024)
Robust Heuristic Algorithm Design with LLMs
di: Karimi, Pantea, et al.
Pubblicazione: (2025)
di: Karimi, Pantea, et al.
Pubblicazione: (2025)
Exploiting Synergistic Cognitive Biases to Bypass Safety in LLMs
di: Yang, Xikang, et al.
Pubblicazione: (2025)
di: Yang, Xikang, et al.
Pubblicazione: (2025)
Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs
di: Han, Pengrui, et al.
Pubblicazione: (2026)
di: Han, Pengrui, et al.
Pubblicazione: (2026)
Large Language Models Cannot Self-Correct Reasoning Yet
di: Huang, Jie, et al.
Pubblicazione: (2023)
di: Huang, Jie, et al.
Pubblicazione: (2023)
Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs
di: Meng, Haoming, et al.
Pubblicazione: (2026)
di: Meng, Haoming, et al.
Pubblicazione: (2026)
Eliciting Better Multilingual Structured Reasoning from LLMs through Code
di: Li, Bryan, et al.
Pubblicazione: (2024)
di: Li, Bryan, et al.
Pubblicazione: (2024)
Hidding the Ghostwriters: An Adversarial Evaluation of AI-Generated Student Essay Detection
di: Peng, Xinlin, et al.
Pubblicazione: (2024)
di: Peng, Xinlin, et al.
Pubblicazione: (2024)
Enhancing Sample Utilization in Noise-Robust Deep Metric Learning With Subgroup-Based Positive-Pair Selection
di: Yu, Zhipeng, et al.
Pubblicazione: (2025)
di: Yu, Zhipeng, et al.
Pubblicazione: (2025)
MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs
di: Wu, Juncheng, et al.
Pubblicazione: (2025)
di: Wu, Juncheng, et al.
Pubblicazione: (2025)
Cannot or Should Not? Automatic Analysis of Refusal Composition in IFT/RLHF Datasets and Refusal Behavior of Black-Box LLMs
di: von Recum, Alexander, et al.
Pubblicazione: (2024)
di: von Recum, Alexander, et al.
Pubblicazione: (2024)
Transformer Block Coupling and its Correlation with Generalization in LLMs
di: Aubry, Murdock, et al.
Pubblicazione: (2024)
di: Aubry, Murdock, et al.
Pubblicazione: (2024)
Unveiling Divergent Inductive Biases of LLMs on Temporal Data
di: Kishore, Sindhu, et al.
Pubblicazione: (2024)
di: Kishore, Sindhu, et al.
Pubblicazione: (2024)
Soft Token Attacks Cannot Reliably Audit Unlearning in Large Language Models
di: Chen, Haokun, et al.
Pubblicazione: (2025)
di: Chen, Haokun, et al.
Pubblicazione: (2025)
Censored LLMs as a Natural Testbed for Secret Knowledge Elicitation
di: Casademunt, Helena, et al.
Pubblicazione: (2026)
di: Casademunt, Helena, et al.
Pubblicazione: (2026)
Strategic Chain-of-Thought: Guiding Accurate Reasoning in LLMs through Strategy Elicitation
di: Wang, Yu, et al.
Pubblicazione: (2024)
di: Wang, Yu, et al.
Pubblicazione: (2024)
LLMs Learn Task Heuristics from Demonstrations: A Heuristic-Driven Prompting Strategy for Document-Level Event Argument Extraction
di: Zhou, Hanzhang, et al.
Pubblicazione: (2023)
di: Zhou, Hanzhang, et al.
Pubblicazione: (2023)
RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models
di: Wang, Zekun Moore, et al.
Pubblicazione: (2023)
di: Wang, Zekun Moore, et al.
Pubblicazione: (2023)
ClinBench-HPB: A Clinical Benchmark for Evaluating LLMs in Hepato-Pancreato-Biliary Diseases
di: Li, Yuchong, et al.
Pubblicazione: (2025)
di: Li, Yuchong, et al.
Pubblicazione: (2025)
MEDEQUALQA: Evaluating Biases in LLMs with Counterfactual Reasoning
di: Ghosh, Rajarshi, et al.
Pubblicazione: (2025)
di: Ghosh, Rajarshi, et al.
Pubblicazione: (2025)
SUPERNOVA: Eliciting General Reasoning in LLMs with Reinforcement Learning on Natural Instructions
di: Suvarna, Ashima, et al.
Pubblicazione: (2026)
di: Suvarna, Ashima, et al.
Pubblicazione: (2026)
Neuronal Group Communication for Efficient Neural representation
di: Pei, Zhengqi, et al.
Pubblicazione: (2025)
di: Pei, Zhengqi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Sequential Manipulation Against Rank Aggregation: Theory and Algorithm
di: Ma, Ke, et al.
Pubblicazione: (2024) -
Divide and Conquer: Heterogeneous Noise Integration for Diffusion-based Adversarial Purification
di: Pei, Gaozheng, et al.
Pubblicazione: (2025) -
Diffusion-based Adversarial Purification from the Perspective of the Frequency Domain
di: Pei, Gaozheng, et al.
Pubblicazione: (2025) -
Can't See the Forest for the Trees: Benchmarking Multimodal Safety Awareness for Multimodal LLMs
di: Wang, Wenxuan, et al.
Pubblicazione: (2025) -
Systematic Analysis of LLM Contributions to Planning: Solver, Verifier, Heuristic
di: Li, Haoming, et al.
Pubblicazione: (2024)