Discourse Heuristics For Paradoxically Moral Self-Correction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Guangliang, Qi, Zimo, Zhang, Xitong, Johnson, Kristen Marie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Diagnosing Moral Reasoning Acquisition in Language Models: Pragmatics and Generalization
von: Liu, Guangliang, et al.
Veröffentlicht: (2025)
von: Liu, Guangliang, et al.
Veröffentlicht: (2025)
Self-correction is Not An Innate Capability in Language Models
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
Smaller Large Language Models Can Do Moral Self-Correction
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
Learning to Diagnose and Correct Errors: Towards Moral Sensitivity Acquisition in Large Language Models
von: Chen, Bocheng, et al.
Veröffentlicht: (2026)
von: Chen, Bocheng, et al.
Veröffentlicht: (2026)
On the Convergence of Moral Self-Correction in Large Language Models
von: Liu, Guangliang, et al.
Veröffentlicht: (2025)
von: Liu, Guangliang, et al.
Veröffentlicht: (2025)
Diagnosing the Performance Trade-off in Moral Alignment: A Case Study on Gender Stereotypes
von: Liu, Guangliang, et al.
Veröffentlicht: (2025)
von: Liu, Guangliang, et al.
Veröffentlicht: (2025)
Pragmatic Inference for Moral Reasoning Acquisition: Generalization via Metapragmatic Links
von: Liu, Guangliang, et al.
Veröffentlicht: (2025)
von: Liu, Guangliang, et al.
Veröffentlicht: (2025)
Intrinsic Self-correction for Enhanced Morality: An Analysis of Internal Mechanisms and the Superficial Hypothesis
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
On the Intrinsic Self-Correction Capability of LLMs: Uncertainty and Latent Concept
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
Towards Understanding Task-agnostic Debiasing Through the Lenses of Intrinsic Bias and Forgetfulness
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
A Survey to Recent Progress Towards Understanding In-Context Learning
von: Mao, Haitao, et al.
Veröffentlicht: (2024)
von: Mao, Haitao, et al.
Veröffentlicht: (2024)
Can Large Language Models Handle Discourse Particles? A Case Study of Colloquial Malay
von: Yusoff, Mariah Al Giptiah Binte, et al.
Veröffentlicht: (2026)
von: Yusoff, Mariah Al Giptiah Binte, et al.
Veröffentlicht: (2026)
Polarization and Morality: Lexical Analysis of Abortion Discourse on Reddit
von: Stanier, Tessa, et al.
Veröffentlicht: (2024)
von: Stanier, Tessa, et al.
Veröffentlicht: (2024)
From Words to Wisdom: Discourse Annotation and Baseline Models for Student Dialogue Understanding
von: Mim, Farjana Sultana, et al.
Veröffentlicht: (2025)
von: Mim, Farjana Sultana, et al.
Veröffentlicht: (2025)
The Compression Paradox in LLM Inference: Provider-Dependent Energy Effects of Prompt Compression
von: Johnson, Warren
Veröffentlicht: (2026)
von: Johnson, Warren
Veröffentlicht: (2026)
Adaptable Moral Stances of Large Language Models on Sexist Content: Implications for Society and Gender Discourse
von: Guo, Rongchen, et al.
Veröffentlicht: (2024)
von: Guo, Rongchen, et al.
Veröffentlicht: (2024)
Which course? Discourse! Teaching Discourse and Generation in the Era of LLMs
von: Li, Junyi Jessy, et al.
Veröffentlicht: (2026)
von: Li, Junyi Jessy, et al.
Veröffentlicht: (2026)
Guardians of Discourse: Evaluating LLMs on Multilingual Offensive Language Detection
von: He, Jianfei, et al.
Veröffentlicht: (2024)
von: He, Jianfei, et al.
Veröffentlicht: (2024)
On the Paradoxical Interference between Instruction-Following and Task Solving
von: Qi, Yunjia, et al.
Veröffentlicht: (2026)
von: Qi, Yunjia, et al.
Veröffentlicht: (2026)
The Perplexity Paradox: Why Code Compresses Better Than Math in LLM Prompts
von: Johnson, Warren
Veröffentlicht: (2026)
von: Johnson, Warren
Veröffentlicht: (2026)
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
von: Aoki, Yoichi, et al.
Veröffentlicht: (2024)
von: Aoki, Yoichi, et al.
Veröffentlicht: (2024)
Understanding the Dark Side of LLMs' Intrinsic Self-Correction
von: Zhang, Qingjie, et al.
Veröffentlicht: (2024)
von: Zhang, Qingjie, et al.
Veröffentlicht: (2024)
Rhea: Role-aware Heuristic Episodic Attention for Conversational LLMs
von: Hong, Wanyang, et al.
Veröffentlicht: (2025)
von: Hong, Wanyang, et al.
Veröffentlicht: (2025)
Context-Aware Counterfactual Data Augmentation for Gender Bias Mitigation in Language Models
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
von: Parihar, Shweta, et al.
Veröffentlicht: (2026)
Self-Explaining Hate Speech Detection with Moral Rationales
von: Vargas, Francielle, et al.
Veröffentlicht: (2026)
von: Vargas, Francielle, et al.
Veröffentlicht: (2026)
MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation
von: He, Junlin, et al.
Veröffentlicht: (2026)
von: He, Junlin, et al.
Veröffentlicht: (2026)
When Can LLMs Actually Correct Their Own Mistakes? A Critical Survey of Self-Correction of LLMs
von: Kamoi, Ryo, et al.
Veröffentlicht: (2024)
von: Kamoi, Ryo, et al.
Veröffentlicht: (2024)
Self-Correction Makes LLMs Better Parsers
von: Zhang, Ziyan, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyan, et al.
Veröffentlicht: (2025)
Automatic Alignment of Discourse Relations of Different Discourse Annotation Frameworks
von: Fu, Yingxue
Veröffentlicht: (2024)
von: Fu, Yingxue
Veröffentlicht: (2024)
Plan of Thoughts: Heuristic-Guided Problem Solving with Large Language Models
von: Liu, Houjun
Veröffentlicht: (2024)
von: Liu, Houjun
Veröffentlicht: (2024)
The Self-Improvement Paradox: Can Language Models Bootstrap Reasoning Capabilities without External Scaffolding?
von: Sun, Yutao, et al.
Veröffentlicht: (2025)
von: Sun, Yutao, et al.
Veröffentlicht: (2025)
Are Bigger Encoders Always Better in Vision Large Models?
von: Li, Bozhou, et al.
Veröffentlicht: (2024)
von: Li, Bozhou, et al.
Veröffentlicht: (2024)
MOKA: Moral Knowledge Augmentation for Moral Event Extraction
von: Zhang, Xinliang Frederick, et al.
Veröffentlicht: (2023)
von: Zhang, Xinliang Frederick, et al.
Veröffentlicht: (2023)
Few-Shot Query Intent Detection via Relation-Aware Prompt Learning
von: Zhang, Liang, et al.
Veröffentlicht: (2025)
von: Zhang, Liang, et al.
Veröffentlicht: (2025)
SuperCorrect: Advancing Small LLM Reasoning with Thought Template Distillation and Self-Correction
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
Joint Modeling of Entities and Discourse Relations for Coherence Assessment
von: Liu, Wei, et al.
Veröffentlicht: (2025)
von: Liu, Wei, et al.
Veröffentlicht: (2025)
CorrectionLM: Self-Corrections with SLM for Dialogue State Tracking
von: Lee, Chia-Hsuan, et al.
Veröffentlicht: (2024)
von: Lee, Chia-Hsuan, et al.
Veröffentlicht: (2024)
On the Role of Context for Discourse Relation Classification in Scientific Writing
von: Wan, Stephen, et al.
Veröffentlicht: (2025)
von: Wan, Stephen, et al.
Veröffentlicht: (2025)
Beyond English: Evaluating Automated Measurement of Moral Foundations in Non-English Discourse with a Chinese Case Study
von: Cheng, Calvin Yixiang, et al.
Veröffentlicht: (2025)
von: Cheng, Calvin Yixiang, et al.
Veröffentlicht: (2025)
Corrective In-Context Learning: Evaluating Self-Correction in Large Language Models
von: Sanz-Guerrero, Mario, et al.
Veröffentlicht: (2025)
von: Sanz-Guerrero, Mario, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Diagnosing Moral Reasoning Acquisition in Language Models: Pragmatics and Generalization
von: Liu, Guangliang, et al.
Veröffentlicht: (2025) -
Self-correction is Not An Innate Capability in Language Models
von: Liu, Guangliang, et al.
Veröffentlicht: (2024) -
Smaller Large Language Models Can Do Moral Self-Correction
von: Liu, Guangliang, et al.
Veröffentlicht: (2024) -
Learning to Diagnose and Correct Errors: Towards Moral Sensitivity Acquisition in Large Language Models
von: Chen, Bocheng, et al.
Veröffentlicht: (2026) -
On the Convergence of Moral Self-Correction in Large Language Models
von: Liu, Guangliang, et al.
Veröffentlicht: (2025)