Exploring Chain-of-Thought Reasoning for Steerable Pluralistic Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Yunfan, McKeown, Kathleen, Muresan, Smaranda |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LiveNewsBench: Evaluating LLM Web Search Capabilities with Freshly Curated News
di: Zhang, Yunfan, et al.
Pubblicazione: (2026)
di: Zhang, Yunfan, et al.
Pubblicazione: (2026)
Forecasting Conversation Derailments Through Generation
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
Layered Insights: Generalizable Analysis of Authorial Style by Leveraging All Transformer Layers
di: Alshomary, Milad, et al.
Pubblicazione: (2025)
di: Alshomary, Milad, et al.
Pubblicazione: (2025)
XAM: Interactive Explainability for Authorship Attribution Models
di: Alshomary, Milad, et al.
Pubblicazione: (2025)
di: Alshomary, Milad, et al.
Pubblicazione: (2025)
LLMs as Science Journalists: Supporting Early-stage Researchers in Communicating Their Science to the Public
di: Alshomary, Milad, et al.
Pubblicazione: (2026)
di: Alshomary, Milad, et al.
Pubblicazione: (2026)
Latent Space Interpretation for Stylistic Analysis and Explainable Authorship Attribution
di: Alshomary, Milad, et al.
Pubblicazione: (2024)
di: Alshomary, Milad, et al.
Pubblicazione: (2024)
Designing and Evaluating Chain-of-Hints for Scientific Question Answering
di: Jangra, Anubhav, et al.
Pubblicazione: (2025)
di: Jangra, Anubhav, et al.
Pubblicazione: (2025)
Parallel Structures in Pre-training Data Yield In-Context Learning
di: Chen, Yanda, et al.
Pubblicazione: (2024)
di: Chen, Yanda, et al.
Pubblicazione: (2024)
On the Relation between Sensitivity and Accuracy in In-context Learning
di: Chen, Yanda, et al.
Pubblicazione: (2022)
di: Chen, Yanda, et al.
Pubblicazione: (2022)
Social Orientation: A New Feature for Dialogue Analysis
di: Morrill, Todd, et al.
Pubblicazione: (2024)
di: Morrill, Todd, et al.
Pubblicazione: (2024)
StyleDistance: Stronger Content-Independent Style Embeddings with Synthetic Parallel Examples
di: Patel, Ajay, et al.
Pubblicazione: (2024)
di: Patel, Ajay, et al.
Pubblicazione: (2024)
Artificial Impressions: Evaluating Large Language Model Behavior Through the Lens of Trait Impressions
di: Deas, Nicholas, et al.
Pubblicazione: (2025)
di: Deas, Nicholas, et al.
Pubblicazione: (2025)
Summarization of Opinionated Political Documents with Varied Perspectives
di: Deas, Nicholas, et al.
Pubblicazione: (2024)
di: Deas, Nicholas, et al.
Pubblicazione: (2024)
Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs
di: Zhang, Xuan, et al.
Pubblicazione: (2024)
di: Zhang, Xuan, et al.
Pubblicazione: (2024)
Getting Serious about Humor: Crafting Humor Datasets with Unfunny Large Language Models
di: Horvitz, Zachary, et al.
Pubblicazione: (2024)
di: Horvitz, Zachary, et al.
Pubblicazione: (2024)
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2023)
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2023)
PluralLLM: Pluralistic Alignment in LLMs via Federated Learning
di: Srewa, Mahmoud, et al.
Pubblicazione: (2025)
di: Srewa, Mahmoud, et al.
Pubblicazione: (2025)
A General Framework for Inference-time Scaling and Steering of Diffusion Models
di: Singhal, Raghav, et al.
Pubblicazione: (2025)
di: Singhal, Raghav, et al.
Pubblicazione: (2025)
Demystifying Long Chain-of-Thought Reasoning in LLMs
di: Yeo, Edward, et al.
Pubblicazione: (2025)
di: Yeo, Edward, et al.
Pubblicazione: (2025)
Understanding Hidden Computations in Chain-of-Thought Reasoning
di: Bharadwaj, Aryasomayajula Ram
Pubblicazione: (2024)
di: Bharadwaj, Aryasomayajula Ram
Pubblicazione: (2024)
Fractured Chain-of-Thought Reasoning
di: Liao, Baohao, et al.
Pubblicazione: (2025)
di: Liao, Baohao, et al.
Pubblicazione: (2025)
Counterfactual Reasoning for Steerable Pluralistic Value Alignment of Large Language Models
di: Guo, Hanze, et al.
Pubblicazione: (2025)
di: Guo, Hanze, et al.
Pubblicazione: (2025)
Pause and Reflect: Conformal Aggregation for Chain-of-Thought Reasoning
di: Gu, Yu, et al.
Pubblicazione: (2026)
di: Gu, Yu, et al.
Pubblicazione: (2026)
Pluralistic Alignment for Healthcare: A Role-Driven Framework
di: Zhong, Jiayou, et al.
Pubblicazione: (2025)
di: Zhong, Jiayou, et al.
Pubblicazione: (2025)
MixDPO: Modeling Preference Strength for Pluralistic Alignment
di: Imai, Saki, et al.
Pubblicazione: (2026)
di: Imai, Saki, et al.
Pubblicazione: (2026)
ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges
di: Qian, Cheng, et al.
Pubblicazione: (2025)
di: Qian, Cheng, et al.
Pubblicazione: (2025)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
di: Ye, Jiacheng, et al.
Pubblicazione: (2024)
di: Ye, Jiacheng, et al.
Pubblicazione: (2024)
Dissecting Long-Chain-of-Thought Reasoning Models: An Empirical Study
di: Mu, Yongyu, et al.
Pubblicazione: (2025)
di: Mu, Yongyu, et al.
Pubblicazione: (2025)
VITAL: A New Dataset for Benchmarking Pluralistic Alignment in Healthcare
di: Shetty, Anudeex, et al.
Pubblicazione: (2025)
di: Shetty, Anudeex, et al.
Pubblicazione: (2025)
VISPA: Pluralistic Alignment via Automatic Value Selection and Activation
di: Zheng, Shenyan, et al.
Pubblicazione: (2026)
di: Zheng, Shenyan, et al.
Pubblicazione: (2026)
Reranking-based Generation for Unbiased Perspective Summarization
di: Ri, Narutatsu, et al.
Pubblicazione: (2025)
di: Ri, Narutatsu, et al.
Pubblicazione: (2025)
iBERT: Interpretable Embeddings via Sense Decomposition
di: Anand, Vishal, et al.
Pubblicazione: (2025)
di: Anand, Vishal, et al.
Pubblicazione: (2025)
Feature Extraction and Steering for Enhanced Chain-of-Thought Reasoning in Language Models
di: Li, Zihao, et al.
Pubblicazione: (2025)
di: Li, Zihao, et al.
Pubblicazione: (2025)
Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning
di: Jia, Jinghan, et al.
Pubblicazione: (2026)
di: Jia, Jinghan, et al.
Pubblicazione: (2026)
Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
di: Arcuschin, Iván, et al.
Pubblicazione: (2025)
di: Arcuschin, Iván, et al.
Pubblicazione: (2025)
Compositional Reasoning with Transformers, RNNs, and Chain of Thought
di: Yehudai, Gilad, et al.
Pubblicazione: (2025)
di: Yehudai, Gilad, et al.
Pubblicazione: (2025)
Scalable Chain of Thoughts via Elastic Reasoning
di: Xu, Yuhui, et al.
Pubblicazione: (2025)
di: Xu, Yuhui, et al.
Pubblicazione: (2025)
Long Chain-of-Thought Reasoning Across Languages
di: Barua, Josh, et al.
Pubblicazione: (2025)
di: Barua, Josh, et al.
Pubblicazione: (2025)
Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning
di: Li, Xintong, et al.
Pubblicazione: (2026)
di: Li, Xintong, et al.
Pubblicazione: (2026)
DYSTIL: Dynamic Strategy Induction with Large Language Models for Reinforcement Learning
di: Wang, Borui, et al.
Pubblicazione: (2025)
di: Wang, Borui, et al.
Pubblicazione: (2025)
Documenti analoghi
-
LiveNewsBench: Evaluating LLM Web Search Capabilities with Freshly Curated News
di: Zhang, Yunfan, et al.
Pubblicazione: (2026) -
Forecasting Conversation Derailments Through Generation
di: Zhang, Yunfan, et al.
Pubblicazione: (2025) -
Layered Insights: Generalizable Analysis of Authorial Style by Leveraging All Transformer Layers
di: Alshomary, Milad, et al.
Pubblicazione: (2025) -
XAM: Interactive Explainability for Authorship Attribution Models
di: Alshomary, Milad, et al.
Pubblicazione: (2025) -
LLMs as Science Journalists: Supporting Early-stage Researchers in Communicating Their Science to the Public
di: Alshomary, Milad, et al.
Pubblicazione: (2026)