The Delusional Hedge Algorithm as a Model of Human Learning from Diverse Opinions
Fuente:
arXiv
Saved in:
| Main Authors: | Chuang, Yun-Shiuan, Zhu, Jerry, Rogers, Timothy T. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Probing LLM World Models: Enhancing Guesstimation with Wisdom of Crowds Decoding
by: Chuang, Yun-Shiuan, et al.
Published: (2025)
by: Chuang, Yun-Shiuan, et al.
Published: (2025)
Characterizing Delusional Spirals through Human-LLM Chat Logs
by: Moore, Jared, et al.
Published: (2026)
by: Moore, Jared, et al.
Published: (2026)
Deep Reinforcement Learning Algorithms for Option Hedging
by: Neagu, Andrei, et al.
Published: (2025)
by: Neagu, Andrei, et al.
Published: (2025)
Sycophantic Chatbots Cause Delusional Spiraling, Even in Ideal Bayesians
by: Chandra, Kartik, et al.
Published: (2026)
by: Chandra, Kartik, et al.
Published: (2026)
Uncovering the Computational Ingredients of Human-Like Representations in LLMs
by: Studdiford, Zach, et al.
Published: (2025)
by: Studdiford, Zach, et al.
Published: (2025)
The Ontological Dissonance Hypothesis: AI-Triggered Delusional Ideation as Folie a Deux Technologique
by: Lipinska, Izabela, et al.
Published: (2025)
by: Lipinska, Izabela, et al.
Published: (2025)
Evaluating Steering Techniques using Human Similarity Judgments
by: Studdiford, Zach, et al.
Published: (2025)
by: Studdiford, Zach, et al.
Published: (2025)
Learning from Noisy Labels for Long-tailed Data via Optimal Transport
by: Li, Mengting, et al.
Published: (2024)
by: Li, Mengting, et al.
Published: (2024)
UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling
by: Chen, Boyu, et al.
Published: (2026)
by: Chen, Boyu, et al.
Published: (2026)
Deep Hedging with Market Impact
by: Neagu, Andrei, et al.
Published: (2024)
by: Neagu, Andrei, et al.
Published: (2024)
Facilitating Opinion Diversity through Hybrid NLP Approaches
by: van der Meer, Michiel
Published: (2024)
by: van der Meer, Michiel
Published: (2024)
Hedging and Non-Affirmation: Quantifying LLM Alignment on Questions of Human Rights
by: Javed, Rafiya, et al.
Published: (2025)
by: Javed, Rafiya, et al.
Published: (2025)
Toward Scalable Verifiable Reward: Proxy State-Based Evaluation for Multi-turn Tool-Calling LLM Agents
by: Chuang, Yun-Shiuan, et al.
Published: (2026)
by: Chuang, Yun-Shiuan, et al.
Published: (2026)
Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble
by: Zhang, Shun, et al.
Published: (2024)
by: Zhang, Shun, et al.
Published: (2024)
SELFDOUBT: Uncertainty Quantification for Reasoning LLMs via the Hedge-to-Verify Ratio
by: Pandey, Satwik, et al.
Published: (2026)
by: Pandey, Satwik, et al.
Published: (2026)
Opinion-Guided Reinforcement Learning
by: Dagenais, Kyanna, et al.
Published: (2024)
by: Dagenais, Kyanna, et al.
Published: (2024)
Mastering Diverse Domains through World Models
by: Hafner, Danijar, et al.
Published: (2023)
by: Hafner, Danijar, et al.
Published: (2023)
Algorithmic Scenario Generation as Quality Diversity Optimization
by: Nikolaidis, Stefanos
Published: (2024)
by: Nikolaidis, Stefanos
Published: (2024)
Hedging Is Not All You Need: A Simple Baseline for Online Learning Under Haphazard Inputs
by: Buckchash, Himanshu, et al.
Published: (2024)
by: Buckchash, Himanshu, et al.
Published: (2024)
Ada-RS: Adaptive Rejection Sampling for Selective Thinking
by: Ge, Yirou, et al.
Published: (2026)
by: Ge, Yirou, et al.
Published: (2026)
Differential Multimodal Transformers
by: Li, Jerry, et al.
Published: (2025)
by: Li, Jerry, et al.
Published: (2025)
MultiHedge: Adaptive Coordination via Retrieval-Augmented Control
by: Bańka, Feliks, et al.
Published: (2026)
by: Bańka, Feliks, et al.
Published: (2026)
Optimizing Large Language Models for Dynamic Constraints through Human-in-the-Loop Discriminators
by: Wei, Timothy, et al.
Published: (2024)
by: Wei, Timothy, et al.
Published: (2024)
Opinion Mining Based Entity Ranking using Fuzzy Logic Algorithmic Approach
by: Kalamkar, Pratik N., et al.
Published: (2025)
by: Kalamkar, Pratik N., et al.
Published: (2025)
How Well Can a Long Sequence Model Model Long Sequences? Comparing Architechtural Inductive Biases on Long-Context Abilities
by: Huang, Jerry
Published: (2024)
by: Huang, Jerry
Published: (2024)
Simulating Opinion Dynamics with Networks of LLM-based Agents
by: Chuang, Yun-Shiuan, et al.
Published: (2023)
by: Chuang, Yun-Shiuan, et al.
Published: (2023)
AlignDiff: Aligning Diverse Human Preferences via Behavior-Customisable Diffusion Model
by: Dong, Zibin, et al.
Published: (2023)
by: Dong, Zibin, et al.
Published: (2023)
Persona-Based Simulation of Human Opinion at Population Scale
by: Li, Mao, et al.
Published: (2026)
by: Li, Mao, et al.
Published: (2026)
Algorithmic Prompt Generation for Diverse Human-like Teaming and Communication with Large Language Models
by: Srikanth, Siddharth, et al.
Published: (2025)
by: Srikanth, Siddharth, et al.
Published: (2025)
Can Large Language Models Capture Public Opinion about Global Warming? An Empirical Assessment of Algorithmic Fidelity and Bias
by: Lee, S., et al.
Published: (2023)
by: Lee, S., et al.
Published: (2023)
diffGHOST: Diffusion based Generative Hedged Oblivious Synthetic Trajectories
by: Guépin, Florent, et al.
Published: (2026)
by: Guépin, Florent, et al.
Published: (2026)
Diverse Transformer Decoding for Offline Reinforcement Learning Using Financial Algorithmic Approaches
by: Elbaz, Dan, et al.
Published: (2025)
by: Elbaz, Dan, et al.
Published: (2025)
Uniform Memory Retrieval with Larger Capacity for Modern Hopfield Models
by: Wu, Dennis, et al.
Published: (2024)
by: Wu, Dennis, et al.
Published: (2024)
NEMO-4-PAYPAL: Leveraging NVIDIA's Nemo Framework for empowering PayPal's Commerce Agent
by: Garg, Sudhanshu, et al.
Published: (2025)
by: Garg, Sudhanshu, et al.
Published: (2025)
Rethinking Diverse Human Preference Learning through Principal Component Analysis
by: Luo, Feng, et al.
Published: (2025)
by: Luo, Feng, et al.
Published: (2025)
Training LLMs to Recognize Hedges in Spontaneous Narratives
by: Paige, Amie J., et al.
Published: (2024)
by: Paige, Amie J., et al.
Published: (2024)
How Diversely Can Language Models Solve Problems? Exploring the Algorithmic Diversity of Model-Generated Code
by: Lee, Seonghyeon, et al.
Published: (2025)
by: Lee, Seonghyeon, et al.
Published: (2025)
Language Models Exhibit Inconsistent Biases Towards Algorithmic Agents and Human Experts
by: Bo, Jessica Y., et al.
Published: (2026)
by: Bo, Jessica Y., et al.
Published: (2026)
Provable Multi-Party Reinforcement Learning with Diverse Human Feedback
by: Zhong, Huiying, et al.
Published: (2024)
by: Zhong, Huiying, et al.
Published: (2024)
Large Language Models estimate fine-grained human color-concept associations
by: Mukherjee, Kushin, et al.
Published: (2024)
by: Mukherjee, Kushin, et al.
Published: (2024)
Similar Items
-
Probing LLM World Models: Enhancing Guesstimation with Wisdom of Crowds Decoding
by: Chuang, Yun-Shiuan, et al.
Published: (2025) -
Characterizing Delusional Spirals through Human-LLM Chat Logs
by: Moore, Jared, et al.
Published: (2026) -
Deep Reinforcement Learning Algorithms for Option Hedging
by: Neagu, Andrei, et al.
Published: (2025) -
Sycophantic Chatbots Cause Delusional Spiraling, Even in Ideal Bayesians
by: Chandra, Kartik, et al.
Published: (2026) -
Uncovering the Computational Ingredients of Human-Like Representations in LLMs
by: Studdiford, Zach, et al.
Published: (2025)