Incentives or Ontology? A Structural Rebuttal to OpenAI's Hallucination Thesis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ackermann, Richard, Emanuilov, Simeon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Perturbation Dose Responses in Recursive LLM Loops: Raw Switching, Stochastic Floors, and Persistent Escape under Append, Replace, and Dialog Updates
von: Kaplanski, Pawel
Veröffentlicht: (2026)
von: Kaplanski, Pawel
Veröffentlicht: (2026)
Behavioural vs. Representational Systematicity in End-to-End Models: An Opinionated Survey
von: Vegner, Ivan, et al.
Veröffentlicht: (2025)
von: Vegner, Ivan, et al.
Veröffentlicht: (2025)
The Drill-Down and Fabricate Test (DDFT): A Protocol for Measuring Epistemic Robustness in Language Models
von: Baxi, Rahul
Veröffentlicht: (2025)
von: Baxi, Rahul
Veröffentlicht: (2025)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
von: Du, Bangde, et al.
Veröffentlicht: (2025)
von: Du, Bangde, et al.
Veröffentlicht: (2025)
A Domain-Agnostic Neurosymbolic Approach for Big Social Data Analysis: Evaluating Mental Health Sentiment on Social Media during COVID-19
von: Khandelwal, Vedant, et al.
Veröffentlicht: (2024)
von: Khandelwal, Vedant, et al.
Veröffentlicht: (2024)
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2023)
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2023)
Aspect-Based Sentiment Analysis for Future Tourism Experiences: A BERT-MoE Framework for Persian User Reviews
von: Taskooh, Hamidreza Kazemi, et al.
Veröffentlicht: (2026)
von: Taskooh, Hamidreza Kazemi, et al.
Veröffentlicht: (2026)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
Automated Circuit Interpretation via Probe Prompting
von: Birardi, Giuseppe
Veröffentlicht: (2025)
von: Birardi, Giuseppe
Veröffentlicht: (2025)
Amazon Nova AI Challenge -- Trusted AI: Advancing secure, AI-assisted software development
von: Sahai, Sattvik, et al.
Veröffentlicht: (2025)
von: Sahai, Sattvik, et al.
Veröffentlicht: (2025)
Prompt Readiness Levels (PRL): a maturity scale and scoring framework for production grade prompt assets
von: Guinard, Sebastien
Veröffentlicht: (2026)
von: Guinard, Sebastien
Veröffentlicht: (2026)
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
von: Sarkar, Nilesh, et al.
Veröffentlicht: (2026)
von: Sarkar, Nilesh, et al.
Veröffentlicht: (2026)
Beyond Hallucinations: A Composite Score for Measuring Reliability in Open-Source Large Language Models
von: Salla, Rohit Kumar, et al.
Veröffentlicht: (2025)
von: Salla, Rohit Kumar, et al.
Veröffentlicht: (2025)
An Automatic Text Classification Method Based on Hierarchical Taxonomies, Neural Networks and Document Embedding: The NETHIC Tool
von: Lomasto, Luigi, et al.
Veröffentlicht: (2026)
von: Lomasto, Luigi, et al.
Veröffentlicht: (2026)
Why Models Know But Don't Say: Chain-of-Thought Faithfulness Divergence Between Thinking Tokens and Answers in Open-Weight Reasoning Models
von: Young, Richard J.
Veröffentlicht: (2026)
von: Young, Richard J.
Veröffentlicht: (2026)
Weakly Supervised Distillation of Hallucination Signals into Transformer Representations
von: Salehmohamed, Shoaib Sadiq, et al.
Veröffentlicht: (2026)
von: Salehmohamed, Shoaib Sadiq, et al.
Veröffentlicht: (2026)
What Teaches Robots to Walk, Teaches Them to Trade too -- Regime Adaptive Execution using Informed Data and LLMs
von: Saqur, Raeid
Veröffentlicht: (2024)
von: Saqur, Raeid
Veröffentlicht: (2024)
SigWavNet: Learning Multiresolution Signal Wavelet Network for Speech Emotion Recognition
von: Nfissi, Alaa, et al.
Veröffentlicht: (2025)
von: Nfissi, Alaa, et al.
Veröffentlicht: (2025)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
Pioneer Agent: Continual Improvement of Small Language Models in Production
von: Atreja, Dhruv, et al.
Veröffentlicht: (2026)
von: Atreja, Dhruv, et al.
Veröffentlicht: (2026)
Reasoning Large Language Model Errors Arise from Hallucinating Critical Problem Features
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
Synthius-Mem: Brain-Inspired Hallucination-Resistant Persona Memory Achieving 94.4% Memory Accuracy and 99.6% Adversarial Robustness on LoCoMo
von: Gadzhiev, Artem, et al.
Veröffentlicht: (2026)
von: Gadzhiev, Artem, et al.
Veröffentlicht: (2026)
Teaching a Language Model to Speak the Language of Tools
von: Emanuilov, Simeon
Veröffentlicht: (2025)
von: Emanuilov, Simeon
Veröffentlicht: (2025)
OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
MODP: Multi Objective Directional Prompting
von: Nema, Aashutosh, et al.
Veröffentlicht: (2025)
von: Nema, Aashutosh, et al.
Veröffentlicht: (2025)
Can AI Read Between The Lines? Benchmarking LLMs On Financial Nuance
von: Kubica, Dominick, et al.
Veröffentlicht: (2025)
von: Kubica, Dominick, et al.
Veröffentlicht: (2025)
Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition
von: Goyal, Prasoon, et al.
Veröffentlicht: (2026)
von: Goyal, Prasoon, et al.
Veröffentlicht: (2026)
Gyan: An Explainable Neuro-Symbolic Language Model
von: Srinivasan, Venkat, et al.
Veröffentlicht: (2026)
von: Srinivasan, Venkat, et al.
Veröffentlicht: (2026)
Stemming Hallucination in Language Models Using a Licensing Oracle
von: Emanuilov, Simeon, et al.
Veröffentlicht: (2025)
von: Emanuilov, Simeon, et al.
Veröffentlicht: (2025)
Eyla: Toward an Identity-Anchored LLM Architecture with Integrated Biological Priors -- Vision, Implementation Attempt, and Lessons from AI-Assisted Development
von: Aditto, Arif
Veröffentlicht: (2026)
von: Aditto, Arif
Veröffentlicht: (2026)
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
von: Yocam, Eric, et al.
Veröffentlicht: (2026)
von: Yocam, Eric, et al.
Veröffentlicht: (2026)
Social Cooperation in Conversational AI Agents
von: Çelikok, Mustafa Mert, et al.
Veröffentlicht: (2025)
von: Çelikok, Mustafa Mert, et al.
Veröffentlicht: (2025)
Structured Prompt Optimization Meets Reinforcement Learning for Global and Local Interpretability over Complex Text
von: Zhou, Tianyang, et al.
Veröffentlicht: (2026)
von: Zhou, Tianyang, et al.
Veröffentlicht: (2026)
Grokking in the Wild: Data Augmentation for Real-World Multi-Hop Reasoning with Transformers
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory
von: Zhou, Wenxuan, et al.
Veröffentlicht: (2025)
von: Zhou, Wenxuan, et al.
Veröffentlicht: (2025)
Towards Ontology-Enhanced Representation Learning for Large Language Models
von: Ronzano, Francesco, et al.
Veröffentlicht: (2024)
von: Ronzano, Francesco, et al.
Veröffentlicht: (2024)
Beyond Deep Learning: Speech Segmentation and Phone Classification with Neural Assemblies
von: Adelson, Trevor, et al.
Veröffentlicht: (2026)
von: Adelson, Trevor, et al.
Veröffentlicht: (2026)
Pareto-Optimized Open-Source LLMs for Healthcare via Context Retrieval
von: Bayarri-Planas, Jordi, et al.
Veröffentlicht: (2024)
von: Bayarri-Planas, Jordi, et al.
Veröffentlicht: (2024)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
von: Costa, Rimom
Veröffentlicht: (2025)
von: Costa, Rimom
Veröffentlicht: (2025)
Ähnliche Einträge
-
Perturbation Dose Responses in Recursive LLM Loops: Raw Switching, Stochastic Floors, and Persistent Escape under Append, Replace, and Dialog Updates
von: Kaplanski, Pawel
Veröffentlicht: (2026) -
Behavioural vs. Representational Systematicity in End-to-End Models: An Opinionated Survey
von: Vegner, Ivan, et al.
Veröffentlicht: (2025) -
The Drill-Down and Fabricate Test (DDFT): A Protocol for Measuring Epistemic Robustness in Language Models
von: Baxi, Rahul
Veröffentlicht: (2025) -
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
von: Du, Bangde, et al.
Veröffentlicht: (2025) -
A Domain-Agnostic Neurosymbolic Approach for Big Social Data Analysis: Evaluating Mental Health Sentiment on Social Media during COVID-19
von: Khandelwal, Vedant, et al.
Veröffentlicht: (2024)