Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement
Fuente:
arXiv
Salvato in:
| Autori principali: | Kersting, Nicholas S., Castelli, Vittorio, Yeh, Chieh Ting, Wang, Xinzhu, Taame, Saad |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
From Black-Box Confidence to Measurable Trust in Clinical AI: A Framework for Evidence, Supervision, and Staged Autonomy
di: Zabolotnii, Serhii, et al.
Pubblicazione: (2026)
di: Zabolotnii, Serhii, et al.
Pubblicazione: (2026)
All in How You Ask for It: Simple Black-Box Method for Jailbreak Attacks
di: Takemoto, Kazuhiro
Pubblicazione: (2024)
di: Takemoto, Kazuhiro
Pubblicazione: (2024)
PatentEdits: Framing Patent Novelty as Textual Entailment
di: Lee, Ryan, et al.
Pubblicazione: (2024)
di: Lee, Ryan, et al.
Pubblicazione: (2024)
Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective
di: Sun, Zhongxiang, et al.
Pubblicazione: (2025)
di: Sun, Zhongxiang, et al.
Pubblicazione: (2025)
SHIELD: Evaluation and Defense Strategies for Copyright Compliance in LLM Text Generation
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
Medical Hallucinations in Foundation Models and Their Impact on Healthcare
di: Kim, Yubin, et al.
Pubblicazione: (2025)
di: Kim, Yubin, et al.
Pubblicazione: (2025)
How Large Language Models are Designed to Hallucinate
di: Ackermann, Richard, et al.
Pubblicazione: (2025)
di: Ackermann, Richard, et al.
Pubblicazione: (2025)
Not Just Novelty: A Longitudinal Study on Utility and Customization of an AI Workflow
di: Long, Tao, et al.
Pubblicazione: (2024)
di: Long, Tao, et al.
Pubblicazione: (2024)
REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading
di: Zhao, Chengshuai, et al.
Pubblicazione: (2026)
di: Zhao, Chengshuai, et al.
Pubblicazione: (2026)
Augmenting Rating-Scale Measures with Text-Derived Items Using the Information-Determined Scoring (IDS) Framework
di: Watson, Joe, et al.
Pubblicazione: (2025)
di: Watson, Joe, et al.
Pubblicazione: (2025)
Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
di: Dahl, Matthew, et al.
Pubblicazione: (2024)
di: Dahl, Matthew, et al.
Pubblicazione: (2024)
H-Neurons: On the Existence, Impact, and Origin of Hallucination-Associated Neurons in LLMs
di: Gao, Cheng, et al.
Pubblicazione: (2025)
di: Gao, Cheng, et al.
Pubblicazione: (2025)
Hallucination Detection: A Probabilistic Framework Using Embeddings Distance Analysis
di: Ricco, Emanuele, et al.
Pubblicazione: (2025)
di: Ricco, Emanuele, et al.
Pubblicazione: (2025)
PICKT: Practical Interlinked Concept Knowledge Tracing for Personalized Learning using Knowledge Map Concept Relations
di: Lee, Wonbeen, et al.
Pubblicazione: (2025)
di: Lee, Wonbeen, et al.
Pubblicazione: (2025)
Place Matters: Comparing LLM Hallucination Rates for Place-Based Legal Queries
di: Curran, Damian, et al.
Pubblicazione: (2025)
di: Curran, Damian, et al.
Pubblicazione: (2025)
What Do LLMs Associate with Your Name? A Human-Centered Black-Box Audit of Personal Data
di: Staufer, Dimitri, et al.
Pubblicazione: (2026)
di: Staufer, Dimitri, et al.
Pubblicazione: (2026)
Thinking Outside the (Gray) Box: A Context-Based Score for Assessing Value and Originality in Neural Text Generation
di: Franceschelli, Giorgio, et al.
Pubblicazione: (2025)
di: Franceschelli, Giorgio, et al.
Pubblicazione: (2025)
Black-Box Hallucination Detection via Consistency Under the Uncertain Expression
di: Joo, Seongho, et al.
Pubblicazione: (2025)
di: Joo, Seongho, et al.
Pubblicazione: (2025)
Identifying Emerging Concepts in Large Corpora
di: Ma, Sibo, et al.
Pubblicazione: (2025)
di: Ma, Sibo, et al.
Pubblicazione: (2025)
Verify when Uncertain: Beyond Self-Consistency in Black Box Hallucination Detection
di: Xue, Yihao, et al.
Pubblicazione: (2025)
di: Xue, Yihao, et al.
Pubblicazione: (2025)
Multilingual and Explainable Text Detoxification with Parallel Corpora
di: Dementieva, Daryna, et al.
Pubblicazione: (2024)
di: Dementieva, Daryna, et al.
Pubblicazione: (2024)
Developmental trajectories of decision making and affective dynamics in large language models
di: Wang, Zhihao, et al.
Pubblicazione: (2025)
di: Wang, Zhihao, et al.
Pubblicazione: (2025)
Measuring Teaching with LLMs
di: Hardy, Michael
Pubblicazione: (2025)
di: Hardy, Michael
Pubblicazione: (2025)
Toxic HallucinAItions: Perturbing Prompts and Tracing LLM Circuits
di: Shimgekar, Soorya Ram, et al.
Pubblicazione: (2026)
di: Shimgekar, Soorya Ram, et al.
Pubblicazione: (2026)
Grounding Text Embeddings in Stakeholder Associations
di: Rystrøm, Jonathan, et al.
Pubblicazione: (2026)
di: Rystrøm, Jonathan, et al.
Pubblicazione: (2026)
Luminol-AIDetect: Fast Zero-shot Machine-Generated Text Detection based on Perplexity under Text Shuffling
di: La Cava, Lucio, et al.
Pubblicazione: (2026)
di: La Cava, Lucio, et al.
Pubblicazione: (2026)
Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation
di: Shen, Hanwen, et al.
Pubblicazione: (2026)
di: Shen, Hanwen, et al.
Pubblicazione: (2026)
Regulating Large Language Models: A Roundtable Report
di: Nicholas, Gabriel, et al.
Pubblicazione: (2024)
di: Nicholas, Gabriel, et al.
Pubblicazione: (2024)
Evaluating the Capabilities of LLMs for Supporting Anticipatory Impact Assessment
di: Allaham, Mowafak, et al.
Pubblicazione: (2024)
di: Allaham, Mowafak, et al.
Pubblicazione: (2024)
From Text to Multimodality: Exploring the Evolution and Impact of Large Language Models in Medical Practice
di: Niu, Qian, et al.
Pubblicazione: (2024)
di: Niu, Qian, et al.
Pubblicazione: (2024)
AI Brown and AI Koditex: LLM-Generated Corpora Comparable to Traditional Corpora of English and Czech Texts
di: Milička, Jiří, et al.
Pubblicazione: (2025)
di: Milička, Jiří, et al.
Pubblicazione: (2025)
Use Sparse Autoencoders to Discover Unknown Concepts, Not to Act on Known Concepts
di: Peng, Kenny, et al.
Pubblicazione: (2025)
di: Peng, Kenny, et al.
Pubblicazione: (2025)
Learned Hallucination Detection in Black-Box LLMs using Token-level Entropy Production Rate
di: Moslonka, Charles, et al.
Pubblicazione: (2025)
di: Moslonka, Charles, et al.
Pubblicazione: (2025)
Computational Measurement of Political Positions: A Review of Text-Based Ideal Point Estimation Algorithms
di: Parschan, Patrick, et al.
Pubblicazione: (2025)
di: Parschan, Patrick, et al.
Pubblicazione: (2025)
ELEPHANT: Measuring and understanding social sycophancy in LLMs
di: Cheng, Myra, et al.
Pubblicazione: (2025)
di: Cheng, Myra, et al.
Pubblicazione: (2025)
Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models
di: Zhang, Yue, et al.
Pubblicazione: (2023)
di: Zhang, Yue, et al.
Pubblicazione: (2023)
AnthroScore: A Computational Linguistic Measure of Anthropomorphism
di: Cheng, Myra, et al.
Pubblicazione: (2024)
di: Cheng, Myra, et al.
Pubblicazione: (2024)
Measuring Human Contribution in AI-Assisted Content Generation
di: Xie, Yueqi, et al.
Pubblicazione: (2024)
di: Xie, Yueqi, et al.
Pubblicazione: (2024)
Measuring Political Preferences in AI Systems: An Integrative Approach
di: Rozado, David
Pubblicazione: (2025)
di: Rozado, David
Pubblicazione: (2025)
Towards Measuring and Modeling "Culture" in LLMs: A Survey
di: Adilazuarda, Muhammad Farid, et al.
Pubblicazione: (2024)
di: Adilazuarda, Muhammad Farid, et al.
Pubblicazione: (2024)
Documenti analoghi
-
From Black-Box Confidence to Measurable Trust in Clinical AI: A Framework for Evidence, Supervision, and Staged Autonomy
di: Zabolotnii, Serhii, et al.
Pubblicazione: (2026) -
All in How You Ask for It: Simple Black-Box Method for Jailbreak Attacks
di: Takemoto, Kazuhiro
Pubblicazione: (2024) -
PatentEdits: Framing Patent Novelty as Textual Entailment
di: Lee, Ryan, et al.
Pubblicazione: (2024) -
Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective
di: Sun, Zhongxiang, et al.
Pubblicazione: (2025) -
SHIELD: Evaluation and Defense Strategies for Copyright Compliance in LLM Text Generation
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)