Salvato in:
| Autore principale: | Coronado-Blázquez, Javier |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2502.19965 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS)
di: Awad, Samer, et al.
Pubblicazione: (2026)
di: Awad, Samer, et al.
Pubblicazione: (2026)
Assessing the Performance of Human-Capable LLMs -- Are LLMs Coming for Your Job?
di: Mavi, John, et al.
Pubblicazione: (2024)
di: Mavi, John, et al.
Pubblicazione: (2024)
Exploring the psychology of LLMs' Moral and Legal Reasoning
di: Almeida, Guilherme F. C. F., et al.
Pubblicazione: (2023)
di: Almeida, Guilherme F. C. F., et al.
Pubblicazione: (2023)
Evaluating book summaries from internal knowledge in Large Language Models: a cross-model and semantic consistency approach
di: Coronado-Blázquez, Javier
Pubblicazione: (2025)
di: Coronado-Blázquez, Javier
Pubblicazione: (2025)
From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law
di: Mavi, John, et al.
Pubblicazione: (2025)
di: Mavi, John, et al.
Pubblicazione: (2025)
A NLP Approach to "Review Bombing" in Metacritic PC Videogames User Ratings
di: Coronado-Blázquez, Javier
Pubblicazione: (2024)
di: Coronado-Blázquez, Javier
Pubblicazione: (2024)
Redefining "Hallucination" in LLMs: Towards a psychology-informed framework for mitigating misinformation
di: Berberette, Elijah, et al.
Pubblicazione: (2024)
di: Berberette, Elijah, et al.
Pubblicazione: (2024)
PsyMem: Fine-grained psychological alignment and Explicit Memory Control for Advanced Role-Playing LLMs
di: Cheng, Xilong, et al.
Pubblicazione: (2025)
di: Cheng, Xilong, et al.
Pubblicazione: (2025)
ITLC at SemEval-2026 Task 11: Normalization and Deterministic Parsing for Formal Reasoning in LLMs
di: Muhamad, Wicaksono Leksono, et al.
Pubblicazione: (2026)
di: Muhamad, Wicaksono Leksono, et al.
Pubblicazione: (2026)
Are LLMs effective psychological assessors? Leveraging adaptive RAG for interpretable mental health screening through psychometric practice
di: Ravenda, Federico, et al.
Pubblicazione: (2025)
di: Ravenda, Federico, et al.
Pubblicazione: (2025)
A Geometric Taxonomy of Hallucinations in LLMs
di: Marín, Javier
Pubblicazione: (2026)
di: Marín, Javier
Pubblicazione: (2026)
Empirical Characterization of Temporal Constraint Processing in LLMs
di: Marín, Javier
Pubblicazione: (2025)
di: Marín, Javier
Pubblicazione: (2025)
Unifying Ontology Construction and Semantic Alignment for Deterministic Enterprise Reasoning at Scale
di: Zhu, Hongyin
Pubblicazione: (2026)
di: Zhu, Hongyin
Pubblicazione: (2026)
<think> So let's replace this phrase with insult... </think> Lessons learned from generation of toxic texts with LLMs
di: Pletenev, Sergey, et al.
Pubblicazione: (2025)
di: Pletenev, Sergey, et al.
Pubblicazione: (2025)
Not all tokens are created equal: Perplexity Attention Weighted Networks for AI generated text detection
di: Miralles-González, Pablo, et al.
Pubblicazione: (2025)
di: Miralles-González, Pablo, et al.
Pubblicazione: (2025)
An evaluation of LLMs for generating movie reviews: GPT-4o, Gemini-2.0 and DeepSeek-V3
di: Sands, Brendan, et al.
Pubblicazione: (2025)
di: Sands, Brendan, et al.
Pubblicazione: (2025)
Can LLMs Evaluate What They Cannot Annotate? Revisiting LLM Reliability in Hate Speech Detection
di: Piot, Paloma, et al.
Pubblicazione: (2025)
di: Piot, Paloma, et al.
Pubblicazione: (2025)
Psycholinguistic Word Features: a New Approach for the Evaluation of LLMs Alignment with Humans
di: Conde, Javier, et al.
Pubblicazione: (2025)
di: Conde, Javier, et al.
Pubblicazione: (2025)
Automated test generation to evaluate tool-augmented LLMs as conversational AI agents
di: Arcadinho, Samuel, et al.
Pubblicazione: (2024)
di: Arcadinho, Samuel, et al.
Pubblicazione: (2024)
What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering
di: Maar, Jim, et al.
Pubblicazione: (2026)
di: Maar, Jim, et al.
Pubblicazione: (2026)
Multiple Choice Questions: Reasoning Makes Large Language Models (LLMs) More Self-Confident, Especially When They are Wrong
di: Fu, Tairan, et al.
Pubblicazione: (2025)
di: Fu, Tairan, et al.
Pubblicazione: (2025)
Auto-Cypher: Improving LLMs on Cypher generation via LLM-supervised generation-verification framework
di: Tiwari, Aman, et al.
Pubblicazione: (2024)
di: Tiwari, Aman, et al.
Pubblicazione: (2024)
Can LLMs Write Faithfully? An Agent-Based Evaluation of LLM-generated Islamic Content
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
CogBench: a large language model walks into a psychology lab
di: Coda-Forno, Julian, et al.
Pubblicazione: (2024)
di: Coda-Forno, Julian, et al.
Pubblicazione: (2024)
Non-Determinism of "Deterministic" LLM Settings
di: Atil, Berk, et al.
Pubblicazione: (2024)
di: Atil, Berk, et al.
Pubblicazione: (2024)
Retrieval-augmented generation in multilingual settings
di: Chirkova, Nadezhda, et al.
Pubblicazione: (2024)
di: Chirkova, Nadezhda, et al.
Pubblicazione: (2024)
Sentiment analysis and random forest to classify LLM versus human source applied to Scientific Texts
di: Sanchez-Medina, Javier J.
Pubblicazione: (2024)
di: Sanchez-Medina, Javier J.
Pubblicazione: (2024)
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
di: Tie, Guiyao, et al.
Pubblicazione: (2025)
di: Tie, Guiyao, et al.
Pubblicazione: (2025)
The Two Sides of the Coin: Hallucination Generation and Detection with LLMs as Evaluators for LLMs
di: Bui, Anh Thu Maria, et al.
Pubblicazione: (2024)
di: Bui, Anh Thu Maria, et al.
Pubblicazione: (2024)
Beyond LLM-as-a-Judge: Deterministic Metrics for Multilingual Generative Text Evaluation
di: Alam, Firoj, et al.
Pubblicazione: (2026)
di: Alam, Firoj, et al.
Pubblicazione: (2026)
A validity-guided workflow for robust large language model research in psychology
di: Lin, Zhicheng
Pubblicazione: (2025)
di: Lin, Zhicheng
Pubblicazione: (2025)
Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs
di: Papi, Sara, et al.
Pubblicazione: (2025)
di: Papi, Sara, et al.
Pubblicazione: (2025)
Ranking LLMs by compression
di: Guo, Peijia, et al.
Pubblicazione: (2024)
di: Guo, Peijia, et al.
Pubblicazione: (2024)
Densing Law of LLMs
di: Xiao, Chaojun, et al.
Pubblicazione: (2024)
di: Xiao, Chaojun, et al.
Pubblicazione: (2024)
The Colorful Future of LLMs: Evaluating and Improving LLMs as Emotional Supporters for Queer Youth
di: Lissak, Shir, et al.
Pubblicazione: (2024)
di: Lissak, Shir, et al.
Pubblicazione: (2024)
Are LLMs Effective Negotiators? Systematic Evaluation of the Multifaceted Capabilities of LLMs in Negotiation Dialogues
di: Kwon, Deuksin, et al.
Pubblicazione: (2024)
di: Kwon, Deuksin, et al.
Pubblicazione: (2024)
Benchmark of stylistic variation in LLM-generated texts
di: Milička, Jiří, et al.
Pubblicazione: (2025)
di: Milička, Jiří, et al.
Pubblicazione: (2025)
Are generative AI text annotations systematically biased?
di: Stolwijk, Sjoerd B., et al.
Pubblicazione: (2025)
di: Stolwijk, Sjoerd B., et al.
Pubblicazione: (2025)
Gender Bias in LLM-generated Interview Responses
di: Kong, Haein, et al.
Pubblicazione: (2024)
di: Kong, Haein, et al.
Pubblicazione: (2024)
Raply: A profanity-mitigated rap generator
di: Bendali, Omar Manil, et al.
Pubblicazione: (2024)
di: Bendali, Omar Manil, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS)
di: Awad, Samer, et al.
Pubblicazione: (2026) -
Assessing the Performance of Human-Capable LLMs -- Are LLMs Coming for Your Job?
di: Mavi, John, et al.
Pubblicazione: (2024) -
Exploring the psychology of LLMs' Moral and Legal Reasoning
di: Almeida, Guilherme F. C. F., et al.
Pubblicazione: (2023) -
Evaluating book summaries from internal knowledge in Large Language Models: a cross-model and semantic consistency approach
di: Coronado-Blázquez, Javier
Pubblicazione: (2025) -
From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law
di: Mavi, John, et al.
Pubblicazione: (2025)