Gespeichert in:
| Hauptverfasser: | Martínez, Héctor Javier Vázquez, Yang, Charles |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.28616 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating Neural Language Models as Cognitive Models of Language Acquisition
von: Martínez, Héctor Javier Vázquez, et al.
Veröffentlicht: (2023)
von: Martínez, Héctor Javier Vázquez, et al.
Veröffentlicht: (2023)
findsylls: A Language-Agnostic Toolkit for Syllable-Level Speech Tokenization and Embedding
von: Martínez, Héctor Javier Vázquez
Veröffentlicht: (2026)
von: Martínez, Héctor Javier Vázquez
Veröffentlicht: (2026)
Measuring Free-Form Decision-Making Inconsistency of Language Models in Military Crisis Simulations
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2024)
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2024)
FactReasoner: A Probabilistic Approach to Long-Form Factuality Assessment for Large Language Models
von: Marinescu, Radu, et al.
Veröffentlicht: (2025)
von: Marinescu, Radu, et al.
Veröffentlicht: (2025)
Use of Retrieval-Augmented Large Language Model Agent for Long-Form COVID-19 Fact-Checking
von: Huang, Jingyi, et al.
Veröffentlicht: (2025)
von: Huang, Jingyi, et al.
Veröffentlicht: (2025)
Abductive Reasoning with Syllogistic Forms in Large Language Models
von: Abe, Hirohiko, et al.
Veröffentlicht: (2026)
von: Abe, Hirohiko, et al.
Veröffentlicht: (2026)
FactCorrector: A Graph-Inspired Approach to Long-Form Factuality Correction of Large Language Models
von: Carnerero-Cano, Javier, et al.
Veröffentlicht: (2026)
von: Carnerero-Cano, Javier, et al.
Veröffentlicht: (2026)
Is There a Case for Conversation Optimized Tokenizers in Large Language Models?
von: Ferrando, Raquel, et al.
Veröffentlicht: (2025)
von: Ferrando, Raquel, et al.
Veröffentlicht: (2025)
Functional Component Ablation Reveals Specialization Patterns in Hybrid Language Model Architectures
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
Paraphrase and Solve: Exploring and Exploiting the Impact of Surface Form on Mathematical Reasoning in Large Language Models
von: Zhou, Yue, et al.
Veröffentlicht: (2024)
von: Zhou, Yue, et al.
Veröffentlicht: (2024)
Do Large Language Models Perform the Way People Expect? Measuring the Human Generalization Function
von: Vafa, Keyon, et al.
Veröffentlicht: (2024)
von: Vafa, Keyon, et al.
Veröffentlicht: (2024)
Detecting Linguistic Indicators for Stereotype Assessment with Large Language Models
von: Görge, Rebekka, et al.
Veröffentlicht: (2025)
von: Görge, Rebekka, et al.
Veröffentlicht: (2025)
Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models
von: Feng, Yujie, et al.
Veröffentlicht: (2026)
von: Feng, Yujie, et al.
Veröffentlicht: (2026)
The Generative Energy Arena (GEA): Incorporating Energy Awareness in Large Language Model (LLM) Human Evaluations
von: Arriaga, Carlos, et al.
Veröffentlicht: (2025)
von: Arriaga, Carlos, et al.
Veröffentlicht: (2025)
From Form(s) to Meaning: Probing the Semantic Depths of Language Models Using Multisense Consistency
von: Ohmer, Xenia, et al.
Veröffentlicht: (2024)
von: Ohmer, Xenia, et al.
Veröffentlicht: (2024)
PROXYQA: An Alternative Framework for Evaluating Long-Form Text Generation with Large Language Models
von: Tan, Haochen, et al.
Veröffentlicht: (2024)
von: Tan, Haochen, et al.
Veröffentlicht: (2024)
Multiple Choice Questions: Reasoning Makes Large Language Models (LLMs) More Self-Confident, Especially When They are Wrong
von: Fu, Tairan, et al.
Veröffentlicht: (2025)
von: Fu, Tairan, et al.
Veröffentlicht: (2025)
NUMCoT: Numerals and Units of Measurement in Chain-of-Thought Reasoning using Large Language Models
von: Xu, Ancheng, et al.
Veröffentlicht: (2024)
von: Xu, Ancheng, et al.
Veröffentlicht: (2024)
Measuring and Benchmarking Large Language Models' Capabilities to Generate Persuasive Language
von: Pauli, Amalie Brogaard, et al.
Veröffentlicht: (2024)
von: Pauli, Amalie Brogaard, et al.
Veröffentlicht: (2024)
Segment-Level Diffusion: A Framework for Controllable Long-Form Generation with Diffusion Language Models
von: Zhu, Xiaochen, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaochen, et al.
Veröffentlicht: (2024)
Measuring Pragmatic Influence in Large Language Model Instructions
von: Geng, Yilin, et al.
Veröffentlicht: (2026)
von: Geng, Yilin, et al.
Veröffentlicht: (2026)
Measuring Representation Robustness in Large Language Models for Geometry
von: Jawandhia, Vedant, et al.
Veröffentlicht: (2026)
von: Jawandhia, Vedant, et al.
Veröffentlicht: (2026)
Measuring and Eliminating Refusals in Military Large Language Models
von: FitzGerald, Jack, et al.
Veröffentlicht: (2026)
von: FitzGerald, Jack, et al.
Veröffentlicht: (2026)
How Do Language Models Compose Functions?
von: Khandelwal, Apoorv, et al.
Veröffentlicht: (2025)
von: Khandelwal, Apoorv, et al.
Veröffentlicht: (2025)
Information Flow Routes: Automatically Interpreting Language Models at Scale
von: Ferrando, Javier, et al.
Veröffentlicht: (2024)
von: Ferrando, Javier, et al.
Veröffentlicht: (2024)
Function Words as Statistical Cues for Language Learning
von: Yang, Xiulin, et al.
Veröffentlicht: (2026)
von: Yang, Xiulin, et al.
Veröffentlicht: (2026)
Towards Measuring the Representation of Subjective Global Opinions in Language Models
von: Durmus, Esin, et al.
Veröffentlicht: (2023)
von: Durmus, Esin, et al.
Veröffentlicht: (2023)
Measuring what Matters: Construct Validity in Large Language Model Benchmarks
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
Atomic Calibration of LLMs in Long-Form Generations
von: Zhang, Caiqi, et al.
Veröffentlicht: (2024)
von: Zhang, Caiqi, et al.
Veröffentlicht: (2024)
The Role of Model Confidence on Bias Effects in Measured Uncertainties for Vision-Language Models
von: Liu, Xinyi, et al.
Veröffentlicht: (2025)
von: Liu, Xinyi, et al.
Veröffentlicht: (2025)
Exploring Graph Representations of Logical Forms for Language Modeling
von: Sullivan, Michael
Veröffentlicht: (2025)
von: Sullivan, Michael
Veröffentlicht: (2025)
Emotion Concepts and their Function in a Large Language Model
von: Sofroniew, Nicholas, et al.
Veröffentlicht: (2026)
von: Sofroniew, Nicholas, et al.
Veröffentlicht: (2026)
Do Influence Functions Work on Large Language Models?
von: Li, Zhe, et al.
Veröffentlicht: (2024)
von: Li, Zhe, et al.
Veröffentlicht: (2024)
The Promises and Pitfalls of Using Language Models to Measure Instruction Quality in Education
von: Xu, Paiheng, et al.
Veröffentlicht: (2024)
von: Xu, Paiheng, et al.
Veröffentlicht: (2024)
FNF: Functional Network Fingerprint for Large Language Models
von: Liu, Yiheng, et al.
Veröffentlicht: (2026)
von: Liu, Yiheng, et al.
Veröffentlicht: (2026)
Bayesian Optimization for Enhanced Language Models: Optimizing Acquisition Functions
von: Bao, Zishuo, et al.
Veröffentlicht: (2025)
von: Bao, Zishuo, et al.
Veröffentlicht: (2025)
Offline Training of Language Model Agents with Functions as Learnable Weights
von: Zhang, Shaokun, et al.
Veröffentlicht: (2024)
von: Zhang, Shaokun, et al.
Veröffentlicht: (2024)
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
von: Fan, Haozhi, et al.
Veröffentlicht: (2026)
von: Fan, Haozhi, et al.
Veröffentlicht: (2026)
The Dog the Cat Chased Stumped the Model: Measuring When Language Models Abandon Structure for Shortcuts
von: Madhusudan, Sangmitra, et al.
Veröffentlicht: (2025)
von: Madhusudan, Sangmitra, et al.
Veröffentlicht: (2025)
Click it or Leave it: Detecting and Spoiling Clickbait with Informativeness Measures and Large Language Models
von: Michaluk, Wojciech, et al.
Veröffentlicht: (2026)
von: Michaluk, Wojciech, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Evaluating Neural Language Models as Cognitive Models of Language Acquisition
von: Martínez, Héctor Javier Vázquez, et al.
Veröffentlicht: (2023) -
findsylls: A Language-Agnostic Toolkit for Syllable-Level Speech Tokenization and Embedding
von: Martínez, Héctor Javier Vázquez
Veröffentlicht: (2026) -
Measuring Free-Form Decision-Making Inconsistency of Language Models in Military Crisis Simulations
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2024) -
FactReasoner: A Probabilistic Approach to Long-Form Factuality Assessment for Large Language Models
von: Marinescu, Radu, et al.
Veröffentlicht: (2025) -
Use of Retrieval-Augmented Large Language Model Agent for Long-Form COVID-19 Fact-Checking
von: Huang, Jingyi, et al.
Veröffentlicht: (2025)