Gespeichert in:
| Hauptverfasser: | Olowe, Emmanuel A., Chitnis, Danial |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2503.02123 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LABIIUM: AI-Enhanced Zero-configuration Measurement Automation System
von: Olowe, Emmanuel A., et al.
Veröffentlicht: (2024)
von: Olowe, Emmanuel A., et al.
Veröffentlicht: (2024)
Policy-Governed LLM Routing with Intent Matching for Instrument Laboratories
von: Olowe, Emmanuel A., et al.
Veröffentlicht: (2026)
von: Olowe, Emmanuel A., et al.
Veröffentlicht: (2026)
LLM-based AI Agent for Sizing of Analog and Mixed Signal Circuit
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
How to Measure the Intelligence of Large Language Models?
von: Körber, Nils, et al.
Veröffentlicht: (2024)
von: Körber, Nils, et al.
Veröffentlicht: (2024)
GenePlan: Evolving Better Generalized PDDL Plans using Large Language Models
von: Murray, Andrew, et al.
Veröffentlicht: (2026)
von: Murray, Andrew, et al.
Veröffentlicht: (2026)
OFCnetLLM: Large Language Model for Network Monitoring and Alertness
von: Yoon, Hong-Jun, et al.
Veröffentlicht: (2025)
von: Yoon, Hong-Jun, et al.
Veröffentlicht: (2025)
PerMedCQA: Benchmarking Large Language Models on Medical Consumer Question Answering in Persian Language
von: Jamali, Naghmeh, et al.
Veröffentlicht: (2025)
von: Jamali, Naghmeh, et al.
Veröffentlicht: (2025)
Quantifying Laziness, Decoding Suboptimality, and Context Degradation in Large Language Models
von: Ma, Yiqing, et al.
Veröffentlicht: (2025)
von: Ma, Yiqing, et al.
Veröffentlicht: (2025)
Problems With Large Language Models for Learner Modelling: Why LLMs Alone Fall Short for Responsible Tutoring in K--12 Education
von: Hooshyar, Danial, et al.
Veröffentlicht: (2025)
von: Hooshyar, Danial, et al.
Veröffentlicht: (2025)
Measuring and Analyzing Intelligence via Contextual Uncertainty in Large Language Models using Information-Theoretic Metrics
von: Shim, Jae Wan
Veröffentlicht: (2025)
von: Shim, Jae Wan
Veröffentlicht: (2025)
Quantifying Self-Preservation Bias in Large Language Models
von: Migliarini, Matteo, et al.
Veröffentlicht: (2026)
von: Migliarini, Matteo, et al.
Veröffentlicht: (2026)
Quantifying Fidelity: A Decisive Feature Approach to Comparing Synthetic and Real Imagery
von: Safaei, Danial, et al.
Veröffentlicht: (2025)
von: Safaei, Danial, et al.
Veröffentlicht: (2025)
Polymetis:Large Language Modeling for Multiple Material Domains
von: Huang, Chao, et al.
Veröffentlicht: (2024)
von: Huang, Chao, et al.
Veröffentlicht: (2024)
DOCS: Quantifying Weight Similarity for Deeper Insights into Large Language Models
von: Min, Zeping, et al.
Veröffentlicht: (2025)
von: Min, Zeping, et al.
Veröffentlicht: (2025)
"Lost-in-the-Later": Framework for Quantifying Contextual Grounding in Large Language Models
von: Tao, Yufei, et al.
Veröffentlicht: (2025)
von: Tao, Yufei, et al.
Veröffentlicht: (2025)
VibeCheck: Discover and Quantify Qualitative Differences in Large Language Models
von: Dunlap, Lisa, et al.
Veröffentlicht: (2024)
von: Dunlap, Lisa, et al.
Veröffentlicht: (2024)
xOffense: An Autonomous Multi-Agent Framework for Penetration Testing with Domain-Adapted Large Language Models
von: Luong, Phung Duc, et al.
Veröffentlicht: (2025)
von: Luong, Phung Duc, et al.
Veröffentlicht: (2025)
BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains
von: Labrak, Yanis, et al.
Veröffentlicht: (2024)
von: Labrak, Yanis, et al.
Veröffentlicht: (2024)
EEspice: A Modular Circuit Simulation Platform with Parallel Device Model Evaluation via Graph Coloring
von: Bao, Xuanhao, et al.
Veröffentlicht: (2026)
von: Bao, Xuanhao, et al.
Veröffentlicht: (2026)
Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2026)
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2026)
Adversarial Moral Stress Testing of Large Language Models
von: Jamshidi, Saeid, et al.
Veröffentlicht: (2026)
von: Jamshidi, Saeid, et al.
Veröffentlicht: (2026)
Monotropic Artificial Intelligence: Toward a Cognitive Taxonomy of Domain-Specialized Language Models
von: Filho, Antonio de Sousa Leitão, et al.
Veröffentlicht: (2026)
von: Filho, Antonio de Sousa Leitão, et al.
Veröffentlicht: (2026)
Large Language Models as Planning Domain Generators
von: Oswald, James, et al.
Veröffentlicht: (2024)
von: Oswald, James, et al.
Veröffentlicht: (2024)
ELATE: Evolutionary Language model for Automated Time-series Engineering
von: Murray, Andrew, et al.
Veröffentlicht: (2025)
von: Murray, Andrew, et al.
Veröffentlicht: (2025)
Large Language Model's Multi-Capability Alignment in Biomedical Domain
von: Wu, Wentao, et al.
Veröffentlicht: (2025)
von: Wu, Wentao, et al.
Veröffentlicht: (2025)
Quantifying Label-Induced Bias in Large Language Model Self- and Cross-Evaluations
von: Saraf, Muskan, et al.
Veröffentlicht: (2025)
von: Saraf, Muskan, et al.
Veröffentlicht: (2025)
Jailbreaking Large Vision Language Models in Intelligent Transportation Systems
von: Das, Badhan Chandra, et al.
Veröffentlicht: (2025)
von: Das, Badhan Chandra, et al.
Veröffentlicht: (2025)
FRoG: Evaluating Fuzzy Reasoning of Generalized Quantifiers in Large Language Models
von: Li, Yiyuan, et al.
Veröffentlicht: (2024)
von: Li, Yiyuan, et al.
Veröffentlicht: (2024)
Quantifying Semantic Emergence in Language Models
von: Chen, Hang, et al.
Veröffentlicht: (2024)
von: Chen, Hang, et al.
Veröffentlicht: (2024)
Large Language Models in the Travel Domain: An Industrial Experience
von: Di Meglio, Sergio, et al.
Veröffentlicht: (2025)
von: Di Meglio, Sergio, et al.
Veröffentlicht: (2025)
Assessing the Capability of Large Language Models for Domain-Specific Ontology Generation
von: Lippolis, Anna Sofia, et al.
Veröffentlicht: (2025)
von: Lippolis, Anna Sofia, et al.
Veröffentlicht: (2025)
Integrating Domain Knowledge into Process Discovery Using Large Language Models
von: Norouzifar, Ali, et al.
Veröffentlicht: (2025)
von: Norouzifar, Ali, et al.
Veröffentlicht: (2025)
Metamorphic Testing of Large Language Models for Natural Language Processing
von: Cho, Steven, et al.
Veröffentlicht: (2025)
von: Cho, Steven, et al.
Veröffentlicht: (2025)
Credal Transformer: A Principled Approach for Quantifying and Mitigating Hallucinations in Large Language Models
von: Ji, Shihao, et al.
Veröffentlicht: (2025)
von: Ji, Shihao, et al.
Veröffentlicht: (2025)
Large Language Models and Scientific Discourse: Where's the Intelligence?
von: Collins, Harry, et al.
Veröffentlicht: (2026)
von: Collins, Harry, et al.
Veröffentlicht: (2026)
Structured Cognitive Loop for Behavioral Intelligence in Large Language Model Agents
von: Kim, Myung Ho
Veröffentlicht: (2025)
von: Kim, Myung Ho
Veröffentlicht: (2025)
DepthCharge: A Domain-Agnostic Framework for Measuring Depth-Dependent Knowledge in Large Language Models
von: Sheppert, Alexander
Veröffentlicht: (2026)
von: Sheppert, Alexander
Veröffentlicht: (2026)
Bayesian Orchestration of Multi-LLM Agents for Cost-Aware Sequential Decision-Making
von: Amin, Danial
Veröffentlicht: (2026)
von: Amin, Danial
Veröffentlicht: (2026)
Quantifying and Understanding Uncertainty in Large Reasoning Models
von: Li, Yangyi, et al.
Veröffentlicht: (2026)
von: Li, Yangyi, et al.
Veröffentlicht: (2026)
Quantifying Hallucinations in Language Language Models on Medical Textbooks
von: Colelough, Brandon C., et al.
Veröffentlicht: (2026)
von: Colelough, Brandon C., et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LABIIUM: AI-Enhanced Zero-configuration Measurement Automation System
von: Olowe, Emmanuel A., et al.
Veröffentlicht: (2024) -
Policy-Governed LLM Routing with Intent Matching for Instrument Laboratories
von: Olowe, Emmanuel A., et al.
Veröffentlicht: (2026) -
LLM-based AI Agent for Sizing of Analog and Mixed Signal Circuit
von: Liu, Chang, et al.
Veröffentlicht: (2025) -
How to Measure the Intelligence of Large Language Models?
von: Körber, Nils, et al.
Veröffentlicht: (2024) -
GenePlan: Evolving Better Generalized PDDL Plans using Large Language Models
von: Murray, Andrew, et al.
Veröffentlicht: (2026)