Evaluating the Performance of Large Language Models for SDG Mapping (Technical Report)
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yin, Hui, Aryani, Amir, Nambiar, Nakul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Technical Report on classification of literature related to children speech disorder
von: Wang, Ziang, et al.
Veröffentlicht: (2025)
von: Wang, Ziang, et al.
Veröffentlicht: (2025)
Arcee Trinity Large Technical Report
von: Singh, Varun, et al.
Veröffentlicht: (2026)
von: Singh, Varun, et al.
Veröffentlicht: (2026)
Combining Knowledge Graphs and Large Language Models
von: Kau, Amanda, et al.
Veröffentlicht: (2024)
von: Kau, Amanda, et al.
Veröffentlicht: (2024)
Evaluating the Performance of Large Language Models via Debates
von: Moniri, Behrad, et al.
Veröffentlicht: (2024)
von: Moniri, Behrad, et al.
Veröffentlicht: (2024)
Comparative Performance Evaluation of Large Language Models for Extracting Molecular Interactions and Pathway Knowledge
von: Park, Gilchan, et al.
Veröffentlicht: (2023)
von: Park, Gilchan, et al.
Veröffentlicht: (2023)
Performance Law of Large Language Models
von: Wu, Chuhan, et al.
Veröffentlicht: (2024)
von: Wu, Chuhan, et al.
Veröffentlicht: (2024)
Technical Report: Small Language Model for Japanese Clinical and Medicine
von: Watanabe, Shogo
Veröffentlicht: (2024)
von: Watanabe, Shogo
Veröffentlicht: (2024)
Compositional Instruction Following with Language Models and Reinforcement Learning
von: Cohen, Vanya, et al.
Veröffentlicht: (2025)
von: Cohen, Vanya, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models with fmeval
von: Schwöbel, Pola, et al.
Veröffentlicht: (2024)
von: Schwöbel, Pola, et al.
Veröffentlicht: (2024)
Data Augmentations for Improved (Large) Language Model Generalization
von: Feder, Amir, et al.
Veröffentlicht: (2023)
von: Feder, Amir, et al.
Veröffentlicht: (2023)
Evaluating Large Language Models at Evaluating Instruction Following
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2023)
Latent Performance Profiling of Large Language Models
von: Chakraborty, Tanmoy, et al.
Veröffentlicht: (2026)
von: Chakraborty, Tanmoy, et al.
Veröffentlicht: (2026)
Baichuan Alignment Technical Report
von: Lin, Mingan, et al.
Veröffentlicht: (2024)
von: Lin, Mingan, et al.
Veröffentlicht: (2024)
Megrez-Omni Technical Report
von: Li, Boxun, et al.
Veröffentlicht: (2025)
von: Li, Boxun, et al.
Veröffentlicht: (2025)
Long Context RAG Performance of Large Language Models
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
MapIQ: Evaluating Multimodal Large Language Models for Map Question Answering
von: Srivastava, Varun, et al.
Veröffentlicht: (2025)
von: Srivastava, Varun, et al.
Veröffentlicht: (2025)
Large Language Model Agent as a Mechanical Designer
von: Jadhav, Yayati, et al.
Veröffentlicht: (2024)
von: Jadhav, Yayati, et al.
Veröffentlicht: (2024)
AMGPT: a Large Language Model for Contextual Querying in Additive Manufacturing
von: Chandrasekhar, Achuth, et al.
Veröffentlicht: (2024)
von: Chandrasekhar, Achuth, et al.
Veröffentlicht: (2024)
Towards Modeling Learner Performance with Large Language Models
von: Neshaei, Seyed Parsa, et al.
Veröffentlicht: (2024)
von: Neshaei, Seyed Parsa, et al.
Veröffentlicht: (2024)
Equivalent Linear Mappings of Large Language Models
von: Golden, James R.
Veröffentlicht: (2025)
von: Golden, James R.
Veröffentlicht: (2025)
Bias in Large Language Models: Origin, Evaluation, and Mitigation
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
MANGO: A Benchmark for Evaluating Mapping and Navigation Abilities of Large Language Models
von: Ding, Peng, et al.
Veröffentlicht: (2024)
von: Ding, Peng, et al.
Veröffentlicht: (2024)
Scaling Laws for Downstream Task Performance of Large Language Models
von: Isik, Berivan, et al.
Veröffentlicht: (2024)
von: Isik, Berivan, et al.
Veröffentlicht: (2024)
InnerQ: Hardware-Aware Tuning-Free Quantization of KV Cache for Large Language Models
von: Hosseini, Sayed Mohammadreza Tayaranian, et al.
Veröffentlicht: (2026)
von: Hosseini, Sayed Mohammadreza Tayaranian, et al.
Veröffentlicht: (2026)
KAT-Coder-V2 Technical Report
von: Li, Fengxiang, et al.
Veröffentlicht: (2026)
von: Li, Fengxiang, et al.
Veröffentlicht: (2026)
Large Language and Reasoning Models are Shallow Disjunctive Reasoners
von: Khalid, Irtaza, et al.
Veröffentlicht: (2025)
von: Khalid, Irtaza, et al.
Veröffentlicht: (2025)
Revisiting Dynamic Evaluation: Online Adaptation for Large Language Models
von: Rannen-Triki, Amal, et al.
Veröffentlicht: (2024)
von: Rannen-Triki, Amal, et al.
Veröffentlicht: (2024)
CEB: Compositional Evaluation Benchmark for Fairness in Large Language Models
von: Wang, Song, et al.
Veröffentlicht: (2024)
von: Wang, Song, et al.
Veröffentlicht: (2024)
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models
von: Qi, Zhen, et al.
Veröffentlicht: (2024)
von: Qi, Zhen, et al.
Veröffentlicht: (2024)
Bayesian Evaluation of Large Language Model Behavior
von: Longjohn, Rachel, et al.
Veröffentlicht: (2025)
von: Longjohn, Rachel, et al.
Veröffentlicht: (2025)
Merging Language and Domain Specific Models: The Impact on Technical Vocabulary Acquisition
von: Rousset, Thibault, et al.
Veröffentlicht: (2025)
von: Rousset, Thibault, et al.
Veröffentlicht: (2025)
H2O-Danube3 Technical Report
von: Pfeiffer, Pascal, et al.
Veröffentlicht: (2024)
von: Pfeiffer, Pascal, et al.
Veröffentlicht: (2024)
Self-Directed Synthetic Dialogues and Revisions Technical Report
von: Lambert, Nathan, et al.
Veröffentlicht: (2024)
von: Lambert, Nathan, et al.
Veröffentlicht: (2024)
Mapping Faithful Reasoning in Language Models
von: Li, Jiazheng, et al.
Veröffentlicht: (2025)
von: Li, Jiazheng, et al.
Veröffentlicht: (2025)
Confidence-Aware Sub-Structure Beam Search (CABS): Mitigating Hallucination in Structured Data Generation with Large Language Models
von: Wei, Chengwei, et al.
Veröffentlicht: (2024)
von: Wei, Chengwei, et al.
Veröffentlicht: (2024)
Flextron: Many-in-One Flexible Large Language Model
von: Cai, Ruisi, et al.
Veröffentlicht: (2024)
von: Cai, Ruisi, et al.
Veröffentlicht: (2024)
Large Language Models in Fire Engineering: An Examination of Technical Questions Against Domain Knowledge
von: Hostetter, Haley, et al.
Veröffentlicht: (2024)
von: Hostetter, Haley, et al.
Veröffentlicht: (2024)
AMALIA Technical Report: A Fully Open Source Large Language Model for European Portuguese
von: Simplício, Afonso, et al.
Veröffentlicht: (2026)
von: Simplício, Afonso, et al.
Veröffentlicht: (2026)
MEIT: Multimodal Electrocardiogram Instruction Tuning on Large Language Models for Report Generation
von: Wan, Zhongwei, et al.
Veröffentlicht: (2024)
von: Wan, Zhongwei, et al.
Veröffentlicht: (2024)
OCEAN: Offline Chain-of-thought Evaluation and Alignment in Large Language Models
von: Wu, Junda, et al.
Veröffentlicht: (2024)
von: Wu, Junda, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Technical Report on classification of literature related to children speech disorder
von: Wang, Ziang, et al.
Veröffentlicht: (2025) -
Arcee Trinity Large Technical Report
von: Singh, Varun, et al.
Veröffentlicht: (2026) -
Combining Knowledge Graphs and Large Language Models
von: Kau, Amanda, et al.
Veröffentlicht: (2024) -
Evaluating the Performance of Large Language Models via Debates
von: Moniri, Behrad, et al.
Veröffentlicht: (2024) -
Comparative Performance Evaluation of Large Language Models for Extracting Molecular Interactions and Pathway Knowledge
von: Park, Gilchan, et al.
Veröffentlicht: (2023)