Evaluating Pixel Language Models on Non-Standardized Languages
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Muñoz-Ortiz, Alberto, Blaschke, Verena, Plank, Barbara |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
von: Bayram, M. Ali, et al.
Veröffentlicht: (2024)
von: Bayram, M. Ali, et al.
Veröffentlicht: (2024)
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference
von: Liu, Aiwei, et al.
Veröffentlicht: (2025)
von: Liu, Aiwei, et al.
Veröffentlicht: (2025)
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023)
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023)
Nested Named Entity Recognition as Single-Pass Sequence Labeling
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2025)
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2025)
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
von: Weigang, Li, et al.
Veröffentlicht: (2025)
von: Weigang, Li, et al.
Veröffentlicht: (2025)
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
von: Yao, Ben, et al.
Veröffentlicht: (2025)
von: Yao, Ben, et al.
Veröffentlicht: (2025)
The Superalignment of Superhuman Intelligence with Large Language Models
von: Huang, Minlie, et al.
Veröffentlicht: (2024)
von: Huang, Minlie, et al.
Veröffentlicht: (2024)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
Exploring State Tracking Capabilities of Large Language Models
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
Distilling Large Language Models for Efficient Clinical Information Extraction
von: Vedula, Karthik S., et al.
Veröffentlicht: (2024)
von: Vedula, Karthik S., et al.
Veröffentlicht: (2024)
A Survey of Text Watermarking in the Era of Large Language Models
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
Adaptive Steering and Remasking for Safe Generation in Diffusion Language Models
von: Lee, Yejin, et al.
Veröffentlicht: (2026)
von: Lee, Yejin, et al.
Veröffentlicht: (2026)
Towards Effective and Efficient Continual Pre-training of Large Language Models
von: Chen, Jie, et al.
Veröffentlicht: (2024)
von: Chen, Jie, et al.
Veröffentlicht: (2024)
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
von: Park, Seungcheol, et al.
Veröffentlicht: (2023)
von: Park, Seungcheol, et al.
Veröffentlicht: (2023)
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
von: Fang, Xi, et al.
Veröffentlicht: (2024)
von: Fang, Xi, et al.
Veröffentlicht: (2024)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
Unveiling Attractor Cycles in Large Language Models: A Dynamical Systems View of Successive Paraphrasing
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
Trusted Uncertainty in Large Language Models: A Unified Framework for Confidence Calibration and Risk-Controlled Refusal
von: Oehri, Markus, et al.
Veröffentlicht: (2025)
von: Oehri, Markus, et al.
Veröffentlicht: (2025)
"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations
von: Choi, Jonghyeon, et al.
Veröffentlicht: (2025)
von: Choi, Jonghyeon, et al.
Veröffentlicht: (2025)
A Survey on Natural Language Counterfactual Generation
von: Wang, Yongjie, et al.
Veröffentlicht: (2024)
von: Wang, Yongjie, et al.
Veröffentlicht: (2024)
Math Natural Language Inference: this should be easy!
von: de Paiva, Valeria, et al.
Veröffentlicht: (2025)
von: de Paiva, Valeria, et al.
Veröffentlicht: (2025)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
von: Pranida, Salsabila Zahirah, et al.
Veröffentlicht: (2025)
von: Pranida, Salsabila Zahirah, et al.
Veröffentlicht: (2025)
Evaluating Input Feature Explanations through a Unified Diagnostic Evaluation Framework
von: Sun, Jingyi, et al.
Veröffentlicht: (2024)
von: Sun, Jingyi, et al.
Veröffentlicht: (2024)
Fine-tuning of Large Language Models for Constituency Parsing Using a Sequence to Sequence Approach
von: Delgado, Francisco Jose Cortes, et al.
Veröffentlicht: (2025)
von: Delgado, Francisco Jose Cortes, et al.
Veröffentlicht: (2025)
The Unlikely Duel: Evaluating Creative Writing in LLMs through a Unique Scenario
von: Gómez-Rodríguez, Carlos, et al.
Veröffentlicht: (2024)
von: Gómez-Rodríguez, Carlos, et al.
Veröffentlicht: (2024)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
von: Miliani, Martina, et al.
Veröffentlicht: (2025)
von: Miliani, Martina, et al.
Veröffentlicht: (2025)
Model Misalignment and Language Change: Traces of AI-Associated Language in Unscripted Spoken English
von: Anderson, Bryce, et al.
Veröffentlicht: (2025)
von: Anderson, Bryce, et al.
Veröffentlicht: (2025)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
von: Yang, Yibo
Veröffentlicht: (2025)
von: Yang, Yibo
Veröffentlicht: (2025)
Profiling German Text Simplification with Interpretable Model-Fingerprints
von: Klöser, Lars, et al.
Veröffentlicht: (2026)
von: Klöser, Lars, et al.
Veröffentlicht: (2026)
On the Robustness of Document-Level Relation Extraction Models to Entity Name Variations
von: Meng, Shiao, et al.
Veröffentlicht: (2024)
von: Meng, Shiao, et al.
Veröffentlicht: (2024)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
Large Language Models for Propaganda Span Annotation
von: Hasanain, Maram, et al.
Veröffentlicht: (2023)
von: Hasanain, Maram, et al.
Veröffentlicht: (2023)
DYNAMICQA: Tracing Internal Knowledge Conflicts in Language Models
von: Marjanović, Sara Vera, et al.
Veröffentlicht: (2024)
von: Marjanović, Sara Vera, et al.
Veröffentlicht: (2024)
A Comprehensive Survey of Compression Algorithms for Language Models
von: Park, Seungcheol, et al.
Veröffentlicht: (2024)
von: Park, Seungcheol, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
von: Bayram, M. Ali, et al.
Veröffentlicht: (2024) -
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference
von: Liu, Aiwei, et al.
Veröffentlicht: (2025) -
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023) -
Nested Named Entity Recognition as Single-Pass Sequence Labeling
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2025) -
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
von: Pan, Leyi, et al.
Veröffentlicht: (2025)