Gespeichert in:
| Hauptverfasser: | Sun, Simeng, Hsieh, Cheng-Ping |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2410.12292 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
L0-Reasoning Bench: Evaluating Procedural Correctness in Language Models via Simple Program Execution
von: Sun, Simeng, et al.
Veröffentlicht: (2025)
von: Sun, Simeng, et al.
Veröffentlicht: (2025)
How much do language models memorize?
von: Morris, John X., et al.
Veröffentlicht: (2025)
von: Morris, John X., et al.
Veröffentlicht: (2025)
An empirical study on the limitation of Transformers in program trace generation
von: Sun, Simeng
Veröffentlicht: (2025)
von: Sun, Simeng
Veröffentlicht: (2025)
RULER: What's the Real Context Size of Your Long-Context Language Models?
von: Hsieh, Cheng-Ping, et al.
Veröffentlicht: (2024)
von: Hsieh, Cheng-Ping, et al.
Veröffentlicht: (2024)
Clinical ModernBERT: An efficient and long context encoder for biomedical text
von: Lee, Simon A., et al.
Veröffentlicht: (2025)
von: Lee, Simon A., et al.
Veröffentlicht: (2025)
Cartridges: Lightweight and general-purpose long context representations via self-study
von: Eyuboglu, Sabri, et al.
Veröffentlicht: (2025)
von: Eyuboglu, Sabri, et al.
Veröffentlicht: (2025)
Injecting Wiktionary to improve token-level contextual representations using contrastive learning
von: Mosolova, Anna, et al.
Veröffentlicht: (2024)
von: Mosolova, Anna, et al.
Veröffentlicht: (2024)
SWAN-GPT: An Efficient and Scalable Approach for Long-Context Language Modeling
von: Puvvada, Krishna C., et al.
Veröffentlicht: (2025)
von: Puvvada, Krishna C., et al.
Veröffentlicht: (2025)
Suri: Multi-constraint Instruction Following for Long-form Text Generation
von: Pham, Chau Minh, et al.
Veröffentlicht: (2024)
von: Pham, Chau Minh, et al.
Veröffentlicht: (2024)
How much do LLMs learn from negative examples?
von: Hamdan, Shadi, et al.
Veröffentlicht: (2025)
von: Hamdan, Shadi, et al.
Veröffentlicht: (2025)
Interpreting the structure of multi-object representations in vision encoders
von: Khajuria, Tarun, et al.
Veröffentlicht: (2024)
von: Khajuria, Tarun, et al.
Veröffentlicht: (2024)
Exploring the encoding of linguistic representations in the Fully-Connected Layer of generative CNNs for Speech
von: Šegedin, Bruno Ferenc, et al.
Veröffentlicht: (2025)
von: Šegedin, Bruno Ferenc, et al.
Veröffentlicht: (2025)
How much reliable is ChatGPT's prediction on Information Extraction under Input Perturbations?
von: Mondal, Ishani, et al.
Veröffentlicht: (2024)
von: Mondal, Ishani, et al.
Veröffentlicht: (2024)
Turbulence-like 5/3 spectral scaling in contextual representations of language as a complex system
von: Yang, Zhongxin, et al.
Veröffentlicht: (2026)
von: Yang, Zhongxin, et al.
Veröffentlicht: (2026)
How much speech data is necessary for ASR in African languages? An evaluation of data scaling in Kinyarwanda and Kikuyu
von: Akera, Benjamin, et al.
Veröffentlicht: (2025)
von: Akera, Benjamin, et al.
Veröffentlicht: (2025)
CLIPPER: Compression enables long-context synthetic data generation
von: Pham, Chau Minh, et al.
Veröffentlicht: (2025)
von: Pham, Chau Minh, et al.
Veröffentlicht: (2025)
Can LLMs reason over extended multilingual contexts? Towards long-context evaluation beyond retrieval and haystacks
von: Hengle, Amey, et al.
Veröffentlicht: (2025)
von: Hengle, Amey, et al.
Veröffentlicht: (2025)
TopicGPT: A Prompt-based Topic Modeling Framework
von: Pham, Chau Minh, et al.
Veröffentlicht: (2023)
von: Pham, Chau Minh, et al.
Veröffentlicht: (2023)
Nationality encoding in language model hidden states: Probing culturally differentiated representations in persona-conditioned academic text
von: Jackson, Paul, et al.
Veröffentlicht: (2026)
von: Jackson, Paul, et al.
Veröffentlicht: (2026)
Does quantization affect models' performance on long-context tasks?
von: Mekala, Anmol, et al.
Veröffentlicht: (2025)
von: Mekala, Anmol, et al.
Veröffentlicht: (2025)
BRoverbs -- Measuring how much LLMs understand Portuguese proverbs
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
Comparing representations of long clinical texts for the task of patient note-identification
von: Alsaidi, Safa, et al.
Veröffentlicht: (2025)
von: Alsaidi, Safa, et al.
Veröffentlicht: (2025)
One ruler to measure them all: Benchmarking multilingual long-context language models
von: Kim, Yekyung, et al.
Veröffentlicht: (2025)
von: Kim, Yekyung, et al.
Veröffentlicht: (2025)
Guideline Learning for In-context Information Extraction
von: Pang, Chaoxu, et al.
Veröffentlicht: (2023)
von: Pang, Chaoxu, et al.
Veröffentlicht: (2023)
Machine learning and emoji prediction: How much accuracy can MARBERT achieve?
von: Shormani, Mohammed Q., et al.
Veröffentlicht: (2026)
von: Shormani, Mohammed Q., et al.
Veröffentlicht: (2026)
Ada-LEval: Evaluating long-context LLMs with length-adaptable benchmarks
von: Wang, Chonghua, et al.
Veröffentlicht: (2024)
von: Wang, Chonghua, et al.
Veröffentlicht: (2024)
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages?
von: Aravapalli, Akhilesh, et al.
Veröffentlicht: (2024)
von: Aravapalli, Akhilesh, et al.
Veröffentlicht: (2024)
Failure of contextual invariance in large language models
von: Kumar, Sagar, et al.
Veröffentlicht: (2026)
von: Kumar, Sagar, et al.
Veröffentlicht: (2026)
Large language models reorganize representational geometry during in-context learning
von: Xiong, Hua-Dong, et al.
Veröffentlicht: (2026)
von: Xiong, Hua-Dong, et al.
Veröffentlicht: (2026)
Semantics or spelling? Probing contextual word embeddings with orthographic noise
von: Matthews, Jacob A., et al.
Veröffentlicht: (2024)
von: Matthews, Jacob A., et al.
Veröffentlicht: (2024)
Extracting domain-specific terms using contextual word embeddings
von: Repar, Andraž, et al.
Veröffentlicht: (2025)
von: Repar, Andraž, et al.
Veröffentlicht: (2025)
HYBRIDMIND: Meta Selection of Natural Language and Symbolic Language for Enhanced LLM Reasoning
von: Han, Simeng, et al.
Veröffentlicht: (2024)
von: Han, Simeng, et al.
Veröffentlicht: (2024)
How much does context affect the accuracy of AI health advice?
von: Garg, Prashant, et al.
Veröffentlicht: (2025)
von: Garg, Prashant, et al.
Veröffentlicht: (2025)
Adjoint sharding for very long context training of state space models
von: Xu, Xingzi, et al.
Veröffentlicht: (2025)
von: Xu, Xingzi, et al.
Veröffentlicht: (2025)
One Thousand and One Pairs: A "novel" challenge for long-context language models
von: Karpinska, Marzena, et al.
Veröffentlicht: (2024)
von: Karpinska, Marzena, et al.
Veröffentlicht: (2024)
AIC CTU@FEVER 8: On-premise fact checking through long context RAG
von: Ullrich, Herbert, et al.
Veröffentlicht: (2025)
von: Ullrich, Herbert, et al.
Veröffentlicht: (2025)
so much depends / upon / a whitespace: Why Whitespace Matters for Poets and LLMs
von: Bhyravajjula, Sriharsh, et al.
Veröffentlicht: (2025)
von: Bhyravajjula, Sriharsh, et al.
Veröffentlicht: (2025)
Phase transition on a context-sensitive random language model with short range interactions
von: Toji, Yuma, et al.
Veröffentlicht: (2026)
von: Toji, Yuma, et al.
Veröffentlicht: (2026)
SugarTextNet: A Transformer-Based Framework for Detecting Sugar Dating-Related Content on Social Media with Context-Aware Focal Loss
von: Wang, Lionel Z., et al.
Veröffentlicht: (2025)
von: Wang, Lionel Z., et al.
Veröffentlicht: (2025)
A graph-based analysis of semantic types and coercion in contextualized word embeddings
von: Chen, Long, et al.
Veröffentlicht: (2026)
von: Chen, Long, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
L0-Reasoning Bench: Evaluating Procedural Correctness in Language Models via Simple Program Execution
von: Sun, Simeng, et al.
Veröffentlicht: (2025) -
How much do language models memorize?
von: Morris, John X., et al.
Veröffentlicht: (2025) -
An empirical study on the limitation of Transformers in program trace generation
von: Sun, Simeng
Veröffentlicht: (2025) -
RULER: What's the Real Context Size of Your Long-Context Language Models?
von: Hsieh, Cheng-Ping, et al.
Veröffentlicht: (2024) -
Clinical ModernBERT: An efficient and long context encoder for biomedical text
von: Lee, Simon A., et al.
Veröffentlicht: (2025)