Guardado en:
| Autores principales: | Matsuzaki, Kosuke, Taniguchi, Masaya, Inui, Kentaro, Sakaguchi, Keisuke |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2402.14411 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FinchGPT: a Transformer based language model for birdsong analysis
por: Kobayashi, Kosei, et al.
Publicado: (2025)
por: Kobayashi, Kosei, et al.
Publicado: (2025)
Can Language Models Handle a Non-Gregorian Calendar? The Case of the Japanese wareki
por: Sasaki, Mutsumi, et al.
Publicado: (2025)
por: Sasaki, Mutsumi, et al.
Publicado: (2025)
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
por: Aoki, Yoichi, et al.
Publicado: (2024)
por: Aoki, Yoichi, et al.
Publicado: (2024)
LLMs Faithfully and Iteratively Compute Answers During CoT: A Systematic Analysis With Multi-step Arithmetics
por: Kudo, Keito, et al.
Publicado: (2024)
por: Kudo, Keito, et al.
Publicado: (2024)
The Curse of Popularity: Popular Entities have Catastrophic Side Effects when Deleting Knowledge from Language Models
por: Takahashi, Ryosuke, et al.
Publicado: (2024)
por: Takahashi, Ryosuke, et al.
Publicado: (2024)
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation
por: Brassard, Ana, et al.
Publicado: (2024)
por: Brassard, Ana, et al.
Publicado: (2024)
Annotating Errors in English Learners' Written Language Production: Advancing Automated Written Feedback Systems
por: Coyne, Steven, et al.
Publicado: (2025)
por: Coyne, Steven, et al.
Publicado: (2025)
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference
por: Kamoda, Go, et al.
Publicado: (2025)
por: Kamoda, Go, et al.
Publicado: (2025)
Nodes Are Early, Edges Are Late: Probing Diagram Representations in Large Vision-Language Models
por: Yoshida, Haruto, et al.
Publicado: (2026)
por: Yoshida, Haruto, et al.
Publicado: (2026)
QuranMorph: Morphologically Annotated Quranic Corpus
por: Akra, Diyam, et al.
Publicado: (2025)
por: Akra, Diyam, et al.
Publicado: (2025)
To Drop or Not to Drop? Predicting Argument Ellipsis Judgments: A Case Study in Japanese
por: Ishizuki, Yukiko, et al.
Publicado: (2024)
por: Ishizuki, Yukiko, et al.
Publicado: (2024)
Repetition Neurons: How Do Language Models Produce Repetitions?
por: Hiraoka, Tatsuya, et al.
Publicado: (2024)
por: Hiraoka, Tatsuya, et al.
Publicado: (2024)
Monotonic Representation of Numeric Properties in Language Models
por: Heinzerling, Benjamin, et al.
Publicado: (2024)
por: Heinzerling, Benjamin, et al.
Publicado: (2024)
Spelling-out is not Straightforward: LLMs' Capability of Tokenization from Token to Characters
por: Hiraoka, Tatsuya, et al.
Publicado: (2025)
por: Hiraoka, Tatsuya, et al.
Publicado: (2025)
Japanese-English Sentence Translation Exercises Dataset for Automatic Grading
por: Miura, Naoki, et al.
Publicado: (2024)
por: Miura, Naoki, et al.
Publicado: (2024)
Analyzing Feed-Forward Blocks in Transformers through the Lens of Attention Maps
por: Kobayashi, Goro, et al.
Publicado: (2023)
por: Kobayashi, Goro, et al.
Publicado: (2023)
RexUniNLU: Recursive Method with Explicit Schema Instructor for Universal NLU
por: Liu, Chengyuan, et al.
Publicado: (2024)
por: Liu, Chengyuan, et al.
Publicado: (2024)
RealTime QA: What's the Answer Right Now?
por: Kasai, Jungo, et al.
Publicado: (2022)
por: Kasai, Jungo, et al.
Publicado: (2022)
Representational Analysis of Binding in Language Models
por: Dai, Qin, et al.
Publicado: (2024)
por: Dai, Qin, et al.
Publicado: (2024)
Rectifying Belief Space via Unlearning to Harness LLMs' Reasoning
por: Niwa, Ayana, et al.
Publicado: (2025)
por: Niwa, Ayana, et al.
Publicado: (2025)
Cell-Based Representation of Relational Binding in Language Models
por: Dai, Qin, et al.
Publicado: (2026)
por: Dai, Qin, et al.
Publicado: (2026)
Unlocking Prompt Infilling Capability for Diffusion Language Models
por: Fujinuma, Yoshinari, et al.
Publicado: (2026)
por: Fujinuma, Yoshinari, et al.
Publicado: (2026)
Understanding and Controlling Repetition Neurons and Induction Heads in In-Context Learning
por: Doan, Nhi Hoai, et al.
Publicado: (2025)
por: Doan, Nhi Hoai, et al.
Publicado: (2025)
A Large Collection of Model-generated Contradictory Responses for Consistency-aware Dialogue Systems
por: Sato, Shiki, et al.
Publicado: (2024)
por: Sato, Shiki, et al.
Publicado: (2024)
Syntactic Learnability of Echo State Neural Language Models at Scale
por: Ueda, Ryo, et al.
Publicado: (2025)
por: Ueda, Ryo, et al.
Publicado: (2025)
TopK Language Models
por: Takahashi, Ryosuke, et al.
Publicado: (2025)
por: Takahashi, Ryosuke, et al.
Publicado: (2025)
Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge
por: Zhang, Ying, et al.
Publicado: (2025)
por: Zhang, Ying, et al.
Publicado: (2025)
Construction of Domain-specified Japanese Large Language Model for Finance through Continual Pre-training
por: Hirano, Masanori, et al.
Publicado: (2024)
por: Hirano, Masanori, et al.
Publicado: (2024)
CommonMorph: Participatory Morphological Documentation Platform
por: Mahmudi, Aso, et al.
Publicado: (2026)
por: Mahmudi, Aso, et al.
Publicado: (2026)
Flee the Flaw: Annotating the Underlying Logic of Fallacious Arguments Through Templates and Slot-filling
por: Robbani, Irfan, et al.
Publicado: (2024)
por: Robbani, Irfan, et al.
Publicado: (2024)
How Individual Traits and Language Styles Shape Preferences In Open-ended User-LLM Interaction: A Preliminary Study
por: Chevi, Rendi, et al.
Publicado: (2025)
por: Chevi, Rendi, et al.
Publicado: (2025)
MorphTok: Morphologically Grounded Tokenization for Indian Languages
por: Brahma, Maharaj, et al.
Publicado: (2025)
por: Brahma, Maharaj, et al.
Publicado: (2025)
Causal Representation Learning with Generative Artificial Intelligence: Application to Texts as Treatments
por: Imai, Kosuke, et al.
Publicado: (2024)
por: Imai, Kosuke, et al.
Publicado: (2024)
Non-commutative linear logic fragments with sub-context-free complexity
por: Nishimiya, Yusaku, et al.
Publicado: (2025)
por: Nishimiya, Yusaku, et al.
Publicado: (2025)
Reducing the Cost: Cross-Prompt Pre-Finetuning for Short Answer Scoring
por: Funayama, Hiroaki, et al.
Publicado: (2024)
por: Funayama, Hiroaki, et al.
Publicado: (2024)
MQM-Chat: Multidimensional Quality Metrics for Chat Translation
por: Li, Yunmeng, et al.
Publicado: (2024)
por: Li, Yunmeng, et al.
Publicado: (2024)
Emergence of Primacy and Recency Effect in Mamba: A Mechanistic Point of View
por: Airlangga, Muhammad Cendekia, et al.
Publicado: (2025)
por: Airlangga, Muhammad Cendekia, et al.
Publicado: (2025)
Linear Representations of Hierarchical Concepts in Language Models
por: Sakata, Masaki, et al.
Publicado: (2026)
por: Sakata, Masaki, et al.
Publicado: (2026)
Why Mean Pooling Works: Quantifying Second-Order Collapse in Text Embeddings
por: Hara, Tomomasa, et al.
Publicado: (2026)
por: Hara, Tomomasa, et al.
Publicado: (2026)
On Entity Identification in Language Models
por: Sakata, Masaki, et al.
Publicado: (2025)
por: Sakata, Masaki, et al.
Publicado: (2025)
Ejemplares similares
-
FinchGPT: a Transformer based language model for birdsong analysis
por: Kobayashi, Kosei, et al.
Publicado: (2025) -
Can Language Models Handle a Non-Gregorian Calendar? The Case of the Japanese wareki
por: Sasaki, Mutsumi, et al.
Publicado: (2025) -
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
por: Aoki, Yoichi, et al.
Publicado: (2024) -
LLMs Faithfully and Iteratively Compute Answers During CoT: A Systematic Analysis With Multi-step Arithmetics
por: Kudo, Keito, et al.
Publicado: (2024) -
The Curse of Popularity: Popular Entities have Catastrophic Side Effects when Deleting Knowledge from Language Models
por: Takahashi, Ryosuke, et al.
Publicado: (2024)