FinchGPT: a Transformer based language model for birdsong analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kobayashi, Kosei, Matsuzaki, Kosuke, Taniguchi, Masaya, Sakaguchi, Keisuke, Inui, Kentaro, Abe, Kentaro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
J-UniMorph: Japanese Morphological Annotation through the Universal Feature Schema
von: Matsuzaki, Kosuke, et al.
Veröffentlicht: (2024)
von: Matsuzaki, Kosuke, et al.
Veröffentlicht: (2024)
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
von: Aoki, Yoichi, et al.
Veröffentlicht: (2024)
von: Aoki, Yoichi, et al.
Veröffentlicht: (2024)
LLMs Faithfully and Iteratively Compute Answers During CoT: A Systematic Analysis With Multi-step Arithmetics
von: Kudo, Keito, et al.
Veröffentlicht: (2024)
von: Kudo, Keito, et al.
Veröffentlicht: (2024)
Can Language Models Handle a Non-Gregorian Calendar? The Case of the Japanese wareki
von: Sasaki, Mutsumi, et al.
Veröffentlicht: (2025)
von: Sasaki, Mutsumi, et al.
Veröffentlicht: (2025)
The Curse of Popularity: Popular Entities have Catastrophic Side Effects when Deleting Knowledge from Language Models
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2024)
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2024)
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference
von: Kamoda, Go, et al.
Veröffentlicht: (2025)
von: Kamoda, Go, et al.
Veröffentlicht: (2025)
Analyzing Feed-Forward Blocks in Transformers through the Lens of Attention Maps
von: Kobayashi, Goro, et al.
Veröffentlicht: (2023)
von: Kobayashi, Goro, et al.
Veröffentlicht: (2023)
Spelling-out is not Straightforward: LLMs' Capability of Tokenization from Token to Characters
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2025)
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2025)
Repetition Neurons: How Do Language Models Produce Repetitions?
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2024)
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2024)
Monotonic Representation of Numeric Properties in Language Models
von: Heinzerling, Benjamin, et al.
Veröffentlicht: (2024)
von: Heinzerling, Benjamin, et al.
Veröffentlicht: (2024)
Nodes Are Early, Edges Are Late: Probing Diagram Representations in Large Vision-Language Models
von: Yoshida, Haruto, et al.
Veröffentlicht: (2026)
von: Yoshida, Haruto, et al.
Veröffentlicht: (2026)
Annotating Errors in English Learners' Written Language Production: Advancing Automated Written Feedback Systems
von: Coyne, Steven, et al.
Veröffentlicht: (2025)
von: Coyne, Steven, et al.
Veröffentlicht: (2025)
MQM-Chat: Multidimensional Quality Metrics for Chat Translation
von: Li, Yunmeng, et al.
Veröffentlicht: (2024)
von: Li, Yunmeng, et al.
Veröffentlicht: (2024)
Rectifying Belief Space via Unlearning to Harness LLMs' Reasoning
von: Niwa, Ayana, et al.
Veröffentlicht: (2025)
von: Niwa, Ayana, et al.
Veröffentlicht: (2025)
Cell-Based Representation of Relational Binding in Language Models
von: Dai, Qin, et al.
Veröffentlicht: (2026)
von: Dai, Qin, et al.
Veröffentlicht: (2026)
Representational Analysis of Binding in Language Models
von: Dai, Qin, et al.
Veröffentlicht: (2024)
von: Dai, Qin, et al.
Veröffentlicht: (2024)
An Investigation of Warning Erroneous Chat Translations in Cross-lingual Communication
von: Li, Yunmeng, et al.
Veröffentlicht: (2024)
von: Li, Yunmeng, et al.
Veröffentlicht: (2024)
Understanding and Controlling Repetition Neurons and Induction Heads in In-Context Learning
von: Doan, Nhi Hoai, et al.
Veröffentlicht: (2025)
von: Doan, Nhi Hoai, et al.
Veröffentlicht: (2025)
Causal Representation Learning with Generative Artificial Intelligence: Application to Texts as Treatments
von: Imai, Kosuke, et al.
Veröffentlicht: (2024)
von: Imai, Kosuke, et al.
Veröffentlicht: (2024)
Tell Me Who Your Students Are: GPT Can Generate Valid Multiple-Choice Questions When Students' (Mis)Understanding Is Hinted
von: Shimmei, Machi, et al.
Veröffentlicht: (2025)
von: Shimmei, Machi, et al.
Veröffentlicht: (2025)
Syntactic Learnability of Echo State Neural Language Models at Scale
von: Ueda, Ryo, et al.
Veröffentlicht: (2025)
von: Ueda, Ryo, et al.
Veröffentlicht: (2025)
TopK Language Models
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2025)
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2025)
Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
A Large Collection of Model-generated Contradictory Responses for Consistency-aware Dialogue Systems
von: Sato, Shiki, et al.
Veröffentlicht: (2024)
von: Sato, Shiki, et al.
Veröffentlicht: (2024)
PheMT: A Phenomenon-wise Dataset for Machine Translation Robustness on User-Generated Contents
von: Fujii, Ryo, et al.
Veröffentlicht: (2020)
von: Fujii, Ryo, et al.
Veröffentlicht: (2020)
How Individual Traits and Language Styles Shape Preferences In Open-ended User-LLM Interaction: A Preliminary Study
von: Chevi, Rendi, et al.
Veröffentlicht: (2025)
von: Chevi, Rendi, et al.
Veröffentlicht: (2025)
RealTime QA: What's the Answer Right Now?
von: Kasai, Jungo, et al.
Veröffentlicht: (2022)
von: Kasai, Jungo, et al.
Veröffentlicht: (2022)
Expressive Power of One-Shot Control Operators and Coroutines
von: Kobayashi, Kentaro, et al.
Veröffentlicht: (2025)
von: Kobayashi, Kentaro, et al.
Veröffentlicht: (2025)
Emergence of Primacy and Recency Effect in Mamba: A Mechanistic Point of View
von: Airlangga, Muhammad Cendekia, et al.
Veröffentlicht: (2025)
von: Airlangga, Muhammad Cendekia, et al.
Veröffentlicht: (2025)
On Entity Identification in Language Models
von: Sakata, Masaki, et al.
Veröffentlicht: (2025)
von: Sakata, Masaki, et al.
Veröffentlicht: (2025)
Linear Representations of Hierarchical Concepts in Language Models
von: Sakata, Masaki, et al.
Veröffentlicht: (2026)
von: Sakata, Masaki, et al.
Veröffentlicht: (2026)
Why Mean Pooling Works: Quantifying Second-Order Collapse in Text Embeddings
von: Hara, Tomomasa, et al.
Veröffentlicht: (2026)
von: Hara, Tomomasa, et al.
Veröffentlicht: (2026)
Reducing the Cost: Cross-Prompt Pre-Finetuning for Short Answer Scoring
von: Funayama, Hiroaki, et al.
Veröffentlicht: (2024)
von: Funayama, Hiroaki, et al.
Veröffentlicht: (2024)
To Drop or Not to Drop? Predicting Argument Ellipsis Judgments: A Case Study in Japanese
von: Ishizuki, Yukiko, et al.
Veröffentlicht: (2024)
von: Ishizuki, Yukiko, et al.
Veröffentlicht: (2024)
Description-based Controllable Text-to-Speech with Cross-Lingual Voice Control
von: Yamamoto, Ryuichi, et al.
Veröffentlicht: (2024)
von: Yamamoto, Ryuichi, et al.
Veröffentlicht: (2024)
Large Language Models Are Human-Like Internally
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
Japanese-English Sentence Translation Exercises Dataset for Automatic Grading
von: Miura, Naoki, et al.
Veröffentlicht: (2024)
von: Miura, Naoki, et al.
Veröffentlicht: (2024)
LLMs Can Compensate for Deficiencies in Visual Representations
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
Measuring AI Reasoning: A Guide for Researchers
von: Nwadike, Munachiso Samuel, et al.
Veröffentlicht: (2026)
von: Nwadike, Munachiso Samuel, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
J-UniMorph: Japanese Morphological Annotation through the Universal Feature Schema
von: Matsuzaki, Kosuke, et al.
Veröffentlicht: (2024) -
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
von: Aoki, Yoichi, et al.
Veröffentlicht: (2024) -
LLMs Faithfully and Iteratively Compute Answers During CoT: A Systematic Analysis With Multi-step Arithmetics
von: Kudo, Keito, et al.
Veröffentlicht: (2024) -
Can Language Models Handle a Non-Gregorian Calendar? The Case of the Japanese wareki
von: Sasaki, Mutsumi, et al.
Veröffentlicht: (2025) -
The Curse of Popularity: Popular Entities have Catastrophic Side Effects when Deleting Knowledge from Language Models
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2024)