Gespeichert in:
| Hauptverfasser: | Kudo, Keito, Aoki, Yoichi, Kuribayashi, Tatsuki, Sone, Shusaku, Taniguchi, Masaya, Brassard, Ana, Sakaguchi, Keisuke, Inui, Kentaro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2412.01113 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
von: Aoki, Yoichi, et al.
Veröffentlicht: (2024)
von: Aoki, Yoichi, et al.
Veröffentlicht: (2024)
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
von: Brassard, Ana, et al.
Veröffentlicht: (2024)
Nodes Are Early, Edges Are Late: Probing Diagram Representations in Large Vision-Language Models
von: Yoshida, Haruto, et al.
Veröffentlicht: (2026)
von: Yoshida, Haruto, et al.
Veröffentlicht: (2026)
J-UniMorph: Japanese Morphological Annotation through the Universal Feature Schema
von: Matsuzaki, Kosuke, et al.
Veröffentlicht: (2024)
von: Matsuzaki, Kosuke, et al.
Veröffentlicht: (2024)
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference
von: Kamoda, Go, et al.
Veröffentlicht: (2025)
von: Kamoda, Go, et al.
Veröffentlicht: (2025)
FinchGPT: a Transformer based language model for birdsong analysis
von: Kobayashi, Kosei, et al.
Veröffentlicht: (2025)
von: Kobayashi, Kosei, et al.
Veröffentlicht: (2025)
Syntactic Learnability of Echo State Neural Language Models at Scale
von: Ueda, Ryo, et al.
Veröffentlicht: (2025)
von: Ueda, Ryo, et al.
Veröffentlicht: (2025)
Analyzing Feed-Forward Blocks in Transformers through the Lens of Attention Maps
von: Kobayashi, Goro, et al.
Veröffentlicht: (2023)
von: Kobayashi, Goro, et al.
Veröffentlicht: (2023)
To Drop or Not to Drop? Predicting Argument Ellipsis Judgments: A Case Study in Japanese
von: Ishizuki, Yukiko, et al.
Veröffentlicht: (2024)
von: Ishizuki, Yukiko, et al.
Veröffentlicht: (2024)
Large Language Models Are Human-Like Internally
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
RealTime QA: What's the Answer Right Now?
von: Kasai, Jungo, et al.
Veröffentlicht: (2022)
von: Kasai, Jungo, et al.
Veröffentlicht: (2022)
On Representational Dissociation of Language and Arithmetic in Large Language Models
von: Kisako, Riku, et al.
Veröffentlicht: (2025)
von: Kisako, Riku, et al.
Veröffentlicht: (2025)
The Curse of Popularity: Popular Entities have Catastrophic Side Effects when Deleting Knowledge from Language Models
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2024)
von: Takahashi, Ryosuke, et al.
Veröffentlicht: (2024)
Does Vision Accelerate Hierarchical Generalization in Neural Language Learners?
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
Repetitive Infection Spreading and Directed Evolution in the Susceptible-Infected-Recovered-Susceptible Model
von: Sakaguchi, Hidetsugu, et al.
Veröffentlicht: (2024)
von: Sakaguchi, Hidetsugu, et al.
Veröffentlicht: (2024)
Can Language Models Handle a Non-Gregorian Calendar? The Case of the Japanese wareki
von: Sasaki, Mutsumi, et al.
Veröffentlicht: (2025)
von: Sasaki, Mutsumi, et al.
Veröffentlicht: (2025)
Spelling-out is not Straightforward: LLMs' Capability of Tokenization from Token to Characters
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2025)
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2025)
A Multi-Agent Probabilistic Inference Framework Inspired by Kairanban-Style CoT System with IdoBata Conversation for Debiasing
von: Ueno, Takato, et al.
Veröffentlicht: (2025)
von: Ueno, Takato, et al.
Veröffentlicht: (2025)
Annotating Errors in English Learners' Written Language Production: Advancing Automated Written Feedback Systems
von: Coyne, Steven, et al.
Veröffentlicht: (2025)
von: Coyne, Steven, et al.
Veröffentlicht: (2025)
Psychometric Predictive Power of Large Language Models
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
Reducing the Cost: Cross-Prompt Pre-Finetuning for Short Answer Scoring
von: Funayama, Hiroaki, et al.
Veröffentlicht: (2024)
von: Funayama, Hiroaki, et al.
Veröffentlicht: (2024)
Rectifying Belief Space via Unlearning to Harness LLMs' Reasoning
von: Niwa, Ayana, et al.
Veröffentlicht: (2025)
von: Niwa, Ayana, et al.
Veröffentlicht: (2025)
Which Word Orders Facilitate Length Generalization in LMs? An Investigation with GCG-Based Artificial Languages
von: El-Naggar, Nadine, et al.
Veröffentlicht: (2025)
von: El-Naggar, Nadine, et al.
Veröffentlicht: (2025)
What Kind of Language is Easy to Language-Model Under Curriculum Learning?
von: El-Naggar, Nadine, et al.
Veröffentlicht: (2026)
von: El-Naggar, Nadine, et al.
Veröffentlicht: (2026)
From Geometry to Culture: An Iterative VLM Layout Framework for Placing Objects in Complex 3D Scene Contexts
von: Asano, Yuto, et al.
Veröffentlicht: (2025)
von: Asano, Yuto, et al.
Veröffentlicht: (2025)
Automatic Feedback Generation for Short Answer Questions using Answer Diagnostic Graphs
von: Furuhashi, Momoka, et al.
Veröffentlicht: (2025)
von: Furuhashi, Momoka, et al.
Veröffentlicht: (2025)
Repetition Neurons: How Do Language Models Produce Repetitions?
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2024)
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2024)
Monotonic Representation of Numeric Properties in Language Models
von: Heinzerling, Benjamin, et al.
Veröffentlicht: (2024)
von: Heinzerling, Benjamin, et al.
Veröffentlicht: (2024)
TNF: Tri-branch Neural Fusion for Multimodal Medical Data Classification
von: Zheng, Tong, et al.
Veröffentlicht: (2024)
von: Zheng, Tong, et al.
Veröffentlicht: (2024)
CoT-based Synthesizer: Enhancing LLM Performance through Answer Synthesis
von: Zhang, Bohan, et al.
Veröffentlicht: (2025)
von: Zhang, Bohan, et al.
Veröffentlicht: (2025)
Diagnosing Harmful Continuation in Answer-Correct Long-CoT Training Traces
von: He, Chen, et al.
Veröffentlicht: (2026)
von: He, Chen, et al.
Veröffentlicht: (2026)
Transformer Key-Value Memories Are Nearly as Interpretable as Sparse Autoencoders
von: Ye, Mengyu, et al.
Veröffentlicht: (2025)
von: Ye, Mengyu, et al.
Veröffentlicht: (2025)
Can Input Attributions Explain Inductive Reasoning in In-Context Learning?
von: Ye, Mengyu, et al.
Veröffentlicht: (2024)
von: Ye, Mengyu, et al.
Veröffentlicht: (2024)
Self-Training Meets Consistency: Improving LLMs' Reasoning with Consistency-Driven Rationale Evaluation
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2024)
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2024)
Dual Alignment Between Language Model Layers and Human Sentence Processing
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2026)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2026)
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
Decentralized Collective World Model for Emergent Communication and Coordination
von: Nomura, Kentaro, et al.
Veröffentlicht: (2025)
von: Nomura, Kentaro, et al.
Veröffentlicht: (2025)
Prior-Free Sample Size Design for Test-and-Roll Experiments
von: Kawato, Kentaro, et al.
Veröffentlicht: (2026)
von: Kawato, Kentaro, et al.
Veröffentlicht: (2026)
LLMs Can Compensate for Deficiencies in Visual Representations
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
Reconsidering Positional Supervision in Masked Diffusion Language Model Training
von: Ye, Mengyu, et al.
Veröffentlicht: (2026)
von: Ye, Mengyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
First Heuristic Then Rational: Dynamic Use of Heuristics in Language Model Reasoning
von: Aoki, Yoichi, et al.
Veröffentlicht: (2024) -
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation
von: Brassard, Ana, et al.
Veröffentlicht: (2024) -
Nodes Are Early, Edges Are Late: Probing Diagram Representations in Large Vision-Language Models
von: Yoshida, Haruto, et al.
Veröffentlicht: (2026) -
J-UniMorph: Japanese Morphological Annotation through the Universal Feature Schema
von: Matsuzaki, Kosuke, et al.
Veröffentlicht: (2024) -
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference
von: Kamoda, Go, et al.
Veröffentlicht: (2025)