Exclusive Unlearning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sasaki, Mutsumi, Nakayama, Kouta, Miyao, Yusuke, Oseki, Yohei, Isonuma, Masaru |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How a Bilingual LM Becomes Bilingual: Tracing Internal Representations with Sparse Autoencoders
von: Inaba, Tatsuro, et al.
Veröffentlicht: (2025)
von: Inaba, Tatsuro, et al.
Veröffentlicht: (2025)
Unlearning Traces the Influential Training Data of Language Models
von: Isonuma, Masaru, et al.
Veröffentlicht: (2024)
von: Isonuma, Masaru, et al.
Veröffentlicht: (2024)
Do LLMs Need to Think in One Language? Correlation between Latent Language and Task Performance
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
Massive Supervised Fine-tuning Experiments Reveal How Data, Layer, and Training Factors Shape LLM Alignment Quality
von: Harada, Yuto, et al.
Veröffentlicht: (2025)
von: Harada, Yuto, et al.
Veröffentlicht: (2025)
Instability in Downstream Task Performance During LLM Pretraining
von: Nishida, Yuto, et al.
Veröffentlicht: (2025)
von: Nishida, Yuto, et al.
Veröffentlicht: (2025)
Towards Transfer Unlearning: Empirical Evidence of Cross-Domain Bias Mitigation
von: Lu, Huimin, et al.
Veröffentlicht: (2024)
von: Lu, Huimin, et al.
Veröffentlicht: (2024)
Is Structure Dependence Shaped for Efficient Communication?: A Case Study on Coordination
von: Kajikawa, Kohei, et al.
Veröffentlicht: (2024)
von: Kajikawa, Kohei, et al.
Veröffentlicht: (2024)
llm-jp-modernbert: A ModernBERT Model Trained on a Large-Scale Japanese Corpus with Long Context Length
von: Sugiura, Issa, et al.
Veröffentlicht: (2025)
von: Sugiura, Issa, et al.
Veröffentlicht: (2025)
What's New in My Data? Novelty Exploration via Contrastive Generation
von: Isonuma, Masaru, et al.
Veröffentlicht: (2024)
von: Isonuma, Masaru, et al.
Veröffentlicht: (2024)
Composition, Attention, or Both?
von: Yoshida, Ryo, et al.
Veröffentlicht: (2022)
von: Yoshida, Ryo, et al.
Veröffentlicht: (2022)
Investigating Training and Generalization in Faithful Self-Explanations of Large Language Models
von: Doi, Tomoki, et al.
Veröffentlicht: (2025)
von: Doi, Tomoki, et al.
Veröffentlicht: (2025)
Comprehensive Evaluation of Large Language Models for Topic Modeling
von: Doi, Tomoki, et al.
Veröffentlicht: (2024)
von: Doi, Tomoki, et al.
Veröffentlicht: (2024)
Developmentally-plausible Working Memory Shapes a Critical Period for Language Acquisition
von: Mita, Masato, et al.
Veröffentlicht: (2025)
von: Mita, Masato, et al.
Veröffentlicht: (2025)
Modeling Human Sentence Processing with Left-Corner Recurrent Neural Network Grammars
von: Yoshida, Ryo, et al.
Veröffentlicht: (2021)
von: Yoshida, Ryo, et al.
Veröffentlicht: (2021)
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision
von: Yoshida, Ryo, et al.
Veröffentlicht: (2024)
von: Yoshida, Ryo, et al.
Veröffentlicht: (2024)
The Imperfective Paradox in Large Language Models
von: Ma, Bolei, et al.
Veröffentlicht: (2026)
von: Ma, Bolei, et al.
Veröffentlicht: (2026)
Analyzing Correlations Between Intrinsic and Extrinsic Bias Metrics of Static Word Embeddings With Their Measuring Biases Aligned
von: Katô, Taisei, et al.
Veröffentlicht: (2024)
von: Katô, Taisei, et al.
Veröffentlicht: (2024)
UniDetox: Universal Detoxification of Large Language Models via Dataset Distillation
von: Lu, Huimin, et al.
Veröffentlicht: (2025)
von: Lu, Huimin, et al.
Veröffentlicht: (2025)
Psychometric Predictive Power of Large Language Models
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2023)
Language Acquisition Device in Large Language Models
von: Mita, Masato, et al.
Veröffentlicht: (2026)
von: Mita, Masato, et al.
Veröffentlicht: (2026)
Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
von: Someya, Taiga, et al.
Veröffentlicht: (2025)
von: Someya, Taiga, et al.
Veröffentlicht: (2025)
Dual Alignment Between Language Model Layers and Human Sentence Processing
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2026)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2026)
Why Are Parsing Actions for Understanding Message Hierarchies Not Random?
von: Kato, Daichi, et al.
Veröffentlicht: (2025)
von: Kato, Daichi, et al.
Veröffentlicht: (2025)
Do Self-Supervised Speech Models Exhibit the Critical Period Effects in Language Acquisition?
von: Koga, Yurie, et al.
Veröffentlicht: (2025)
von: Koga, Yurie, et al.
Veröffentlicht: (2025)
Are Checklists Really Useful for Automatic Evaluation of Generative Tasks?
von: Furuhashi, Momoka, et al.
Veröffentlicht: (2025)
von: Furuhashi, Momoka, et al.
Veröffentlicht: (2025)
An Existence Proof for Neural Language Models That Can Explain Garden-Path Effects via Surprisal
von: Yoshida, Ryo, et al.
Veröffentlicht: (2026)
von: Yoshida, Ryo, et al.
Veröffentlicht: (2026)
Rethinking the Relationship between the Power Law and Hierarchical Structures
von: Nakaishi, Kai, et al.
Veröffentlicht: (2025)
von: Nakaishi, Kai, et al.
Veröffentlicht: (2025)
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency
von: Haga, Akari, et al.
Veröffentlicht: (2024)
von: Haga, Akari, et al.
Veröffentlicht: (2024)
A Comparative Analysis of LLM Memorization at Statistical and Internal Levels: Cross-Model Commonalities and Model-Specific Signatures
von: Chen, Bowen, et al.
Veröffentlicht: (2026)
von: Chen, Bowen, et al.
Veröffentlicht: (2026)
A Multi-Perspective Analysis of Memorization in Large Language Models
von: Chen, Bowen, et al.
Veröffentlicht: (2024)
von: Chen, Bowen, et al.
Veröffentlicht: (2024)
A Statistical and Multi-Perspective Revisiting of the Membership Inference Attack in Large Language Models
von: Chen, Bowen, et al.
Veröffentlicht: (2024)
von: Chen, Bowen, et al.
Veröffentlicht: (2024)
Large Language Models Are Human-Like Internally
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2025)
Can Language Models Learn Typologically Implausible Languages?
von: Xu, Tianyang, et al.
Veröffentlicht: (2025)
von: Xu, Tianyang, et al.
Veröffentlicht: (2025)
Human-Grounded Multimodal Benchmark with 900K-Scale Aggregated Student Response Distributions from Japan's National Assessment of Academic Ability
von: Takami, Kyosuke, et al.
Veröffentlicht: (2026)
von: Takami, Kyosuke, et al.
Veröffentlicht: (2026)
Improving Unsupervised Constituency Parsing via Maximizing Semantic Information
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
Unsupervised Parsing by Searching for Frequent Word Sequences among Sentences with Equivalent Predicate-Argument Structures
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
Textless Dependency Parsing by Labeled Sequence Prediction
von: Kando, Shunsuke, et al.
Veröffentlicht: (2024)
von: Kando, Shunsuke, et al.
Veröffentlicht: (2024)
Exploring the Effect of Segmentation and Vocabulary Size on Speech Tokenization for Speech Language Models
von: Kando, Shunsuke, et al.
Veröffentlicht: (2025)
von: Kando, Shunsuke, et al.
Veröffentlicht: (2025)
Emergent Word Order Universals from Cognitively-Motivated Language Models
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2024)
von: Kuribayashi, Tatsuki, et al.
Veröffentlicht: (2024)
If Attention Serves as a Cognitive Model of Human Memory Retrieval, What is the Plausible Memory Representation?
von: Yoshida, Ryo, et al.
Veröffentlicht: (2025)
von: Yoshida, Ryo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
How a Bilingual LM Becomes Bilingual: Tracing Internal Representations with Sparse Autoencoders
von: Inaba, Tatsuro, et al.
Veröffentlicht: (2025) -
Unlearning Traces the Influential Training Data of Language Models
von: Isonuma, Masaru, et al.
Veröffentlicht: (2024) -
Do LLMs Need to Think in One Language? Correlation between Latent Language and Task Performance
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025) -
Massive Supervised Fine-tuning Experiments Reveal How Data, Layer, and Training Factors Shape LLM Alignment Quality
von: Harada, Yuto, et al.
Veröffentlicht: (2025) -
Instability in Downstream Task Performance During LLM Pretraining
von: Nishida, Yuto, et al.
Veröffentlicht: (2025)