Do Neurons Dream of Primitive Operators? Wake-Sleep Compression Rediscovers Schank's Event Semantics
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Balogh, Peter |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Anxiety of Influence: Bloom Filters in Transformer Attention Heads
von: Balogh, Peter
Veröffentlicht: (2026)
von: Balogh, Peter
Veröffentlicht: (2026)
Do Androids Know They're Only Dreaming of Electric Sheep?
von: CH-Wang, Sky, et al.
Veröffentlicht: (2023)
von: CH-Wang, Sky, et al.
Veröffentlicht: (2023)
MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses
von: Yang, Zonglin, et al.
Veröffentlicht: (2024)
von: Yang, Zonglin, et al.
Veröffentlicht: (2024)
Language-specific Neurons Do Not Facilitate Cross-Lingual Transfer
von: Mondal, Soumen Kumar, et al.
Veröffentlicht: (2025)
von: Mondal, Soumen Kumar, et al.
Veröffentlicht: (2025)
Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution
von: Haider, Muhammad Umair, et al.
Veröffentlicht: (2025)
von: Haider, Muhammad Umair, et al.
Veröffentlicht: (2025)
Dreaming in Code for Curriculum Learning in Open-Ended Worlds
von: Mitsides, Konstantinos, et al.
Veröffentlicht: (2026)
von: Mitsides, Konstantinos, et al.
Veröffentlicht: (2026)
NeuronTune: Fine-Grained Neuron Modulation for Balanced Safety-Utility Alignment in LLMs
von: Pan, Birong, et al.
Veröffentlicht: (2025)
von: Pan, Birong, et al.
Veröffentlicht: (2025)
Confidence Regulation Neurons in Language Models
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
DreamNet: A Multimodal Framework for Semantic and Emotional Analysis of Sleep Narratives
von: Panchagnula, Tapasvi
Veröffentlicht: (2025)
von: Panchagnula, Tapasvi
Veröffentlicht: (2025)
CoSy: Evaluating Textual Explanations of Neurons
von: Kopf, Laura, et al.
Veröffentlicht: (2024)
von: Kopf, Laura, et al.
Veröffentlicht: (2024)
Decomposing Attention To Find Context-Sensitive Neurons
von: Gibson, Alex
Veröffentlicht: (2025)
von: Gibson, Alex
Veröffentlicht: (2025)
Universal Neurons in GPT2 Language Models
von: Gurnee, Wes, et al.
Veröffentlicht: (2024)
von: Gurnee, Wes, et al.
Veröffentlicht: (2024)
NEAT: Concept driven Neuron Attribution in LLMs
von: Kavuri, Vivek Hruday, et al.
Veröffentlicht: (2025)
von: Kavuri, Vivek Hruday, et al.
Veröffentlicht: (2025)
Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect
von: Presa, Joao Paulo Cavalcante, et al.
Veröffentlicht: (2026)
von: Presa, Joao Paulo Cavalcante, et al.
Veröffentlicht: (2026)
DreamPRM-Code: Function-as-Step Process Reward Model with Label Correction for LLM Coding
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2025)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2025)
SemanticZip: A Pilot Framework for Lossy Text Compression with LLMs as Semantic Decompressors
von: Trukhina, Natalia, et al.
Veröffentlicht: (2026)
von: Trukhina, Natalia, et al.
Veröffentlicht: (2026)
Finding Culture-Sensitive Neurons in Vision-Language Models
von: Zhao, Xiutian, et al.
Veröffentlicht: (2025)
von: Zhao, Xiutian, et al.
Veröffentlicht: (2025)
Projected Compression: Trainable Projection for Efficient Transformer Compression
von: Stefaniak, Maciej, et al.
Veröffentlicht: (2025)
von: Stefaniak, Maciej, et al.
Veröffentlicht: (2025)
Wake-Sleep Consolidated Learning
von: Sorrenti, Amelia, et al.
Veröffentlicht: (2023)
von: Sorrenti, Amelia, et al.
Veröffentlicht: (2023)
DreamCraft: Text-Guided Generation of Functional 3D Environments in Minecraft
von: Earle, Sam, et al.
Veröffentlicht: (2024)
von: Earle, Sam, et al.
Veröffentlicht: (2024)
A Single Neuron Is Sufficient to Bypass Safety Alignment in Large Language Models
von: Kazemi, Hamid, et al.
Veröffentlicht: (2026)
von: Kazemi, Hamid, et al.
Veröffentlicht: (2026)
NeuroAda: Activating Each Neuron's Potential for Parameter-Efficient Fine-Tuning
von: Zhang, Zhi, et al.
Veröffentlicht: (2025)
von: Zhang, Zhi, et al.
Veröffentlicht: (2025)
Zeroth-Order Adaptive Neuron Alignment Based Pruning without Re-Training
von: Cunegatti, Elia, et al.
Veröffentlicht: (2024)
von: Cunegatti, Elia, et al.
Veröffentlicht: (2024)
Mitigating Biases for Instruction-following Language Models via Bias Neurons Elimination
von: Yang, Nakyeong, et al.
Veröffentlicht: (2023)
von: Yang, Nakyeong, et al.
Veröffentlicht: (2023)
SPIN: Sparsifying and Integrating Internal Neurons in Large Language Models for Text Classification
von: Jiao, Difan, et al.
Veröffentlicht: (2023)
von: Jiao, Difan, et al.
Veröffentlicht: (2023)
Towards Understanding Safety Alignment: A Mechanistic Perspective from Safety Neurons
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
Event-Keyed Summarization
von: Gantt, William, et al.
Veröffentlicht: (2024)
von: Gantt, William, et al.
Veröffentlicht: (2024)
Polysemy of Synthetic Neurons Towards a New Type of Explanatory Categorical Vector Spaces
von: Pichat, Michael, et al.
Veröffentlicht: (2025)
von: Pichat, Michael, et al.
Veröffentlicht: (2025)
The Transfer Neurons Hypothesis: An Underlying Mechanism for Language Latent Space Transitions in Multilingual LLMs
von: Tezuka, Hinata, et al.
Veröffentlicht: (2025)
von: Tezuka, Hinata, et al.
Veröffentlicht: (2025)
KVSculpt: KV Cache Compression as Distillation
von: Jiang, Bo, et al.
Veröffentlicht: (2026)
von: Jiang, Bo, et al.
Veröffentlicht: (2026)
On the Compressibility of Quantized Large Language Models
von: Mao, Yu, et al.
Veröffentlicht: (2024)
von: Mao, Yu, et al.
Veröffentlicht: (2024)
Strategic Fusion Optimizes Transformer Compression
von: Rahman, Md Shoaibur
Veröffentlicht: (2025)
von: Rahman, Md Shoaibur
Veröffentlicht: (2025)
Activation Transport Operators
von: Szablewski, Andrzej, et al.
Veröffentlicht: (2025)
von: Szablewski, Andrzej, et al.
Veröffentlicht: (2025)
LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety
von: Yang, Junxiao, et al.
Veröffentlicht: (2026)
von: Yang, Junxiao, et al.
Veröffentlicht: (2026)
Fast Vocabulary Transfer for Language Model Compression
von: Gee, Leonidas, et al.
Veröffentlicht: (2024)
von: Gee, Leonidas, et al.
Veröffentlicht: (2024)
Communication Compression for Tensor Parallel LLM Inference
von: Hansen-Palmus, Jan, et al.
Veröffentlicht: (2024)
von: Hansen-Palmus, Jan, et al.
Veröffentlicht: (2024)
Learning to Compress Prompt in Natural Language Formats
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
Language Models Do Hard Arithmetic Tasks Easily and Hardly Do Easy Arithmetic Tasks
von: Gambardella, Andrew, et al.
Veröffentlicht: (2024)
von: Gambardella, Andrew, et al.
Veröffentlicht: (2024)
Semantic-Preserving Transformations as Mutation Operators: A Study on Their Effectiveness in Defect Detection
von: Hort, Max, et al.
Veröffentlicht: (2025)
von: Hort, Max, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Anxiety of Influence: Bloom Filters in Transformer Attention Heads
von: Balogh, Peter
Veröffentlicht: (2026) -
Do Androids Know They're Only Dreaming of Electric Sheep?
von: CH-Wang, Sky, et al.
Veröffentlicht: (2023) -
MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses
von: Yang, Zonglin, et al.
Veröffentlicht: (2024) -
Language-specific Neurons Do Not Facilitate Cross-Lingual Transfer
von: Mondal, Soumen Kumar, et al.
Veröffentlicht: (2025) -
Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination
von: Huang, Jerry, et al.
Veröffentlicht: (2024)