Rethinking Circuit Completeness in Language Models: AND, OR, and ADDER Gates
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Hang, Zhu, Jiaying, Yang, Xinyu, Wang, Wenya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CLUE: Conflict-guided Localization for LLM Unlearning Framework
von: Chen, Hang, et al.
Veröffentlicht: (2025)
von: Chen, Hang, et al.
Veröffentlicht: (2025)
Skill Path: Unveiling Language Skills from Circuit Graphs
von: Chen, Hang, et al.
Veröffentlicht: (2024)
von: Chen, Hang, et al.
Veröffentlicht: (2024)
SSL Framework for Causal Inconsistency between Structures and Representations
von: Chen, Hang, et al.
Veröffentlicht: (2023)
von: Chen, Hang, et al.
Veröffentlicht: (2023)
Quantifying Semantic Emergence in Language Models
von: Chen, Hang, et al.
Veröffentlicht: (2024)
von: Chen, Hang, et al.
Veröffentlicht: (2024)
Towards Causal Relationship in Indefinite Data: Baseline Model and New Datasets
von: Chen, Hang, et al.
Veröffentlicht: (2024)
von: Chen, Hang, et al.
Veröffentlicht: (2024)
Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete
von: Li, Qian, et al.
Veröffentlicht: (2026)
von: Li, Qian, et al.
Veröffentlicht: (2026)
Rethinking Probabilistic Circuit Parameter Learning
von: Liu, Anji, et al.
Veröffentlicht: (2025)
von: Liu, Anji, et al.
Veröffentlicht: (2025)
Circuit Representation Learning with Masked Gate Modeling and Verilog-AIG Alignment
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
Rethinking Language Model Scaling under Transferable Hypersphere Optimization
von: Ren, Liliang, et al.
Veröffentlicht: (2026)
von: Ren, Liliang, et al.
Veröffentlicht: (2026)
Rethinking Machine Unlearning for Large Language Models
von: Liu, Sijia, et al.
Veröffentlicht: (2024)
von: Liu, Sijia, et al.
Veröffentlicht: (2024)
Rethinking Key-Value Cache Compression Techniques for Large Language Model Serving
von: Gao, Wei, et al.
Veröffentlicht: (2025)
von: Gao, Wei, et al.
Veröffentlicht: (2025)
Rethinking Multinomial Logistic Mixture of Experts with Sigmoid Gating Function
von: Pham, Tuan Minh, et al.
Veröffentlicht: (2026)
von: Pham, Tuan Minh, et al.
Veröffentlicht: (2026)
Rethinking Gating Mechanism in Sparse MoE: Handling Arbitrary Modality Inputs with Confidence-Guided Gate
von: Zheng, Liangwei Nathan, et al.
Veröffentlicht: (2025)
von: Zheng, Liangwei Nathan, et al.
Veröffentlicht: (2025)
Towards the Worst-case Robustness of Large Language Models
von: Chen, Huanran, et al.
Veröffentlicht: (2025)
von: Chen, Huanran, et al.
Veröffentlicht: (2025)
Quantifying Gate Contribution in Quantum Feature Maps for Scalable Circuit Optimization
von: Rodríguez-Díaz, F., et al.
Veröffentlicht: (2026)
von: Rodríguez-Díaz, F., et al.
Veröffentlicht: (2026)
Rethinking Regularization Methods for Knowledge Graph Completion
von: Li, Linyu, et al.
Veröffentlicht: (2025)
von: Li, Linyu, et al.
Veröffentlicht: (2025)
SPAC-Net: Rethinking Point Cloud Completion with Structural Prior
von: Wu, Zizhao, et al.
Veröffentlicht: (2024)
von: Wu, Zizhao, et al.
Veröffentlicht: (2024)
Multiverse: Your Language Models Secretly Decide How to Parallelize and Merge Generation
von: Yang, Xinyu, et al.
Veröffentlicht: (2025)
von: Yang, Xinyu, et al.
Veröffentlicht: (2025)
Enhancing Circuit Trainability with Selective Gate Activation Strategy
von: Cho, Jeihee, et al.
Veröffentlicht: (2025)
von: Cho, Jeihee, et al.
Veröffentlicht: (2025)
Rethinking the Role of Temperature in Large Language Model Distillation
von: Luong, Hoang-Chau, et al.
Veröffentlicht: (2026)
von: Luong, Hoang-Chau, et al.
Veröffentlicht: (2026)
DeepGate3: Towards Scalable Circuit Representation Learning
von: Shi, Zhengyuan, et al.
Veröffentlicht: (2024)
von: Shi, Zhengyuan, et al.
Veröffentlicht: (2024)
Efficient Learning for Linear Properties of Bounded-Gate Quantum Circuits
von: Du, Yuxuan, et al.
Veröffentlicht: (2024)
von: Du, Yuxuan, et al.
Veröffentlicht: (2024)
Rethinking Federated Graph Foundation Models: A Graph-Language Alignment-based Approach
von: Zhu, Yinlin, et al.
Veröffentlicht: (2026)
von: Zhu, Yinlin, et al.
Veröffentlicht: (2026)
LLM-MINE: Large Language Model based Alzheimer's Disease and Related Dementias Phenotypes Mining from Clinical Notes
von: Shao, Mingchen, et al.
Veröffentlicht: (2026)
von: Shao, Mingchen, et al.
Veröffentlicht: (2026)
Navigating by Old Maps: The Pitfalls of Static Mechanistic Localization in LLM Post-Training
von: Chen, Hang, et al.
Veröffentlicht: (2026)
von: Chen, Hang, et al.
Veröffentlicht: (2026)
Divide, Harmonize, Then Conquer It: Shooting Multi-Commodity Flow Problems with Multimodal Language Models
von: Yuan, Xinyu, et al.
Veröffentlicht: (2026)
von: Yuan, Xinyu, et al.
Veröffentlicht: (2026)
Rethinking Memorization Measures and their Implications in Large Language Models
von: Ghosh, Bishwamittra, et al.
Veröffentlicht: (2025)
von: Ghosh, Bishwamittra, et al.
Veröffentlicht: (2025)
Rethinking Token Prediction: Tree-Structured Diffusion Language Model
von: Wu, Zihao, et al.
Veröffentlicht: (2026)
von: Wu, Zihao, et al.
Veröffentlicht: (2026)
Rethinking Client-oriented Federated Graph Learning
von: Chen, Zekai, et al.
Veröffentlicht: (2025)
von: Chen, Zekai, et al.
Veröffentlicht: (2025)
Learning-driven Physically-aware Large-scale Circuit Gate Sizing
von: Ye, Yuyang, et al.
Veröffentlicht: (2024)
von: Ye, Yuyang, et al.
Veröffentlicht: (2024)
Gradient-Gated DPO: Stabilizing Preference Optimization in Language Models
von: Mouiche, Inoussa
Veröffentlicht: (2026)
von: Mouiche, Inoussa
Veröffentlicht: (2026)
Probabilistic Design of Parametrized Quantum Circuits through Local Gate Modifications
von: Jones, Grier M., et al.
Veröffentlicht: (2026)
von: Jones, Grier M., et al.
Veröffentlicht: (2026)
VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
When Continual Learning Moves to Memory: A Study of Experience Reuse in LLM Agents
von: Hu, Qisheng, et al.
Veröffentlicht: (2026)
von: Hu, Qisheng, et al.
Veröffentlicht: (2026)
Beyond Efficiency: A Systematic Survey of Resource-Efficient Large Language Models
von: Bai, Guangji, et al.
Veröffentlicht: (2024)
von: Bai, Guangji, et al.
Veröffentlicht: (2024)
Rethinking Data Mixing from the Perspective of Large Language Models
von: Xu, Yuanjian, et al.
Veröffentlicht: (2026)
von: Xu, Yuanjian, et al.
Veröffentlicht: (2026)
STGCN-LSTM for Olympic Medal Prediction: Dynamic Power Modeling and Causal Policy Optimization
von: Wang, Yiquan, et al.
Veröffentlicht: (2025)
von: Wang, Yiquan, et al.
Veröffentlicht: (2025)
FFN Fusion: Rethinking Sequential Computation in Large Language Models
von: Bercovich, Akhiad, et al.
Veröffentlicht: (2025)
von: Bercovich, Akhiad, et al.
Veröffentlicht: (2025)
NovoMolGen: Rethinking Molecular Language Model Pretraining
von: Chitsaz, Kamran, et al.
Veröffentlicht: (2025)
von: Chitsaz, Kamran, et al.
Veröffentlicht: (2025)
Unveiling the Basin-Like Loss Landscape in Large Language Models
von: Chen, Huanran, et al.
Veröffentlicht: (2025)
von: Chen, Huanran, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CLUE: Conflict-guided Localization for LLM Unlearning Framework
von: Chen, Hang, et al.
Veröffentlicht: (2025) -
Skill Path: Unveiling Language Skills from Circuit Graphs
von: Chen, Hang, et al.
Veröffentlicht: (2024) -
SSL Framework for Causal Inconsistency between Structures and Representations
von: Chen, Hang, et al.
Veröffentlicht: (2023) -
Quantifying Semantic Emergence in Language Models
von: Chen, Hang, et al.
Veröffentlicht: (2024) -
Towards Causal Relationship in Indefinite Data: Baseline Model and New Datasets
von: Chen, Hang, et al.
Veröffentlicht: (2024)