Dead Weights, Live Signals: Feedforward Graphs of Frozen Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Armstrong, Marcus, Ayoobi, Navid, Mukherjee, Arjun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Thinking in Different Spaces: Domain-Specific Latent Geometry Survives Cross-Architecture Translation
di: Armstrong, Marcus, et al.
Pubblicazione: (2026)
di: Armstrong, Marcus, et al.
Pubblicazione: (2026)
The Erosion of LLM Signatures: Can We Still Distinguish Human and LLM-Generated Scientific Ideas After Iterative Paraphrasing?
di: Shahriar, Sadat, et al.
Pubblicazione: (2025)
di: Shahriar, Sadat, et al.
Pubblicazione: (2025)
Say Anything but This: When Tokenizer Betrays Reasoning in LLMs
di: Ayoobi, Navid, et al.
Pubblicazione: (2026)
di: Ayoobi, Navid, et al.
Pubblicazione: (2026)
Seeing Through AI's Lens: Enhancing Human Skepticism Towards LLM-Generated Fake News
di: Ayoobi, Navid, et al.
Pubblicazione: (2024)
di: Ayoobi, Navid, et al.
Pubblicazione: (2024)
Argumentative Debates for Transparent Bias Detection [Technical Report]
di: Ayoobi, Hamed, et al.
Pubblicazione: (2025)
di: Ayoobi, Hamed, et al.
Pubblicazione: (2025)
ChatGPT or A Silent Everywhere Helper: A Survey of Large Language Models
di: Akhtarshenas, Azim, et al.
Pubblicazione: (2025)
di: Akhtarshenas, Azim, et al.
Pubblicazione: (2025)
Compression Repair for Feedforward Neural Networks Based on Model Equivalence Evaluation
di: Mo, Zihao, et al.
Pubblicazione: (2024)
di: Mo, Zihao, et al.
Pubblicazione: (2024)
Exposing Pink Slime Journalism: Linguistic Signatures and Robust Detection Against LLM-Generated Threats
di: Shahriar, Sadat, et al.
Pubblicazione: (2025)
di: Shahriar, Sadat, et al.
Pubblicazione: (2025)
Deriving Equivalent Symbol-Based Decision Models from Feedforward Neural Networks
di: Seidel, Sebastian, et al.
Pubblicazione: (2025)
di: Seidel, Sebastian, et al.
Pubblicazione: (2025)
SplitFrozen: Split Learning with Device-side Model Frozen for Fine-Tuning LLM on Heterogeneous Resource-Constrained Devices
di: Ma, Jian, et al.
Pubblicazione: (2025)
di: Ma, Jian, et al.
Pubblicazione: (2025)
LORA-CRAFT: Cross-layer Rank Adaptation via Frozen Tucker Decomposition of Pre-trained Attention Weights
di: Dewage, Kasun, et al.
Pubblicazione: (2026)
di: Dewage, Kasun, et al.
Pubblicazione: (2026)
OmniCellTOSG: The First Cell Text-Omic Signaling Graphs Dataset for Graph Language Foundation Modeling
di: Zhang, Heming, et al.
Pubblicazione: (2025)
di: Zhang, Heming, et al.
Pubblicazione: (2025)
Weight-Entanglement Meets Gradient-Based Neural Architecture Search
di: Sukthanker, Rhea Sanjay, et al.
Pubblicazione: (2023)
di: Sukthanker, Rhea Sanjay, et al.
Pubblicazione: (2023)
Sysformer: Safeguarding Frozen Large Language Models with Adaptive System Prompts
di: Sharma, Kartik, et al.
Pubblicazione: (2025)
di: Sharma, Kartik, et al.
Pubblicazione: (2025)
Evaluating Computational Accuracy of Large Language Models in Numerical Reasoning Tasks for Healthcare Applications
di: Malghan, Arjun R.
Pubblicazione: (2025)
di: Malghan, Arjun R.
Pubblicazione: (2025)
More Expressive Feedforward Layers: Part I. Token-Adaptive Mixing of Activations
di: Wang, Mingze, et al.
Pubblicazione: (2026)
di: Wang, Mingze, et al.
Pubblicazione: (2026)
Weight Ensembling Improves Reasoning in Language Models
di: Dang, Xingyu, et al.
Pubblicazione: (2025)
di: Dang, Xingyu, et al.
Pubblicazione: (2025)
IAA: Inner-Adaptor Architecture Empowers Frozen Large Language Model with Multimodal Capabilities
di: Wang, Bin, et al.
Pubblicazione: (2024)
di: Wang, Bin, et al.
Pubblicazione: (2024)
Bridging Interpretability and Robustness Using LIME-Guided Model Refinement
di: Nayyem, Navid, et al.
Pubblicazione: (2024)
di: Nayyem, Navid, et al.
Pubblicazione: (2024)
Llamba: Scaling Distilled Recurrent Models for Efficient Language Processing
di: Bick, Aviv, et al.
Pubblicazione: (2025)
di: Bick, Aviv, et al.
Pubblicazione: (2025)
QF: Quick Feedforward AI Model Training without Gradient Back Propagation
di: Qi, Feng
Pubblicazione: (2025)
di: Qi, Feng
Pubblicazione: (2025)
Form Follows Function: Recursive Stem Model
di: Hakimi, Navid
Pubblicazione: (2026)
di: Hakimi, Navid
Pubblicazione: (2026)
Trained Persistent Memory for Frozen Decoder-Only LLMs
di: Jeong, Hong
Pubblicazione: (2026)
di: Jeong, Hong
Pubblicazione: (2026)
BOHM: Zero-Cost Hierarchical Attribution for Compound AI Systems
di: Armstrong, Joss
Pubblicazione: (2026)
di: Armstrong, Joss
Pubblicazione: (2026)
Partition Tree Weighting for Non-Stationary Stochastic Bandits
di: Veness, Joel, et al.
Pubblicazione: (2025)
di: Veness, Joel, et al.
Pubblicazione: (2025)
HardNet: Hard-Constrained Neural Networks with Universal Approximation Guarantees
di: Min, Youngjae, et al.
Pubblicazione: (2024)
di: Min, Youngjae, et al.
Pubblicazione: (2024)
Graph-Aware Diffusion for Signal Generation
di: Rozada, Sergio, et al.
Pubblicazione: (2025)
di: Rozada, Sergio, et al.
Pubblicazione: (2025)
Whispering to a Blackbox: Bootstrapping Frozen OCR with Visual Prompts
di: Samandarov, Samandar, et al.
Pubblicazione: (2026)
di: Samandarov, Samandar, et al.
Pubblicazione: (2026)
Frozen Layers: Memory-efficient Many-fidelity Hyperparameter Optimization
di: Carstensen, Timur, et al.
Pubblicazione: (2025)
di: Carstensen, Timur, et al.
Pubblicazione: (2025)
Federated Learning: A Cutting-Edge Survey of the Latest Advancements and Applications
di: Akhtarshenas, Azim, et al.
Pubblicazione: (2023)
di: Akhtarshenas, Azim, et al.
Pubblicazione: (2023)
Beyond Gaussian Initializations: Signal Preserving Weight Initialization for Odd-Sigmoid Activations
di: Lee, Hyunwoo, et al.
Pubblicazione: (2025)
di: Lee, Hyunwoo, et al.
Pubblicazione: (2025)
Graph Foundation Models: Bridging Language Model Paradigms and Graph Optimization
di: Liang, Yunhao, et al.
Pubblicazione: (2025)
di: Liang, Yunhao, et al.
Pubblicazione: (2025)
SwiftPrune: Hessian-Free Weight Pruning for Large Language Models
di: Kang, Yuhan, et al.
Pubblicazione: (2025)
di: Kang, Yuhan, et al.
Pubblicazione: (2025)
wd1: Weighted Policy Optimization for Reasoning in Diffusion Language Models
di: Tang, Xiaohang, et al.
Pubblicazione: (2025)
di: Tang, Xiaohang, et al.
Pubblicazione: (2025)
Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies
di: Zhao, Guangyu, et al.
Pubblicazione: (2024)
di: Zhao, Guangyu, et al.
Pubblicazione: (2024)
Partially Frozen Random Networks Contain Compact Strong Lottery Tickets
di: Otsuka, Hikari, et al.
Pubblicazione: (2024)
di: Otsuka, Hikari, et al.
Pubblicazione: (2024)
Deep Learning without Weight Symmetry
di: Ji-An, Li, et al.
Pubblicazione: (2024)
di: Ji-An, Li, et al.
Pubblicazione: (2024)
GraphEdit: Large Language Models for Graph Structure Learning
di: Guo, Zirui, et al.
Pubblicazione: (2024)
di: Guo, Zirui, et al.
Pubblicazione: (2024)
Vertex-Softmax: Tight Transformer Verification via Exact Softmax Optimization
di: Rezazadeh, Navid, et al.
Pubblicazione: (2026)
di: Rezazadeh, Navid, et al.
Pubblicazione: (2026)
Quantifying Representation Reliability in Self-Supervised Learning Models
di: Park, Young-Jin, et al.
Pubblicazione: (2023)
di: Park, Young-Jin, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Thinking in Different Spaces: Domain-Specific Latent Geometry Survives Cross-Architecture Translation
di: Armstrong, Marcus, et al.
Pubblicazione: (2026) -
The Erosion of LLM Signatures: Can We Still Distinguish Human and LLM-Generated Scientific Ideas After Iterative Paraphrasing?
di: Shahriar, Sadat, et al.
Pubblicazione: (2025) -
Say Anything but This: When Tokenizer Betrays Reasoning in LLMs
di: Ayoobi, Navid, et al.
Pubblicazione: (2026) -
Seeing Through AI's Lens: Enhancing Human Skepticism Towards LLM-Generated Fake News
di: Ayoobi, Navid, et al.
Pubblicazione: (2024) -
Argumentative Debates for Transparent Bias Detection [Technical Report]
di: Ayoobi, Hamed, et al.
Pubblicazione: (2025)