vmintf/WsFFN: WsFFN Pre Alpha v0.0.2a-e Experimental Version Release
Fuente:
Zenodo
Saved in:
| Main Author: | 민성 Skystarry |
|---|---|
| Format: | Recurso digital |
| Published: |
Zenodo
2025
|
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Withania somnifera osmotin (WsOsm) confers stress tolerance in tobacco and establishes novel interactions with the defensin protein (WsDF)
by: Varinder Singh, et al.
Published: (2024)
by: Varinder Singh, et al.
Published: (2024)
Spark Transformer: Reactivating Sparsity in FFN and Attention
by: You, Chong, et al.
Published: (2025)
by: You, Chong, et al.
Published: (2025)
UMoE: Unifying Attention and FFN with Shared Experts
by: Yang, Yuanhang, et al.
Published: (2025)
by: Yang, Yuanhang, et al.
Published: (2025)
EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents
by: Sharma, Praval, et al.
Published: (2026)
by: Sharma, Praval, et al.
Published: (2026)
FFN Fusion: Rethinking Sequential Computation in Large Language Models
by: Bercovich, Akhiad, et al.
Published: (2025)
by: Bercovich, Akhiad, et al.
Published: (2025)
Fast Forward: Accelerating LLM Prefill with Predictive FFN Sparsity
by: Gautam, Aayush, et al.
Published: (2026)
by: Gautam, Aayush, et al.
Published: (2026)
LookupFFN: Making Transformers Compute-lite for CPU inference
by: Zeng, Zhanpeng, et al.
Published: (2024)
by: Zeng, Zhanpeng, et al.
Published: (2024)
Automated Evaluation of Classroom Instructional Support with LLMs and BoWs: Connecting Global Predictions to Specific Feedback
by: Whitehill, Jacob, et al.
Published: (2023)
by: Whitehill, Jacob, et al.
Published: (2023)
Longitudinal Piezoelectricity and Polarization‐Insensitive Oxidation in Janus vdWs Nb3SeI7
by: Jiapeng Wang, et al.
Published: (2024)
by: Jiapeng Wang, et al.
Published: (2024)
Analytical FFN-to-MoE Restructuring via Activation Pattern Analysis
by: Pei, Zehua, et al.
Published: (2025)
by: Pei, Zehua, et al.
Published: (2025)
FFN: a Fine-grained Chinese-English Financial Domain Parallel Corpus
by: Fu, Yuxin, et al.
Published: (2024)
by: Fu, Yuxin, et al.
Published: (2024)
EnvSSLAM-FFN: Lightweight Layer-Fused System for ESDD 2026 Challenge
by: Guo, Xiaoxuan, et al.
Published: (2025)
by: Guo, Xiaoxuan, et al.
Published: (2025)
Sparsity Moves Computation: How FFN Architecture Reshapes Attention in Small Transformers
by: Smithline, Gabriel, et al.
Published: (2026)
by: Smithline, Gabriel, et al.
Published: (2026)
Analytical Provisioning for Attention-FFN Disaggregated LLM Serving under Stochastic Workloads
by: Song, Chendong, et al.
Published: (2026)
by: Song, Chendong, et al.
Published: (2026)
pysal/gwlearn: v0.0.2
by: Martin Fleischmann, et al.
Published: (2026)
by: Martin Fleischmann, et al.
Published: (2026)
sunipkm/glowpython2: v0.0.2
by: Sunip Mukherjee
Published: (2025)
by: Sunip Mukherjee
Published: (2025)
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN
by: Xu, Yao, et al.
Published: (2025)
by: Xu, Yao, et al.
Published: (2025)
VersatileFFN: Achieving Parameter Efficiency in LLMs via Adaptive Wide-and-Deep Reuse
by: Nie, Ying, et al.
Published: (2025)
by: Nie, Ying, et al.
Published: (2025)
CAPA: Contribution-Aware Pruning and FFN Approximation for Efficient Large Vision-Language Models
by: Jha, Samyak, et al.
Published: (2026)
by: Jha, Samyak, et al.
Published: (2026)
Revealing the Challenges of Attention-FFN Disaggregation for Modern MoE Models and Hardware Systems
by: Liu, Guowei, et al.
Published: (2026)
by: Liu, Guowei, et al.
Published: (2026)
UniBias: Unveiling and Mitigating LLM Bias through Internal Attention and FFN Manipulation
by: Zhou, Hanzhang, et al.
Published: (2024)
by: Zhou, Hanzhang, et al.
Published: (2024)
gpostill/Unmet-Healthcare-Needs: Version 0.0.2
by: Gemma Postill
Published: (2026)
by: Gemma Postill
Published: (2026)
FFN-SkipLLM: A Hidden Gem for Autoregressive Decoding with Adaptive Feed Forward Skipping
by: Jaiswal, Ajay, et al.
Published: (2024)
by: Jaiswal, Ajay, et al.
Published: (2024)
Sparse-VQ Transformer: An FFN-Free Framework with Vector Quantization for Enhanced Time Series Forecasting
by: Zhao, Yanjun, et al.
Published: (2024)
by: Zhao, Yanjun, et al.
Published: (2024)
The Five Ws of Multi-Agent Communication: Who Talks to Whom, When, What, and Why -- A Survey from MARL to Emergent Language and LLMs
by: Chen, Jingdi, et al.
Published: (2026)
by: Chen, Jingdi, et al.
Published: (2026)
waudbygroup/nmr-sample-schema: v0.0.2
by: Chris Waudby
Published: (2025)
by: Chris Waudby
Published: (2025)
RevFFN: Memory-Efficient Full-Parameter Fine-Tuning of Mixture-of-Experts LLMs with Reversible Blocks
by: Liu, Ningyuan, et al.
Published: (2025)
by: Liu, Ningyuan, et al.
Published: (2025)
TreeGPT: Pure TreeFFN Encoder-Decoder Architecture for Structured Reasoning Without Attention Mechanisms
by: Li, Zixi
Published: (2025)
by: Li, Zixi
Published: (2025)
BlockFFN: Towards End-Side Acceleration-Friendly Mixture-of-Experts with Chunk-Level Activation Sparsity
by: Song, Chenyang, et al.
Published: (2025)
by: Song, Chenyang, et al.
Published: (2025)
Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs
by: Liu, Hongliang, et al.
Published: (2026)
by: Liu, Hongliang, et al.
Published: (2026)
EfficientASR: Speech Recognition Network Compression via Attention Redundancy and Chunk-Level FFN Optimization
by: Wang, Jianzong, et al.
Published: (2024)
by: Wang, Jianzong, et al.
Published: (2024)
Ayoub-dv/SIGN-AI: v0.0.2-alpha
by: Ayoub-dv, et al.
Published: (2025)
by: Ayoub-dv, et al.
Published: (2025)
Observing Weak Interchain Coupling in 1D vdWs Ternary Mo6Se2I8 to Achieve Probe Exfoliation of Ultrathin Molecular Chains
by: Yanchao Guan, et al.
Published: (2024)
by: Yanchao Guan, et al.
Published: (2024)
How does Architecture Influence the Base Capabilities of Pre-trained Language Models? A Case Study Based on FFN-Wider and MoE Transformers
by: Lu, Xin, et al.
Published: (2024)
by: Lu, Xin, et al.
Published: (2024)
umati/umatiGateway: Pre-Release v1.0.0-rc9
by: Matthias Dornaus, et al.
Published: (2026)
by: Matthias Dornaus, et al.
Published: (2026)
Time series of surface pressure at station AK-002-002, 2008-2013, Version 2.0
by: Rennermalm, Asa K, et al.
Published: (2014)
by: Rennermalm, Asa K, et al.
Published: (2014)
How Far Can Disaggregation Go? A Design-Space Exploration of Attention-FFN Disaggregation for Efficient MoE LLM Serving
by: Wu, Hanjiang, et al.
Published: (2026)
by: Wu, Hanjiang, et al.
Published: (2026)
PQC Benchmarks — Methodology Release (v0.0)
by: Sivasubramani, Santhosh, et al.
Published: (2026)
by: Sivasubramani, Santhosh, et al.
Published: (2026)
merszym/alphabet: Version v0.2
by: Merlin Szymanski
Published: (2026)
by: Merlin Szymanski
Published: (2026)
tamada/sibling: Release v2.0.4
by: Haruaki Tamada, et al.
Published: (2026)
by: Haruaki Tamada, et al.
Published: (2026)
Similar Items
-
Withania somnifera osmotin (WsOsm) confers stress tolerance in tobacco and establishes novel interactions with the defensin protein (WsDF)
by: Varinder Singh, et al.
Published: (2024) -
Spark Transformer: Reactivating Sparsity in FFN and Attention
by: You, Chong, et al.
Published: (2025) -
UMoE: Unifying Attention and FFN with Shared Experts
by: Yang, Yuanhang, et al.
Published: (2025) -
EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents
by: Sharma, Praval, et al.
Published: (2026) -
FFN Fusion: Rethinking Sequential Computation in Large Language Models
by: Bercovich, Akhiad, et al.
Published: (2025)