Sparsified State-Space Models are Efficient Highway Networks
Fuente:
arXiv
Salvato in:
| Autori principali: | Song, Woomin, Tack, Jihoon, Mo, Sangwoo, Oh, Seunghyuk, Shin, Jinwoo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ReVISE: Learning to Refine at Test-Time via Intrinsic Self-Verification
di: Lee, Hyunseok, et al.
Pubblicazione: (2025)
di: Lee, Hyunseok, et al.
Pubblicazione: (2025)
Hierarchical Context Merging: Better Long Context Understanding for Pre-trained LLMs
di: Song, Woomin, et al.
Pubblicazione: (2024)
di: Song, Woomin, et al.
Pubblicazione: (2024)
Optimized Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
ReMoDetect: Reward Models Recognize Aligned LLM's Generations
di: Lee, Hyunseok, et al.
Pubblicazione: (2024)
di: Lee, Hyunseok, et al.
Pubblicazione: (2024)
Tabular Transfer Learning via Prompting LLMs
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
Beyond Correctness: Learning Robust Reasoning via Transfer
di: Lee, Hyunseok, et al.
Pubblicazione: (2026)
di: Lee, Hyunseok, et al.
Pubblicazione: (2026)
Think Clearly: Improving Reasoning via Redundant Token Pruning
di: Choi, Daewon, et al.
Pubblicazione: (2025)
di: Choi, Daewon, et al.
Pubblicazione: (2025)
Online Adaptation of Language Models with a Memory of Amortized Contexts
di: Tack, Jihoon, et al.
Pubblicazione: (2024)
di: Tack, Jihoon, et al.
Pubblicazione: (2024)
ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search
di: Lee, Hyunseok, et al.
Pubblicazione: (2025)
di: Lee, Hyunseok, et al.
Pubblicazione: (2025)
Discovering and Mitigating Visual Biases through Keyword Explanation
di: Kim, Younghyun, et al.
Pubblicazione: (2023)
di: Kim, Younghyun, et al.
Pubblicazione: (2023)
Mamba Drafters for Speculative Decoding
di: Choi, Daewon, et al.
Pubblicazione: (2025)
di: Choi, Daewon, et al.
Pubblicazione: (2025)
Efficient Covariance Estimation for Sparsified Functional Data
di: Zheng, Sijie, et al.
Pubblicazione: (2025)
di: Zheng, Sijie, et al.
Pubblicazione: (2025)
Compress, Gather, and Recompute: REFORMing Long-Context Processing in Transformers
di: Song, Woomin, et al.
Pubblicazione: (2025)
di: Song, Woomin, et al.
Pubblicazione: (2025)
Sparsifying Parametric Models with L0 Regularization
di: Botteghi, Nicolò, et al.
Pubblicazione: (2024)
di: Botteghi, Nicolò, et al.
Pubblicazione: (2024)
Energy-Efficient Wireless LLM Inference via Uncertainty and Importance-Aware Speculative Decoding
di: Park, Jihoon, et al.
Pubblicazione: (2025)
di: Park, Jihoon, et al.
Pubblicazione: (2025)
One-Pass Sparsified Gaussian Mixtures
di: Kightley, Eric, et al.
Pubblicazione: (2019)
di: Kightley, Eric, et al.
Pubblicazione: (2019)
SuRe: Summarizing Retrievals using Answer Candidates for Open-domain QA of LLMs
di: Kim, Jaehyung, et al.
Pubblicazione: (2024)
di: Kim, Jaehyung, et al.
Pubblicazione: (2024)
Data-Efficient Molecular Generation with Hierarchical Textual Inversion
di: Kim, Seojin, et al.
Pubblicazione: (2024)
di: Kim, Seojin, et al.
Pubblicazione: (2024)
Highway Value Iteration Networks
di: Wang, Yuhui, et al.
Pubblicazione: (2024)
di: Wang, Yuhui, et al.
Pubblicazione: (2024)
Accelerating Reinforcement Learning with Value-Conditional State Entropy Exploration
di: Kim, Dongyoung, et al.
Pubblicazione: (2023)
di: Kim, Dongyoung, et al.
Pubblicazione: (2023)
Self-Refining Language Model Anonymizers via Adversarial Distillation
di: Kim, Kyuyoung, et al.
Pubblicazione: (2025)
di: Kim, Kyuyoung, et al.
Pubblicazione: (2025)
Formalizing the Sampling Design Space of Diffusion-Based Generative Models via Adaptive Solvers and Wasserstein-Bounded Timesteps
di: Jo, Sangwoo, et al.
Pubblicazione: (2026)
di: Jo, Sangwoo, et al.
Pubblicazione: (2026)
Two Sparse Matrices are Better than One: Sparsifying Neural Networks with Double Sparse Factorization
di: Boža, Vladimír, et al.
Pubblicazione: (2024)
di: Boža, Vladimír, et al.
Pubblicazione: (2024)
TorchSISSO: A PyTorch-Based Implementation of the Sure Independence Screening and Sparsifying Operator for Efficient and Interpretable Model Discovery
di: Muthyala, Madhav, et al.
Pubblicazione: (2024)
di: Muthyala, Madhav, et al.
Pubblicazione: (2024)
MAST: Model-Agnostic Sparsified Training
di: Demidovich, Yury, et al.
Pubblicazione: (2023)
di: Demidovich, Yury, et al.
Pubblicazione: (2023)
Stay Fair! Ensuring Group Fairness in Diffusion Models Across Guidance Scales
di: Kim, Myeongsoo, et al.
Pubblicazione: (2026)
di: Kim, Myeongsoo, et al.
Pubblicazione: (2026)
From Black-Box to White-Box: Control-Theoretic Neural Network Interpretability
di: Moon, Jihoon
Pubblicazione: (2025)
di: Moon, Jihoon
Pubblicazione: (2025)
MeSH: Memory-as-State-Highways for Recursive Transformers
di: Yu, Chengting, et al.
Pubblicazione: (2025)
di: Yu, Chengting, et al.
Pubblicazione: (2025)
Non-linear Interventions on Large Language Models
di: Kim, Sangwoo
Pubblicazione: (2026)
di: Kim, Sangwoo
Pubblicazione: (2026)
Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
di: Kim, Dongyoung, et al.
Pubblicazione: (2024)
di: Kim, Dongyoung, et al.
Pubblicazione: (2024)
Learning to Sparsify Stochastic Linear Bandits
di: Wang, Zhengmiao, et al.
Pubblicazione: (2026)
di: Wang, Zhengmiao, et al.
Pubblicazione: (2026)
Retrieval Visual Contrastive Decoding to Mitigate Object Hallucinations in Large Vision-Language Models
di: Lee, Jihoon, et al.
Pubblicazione: (2025)
di: Lee, Jihoon, et al.
Pubblicazione: (2025)
Feature Unlearning for Pre-trained GANs and VAEs
di: Moon, Saemi, et al.
Pubblicazione: (2023)
di: Moon, Saemi, et al.
Pubblicazione: (2023)
Model Medicine: A Clinical Framework for Understanding, Diagnosing, and Treating AI Models
di: Jeong, Jihoon
Pubblicazione: (2026)
di: Jeong, Jihoon
Pubblicazione: (2026)
Sparsifying Suprema of Gaussian Processes
di: De, Anindya, et al.
Pubblicazione: (2024)
di: De, Anindya, et al.
Pubblicazione: (2024)
Information-Theoretic Discrete Diffusion
di: Jeon, Moongyu, et al.
Pubblicazione: (2025)
di: Jeon, Moongyu, et al.
Pubblicazione: (2025)
Sparsified Simultaneous Confidence Intervals for High-Dimensional Linear Models
di: Zhu, Xiaorui, et al.
Pubblicazione: (2023)
di: Zhu, Xiaorui, et al.
Pubblicazione: (2023)
Efficient Process Reward Modeling via Contrastive Mutual Information
di: Lee, Nakyung, et al.
Pubblicazione: (2026)
di: Lee, Nakyung, et al.
Pubblicazione: (2026)
Elastic Spectral State Space Models for Budgeted Inference
di: Song, Dachuan, et al.
Pubblicazione: (2026)
di: Song, Dachuan, et al.
Pubblicazione: (2026)
OrthoRank: Token Selection via Sink Token Orthogonality for Efficient LLM inference
di: Shin, Seungjun, et al.
Pubblicazione: (2025)
di: Shin, Seungjun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
ReVISE: Learning to Refine at Test-Time via Intrinsic Self-Verification
di: Lee, Hyunseok, et al.
Pubblicazione: (2025) -
Hierarchical Context Merging: Better Long Context Understanding for Pre-trained LLMs
di: Song, Woomin, et al.
Pubblicazione: (2024) -
Optimized Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning
di: Nam, Jaehyun, et al.
Pubblicazione: (2024) -
ReMoDetect: Reward Models Recognize Aligned LLM's Generations
di: Lee, Hyunseok, et al.
Pubblicazione: (2024) -
Tabular Transfer Learning via Prompting LLMs
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)