SAS: Simulated Attention Score
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Chuanyang, Sun, Jiankai, Gao, Yihang, Wang, Yuehao, Wang, Peihao, Xiong, Jing, Ren, Liliang, Cheng, Hao, Kulkarni, Janardhan, Shen, Yelong, Wang, Atlas, Schwager, Mac, Schneider, Anderson, Liu, Xiaodong, Gao, Jianfeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cubit: Token Mixer with Kernel Ridge Regression
by: Zheng, Chuanyang, et al.
Published: (2026)
by: Zheng, Chuanyang, et al.
Published: (2026)
GeoNorm: Unify Pre-Norm and Post-Norm with Geodesic Optimization
by: Zheng, Chuanyang, et al.
Published: (2026)
by: Zheng, Chuanyang, et al.
Published: (2026)
Understanding the Mixture-of-Experts with Nadaraya-Watson Kernel
by: Zheng, Chuanyang, et al.
Published: (2025)
by: Zheng, Chuanyang, et al.
Published: (2025)
DAPE V2: Process Attention Score as Feature Map for Length Extrapolation
by: Zheng, Chuanyang, et al.
Published: (2024)
by: Zheng, Chuanyang, et al.
Published: (2024)
Aria-NeRF: Multimodal Egocentric View Synthesis
by: Sun, Jiankai, et al.
Published: (2023)
by: Sun, Jiankai, et al.
Published: (2023)
FAST-Splat: Fast, Ambiguity-Free Semantics Transfer in Gaussian Splatting
by: Shorinwa, Ola, et al.
Published: (2024)
by: Shorinwa, Ola, et al.
Published: (2024)
Rethinking Language Model Scaling under Transferable Hypersphere Optimization
by: Ren, Liliang, et al.
Published: (2026)
by: Ren, Liliang, et al.
Published: (2026)
SIREN: Semantic, Initialization-Free Registration of Multi-Robot Gaussian Splatting Maps
by: Shorinwa, Ola, et al.
Published: (2025)
by: Shorinwa, Ola, et al.
Published: (2025)
Early Discoveries of Algorithmist I: Promise of Provable Algorithm Synthesis at Scale
by: Kulkarni, Janardhan
Published: (2026)
by: Kulkarni, Janardhan
Published: (2026)
GRaD-Nav: Efficiently Learning Visual Drone Navigation with Gaussian Radiance Fields and Differentiable Dynamics
by: Chen, Qianzhong, et al.
Published: (2025)
by: Chen, Qianzhong, et al.
Published: (2025)
Decoder-Hybrid-Decoder Architecture for Efficient Reasoning with Long Generation
by: Ren, Liliang, et al.
Published: (2025)
by: Ren, Liliang, et al.
Published: (2025)
ParticleFormer: A 3D Point Cloud World Model for Multi-Object, Multi-Material Robotic Manipulation
by: Huang, Suning, et al.
Published: (2025)
by: Huang, Suning, et al.
Published: (2025)
PERRY: Policy Evaluation with Confidence Intervals using Auxiliary Data
by: Mandyam, Aishwarya, et al.
Published: (2025)
by: Mandyam, Aishwarya, et al.
Published: (2025)
GRaD-Nav++: Vision-Language Model Enabled Visual Drone Navigation with Gaussian Radiance Fields and Differentiable Dynamics
by: Chen, Qianzhong, et al.
Published: (2025)
by: Chen, Qianzhong, et al.
Published: (2025)
Dyna-Mind: Learning to Simulate from Experience for Better AI Agents
by: Yu, Xiao, et al.
Published: (2025)
by: Yu, Xiao, et al.
Published: (2025)
The Linear Attention Resurrection in Vision Transformer
by: Zheng, Chuanyang
Published: (2025)
by: Zheng, Chuanyang
Published: (2025)
Routing Mamba: Scaling State Space Models with Mixture-of-Experts Projection
by: Zhan, Zheng, et al.
Published: (2025)
by: Zhan, Zheng, et al.
Published: (2025)
Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling
by: Ren, Liliang, et al.
Published: (2024)
by: Ren, Liliang, et al.
Published: (2024)
Shuffle the Context: RoPE-Perturbed Self-Distillation for Long-Context Adaptation
by: Li, Zichong, et al.
Published: (2026)
by: Li, Zichong, et al.
Published: (2026)
LEGS: Fine-Tuning Teleop-Free VLAs for Humanoid Loco-manipulation in an Embodied Gaussian Splatting World
by: Kim, Hojune, et al.
Published: (2026)
by: Kim, Hojune, et al.
Published: (2026)
Reinforcement Learning for Reasoning in Large Language Models with One Training Example
by: Wang, Yiping, et al.
Published: (2025)
by: Wang, Yiping, et al.
Published: (2025)
Self-Adjust Softmax
by: Zheng, Chuanyang, et al.
Published: (2025)
by: Zheng, Chuanyang, et al.
Published: (2025)
Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training
by: Huang, Suning, et al.
Published: (2026)
by: Huang, Suning, et al.
Published: (2026)
Oscillation Inversion: Understand the structure of Large Flow Model through the Lens of Inversion Method
by: Zheng, Yan, et al.
Published: (2024)
by: Zheng, Yan, et al.
Published: (2024)
Understanding and Mitigating Bottlenecks of State Space Models through the Lens of Recency and Over-smoothing
by: Wang, Peihao, et al.
Published: (2024)
by: Wang, Peihao, et al.
Published: (2024)
Phi-4-Mini-Reasoning: Exploring the Limits of Small Reasoning Language Models in Math
by: Xu, Haoran, et al.
Published: (2025)
by: Xu, Haoran, et al.
Published: (2025)
Distributed Conjugate Gradient Method via Conjugate Direction Tracking
by: Shorinwa, Ola, et al.
Published: (2023)
by: Shorinwa, Ola, et al.
Published: (2023)
Distributed Quasi-Newton Method for Multi-Agent Optimization
by: Shorinwa, Ola, et al.
Published: (2024)
by: Shorinwa, Ola, et al.
Published: (2024)
Guarantees on Robot System Performance Using Stochastic Simulation Rollouts
by: Vincent, Joseph A., et al.
Published: (2023)
by: Vincent, Joseph A., et al.
Published: (2023)
DAPE: Data-Adaptive Positional Encoding for Length Extrapolation
by: Zheng, Chuanyang, et al.
Published: (2024)
by: Zheng, Chuanyang, et al.
Published: (2024)
Sequential Bayesian parameter-state estimation in dynamical systems with noisy and incomplete observations via a variational framework
by: Wang, Liliang, et al.
Published: (2025)
by: Wang, Liliang, et al.
Published: (2025)
Coverage Optimization for Camera View Selection
by: Chen, Timothy, et al.
Published: (2026)
by: Chen, Timothy, et al.
Published: (2026)
SpatMCDA : An R package for assessing areas at risk of infectious diseases based on spatial multi‐criteria decision analysis
by: Haoran Wang, et al.
Published: (2024)
by: Haoran Wang, et al.
Published: (2024)
Reachable Polyhedral Marching (RPM): An Exact Analysis Tool for Deep-Learned Control Systems
by: Vincent, Joseph A., et al.
Published: (2022)
by: Vincent, Joseph A., et al.
Published: (2022)
FlexGS: Train Once, Deploy Everywhere with Many-in-One Flexible 3D Gaussian Splatting
by: Liu, Hengyu, et al.
Published: (2025)
by: Liu, Hengyu, et al.
Published: (2025)
Why Neural Network Can Discover Symbolic Structures with Gradient-based Training: An Algebraic and Geometric Foundation for Neurosymbolic Reasoning
by: Wang, Peihao, et al.
Published: (2025)
by: Wang, Peihao, et al.
Published: (2025)
Protein and Peptide‐Based Strategies for Advanced Cryopreservation
by: Yihang Gao, et al.
Published: (2026)
by: Yihang Gao, et al.
Published: (2026)
Pruned Convolutional Attention Network Based Wideband Spectrum Sensing with Sub-Nyquist Sampling
by: Dong, Peihao, et al.
Published: (2024)
by: Dong, Peihao, et al.
Published: (2024)
A global Lipschitz stability perspective for understanding approximate approaches in Bayesian sequential learning
by: Wang, Liliang, et al.
Published: (2025)
by: Wang, Liliang, et al.
Published: (2025)
SAS-Bench: A Fine-Grained Benchmark for Evaluating Short Answer Scoring with Large Language Models
by: Lai, Peichao, et al.
Published: (2025)
by: Lai, Peichao, et al.
Published: (2025)
Similar Items
-
Cubit: Token Mixer with Kernel Ridge Regression
by: Zheng, Chuanyang, et al.
Published: (2026) -
GeoNorm: Unify Pre-Norm and Post-Norm with Geodesic Optimization
by: Zheng, Chuanyang, et al.
Published: (2026) -
Understanding the Mixture-of-Experts with Nadaraya-Watson Kernel
by: Zheng, Chuanyang, et al.
Published: (2025) -
DAPE V2: Process Attention Score as Feature Map for Length Extrapolation
by: Zheng, Chuanyang, et al.
Published: (2024) -
Aria-NeRF: Multimodal Egocentric View Synthesis
by: Sun, Jiankai, et al.
Published: (2023)