A Method on Searching Better Activation Functions
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Haoyuan, Wu, Zihao, Xia, Bo, Chang, Pu, Dong, Zibin, Yuan, Yifu, Chang, Yongzhe, Wang, Xueqian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Wavelet Fourier Diffuser: Frequency-Aware Diffusion Model for Reinforcement Learning
by: Luo, Yifu, et al.
Published: (2025)
by: Luo, Yifu, et al.
Published: (2025)
DEER: A Delay-Resilient Framework for Reinforcement Learning with Variable Delays
by: Xia, Bo, et al.
Published: (2024)
by: Xia, Bo, et al.
Published: (2024)
UACER: An Uncertainty-Adaptive Critic Ensemble Framework for Robust Adversarial Reinforcement Learning
by: Wu, Jiaxi, et al.
Published: (2025)
by: Wu, Jiaxi, et al.
Published: (2025)
MODULI: Unlocking Preference Generalization via Diffusion Models for Offline Multi-Objective Reinforcement Learning
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
Distribution Preference Optimization: A Fine-grained Perspective for LLM Unlearning
by: Qin, Kai, et al.
Published: (2025)
by: Qin, Kai, et al.
Published: (2025)
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models
by: Sun, Haoyuan, et al.
Published: (2025)
by: Sun, Haoyuan, et al.
Published: (2025)
CleanDiffuser: An Easy-to-use Modularized Library for Diffusion Models in Decision Making
by: Dong, Zibin, et al.
Published: (2024)
by: Dong, Zibin, et al.
Published: (2024)
Explaining the Unexplained: Revealing Hidden Correlations for Better Interpretability
by: Jiang, Wen-Dong, et al.
Published: (2024)
by: Jiang, Wen-Dong, et al.
Published: (2024)
Efficient Search for Customized Activation Functions with Gradient Descent
by: Strack, Lukas, et al.
Published: (2024)
by: Strack, Lukas, et al.
Published: (2024)
A Historical Trajectory Assisted Optimization Method for Zeroth-Order Federated Learning
by: Wu, Chenlin, et al.
Published: (2024)
by: Wu, Chenlin, et al.
Published: (2024)
ANAct: Adaptive Normalization for Activation Functions
by: Peiwen, Yuan, et al.
Published: (2022)
by: Peiwen, Yuan, et al.
Published: (2022)
Generalizing Alignment Paradigm of Text-to-Image Generation with Preferences through $f$-divergence Minimization
by: Sun, Haoyuan, et al.
Published: (2024)
by: Sun, Haoyuan, et al.
Published: (2024)
Learning Multi-Indicator Weights for Data Selection: A Joint Task-Model Adaptation Framework with Efficient Proxies
by: Song, Jingze, et al.
Published: (2026)
by: Song, Jingze, et al.
Published: (2026)
ED2: Environment Dynamics Decomposition World Models for Continuous Control
by: Hao, Jianye, et al.
Published: (2021)
by: Hao, Jianye, et al.
Published: (2021)
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
PPNet: A Two-Stage Neural Network for End-to-end Path Planning
by: Meng, Qinglong, et al.
Published: (2024)
by: Meng, Qinglong, et al.
Published: (2024)
Towards A Transferable Acceleration Method for Density Functional Theory
by: Liu, Zhe, et al.
Published: (2025)
by: Liu, Zhe, et al.
Published: (2025)
R-LoRA: Randomized Multi-Head LoRA for Efficient Multi-Task Learning
by: Liu, Jinda, et al.
Published: (2025)
by: Liu, Jinda, et al.
Published: (2025)
REAct: Rational Exponential Activation for Better Learning and Generalization in PINNs
by: Mishra, Sourav, et al.
Published: (2025)
by: Mishra, Sourav, et al.
Published: (2025)
Oversmoothing: A Nightmare for Graph Contrastive Learning?
by: Li, Jintang, et al.
Published: (2023)
by: Li, Jintang, et al.
Published: (2023)
On the Role of Transformer Feed-Forward Layers in Nonlinear In-Context Learning
by: Sun, Haoyuan, et al.
Published: (2025)
by: Sun, Haoyuan, et al.
Published: (2025)
Global Convergence in Neural ODEs: Impact of Activation Functions
by: Gao, Tianxiang, et al.
Published: (2025)
by: Gao, Tianxiang, et al.
Published: (2025)
From Projection to Prediction: Beyond Logits for Scalable Language Models
by: Dong, Jianbing, et al.
Published: (2025)
by: Dong, Jianbing, et al.
Published: (2025)
CauchyNet: Compact and Data-Efficient Learning using Holomorphic Activation Functions
by: Zhang, Hong-Kun, et al.
Published: (2025)
by: Zhang, Hong-Kun, et al.
Published: (2025)
A learning-driven automatic planning framework for proton PBS treatments of H&N cancers
by: Wang, Qingqing, et al.
Published: (2025)
by: Wang, Qingqing, et al.
Published: (2025)
Improving Offline Reinforcement Learning with Inaccurate Simulators
by: Hou, Yiwen, et al.
Published: (2024)
by: Hou, Yiwen, et al.
Published: (2024)
DFWLayer: Differentiable Frank-Wolfe Optimization Layer
by: Liu, Zixuan, et al.
Published: (2023)
by: Liu, Zixuan, et al.
Published: (2023)
BACE-RUL: A Bi-directional Adversarial Network with Covariate Encoding for Machine Remaining Useful Life Prediction
by: Zhang, Zekai, et al.
Published: (2025)
by: Zhang, Zekai, et al.
Published: (2025)
The Impact of Machine Learning Uncertainty on the Robustness of Counterfactual Explanations
by: Christodoulou, Leonidas, et al.
Published: (2026)
by: Christodoulou, Leonidas, et al.
Published: (2026)
DARE: Diffusion Language Model Activation Reuse for Efficient Inference
by: Frumkin, Natalia, et al.
Published: (2026)
by: Frumkin, Natalia, et al.
Published: (2026)
Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI
by: Cui, Haiqin, et al.
Published: (2025)
by: Cui, Haiqin, et al.
Published: (2025)
Regret-Guided Search Control for Efficient Learning in AlphaZero
by: Tsai, Yun-Jui, et al.
Published: (2026)
by: Tsai, Yun-Jui, et al.
Published: (2026)
QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search
by: Lin, Zongyu, et al.
Published: (2025)
by: Lin, Zongyu, et al.
Published: (2025)
Structured Progressive Knowledge Activation for LLM-Driven Neural Architecture Search
by: Liu, Zhen, et al.
Published: (2026)
by: Liu, Zhen, et al.
Published: (2026)
Mining Generalizable Activation Functions
by: Vitvitskyi, Alex, et al.
Published: (2026)
by: Vitvitskyi, Alex, et al.
Published: (2026)
EasyQuant: An Efficient Data-free Quantization Algorithm for LLMs
by: Tang, Hanlin, et al.
Published: (2024)
by: Tang, Hanlin, et al.
Published: (2024)
ScoresActivation: A New Activation Function for Model Agnostic Global Explainability by Design
by: Covaci, Emanuel, et al.
Published: (2025)
by: Covaci, Emanuel, et al.
Published: (2025)
Spend Less, Reason Better: Budget-Aware Value Tree Search for LLM Agents
by: Li, Yushu, et al.
Published: (2026)
by: Li, Yushu, et al.
Published: (2026)
Similar Items
-
Wavelet Fourier Diffuser: Frequency-Aware Diffusion Model for Reinforcement Learning
by: Luo, Yifu, et al.
Published: (2025) -
DEER: A Delay-Resilient Framework for Reinforcement Learning with Variable Delays
by: Xia, Bo, et al.
Published: (2024) -
UACER: An Uncertainty-Adaptive Critic Ensemble Framework for Robust Adversarial Reinforcement Learning
by: Wu, Jiaxi, et al.
Published: (2025) -
MODULI: Unlocking Preference Generalization via Diffusion Models for Offline Multi-Objective Reinforcement Learning
by: Yuan, Yifu, et al.
Published: (2024) -
Distribution Preference Optimization: A Fine-grained Perspective for LLM Unlearning
by: Qin, Kai, et al.
Published: (2025)