Shared DIFF Transformer
Fuente:
arXiv
Saved in:
| Main Authors: | Cang, Yueyang, Liu, Yuhang, Zhang, Xiaoteng, Shi, Li, Que, Wenge |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DINT Transformer
by: Cang, Yueyang, et al.
Published: (2025)
by: Cang, Yueyang, et al.
Published: (2025)
Graph-GRPO: Stabilizing Multi-Agent Topology Learning via Group Relative Policy Optimization
by: Cang, Yueyang, et al.
Published: (2026)
by: Cang, Yueyang, et al.
Published: (2026)
RetCompletion:High-Speed Inference Image Completion with Retentive Network
by: Cang, Yueyang, et al.
Published: (2024)
by: Cang, Yueyang, et al.
Published: (2024)
Conformal Risk Control under Non-Monotone Losses: Theory and Finite-Sample Guarantees
by: Aldirawi, Tareq, et al.
Published: (2026)
by: Aldirawi, Tareq, et al.
Published: (2026)
Intelligent Pathological Diagnosis of Gestational Trophoblastic Diseases via Visual-Language Deep Learning Model
by: Liu, Yuhang, et al.
Published: (2026)
by: Liu, Yuhang, et al.
Published: (2026)
Conformal Risk Control for Ordinal Classification
by: Xu, Yunpeng, et al.
Published: (2024)
by: Xu, Yunpeng, et al.
Published: (2024)
Mildly Conservative Q-Learning for Offline Reinforcement Learning
by: Lyu, Jiafei, et al.
Published: (2022)
by: Lyu, Jiafei, et al.
Published: (2022)
Non-Stationary Bandit Learning via Predictive Sampling
by: Liu, Yueyang, et al.
Published: (2022)
by: Liu, Yueyang, et al.
Published: (2022)
Selective Conformal Risk Control
by: Xu, Yunpeng, et al.
Published: (2025)
by: Xu, Yunpeng, et al.
Published: (2025)
AED: An black-box NLP classifier model attacker
by: Liu, Yueyang, et al.
Published: (2021)
by: Liu, Yueyang, et al.
Published: (2021)
Transfer Learning for Assessing Heavy Metal Pollution in Seaports Sediments
by: Lai, Tin, et al.
Published: (2025)
by: Lai, Tin, et al.
Published: (2025)
Label Unbalance in High-frequency Trading
by: Zhao, Zijian, et al.
Published: (2025)
by: Zhao, Zijian, et al.
Published: (2025)
Distribution-free Conformal Prediction for Ordinal Classification
by: Chakraborty, Subhrasish, et al.
Published: (2024)
by: Chakraborty, Subhrasish, et al.
Published: (2024)
Uncertainty Quantification With Multiple Sources
by: Ying, Mufang, et al.
Published: (2024)
by: Ying, Mufang, et al.
Published: (2024)
SEABO: A Simple Search-Based Method for Offline Imitation Learning
by: Lyu, Jiafei, et al.
Published: (2024)
by: Lyu, Jiafei, et al.
Published: (2024)
Single-Trajectory Distributionally Robust Reinforcement Learning
by: Liang, Zhipeng, et al.
Published: (2023)
by: Liang, Zhipeng, et al.
Published: (2023)
A Review of Data Mining in Personalized Education: Current Trends and Future Prospects
by: Xiong, Zhang, et al.
Published: (2024)
by: Xiong, Zhang, et al.
Published: (2024)
TIDFormer: Exploiting Temporal and Interactive Dynamics Makes A Great Dynamic Graph Transformer
by: Peng, Jie, et al.
Published: (2025)
by: Peng, Jie, et al.
Published: (2025)
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
by: Peng, Kunyu, et al.
Published: (2025)
by: Peng, Kunyu, et al.
Published: (2025)
Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R
by: Guerrero, Pablo Robin, et al.
Published: (2025)
by: Guerrero, Pablo Robin, et al.
Published: (2025)
When Less is More: The LLM Scaling Paradox in Context Compression
by: Guo, Ruishan, et al.
Published: (2026)
by: Guo, Ruishan, et al.
Published: (2026)
QUARK: Quantization-Enabled Circuit Sharing for Transformer Acceleration by Exploiting Common Patterns in Nonlinear Operations
by: Zhao, Zhixiong, et al.
Published: (2025)
by: Zhao, Zhixiong, et al.
Published: (2025)
SFi-Former: Sparse Flow Induced Attention for Graph Transformer
by: Li, Zhonghao, et al.
Published: (2025)
by: Li, Zhonghao, et al.
Published: (2025)
Efficient Multi-agent Reinforcement Learning by Planning
by: Liu, Qihan, et al.
Published: (2024)
by: Liu, Qihan, et al.
Published: (2024)
Causality-aware Graph Aggregation Weight Estimator for Popularity Debiasing in Top-K Recommendation
by: Que, Yue, et al.
Published: (2025)
by: Que, Yue, et al.
Published: (2025)
Two-stage Risk Control with Application to Ranked Retrieval
by: Xu, Yunpeng, et al.
Published: (2024)
by: Xu, Yunpeng, et al.
Published: (2024)
Training Machine Learning Models on Human Spatio-temporal Mobility Data: An Experimental Study [Experiment Paper]
by: Liu, Yueyang, et al.
Published: (2025)
by: Liu, Yueyang, et al.
Published: (2025)
Strategic Advice in the Age of Personal AI
by: Liu, Yueyang, et al.
Published: (2026)
by: Liu, Yueyang, et al.
Published: (2026)
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
by: Ma, Xiaoteng, et al.
Published: (2020)
by: Ma, Xiaoteng, et al.
Published: (2020)
LLM-OptiRA: LLM-Driven Optimization of Resource Allocation for Non-Convex Problems in Wireless Communications
by: Peng, Xinyue, et al.
Published: (2025)
by: Peng, Xinyue, et al.
Published: (2025)
TesseraQ: Ultra Low-Bit LLM Post-Training Quantization with Block Reconstruction
by: Li, Yuhang, et al.
Published: (2024)
by: Li, Yuhang, et al.
Published: (2024)
ChronoFormer: Time-Aware Transformer Architectures for Structured Clinical Event Modeling
by: Zhang, Yuanyun, et al.
Published: (2025)
by: Zhang, Yuanyun, et al.
Published: (2025)
Boosting LLM Reasoning via Human-Inspired Reward Shaping
by: Lin, Wenze, et al.
Published: (2026)
by: Lin, Wenze, et al.
Published: (2026)
A Comparative Study of Deep Reinforcement Learning for Crop Production Management
by: Balderas, Joseph, et al.
Published: (2024)
by: Balderas, Joseph, et al.
Published: (2024)
Rethinking Hebbian Principle: Low-Dimensional Structural Projection for Unsupervised Learning
by: Deng, Shikuang, et al.
Published: (2025)
by: Deng, Shikuang, et al.
Published: (2025)
Feature Bank Enhancement for Distance-based Out-of-Distribution Detection
by: Liu, Yuhang, et al.
Published: (2025)
by: Liu, Yuhang, et al.
Published: (2025)
ResidualTransformer: Residual Low-Rank Learning with Weight-Sharing for Transformer Layers
by: Wang, Yiming, et al.
Published: (2023)
by: Wang, Yiming, et al.
Published: (2023)
Basis Sharing: Cross-Layer Parameter Sharing for Large Language Model Compression
by: Wang, Jingcun, et al.
Published: (2024)
by: Wang, Jingcun, et al.
Published: (2024)
Task-Aware Harmony Multi-Task Decision Transformer for Offline Reinforcement Learning
by: Fan, Ziqing, et al.
Published: (2024)
by: Fan, Ziqing, et al.
Published: (2024)
InvDec: Inverted Decoder for Multivariate Time Series Forecasting with Separated Temporal and Variate Modeling
by: Wang, Yuhang
Published: (2025)
by: Wang, Yuhang
Published: (2025)
Similar Items
-
DINT Transformer
by: Cang, Yueyang, et al.
Published: (2025) -
Graph-GRPO: Stabilizing Multi-Agent Topology Learning via Group Relative Policy Optimization
by: Cang, Yueyang, et al.
Published: (2026) -
RetCompletion:High-Speed Inference Image Completion with Retentive Network
by: Cang, Yueyang, et al.
Published: (2024) -
Conformal Risk Control under Non-Monotone Losses: Theory and Finite-Sample Guarantees
by: Aldirawi, Tareq, et al.
Published: (2026) -
Intelligent Pathological Diagnosis of Gestational Trophoblastic Diseases via Visual-Language Deep Learning Model
by: Liu, Yuhang, et al.
Published: (2026)