Saved in:
| Main Authors: | Sun, Daran, Kan, Bowen, Long, Haoquan, Zhao, Hairui, Li, Haoxu, Liu, Yicheng, Zhou, Pengyu, Feng, Ankang, Huang, Wenjing, Gu, Yida, Li, Zhenyu, Shang, Honghui, Zhang, Yunquan, Tao, Dingwen, Sun, Ninghui, Tan, Guangming |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.15768 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TSUE: A Two-Stage Data Update Method for an Erasure Coded Cluster File System
by: Wei, Zheng, et al.
Published: (2025)
by: Wei, Zheng, et al.
Published: (2025)
CCL-D: A High-Precision Diagnostic System for Slow and Hang Anomalies in Large-Scale Model Training
by: Gu, Yida, et al.
Published: (2026)
by: Gu, Yida, et al.
Published: (2026)
KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving
by: Liu, Zedong, et al.
Published: (2026)
by: Liu, Zedong, et al.
Published: (2026)
Bridging Evolutionary Multiobjective Optimization and GPU Acceleration via Tensorization
by: Liang, Zhenyu, et al.
Published: (2025)
by: Liang, Zhenyu, et al.
Published: (2025)
The (Exact) Price of Cardinality for Indivisible Goods: A Parametric Perspective
by: Lam, Alexander, et al.
Published: (2025)
by: Lam, Alexander, et al.
Published: (2025)
Fair Orientations: Proportionality and Equitability
by: Sun, Ankang, et al.
Published: (2026)
by: Sun, Ankang, et al.
Published: (2026)
Overcoming Memory Constraints in Quantum Circuit Simulation with a High-Fidelity Compression Framework
by: Zhang, Boyuan, et al.
Published: (2024)
by: Zhang, Boyuan, et al.
Published: (2024)
TACO: Efficient Communication Compression of Intermediate Tensors for Scalable Tensor-Parallel LLM Training
by: Liu, Man, et al.
Published: (2026)
by: Liu, Man, et al.
Published: (2026)
A Fair Allocation is Approximately Optimal for Indivisible Chores, or Is It?
by: Li, Bo, et al.
Published: (2024)
by: Li, Bo, et al.
Published: (2024)
Randomized Strategyproof Mechanisms with Best of Both Worlds Fairness and Efficiency
by: Sun, Ankang, et al.
Published: (2024)
by: Sun, Ankang, et al.
Published: (2024)
Pipelined Dense Symmetric Eigenvalue Decomposition on Multi-GPU Architectures
by: Wang, Hansheng, et al.
Published: (2025)
by: Wang, Hansheng, et al.
Published: (2025)
Multi-Agent Non-Discriminatory Contracts
by: Ding, Ke, et al.
Published: (2026)
by: Ding, Ke, et al.
Published: (2026)
On the Subsidy of Envy-Free Orientations in Graphs
by: Li, Bo, et al.
Published: (2025)
by: Li, Bo, et al.
Published: (2025)
Bin Packing and Covering: Pushing the Frontier on the Maximin Share Fairness
by: Li, Bo, et al.
Published: (2025)
by: Li, Bo, et al.
Published: (2025)
Towards Scalable GPU-Accelerated SNN Training via Temporal Fusion
by: Li, Yanchen, et al.
Published: (2024)
by: Li, Yanchen, et al.
Published: (2024)
Fully GPU-Accelerated, Matrix-Free Immersed Boundary Method for Complex Fiber-reinforced Hyperelastic Cardiac Models
by: Ma, Pengfei, et al.
Published: (2025)
by: Ma, Pengfei, et al.
Published: (2025)
GPU-accelerated Evolutionary Multiobjective Optimization Using Tensorized RVEA
by: Liang, Zhenyu, et al.
Published: (2024)
by: Liang, Zhenyu, et al.
Published: (2024)
GPU Acceleration of Sparse Fully Homomorphic Encrypted DNNs
by: D'Agata, Lara, et al.
Published: (2026)
by: D'Agata, Lara, et al.
Published: (2026)
ENEC: A Lossless AI Model Compression Method Enabling Fast Inference on Ascend NPUs
by: Yang, Jinwu, et al.
Published: (2026)
by: Yang, Jinwu, et al.
Published: (2026)
A Collaborative Jade Recognition System for Mobile Devices Based on Lightweight and Large Models
by: Wang, Zhenyu, et al.
Published: (2025)
by: Wang, Zhenyu, et al.
Published: (2025)
Enabling Population-Level Parallelism in Tree-Based Genetic Programming for GPU Acceleration
by: Wu, Zhihong, et al.
Published: (2025)
by: Wu, Zhihong, et al.
Published: (2025)
Tensorized NeuroEvolution of Augmenting Topologies for GPU Acceleration
by: Wang, Lishuang, et al.
Published: (2024)
by: Wang, Lishuang, et al.
Published: (2024)
ElasticMM: Efficient Multimodal LLMs Serving with Elastic Multimodal Parallelism
by: Liu, Zedong, et al.
Published: (2025)
by: Liu, Zedong, et al.
Published: (2025)
Efficient GPU-Accelerated Training of a Neuroevolution Potential with Analytical Gradients
by: Huang, Hongfu, et al.
Published: (2025)
by: Huang, Hongfu, et al.
Published: (2025)
Making Sense of Scams: Understanding Scam Conversations Through Multi-Level Alignment
by: Mao, Zhenyu, et al.
Published: (2026)
by: Mao, Zhenyu, et al.
Published: (2026)
Configurable Holography: Towards Display and Scene Adaptation
by: Zhan, Yicheng, et al.
Published: (2024)
by: Zhan, Yicheng, et al.
Published: (2024)
Speeding up Local Optimization in Vehicle Routing with Tensor-based GPU Acceleration
by: Lei, Zhenyu, et al.
Published: (2025)
by: Lei, Zhenyu, et al.
Published: (2025)
Accelerating Biclique Counting on GPU
by: Qiu, Linshan, et al.
Published: (2024)
by: Qiu, Linshan, et al.
Published: (2024)
Prompt-Induced Score Variance in Zero-Shot Binary Vision-Language Safety Classification
by: Weng, Charles, et al.
Published: (2026)
by: Weng, Charles, et al.
Published: (2026)
GPU-MetaD: Full-Life-Cycle GPU Accelerated Metadynamics with Machine Learning Potentials
by: Zhang, Haoting, et al.
Published: (2025)
by: Zhang, Haoting, et al.
Published: (2025)
Plastic instability of annular crystalline membrane in circular confinement
by: Sun, Honghui, et al.
Published: (2024)
by: Sun, Honghui, et al.
Published: (2024)
A GPU-Accelerated Hybrid Method for a Class of Multi-Depot Vehicle Routing Problems
by: Lei, Zhenyu, et al.
Published: (2026)
by: Lei, Zhenyu, et al.
Published: (2026)
A2GC: Asymmetric Aggregation with Geometric Constraints for Locally Aggregated Descriptors
by: Li, Zhenyu, et al.
Published: (2025)
by: Li, Zhenyu, et al.
Published: (2025)
Riemannian and Symplectic Geometry for Hierarchical Text-Driven Place Recognition
by: Shang, Tianyi, et al.
Published: (2026)
by: Shang, Tianyi, et al.
Published: (2026)
Adversarial Attacks on Robot Localization Systems via Deep Feature Perturbation
by: Li, Zhenyu, et al.
Published: (2026)
by: Li, Zhenyu, et al.
Published: (2026)
GPU-accelerated Evolutionary Many-objective Optimization Using Tensorized NSGA-III
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Sparsity-Accelerated Training for Large Language Models
by: Ma, Da, et al.
Published: (2024)
by: Ma, Da, et al.
Published: (2024)
FastCLIP: A Suite of Optimization Techniques to Accelerate CLIP Training with Limited Resources
by: Wei, Xiyuan, et al.
Published: (2024)
by: Wei, Xiyuan, et al.
Published: (2024)
Unfolding an Atomistic World: Atomistic Simulation of Reactor Pressure Vessel Steel Across Year-and-Meter Scales
by: Han, Haozhi, et al.
Published: (2026)
by: Han, Haozhi, et al.
Published: (2026)
Self-Cooperation Knowledge Distillation for Novel Class Discovery
by: Wang, Yuzheng, et al.
Published: (2024)
by: Wang, Yuzheng, et al.
Published: (2024)
Similar Items
-
TSUE: A Two-Stage Data Update Method for an Erasure Coded Cluster File System
by: Wei, Zheng, et al.
Published: (2025) -
CCL-D: A High-Precision Diagnostic System for Slow and Hang Anomalies in Large-Scale Model Training
by: Gu, Yida, et al.
Published: (2026) -
KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving
by: Liu, Zedong, et al.
Published: (2026) -
Bridging Evolutionary Multiobjective Optimization and GPU Acceleration via Tensorization
by: Liang, Zhenyu, et al.
Published: (2025) -
The (Exact) Price of Cardinality for Indivisible Goods: A Parametric Perspective
by: Lam, Alexander, et al.
Published: (2025)