Guardado en:
| Autores principales: | Liu, Pengwei, Hao, Zhongkai, Ren, Xingyu, Yuan, Hangjie, Ren, Jiayang, Ni, Dong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2407.05232 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Discovering Physical Directions in Weight Space: Composing Neural PDE Experts
por: Wang, Pengkai, et al.
Publicado: (2026)
por: Wang, Pengkai, et al.
Publicado: (2026)
An Efficient Graph-Transformer Operator for Learning Physical Dynamics with Manifolds Embedding
por: Liu, Pengwei, et al.
Publicado: (2025)
por: Liu, Pengwei, et al.
Publicado: (2025)
Quantum Semi-Random Forests for Qubit-Efficient Recommender Systems
por: Alavi, Azadeh, et al.
Publicado: (2025)
por: Alavi, Azadeh, et al.
Publicado: (2025)
Where We Have Arrived in Proving the Emergence of Sparse Symbolic Concepts in AI Models
por: Ren, Qihan, et al.
Publicado: (2023)
por: Ren, Qihan, et al.
Publicado: (2023)
A Global Optimization Algorithm for K-Center Clustering of One Billion Samples
por: Ren, Jiayang, et al.
Publicado: (2022)
por: Ren, Jiayang, et al.
Publicado: (2022)
Learning-driven Physically-aware Large-scale Circuit Gate Sizing
por: Ye, Yuyang, et al.
Publicado: (2024)
por: Ye, Yuyang, et al.
Publicado: (2024)
Correcting Mean Bias in Text Embeddings: A Refined Renormalization with Training-Free Improvements on MMTEB
por: Ren, Xingyu, et al.
Publicado: (2025)
por: Ren, Xingyu, et al.
Publicado: (2025)
Task Aware Dreamer for Task Generalization in Reinforcement Learning
por: Ying, Chengyang, et al.
Publicado: (2023)
por: Ying, Chengyang, et al.
Publicado: (2023)
Rethinking the Stability-Plasticity Trade-off in Continual Learning from an Architectural Perspective
por: Lu, Aojun, et al.
Publicado: (2025)
por: Lu, Aojun, et al.
Publicado: (2025)
Proxy Compression for Language Modeling
por: Zheng, Lin, et al.
Publicado: (2026)
por: Zheng, Lin, et al.
Publicado: (2026)
On the Reuse Bias in Off-Policy Reinforcement Learning
por: Ying, Chengyang, et al.
Publicado: (2022)
por: Ying, Chengyang, et al.
Publicado: (2022)
A Physics-preserved Transfer Learning Method for Differential Equations
por: Yang, Hao-Ran, et al.
Publicado: (2025)
por: Yang, Hao-Ran, et al.
Publicado: (2025)
Preconditioning for Physics-Informed Neural Networks
por: Liu, Songming, et al.
Publicado: (2024)
por: Liu, Songming, et al.
Publicado: (2024)
Task-Oriented Multimodal Token Transmission in Resource-Constrained Multiuser Networks
por: Zhang, Junhe, et al.
Publicado: (2025)
por: Zhang, Junhe, et al.
Publicado: (2025)
A Faster Path to Continual Learning
por: Li, Wei, et al.
Publicado: (2026)
por: Li, Wei, et al.
Publicado: (2026)
Learning Semantic Proxies from Visual Prompts for Parameter-Efficient Fine-Tuning in Deep Metric Learning
por: Ren, Li, et al.
Publicado: (2024)
por: Ren, Li, et al.
Publicado: (2024)
Towards Mitigating Excessive Forgetting in LLM Unlearning via Entanglement-Guidance with Proxy Constraint
por: Liu, Zhihao, et al.
Publicado: (2025)
por: Liu, Zhihao, et al.
Publicado: (2025)
A Survey on Diffusion Models for Anomaly Detection
por: Liu, Jing, et al.
Publicado: (2025)
por: Liu, Jing, et al.
Publicado: (2025)
Why Does RL Generalize Better Than SFT? A Data-Centric Perspective on VLM Post-Training
por: Lu, Aojun, et al.
Publicado: (2026)
por: Lu, Aojun, et al.
Publicado: (2026)
Curriculum Sampling: A Two-Phase Curriculum for Efficient Training of Flow Matching
por: Sun, Pengwei
Publicado: (2026)
por: Sun, Pengwei
Publicado: (2026)
Uncertainty-aware Knowledge Tracing
por: Cheng, Weihua, et al.
Publicado: (2025)
por: Cheng, Weihua, et al.
Publicado: (2025)
Verbalized Graph Representation Learning: A Fully Interpretable Graph Model Based on Large Language Models Throughout the Entire Process
por: Ji, Xingyu, et al.
Publicado: (2024)
por: Ji, Xingyu, et al.
Publicado: (2024)
Cognitive Activation and Chaotic Dynamics in Large Language Models: A Quasi-Lyapunov Analysis of Reasoning Mechanisms
por: Li, Xiaojian, et al.
Publicado: (2025)
por: Li, Xiaojian, et al.
Publicado: (2025)
Are Your Reasoning Models Reasoning or Guessing? A Mechanistic Analysis of Hierarchical Reasoning Models
por: Ren, Zirui, et al.
Publicado: (2026)
por: Ren, Zirui, et al.
Publicado: (2026)
SAMora: Enhancing SAM through Hierarchical Self-Supervised Pre-Training for Medical Images
por: Chen, Shuhang, et al.
Publicado: (2025)
por: Chen, Shuhang, et al.
Publicado: (2025)
Your Diffusion Model is Secretly a Certifiably Robust Classifier
por: Chen, Huanran, et al.
Publicado: (2024)
por: Chen, Huanran, et al.
Publicado: (2024)
Proxy-Guided Measurement Calibration
por: Vishnubhatla, Saketh, et al.
Publicado: (2026)
por: Vishnubhatla, Saketh, et al.
Publicado: (2026)
Generating In-Distribution Proxy Graphs for Explaining Graph Neural Networks
por: Chen, Zhuomin, et al.
Publicado: (2024)
por: Chen, Zhuomin, et al.
Publicado: (2024)
Revisiting Neural Networks for Continual Learning: An Architectural Perspective
por: Lu, Aojun, et al.
Publicado: (2024)
por: Lu, Aojun, et al.
Publicado: (2024)
Adapt before Continual Learning
por: Lu, Aojun, et al.
Publicado: (2025)
por: Lu, Aojun, et al.
Publicado: (2025)
Sharpness-aware Federated Graph Learning
por: Li, Ruiyu, et al.
Publicado: (2025)
por: Li, Ruiyu, et al.
Publicado: (2025)
Improved Operator Learning by Orthogonal Attention
por: Xiao, Zipeng, et al.
Publicado: (2023)
por: Xiao, Zipeng, et al.
Publicado: (2023)
Human-AI Collaborative Autonomous Experimentation With Proxy Modeling for Comparative Observation
por: Biswas, Arpan, et al.
Publicado: (2026)
por: Biswas, Arpan, et al.
Publicado: (2026)
Linking Process to Outcome: Conditional Reward Modeling for LLM Reasoning
por: Zhang, Zheng, et al.
Publicado: (2025)
por: Zhang, Zheng, et al.
Publicado: (2025)
Exploratory Diffusion Model for Unsupervised Reinforcement Learning
por: Ying, Chengyang, et al.
Publicado: (2025)
por: Ying, Chengyang, et al.
Publicado: (2025)
Modeling Latent Non-Linear Dynamical System over Time Series
por: Fujiwara, Ren, et al.
Publicado: (2024)
por: Fujiwara, Ren, et al.
Publicado: (2024)
Amortized Network Intervention to Steer the Excitatory Point Processes
por: Song, Zitao, et al.
Publicado: (2023)
por: Song, Zitao, et al.
Publicado: (2023)
Bi-directional Model Cascading with Proxy Confidence
por: Warren, David, et al.
Publicado: (2025)
por: Warren, David, et al.
Publicado: (2025)
GIT-Mol: A Multi-modal Large Language Model for Molecular Science with Graph, Image, and Text
por: Liu, Pengfei, et al.
Publicado: (2023)
por: Liu, Pengfei, et al.
Publicado: (2023)
ProxyKV: Cross-Model Proxy Pruning for Efficient Long-Context LLM Inference
por: Li, Junjie, et al.
Publicado: (2026)
por: Li, Junjie, et al.
Publicado: (2026)
Ejemplares similares
-
Discovering Physical Directions in Weight Space: Composing Neural PDE Experts
por: Wang, Pengkai, et al.
Publicado: (2026) -
An Efficient Graph-Transformer Operator for Learning Physical Dynamics with Manifolds Embedding
por: Liu, Pengwei, et al.
Publicado: (2025) -
Quantum Semi-Random Forests for Qubit-Efficient Recommender Systems
por: Alavi, Azadeh, et al.
Publicado: (2025) -
Where We Have Arrived in Proving the Emergence of Sparse Symbolic Concepts in AI Models
por: Ren, Qihan, et al.
Publicado: (2023) -
A Global Optimization Algorithm for K-Center Clustering of One Billion Samples
por: Ren, Jiayang, et al.
Publicado: (2022)