M$^3$: Reframing Training Measures for Discretized Physical Simulations
Fuente:
arXiv
Guardado en:
| Autores principales: | Mei, Yuan, Song, Xingyu, Song, Xiaowen, Takeishi, Naoya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Deterministic Decomposition of Stochastic Generative Dynamics
por: Song, Xingyu, et al.
Publicado: (2026)
por: Song, Xingyu, et al.
Publicado: (2026)
U-Nets as Belief Propagation: Efficient Classification, Denoising, and Diffusion in Generative Hierarchical Models
por: Mei, Song
Publicado: (2024)
por: Mei, Song
Publicado: (2024)
Estimating counterfactual treatment outcomes over time in complex multiagent scenarios
por: Fujii, Keisuke, et al.
Publicado: (2022)
por: Fujii, Keisuke, et al.
Publicado: (2022)
Data-driven simulator of multi-animal behavior with unknown dynamics via offline and online reinforcement learning
por: Fujii, Keisuke, et al.
Publicado: (2025)
por: Fujii, Keisuke, et al.
Publicado: (2025)
Training-free Heterogeneous Model Merging
por: Xu, Zhengqi, et al.
Publicado: (2024)
por: Xu, Zhengqi, et al.
Publicado: (2024)
Generalized Discrete Diffusion with Self-Correction
por: Wang, Linxuan, et al.
Publicado: (2026)
por: Wang, Linxuan, et al.
Publicado: (2026)
A Neural Model of Rule Discovery with Relatively Short-Term Sequence Memory
por: Arakawa, Naoya
Publicado: (2024)
por: Arakawa, Naoya
Publicado: (2024)
Unified Molecule Pre-training with Flexible 2D and 3D Modalities: Single and Paired Modality Integration
por: Song, Tengwei, et al.
Publicado: (2025)
por: Song, Tengwei, et al.
Publicado: (2025)
Sparse Training of Discrete Diffusion Models for Graph Generation
por: Qin, Yiming, et al.
Publicado: (2023)
por: Qin, Yiming, et al.
Publicado: (2023)
L-MoE: End-to-End Training of a Lightweight Mixture of Low-Rank Adaptation Experts
por: Ji, Shihao, et al.
Publicado: (2025)
por: Ji, Shihao, et al.
Publicado: (2025)
On Measuring Long-Range Interactions in Graph Neural Networks
por: Bamberger, Jacob, et al.
Publicado: (2025)
por: Bamberger, Jacob, et al.
Publicado: (2025)
Deep Generative Models for Discrete Genotype Simulation
por: Xie, Sihan, et al.
Publicado: (2025)
por: Xie, Sihan, et al.
Publicado: (2025)
Physics in Next-token Prediction
por: An, Hongjun, et al.
Publicado: (2024)
por: An, Hongjun, et al.
Publicado: (2024)
Learning Scenario Reduction for Two-Stage Robust Optimization with Discrete Uncertainty
por: Lin, Tianjue, et al.
Publicado: (2026)
por: Lin, Tianjue, et al.
Publicado: (2026)
SlimPipe: Memory-Thrifty and Efficient Pipeline Parallelism for Long-Context LLM Training
por: Li, Zhouyang, et al.
Publicado: (2025)
por: Li, Zhouyang, et al.
Publicado: (2025)
Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation
por: Wu, Yecheng, et al.
Publicado: (2026)
por: Wu, Yecheng, et al.
Publicado: (2026)
Discrete Markov Bridge
por: Li, Hengli, et al.
Publicado: (2025)
por: Li, Hengli, et al.
Publicado: (2025)
A Theoretical Analysis of Discrete Flow Matching Generative Models
por: Su, Maojiang, et al.
Publicado: (2025)
por: Su, Maojiang, et al.
Publicado: (2025)
UNITE-FND: Reframing Multimodal Fake News Detection through Unimodal Scene Translation
por: Mukherjee, Arka, et al.
Publicado: (2025)
por: Mukherjee, Arka, et al.
Publicado: (2025)
Training Data Selection with Gradient Orthogonality for Efficient Domain Adaptation
por: Zhang, Xiyang, et al.
Publicado: (2026)
por: Zhang, Xiyang, et al.
Publicado: (2026)
GPU-accelerated simulated annealing based on p-bits with real-world device-variability modeling
por: Onizawa, Naoya, et al.
Publicado: (2026)
por: Onizawa, Naoya, et al.
Publicado: (2026)
On the Error-Correcting Effects of Stochasticity in Discrete Diffusion
por: Yuan, William, et al.
Publicado: (2026)
por: Yuan, William, et al.
Publicado: (2026)
Training Greedy Policy for Proposal Batch Selection in Expensive Multi-Objective Combinatorial Optimization
por: Lee, Deokjae, et al.
Publicado: (2024)
por: Lee, Deokjae, et al.
Publicado: (2024)
HARP: Hesitation-Aware Reframing in Transformer Inference Pass
por: Storaï, Romain, et al.
Publicado: (2024)
por: Storaï, Romain, et al.
Publicado: (2024)
Equivariant Spatio-Temporal Attentive Graph Networks to Simulate Physical Dynamics
por: Wu, Liming, et al.
Publicado: (2024)
por: Wu, Liming, et al.
Publicado: (2024)
COLA: Cross-city Mobility Transformer for Human Trajectory Simulation
por: Wang, Yu, et al.
Publicado: (2024)
por: Wang, Yu, et al.
Publicado: (2024)
Towards Quantifying Long-Range Interactions in Graph Machine Learning: a Large Graph Dataset and a Measurement
por: Liang, Huidong, et al.
Publicado: (2025)
por: Liang, Huidong, et al.
Publicado: (2025)
Simulating Environments with Reasoning Models for Agent Training
por: Li, Yuetai, et al.
Publicado: (2025)
por: Li, Yuetai, et al.
Publicado: (2025)
Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code
por: Song, Zhenghan, et al.
Publicado: (2026)
por: Song, Zhenghan, et al.
Publicado: (2026)
Training-Free Message Passing for Learning on Hypergraphs
por: Tang, Bohan, et al.
Publicado: (2024)
por: Tang, Bohan, et al.
Publicado: (2024)
Enhancing Stability for Large Language Models Training in Constrained Bandwidth Networks
por: Dai, Yun, et al.
Publicado: (2024)
por: Dai, Yun, et al.
Publicado: (2024)
Enhancing Pre-Trained Model-Based Class-Incremental Learning through Neural Collapse
por: He, Kun, et al.
Publicado: (2025)
por: He, Kun, et al.
Publicado: (2025)
Unified Algorithms for RL with Decision-Estimation Coefficients: PAC, Reward-Free, Preference-Based Learning, and Beyond
por: Chen, Fan, et al.
Publicado: (2022)
por: Chen, Fan, et al.
Publicado: (2022)
A Conditional Independence Test in the Presence of Discretization
por: Sun, Boyang, et al.
Publicado: (2024)
por: Sun, Boyang, et al.
Publicado: (2024)
Entropy-Gated Selective Policy Optimization:Token-Level Gradient Allocation for Hybrid Training of Large Language Models
por: Hu, Yuelin, et al.
Publicado: (2026)
por: Hu, Yuelin, et al.
Publicado: (2026)
GAC: Noise-Aware Adaptive Mixing for Hybrid SFT-RL Post-Training
por: Hu, Yuelin, et al.
Publicado: (2026)
por: Hu, Yuelin, et al.
Publicado: (2026)
A Framework for Quantifying How Pre-Training and Context Benefit In-Context Learning
por: Song, Bingqing, et al.
Publicado: (2025)
por: Song, Bingqing, et al.
Publicado: (2025)
DEMO: Reframing Dialogue Interaction with Fine-grained Element Modeling
por: Wang, Minzheng, et al.
Publicado: (2024)
por: Wang, Minzheng, et al.
Publicado: (2024)
Pre-Training Protein Bi-level Representation Through Span Mask Strategy On 3D Protein Chains
por: Zhao, Jiale, et al.
Publicado: (2024)
por: Zhao, Jiale, et al.
Publicado: (2024)
Achieving binary weight and activation for LLMs using Post-Training Quantization
por: Song, Siqing, et al.
Publicado: (2025)
por: Song, Siqing, et al.
Publicado: (2025)
Ejemplares similares
-
Deterministic Decomposition of Stochastic Generative Dynamics
por: Song, Xingyu, et al.
Publicado: (2026) -
U-Nets as Belief Propagation: Efficient Classification, Denoising, and Diffusion in Generative Hierarchical Models
por: Mei, Song
Publicado: (2024) -
Estimating counterfactual treatment outcomes over time in complex multiagent scenarios
por: Fujii, Keisuke, et al.
Publicado: (2022) -
Data-driven simulator of multi-animal behavior with unknown dynamics via offline and online reinforcement learning
por: Fujii, Keisuke, et al.
Publicado: (2025) -
Training-free Heterogeneous Model Merging
por: Xu, Zhengqi, et al.
Publicado: (2024)