Improving Generalization by Permutation Routing Across Model Copies
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kashiwamura, Shuhei, Leleu, Timothee |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
High-Dimensional Learning Dynamics of Quantized Models with Straight-Through Estimator
von: Ichikawa, Yuma, et al.
Veröffentlicht: (2025)
von: Ichikawa, Yuma, et al.
Veröffentlicht: (2025)
Contrastive Concept-Tree Search for LLM-Assisted Algorithm Discovery
von: Leleu, Timothee, et al.
Veröffentlicht: (2026)
von: Leleu, Timothee, et al.
Veröffentlicht: (2026)
AI Methods for Permutation Circuit Synthesis Across Generic Topologies
von: Villar, Victor, et al.
Veröffentlicht: (2025)
von: Villar, Victor, et al.
Veröffentlicht: (2025)
Effect of Weight Quantization on Learning Models by Typical Case Analysis
von: Kashiwamura, Shuhei, et al.
Veröffentlicht: (2024)
von: Kashiwamura, Shuhei, et al.
Veröffentlicht: (2024)
Improving Generalization of Neural Vehicle Routing Problem Solvers Through the Lens of Model Architecture
von: Xiao, Yubin, et al.
Veröffentlicht: (2024)
von: Xiao, Yubin, et al.
Veröffentlicht: (2024)
SwinGNN: Rethinking Permutation Invariance in Diffusion Models for Graph Generation
von: Yan, Qi, et al.
Veröffentlicht: (2023)
von: Yan, Qi, et al.
Veröffentlicht: (2023)
Derivation of Closed Form of Expected Improvement for Gaussian Process Trained on Log-Transformed Objective
von: Watanabe, Shuhei
Veröffentlicht: (2024)
von: Watanabe, Shuhei
Veröffentlicht: (2024)
Tree-Structured Parzen Estimator: Understanding Its Algorithm Components and Their Roles for Better Empirical Performance
von: Watanabe, Shuhei
Veröffentlicht: (2023)
von: Watanabe, Shuhei
Veröffentlicht: (2023)
Approximation of Box Decomposition Algorithm for Fast Hypervolume-Based Multi-Objective Optimization
von: Watanabe, Shuhei
Veröffentlicht: (2025)
von: Watanabe, Shuhei
Veröffentlicht: (2025)
Derivation of Output Correlation Inferences for Multi-Output (aka Multi-Task) Gaussian Process
von: Watanabe, Shuhei
Veröffentlicht: (2025)
von: Watanabe, Shuhei
Veröffentlicht: (2025)
Language Models "Grok" to Copy
von: Lv, Ang, et al.
Veröffentlicht: (2024)
von: Lv, Ang, et al.
Veröffentlicht: (2024)
The Impossibility of Inverse Permutation Learning in Transformer Models
von: Alur, Rohan, et al.
Veröffentlicht: (2025)
von: Alur, Rohan, et al.
Veröffentlicht: (2025)
eyeballvul: a future-proof benchmark for vulnerability detection in the wild
von: Chauvin, Timothee
Veröffentlicht: (2024)
von: Chauvin, Timothee
Veröffentlicht: (2024)
Scalable Permutation-Aware Modeling for Temporal Set Prediction
von: Ranjan, Ashish, et al.
Veröffentlicht: (2025)
von: Ranjan, Ashish, et al.
Veröffentlicht: (2025)
Monte Carlo Permutation Search
von: Cazenave, Tristan
Veröffentlicht: (2025)
von: Cazenave, Tristan
Veröffentlicht: (2025)
Neural Inhibition Improves Dynamic Routing and Mixture of Experts
von: Zou, Will Y., et al.
Veröffentlicht: (2025)
von: Zou, Will Y., et al.
Veröffentlicht: (2025)
Uncertainty-Aware Sparse Identification of Dynamical Systems via Bayesian Model Averaging
von: Kashiwamura, Shuhei, et al.
Veröffentlicht: (2026)
von: Kashiwamura, Shuhei, et al.
Veröffentlicht: (2026)
Permutation Equivariant Model-based Offline Reinforcement Learning for Auto-bidding
von: Mou, Zhiyu, et al.
Veröffentlicht: (2025)
von: Mou, Zhiyu, et al.
Veröffentlicht: (2025)
c-TPE: Tree-structured Parzen Estimator with Inequality Constraints for Expensive Hyperparameter Optimization
von: Watanabe, Shuhei, et al.
Veröffentlicht: (2022)
von: Watanabe, Shuhei, et al.
Veröffentlicht: (2022)
Towards Generalization-Oriented Models for Vehicle Routing Problems with Mixture-of-Experts
von: Miao, Changhao, et al.
Veröffentlicht: (2026)
von: Miao, Changhao, et al.
Veröffentlicht: (2026)
Prompt Learning for Generalized Vehicle Routing
von: Liu, Fei, et al.
Veröffentlicht: (2024)
von: Liu, Fei, et al.
Veröffentlicht: (2024)
Learning Unbiased Permutations via Flow Matching
von: Min, Yimeng, et al.
Veröffentlicht: (2026)
von: Min, Yimeng, et al.
Veröffentlicht: (2026)
Permutation Picture of Graph Combinatorial Optimization Problems
von: Min, Yimeng
Veröffentlicht: (2024)
von: Min, Yimeng
Veröffentlicht: (2024)
PermLLM: Learnable Channel Permutation for N:M Sparse Large Language Models
von: Zou, Lancheng, et al.
Veröffentlicht: (2025)
von: Zou, Lancheng, et al.
Veröffentlicht: (2025)
Gaussian Match-and-Copy: A Minimalist Benchmark for Studying Transformer Induction
von: Gonon, Antoine, et al.
Veröffentlicht: (2026)
von: Gonon, Antoine, et al.
Veröffentlicht: (2026)
ProxRouter: Proximity-Weighted LLM Query Routing for Improved Robustness to Outliers
von: Patel, Shivam, et al.
Veröffentlicht: (2025)
von: Patel, Shivam, et al.
Veröffentlicht: (2025)
Learning Permutation Distributions via Reflected Diffusion on Ranks
von: He, Sizhuang, et al.
Veröffentlicht: (2026)
von: He, Sizhuang, et al.
Veröffentlicht: (2026)
Permutation Invariant Learning with High-Dimensional Particle Filters
von: Boopathy, Akhilan, et al.
Veröffentlicht: (2024)
von: Boopathy, Akhilan, et al.
Veröffentlicht: (2024)
Structure As Search: Unsupervised Permutation Learning for Combinatorial Optimization
von: Min, Yimeng, et al.
Veröffentlicht: (2025)
von: Min, Yimeng, et al.
Veröffentlicht: (2025)
More Than Routing: Joint GPS and Route Modeling for Refine Trajectory Representation Learning
von: Ma, Zhipeng, et al.
Veröffentlicht: (2024)
von: Ma, Zhipeng, et al.
Veröffentlicht: (2024)
One Permutation Is All You Need: Fast, Reliable Variable Importance and Model Stress-Testing
von: Dorador, Albert
Veröffentlicht: (2025)
von: Dorador, Albert
Veröffentlicht: (2025)
General Demographic Foundation Models for Enhancing Predictive Performance Across Diseases and Populations
von: Chen, Li-Chin, et al.
Veröffentlicht: (2025)
von: Chen, Li-Chin, et al.
Veröffentlicht: (2025)
Keep Everyone Happy: Online Fair Division of Numerous Items with Few Copies
von: Verma, Arun, et al.
Veröffentlicht: (2024)
von: Verma, Arun, et al.
Veröffentlicht: (2024)
Conditional PED-ANOVA: Hyperparameter Importance in Hierarchical & Dynamic Search Spaces
von: Baba, Kaito, et al.
Veröffentlicht: (2026)
von: Baba, Kaito, et al.
Veröffentlicht: (2026)
Tree-Structured Parzen Estimator Can Solve Black-Box Combinatorial Optimization More Efficiently
von: Abe, Kenshin, et al.
Veröffentlicht: (2025)
von: Abe, Kenshin, et al.
Veröffentlicht: (2025)
Batch Acquisition Function Evaluations and Decouple Optimizer Updates for Faster Bayesian Optimization
von: Irie, Kaichi, et al.
Veröffentlicht: (2025)
von: Irie, Kaichi, et al.
Veröffentlicht: (2025)
OptunaHub: A Platform for Black-Box Optimization
von: Ozaki, Yoshihiko, et al.
Veröffentlicht: (2025)
von: Ozaki, Yoshihiko, et al.
Veröffentlicht: (2025)
DeepSpeed Data Efficiency: Improving Deep Learning Model Quality and Training Efficiency via Efficient Data Sampling and Routing
von: Li, Conglong, et al.
Veröffentlicht: (2022)
von: Li, Conglong, et al.
Veröffentlicht: (2022)
Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations
von: Sverdlov, Yonatan, et al.
Veröffentlicht: (2024)
von: Sverdlov, Yonatan, et al.
Veröffentlicht: (2024)
Toward Efficient Permutation for Hierarchical N:M Sparsity on GPUs
von: Yu, Seungmin, et al.
Veröffentlicht: (2024)
von: Yu, Seungmin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
High-Dimensional Learning Dynamics of Quantized Models with Straight-Through Estimator
von: Ichikawa, Yuma, et al.
Veröffentlicht: (2025) -
Contrastive Concept-Tree Search for LLM-Assisted Algorithm Discovery
von: Leleu, Timothee, et al.
Veröffentlicht: (2026) -
AI Methods for Permutation Circuit Synthesis Across Generic Topologies
von: Villar, Victor, et al.
Veröffentlicht: (2025) -
Effect of Weight Quantization on Learning Models by Typical Case Analysis
von: Kashiwamura, Shuhei, et al.
Veröffentlicht: (2024) -
Improving Generalization of Neural Vehicle Routing Problem Solvers Through the Lens of Model Architecture
von: Xiao, Yubin, et al.
Veröffentlicht: (2024)