SIMPLE: A Gradient Estimator for $k$-Subset Sampling
Fuente:
arXiv
Saved in:
| Main Authors: | Ahmed, Kareem, Zeng, Zhe, Niepert, Mathias, Broeck, Guy Van den |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Collapsed Inference for Bayesian Deep Learning
by: Zeng, Zhe, et al.
Published: (2023)
by: Zeng, Zhe, et al.
Published: (2023)
A Pseudo-Semantic Loss for Autoregressive Models with Logical Constraints
by: Ahmed, Kareem, et al.
Published: (2023)
by: Ahmed, Kareem, et al.
Published: (2023)
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
by: Zhao, Heng, et al.
Published: (2026)
by: Zhao, Heng, et al.
Published: (2026)
Rao-Blackwell Gradient Estimators for Equivariant Denoising Diffusion
by: Tong, Vinh, et al.
Published: (2025)
by: Tong, Vinh, et al.
Published: (2025)
On the Relationship Between Monotone and Squared Probabilistic Circuits
by: Wang, Benjie, et al.
Published: (2024)
by: Wang, Benjie, et al.
Published: (2024)
Preference-Based Gradient Estimation for ML-Guided Approximate Combinatorial Optimization
by: Mielke, Arman, et al.
Published: (2025)
by: Mielke, Arman, et al.
Published: (2025)
Probabilistically Rewired Message-Passing Neural Networks
by: Qian, Chendi, et al.
Published: (2023)
by: Qian, Chendi, et al.
Published: (2023)
Image Inpainting via Tractable Steering of Diffusion Models
by: Liu, Anji, et al.
Published: (2023)
by: Liu, Anji, et al.
Published: (2023)
Scaling Up Probabilistic Circuits by Latent Variable Distillation
by: Liu, Anji, et al.
Published: (2022)
by: Liu, Anji, et al.
Published: (2022)
How to Marginalize in Causal Structure Learning?
by: Zhao, William, et al.
Published: (2025)
by: Zhao, William, et al.
Published: (2025)
A Tractable Inference Perspective of Offline RL
by: Liu, Xuejie, et al.
Published: (2023)
by: Liu, Xuejie, et al.
Published: (2023)
Discrete Copula Diffusion
by: Liu, Anji, et al.
Published: (2024)
by: Liu, Anji, et al.
Published: (2024)
L2XGNN: Learning to Explain Graph Neural Networks
by: Serra, Giuseppe, et al.
Published: (2022)
by: Serra, Giuseppe, et al.
Published: (2022)
Tractable Probabilistic Graph Representation Learning with Graph-Induced Sum-Product Networks
by: Errica, Federico, et al.
Published: (2023)
by: Errica, Federico, et al.
Published: (2023)
Enabling Autoregressive Models to Fill In Masked Tokens
by: Israel, Daniel, et al.
Published: (2025)
by: Israel, Daniel, et al.
Published: (2025)
Restructuring Tractable Probabilistic Circuits
by: Zhang, Honghua, et al.
Published: (2024)
by: Zhang, Honghua, et al.
Published: (2024)
Scaling Tractable Probabilistic Circuits: A Systems Perspective
by: Liu, Anji, et al.
Published: (2024)
by: Liu, Anji, et al.
Published: (2024)
Accelerating Diffusion LLMs via Adaptive Parallel Decoding
by: Israel, Daniel, et al.
Published: (2025)
by: Israel, Daniel, et al.
Published: (2025)
Adversarial Tokenization
by: Geh, Renato Lui, et al.
Published: (2025)
by: Geh, Renato Lui, et al.
Published: (2025)
Adaptive Physics-informed Neural Networks: A Survey
by: Torres, Edgar, et al.
Published: (2025)
by: Torres, Edgar, et al.
Published: (2025)
MolMix: A Simple Yet Effective Baseline for Multimodal Molecular Representation Learning
by: Manolache, Andrei, et al.
Published: (2024)
by: Manolache, Andrei, et al.
Published: (2024)
Prepacking: A Simple Method for Fast Prefilling and Increased Throughput in Large Language Models
by: Zhao, Siyan, et al.
Published: (2024)
by: Zhao, Siyan, et al.
Published: (2024)
Probabilistic Circuits for Cumulative Distribution Functions
by: Broadrick, Oliver, et al.
Published: (2024)
by: Broadrick, Oliver, et al.
Published: (2024)
The Pitfalls of KV Cache Compression
by: Chen, Alex, et al.
Published: (2025)
by: Chen, Alex, et al.
Published: (2025)
Learning the Neighborhood: Contrast-Free Multimodal Self-Supervised Molecular Graph Pretraining
by: Ariguib, Boshra, et al.
Published: (2025)
by: Ariguib, Boshra, et al.
Published: (2025)
Breaking the Factorization Barrier in Diffusion Language Models
by: Li, Ian, et al.
Published: (2026)
by: Li, Ian, et al.
Published: (2026)
Learning to Discretize Denoising Diffusion ODEs
by: Tong, Vinh, et al.
Published: (2024)
by: Tong, Vinh, et al.
Published: (2024)
A Compositional Atlas for Algebraic Circuits
by: Wang, Benjie, et al.
Published: (2024)
by: Wang, Benjie, et al.
Published: (2024)
Learning (Approximately) Equivariant Networks via Constrained Optimization
by: Manolache, Andrei, et al.
Published: (2025)
by: Manolache, Andrei, et al.
Published: (2025)
Controllable Generation via Locally Constrained Resampling
by: Ahmed, Kareem, et al.
Published: (2024)
by: Ahmed, Kareem, et al.
Published: (2024)
SMART: Scalable Mesh-free Aerodynamic Simulations from Raw Geometries using a Transformer-based Surrogate Model
by: Hagnberger, Jan, et al.
Published: (2026)
by: Hagnberger, Jan, et al.
Published: (2026)
Adaptive Width Neural Networks
by: Errica, Federico, et al.
Published: (2025)
by: Errica, Federico, et al.
Published: (2025)
Tractable Transformers for Flexible Conditional Generation
by: Liu, Anji, et al.
Published: (2025)
by: Liu, Anji, et al.
Published: (2025)
Zero-Variance Gradients for Variational Autoencoders
by: Shao, Zilei, et al.
Published: (2025)
by: Shao, Zilei, et al.
Published: (2025)
LOGLO-FNO: Efficient Learning of Local and Global Features in Fourier Neural Operators
by: Kalimuthu, Marimuthu, et al.
Published: (2025)
by: Kalimuthu, Marimuthu, et al.
Published: (2025)
SymDrift: One-Shot Generative Modeling under Symmetries
by: Darouich, Samir, et al.
Published: (2026)
by: Darouich, Samir, et al.
Published: (2026)
Deep Generative Models with Hard Linear Equality Constraints
by: Li, Ruoyan, et al.
Published: (2025)
by: Li, Ruoyan, et al.
Published: (2025)
CALM-PDE: Continuous and Adaptive Convolutions for Latent Space Modeling of Time-dependent PDEs
by: Hagnberger, Jan, et al.
Published: (2025)
by: Hagnberger, Jan, et al.
Published: (2025)
Where is the signal in tokenization space?
by: Geh, Renato Lui, et al.
Published: (2024)
by: Geh, Renato Lui, et al.
Published: (2024)
Variational Learning of Gaussian Process Latent Variable Models through Stochastic Gradient Annealed Importance Sampling
by: Xu, Jian, et al.
Published: (2024)
by: Xu, Jian, et al.
Published: (2024)
Similar Items
-
Collapsed Inference for Bayesian Deep Learning
by: Zeng, Zhe, et al.
Published: (2023) -
A Pseudo-Semantic Loss for Autoregressive Models with Logical Constraints
by: Ahmed, Kareem, et al.
Published: (2023) -
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
by: Zhao, Heng, et al.
Published: (2026) -
Rao-Blackwell Gradient Estimators for Equivariant Denoising Diffusion
by: Tong, Vinh, et al.
Published: (2025) -
On the Relationship Between Monotone and Squared Probabilistic Circuits
by: Wang, Benjie, et al.
Published: (2024)