Improving Discrete Optimisation Via Decoupled Straight-Through Estimator
Fuente:
arXiv
Saved in:
| Main Authors: | Shah, Rushi, Yan, Mingyuan, Mozer, Michael Curtis, Liu, Dianbo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Representation Collapsing Problems in Vector Quantization
by: Zhao, Wenhao, et al.
Published: (2024)
by: Zhao, Wenhao, et al.
Published: (2024)
Gaussian Mixture Vector Quantization with Aggregated Categorical Posterior
by: Yan, Mingyuan, et al.
Published: (2024)
by: Yan, Mingyuan, et al.
Published: (2024)
Early Quantization Shrinks Codebook: A Simple Fix for Diversity-Preserving Tokenization
by: Zhao, Wenhao, et al.
Published: (2026)
by: Zhao, Wenhao, et al.
Published: (2026)
VQSynery: Robust Drug Synergy Prediction With Vector Quantization Mechanism
by: Wu, Jiawei, et al.
Published: (2024)
by: Wu, Jiawei, et al.
Published: (2024)
Explore-Execute Chain: Towards an Efficient Structured Reasoning Paradigm
by: Yang, Kaisen, et al.
Published: (2025)
by: Yang, Kaisen, et al.
Published: (2025)
Deconstructing Generative Diversity: An Information Bottleneck Analysis of Discrete Latent Generative Models
by: Wu, Yudi, et al.
Published: (2025)
by: Wu, Yudi, et al.
Published: (2025)
Performance Asymmetry in Model-Based Reinforcement Learning
by: Lim, Jing Yu, et al.
Published: (2025)
by: Lim, Jing Yu, et al.
Published: (2025)
Decoupling the "What" and "Where" With Polar Coordinate Positional Embeddings
by: Gopalakrishnan, Anand, et al.
Published: (2025)
by: Gopalakrishnan, Anand, et al.
Published: (2025)
Extending Straight-Through Estimation for Robust Neural Networks on Analog CIM Hardware
by: Feng, Yuannuo, et al.
Published: (2025)
by: Feng, Yuannuo, et al.
Published: (2025)
SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation
by: Vetcha, Nitin, et al.
Published: (2026)
by: Vetcha, Nitin, et al.
Published: (2026)
The Topological Trouble With Transformers
by: Mozer, Michael C., et al.
Published: (2026)
by: Mozer, Michael C., et al.
Published: (2026)
From Dormant to Deleted: Tamper-Resistant Unlearning Through Weight-Space Regularization
by: Siddiqui, Shoaib Ahmed, et al.
Published: (2025)
by: Siddiqui, Shoaib Ahmed, et al.
Published: (2025)
Masked Generative Priors Improve World Models Sequence Modelling Capabilities
by: Meo, Cristian, et al.
Published: (2024)
by: Meo, Cristian, et al.
Published: (2024)
On the Structural Limitations of Weight-Based Neural Adaptation and the Role of Reversible Behavioral Learning
by: Konduru, Pardhu Sri Rushi Varma
Published: (2026)
by: Konduru, Pardhu Sri Rushi Varma
Published: (2026)
Evolution Guided Generative Flow Networks
by: Ikram, Zarif, et al.
Published: (2024)
by: Ikram, Zarif, et al.
Published: (2024)
Expected Return Causes Outcome-Level Mode Collapse in Reinforcement Learning and How to Fix It with Inverse Probability Scaling
by: Sinha, Abhijeet, et al.
Published: (2026)
by: Sinha, Abhijeet, et al.
Published: (2026)
High-Dimensional Learning Dynamics of Quantized Models with Straight-Through Estimator
by: Ichikawa, Yuma, et al.
Published: (2025)
by: Ichikawa, Yuma, et al.
Published: (2025)
Decoupled Hierarchical Reinforcement Learning with State Abstraction for Discrete Grids
by: Xiao, Qingyu, et al.
Published: (2025)
by: Xiao, Qingyu, et al.
Published: (2025)
Unlearning via Sparse Representations
by: Shah, Vedant, et al.
Published: (2023)
by: Shah, Vedant, et al.
Published: (2023)
STORI: A Benchmark and Taxonomy for Stochastic Environments
by: Barsainyan, Aryan Amit, et al.
Published: (2025)
by: Barsainyan, Aryan Amit, et al.
Published: (2025)
Quotient DAGs for Off-Policy Evaluation:Forward-Flow Importance Sampling and Exact Slate Propensities
by: Xie, Ziwen, et al.
Published: (2026)
by: Xie, Ziwen, et al.
Published: (2026)
Straight-Through meets Sparse Recovery: the Support Exploration Algorithm
by: Mohamed, Mimoun, et al.
Published: (2023)
by: Mohamed, Mimoun, et al.
Published: (2023)
SPI-GAN: Denoising Diffusion GANs with Straight-Path Interpolations
by: Jeon, Jinsung, et al.
Published: (2022)
by: Jeon, Jinsung, et al.
Published: (2022)
Auto-Discovery-Bench: Diagnosing Structured State Tracking in Oracle-Guided Discovery
by: Chen, Tingting, et al.
Published: (2025)
by: Chen, Tingting, et al.
Published: (2025)
Straight-Line Diffusion Model for Efficient 3D Molecular Generation
by: Ni, Yuyan, et al.
Published: (2025)
by: Ni, Yuyan, et al.
Published: (2025)
Enhancing Uncertainty Estimation and Interpretability via Bayesian Non-negative Decision Layer
by: Hu, Xinyue, et al.
Published: (2025)
by: Hu, Xinyue, et al.
Published: (2025)
Attention Schema-based Attention Control (ASAC): A Cognitive-Inspired Approach for Attention Management in Transformers
by: Saxena, Krati, et al.
Published: (2025)
by: Saxena, Krati, et al.
Published: (2025)
Procedural Fairness Through Decoupling Objectionable Data Generating Components
by: Tang, Zeyu, et al.
Published: (2023)
by: Tang, Zeyu, et al.
Published: (2023)
Golden Layers and Where to Find Them: Improved Knowledge Editing for Large Language Models Via Layer Gradient Analysis
by: Datta, Shrestha, et al.
Published: (2026)
by: Datta, Shrestha, et al.
Published: (2026)
QuantFPFlow: Quantum Amplitude Estimation for Fokker--Planck Policy Optimisation in Continuous Reinforcement Learning
by: Weinberg, Abraham Itzhak
Published: (2026)
by: Weinberg, Abraham Itzhak
Published: (2026)
Holder Policy Optimisation
by: Chen, Yuxiang, et al.
Published: (2026)
by: Chen, Yuxiang, et al.
Published: (2026)
Improving Data Efficiency for LLM Reinforcement Fine-tuning Through Difficulty-targeted Online Data Selection and Rollout Replay
by: Sun, Yifan, et al.
Published: (2025)
by: Sun, Yifan, et al.
Published: (2025)
AI-Assisted Generation of Difficult Math Questions
by: Shah, Vedant, et al.
Published: (2024)
by: Shah, Vedant, et al.
Published: (2024)
Deciphering Invariant Feature Decoupling in Source-free Time Series Forecasting with Proxy Denoising
by: Yan, Kangjia, et al.
Published: (2025)
by: Yan, Kangjia, et al.
Published: (2025)
Improving the Straight-Through Estimator with Zeroth-Order Information
by: Yang, Ningfeng, et al.
Published: (2025)
by: Yang, Ningfeng, et al.
Published: (2025)
Decoupled Prioritized Resampling for Offline RL
by: Yue, Yang, et al.
Published: (2023)
by: Yue, Yang, et al.
Published: (2023)
REPEAT: Improving Uncertainty Estimation in Representation Learning Explainability
by: Wickstrøm, Kristoffer K., et al.
Published: (2024)
by: Wickstrøm, Kristoffer K., et al.
Published: (2024)
Optimisation in Neurosymbolic Learning Systems
by: van Krieken, Emile
Published: (2024)
by: van Krieken, Emile
Published: (2024)
Optimisation Is Not What You Need
by: Ibias, Alfredo
Published: (2025)
by: Ibias, Alfredo
Published: (2025)
DRPO: Efficient Reasoning via Decoupled Reward Policy Optimization
by: Li, Gang, et al.
Published: (2025)
by: Li, Gang, et al.
Published: (2025)
Similar Items
-
Representation Collapsing Problems in Vector Quantization
by: Zhao, Wenhao, et al.
Published: (2024) -
Gaussian Mixture Vector Quantization with Aggregated Categorical Posterior
by: Yan, Mingyuan, et al.
Published: (2024) -
Early Quantization Shrinks Codebook: A Simple Fix for Diversity-Preserving Tokenization
by: Zhao, Wenhao, et al.
Published: (2026) -
VQSynery: Robust Drug Synergy Prediction With Vector Quantization Mechanism
by: Wu, Jiawei, et al.
Published: (2024) -
Explore-Execute Chain: Towards an Efficient Structured Reasoning Paradigm
by: Yang, Kaisen, et al.
Published: (2025)