BayesQ: Uncertainty-Guided Bayesian Quantization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lamaakal, Ismail, Yahyati, Chaymae, Maleh, Yassine, Makkaoui, Khalid El, Ouahbi, Ibrahim |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
T3C: Test-Time Tensor Compression with Consistency Guarantees
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2026)
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2026)
When Bits Break Recourse: Counterfactual-Faithful Quantization
von: Yahyati, Chaymae, et al.
Veröffentlicht: (2026)
von: Yahyati, Chaymae, et al.
Veröffentlicht: (2026)
SNAP-UQ: Self-supervised Next-Activation Prediction for Single-Pass Uncertainty in TinyML
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2025)
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2025)
TCUQ: Single-Pass Uncertainty Quantification from Temporal Consistency with Streaming Conformal Calibration for TinyML
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2025)
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2025)
Drift-to-Action Controllers: Budgeted Interventions with Online Risk Certificates
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2026)
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2026)
Simplex-FEM Networks (SiFEN): Learning A Triangulated Function Approximator
von: Yahyati, Chaymae, et al.
Veröffentlicht: (2025)
von: Yahyati, Chaymae, et al.
Veröffentlicht: (2025)
Motion-Compensated Weight Compression
von: Lamaakal, Ismail
Veröffentlicht: (2026)
von: Lamaakal, Ismail
Veröffentlicht: (2026)
ParetoQ: Improving Scaling Laws in Extremely Low-bit LLM Quantization
von: Liu, Zechun, et al.
Veröffentlicht: (2025)
von: Liu, Zechun, et al.
Veröffentlicht: (2025)
Technical Report: Activation Residual Hessian Quantization (ARHQ) for Low-Bit LLM Quantization
von: Wang, YiFeng, et al.
Veröffentlicht: (2026)
von: Wang, YiFeng, et al.
Veröffentlicht: (2026)
Uncertainty-Guided Alignment for Unsupervised Domain Adaptation in Regression
von: Nejjar, Ismail, et al.
Veröffentlicht: (2024)
von: Nejjar, Ismail, et al.
Veröffentlicht: (2024)
Enhancing Post-Training Quantization via Future Activation Awareness
von: Lv, Zheqi, et al.
Veröffentlicht: (2026)
von: Lv, Zheqi, et al.
Veröffentlicht: (2026)
Fine-tuning Quantized Neural Networks with Zeroth-order Optimization
von: Shang, Sifeng, et al.
Veröffentlicht: (2025)
von: Shang, Sifeng, et al.
Veröffentlicht: (2025)
Transformer-VQ: Linear-Time Transformers via Vector Quantization
von: Lingle, Lucas D.
Veröffentlicht: (2023)
von: Lingle, Lucas D.
Veröffentlicht: (2023)
BayesDiff: Estimating Pixel-wise Uncertainty in Diffusion via Bayesian Inference
von: Kou, Siqi, et al.
Veröffentlicht: (2023)
von: Kou, Siqi, et al.
Veröffentlicht: (2023)
QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Dual Precision Quantization for Efficient and Accurate Deep Neural Networks Inference
von: Gafni, Tomer, et al.
Veröffentlicht: (2025)
von: Gafni, Tomer, et al.
Veröffentlicht: (2025)
TWEO: Transformers Without Extreme Outliers Enables FP8 Training And Quantization For Dummies
von: Liang, Guang, et al.
Veröffentlicht: (2025)
von: Liang, Guang, et al.
Veröffentlicht: (2025)
Q-GroundCAM: Quantifying Grounding in Vision Language Models via GradCAM
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Exploratory Direct Preference Optimization for Multimodal Large Language Models
von: Zhang, Huatian, et al.
Veröffentlicht: (2026)
von: Zhang, Huatian, et al.
Veröffentlicht: (2026)
Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models
von: Groot, Tobias, et al.
Veröffentlicht: (2024)
von: Groot, Tobias, et al.
Veröffentlicht: (2024)
Fine-Grained Alignment in Vision-and-Language Navigation through Bayesian Optimization
von: Song, Yuhang, et al.
Veröffentlicht: (2024)
von: Song, Yuhang, et al.
Veröffentlicht: (2024)
Patch-Prompt Aligned Bayesian Prompt Tuning for Vision-Language Models
von: Liu, Xinyang, et al.
Veröffentlicht: (2023)
von: Liu, Xinyang, et al.
Veröffentlicht: (2023)
VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation
von: Kumar, Divake, et al.
Veröffentlicht: (2026)
von: Kumar, Divake, et al.
Veröffentlicht: (2026)
A Practitioner's Guide to Continual Multimodal Pretraining
von: Roth, Karsten, et al.
Veröffentlicht: (2024)
von: Roth, Karsten, et al.
Veröffentlicht: (2024)
Q-SENN: Quantized Self-Explaining Neural Networks
von: Norrenbrock, Thomas, et al.
Veröffentlicht: (2023)
von: Norrenbrock, Thomas, et al.
Veröffentlicht: (2023)
A Language Anchor-Guided Method for Robust Noisy Domain Generalization
von: Dai, Zilin, et al.
Veröffentlicht: (2025)
von: Dai, Zilin, et al.
Veröffentlicht: (2025)
ChartGaze: Enhancing Chart Understanding in LVLMs with Eye-Tracking Guided Attention Refinement
von: Salamatian, Ali, et al.
Veröffentlicht: (2025)
von: Salamatian, Ali, et al.
Veröffentlicht: (2025)
Pico-Banana-400K: A Large-Scale Dataset for Text-Guided Image Editing
von: Qian, Yusu, et al.
Veröffentlicht: (2025)
von: Qian, Yusu, et al.
Veröffentlicht: (2025)
D$^{3}$ToM: Decider-Guided Dynamic Token Merging for Accelerating Diffusion MLLMs
von: Chang, Shuochen, et al.
Veröffentlicht: (2025)
von: Chang, Shuochen, et al.
Veröffentlicht: (2025)
Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
Q-Drift: Quantization-Aware Drift Correction for Diffusion Model Sampling
von: Ryu, Sooyoung, et al.
Veröffentlicht: (2026)
von: Ryu, Sooyoung, et al.
Veröffentlicht: (2026)
SPEGNet: Synergistic Perception-Guided Network for Camouflaged Object Detection
von: Jan, Baber, et al.
Veröffentlicht: (2025)
von: Jan, Baber, et al.
Veröffentlicht: (2025)
Prompt as Free Lunch: Enhancing Diversity in Source-Free Cross-domain Few-shot Learning through Semantic-Guided Prompting
von: Zhuo, Linhai, et al.
Veröffentlicht: (2024)
von: Zhuo, Linhai, et al.
Veröffentlicht: (2024)
On Structured State-Space Duality
von: Hu, Jerry Yao-Chieh, et al.
Veröffentlicht: (2025)
von: Hu, Jerry Yao-Chieh, et al.
Veröffentlicht: (2025)
DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers
von: Sharify, Sayeh, et al.
Veröffentlicht: (2026)
von: Sharify, Sayeh, et al.
Veröffentlicht: (2026)
Leveraging NTPs for Efficient Hallucination Detection in VLMs
von: Azachi, Ofir, et al.
Veröffentlicht: (2025)
von: Azachi, Ofir, et al.
Veröffentlicht: (2025)
First-Order Error Matters: Accurate Compensation for Quantized Large Language Models
von: Zheng, Xingyu, et al.
Veröffentlicht: (2025)
von: Zheng, Xingyu, et al.
Veröffentlicht: (2025)
Extracting Usable Predictions from Quantized Networks through Uncertainty Quantification for OOD Detection
von: Singhal, Rishi, et al.
Veröffentlicht: (2024)
von: Singhal, Rishi, et al.
Veröffentlicht: (2024)
Can Bayesian Neural Networks Explicitly Model Input Uncertainty?
von: Valdenegro-Toro, Matias, et al.
Veröffentlicht: (2025)
von: Valdenegro-Toro, Matias, et al.
Veröffentlicht: (2025)
RAVEN: Query-Guided Representation Alignment for Question Answering over Audio, Video, Embedded Sensors, and Natural Language
von: Biswas, Subrata, et al.
Veröffentlicht: (2025)
von: Biswas, Subrata, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
T3C: Test-Time Tensor Compression with Consistency Guarantees
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2026) -
When Bits Break Recourse: Counterfactual-Faithful Quantization
von: Yahyati, Chaymae, et al.
Veröffentlicht: (2026) -
SNAP-UQ: Self-supervised Next-Activation Prediction for Single-Pass Uncertainty in TinyML
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2025) -
TCUQ: Single-Pass Uncertainty Quantification from Temporal Consistency with Streaming Conformal Calibration for TinyML
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2025) -
Drift-to-Action Controllers: Budgeted Interventions with Online Risk Certificates
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2026)