Conformal Sparsification for Bandwidth-Efficient Edge-Cloud Speculative Decoding
Fuente:
arXiv
Saved in:
| Main Authors: | Bhattacharjee, Payel, Tian, Fengwei, Zhong, Meiyu, Zhang, Guangyi, Simeone, Osvaldo, Tandon, Ravi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MARS: Margin and Semantic-Aware Data Augmentation for Reward Modeling
by: Bhattacharjee, Payel, et al.
Published: (2026)
by: Bhattacharjee, Payel, et al.
Published: (2026)
Intrinsic Fairness-Accuracy Tradeoffs under Equalized Odds
by: Zhong, Meiyu, et al.
Published: (2024)
by: Zhong, Meiyu, et al.
Published: (2024)
Speeding up Speculative Decoding via Sequential Approximate Verification
by: Zhong, Meiyu, et al.
Published: (2025)
by: Zhong, Meiyu, et al.
Published: (2025)
PROPS: Progressively Private Self-alignment of Large Language Models
by: Teku, Noel, et al.
Published: (2025)
by: Teku, Noel, et al.
Published: (2025)
CURATE: Scaling-up Differentially Private Causal Graph Discovery
by: Bhattacharjee, Payel, et al.
Published: (2024)
by: Bhattacharjee, Payel, et al.
Published: (2024)
SPLITZ: Certifiable Robustness via Split Lipschitz Randomized Smoothing
by: Zhong, Meiyu, et al.
Published: (2024)
by: Zhong, Meiyu, et al.
Published: (2024)
Generalization and Informativeness of Conformal Prediction
by: Zecchin, Matteo, et al.
Published: (2024)
by: Zecchin, Matteo, et al.
Published: (2024)
Inference Privacy: Properties and Mechanisms
by: Tian, Fengwei, et al.
Published: (2024)
by: Tian, Fengwei, et al.
Published: (2024)
Generalization and Informativeness of Weighted Conformal Risk Control Under Covariate Shift
by: Zecchin, Matteo, et al.
Published: (2025)
by: Zecchin, Matteo, et al.
Published: (2025)
Localized Adaptive Risk Control
by: Zecchin, Matteo, et al.
Published: (2024)
by: Zecchin, Matteo, et al.
Published: (2024)
Uncertainty Quantification and Data Efficiency in AI: An Information-Theoretic Perspective
by: Simeone, Osvaldo, et al.
Published: (2025)
by: Simeone, Osvaldo, et al.
Published: (2025)
Prompt Fairness: Sub-group Disparities in LLMs
by: Zhong, Meiyu, et al.
Published: (2025)
by: Zhong, Meiyu, et al.
Published: (2025)
STAMP: Selective Task-Aware Mechanism for Text Privacy
by: Tian, Fengwei, et al.
Published: (2026)
by: Tian, Fengwei, et al.
Published: (2026)
Adaptive Learn-then-Test: Statistically Valid and Efficient Hyperparameter Selection
by: Zecchin, Matteo, et al.
Published: (2024)
by: Zecchin, Matteo, et al.
Published: (2024)
In-Context Learning for MIMO Equalization Using Transformer-Based Sequence Models
by: Zecchin, Matteo, et al.
Published: (2023)
by: Zecchin, Matteo, et al.
Published: (2023)
Trustworthy Actionable Perturbations
by: Friedbaum, Jesse, et al.
Published: (2024)
by: Friedbaum, Jesse, et al.
Published: (2024)
Synthetic Counterfactual Labels for Efficient Conformal Counterfactual Inference
by: Farzaneh, Amirmohammad, et al.
Published: (2025)
by: Farzaneh, Amirmohammad, et al.
Published: (2025)
WISV: Wireless-Informed Semantic Verification for Distributed Speculative Decoding in Device-Edge LLM Inference
by: Liu, Zixuan, et al.
Published: (2026)
by: Liu, Zixuan, et al.
Published: (2026)
Improving Epidemic Analyses with Privacy-Preserving Integration of Sensitive Data
by: Guan, Zihan, et al.
Published: (2025)
by: Guan, Zihan, et al.
Published: (2025)
Context-Aware Online Conformal Anomaly Detection with Prediction-Powered Data Acquisition
by: Farzaneh, Amirmohammad, et al.
Published: (2025)
by: Farzaneh, Amirmohammad, et al.
Published: (2025)
Beyond Fixed False Discovery Rates: Post-Hoc Conformal Selection with E-Variables
by: Zhu, Meiyi, et al.
Published: (2026)
by: Zhu, Meiyi, et al.
Published: (2026)
Towards Efficient and Reliable AI Through Neuromorphic Principles
by: Rajendran, Bipin, et al.
Published: (2023)
by: Rajendran, Bipin, et al.
Published: (2023)
Filtered Randomized Smoothing: A New Defense for Robust Modulation Classification
by: Zhang, Wenhan, et al.
Published: (2024)
by: Zhang, Wenhan, et al.
Published: (2024)
Efficient Federated Conformal Prediction with Group-Conditional Guarantees
by: Wen, Haifeng, et al.
Published: (2026)
by: Wen, Haifeng, et al.
Published: (2026)
Collaborative Edge AI Inference over Cloud-RAN
by: Zhang, Pengfei, et al.
Published: (2024)
by: Zhang, Pengfei, et al.
Published: (2024)
Bayesian Optimization with Formal Safety Guarantees via Online Conformal Prediction
by: Zhang, Yunchuan, et al.
Published: (2023)
by: Zhang, Yunchuan, et al.
Published: (2023)
Learning to Diagnose Privately: DP-Powered LLMs for Radiology Report Classification
by: Bhattacharjee, Payel, et al.
Published: (2025)
by: Bhattacharjee, Payel, et al.
Published: (2025)
Confidence-Based Decoding is Provably Efficient for Diffusion Language Models
by: Cai, Changxiao, et al.
Published: (2026)
by: Cai, Changxiao, et al.
Published: (2026)
REST: Retrieval-Based Speculative Decoding
by: He, Zhenyu, et al.
Published: (2023)
by: He, Zhenyu, et al.
Published: (2023)
Energy-Efficient Edge Learning via Joint Data Deepening-and-Prefetching
by: Kook, Sujin, et al.
Published: (2024)
by: Kook, Sujin, et al.
Published: (2024)
Conformal Calibration: Ensuring the Reliability of Black-Box AI in Wireless Systems
by: Simeone, Osvaldo, et al.
Published: (2025)
by: Simeone, Osvaldo, et al.
Published: (2025)
Distributed Conformal Prediction via Message Passing
by: Wen, Haifeng, et al.
Published: (2025)
by: Wen, Haifeng, et al.
Published: (2025)
Generalization Bounds for Neural Belief Propagation Decoders
by: Adiga, Sudarshan, et al.
Published: (2023)
by: Adiga, Sudarshan, et al.
Published: (2023)
Modern Neuromorphic AI: From Intra-Token to Inter-Token Processing
by: Simeone, Osvaldo
Published: (2026)
by: Simeone, Osvaldo
Published: (2026)
Neural Polar Decoders for Deletion Channels
by: Aharoni, Ziv, et al.
Published: (2025)
by: Aharoni, Ziv, et al.
Published: (2025)
Attention-Based Feature Online Conformal Prediction for Time Series
by: Zhu, Meiyi, et al.
Published: (2025)
by: Zhu, Meiyi, et al.
Published: (2025)
Score Based Error Correcting Code Decoder
by: Helvits, Alon, et al.
Published: (2026)
by: Helvits, Alon, et al.
Published: (2026)
Post-Selection Distributional Model Evaluation
by: Farzaneh, Amirmohammad, et al.
Published: (2026)
by: Farzaneh, Amirmohammad, et al.
Published: (2026)
Multi-Objective Hyperparameter Selection via Hypothesis Testing on Reliability Graphs
by: Farzaneh, Amirmohammad, et al.
Published: (2025)
by: Farzaneh, Amirmohammad, et al.
Published: (2025)
Statistically Valid Information Bottleneck via Multiple Hypothesis Testing
by: Farzaneh, Amirmohammad, et al.
Published: (2024)
by: Farzaneh, Amirmohammad, et al.
Published: (2024)
Similar Items
-
MARS: Margin and Semantic-Aware Data Augmentation for Reward Modeling
by: Bhattacharjee, Payel, et al.
Published: (2026) -
Intrinsic Fairness-Accuracy Tradeoffs under Equalized Odds
by: Zhong, Meiyu, et al.
Published: (2024) -
Speeding up Speculative Decoding via Sequential Approximate Verification
by: Zhong, Meiyu, et al.
Published: (2025) -
PROPS: Progressively Private Self-alignment of Large Language Models
by: Teku, Noel, et al.
Published: (2025) -
CURATE: Scaling-up Differentially Private Causal Graph Discovery
by: Bhattacharjee, Payel, et al.
Published: (2024)