When Do Early-Exit Networks Generalize? A PAC-Bayesian Theory of Adaptive Depth
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guo, Dongxin, Wu, Jikun, Yiu, Siu Ming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Geometric Metrics for MoE Specialization: From Fisher Information to Early Failure Detection
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
SigGate-GT: Taming Over-Smoothing in Graph Transformers via Sigmoid-Gated Attention
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Closing the Theory-Practice Gap in Spiking Transformers via Effective Dimension
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Generalization and Feature Attribution in Machine Learning Models for Crop Yield and Anomaly Prediction in Germany
von: Baatz, Roland
Veröffentlicht: (2025)
von: Baatz, Roland
Veröffentlicht: (2025)
When Can Human-AI Teams Outperform Individuals? Tight Bounds with Impossibility Guarantees
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Parameter-Efficient Neuroevolution for Diverse LLM Generation: Quality-Diversity Optimization via Prompt Embedding Evolution
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Attribution Projection Calculus: A Novel Framework for Causal Inference in Bayesian Networks
von: Amin, M Ruhul
Veröffentlicht: (2025)
von: Amin, M Ruhul
Veröffentlicht: (2025)
Emotion-Inspired Learning Signals (EILS): A Homeostatic Framework for Adaptive Autonomous Agents
von: Tiwari, Dhruv
Veröffentlicht: (2025)
von: Tiwari, Dhruv
Veröffentlicht: (2025)
Causal Direction from Convergence Time: Faster Training in the True Causal Direction
von: Tamim, Abdulrahman
Veröffentlicht: (2026)
von: Tamim, Abdulrahman
Veröffentlicht: (2026)
SCOPE: Selective Conformal Optimized Pairwise LLM Judging
von: Badshah, Sher, et al.
Veröffentlicht: (2026)
von: Badshah, Sher, et al.
Veröffentlicht: (2026)
EvoPref: Multi-Objective Evolutionary Optimization Discovers Diverse LLM Alignments Beyond Gradient Descent
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
From Tail Universality to Bernstein-von Mises: A Unified Statistical Theory of Semi-Implicit Variational Inference
von: Plummer, Sean
Veröffentlicht: (2025)
von: Plummer, Sean
Veröffentlicht: (2025)
Quantum computer formulation of the FKP-operator eigenvalue problem for probabilistic learning on manifolds
von: Soize, Christian, et al.
Veröffentlicht: (2025)
von: Soize, Christian, et al.
Veröffentlicht: (2025)
Understanding and Tackling Over-Dilution in Graph Neural Networks
von: Lee, Junhyun, et al.
Veröffentlicht: (2025)
von: Lee, Junhyun, et al.
Veröffentlicht: (2025)
The Deterministic Horizon: When Extended Reasoning Fails and Tool Delegation Becomes Necessary
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Probabilistic Digital Twins of Users: Latent Representation Learning with Statistically Validated Semantics
von: David, Daniel
Veröffentlicht: (2025)
von: David, Daniel
Veröffentlicht: (2025)
Bias by Necessity: Impossibility Theorems for Sequential Processing with Convergent AI and Human Validation
von: Wu, Jikun, et al.
Veröffentlicht: (2026)
von: Wu, Jikun, et al.
Veröffentlicht: (2026)
EARCP: Self-Regulating Coherence-Aware Ensemble Architecture for Sequential Decision Making -- Ensemble Auto-Regule par Coherence et Performance
von: Amega, Mike
Veröffentlicht: (2026)
von: Amega, Mike
Veröffentlicht: (2026)
Constant-Target Energy Matching: A Unified Framework for Continuous and Discrete Density Estimation
von: Zeng, Zhijun, et al.
Veröffentlicht: (2026)
von: Zeng, Zhijun, et al.
Veröffentlicht: (2026)
metabeta -- A fast neural model for Bayesian mixed-effects regression
von: Kipnis, Alex, et al.
Veröffentlicht: (2025)
von: Kipnis, Alex, et al.
Veröffentlicht: (2025)
CDFlow: Building Invertible Layers with Circulant and Diagonal Matrices
von: Feng, Xuchen, et al.
Veröffentlicht: (2025)
von: Feng, Xuchen, et al.
Veröffentlicht: (2025)
Ice Cream Doesn't Cause Drowning: Benchmarking LLMs Against Statistical Pitfalls in Causal Inference
von: Du, Jin, et al.
Veröffentlicht: (2025)
von: Du, Jin, et al.
Veröffentlicht: (2025)
Can Computational Reducibility Lead to Transferable Models for Graph Combinatorial Optimization?
von: Cantürk, Semih, et al.
Veröffentlicht: (2026)
von: Cantürk, Semih, et al.
Veröffentlicht: (2026)
BLISS: Bandit Layer Importance Sampling Strategy for Efficient Training of Graph Neural Networks
von: Alsaqa, Omar, et al.
Veröffentlicht: (2025)
von: Alsaqa, Omar, et al.
Veröffentlicht: (2025)
Closed-Form Beta Distribution Estimation from Sparse Statistics with Random Forest Implicit Regularization
von: Landers, Jonathan R.
Veröffentlicht: (2025)
von: Landers, Jonathan R.
Veröffentlicht: (2025)
Neuro-Symbolic Learning for Galois Groups: Unveiling Probabilistic Trends in Polynomials
von: Shaska, Elira, et al.
Veröffentlicht: (2025)
von: Shaska, Elira, et al.
Veröffentlicht: (2025)
ProactBench: Beyond What The User Asked For
von: Harfi, Sepehr, et al.
Veröffentlicht: (2026)
von: Harfi, Sepehr, et al.
Veröffentlicht: (2026)
VORT: Adaptive Power-Law Memory for NLP Transformers
von: Mlaiki, Nabil
Veröffentlicht: (2026)
von: Mlaiki, Nabil
Veröffentlicht: (2026)
RF-BayesPhysNet: A Bayesian rPPG Uncertainty Estimation Method for Complex Scenarios
von: Ma, Rufei, et al.
Veröffentlicht: (2025)
von: Ma, Rufei, et al.
Veröffentlicht: (2025)
DRO-InstructZero: Distributionally Robust Prompt Optimization for Large Language Models
von: Li, Yangyang
Veröffentlicht: (2025)
von: Li, Yangyang
Veröffentlicht: (2025)
Revisiting Unbiased Implicit Variational Inference
von: Pielok, Tobias, et al.
Veröffentlicht: (2025)
von: Pielok, Tobias, et al.
Veröffentlicht: (2025)
Urban Spatio-Temporal Foundation Models for Climate-Resilient Housing: Scaling Diffusion Transformers for Disaster Risk Prediction
von: Imanov, Olaf Yunus Laitinen, et al.
Veröffentlicht: (2026)
von: Imanov, Olaf Yunus Laitinen, et al.
Veröffentlicht: (2026)
Model Collapse as Cultural Evolution
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Quantum Circuits for Quantum Convolutions: A Quantum Convolutional Autoencoder
von: Orduz, Javier, et al.
Veröffentlicht: (2025)
von: Orduz, Javier, et al.
Veröffentlicht: (2025)
Evaluating the Quality of the Quantified Uncertainty for (Re)Calibration of Data-Driven Regression Models
von: Wibbeke, Jelke, et al.
Veröffentlicht: (2025)
von: Wibbeke, Jelke, et al.
Veröffentlicht: (2025)
Gaussian Ensemble Belief Propagation for Efficient Inference in High-Dimensional Systems
von: MacKinlay, Dan, et al.
Veröffentlicht: (2024)
von: MacKinlay, Dan, et al.
Veröffentlicht: (2024)
On the Accuracy of Newton Step and Influence Function Data Attributions
von: Rubinstein, Ittai, et al.
Veröffentlicht: (2025)
von: Rubinstein, Ittai, et al.
Veröffentlicht: (2025)
Robustness Auditing for Linear Regression: To Singularity and Beyond
von: Rubinstein, Ittai, et al.
Veröffentlicht: (2024)
von: Rubinstein, Ittai, et al.
Veröffentlicht: (2024)
ComplianceNLP: Knowledge-Graph-Augmented RAG for Multi-Framework Regulatory Gap Detection
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Geometric Metrics for MoE Specialization: From Fisher Information to Early Failure Detection
von: Guo, Dongxin, et al.
Veröffentlicht: (2026) -
SigGate-GT: Taming Over-Smoothing in Graph Transformers via Sigmoid-Gated Attention
von: Guo, Dongxin, et al.
Veröffentlicht: (2026) -
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
von: Guo, Dongxin, et al.
Veröffentlicht: (2026) -
Closing the Theory-Practice Gap in Spiking Transformers via Effective Dimension
von: Guo, Dongxin, et al.
Veröffentlicht: (2026) -
Generalization and Feature Attribution in Machine Learning Models for Crop Yield and Anomaly Prediction in Germany
von: Baatz, Roland
Veröffentlicht: (2025)