Learning Interpretable Models Using Uncertainty Oracles
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ghose, Abhishek, Ravindran, Balaraman |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2019
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generalized Adaptive Transfer Network: Enhancing Transfer Learning in Reinforcement Learning Across Domains
von: Verma, Abhishek, et al.
Veröffentlicht: (2025)
von: Verma, Abhishek, et al.
Veröffentlicht: (2025)
Adaptive Action Duration with Contextual Bandits for Deep Reinforcement Learning in Dynamic Environments
von: Verma, Abhishek, et al.
Veröffentlicht: (2025)
von: Verma, Abhishek, et al.
Veröffentlicht: (2025)
Data Selection: A General Principle for Building Small Interpretable Models
von: Ghose, Abhishek
Veröffentlicht: (2022)
von: Ghose, Abhishek
Veröffentlicht: (2022)
Unifying Model-Free Efficiency and Model-Based Representations via Latent Dynamics
von: Acharjee, Jashaswimalya, et al.
Veröffentlicht: (2026)
von: Acharjee, Jashaswimalya, et al.
Veröffentlicht: (2026)
QuAKE: Speeding up Model Inference Using Quick and Approximate Kernels for Exponential Non-Linearities
von: Narayanaswami, Sai Kiran, et al.
Veröffentlicht: (2024)
von: Narayanaswami, Sai Kiran, et al.
Veröffentlicht: (2024)
MABL: Bi-Level Latent-Variable World Model for Sample-Efficient Multi-Agent Reinforcement Learning
von: Venugopal, Aravind, et al.
Veröffentlicht: (2023)
von: Venugopal, Aravind, et al.
Veröffentlicht: (2023)
SafeMIL: Learning Offline Safe Imitation Policy from Non-Preferred Trajectories
von: Burnwal, Returaj, et al.
Veröffentlicht: (2025)
von: Burnwal, Returaj, et al.
Veröffentlicht: (2025)
OSIL: Learning Offline Safe Imitation Policies with Safety Inferred from Non-preferred Trajectories
von: Burnwal, Returaj, et al.
Veröffentlicht: (2026)
von: Burnwal, Returaj, et al.
Veröffentlicht: (2026)
Pack and Detect: Fast Object Detection in Videos Using Region-of-Interest Packing
von: Kumar, Athindran Ramesh, et al.
Veröffentlicht: (2018)
von: Kumar, Athindran Ramesh, et al.
Veröffentlicht: (2018)
SWAN: Sparse Winnowed Attention for Reduced Inference Memory via Decompression-Free KV-Cache Compression
von: S, Santhosh G, et al.
Veröffentlicht: (2025)
von: S, Santhosh G, et al.
Veröffentlicht: (2025)
AQUA: Attention via QUery mAgnitudes for Memory and Compute Efficient Inference in LLMs
von: S, Santhosh G, et al.
Veröffentlicht: (2025)
von: S, Santhosh G, et al.
Veröffentlicht: (2025)
Know your Trajectory -- Trustworthy Reinforcement Learning deployment through Importance-Based Trajectory Analysis
von: F, Clifford, et al.
Veröffentlicht: (2025)
von: F, Clifford, et al.
Veröffentlicht: (2025)
Learning from Observation: A Survey of Recent Advances
von: Burnwal, Returaj, et al.
Veröffentlicht: (2025)
von: Burnwal, Returaj, et al.
Veröffentlicht: (2025)
PREFINE: Preference-Based Implicit Reward and Cost Fine-Tuning for Safety Alignment
von: Verma, Richa, et al.
Veröffentlicht: (2026)
von: Verma, Richa, et al.
Veröffentlicht: (2026)
On the Fragility of Active Learners for Text Classification
von: Ghose, Abhishek, et al.
Veröffentlicht: (2024)
von: Ghose, Abhishek, et al.
Veröffentlicht: (2024)
Oracle-Efficient Differentially Private Learning with Public Data
von: Block, Adam, et al.
Veröffentlicht: (2024)
von: Block, Adam, et al.
Veröffentlicht: (2024)
Inherently Interpretable and Uncertainty-Aware Models for Online Learning in Cyber-Security Problems
von: Kolicic, Benjamin, et al.
Veröffentlicht: (2024)
von: Kolicic, Benjamin, et al.
Veröffentlicht: (2024)
Imputation Uncertainty in Interpretable Machine Learning Methods
von: Golchian, Pegah, et al.
Veröffentlicht: (2025)
von: Golchian, Pegah, et al.
Veröffentlicht: (2025)
Learning Bayesian and Markov Networks with an Unreliable Oracle
von: Harviainen, Juha, et al.
Veröffentlicht: (2026)
von: Harviainen, Juha, et al.
Veröffentlicht: (2026)
OFAL: An Oracle-Free Active Learning Framework
von: Khorsand, Hadi, et al.
Veröffentlicht: (2025)
von: Khorsand, Hadi, et al.
Veröffentlicht: (2025)
Oracle-efficient Hybrid Learning with Constrained Adversaries
von: Okoroafor, Princewill, et al.
Veröffentlicht: (2026)
von: Okoroafor, Princewill, et al.
Veröffentlicht: (2026)
Zero-shot Active Learning Using Self Supervised Learning
von: Sinha, Abhishek, et al.
Veröffentlicht: (2024)
von: Sinha, Abhishek, et al.
Veröffentlicht: (2024)
Multilinguality in LLM-Designed Reward Functions for Restless Bandits: Effects on Task Performance and Fairness
von: Parthasarathy, Ambreesh, et al.
Veröffentlicht: (2025)
von: Parthasarathy, Ambreesh, et al.
Veröffentlicht: (2025)
Regret-Oracle Complexity Tradeoffs in Agnostic Online Learning
von: Attias, Idan, et al.
Veröffentlicht: (2026)
von: Attias, Idan, et al.
Veröffentlicht: (2026)
Data-dependent and Oracle Bounds on Forgetting in Continual Learning
von: Friedman, Lior, et al.
Veröffentlicht: (2024)
von: Friedman, Lior, et al.
Veröffentlicht: (2024)
Oracle-Efficient Hybrid Online Learning with Unknown Distribution
von: Wu, Changlong, et al.
Veröffentlicht: (2024)
von: Wu, Changlong, et al.
Veröffentlicht: (2024)
Model-Based Reinforcement Learning with Double Oracle Efficiency in Policy Optimization and Offline Estimation
von: Hu, Haichen, et al.
Veröffentlicht: (2026)
von: Hu, Haichen, et al.
Veröffentlicht: (2026)
Is Oracle Pruning the True Oracle?
von: Feng, Sicheng, et al.
Veröffentlicht: (2024)
von: Feng, Sicheng, et al.
Veröffentlicht: (2024)
Oracle-Robust Online Alignment for Large Language Models
von: Li, Zimeng, et al.
Veröffentlicht: (2026)
von: Li, Zimeng, et al.
Veröffentlicht: (2026)
Active Learning of Discrete-Time Dynamics for Uncertainty-Aware Model Predictive Control
von: Saviolo, Alessandro, et al.
Veröffentlicht: (2022)
von: Saviolo, Alessandro, et al.
Veröffentlicht: (2022)
Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models
von: Shoaeinaeini, Maryam, et al.
Veröffentlicht: (2024)
von: Shoaeinaeini, Maryam, et al.
Veröffentlicht: (2024)
SmartOracle -- An Agentic Approach to Mitigate Noise in Differential Oracles
von: Srinivasan, Srinath, et al.
Veröffentlicht: (2026)
von: Srinivasan, Srinath, et al.
Veröffentlicht: (2026)
Oracle-Efficient Smoothed Online Learning for Piecewise Continuous Decision Making
von: Block, Adam, et al.
Veröffentlicht: (2023)
von: Block, Adam, et al.
Veröffentlicht: (2023)
Autoregressive Learning in Joint KL: Sharp Oracle Bounds and Lower Bounds
von: Xu, Yunbei, et al.
Veröffentlicht: (2026)
von: Xu, Yunbei, et al.
Veröffentlicht: (2026)
Guided Uncertainty Learning Using a Post-Hoc Evidential Meta-Model
von: Barker, Charmaine, et al.
Veröffentlicht: (2025)
von: Barker, Charmaine, et al.
Veröffentlicht: (2025)
Global Climate Model Bias Correction Using Deep Learning
von: Pasula, Abhishek, et al.
Veröffentlicht: (2025)
von: Pasula, Abhishek, et al.
Veröffentlicht: (2025)
Code-Space Response Oracles: Generating Interpretable Multi-Agent Policies with Large Language Models
von: Hennes, Daniel, et al.
Veröffentlicht: (2026)
von: Hennes, Daniel, et al.
Veröffentlicht: (2026)
Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers
von: Ravindran, Santhosh Kumar
Veröffentlicht: (2025)
von: Ravindran, Santhosh Kumar
Veröffentlicht: (2025)
Oracle-Guided Masked Contrastive Reinforcement Learning for Visuomotor Policies
von: Zhang, Yuhang, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhang, et al.
Veröffentlicht: (2025)
Is Efficient PAC Learning Possible with an Oracle That Responds 'Yes' or 'No'?
von: Daskalakis, Constantinos, et al.
Veröffentlicht: (2024)
von: Daskalakis, Constantinos, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Generalized Adaptive Transfer Network: Enhancing Transfer Learning in Reinforcement Learning Across Domains
von: Verma, Abhishek, et al.
Veröffentlicht: (2025) -
Adaptive Action Duration with Contextual Bandits for Deep Reinforcement Learning in Dynamic Environments
von: Verma, Abhishek, et al.
Veröffentlicht: (2025) -
Data Selection: A General Principle for Building Small Interpretable Models
von: Ghose, Abhishek
Veröffentlicht: (2022) -
Unifying Model-Free Efficiency and Model-Based Representations via Latent Dynamics
von: Acharjee, Jashaswimalya, et al.
Veröffentlicht: (2026) -
QuAKE: Speeding up Model Inference Using Quick and Approximate Kernels for Exponential Non-Linearities
von: Narayanaswami, Sai Kiran, et al.
Veröffentlicht: (2024)