Active Few-Shot Fine-Tuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hübotter, Jonas, Sukhija, Bhavya, Treven, Lenart, As, Yarden, Krause, Andreas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Transductive Active Learning: Theory and Applications
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024)
Sample-efficient and Scalable Exploration in Continuous-Time RL
von: Iten, Klemens, et al.
Veröffentlicht: (2025)
von: Iten, Klemens, et al.
Veröffentlicht: (2025)
Simulation Priors for Data-Efficient Deep Learning
von: Treven, Lenart, et al.
Veröffentlicht: (2025)
von: Treven, Lenart, et al.
Veröffentlicht: (2025)
Safe Exploration Using Bayesian World Models and Log-Barrier Optimization
von: As, Yarden, et al.
Veröffentlicht: (2024)
von: As, Yarden, et al.
Veröffentlicht: (2024)
When to Sense and Control? A Time-adaptive Approach for Continuous-Time RL
von: Treven, Lenart, et al.
Veröffentlicht: (2024)
von: Treven, Lenart, et al.
Veröffentlicht: (2024)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
von: As, Yarden, et al.
Veröffentlicht: (2024)
von: As, Yarden, et al.
Veröffentlicht: (2024)
Efficiently Learning at Test-Time: Active Fine-Tuning of LLMs
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024)
NeoRL: Efficient Exploration for Nonepisodic RL
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024)
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024)
Bridging the Sim-to-Real Gap with Bayesian Inference
von: Rothfuss, Jonas, et al.
Veröffentlicht: (2024)
von: Rothfuss, Jonas, et al.
Veröffentlicht: (2024)
Probabilistic Artificial Intelligence
von: Krause, Andreas, et al.
Veröffentlicht: (2025)
von: Krause, Andreas, et al.
Veröffentlicht: (2025)
SOMBRL: Scalable and Optimistic Model-Based RL
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2025)
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2025)
Model-Based Reinforcement Learning for Control under Time-Varying Dynamics
von: Iten, Klemens, et al.
Veröffentlicht: (2026)
von: Iten, Klemens, et al.
Veröffentlicht: (2026)
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024)
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024)
Active Fine-Tuning of Multi-Task Policies
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
Local Mixtures of Experts: Essentially Free Test-Time Training via Model Merging
von: Bertolissi, Ryo, et al.
Veröffentlicht: (2025)
von: Bertolissi, Ryo, et al.
Veröffentlicht: (2025)
TARC: Time-Adaptive Robotic Control
von: Sukhija, Arnav, et al.
Veröffentlicht: (2025)
von: Sukhija, Arnav, et al.
Veröffentlicht: (2025)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
Learning Safety Constraints for Large Language Models
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
Learning on the Job: Test-Time Curricula for Targeted Reinforcement Learning
von: Hübotter, Jonas, et al.
Veröffentlicht: (2025)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2025)
Specialization after Generalization: Towards Understanding Test-Time Training in Foundation Models
von: Hübotter, Jonas, et al.
Veröffentlicht: (2025)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2025)
Data-Efficient Task Generalization via Probabilistic Model-based Meta Reinforcement Learning
von: Bhardwaj, Arjun, et al.
Veröffentlicht: (2023)
von: Bhardwaj, Arjun, et al.
Veröffentlicht: (2023)
Aligning Language Models from User Interactions
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2026)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2026)
Thompson Sampling via Fine-Tuning of LLMs
von: Menet, Nicolas, et al.
Veröffentlicht: (2025)
von: Menet, Nicolas, et al.
Veröffentlicht: (2025)
Symmetry-Guided Memory Augmentation for Efficient Locomotion Learning
von: Bao, Kaixi, et al.
Veröffentlicht: (2025)
von: Bao, Kaixi, et al.
Veröffentlicht: (2025)
SGPT: Few-Shot Prompt Tuning for Signed Graphs
von: Zhai, Zian, et al.
Veröffentlicht: (2024)
von: Zhai, Zian, et al.
Veröffentlicht: (2024)
CTIGuardian: A Few-Shot Framework for Mitigating Privacy Leakage in Fine-Tuned LLMs
von: Arachchige, Shashie Dilhara Batan, et al.
Veröffentlicht: (2025)
von: Arachchige, Shashie Dilhara Batan, et al.
Veröffentlicht: (2025)
Learning Soft Robotic Dynamics with Active Exploration
von: Zheng, Hehui, et al.
Veröffentlicht: (2025)
von: Zheng, Hehui, et al.
Veröffentlicht: (2025)
HyperFlow: Gradient-Free Emulation of Few-Shot Fine-Tuning
von: Kim, Donggyun, et al.
Veröffentlicht: (2025)
von: Kim, Donggyun, et al.
Veröffentlicht: (2025)
Safe Exploration via Policy Priors
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
Quantum-Inspired Fine-Tuning for Few-Shot AIGC Detection via Phase-Structured Reparameterization
von: Xing, Kaiyang, et al.
Veröffentlicht: (2026)
von: Xing, Kaiyang, et al.
Veröffentlicht: (2026)
Silent Sabotage During Fine-Tuning: Few-Shot Rationale Poisoning of Compact Medical LLMs
von: Xie, Jingyuan, et al.
Veröffentlicht: (2026)
von: Xie, Jingyuan, et al.
Veröffentlicht: (2026)
Prompt Tuning with Diffusion for Few-Shot Pre-trained Policy Generalization
von: Hu, Shengchao, et al.
Veröffentlicht: (2024)
von: Hu, Shengchao, et al.
Veröffentlicht: (2024)
Context-Aware Adapter Tuning for Few-Shot Relation Learning in Knowledge Graphs
von: Liu, Ran, et al.
Veröffentlicht: (2024)
von: Liu, Ran, et al.
Veröffentlicht: (2024)
Reinforcement Learning via Self-Distillation
von: Hübotter, Jonas, et al.
Veröffentlicht: (2026)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2026)
Adaptive Prompt Tuning: Vision Guided Prompt Tuning with Cross-Attention for Fine-Grained Few-Shot Learning
von: Brouwer, Eric, et al.
Veröffentlicht: (2024)
von: Brouwer, Eric, et al.
Veröffentlicht: (2024)
Sampling-Based Safe Reinforcement Learning
von: Vignola, Luca, et al.
Veröffentlicht: (2026)
von: Vignola, Luca, et al.
Veröffentlicht: (2026)
Pin-Tuning: Parameter-Efficient In-Context Tuning for Few-Shot Molecular Property Prediction
von: Wang, Liang, et al.
Veröffentlicht: (2024)
von: Wang, Liang, et al.
Veröffentlicht: (2024)
Safety at One Shot: Patching Fine-Tuned LLMs with A Single Instance
von: Zhang, Jiawen, et al.
Veröffentlicht: (2026)
von: Zhang, Jiawen, et al.
Veröffentlicht: (2026)
Predicting Intermittent Job Failure Categories for Diagnosis Using Few-Shot Fine-Tuned Language Models
von: Aïdasso, Henri, et al.
Veröffentlicht: (2026)
von: Aïdasso, Henri, et al.
Veröffentlicht: (2026)
Test-Time Tuned Language Models Enable End-to-end De Novo Molecular Structure Generation from MS/MS Spectra
von: Mismetti, Laura, et al.
Veröffentlicht: (2025)
von: Mismetti, Laura, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Transductive Active Learning: Theory and Applications
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024) -
Sample-efficient and Scalable Exploration in Continuous-Time RL
von: Iten, Klemens, et al.
Veröffentlicht: (2025) -
Simulation Priors for Data-Efficient Deep Learning
von: Treven, Lenart, et al.
Veröffentlicht: (2025) -
Safe Exploration Using Bayesian World Models and Log-Barrier Optimization
von: As, Yarden, et al.
Veröffentlicht: (2024) -
When to Sense and Control? A Time-adaptive Approach for Continuous-Time RL
von: Treven, Lenart, et al.
Veröffentlicht: (2024)