Comparing Bottom-Up and Top-Down Steering Approaches on In-Context Learning Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Brumley, Madeline, Kwon, Joe, Krueger, David, Krasheninnikov, Dmitrii, Anwar, Usman |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Understanding (Un)Reliability of Steering Vectors in Language Models
von: Braun, Joschka, et al.
Veröffentlicht: (2025)
von: Braun, Joschka, et al.
Veröffentlicht: (2025)
Implicit meta-learning may lead language models to trust more reliable sources
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2023)
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2023)
Fresh in memory: Training-order recency is linearly encoded in language model activations
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2025)
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2025)
Stress-Testing Capability Elicitation With Password-Locked Models
von: Greenblatt, Ryan, et al.
Veröffentlicht: (2024)
von: Greenblatt, Ryan, et al.
Veröffentlicht: (2024)
Hierarchical Synthetic Tabular Data Generation: A Hybrid Top-Down and Bottom-Up Framework
von: Nie, Junfeng, et al.
Veröffentlicht: (2026)
von: Nie, Junfeng, et al.
Veröffentlicht: (2026)
Understanding In-Context Learning of Linear Models in Transformers Through an Adversarial Lens
von: Anwar, Usman, et al.
Veröffentlicht: (2024)
von: Anwar, Usman, et al.
Veröffentlicht: (2024)
Defining and Characterizing Reward Hacking
von: Skalse, Joar, et al.
Veröffentlicht: (2022)
von: Skalse, Joar, et al.
Veröffentlicht: (2022)
Reward Model Ensembles Help Mitigate Overoptimization
von: Coste, Thomas, et al.
Veröffentlicht: (2023)
von: Coste, Thomas, et al.
Veröffentlicht: (2023)
ReeFRAME: Reeb Graph based Trajectory Analysis Framework to Capture Top-Down and Bottom-Up Patterns of Life
von: Gudavalli, Chandrakanth, et al.
Veröffentlicht: (2024)
von: Gudavalli, Chandrakanth, et al.
Veröffentlicht: (2024)
Interpreting Emergent Planning in Model-Free Reinforcement Learning
von: Bush, Thomas, et al.
Veröffentlicht: (2025)
von: Bush, Thomas, et al.
Veröffentlicht: (2025)
Noisy Zero-Shot Coordination: Breaking The Common Knowledge Assumption In Zero-Shot Coordination Games
von: Anwar, Usman, et al.
Veröffentlicht: (2024)
von: Anwar, Usman, et al.
Veröffentlicht: (2024)
Detecting High-Stakes Interactions with Activation Probes
von: McKenzie, Alex, et al.
Veröffentlicht: (2025)
von: McKenzie, Alex, et al.
Veröffentlicht: (2025)
Learning to Forget using Hypernetworks
von: Rangel, Jose Miguel Lara, et al.
Veröffentlicht: (2024)
von: Rangel, Jose Miguel Lara, et al.
Veröffentlicht: (2024)
Eliciting Latent Knowledge from Quirky Language Models
von: Mallen, Alex, et al.
Veröffentlicht: (2023)
von: Mallen, Alex, et al.
Veröffentlicht: (2023)
Bottom-Up Synthesis of Knowledge-Grounded Task-Oriented Dialogues with Iteratively Self-Refined Prompts
von: Qian, Kun, et al.
Veröffentlicht: (2025)
von: Qian, Kun, et al.
Veröffentlicht: (2025)
Forward Learning with Top-Down Feedback: Empirical and Analytical Characterization
von: Srinivasan, Ravi, et al.
Veröffentlicht: (2023)
von: Srinivasan, Ravi, et al.
Veröffentlicht: (2023)
Chapter Comparing Top-Down And Bottom-Up Approaches. Maps Of Cultural Landscape Digitisation Processes
von: Vedoà, Marco
Veröffentlicht: (2022)
von: Vedoà, Marco
Veröffentlicht: (2022)
Distance-Aware Error for Spline Networks: A Bottom-Up Approach to Uncertainty
von: Ataei, Masoud, et al.
Veröffentlicht: (2025)
von: Ataei, Masoud, et al.
Veröffentlicht: (2025)
Mitigating Goal Misgeneralization via Minimax Regret
von: Sadek, Karim Abdel, et al.
Veröffentlicht: (2025)
von: Sadek, Karim Abdel, et al.
Veröffentlicht: (2025)
Attention Down-Sampling Transformer, Relative Ranking and Self-Consistency for Blind Image Quality Assessment
von: Alsaafin, Mohammed, et al.
Veröffentlicht: (2024)
von: Alsaafin, Mohammed, et al.
Veröffentlicht: (2024)
Can Mamba Learn How to Learn? A Comparative Study on In-Context Learning Tasks
von: Park, Jongho, et al.
Veröffentlicht: (2024)
von: Park, Jongho, et al.
Veröffentlicht: (2024)
Top-Down Bayesian Posterior Sampling for Sum-Product Networks
von: Yokoi, Soma, et al.
Veröffentlicht: (2024)
von: Yokoi, Soma, et al.
Veröffentlicht: (2024)
BUILD with Precision: Bottom-Up Inference of Linear DAGs
von: Ajorlou, Hamed, et al.
Veröffentlicht: (2025)
von: Ajorlou, Hamed, et al.
Veröffentlicht: (2025)
Productively Deploying Emerging Models on Emerging Platforms: A Top-Down Approach for Testing and Debugging
von: Feng, Siyuan, et al.
Veröffentlicht: (2024)
von: Feng, Siyuan, et al.
Veröffentlicht: (2024)
Open-world Instance Segmentation: Top-down Learning with Bottom-up Supervision
von: Kalluri, Tarun, et al.
Veröffentlicht: (2023)
von: Kalluri, Tarun, et al.
Veröffentlicht: (2023)
Transformers as Implicit State Estimators: In-Context Learning in Dynamical Systems
von: Akram, Usman, et al.
Veröffentlicht: (2024)
von: Akram, Usman, et al.
Veröffentlicht: (2024)
Contextual Feedback Loops: Amplifying Deep Reasoning with Iterative Top-Down Feedback
von: Fein-Ashley, Jacob, et al.
Veröffentlicht: (2024)
von: Fein-Ashley, Jacob, et al.
Veröffentlicht: (2024)
Representation Engineering: A Top-Down Approach to AI Transparency
von: Zou, Andy, et al.
Veröffentlicht: (2023)
von: Zou, Andy, et al.
Veröffentlicht: (2023)
Taking the Next Step: Bottom‐Up and Top‐Down Approaches for Mental Health Equity
von: Alexandria G. Bauer, et al.
Veröffentlicht: (2026)
von: Alexandria G. Bauer, et al.
Veröffentlicht: (2026)
IM-Context: In-Context Learning for Imbalanced Regression Tasks
von: Nejjar, Ismail, et al.
Veröffentlicht: (2024)
von: Nejjar, Ismail, et al.
Veröffentlicht: (2024)
Context-Scaling versus Task-Scaling in In-Context Learning
von: Abedsoltan, Amirhesam, et al.
Veröffentlicht: (2024)
von: Abedsoltan, Amirhesam, et al.
Veröffentlicht: (2024)
The Impact of Off-Policy Training Data on Probe Generalisation
von: Kirch, Nathalie, et al.
Veröffentlicht: (2025)
von: Kirch, Nathalie, et al.
Veröffentlicht: (2025)
Eye Gaze-Informed and Context-Aware Pedestrian Trajectory Prediction in Shared Spaces with Automated Shuttles: A Virtual Reality Study
von: Li, Danya, et al.
Veröffentlicht: (2026)
von: Li, Danya, et al.
Veröffentlicht: (2026)
Think Outside the Policy: In-Context Steered Policy Optimization
von: Huang, Hsiu-Yuan, et al.
Veröffentlicht: (2025)
von: Huang, Hsiu-Yuan, et al.
Veröffentlicht: (2025)
Position: AI Scaling: From Up to Down and Out
von: Wang, Yunke, et al.
Veröffentlicht: (2025)
von: Wang, Yunke, et al.
Veröffentlicht: (2025)
Split-on-Share: Mixture of Sparse Experts for Task-Agnostic Continual Learning
von: Siddika, Fatema, et al.
Veröffentlicht: (2026)
von: Siddika, Fatema, et al.
Veröffentlicht: (2026)
COLD-Steer: Steering Large Language Models via In-Context One-step Learning Dynamics
von: Sharma, Kartik, et al.
Veröffentlicht: (2026)
von: Sharma, Kartik, et al.
Veröffentlicht: (2026)
Enhancing Time-Series Anomaly Detection by Integrating Spectral-Residual Bottom-Up Attention with Reservoir Computing
von: Nihei, Hayato, et al.
Veröffentlicht: (2025)
von: Nihei, Hayato, et al.
Veröffentlicht: (2025)
Beyond Top-Down and Bottom-Up- Cognition as Context-Dependent Emergence in a Complex Adaptive Brain
von: Topcular, Baris, et al.
Veröffentlicht: (2025)
von: Topcular, Baris, et al.
Veröffentlicht: (2025)
Learning to Order: Task Sequencing as In-Context Optimization
von: Kobiolka, Jan, et al.
Veröffentlicht: (2026)
von: Kobiolka, Jan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Understanding (Un)Reliability of Steering Vectors in Language Models
von: Braun, Joschka, et al.
Veröffentlicht: (2025) -
Implicit meta-learning may lead language models to trust more reliable sources
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2023) -
Fresh in memory: Training-order recency is linearly encoded in language model activations
von: Krasheninnikov, Dmitrii, et al.
Veröffentlicht: (2025) -
Stress-Testing Capability Elicitation With Password-Locked Models
von: Greenblatt, Ryan, et al.
Veröffentlicht: (2024) -
Hierarchical Synthetic Tabular Data Generation: A Hybrid Top-Down and Bottom-Up Framework
von: Nie, Junfeng, et al.
Veröffentlicht: (2026)