On the Relationship Between the Choice of Representation and In-Context Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Marinescu, Ioana, Cho, Kyunghyun, Oermann, Eric Karl |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Clinically Grounded Agent-based Report Evaluation: An Interpretable Metric for Radiology Report Generation
por: Dua, Radhika, et al.
Publicado: (2025)
por: Dua, Radhika, et al.
Publicado: (2025)
Hyperparameter Loss Surfaces Are Simple Near their Optima
por: Lourie, Nicholas, et al.
Publicado: (2025)
por: Lourie, Nicholas, et al.
Publicado: (2025)
Show Your Work with Confidence: Confidence Bands for Tuning Curves
por: Lourie, Nicholas, et al.
Publicado: (2023)
por: Lourie, Nicholas, et al.
Publicado: (2023)
Scaling Laws Are Unreliable for Downstream Tasks: A Reality Check
por: Lourie, Nicholas, et al.
Publicado: (2025)
por: Lourie, Nicholas, et al.
Publicado: (2025)
Efficient semantic uncertainty quantification in language models via diversity-steered sampling
por: Park, Ji Won, et al.
Publicado: (2025)
por: Park, Ji Won, et al.
Publicado: (2025)
Generalization in Healthcare AI: Evaluation of a Clinical Large Language Model
por: Rahman, Salman, et al.
Publicado: (2024)
por: Rahman, Salman, et al.
Publicado: (2024)
Jointly Modeling Inter- & Intra-Modality Dependencies for Multi-modal Learning
por: Madaan, Divyam, et al.
Publicado: (2024)
por: Madaan, Divyam, et al.
Publicado: (2024)
Temporal Generalization: A Reality Check
por: Madaan, Divyam, et al.
Publicado: (2025)
por: Madaan, Divyam, et al.
Publicado: (2025)
Characterizing the Predictive Impact of Modalities with Supervised Latent-Variable Modeling
por: Madaan, Divyam, et al.
Publicado: (2026)
por: Madaan, Divyam, et al.
Publicado: (2026)
Paradox of De-identification: A Critique of HIPAA Safe Harbour in the Age of LLMs
por: Jiang, Lavender Y., et al.
Publicado: (2026)
por: Jiang, Lavender Y., et al.
Publicado: (2026)
Learning Task Representations from In-Context Learning
por: Saglam, Baturay, et al.
Publicado: (2025)
por: Saglam, Baturay, et al.
Publicado: (2025)
Always Learning, Always Mixing: Efficient and Simple Data Mixing All The Time
por: Hu, Michael Y., et al.
Publicado: (2026)
por: Hu, Michael Y., et al.
Publicado: (2026)
Language Models as Causal Effect Generators
por: Bynum, Lucius E. J., et al.
Publicado: (2024)
por: Bynum, Lucius E. J., et al.
Publicado: (2024)
Neural Neural Scaling Laws
por: Hu, Michael Y., et al.
Publicado: (2026)
por: Hu, Michael Y., et al.
Publicado: (2026)
Multi-modal Data Spectrum: Multi-modal Datasets are Multi-dimensional
por: Madaan, Divyam, et al.
Publicado: (2025)
por: Madaan, Divyam, et al.
Publicado: (2025)
In-Context Learning Learns Label Relationships but Is Not Conventional Learning
por: Kossen, Jannik, et al.
Publicado: (2023)
por: Kossen, Jannik, et al.
Publicado: (2023)
Preference Learning Algorithms Do Not Learn Preference Rankings
por: Chen, Angelica, et al.
Publicado: (2024)
por: Chen, Angelica, et al.
Publicado: (2024)
Generalist Foundation Models Are Not Clinical Enough for Hospital Operations
por: Jiang, Lavender Y., et al.
Publicado: (2025)
por: Jiang, Lavender Y., et al.
Publicado: (2025)
ICLR: In-Context Learning of Representations
por: Park, Core Francisco, et al.
Publicado: (2024)
por: Park, Core Francisco, et al.
Publicado: (2024)
Aioli: A Unified Optimization Framework for Language Model Data Mixing
por: Chen, Mayee F., et al.
Publicado: (2024)
por: Chen, Mayee F., et al.
Publicado: (2024)
Patent Representation Learning via Self-supervision
por: Zuo, You, et al.
Publicado: (2025)
por: Zuo, You, et al.
Publicado: (2025)
Learning to Represent Individual Differences for Choice Decision Making
por: Chen, Yan-Ying, et al.
Publicado: (2025)
por: Chen, Yan-Ying, et al.
Publicado: (2025)
Human and Automatic Interpretation of Romanian Noun Compounds
por: Marinescu, Ioana, et al.
Publicado: (2024)
por: Marinescu, Ioana, et al.
Publicado: (2024)
Machine Learning: a Lecture Note
por: Cho, Kyunghyun
Publicado: (2025)
por: Cho, Kyunghyun
Publicado: (2025)
Translating Hanja Historical Documents to Contemporary Korean and English
por: Son, Juhee, et al.
Publicado: (2022)
por: Son, Juhee, et al.
Publicado: (2022)
Hyperparameters in Continual Learning: A Reality Check
por: Cha, Sungmin, et al.
Publicado: (2024)
por: Cha, Sungmin, et al.
Publicado: (2024)
A Brief Introduction to Causal Inference in Machine Learning
por: Cho, Kyunghyun
Publicado: (2024)
por: Cho, Kyunghyun
Publicado: (2024)
On the Loss of Context-awareness in General Instruction Fine-tuning
por: Wang, Yihan, et al.
Publicado: (2024)
por: Wang, Yihan, et al.
Publicado: (2024)
Automatic Replication of LLM Mistakes in Medical Conversations
por: Proniakin, Oleksii, et al.
Publicado: (2025)
por: Proniakin, Oleksii, et al.
Publicado: (2025)
Uncovering Emergent Physics Representations Learned In-Context by Large Language Models
por: Song, Yeongwoo, et al.
Publicado: (2025)
por: Song, Yeongwoo, et al.
Publicado: (2025)
Relationships are Complicated! An Analysis of Relationships Between Datasets on the Web
por: Lin, Kate, et al.
Publicado: (2024)
por: Lin, Kate, et al.
Publicado: (2024)
Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs
por: Venkatesh, Sohan
Publicado: (2026)
por: Venkatesh, Sohan
Publicado: (2026)
Brewing Knowledge in Context: Distillation Perspectives on In-Context Learning
por: Li, Chengye, et al.
Publicado: (2025)
por: Li, Chengye, et al.
Publicado: (2025)
Training Language Models with Language Feedback at Scale
por: Scheurer, Jérémy, et al.
Publicado: (2023)
por: Scheurer, Jérémy, et al.
Publicado: (2023)
Gateformer: Advancing Multivariate Time Series Forecasting through Temporal and Variate-Wise Attention with Gated Representations
por: Lan, Yu-Hsiang, et al.
Publicado: (2025)
por: Lan, Yu-Hsiang, et al.
Publicado: (2025)
Large Language Models Predict Functional Outcomes after Acute Ischemic Stroke
por: Kapoor, Anjali K., et al.
Publicado: (2026)
por: Kapoor, Anjali K., et al.
Publicado: (2026)
In-Context Learning Dynamics with Random Binary Sequences
por: Bigelow, Eric J., et al.
Publicado: (2023)
por: Bigelow, Eric J., et al.
Publicado: (2023)
On the Mathematical Relationship Between Layer Normalization and Dynamic Activation Functions
por: Stollenwerk, Felix
Publicado: (2025)
por: Stollenwerk, Felix
Publicado: (2025)
Towards Better Understanding of In-Context Learning Ability from In-Context Uncertainty Quantification
por: Liu, Shang, et al.
Publicado: (2024)
por: Liu, Shang, et al.
Publicado: (2024)
Stronger Random Baselines for In-Context Learning
por: Yauney, Gregory, et al.
Publicado: (2024)
por: Yauney, Gregory, et al.
Publicado: (2024)
Ejemplares similares
-
Clinically Grounded Agent-based Report Evaluation: An Interpretable Metric for Radiology Report Generation
por: Dua, Radhika, et al.
Publicado: (2025) -
Hyperparameter Loss Surfaces Are Simple Near their Optima
por: Lourie, Nicholas, et al.
Publicado: (2025) -
Show Your Work with Confidence: Confidence Bands for Tuning Curves
por: Lourie, Nicholas, et al.
Publicado: (2023) -
Scaling Laws Are Unreliable for Downstream Tasks: A Reality Check
por: Lourie, Nicholas, et al.
Publicado: (2025) -
Efficient semantic uncertainty quantification in language models via diversity-steered sampling
por: Park, Ji Won, et al.
Publicado: (2025)