ICLR: In-Context Learning of Representations
Fuente:
arXiv
Guardado en:
| Autores principales: | Park, Core Francisco, Lee, Andrew, Lubana, Ekdeep Singh, Yang, Yongyi, Okawa, Maya, Nishi, Kento, Wattenberg, Martin, Tanaka, Hidenori |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Emergence of Hidden Capabilities: Exploring Learning Dynamics in Concept Space
por: Park, Core Francisco, et al.
Publicado: (2024)
por: Park, Core Francisco, et al.
Publicado: (2024)
Competition Dynamics Shape Algorithmic Phases of In-Context Learning
por: Park, Core Francisco, et al.
Publicado: (2024)
por: Park, Core Francisco, et al.
Publicado: (2024)
Swing-by Dynamics in Concept Learning and Compositional Generalization
por: Yang, Yongyi, et al.
Publicado: (2024)
por: Yang, Yongyi, et al.
Publicado: (2024)
In-Context Learning Strategies Emerge Rationally
por: Wurgaft, Daniel, et al.
Publicado: (2025)
por: Wurgaft, Daniel, et al.
Publicado: (2025)
Towards an Understanding of Stepwise Inference in Transformers: A Synthetic Graph Navigation Model
por: Khona, Mikail, et al.
Publicado: (2024)
por: Khona, Mikail, et al.
Publicado: (2024)
Emergence of Hierarchical Emotion Organization in Large Language Models
por: Zhao, Bo, et al.
Publicado: (2025)
por: Zhao, Bo, et al.
Publicado: (2025)
In-Context Learning Dynamics with Random Binary Sequences
por: Bigelow, Eric J., et al.
Publicado: (2023)
por: Bigelow, Eric J., et al.
Publicado: (2023)
Representation Shattering in Transformers: A Synthetic Study with Knowledge Editing
por: Nishi, Kento, et al.
Publicado: (2024)
por: Nishi, Kento, et al.
Publicado: (2024)
Belief Dynamics Reveal the Dual Nature of In-Context Learning and Activation Steering
por: Bigelow, Eric, et al.
Publicado: (2025)
por: Bigelow, Eric, et al.
Publicado: (2025)
$\textit{New News}$: System-2 Fine-tuning for Robust Integration of New Knowledge
por: Park, Core Francisco, et al.
Publicado: (2025)
por: Park, Core Francisco, et al.
Publicado: (2025)
How Do LLMs Persuade? Linear Probes Can Uncover Persuasion Dynamics in Multi-Turn Conversations
por: Jaipersaud, Brandon, et al.
Publicado: (2025)
por: Jaipersaud, Brandon, et al.
Publicado: (2025)
A Percolation Model of Emergence: Analyzing Transformers Trained on a Formal Language
por: Lubana, Ekdeep Singh, et al.
Publicado: (2024)
por: Lubana, Ekdeep Singh, et al.
Publicado: (2024)
Are language models aware of the road not taken? Token-level uncertainty and hidden state dynamics
por: Zur, Amir, et al.
Publicado: (2025)
por: Zur, Amir, et al.
Publicado: (2025)
Towards Reliable Evaluation of Behavior Steering Interventions in LLMs
por: Pres, Itamar, et al.
Publicado: (2024)
por: Pres, Itamar, et al.
Publicado: (2024)
Compositional Abilities Emerge Multiplicatively: Exploring Diffusion Models on a Synthetic Task
por: Okawa, Maya, et al.
Publicado: (2023)
por: Okawa, Maya, et al.
Publicado: (2023)
From Isolation to Entanglement: When Do Interpretability Methods Identify and Disentangle Known Concepts?
por: Mueller, Aaron, et al.
Publicado: (2025)
por: Mueller, Aaron, et al.
Publicado: (2025)
Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space
por: Bigelow, Eric, et al.
Publicado: (2026)
por: Bigelow, Eric, et al.
Publicado: (2026)
Convergent World Representations and Divergent Tasks
por: Park, Core Francisco
Publicado: (2026)
por: Park, Core Francisco
Publicado: (2026)
Decomposing Elements of Problem Solving: What "Math" Does RL Teach?
por: Qin, Tian, et al.
Publicado: (2025)
por: Qin, Tian, et al.
Publicado: (2025)
What Does it Mean for a Neural Network to Learn a "World Model"?
por: Li, Kenneth, et al.
Publicado: (2025)
por: Li, Kenneth, et al.
Publicado: (2025)
A Mechanistic Understanding of Alignment Algorithms: A Case Study on DPO and Toxicity
por: Lee, Andrew, et al.
Publicado: (2024)
por: Lee, Andrew, et al.
Publicado: (2024)
Uncovering Conceptual Blindspots in Generative Image Models Using Sparse Autoencoders
por: Bohacek, Matyas, et al.
Publicado: (2025)
por: Bohacek, Matyas, et al.
Publicado: (2025)
Projecting Assumptions: The Duality Between Sparse Autoencoders and Concept Geometry
por: Hindupur, Sai Sumedh R., et al.
Publicado: (2025)
por: Hindupur, Sai Sumedh R., et al.
Publicado: (2025)
Into the Rabbit Hull: From Task-Relevant Concepts in DINO to Minkowski Geometry
por: Fel, Thomas, et al.
Publicado: (2025)
por: Fel, Thomas, et al.
Publicado: (2025)
Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task
por: Li, Kenneth, et al.
Publicado: (2022)
por: Li, Kenneth, et al.
Publicado: (2022)
Distinct Computations Emerge From Compositional Curricula in In-Context Learning
por: Lee, Jin Hwa, et al.
Publicado: (2025)
por: Lee, Jin Hwa, et al.
Publicado: (2025)
Improving Multi-hop Logical Reasoning in Knowledge Graphs with Context-Aware Query Representation Learning
por: Kim, Jeonghoon, et al.
Publicado: (2024)
por: Kim, Jeonghoon, et al.
Publicado: (2024)
Dialogue Action Tokens: Steering Language Models in Goal-Directed Dialogue with a Multi-Turn Planner
por: Li, Kenneth, et al.
Publicado: (2024)
por: Li, Kenneth, et al.
Publicado: (2024)
When Bad Data Leads to Good Models
por: Li, Kenneth, et al.
Publicado: (2025)
por: Li, Kenneth, et al.
Publicado: (2025)
The Impact of Off-Policy Training Data on Probe Generalisation
por: Kirch, Nathalie, et al.
Publicado: (2025)
por: Kirch, Nathalie, et al.
Publicado: (2025)
Provable Low-Frequency Bias of In-Context Learning of Representations
por: Yang, Yongyi, et al.
Publicado: (2025)
por: Yang, Yongyi, et al.
Publicado: (2025)
Context Structure Reshapes the Representational Geometry of Language Models
por: Hosseini, Eghbal A., et al.
Publicado: (2026)
por: Hosseini, Eghbal A., et al.
Publicado: (2026)
Enhancing Robustness of Retrieval-Augmented Language Models with In-Context Learning
por: Park, Seong-Il, et al.
Publicado: (2024)
por: Park, Seong-Il, et al.
Publicado: (2024)
Language Models Struggle to Use Representations Learned In-Context
por: Lepori, Michael A., et al.
Publicado: (2026)
por: Lepori, Michael A., et al.
Publicado: (2026)
Inference-Time Intervention: Eliciting Truthful Answers from a Language Model
por: Li, Kenneth, et al.
Publicado: (2023)
por: Li, Kenneth, et al.
Publicado: (2023)
In-Context Learning for Long-Context Sentiment Analysis on Infrastructure Project Opinions
por: Shamshiri, Alireza, et al.
Publicado: (2024)
por: Shamshiri, Alireza, et al.
Publicado: (2024)
Forking Paths in Neural Text Generation
por: Bigelow, Eric, et al.
Publicado: (2024)
por: Bigelow, Eric, et al.
Publicado: (2024)
Enhancing Hallucination Detection via Future Context
por: Lee, Joosung, et al.
Publicado: (2025)
por: Lee, Joosung, et al.
Publicado: (2025)
Can LLM feedback enhance review quality? A randomized study of 20K reviews at ICLR 2025
por: Thakkar, Nitya, et al.
Publicado: (2025)
por: Thakkar, Nitya, et al.
Publicado: (2025)
Insights from the ICLR Peer Review and Rebuttal Process
por: Kargaran, Amir Hossein, et al.
Publicado: (2025)
por: Kargaran, Amir Hossein, et al.
Publicado: (2025)
Ejemplares similares
-
Emergence of Hidden Capabilities: Exploring Learning Dynamics in Concept Space
por: Park, Core Francisco, et al.
Publicado: (2024) -
Competition Dynamics Shape Algorithmic Phases of In-Context Learning
por: Park, Core Francisco, et al.
Publicado: (2024) -
Swing-by Dynamics in Concept Learning and Compositional Generalization
por: Yang, Yongyi, et al.
Publicado: (2024) -
In-Context Learning Strategies Emerge Rationally
por: Wurgaft, Daniel, et al.
Publicado: (2025) -
Towards an Understanding of Stepwise Inference in Transformers: A Synthetic Graph Navigation Model
por: Khona, Mikail, et al.
Publicado: (2024)