Getting aligned on representational alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Sucholutsky, Ilia, Muttenthaler, Lukas, Weller, Adrian, Peng, Andi, Bobu, Andreea, Kim, Been, Love, Bradley C., Cueva, Christopher J., Grant, Erin, Groen, Iris, Achterberg, Jascha, Tenenbaum, Joshua B., Collins, Katherine M., Hermann, Katherine L., Oktar, Kerem, Greff, Klaus, Hebart, Martin N., Cloos, Nathan, Kriegeskorte, Nikolaus, Jacoby, Nori, Zhang, Qiuyi, Marjieh, Raja, Geirhos, Robert, Chen, Sherol, Kornblith, Simon, Rane, Sunayana, Konkle, Talia, O'Connell, Thomas P., Unterthiner, Thomas, Lampinen, Andrew K., Müller, Klaus-Robert, Toneva, Mariya, Griffiths, Thomas L. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Aligning Machine and Human Visual Representations across Abstraction Levels
by: Muttenthaler, Lukas, et al.
Published: (2024)
by: Muttenthaler, Lukas, et al.
Published: (2024)
Dimensions underlying the representational alignment of deep neural networks with humans
by: Mahner, Florian P., et al.
Published: (2024)
by: Mahner, Florian P., et al.
Published: (2024)
Human alignment of neural network representations
by: Muttenthaler, Lukas, et al.
Published: (2022)
by: Muttenthaler, Lukas, et al.
Published: (2022)
Set Learning for Accurate and Calibrated Models
by: Muttenthaler, Lukas, et al.
Published: (2023)
by: Muttenthaler, Lukas, et al.
Published: (2023)
Concept Alignment
by: Rane, Sunayana, et al.
Published: (2024)
by: Rane, Sunayana, et al.
Published: (2024)
The Reasonable Person Standard for AI
by: Rane, Sunayana
Published: (2024)
by: Rane, Sunayana
Published: (2024)
What is a Number, That a Large Language Model May Know It?
by: Marjieh, Raja, et al.
Published: (2025)
by: Marjieh, Raja, et al.
Published: (2025)
Language models and brains align due to more than next-word prediction and word-level information
by: Merlin, Gabriele, et al.
Published: (2022)
by: Merlin, Gabriele, et al.
Published: (2022)
Why Human Guidance Matters in Collaborative Vibe Coding
by: Hu, Haoyu, et al.
Published: (2026)
by: Hu, Haoyu, et al.
Published: (2026)
Investigating Concept Alignment Using Implausible Category Members
by: Rane, Sunayana, et al.
Published: (2026)
by: Rane, Sunayana, et al.
Published: (2026)
We Can't Understand AI Using our Existing Vocabulary
by: Hewitt, John, et al.
Published: (2025)
by: Hewitt, John, et al.
Published: (2025)
Convolutional Neural Networks Can (Meta-)Learn the Same-Different Relation
by: Gupta, Max, et al.
Published: (2025)
by: Gupta, Max, et al.
Published: (2025)
Understanding Visual Feature Reliance through the Lens of Complexity
by: Fel, Thomas, et al.
Published: (2024)
by: Fel, Thomas, et al.
Published: (2024)
Are aligned neural networks adversarially aligned?
by: Carlini, Nicholas, et al.
Published: (2023)
by: Carlini, Nicholas, et al.
Published: (2023)
Human-AI Synergy Supports Collective Creative Search
by: Li, Chenyi, et al.
Published: (2026)
by: Li, Chenyi, et al.
Published: (2026)
Identifying, Evaluating, and Mitigating Risks of AI Thought Partnerships
by: Oktar, Kerem, et al.
Published: (2025)
by: Oktar, Kerem, et al.
Published: (2025)
MAPS: Masked Attribution-based Probing of Strategies- A computational framework to align human and model explanations
by: Muzellec, Sabine, et al.
Published: (2025)
by: Muzellec, Sabine, et al.
Published: (2025)
Neologism Learning for Controllability and Self-Verbalization
by: Hewitt, John, et al.
Published: (2025)
by: Hewitt, John, et al.
Published: (2025)
In‐Plane versus Through‐Plane Thermal Gradients during Cyclic Aging of Lithium‐Ion Batteries: An Experimental Study
by: Lisa Cloos, et al.
Published: (2025)
by: Lisa Cloos, et al.
Published: (2025)
A Rational Analysis of the Speech-to-Song Illusion
by: Marjieh, Raja, et al.
Published: (2024)
by: Marjieh, Raja, et al.
Published: (2024)
Characterizing the Large‐Scale Structure of Multimodal Semantic Networks
by: Raja Marjieh, et al.
Published: (2025)
by: Raja Marjieh, et al.
Published: (2025)
DyST: Towards Dynamic Neural Scene Representations on Real-World Videos
by: Seitzer, Maximilian, et al.
Published: (2023)
by: Seitzer, Maximilian, et al.
Published: (2023)
Reduced density fluctuations via anti-aligning in active matter
by: Boltz, Horst-Holger, et al.
Published: (2024)
by: Boltz, Horst-Holger, et al.
Published: (2024)
Language Model Teams as Distributed Systems
by: Mieczkowski, Elizabeth, et al.
Published: (2026)
by: Mieczkowski, Elizabeth, et al.
Published: (2026)
Objective drives the consistency of representational similarity across datasets
by: Ciernik, Laure, et al.
Published: (2024)
by: Ciernik, Laure, et al.
Published: (2024)
Don't trust your eyes: on the (un)reliability of feature visualizations
by: Geirhos, Robert, et al.
Published: (2023)
by: Geirhos, Robert, et al.
Published: (2023)
Revisiting Rogers' Paradox in the Context of Human-AI Interaction
by: Collins, Katherine M., et al.
Published: (2025)
by: Collins, Katherine M., et al.
Published: (2025)
TERRITÓRIOS ÉTNICOS NO PÓS-ABOLIÇÃO: O CASO DO QUILOMBO DA MORMAÇA (RS)
by: Sherol dos Santos
Published: (2009)
by: Sherol dos Santos
Published: (2009)
A computing machinery using a continuous memory tape
by: Oktar, Yigit
Published: (2023)
by: Oktar, Yigit
Published: (2023)
emcramer/align_spheroid: align_spheroid
by: Eric Cramer
Published: (2025)
by: Eric Cramer
Published: (2025)
Non-reciprocal anti-aligning active mixtures: deriving the exact Boltzmann collision operator
by: Mihatsch, Jakob, et al.
Published: (2025)
by: Mihatsch, Jakob, et al.
Published: (2025)
Characterizing Similarities and Divergences in Conversational Tones in Humans and LLMs by Sampling with People
by: Huang, Dun-Ming, et al.
Published: (2024)
by: Huang, Dun-Ming, et al.
Published: (2024)
Woke women of the 90s return: Toxic white feminism and the Murphy Brown and Roseanne reboots
by: Amanda Konkle
Published: (2024)
by: Amanda Konkle
Published: (2024)
Gradient extremals, talwegs, valleys, and directional alignment for generic gradient descent
by: Bégout, Pascal, et al.
Published: (2026)
by: Bégout, Pascal, et al.
Published: (2026)
B-modes from galaxy cluster alignments in future surveys
by: Georgiou, Christos, et al.
Published: (2023)
by: Georgiou, Christos, et al.
Published: (2023)
How Aligned are Different Alignment Metrics?
by: Ahlert, Jannis, et al.
Published: (2024)
by: Ahlert, Jannis, et al.
Published: (2024)
Three-point intrinsic alignments of galaxies and haloes in the FLAMINGO simulations
by: Vedder, Casper, et al.
Published: (2026)
by: Vedder, Casper, et al.
Published: (2026)
Mesoscopic theory of flocking with alignment and anti-alignment copying
by: Zheng, Chunming
Published: (2026)
by: Zheng, Chunming
Published: (2026)
Growth of aligned and twisted hexagonal boron nitride on Ir(110)
by: Michely, Thomas, et al.
Published: (2023)
by: Michely, Thomas, et al.
Published: (2023)
Multi-sequence alignment using the Quantum Approximate Optimization Algorithm
by: Madsen, Sebastian Yde, et al.
Published: (2023)
by: Madsen, Sebastian Yde, et al.
Published: (2023)
Similar Items
-
Aligning Machine and Human Visual Representations across Abstraction Levels
by: Muttenthaler, Lukas, et al.
Published: (2024) -
Dimensions underlying the representational alignment of deep neural networks with humans
by: Mahner, Florian P., et al.
Published: (2024) -
Human alignment of neural network representations
by: Muttenthaler, Lukas, et al.
Published: (2022) -
Set Learning for Accurate and Calibrated Models
by: Muttenthaler, Lukas, et al.
Published: (2023) -
Concept Alignment
by: Rane, Sunayana, et al.
Published: (2024)