Identifying and interpreting non-aligned human conceptual representations using language modeling
Fuente:
arXiv
Guardado en:
| Autores principales: | Bao, Wanqian, Hasson, Uri |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
TouchAI: Exploring human-AI perceptual alignment in touch through language model representations
por: Zhong, Shu, et al.
Publicado: (2024)
por: Zhong, Shu, et al.
Publicado: (2024)
Revealing emergent human-like conceptual representations from language prediction
por: Xu, Ningyu, et al.
Publicado: (2025)
por: Xu, Ningyu, et al.
Publicado: (2025)
Strong and weak alignment of large language models with human values
por: Khamassi, Mehdi, et al.
Publicado: (2024)
por: Khamassi, Mehdi, et al.
Publicado: (2024)
Can LLMs interpret figurative language as humans do?: surface-level vs representational similarity
por: Bollepally, Samhita, et al.
Publicado: (2026)
por: Bollepally, Samhita, et al.
Publicado: (2026)
Do large language models resemble humans in language use?
por: Cai, Zhenguang G., et al.
Publicado: (2023)
por: Cai, Zhenguang G., et al.
Publicado: (2023)
Brains and language models converge on a shared conceptual space across different languages
por: Zada, Zaid, et al.
Publicado: (2025)
por: Zada, Zaid, et al.
Publicado: (2025)
Distinct social-linguistic processing between humans and large audio-language models: Evidence from model-brain alignment
por: Wu, Hanlin, et al.
Publicado: (2025)
por: Wu, Hanlin, et al.
Publicado: (2025)
Context informs pragmatic interpretation in vision-language models
por: Tan, Alvin Wei Ming, et al.
Publicado: (2025)
por: Tan, Alvin Wei Ming, et al.
Publicado: (2025)
Assessing the alignment between infants' visual and linguistic experience using multimodal language models
por: Tan, Alvin Wei Ming, et al.
Publicado: (2025)
por: Tan, Alvin Wei Ming, et al.
Publicado: (2025)
Studies with impossible languages falsify LMs as models of human language
por: Bowers, Jeffrey S., et al.
Publicado: (2025)
por: Bowers, Jeffrey S., et al.
Publicado: (2025)
Visual representations in the human brain are aligned with large language models
por: Doerig, Adrien, et al.
Publicado: (2022)
por: Doerig, Adrien, et al.
Publicado: (2022)
InfAlign: Inference-aware language model alignment
por: Balashankar, Ananth, et al.
Publicado: (2024)
por: Balashankar, Ananth, et al.
Publicado: (2024)
Human-interpretable clustering of short-text using large language models
por: Miller, Justin K., et al.
Publicado: (2024)
por: Miller, Justin K., et al.
Publicado: (2024)
Multilingual large language models leak human stereotypes across language boundaries
por: Cao, Yang Trista, et al.
Publicado: (2023)
por: Cao, Yang Trista, et al.
Publicado: (2023)
The American Ghost in the Machine: How language models align culturally and the effects of cultural prompting
por: Luther, James, et al.
Publicado: (2025)
por: Luther, James, et al.
Publicado: (2025)
Aligning language models with human preferences
por: Korbak, Tomasz
Publicado: (2024)
por: Korbak, Tomasz
Publicado: (2024)
Language models align with human judgments on key grammatical constructions
por: Hu, Jennifer, et al.
Publicado: (2024)
por: Hu, Jennifer, et al.
Publicado: (2024)
Cognitive models can reveal interpretable value trade-offs in language models
por: Murthy, Sonia K., et al.
Publicado: (2025)
por: Murthy, Sonia K., et al.
Publicado: (2025)
Transferable speech-to-text large language model alignment module
por: Wu, Boyong, et al.
Publicado: (2024)
por: Wu, Boyong, et al.
Publicado: (2024)
Do self-supervised speech and language models extract similar representations as human brain?
por: Chen, Peili, et al.
Publicado: (2023)
por: Chen, Peili, et al.
Publicado: (2023)
Just-in-time and distributed task representations in language models
por: Li, Yuxuan, et al.
Publicado: (2025)
por: Li, Yuxuan, et al.
Publicado: (2025)
Are they human? Detecting large language models by probing human memory constraints
por: Schug, Simon, et al.
Publicado: (2026)
por: Schug, Simon, et al.
Publicado: (2026)
Evolution and compression in LLMs: On the emergence of human-aligned categorization
por: Imel, Nathaniel, et al.
Publicado: (2025)
por: Imel, Nathaniel, et al.
Publicado: (2025)
Effect-driven interpretation: Functors for natural language composition
por: Bumford, Dylan, et al.
Publicado: (2025)
por: Bumford, Dylan, et al.
Publicado: (2025)
Large language models have learned to use language
por: Lupyan, Gary
Publicado: (2025)
por: Lupyan, Gary
Publicado: (2025)
SemPool: Simple, robust, and interpretable KG pooling for enhancing language models
por: Mavromatis, Costas, et al.
Publicado: (2024)
por: Mavromatis, Costas, et al.
Publicado: (2024)
One fish, two fish, but not the whole sea: Alignment reduces language models' conceptual diversity
por: Murthy, Sonia K., et al.
Publicado: (2024)
por: Murthy, Sonia K., et al.
Publicado: (2024)
Annotation alignment: Comparing LLM and human annotations of conversational safety
por: Movva, Rajiv, et al.
Publicado: (2024)
por: Movva, Rajiv, et al.
Publicado: (2024)
Fine-tuning with HED-IT: The impact of human post-editing for dialogical language models
por: Occhipinti, Daniela, et al.
Publicado: (2024)
por: Occhipinti, Daniela, et al.
Publicado: (2024)
Beyond topography: Topographic regularization improves robustness and reshapes representations in convolutional neural networks
por: Truong, Nhut, et al.
Publicado: (2025)
por: Truong, Nhut, et al.
Publicado: (2025)
Logical forms complement probability in understanding language model (and human) performance
por: Wang, Yixuan, et al.
Publicado: (2025)
por: Wang, Yixuan, et al.
Publicado: (2025)
Leveraging large language models for efficient representation learning for entity resolution
por: Xu, Xiaowei, et al.
Publicado: (2024)
por: Xu, Xiaowei, et al.
Publicado: (2024)
ARC-Encoder: learning compressed text representations for large language models
por: Pilchen, Hippolyte, et al.
Publicado: (2025)
por: Pilchen, Hippolyte, et al.
Publicado: (2025)
XLM: A Python package for non-autoregressive language models
por: Patel, Dhruvesh, et al.
Publicado: (2025)
por: Patel, Dhruvesh, et al.
Publicado: (2025)
How desirable is alignment between LLMs and linguistically diverse human users?
por: Knoeferle, Pia, et al.
Publicado: (2025)
por: Knoeferle, Pia, et al.
Publicado: (2025)
Comparing large language models and human programmers for generating programming code
por: Hou, Wenpin, et al.
Publicado: (2024)
por: Hou, Wenpin, et al.
Publicado: (2024)
Explaining Human Comparisons using Alignment-Importance Heatmaps
por: Truong, Nhut, et al.
Publicado: (2024)
por: Truong, Nhut, et al.
Publicado: (2024)
Generics in science communication: Misaligned interpretations across laypeople, scientists, and large language models
por: Peters, Uwe, et al.
Publicado: (2026)
por: Peters, Uwe, et al.
Publicado: (2026)
When the LM misunderstood the human chuckled: Analyzing garden path effects in humans and language models
por: Amouyal, Samuel Joseph, et al.
Publicado: (2025)
por: Amouyal, Samuel Joseph, et al.
Publicado: (2025)
InstructPatentGPT: Training patent language models to follow instructions with human feedback
por: Lee, Jieh-Sheng
Publicado: (2024)
por: Lee, Jieh-Sheng
Publicado: (2024)
Ejemplares similares
-
TouchAI: Exploring human-AI perceptual alignment in touch through language model representations
por: Zhong, Shu, et al.
Publicado: (2024) -
Revealing emergent human-like conceptual representations from language prediction
por: Xu, Ningyu, et al.
Publicado: (2025) -
Strong and weak alignment of large language models with human values
por: Khamassi, Mehdi, et al.
Publicado: (2024) -
Can LLMs interpret figurative language as humans do?: surface-level vs representational similarity
por: Bollepally, Samhita, et al.
Publicado: (2026) -
Do large language models resemble humans in language use?
por: Cai, Zhenguang G., et al.
Publicado: (2023)