Abstraction Alignment: Comparing Model-Learned and Human-Encoded Conceptual Relationships
Fuente:
arXiv
Saved in:
| Main Authors: | Boggust, Angie, Bang, Hyemin, Strobelt, Hendrik, Satyanarayan, Arvind |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Compress and Compare: Interactively Evaluating Efficiency and Behavior Across ML Model Compression Experiments
by: Boggust, Angie, et al.
Published: (2024)
by: Boggust, Angie, et al.
Published: (2024)
Creo: From One-Shot Image Generation to Progressive, Co-Creative Ideation
by: De Simone, Zoe, et al.
Published: (2026)
by: De Simone, Zoe, et al.
Published: (2026)
Value Profiles for Encoding Human Variation
by: Sorensen, Taylor, et al.
Published: (2025)
by: Sorensen, Taylor, et al.
Published: (2025)
Diffusion Explainer: Visual Explanation for Text-to-image Stable Diffusion
by: Lee, Seongmin, et al.
Published: (2023)
by: Lee, Seongmin, et al.
Published: (2023)
Toward Cultural Interpretability: A Linguistic Anthropological Framework for Describing and Evaluating Large Language Models (LLMs)
by: Jones, Graham M., et al.
Published: (2024)
by: Jones, Graham M., et al.
Published: (2024)
Heterogeneous Value Alignment Evaluation for Large Language Models
by: Zhang, Zhaowei, et al.
Published: (2023)
by: Zhang, Zhaowei, et al.
Published: (2023)
Comparing Exploration-Exploitation Strategies of LLMs and Humans: Insights from Standard Multi-armed Bandit Experiments
by: Zhang, Ziyuan, et al.
Published: (2025)
by: Zhang, Ziyuan, et al.
Published: (2025)
MotionTeller: Multi-modal Integration of Wearable Time-Series with LLMs for Health and Behavioral Understanding
by: Zhang, Aiwei, et al.
Published: (2025)
by: Zhang, Aiwei, et al.
Published: (2025)
Value-Action Alignment in Large Language Models under Privacy-Prosocial Conflict
by: Chen, Guanyu, et al.
Published: (2026)
by: Chen, Guanyu, et al.
Published: (2026)
DiffusionWorldViewer: Exposing and Broadening the Worldview Reflected by Generative Text-to-Image Models
by: De Simone, Zoe, et al.
Published: (2023)
by: De Simone, Zoe, et al.
Published: (2023)
LLM Comparator: Visual Analytics for Side-by-Side Evaluation of Large Language Models
by: Kahng, Minsuk, et al.
Published: (2024)
by: Kahng, Minsuk, et al.
Published: (2024)
Automating Customer Needs Analysis: A Comparative Study of Large Language Models in the Travel Industry
by: Barandoni, Simone, et al.
Published: (2024)
by: Barandoni, Simone, et al.
Published: (2024)
Understanding Large Language Model Behaviors through Interactive Counterfactual Generation and Analysis
by: Cheng, Furui, et al.
Published: (2024)
by: Cheng, Furui, et al.
Published: (2024)
Augmenting Human Evaluation with LLM Judges: How Many Human Reviews Do You Need?
by: Kim, Jane Paik
Published: (2026)
by: Kim, Jane Paik
Published: (2026)
Learning to Decide with AI Assistance under Human-Alignment
by: Benz, Nina Corvelo, et al.
Published: (2026)
by: Benz, Nina Corvelo, et al.
Published: (2026)
UniAutoML: A Human-Centered Framework for Unified Discriminative and Generative AutoML with Large Language Models
by: Guo, Jiayi, et al.
Published: (2024)
by: Guo, Jiayi, et al.
Published: (2024)
Transformer Explainer: Interactive Learning of Text-Generative Models
by: Cho, Aeree, et al.
Published: (2024)
by: Cho, Aeree, et al.
Published: (2024)
Addressing the Ecological Fallacy in Larger LMs with Human Context
by: Soni, Nikita, et al.
Published: (2026)
by: Soni, Nikita, et al.
Published: (2026)
Almost AI, Almost Human: The Challenge of Detecting AI-Polished Writing
by: Saha, Shoumik, et al.
Published: (2025)
by: Saha, Shoumik, et al.
Published: (2025)
Vocal Sandbox: Continual Learning and Adaptation for Situated Human-Robot Collaboration
by: Grannen, Jennifer, et al.
Published: (2024)
by: Grannen, Jennifer, et al.
Published: (2024)
HybridQuestion: Human-AI Collaboration for Identifying High-Impact Research Questions
by: Zhao, Keyu, et al.
Published: (2025)
by: Zhao, Keyu, et al.
Published: (2025)
Collaborative Causal Sensemaking: Closing the Complementarity Gap in Human-AI Decision Support
by: Jain, Raunak
Published: (2025)
by: Jain, Raunak
Published: (2025)
Evaluation of LLMs-based Hidden States as Author Representations for Psychological Human-Centered NLP Tasks
by: Soni, Nikita, et al.
Published: (2025)
by: Soni, Nikita, et al.
Published: (2025)
Prompting in the Dark: Assessing Human Performance in Prompt Engineering for Data Labeling When Gold Labels Are Absent
by: He, Zeyu, et al.
Published: (2025)
by: He, Zeyu, et al.
Published: (2025)
Leading Across the Spectrum of Human-AI Relationships: A Conceptual Framework for Increasingly Heterogeneous Teams
by: Jadad, Alejandro R.
Published: (2026)
by: Jadad, Alejandro R.
Published: (2026)
ProgressGym: Alignment with a Millennium of Moral Progress
by: Qiu, Tianyi, et al.
Published: (2024)
by: Qiu, Tianyi, et al.
Published: (2024)
Lifelong and Continual Learning Dialogue Systems
by: Mazumder, Sahisnu, et al.
Published: (2022)
by: Mazumder, Sahisnu, et al.
Published: (2022)
Reinforcement Learning for Personalized Dialogue Management
by: Hengst, Floris den, et al.
Published: (2019)
by: Hengst, Floris den, et al.
Published: (2019)
TalkWithMachines: Enhancing Human-Robot Interaction for Interpretable Industrial Robotics Through Large/Vision Language Models
by: Abbas, Ammar N., et al.
Published: (2024)
by: Abbas, Ammar N., et al.
Published: (2024)
Rationalize: Shared Semantic Reasoning for Human-AI Alignment
by: Dasgupta, Aritra, et al.
Published: (2026)
by: Dasgupta, Aritra, et al.
Published: (2026)
RuleAlign: Making Large Language Models Better Physicians with Diagnostic Rule Alignment
by: Wang, Xiaohan, et al.
Published: (2024)
by: Wang, Xiaohan, et al.
Published: (2024)
Self-Reported Confidence of Large Language Models in Gastroenterology: Analysis of Commercial, Open-Source, and Quantized Models
by: Naderi, Nariman, et al.
Published: (2025)
by: Naderi, Nariman, et al.
Published: (2025)
Augmenting Automation: Intent-Based User Instruction Classification with Machine Learning
by: Basyal, Lochan, et al.
Published: (2024)
by: Basyal, Lochan, et al.
Published: (2024)
Advanced Machine Learning Techniques for Social Support Detection on Social Media
by: Kolesnikova, Olga, et al.
Published: (2025)
by: Kolesnikova, Olga, et al.
Published: (2025)
Stress Detection on Code-Mixed Texts in Dravidian Languages using Machine Learning
by: Ramos, L., et al.
Published: (2024)
by: Ramos, L., et al.
Published: (2024)
Hierarchical Reward Design from Language: Enhancing Alignment of Agent Behavior with Human Specifications
by: Qian, Zhiqin, et al.
Published: (2026)
by: Qian, Zhiqin, et al.
Published: (2026)
Deep Neural Networks and Brain Alignment: Brain Encoding and Decoding (Survey)
by: Oota, Subba Reddy, et al.
Published: (2023)
by: Oota, Subba Reddy, et al.
Published: (2023)
HumanAgencyBench: Scalable Evaluation of Human Agency Support in AI Assistants
by: Sturgeon, Benjamin, et al.
Published: (2025)
by: Sturgeon, Benjamin, et al.
Published: (2025)
Benchmarking Gender and Political Bias in Large Language Models
by: Yang, Jinrui, et al.
Published: (2025)
by: Yang, Jinrui, et al.
Published: (2025)
Wordflow: Social Prompt Engineering for Large Language Models
by: Wang, Zijie J., et al.
Published: (2024)
by: Wang, Zijie J., et al.
Published: (2024)
Similar Items
-
Compress and Compare: Interactively Evaluating Efficiency and Behavior Across ML Model Compression Experiments
by: Boggust, Angie, et al.
Published: (2024) -
Creo: From One-Shot Image Generation to Progressive, Co-Creative Ideation
by: De Simone, Zoe, et al.
Published: (2026) -
Value Profiles for Encoding Human Variation
by: Sorensen, Taylor, et al.
Published: (2025) -
Diffusion Explainer: Visual Explanation for Text-to-image Stable Diffusion
by: Lee, Seongmin, et al.
Published: (2023) -
Toward Cultural Interpretability: A Linguistic Anthropological Framework for Describing and Evaluating Large Language Models (LLMs)
by: Jones, Graham M., et al.
Published: (2024)