KnowledgeVIS: Interpreting Language Models by Comparing Fill-in-the-Blank Prompts
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Coscia, Adam, Endert, Alex |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
iScore: Visual Analytics for Interpreting How Language Models Automatically Score Summaries
par: Coscia, Adam, et autres
Publié: (2024)
par: Coscia, Adam, et autres
Publié: (2024)
OnGoal: Tracking and Visualizing Conversational Goals in Multi-Turn Dialogue with Large Language Models
par: Coscia, Adam, et autres
Publié: (2025)
par: Coscia, Adam, et autres
Publié: (2025)
VisPile: A Visual Analytics System for Analyzing Multiple Text Documents With Large Language Models and Knowledge Graphs
par: Coscia, Adam, et autres
Publié: (2025)
par: Coscia, Adam, et autres
Publié: (2025)
VAE Explainer: Supplement Learning Variational Autoencoders with Interactive Visualization
par: Bertucci, Donald, et autres
Publié: (2024)
par: Bertucci, Donald, et autres
Publié: (2024)
Fair Knowledge Tracing in Second Language Acquisition
par: Tang, Weitao, et autres
Publié: (2024)
par: Tang, Weitao, et autres
Publié: (2024)
When "A Helpful Assistant" Is Not Really Helpful: Personas in System Prompts Do Not Improve Performances of Large Language Models
par: Zheng, Mingqian, et autres
Publié: (2023)
par: Zheng, Mingqian, et autres
Publié: (2023)
AI Meets the Classroom: When Do Large Language Models Harm Learning?
par: Lehmann, Matthias, et autres
Publié: (2024)
par: Lehmann, Matthias, et autres
Publié: (2024)
Is it Still Fair? A Comparative Evaluation of Fairness Algorithms through the Lens of Covariate Drift
par: Deho, Oscar Blessed, et autres
Publié: (2024)
par: Deho, Oscar Blessed, et autres
Publié: (2024)
Sociodemographic Prompting is Not Yet an Effective Approach for Simulating Subjective Judgments with LLMs
par: Sun, Huaman, et autres
Publié: (2023)
par: Sun, Huaman, et autres
Publié: (2023)
Prediction-Powered Ranking of Large Language Models
par: Chatzi, Ivi, et autres
Publié: (2024)
par: Chatzi, Ivi, et autres
Publié: (2024)
A Nested Model for AI Design and Validation
par: Dubey, Akshat, et autres
Publié: (2024)
par: Dubey, Akshat, et autres
Publié: (2024)
Implicit Personalization in Language Models: A Systematic Study
par: Jin, Zhijing, et autres
Publié: (2024)
par: Jin, Zhijing, et autres
Publié: (2024)
Ethics and Technical Aspects of Generative AI Models in Digital Content Creation
par: Karagoz, Atahan
Publié: (2024)
par: Karagoz, Atahan
Publié: (2024)
The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested
par: Vishwarupe, Varad, et autres
Publié: (2026)
par: Vishwarupe, Varad, et autres
Publié: (2026)
Linear Representations of Political Perspective Emerge in Large Language Models
par: Kim, Junsol, et autres
Publié: (2025)
par: Kim, Junsol, et autres
Publié: (2025)
Corn Yield Prediction Model with Deep Neural Networks for Smallholder Farmer Decision Support System
par: Olisah, Chollette C., et autres
Publié: (2024)
par: Olisah, Chollette C., et autres
Publié: (2024)
AI's Regimes of Representation: A Community-centered Study of Text-to-Image Models in South Asia
par: Qadri, Rida, et autres
Publié: (2023)
par: Qadri, Rida, et autres
Publié: (2023)
Laboratory-Scale AI: Open-Weight Models are Competitive with ChatGPT Even in Low-Resource Settings
par: Wolfe, Robert, et autres
Publié: (2024)
par: Wolfe, Robert, et autres
Publié: (2024)
Can Large Language Models Unlock Novel Scientific Research Ideas?
par: Kumar, Sandeep, et autres
Publié: (2024)
par: Kumar, Sandeep, et autres
Publié: (2024)
Evaluating Multimodal Language Models as Visual Assistants for Visually Impaired Users
par: Karamolegkou, Antonia, et autres
Publié: (2025)
par: Karamolegkou, Antonia, et autres
Publié: (2025)
LLM4PM: A case study on using Large Language Models for Process Modeling in Enterprise Organizations
par: Ziche, Clara, et autres
Publié: (2024)
par: Ziche, Clara, et autres
Publié: (2024)
Computational Approaches to Understanding Large Language Model Impact on Writing and Information Ecosystems
par: Liang, Weixin
Publié: (2025)
par: Liang, Weixin
Publié: (2025)
Large Language Models Can Infer Personality from Free-Form User Interactions
par: Peters, Heinrich, et autres
Publié: (2024)
par: Peters, Heinrich, et autres
Publié: (2024)
MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes
par: Chiu, Yu Ying, et autres
Publié: (2025)
par: Chiu, Yu Ying, et autres
Publié: (2025)
Preliminary Guidelines For Combining Data Integration and Visual Data Analysis
par: Coscia, Adam, et autres
Publié: (2024)
par: Coscia, Adam, et autres
Publié: (2024)
The Lock-in Hypothesis: Stagnation by Algorithm
par: Qiu, Tianyi Alex, et autres
Publié: (2025)
par: Qiu, Tianyi Alex, et autres
Publié: (2025)
"Filling the Blanks'': Identifying Micro-activities that Compose Complex Human Activities of Daily Living
par: Chatterjee, Soumyajit, et autres
Publié: (2023)
par: Chatterjee, Soumyajit, et autres
Publié: (2023)
Why am I Still Seeing This: Measuring the Effectiveness Of Ad Controls and Explanations in AI-Mediated Ad Targeting Systems
par: Castleman, Jane, et autres
Publié: (2024)
par: Castleman, Jane, et autres
Publié: (2024)
PsyDI: Towards a Personalized and Progressively In-depth Chatbot for Psychological Measurements
par: Li, Xueyan, et autres
Publié: (2024)
par: Li, Xueyan, et autres
Publié: (2024)
The Digital Transformation in Health: How AI Can Improve the Performance of Health Systems
par: Periáñez, África, et autres
Publié: (2024)
par: Periáñez, África, et autres
Publié: (2024)
M3BAT: Unsupervised Domain Adaptation for Multimodal Mobile Sensing with Multi-Branch Adversarial Training
par: Meegahapola, Lakmal, et autres
Publié: (2024)
par: Meegahapola, Lakmal, et autres
Publié: (2024)
A Comprehensive Survey and Classification of Evaluation Criteria for Trustworthy Artificial Intelligence
par: McCormack, Louise, et autres
Publié: (2024)
par: McCormack, Louise, et autres
Publié: (2024)
Automated Assessment of Encouragement and Warmth in Classrooms Leveraging Multimodal Emotional Features and ChatGPT
par: Hou, Ruikun, et autres
Publié: (2024)
par: Hou, Ruikun, et autres
Publié: (2024)
Generative AI in Medicine
par: Shanmugam, Divya, et autres
Publié: (2024)
par: Shanmugam, Divya, et autres
Publié: (2024)
Creative Loss: Ambiguity, Uncertainty and Indeterminacy
par: Holberton, Tom
Publié: (2024)
par: Holberton, Tom
Publié: (2024)
Human-in-the-Loop AI for Cheating Ring Detection
par: Shih, Yong-Siang, et autres
Publié: (2024)
par: Shih, Yong-Siang, et autres
Publié: (2024)
Farsight: Fostering Responsible AI Awareness During AI Application Prototyping
par: Wang, Zijie J., et autres
Publié: (2024)
par: Wang, Zijie J., et autres
Publié: (2024)
Hybrid Forecasting of Geopolitical Events
par: Benjamin, Daniel M., et autres
Publié: (2024)
par: Benjamin, Daniel M., et autres
Publié: (2024)
Learning to Assist Humans without Inferring Rewards
par: Myers, Vivek, et autres
Publié: (2024)
par: Myers, Vivek, et autres
Publié: (2024)
Participation in the age of foundation models
par: Suresh, Harini, et autres
Publié: (2024)
par: Suresh, Harini, et autres
Publié: (2024)
Documents similaires
-
iScore: Visual Analytics for Interpreting How Language Models Automatically Score Summaries
par: Coscia, Adam, et autres
Publié: (2024) -
OnGoal: Tracking and Visualizing Conversational Goals in Multi-Turn Dialogue with Large Language Models
par: Coscia, Adam, et autres
Publié: (2025) -
VisPile: A Visual Analytics System for Analyzing Multiple Text Documents With Large Language Models and Knowledge Graphs
par: Coscia, Adam, et autres
Publié: (2025) -
VAE Explainer: Supplement Learning Variational Autoencoders with Interactive Visualization
par: Bertucci, Donald, et autres
Publié: (2024) -
Fair Knowledge Tracing in Second Language Acquisition
par: Tang, Weitao, et autres
Publié: (2024)