VibeCheck: Discover and Quantify Qualitative Differences in Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Dunlap, Lisa, Mandal, Krishna, Darrell, Trevor, Steinhardt, Jacob, Gonzalez, Joseph E |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning a Generative Meta-Model of LLM Activations
di: Luo, Grace, et al.
Pubblicazione: (2026)
di: Luo, Grace, et al.
Pubblicazione: (2026)
Discovering Latent Knowledge in Language Models Without Supervision
di: Burns, Collin, et al.
Pubblicazione: (2022)
di: Burns, Collin, et al.
Pubblicazione: (2022)
Describing Differences in Image Sets with Natural Language
di: Dunlap, Lisa, et al.
Pubblicazione: (2023)
di: Dunlap, Lisa, et al.
Pubblicazione: (2023)
How do Language Models Bind Entities in Context?
di: Feng, Jiahai, et al.
Pubblicazione: (2023)
di: Feng, Jiahai, et al.
Pubblicazione: (2023)
ClaimCheck: Real-Time Fact-Checking with Small Language Models
di: Putta, Akshith Reddy, et al.
Pubblicazione: (2025)
di: Putta, Akshith Reddy, et al.
Pubblicazione: (2025)
Overthinking the Truth: Understanding how Language Models Process False Demonstrations
di: Halawi, Danny, et al.
Pubblicazione: (2023)
di: Halawi, Danny, et al.
Pubblicazione: (2023)
Recon: Reconstruction-Guided Reasoning Synthesis for User Modeling
di: Zhu, Alan, et al.
Pubblicazione: (2026)
di: Zhu, Alan, et al.
Pubblicazione: (2026)
Feedback Loops With Language Models Drive In-Context Reward Hacking
di: Pan, Alexander, et al.
Pubblicazione: (2024)
di: Pan, Alexander, et al.
Pubblicazione: (2024)
Explaining Datasets in Words: Statistical Models with Natural Language Parameters
di: Zhong, Ruiqi, et al.
Pubblicazione: (2024)
di: Zhong, Ruiqi, et al.
Pubblicazione: (2024)
Which Attention Heads Matter for In-Context Learning?
di: Yin, Kayo, et al.
Pubblicazione: (2025)
di: Yin, Kayo, et al.
Pubblicazione: (2025)
Language Model Circuits Are Sparse in the Neuron Basis
di: Arora, Aryaman, et al.
Pubblicazione: (2026)
di: Arora, Aryaman, et al.
Pubblicazione: (2026)
Training Language Models to Explain Their Own Computations
di: Li, Belinda Z., et al.
Pubblicazione: (2025)
di: Li, Belinda Z., et al.
Pubblicazione: (2025)
Eliciting Language Model Behaviors with Investigator Agents
di: Li, Xiang Lisa, et al.
Pubblicazione: (2025)
di: Li, Xiang Lisa, et al.
Pubblicazione: (2025)
Compositional Chain-of-Thought Prompting for Large Multimodal Models
di: Mitra, Chancharik, et al.
Pubblicazione: (2023)
di: Mitra, Chancharik, et al.
Pubblicazione: (2023)
Approaching Human-Level Forecasting with Language Models
di: Halawi, Danny, et al.
Pubblicazione: (2024)
di: Halawi, Danny, et al.
Pubblicazione: (2024)
Multimodal Large Language Models to Support Real-World Fact-Checking
di: Geng, Jiahui, et al.
Pubblicazione: (2024)
di: Geng, Jiahui, et al.
Pubblicazione: (2024)
Learning to Check: Unleashing Potentials for Self-Correction in Large Language Models
di: Zhang, Che, et al.
Pubblicazione: (2024)
di: Zhang, Che, et al.
Pubblicazione: (2024)
Fact-Checking with Large Language Models via Probabilistic Certainty and Consistency
di: Wang, Haoran, et al.
Pubblicazione: (2026)
di: Wang, Haoran, et al.
Pubblicazione: (2026)
Self-Discover: Large Language Models Self-Compose Reasoning Structures
di: Zhou, Pei, et al.
Pubblicazione: (2024)
di: Zhou, Pei, et al.
Pubblicazione: (2024)
Iterative Label Refinement Matters More than Preference Optimization under Weak Supervision
di: Ye, Yaowen, et al.
Pubblicazione: (2025)
di: Ye, Yaowen, et al.
Pubblicazione: (2025)
The Vibe-Check Protocol: Quantifying Cognitive Offloading in AI Programming
di: Aiersilan, Aizierjiang
Pubblicazione: (2026)
di: Aiersilan, Aizierjiang
Pubblicazione: (2026)
Learning Adaptive Parallel Reasoning with Language Models
di: Pan, Jiayi, et al.
Pubblicazione: (2025)
di: Pan, Jiayi, et al.
Pubblicazione: (2025)
Atomic Fact-Checking Increases Clinician Trust in Large Language Model Recommendations for Oncology Decision Support: A Randomized Controlled Trial
di: Adams, Lisa C., et al.
Pubblicazione: (2026)
di: Adams, Lisa C., et al.
Pubblicazione: (2026)
Vibe Check: Understanding the Effects of LLM-Based Conversational Agents' Personality and Alignment on User Perceptions in Goal-Oriented Tasks
di: Rahman, Hasibur, et al.
Pubblicazione: (2025)
di: Rahman, Hasibur, et al.
Pubblicazione: (2025)
Neuron Empirical Gradient: Discovering and Quantifying Neurons Global Linear Controllability
di: Zhao, Xin, et al.
Pubblicazione: (2024)
di: Zhao, Xin, et al.
Pubblicazione: (2024)
LEAF: Learning and Evaluation Augmented by Fact-Checking to Improve Factualness in Large Language Models
di: Tran, Hieu, et al.
Pubblicazione: (2024)
di: Tran, Hieu, et al.
Pubblicazione: (2024)
TrumorGPT: Graph-Based Retrieval-Augmented Large Language Model for Fact-Checking
di: Hang, Ching Nam, et al.
Pubblicazione: (2025)
di: Hang, Ching Nam, et al.
Pubblicazione: (2025)
DOCS: Quantifying Weight Similarity for Deeper Insights into Large Language Models
di: Min, Zeping, et al.
Pubblicazione: (2025)
di: Min, Zeping, et al.
Pubblicazione: (2025)
"Lost-in-the-Later": Framework for Quantifying Contextual Grounding in Large Language Models
di: Tao, Yufei, et al.
Pubblicazione: (2025)
di: Tao, Yufei, et al.
Pubblicazione: (2025)
Measuring Representation Robustness in Large Language Models for Geometry
di: Jawandhia, Vedant, et al.
Pubblicazione: (2026)
di: Jawandhia, Vedant, et al.
Pubblicazione: (2026)
TurQUaz at CheckThat! 2025: Debating Large Language Models for Scientific Web Discourse Detection
di: Saraç, Tarık, et al.
Pubblicazione: (2025)
di: Saraç, Tarık, et al.
Pubblicazione: (2025)
MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts
di: He, Jiayi, et al.
Pubblicazione: (2025)
di: He, Jiayi, et al.
Pubblicazione: (2025)
Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B
di: Xu, Sen, et al.
Pubblicazione: (2025)
di: Xu, Sen, et al.
Pubblicazione: (2025)
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features
di: Mitra, Chancharik, et al.
Pubblicazione: (2024)
di: Mitra, Chancharik, et al.
Pubblicazione: (2024)
Quantifying Label-Induced Bias in Large Language Model Self- and Cross-Evaluations
di: Saraf, Muskan, et al.
Pubblicazione: (2025)
di: Saraf, Muskan, et al.
Pubblicazione: (2025)
FRoG: Evaluating Fuzzy Reasoning of Generalized Quantifiers in Large Language Models
di: Li, Yiyuan, et al.
Pubblicazione: (2024)
di: Li, Yiyuan, et al.
Pubblicazione: (2024)
HYBRINFOX at CheckThat! 2024 -- Task 1: Enhancing Language Models with Structured Information for Check-Worthiness Estimation
di: Faye, Géraud, et al.
Pubblicazione: (2024)
di: Faye, Géraud, et al.
Pubblicazione: (2024)
Meta-Reasoning Improves Tool Use in Large Language Models
di: Alazraki, Lisa, et al.
Pubblicazione: (2024)
di: Alazraki, Lisa, et al.
Pubblicazione: (2024)
RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking
di: Yang, Shuo, et al.
Pubblicazione: (2025)
di: Yang, Shuo, et al.
Pubblicazione: (2025)
ALOHa: A New Measure for Hallucination in Captioning Models
di: Petryk, Suzanne, et al.
Pubblicazione: (2024)
di: Petryk, Suzanne, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Learning a Generative Meta-Model of LLM Activations
di: Luo, Grace, et al.
Pubblicazione: (2026) -
Discovering Latent Knowledge in Language Models Without Supervision
di: Burns, Collin, et al.
Pubblicazione: (2022) -
Describing Differences in Image Sets with Natural Language
di: Dunlap, Lisa, et al.
Pubblicazione: (2023) -
How do Language Models Bind Entities in Context?
di: Feng, Jiahai, et al.
Pubblicazione: (2023) -
ClaimCheck: Real-Time Fact-Checking with Small Language Models
di: Putta, Akshith Reddy, et al.
Pubblicazione: (2025)