Evaluating the World Model Implicit in a Generative Model
Fuente:
arXiv
Saved in:
| Main Authors: | Vafa, Keyon, Chen, Justin Y., Rambachan, Ashesh, Kleinberg, Jon, Mullainathan, Sendhil |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do Large Language Models Perform the Way People Expect? Measuring the Human Generalization Function
by: Vafa, Keyon, et al.
Published: (2024)
by: Vafa, Keyon, et al.
Published: (2024)
What Has a Foundation Model Found? Using Inductive Bias to Probe for World Models
by: Vafa, Keyon, et al.
Published: (2025)
by: Vafa, Keyon, et al.
Published: (2025)
What's Producible May Not Be Reachable: Measuring the Steerability of Generative Models
by: Vafa, Keyon, et al.
Published: (2025)
by: Vafa, Keyon, et al.
Published: (2025)
Potemkin Understanding in Large Language Models
by: Mancoridis, Marina, et al.
Published: (2025)
by: Mancoridis, Marina, et al.
Published: (2025)
Language Generation in the Limit
by: Kleinberg, Jon, et al.
Published: (2024)
by: Kleinberg, Jon, et al.
Published: (2024)
Large Language Models: An Applied Econometric Framework
by: Ludwig, Jens, et al.
Published: (2024)
by: Ludwig, Jens, et al.
Published: (2024)
From Predictive Algorithms to Automatic Generation of Anomalies
by: Mullainathan, Sendhil, et al.
Published: (2024)
by: Mullainathan, Sendhil, et al.
Published: (2024)
Large Language Models in Cryptocurrency Securities Cases: Can a GPT Model Meaningfully Assist Lawyers?
by: Trozze, Arianna, et al.
Published: (2023)
by: Trozze, Arianna, et al.
Published: (2023)
How Many Features Can a Language Model Store Under the Linear Representation Hypothesis?
by: Garg, Nikhil, et al.
Published: (2026)
by: Garg, Nikhil, et al.
Published: (2026)
Sparse Autoencoders for Hypothesis Generation
by: Movva, Rajiv, et al.
Published: (2025)
by: Movva, Rajiv, et al.
Published: (2025)
MDIT-Bench: Evaluating the Dual-Implicit Toxicity in Large Multimodal Models
by: Jin, Bohan, et al.
Published: (2025)
by: Jin, Bohan, et al.
Published: (2025)
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
by: Wen, Yuchen, et al.
Published: (2024)
by: Wen, Yuchen, et al.
Published: (2024)
On Language Generation in the Limit with Bounded Memory
by: Kleinberg, Jon, et al.
Published: (2026)
by: Kleinberg, Jon, et al.
Published: (2026)
Critical Thinking: Which Kinds of Complexity Govern Optimal Reasoning Length?
by: Lee, Celine, et al.
Published: (2025)
by: Lee, Celine, et al.
Published: (2025)
Text2World: Benchmarking Large Language Models for Symbolic World Model Generation
by: Hu, Mengkang, et al.
Published: (2025)
by: Hu, Mengkang, et al.
Published: (2025)
SwissNYF: Tool Grounded LLM Agents for Black Box Setting
by: Kumar, Somnath Sendhil, et al.
Published: (2024)
by: Kumar, Somnath Sendhil, et al.
Published: (2024)
Effective faking of verbal deception detection with target-aligned adversarial attacks
by: Kleinberg, Bennett, et al.
Published: (2025)
by: Kleinberg, Bennett, et al.
Published: (2025)
Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model
by: Liu, Runheng, et al.
Published: (2026)
by: Liu, Runheng, et al.
Published: (2026)
Use Sparse Autoencoders to Discover Unknown Concepts, Not to Act on Known Concepts
by: Peng, Kenny, et al.
Published: (2025)
by: Peng, Kenny, et al.
Published: (2025)
Generative Emergent Communication: Large Language Model is a Collective World Model
by: Taniguchi, Tadahiro, et al.
Published: (2024)
by: Taniguchi, Tadahiro, et al.
Published: (2024)
iCLP: Large Language Model Reasoning with Implicit Cognition Latent Planning
by: Chen, Sijia, et al.
Published: (2025)
by: Chen, Sijia, et al.
Published: (2025)
ImplicitRM: Unbiased Reward Modeling from Implicit Preference Data for LLM alignment
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Multi-Domain ABSA Conversation Dataset Generation via LLMs for Real-World Evaluation and Model Comparison
by: Pandit, Tejul, et al.
Published: (2025)
by: Pandit, Tejul, et al.
Published: (2025)
ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues
by: Vedula, Bhaskara Hanuma, et al.
Published: (2026)
by: Vedula, Bhaskara Hanuma, et al.
Published: (2026)
Fine-Tuning Games: Bargaining and Adaptation for General-Purpose Models
by: Laufer, Benjamin, et al.
Published: (2023)
by: Laufer, Benjamin, et al.
Published: (2023)
Evaluating Large Language Models for Real-World Engineering Tasks
by: Heesch, Rene, et al.
Published: (2025)
by: Heesch, Rene, et al.
Published: (2025)
Word2World: Generating Stories and Worlds through Large Language Models
by: Nasir, Muhammad U., et al.
Published: (2024)
by: Nasir, Muhammad U., et al.
Published: (2024)
Contextual Feature Extraction Hierarchies Converge in Large Language Models and the Brain
by: Mischler, Gavin, et al.
Published: (2024)
by: Mischler, Gavin, et al.
Published: (2024)
ImF: Implicit Fingerprint for Large Language Models
by: Wu, Jiaxuan, et al.
Published: (2025)
by: Wu, Jiaxuan, et al.
Published: (2025)
Agent2World: Learning to Generate Symbolic World Models via Adaptive Multi-Agent Feedback
by: Hu, Mengkang, et al.
Published: (2025)
by: Hu, Mengkang, et al.
Published: (2025)
Grammatical Error Feedback: An Implicit Evaluation Approach
by: Bannò, Stefano, et al.
Published: (2024)
by: Bannò, Stefano, et al.
Published: (2024)
LABOR-LLM: Language-Based Occupational Representations with Large Language Models
by: Athey, Susan, et al.
Published: (2024)
by: Athey, Susan, et al.
Published: (2024)
Evaluating Clinical Competencies of Large Language Models with a General Practice Benchmark
by: Li, Zheqing, et al.
Published: (2025)
by: Li, Zheqing, et al.
Published: (2025)
Comparing Human and Large Language Model Interpretation of Implicit Information
by: De Santis, Antonio, et al.
Published: (2026)
by: De Santis, Antonio, et al.
Published: (2026)
Implicit Reasoning in Large Language Models: A Comprehensive Survey
by: Li, Jindong, et al.
Published: (2025)
by: Li, Jindong, et al.
Published: (2025)
Quantifying Label-Induced Bias in Large Language Model Self- and Cross-Evaluations
by: Saraf, Muskan, et al.
Published: (2025)
by: Saraf, Muskan, et al.
Published: (2025)
SimpleStrat: Diversifying Language Model Generation with Stratification
by: Wong, Justin, et al.
Published: (2024)
by: Wong, Justin, et al.
Published: (2024)
ConsintBench: Evaluating Language Models on Real-World Consumer Intent Understanding
by: Li, Xiaozhe, et al.
Published: (2025)
by: Li, Xiaozhe, et al.
Published: (2025)
Beyond the Mean: Within-Model Reliable Change Detection for LLM Evaluation
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Promises, Outlooks and Challenges of Diffusion Language Modeling
by: Deschenaux, Justin, et al.
Published: (2024)
by: Deschenaux, Justin, et al.
Published: (2024)
Similar Items
-
Do Large Language Models Perform the Way People Expect? Measuring the Human Generalization Function
by: Vafa, Keyon, et al.
Published: (2024) -
What Has a Foundation Model Found? Using Inductive Bias to Probe for World Models
by: Vafa, Keyon, et al.
Published: (2025) -
What's Producible May Not Be Reachable: Measuring the Steerability of Generative Models
by: Vafa, Keyon, et al.
Published: (2025) -
Potemkin Understanding in Large Language Models
by: Mancoridis, Marina, et al.
Published: (2025) -
Language Generation in the Limit
by: Kleinberg, Jon, et al.
Published: (2024)