Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
Fuente:
arXiv
Guardado en:
| Autores principales: | Qiu, Linlu, Jiang, Liwei, Lu, Ximing, Sclar, Melanie, Pyatkin, Valentina, Bhagavatula, Chandra, Wang, Bailin, Kim, Yoon, Choi, Yejin, Dziri, Nouha, Ren, Xiang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
por: Sorensen, Taylor, et al.
Publicado: (2023)
por: Sorensen, Taylor, et al.
Publicado: (2023)
TurnWise: The Gap between Single- and Multi-turn Language Model Capabilities
por: Graf, Victoria, et al.
Publicado: (2026)
por: Graf, Victoria, et al.
Publicado: (2026)
What Makes it Ok to Set a Fire? Iterative Self-distillation of Contexts and Rationales for Disambiguating Defeasible Social and Moral Situations
por: Rao, Kavel, et al.
Publicado: (2023)
por: Rao, Kavel, et al.
Publicado: (2023)
AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
por: Lu, Ximing, et al.
Publicado: (2024)
por: Lu, Ximing, et al.
Publicado: (2024)
WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild
por: Lin, Bill Yuchen, et al.
Publicado: (2024)
por: Lin, Bill Yuchen, et al.
Publicado: (2024)
Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models
por: Kim, Hyunwoo, et al.
Publicado: (2025)
por: Kim, Hyunwoo, et al.
Publicado: (2025)
Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
por: Sclar, Melanie, et al.
Publicado: (2023)
por: Sclar, Melanie, et al.
Publicado: (2023)
SafetyAnalyst: Interpretable, Transparent, and Steerable Safety Moderation for AI Behavior
por: Li, Jing-Jing, et al.
Publicado: (2024)
por: Li, Jing-Jing, et al.
Publicado: (2024)
Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks
por: Wu, Zhaofeng, et al.
Publicado: (2023)
por: Wu, Zhaofeng, et al.
Publicado: (2023)
WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
por: Jiang, Liwei, et al.
Publicado: (2024)
por: Jiang, Liwei, et al.
Publicado: (2024)
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
por: Han, Seungju, et al.
Publicado: (2024)
por: Han, Seungju, et al.
Publicado: (2024)
Multi-Attribute Constraint Satisfaction via Language Model Rewriting
por: Baheti, Ashutosh, et al.
Publicado: (2024)
por: Baheti, Ashutosh, et al.
Publicado: (2024)
A Roadmap to Pluralistic Alignment
por: Sorensen, Taylor, et al.
Publicado: (2024)
por: Sorensen, Taylor, et al.
Publicado: (2024)
RewardBench: Evaluating Reward Models for Language Modeling
por: Lambert, Nathan, et al.
Publicado: (2024)
por: Lambert, Nathan, et al.
Publicado: (2024)
PlaSma: Making Small Language Models Better Procedural Knowledge Models for (Counterfactual) Planning
por: Brahman, Faeze, et al.
Publicado: (2023)
por: Brahman, Faeze, et al.
Publicado: (2023)
Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)
por: Jiang, Liwei, et al.
Publicado: (2025)
por: Jiang, Liwei, et al.
Publicado: (2025)
Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
por: Ravichander, Abhilasha, et al.
Publicado: (2025)
por: Ravichander, Abhilasha, et al.
Publicado: (2025)
Finding Flawed Fictions: Evaluating Complex Reasoning in Language Models via Plot Hole Detection
por: Ahuja, Kabir, et al.
Publicado: (2025)
por: Ahuja, Kabir, et al.
Publicado: (2025)
The Art of Saying No: Contextual Noncompliance in Language Models
por: Brahman, Faeze, et al.
Publicado: (2024)
por: Brahman, Faeze, et al.
Publicado: (2024)
CULTURE-GEN: Revealing Global Cultural Perception in Language Models through Natural Language Prompting
por: Li, Huihan, et al.
Publicado: (2024)
por: Li, Huihan, et al.
Publicado: (2024)
Surfacing Semantic Orthogonality Across Model Safety Benchmarks: A Multi-Dimensional Analysis
por: Bennion, Jonathan, et al.
Publicado: (2025)
por: Bennion, Jonathan, et al.
Publicado: (2025)
The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage
por: Hallinan, Skyler, et al.
Publicado: (2025)
por: Hallinan, Skyler, et al.
Publicado: (2025)
Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions
por: Acuna, David, et al.
Publicado: (2025)
por: Acuna, David, et al.
Publicado: (2025)
Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning
por: Sclar, Melanie, et al.
Publicado: (2024)
por: Sclar, Melanie, et al.
Publicado: (2024)
JAMDEC: Unsupervised Authorship Obfuscation using Constrained Decoding over Small Language Models
por: Fisher, Jillian, et al.
Publicado: (2024)
por: Fisher, Jillian, et al.
Publicado: (2024)
Bias Amplification in Language Model Evolution: An Iterated Learning Perspective
por: Ren, Yi, et al.
Publicado: (2024)
por: Ren, Yi, et al.
Publicado: (2024)
Can Language Models Reason about Individualistic Human Values and Preferences?
por: Jiang, Liwei, et al.
Publicado: (2024)
por: Jiang, Liwei, et al.
Publicado: (2024)
Hypothesis Search: Inductive Reasoning with Language Models
por: Wang, Ruocheng, et al.
Publicado: (2023)
por: Wang, Ruocheng, et al.
Publicado: (2023)
THE HERESY OF JACOB FRANK: FROM JEWISH MESSIANISM TO ESOTERIC MYTH. By JayMichaelson. Oxford: Oxford University Press, 2022. Pp. 272. Cloth, $120; Paper, $34.99.
por: David Sclar
Publicado: (2024)
por: David Sclar
Publicado: (2024)
FIRST IMPRESSIONS: SEFER HASIDIM AND EARLY MODERN HEBREW PRINTING. By Joseph A.Skloot. Waltham, MA: Brandeis University Press, 2023. Pp. 268. Cloth, $130; paper, $40.
por: David Sclar
Publicado: (2024)
por: David Sclar
Publicado: (2024)
The Complex Brain Hypothesis: Resolving the Entropy-Content Conundrum in Minimal Phenomenal Experience
por: Mago, Jonas, et al.
Publicado: (2026)
por: Mago, Jonas, et al.
Publicado: (2026)
Climbing the Ladder of Reasoning: What LLMs Can-and Still Can't-Solve after SFT?
por: Sun, Yiyou, et al.
Publicado: (2025)
por: Sun, Yiyou, et al.
Publicado: (2025)
OMEGA: Can LLMs Reason Outside the Box in Math? Evaluating Exploratory, Compositional, and Transformative Generalization
por: Sun, Yiyou, et al.
Publicado: (2025)
por: Sun, Yiyou, et al.
Publicado: (2025)
Bayesian Teaching Enables Probabilistic Reasoning in Large Language Models
por: Qiu, Linlu, et al.
Publicado: (2025)
por: Qiu, Linlu, et al.
Publicado: (2025)
On the Same Wavelength? Evaluating Pragmatic Reasoning in Language Models across Broad Concepts
por: Qiu, Linlu, et al.
Publicado: (2025)
por: Qiu, Linlu, et al.
Publicado: (2025)
Information-Theoretic Distillation for Reference-less Summarization
por: Jung, Jaehun, et al.
Publicado: (2024)
por: Jung, Jaehun, et al.
Publicado: (2024)
DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life
por: Chiu, Yu Ying, et al.
Publicado: (2024)
por: Chiu, Yu Ying, et al.
Publicado: (2024)
SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning
por: Limozin, Alexis, et al.
Publicado: (2026)
por: Limozin, Alexis, et al.
Publicado: (2026)
Promptly Predicting Structures: The Return of Inference
por: Mehta, Maitrey, et al.
Publicado: (2024)
por: Mehta, Maitrey, et al.
Publicado: (2024)
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing
por: Jung, Jaehun, et al.
Publicado: (2023)
por: Jung, Jaehun, et al.
Publicado: (2023)
Ejemplares similares
-
Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
por: Sorensen, Taylor, et al.
Publicado: (2023) -
TurnWise: The Gap between Single- and Multi-turn Language Model Capabilities
por: Graf, Victoria, et al.
Publicado: (2026) -
What Makes it Ok to Set a Fire? Iterative Self-distillation of Contexts and Rationales for Disambiguating Defeasible Social and Moral Situations
por: Rao, Kavel, et al.
Publicado: (2023) -
AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
por: Lu, Ximing, et al.
Publicado: (2024) -
WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild
por: Lin, Bill Yuchen, et al.
Publicado: (2024)