Influence of Solution Efficiency and Valence of Instruction on Additive and Subtractive Solution Strategies in Humans and GPT-4
Fuente:
arXiv
Guardado en:
| Autores principales: | Uhler, Lydia, Jordan, Verena, Buder, Jürgen, Huff, Markus, Papenmeier, Frank |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards a Psychology of Machines: Large Language Models Predict Human Memory
por: Huff, Markus, et al.
Publicado: (2024)
por: Huff, Markus, et al.
Publicado: (2024)
Replicating Human Social Perception in Generative AI: Evaluating the Valence-Dominance Model
por: Gurkan, Necdet, et al.
Publicado: (2025)
por: Gurkan, Necdet, et al.
Publicado: (2025)
ChatGPT Alternative Solutions: Large Language Models Survey
por: Alipour, Hanieh, et al.
Publicado: (2024)
por: Alipour, Hanieh, et al.
Publicado: (2024)
ArchiveGPT: A human-centered evaluation of using a vision language model for image cataloguing
por: Abele, Line, et al.
Publicado: (2025)
por: Abele, Line, et al.
Publicado: (2025)
Evaluation of GPT-4o and GPT-4o-mini's Vision Capabilities for Compositional Analysis from Dried Solution Drops
por: Dangi, Deven B., et al.
Publicado: (2024)
por: Dangi, Deven B., et al.
Publicado: (2024)
Exploring the features used for summary evaluation by Human and GPT
por: Sadeghi, Zahra, et al.
Publicado: (2025)
por: Sadeghi, Zahra, et al.
Publicado: (2025)
GPTEval: A Survey on Assessments of ChatGPT and GPT-4
por: Mao, Rui, et al.
Publicado: (2023)
por: Mao, Rui, et al.
Publicado: (2023)
GraphGPT: Graph Instruction Tuning for Large Language Models
por: Tang, Jiabin, et al.
Publicado: (2023)
por: Tang, Jiabin, et al.
Publicado: (2023)
Thinking by Subtraction: Confidence-Driven Contrastive Decoding for LLM Reasoning
por: Tang, Lexiang, et al.
Publicado: (2026)
por: Tang, Lexiang, et al.
Publicado: (2026)
Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4
por: Bsharat, Sondos Mahmoud, et al.
Publicado: (2023)
por: Bsharat, Sondos Mahmoud, et al.
Publicado: (2023)
DeepSolution: Boosting Complex Engineering Solution Design via Tree-based Exploration and Bi-point Thinking
por: Li, Zhuoqun, et al.
Publicado: (2025)
por: Li, Zhuoqun, et al.
Publicado: (2025)
Strategy-Induct: Task-Level Strategy Induction for Instruction Generation
por: Chen, Po-Chun, et al.
Publicado: (2026)
por: Chen, Po-Chun, et al.
Publicado: (2026)
The Human and the Mechanical: logos, truthfulness, and ChatGPT
por: Giannakidou, Anastasia, et al.
Publicado: (2024)
por: Giannakidou, Anastasia, et al.
Publicado: (2024)
Is GPT-4 a reliable rater? Evaluating Consistency in GPT-4 Text Ratings
por: Hackl, Veronika, et al.
Publicado: (2023)
por: Hackl, Veronika, et al.
Publicado: (2023)
GPT-4 Technical Report
por: OpenAI, et al.
Publicado: (2023)
por: OpenAI, et al.
Publicado: (2023)
Measuring Pragmatic Influence in Large Language Model Instructions
por: Geng, Yilin, et al.
Publicado: (2026)
por: Geng, Yilin, et al.
Publicado: (2026)
Finding your MUSE: Mining Unexpected Solutions Engine
por: Sweed, Nir, et al.
Publicado: (2025)
por: Sweed, Nir, et al.
Publicado: (2025)
Reasons and Solutions for the Decline in Model Performance after Editing
por: Huang, Xiusheng, et al.
Publicado: (2024)
por: Huang, Xiusheng, et al.
Publicado: (2024)
GenKnowSub: Improving Modularity and Reusability of LLMs through General Knowledge Subtraction
por: Bagherifard, Mohammadtaha, et al.
Publicado: (2025)
por: Bagherifard, Mohammadtaha, et al.
Publicado: (2025)
Benchmarking GPT-4 on Algorithmic Problems: A Systematic Evaluation of Prompting Strategies
por: Petruzzellis, Flavio, et al.
Publicado: (2024)
por: Petruzzellis, Flavio, et al.
Publicado: (2024)
Benchmarking GPT-4 against Human Translators: A Comprehensive Evaluation Across Languages, Domains, and Expertise Levels
por: Yan, Jianhao, et al.
Publicado: (2024)
por: Yan, Jianhao, et al.
Publicado: (2024)
Chasing Random: Instruction Selection Strategies Fail to Generalize
por: Diddee, Harshita, et al.
Publicado: (2024)
por: Diddee, Harshita, et al.
Publicado: (2024)
Differentiate ChatGPT-generated and Human-written Medical Texts
por: Liao, Wenxiong, et al.
Publicado: (2023)
por: Liao, Wenxiong, et al.
Publicado: (2023)
Assessing the Creativity of LLMs in Proposing Novel Solutions to Mathematical Problems
por: Ye, Junyi, et al.
Publicado: (2024)
por: Ye, Junyi, et al.
Publicado: (2024)
CEC-Zero: Chinese Error Correction Solution Based on LLM
por: Zhang, Sophie, et al.
Publicado: (2025)
por: Zhang, Sophie, et al.
Publicado: (2025)
LLM-First Search: Self-Guided Exploration of the Solution Space
por: Herr, Nathan, et al.
Publicado: (2025)
por: Herr, Nathan, et al.
Publicado: (2025)
Towards Robust Universal Information Extraction: Benchmark, Evaluation, and Solution
por: Zhu, Jizhao, et al.
Publicado: (2025)
por: Zhu, Jizhao, et al.
Publicado: (2025)
Comparative Analysis of ChatGPT, GPT-4, and Microsoft Bing Chatbots for GRE Test
por: Abu-Haifa, Mohammad, et al.
Publicado: (2023)
por: Abu-Haifa, Mohammad, et al.
Publicado: (2023)
Reasoning: From Reflection to Solution
por: Li, Zixi
Publicado: (2025)
por: Li, Zixi
Publicado: (2025)
Inductive-Deductive Strategy Reuse for Multi-Turn Instructional Dialogues
por: Ou, Jiao, et al.
Publicado: (2024)
por: Ou, Jiao, et al.
Publicado: (2024)
Does GPT-4 pass the Turing test?
por: Jones, Cameron R., et al.
Publicado: (2023)
por: Jones, Cameron R., et al.
Publicado: (2023)
Advancing AI with Integrity: Ethical Challenges and Solutions in Neural Machine Translation
por: Kimera, Richard, et al.
Publicado: (2024)
por: Kimera, Richard, et al.
Publicado: (2024)
Is Fine-Tuning an Effective Solution? Reassessing Knowledge Editing for Unstructured Data
por: Xiong, Hao, et al.
Publicado: (2025)
por: Xiong, Hao, et al.
Publicado: (2025)
A Rule Based Solution to Co-reference Resolution in Clinical Text
por: Chen, Ping, et al.
Publicado: (2025)
por: Chen, Ping, et al.
Publicado: (2025)
Using Combinatorial Optimization to Design a High quality LLM Solution
por: Ackerman, Samuel, et al.
Publicado: (2024)
por: Ackerman, Samuel, et al.
Publicado: (2024)
B4: Towards Optimal Assessment of Plausible Code Solutions with Plausible Tests
por: Chen, Mouxiang, et al.
Publicado: (2024)
por: Chen, Mouxiang, et al.
Publicado: (2024)
Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
por: Li, Hang, et al.
Publicado: (2025)
por: Li, Hang, et al.
Publicado: (2025)
A Looming Replication Crisis in Evaluating Behavior in Language Models? Evidence and Solutions
por: Vaugrante, Laurène, et al.
Publicado: (2024)
por: Vaugrante, Laurène, et al.
Publicado: (2024)
Building another Spanish dictionary, this time with GPT-4
por: Ortega-Martín, Miguel, et al.
Publicado: (2024)
por: Ortega-Martín, Miguel, et al.
Publicado: (2024)
Human-Instruction-Free LLM Self-Alignment with Limited Samples
por: Guo, Hongyi, et al.
Publicado: (2024)
por: Guo, Hongyi, et al.
Publicado: (2024)
Ejemplares similares
-
Towards a Psychology of Machines: Large Language Models Predict Human Memory
por: Huff, Markus, et al.
Publicado: (2024) -
Replicating Human Social Perception in Generative AI: Evaluating the Valence-Dominance Model
por: Gurkan, Necdet, et al.
Publicado: (2025) -
ChatGPT Alternative Solutions: Large Language Models Survey
por: Alipour, Hanieh, et al.
Publicado: (2024) -
ArchiveGPT: A human-centered evaluation of using a vision language model for image cataloguing
por: Abele, Line, et al.
Publicado: (2025) -
Evaluation of GPT-4o and GPT-4o-mini's Vision Capabilities for Compositional Analysis from Dried Solution Drops
por: Dangi, Deven B., et al.
Publicado: (2024)