Influence of Solution Efficiency and Valence of Instruction on Additive and Subtractive Solution Strategies in Humans and GPT-4
Fuente:
arXiv
Salvato in:
| Autori principali: | Uhler, Lydia, Jordan, Verena, Buder, Jürgen, Huff, Markus, Papenmeier, Frank |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards a Psychology of Machines: Large Language Models Predict Human Memory
di: Huff, Markus, et al.
Pubblicazione: (2024)
di: Huff, Markus, et al.
Pubblicazione: (2024)
Replicating Human Social Perception in Generative AI: Evaluating the Valence-Dominance Model
di: Gurkan, Necdet, et al.
Pubblicazione: (2025)
di: Gurkan, Necdet, et al.
Pubblicazione: (2025)
ChatGPT Alternative Solutions: Large Language Models Survey
di: Alipour, Hanieh, et al.
Pubblicazione: (2024)
di: Alipour, Hanieh, et al.
Pubblicazione: (2024)
ArchiveGPT: A human-centered evaluation of using a vision language model for image cataloguing
di: Abele, Line, et al.
Pubblicazione: (2025)
di: Abele, Line, et al.
Pubblicazione: (2025)
Evaluation of GPT-4o and GPT-4o-mini's Vision Capabilities for Compositional Analysis from Dried Solution Drops
di: Dangi, Deven B., et al.
Pubblicazione: (2024)
di: Dangi, Deven B., et al.
Pubblicazione: (2024)
Exploring the features used for summary evaluation by Human and GPT
di: Sadeghi, Zahra, et al.
Pubblicazione: (2025)
di: Sadeghi, Zahra, et al.
Pubblicazione: (2025)
GPTEval: A Survey on Assessments of ChatGPT and GPT-4
di: Mao, Rui, et al.
Pubblicazione: (2023)
di: Mao, Rui, et al.
Pubblicazione: (2023)
GraphGPT: Graph Instruction Tuning for Large Language Models
di: Tang, Jiabin, et al.
Pubblicazione: (2023)
di: Tang, Jiabin, et al.
Pubblicazione: (2023)
Thinking by Subtraction: Confidence-Driven Contrastive Decoding for LLM Reasoning
di: Tang, Lexiang, et al.
Pubblicazione: (2026)
di: Tang, Lexiang, et al.
Pubblicazione: (2026)
Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4
di: Bsharat, Sondos Mahmoud, et al.
Pubblicazione: (2023)
di: Bsharat, Sondos Mahmoud, et al.
Pubblicazione: (2023)
DeepSolution: Boosting Complex Engineering Solution Design via Tree-based Exploration and Bi-point Thinking
di: Li, Zhuoqun, et al.
Pubblicazione: (2025)
di: Li, Zhuoqun, et al.
Pubblicazione: (2025)
Strategy-Induct: Task-Level Strategy Induction for Instruction Generation
di: Chen, Po-Chun, et al.
Pubblicazione: (2026)
di: Chen, Po-Chun, et al.
Pubblicazione: (2026)
The Human and the Mechanical: logos, truthfulness, and ChatGPT
di: Giannakidou, Anastasia, et al.
Pubblicazione: (2024)
di: Giannakidou, Anastasia, et al.
Pubblicazione: (2024)
Is GPT-4 a reliable rater? Evaluating Consistency in GPT-4 Text Ratings
di: Hackl, Veronika, et al.
Pubblicazione: (2023)
di: Hackl, Veronika, et al.
Pubblicazione: (2023)
GPT-4 Technical Report
di: OpenAI, et al.
Pubblicazione: (2023)
di: OpenAI, et al.
Pubblicazione: (2023)
Measuring Pragmatic Influence in Large Language Model Instructions
di: Geng, Yilin, et al.
Pubblicazione: (2026)
di: Geng, Yilin, et al.
Pubblicazione: (2026)
Finding your MUSE: Mining Unexpected Solutions Engine
di: Sweed, Nir, et al.
Pubblicazione: (2025)
di: Sweed, Nir, et al.
Pubblicazione: (2025)
Reasons and Solutions for the Decline in Model Performance after Editing
di: Huang, Xiusheng, et al.
Pubblicazione: (2024)
di: Huang, Xiusheng, et al.
Pubblicazione: (2024)
GenKnowSub: Improving Modularity and Reusability of LLMs through General Knowledge Subtraction
di: Bagherifard, Mohammadtaha, et al.
Pubblicazione: (2025)
di: Bagherifard, Mohammadtaha, et al.
Pubblicazione: (2025)
Benchmarking GPT-4 on Algorithmic Problems: A Systematic Evaluation of Prompting Strategies
di: Petruzzellis, Flavio, et al.
Pubblicazione: (2024)
di: Petruzzellis, Flavio, et al.
Pubblicazione: (2024)
Benchmarking GPT-4 against Human Translators: A Comprehensive Evaluation Across Languages, Domains, and Expertise Levels
di: Yan, Jianhao, et al.
Pubblicazione: (2024)
di: Yan, Jianhao, et al.
Pubblicazione: (2024)
Chasing Random: Instruction Selection Strategies Fail to Generalize
di: Diddee, Harshita, et al.
Pubblicazione: (2024)
di: Diddee, Harshita, et al.
Pubblicazione: (2024)
Differentiate ChatGPT-generated and Human-written Medical Texts
di: Liao, Wenxiong, et al.
Pubblicazione: (2023)
di: Liao, Wenxiong, et al.
Pubblicazione: (2023)
Assessing the Creativity of LLMs in Proposing Novel Solutions to Mathematical Problems
di: Ye, Junyi, et al.
Pubblicazione: (2024)
di: Ye, Junyi, et al.
Pubblicazione: (2024)
CEC-Zero: Chinese Error Correction Solution Based on LLM
di: Zhang, Sophie, et al.
Pubblicazione: (2025)
di: Zhang, Sophie, et al.
Pubblicazione: (2025)
LLM-First Search: Self-Guided Exploration of the Solution Space
di: Herr, Nathan, et al.
Pubblicazione: (2025)
di: Herr, Nathan, et al.
Pubblicazione: (2025)
Towards Robust Universal Information Extraction: Benchmark, Evaluation, and Solution
di: Zhu, Jizhao, et al.
Pubblicazione: (2025)
di: Zhu, Jizhao, et al.
Pubblicazione: (2025)
Comparative Analysis of ChatGPT, GPT-4, and Microsoft Bing Chatbots for GRE Test
di: Abu-Haifa, Mohammad, et al.
Pubblicazione: (2023)
di: Abu-Haifa, Mohammad, et al.
Pubblicazione: (2023)
Reasoning: From Reflection to Solution
di: Li, Zixi
Pubblicazione: (2025)
di: Li, Zixi
Pubblicazione: (2025)
Inductive-Deductive Strategy Reuse for Multi-Turn Instructional Dialogues
di: Ou, Jiao, et al.
Pubblicazione: (2024)
di: Ou, Jiao, et al.
Pubblicazione: (2024)
Does GPT-4 pass the Turing test?
di: Jones, Cameron R., et al.
Pubblicazione: (2023)
di: Jones, Cameron R., et al.
Pubblicazione: (2023)
Advancing AI with Integrity: Ethical Challenges and Solutions in Neural Machine Translation
di: Kimera, Richard, et al.
Pubblicazione: (2024)
di: Kimera, Richard, et al.
Pubblicazione: (2024)
Is Fine-Tuning an Effective Solution? Reassessing Knowledge Editing for Unstructured Data
di: Xiong, Hao, et al.
Pubblicazione: (2025)
di: Xiong, Hao, et al.
Pubblicazione: (2025)
A Rule Based Solution to Co-reference Resolution in Clinical Text
di: Chen, Ping, et al.
Pubblicazione: (2025)
di: Chen, Ping, et al.
Pubblicazione: (2025)
Using Combinatorial Optimization to Design a High quality LLM Solution
di: Ackerman, Samuel, et al.
Pubblicazione: (2024)
di: Ackerman, Samuel, et al.
Pubblicazione: (2024)
B4: Towards Optimal Assessment of Plausible Code Solutions with Plausible Tests
di: Chen, Mouxiang, et al.
Pubblicazione: (2024)
di: Chen, Mouxiang, et al.
Pubblicazione: (2024)
Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
di: Li, Hang, et al.
Pubblicazione: (2025)
di: Li, Hang, et al.
Pubblicazione: (2025)
A Looming Replication Crisis in Evaluating Behavior in Language Models? Evidence and Solutions
di: Vaugrante, Laurène, et al.
Pubblicazione: (2024)
di: Vaugrante, Laurène, et al.
Pubblicazione: (2024)
Building another Spanish dictionary, this time with GPT-4
di: Ortega-Martín, Miguel, et al.
Pubblicazione: (2024)
di: Ortega-Martín, Miguel, et al.
Pubblicazione: (2024)
Human-Instruction-Free LLM Self-Alignment with Limited Samples
di: Guo, Hongyi, et al.
Pubblicazione: (2024)
di: Guo, Hongyi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Towards a Psychology of Machines: Large Language Models Predict Human Memory
di: Huff, Markus, et al.
Pubblicazione: (2024) -
Replicating Human Social Perception in Generative AI: Evaluating the Valence-Dominance Model
di: Gurkan, Necdet, et al.
Pubblicazione: (2025) -
ChatGPT Alternative Solutions: Large Language Models Survey
di: Alipour, Hanieh, et al.
Pubblicazione: (2024) -
ArchiveGPT: A human-centered evaluation of using a vision language model for image cataloguing
di: Abele, Line, et al.
Pubblicazione: (2025) -
Evaluation of GPT-4o and GPT-4o-mini's Vision Capabilities for Compositional Analysis from Dried Solution Drops
di: Dangi, Deven B., et al.
Pubblicazione: (2024)