Impact of Noise on LLM-Models Performance in Abstraction and Reasoning Corpus (ARC) Tasks with Model Temperature Considerations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Khandalkar, Nikhil, Yadav, Pavan, Shinde, Krishna, Ramegowda, Lokesh B., Das, Rajarshi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring Next Token Prediction in Theory of Mind (ToM) Tasks: Comparative Experiments with GPT-2 and LLaMA-2 AI Models
von: Yadav, Pavan, et al.
Veröffentlicht: (2025)
von: Yadav, Pavan, et al.
Veröffentlicht: (2025)
H-ARC: A Robust Estimate of Human Performance on the Abstraction and Reasoning Corpus Benchmark
von: LeGris, Solim, et al.
Veröffentlicht: (2024)
von: LeGris, Solim, et al.
Veröffentlicht: (2024)
ARC-NCA: Towards Developmental Solutions to the Abstraction and Reasoning Corpus
von: Guichard, Etienne, et al.
Veröffentlicht: (2025)
von: Guichard, Etienne, et al.
Veröffentlicht: (2025)
ARC-GEN: A Mimetic Procedural Benchmark Generator for the Abstraction and Reasoning Corpus
von: Moffitt, Michael D.
Veröffentlicht: (2025)
von: Moffitt, Michael D.
Veröffentlicht: (2025)
Generalized Planning for the Abstraction and Reasoning Corpus
von: Lei, Chao, et al.
Veröffentlicht: (2024)
von: Lei, Chao, et al.
Veröffentlicht: (2024)
JaxARC: A High-Performance JAX-based Environment for Abstraction and Reasoning Research
von: Aadam, et al.
Veröffentlicht: (2026)
von: Aadam, et al.
Veröffentlicht: (2026)
Abductive Symbolic Solver on Abstraction and Reasoning Corpus
von: Lim, Mintaek, et al.
Veröffentlicht: (2024)
von: Lim, Mintaek, et al.
Veröffentlicht: (2024)
Enhancing Analogical Reasoning in the Abstraction and Reasoning Corpus via Model-Based RL
von: Lee, Jihwan, et al.
Veröffentlicht: (2024)
von: Lee, Jihwan, et al.
Veröffentlicht: (2024)
Vector Symbolic Algebras for the Abstraction and Reasoning Corpus
von: Joffe, Isaac, et al.
Veröffentlicht: (2025)
von: Joffe, Isaac, et al.
Veröffentlicht: (2025)
Capturing Sparks of Abstraction for the ARC Challenge
von: Andrews, Martin
Veröffentlicht: (2024)
von: Andrews, Martin
Veröffentlicht: (2024)
Do LLMs Build Spatial World Models? Evidence from Grid-World Maze Tasks
von: Li, Weijiang, et al.
Veröffentlicht: (2026)
von: Li, Weijiang, et al.
Veröffentlicht: (2026)
Reasoning Abilities of Large Language Models: In-Depth Analysis on the Abstraction and Reasoning Corpus
von: Lee, Seungpil, et al.
Veröffentlicht: (2024)
von: Lee, Seungpil, et al.
Veröffentlicht: (2024)
Addressing the Abstraction and Reasoning Corpus via Procedural Example Generation
von: Hodel, Michael
Veröffentlicht: (2024)
von: Hodel, Michael
Veröffentlicht: (2024)
ARCLE: The Abstraction and Reasoning Corpus Learning Environment for Reinforcement Learning
von: Lee, Hosung, et al.
Veröffentlicht: (2024)
von: Lee, Hosung, et al.
Veröffentlicht: (2024)
LLMs and the Abstraction and Reasoning Corpus: Successes, Failures, and the Importance of Object-based Representations
von: Xu, Yudong, et al.
Veröffentlicht: (2023)
von: Xu, Yudong, et al.
Veröffentlicht: (2023)
ARC-TGI: Human-Validated Task Generators with Reasoning Chain Templates for ARC-AGI
von: Lehmann, Jens, et al.
Veröffentlicht: (2026)
von: Lehmann, Jens, et al.
Veröffentlicht: (2026)
Graph-Based Exploration for ARC-AGI-3 Interactive Reasoning Tasks
von: Rudakov, Evgenii, et al.
Veröffentlicht: (2025)
von: Rudakov, Evgenii, et al.
Veröffentlicht: (2025)
The ARC of Progress towards AGI: A Living Survey of Abstraction and Reasoning
von: Vahdati, Sahar, et al.
Veröffentlicht: (2026)
von: Vahdati, Sahar, et al.
Veröffentlicht: (2026)
Program Synthesis using Inductive Logic Programming for the Abstraction and Reasoning Corpus
von: Rocha, Filipe Marinho, et al.
Veröffentlicht: (2024)
von: Rocha, Filipe Marinho, et al.
Veröffentlicht: (2024)
CausalARC: Abstract Reasoning with Causal World Models
von: Maasch, Jacqueline, et al.
Veröffentlicht: (2025)
von: Maasch, Jacqueline, et al.
Veröffentlicht: (2025)
LLM-ARC: Enhancing LLMs with an Automated Reasoning Critic
von: Kalyanpur, Aditya, et al.
Veröffentlicht: (2024)
von: Kalyanpur, Aditya, et al.
Veröffentlicht: (2024)
Analyzing the Performance of Large Language Models on Code Summarization
von: Haldar, Rajarshi, et al.
Veröffentlicht: (2024)
von: Haldar, Rajarshi, et al.
Veröffentlicht: (2024)
Improving Noise Robustness through Abstractions and its Impact on Machine Learning
von: Ibias, Alfredo, et al.
Veröffentlicht: (2024)
von: Ibias, Alfredo, et al.
Veröffentlicht: (2024)
PEA: Enhancing LLM Performance on Computational-Reasoning Tasks
von: Wang, Zi, et al.
Veröffentlicht: (2025)
von: Wang, Zi, et al.
Veröffentlicht: (2025)
Abstraction-of-Thought Makes Language Models Better Reasoners
von: Hong, Ruixin, et al.
Veröffentlicht: (2024)
von: Hong, Ruixin, et al.
Veröffentlicht: (2024)
Exploring Human Behavior During Abstract Rule Inference and Problem Solving with the Cognitive Abstraction and Reasoning Corpus
von: Ahn, Caroline, et al.
Veröffentlicht: (2026)
von: Ahn, Caroline, et al.
Veröffentlicht: (2026)
Understanding LLMs' Fluid Intelligence Deficiency: An Analysis of the ARC Task
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
Is Large Language Model Performance on Reasoning Tasks Impacted by Different Ways Questions Are Asked?
von: Song, Seok Hwan, et al.
Veröffentlicht: (2025)
von: Song, Seok Hwan, et al.
Veröffentlicht: (2025)
Tackling the Abstraction and Reasoning Corpus with Vision Transformers: the Importance of 2D Representation, Positions, and Objects
von: Li, Wenhao, et al.
Veröffentlicht: (2024)
von: Li, Wenhao, et al.
Veröffentlicht: (2024)
Compositional-ARC: Assessing Systematic Generalization in Abstract Spatial Reasoning
von: Mondorf, Philipp, et al.
Veröffentlicht: (2025)
von: Mondorf, Philipp, et al.
Veröffentlicht: (2025)
Same Task, More Tokens: the Impact of Input Length on the Reasoning Performance of Large Language Models
von: Levy, Mosh, et al.
Veröffentlicht: (2024)
von: Levy, Mosh, et al.
Veröffentlicht: (2024)
A Multi-Stage Workflow for the Review of Marketing Content with Reasoning Large Language Models
von: Purpura, Alberto, et al.
Veröffentlicht: (2025)
von: Purpura, Alberto, et al.
Veröffentlicht: (2025)
GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning
von: Peltonen, Saku, et al.
Veröffentlicht: (2026)
von: Peltonen, Saku, et al.
Veröffentlicht: (2026)
Enhancing Structural Mapping with LLM-derived Abstractions for Analogical Reasoning in Narratives
von: Khojasteh, Mohammadhossein, et al.
Veröffentlicht: (2026)
von: Khojasteh, Mohammadhossein, et al.
Veröffentlicht: (2026)
Lost-in-Distance: Impact of Contextual Proximity on LLM Performance in Graph Tasks
von: Firooz, Hamed, et al.
Veröffentlicht: (2024)
von: Firooz, Hamed, et al.
Veröffentlicht: (2024)
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Parallel LLM Reasoning for Bias-Resilient, Robust Conceptual Abstraction
von: Adeseye, Aisvarya, et al.
Veröffentlicht: (2026)
von: Adeseye, Aisvarya, et al.
Veröffentlicht: (2026)
Executable World Models for ARC-AGI-3 in the Era of Coding Agents
von: Rodionov, Sergey
Veröffentlicht: (2026)
von: Rodionov, Sergey
Veröffentlicht: (2026)
ARC-AGI-2: A New Challenge for Frontier AI Reasoning Systems
von: Chollet, Francois, et al.
Veröffentlicht: (2025)
von: Chollet, Francois, et al.
Veröffentlicht: (2025)
From Reasoning to Generalization: Knowledge-Augmented LLMs for ARC Benchmark
von: Lei, Chao, et al.
Veröffentlicht: (2025)
von: Lei, Chao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Exploring Next Token Prediction in Theory of Mind (ToM) Tasks: Comparative Experiments with GPT-2 and LLaMA-2 AI Models
von: Yadav, Pavan, et al.
Veröffentlicht: (2025) -
H-ARC: A Robust Estimate of Human Performance on the Abstraction and Reasoning Corpus Benchmark
von: LeGris, Solim, et al.
Veröffentlicht: (2024) -
ARC-NCA: Towards Developmental Solutions to the Abstraction and Reasoning Corpus
von: Guichard, Etienne, et al.
Veröffentlicht: (2025) -
ARC-GEN: A Mimetic Procedural Benchmark Generator for the Abstraction and Reasoning Corpus
von: Moffitt, Michael D.
Veröffentlicht: (2025) -
Generalized Planning for the Abstraction and Reasoning Corpus
von: Lei, Chao, et al.
Veröffentlicht: (2024)