GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Havrilla, Alex, Raparthy, Sharath, Nalmpantis, Christoforus, Dwivedi-Yu, Jane, Zhuravinskyi, Maksym, Hambro, Eric, Raileanu, Roberta |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Teaching Large Language Models to Reason with Reinforcement Learning
par: Havrilla, Alex, et autres
Publié: (2024)
par: Havrilla, Alex, et autres
Publié: (2024)
Understanding the Effects of RLHF on LLM Generalisation and Diversity
par: Kirk, Robert, et autres
Publié: (2023)
par: Kirk, Robert, et autres
Publié: (2023)
Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts
par: Samvelyan, Mikayel, et autres
Publié: (2024)
par: Samvelyan, Mikayel, et autres
Publié: (2024)
GLoD: Composing Global Contexts and Local Details in Image Generation
par: Yamada, Moyuru
Publié: (2024)
par: Yamada, Moyuru
Publié: (2024)
Know When To Stop: A Study of Semantic Drift in Text Generation
par: Spataru, Ava, et autres
Publié: (2024)
par: Spataru, Ava, et autres
Publié: (2024)
Understanding Scaling Laws with Statistical and Approximation Theory for Transformer Neural Networks on Intrinsically Low-dimensional Data
par: Havrilla, Alex, et autres
Publié: (2024)
par: Havrilla, Alex, et autres
Publié: (2024)
Understanding the Effect of Noise in LLM Training Data with Algorithmic Chains of Thought
par: Havrilla, Alex, et autres
Publié: (2024)
par: Havrilla, Alex, et autres
Publié: (2024)
TOOLVERIFIER: Generalization to New Tools via Self-Verification
par: Mekala, Dheeraj, et autres
Publié: (2024)
par: Mekala, Dheeraj, et autres
Publié: (2024)
SPARQ: Synthetic Problem Generation for Reasoning via Quality-Diversity Algorithms
par: Havrilla, Alex, et autres
Publié: (2025)
par: Havrilla, Alex, et autres
Publié: (2025)
GLoRE: Evaluating Logical Reasoning of Large Language Models
par: liu, Hanmeng, et autres
Publié: (2023)
par: liu, Hanmeng, et autres
Publié: (2023)
Source2Synth: Synthetic Data Generation and Curation Grounded in Real Data Sources
par: Lupidi, Alisia, et autres
Publié: (2024)
par: Lupidi, Alisia, et autres
Publié: (2024)
LLM-First Search: Self-Guided Exploration of the Solution Space
par: Herr, Nathan, et autres
Publié: (2025)
par: Herr, Nathan, et autres
Publié: (2025)
Progress: A Post-AI Manifesto
par: Haryanto, Christoforus Yoga
Publié: (2024)
par: Haryanto, Christoforus Yoga
Publié: (2024)
LLAssist: Simple Tools for Automating Literature Review Using Large Language Models
par: Haryanto, Christoforus Yoga
Publié: (2024)
par: Haryanto, Christoforus Yoga
Publié: (2024)
Khinchin-type inequalities via Hadamard's factorisation
par: Havrilla, Alex, et autres
Publié: (2021)
par: Havrilla, Alex, et autres
Publié: (2021)
GLoCIM: Global-view Long Chain Interest Modeling for news recommendation
par: Yang, Zhen, et autres
Publié: (2024)
par: Yang, Zhen, et autres
Publié: (2024)
Epistemic Dissonance and Modal Boundaries
par: Raileanu, Dragos
Publié: (2025)
par: Raileanu, Dragos
Publié: (2025)
IGDA: Interactive Graph Discovery through Large Language Model Agents
par: Havrilla, Alex, et autres
Publié: (2025)
par: Havrilla, Alex, et autres
Publié: (2025)
Global versus Local: Evaluating AlexNet Architectures for Tropical Cyclone Intensity Estimation
par: Dwivedi, Vikas
Publié: (2024)
par: Dwivedi, Vikas
Publié: (2024)
The Generalization Gap in Offline Reinforcement Learning
par: Mediratta, Ishita, et autres
Publié: (2023)
par: Mediratta, Ishita, et autres
Publié: (2023)
DFU: scale-robust diffusion model for zero-shot super-resolution image generation
par: Havrilla, Alex, et autres
Publié: (2023)
par: Havrilla, Alex, et autres
Publié: (2023)
Cognitive Silicon: An Architectural Blueprint for Post-Industrial Computing Systems
par: Haryanto, Christoforus Yoga, et autres
Publié: (2025)
par: Haryanto, Christoforus Yoga, et autres
Publié: (2025)
Pramana: Fine-Tuning Large Language Models for Epistemic Reasoning through Navya-Nyaya
par: Sathish, Sharath
Publié: (2026)
par: Sathish, Sharath
Publié: (2026)
Improving Language Plasticity via Pretraining with Active Forgetting
par: Chen, Yihong, et autres
Publié: (2023)
par: Chen, Yihong, et autres
Publié: (2023)
GLoW: novel methods for wave-optics phenomena in gravitational lensing
par: Villarrubia-Rojo, Hector, et autres
Publié: (2024)
par: Villarrubia-Rojo, Hector, et autres
Publié: (2024)
GLoRIA: Gated Low-Rank Interpretable Adaptation for Dialectal ASR
par: Mehralian, Pouya, et autres
Publié: (2026)
par: Mehralian, Pouya, et autres
Publié: (2026)
GLoSS: Generative Language Models with Semantic Search for Sequential Recommendation
par: Acharya, Krishna, et autres
Publié: (2025)
par: Acharya, Krishna, et autres
Publié: (2025)
DreamCraft: Text-Guided Generation of Functional 3D Environments in Minecraft
par: Earle, Sam, et autres
Publié: (2024)
par: Earle, Sam, et autres
Publié: (2024)
Where, When, and How? Integrating Spatiotemporal Cues in Cell Division
par: Luca Cirillo, et autres
Publié: (2025)
par: Luca Cirillo, et autres
Publié: (2025)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
par: Shahin, Nada, et autres
Publié: (2025)
par: Shahin, Nada, et autres
Publié: (2025)
Transformers for Learning on Noisy and Task-Level Manifolds: Approximation and Generalization Insights
par: Shen, Zhaiming, et autres
Publié: (2025)
par: Shen, Zhaiming, et autres
Publié: (2025)
FairPair: A Robust Evaluation of Biases in Language Models through Paired Perturbations
par: Dwivedi-Yu, Jane, et autres
Publié: (2024)
par: Dwivedi-Yu, Jane, et autres
Publié: (2024)
Music Proofreading with RefinPaint: Where and How to Modify Compositions given Context
par: Ramoneda, Pedro, et autres
Publié: (2024)
par: Ramoneda, Pedro, et autres
Publié: (2024)
What Are the Most Important Contributors to Arctic Precipitation—When, Where, and How?
par: Melanie Lauer, et autres
Publié: (2025)
par: Melanie Lauer, et autres
Publié: (2025)
Are Large Language Models Strategic Decision Makers? A Study of Performance and Bias in Two-Player Non-Zero-Sum Games
par: Herr, Nathan, et autres
Publié: (2024)
par: Herr, Nathan, et autres
Publié: (2024)
Transparent AI Disclosure Obligations: Who, What, When, Where, Why, How
par: Ali, Abdallah El, et autres
Publié: (2024)
par: Ali, Abdallah El, et autres
Publié: (2024)
Research Methods in a Nutshell: What, Why, When, Where, Who, and How?
par: Justin Paul, et autres
Publié: (2025)
par: Justin Paul, et autres
Publié: (2025)
GLoG-CSUnet: Enhancing Vision Transformers with Adaptable Radiomic Features for Medical Image Segmentation
par: Eghbali, Niloufar, et autres
Publié: (2025)
par: Eghbali, Niloufar, et autres
Publié: (2025)
Stable Cinemetrics : Structured Taxonomy and Evaluation for Professional Video Generation
par: Chatterjee, Agneet, et autres
Publié: (2025)
par: Chatterjee, Agneet, et autres
Publié: (2025)
When and Where Localization Fails: An Analysis of the Iterative Closest Point in Evolving Environment
par: Dannaoui, Abdel-Raouf, et autres
Publié: (2025)
par: Dannaoui, Abdel-Raouf, et autres
Publié: (2025)
Documents similaires
-
Teaching Large Language Models to Reason with Reinforcement Learning
par: Havrilla, Alex, et autres
Publié: (2024) -
Understanding the Effects of RLHF on LLM Generalisation and Diversity
par: Kirk, Robert, et autres
Publié: (2023) -
Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts
par: Samvelyan, Mikayel, et autres
Publié: (2024) -
GLoD: Composing Global Contexts and Local Details in Image Generation
par: Yamada, Moyuru
Publié: (2024) -
Know When To Stop: A Study of Semantic Drift in Text Generation
par: Spataru, Ava, et autres
Publié: (2024)