OccamLLM: Fast and Exact Language Model Arithmetic in a Single Step
Fuente:
arXiv
Salvato in:
| Autori principali: | Dugan, Owen, Beneto, Donato Manuel Jimenez, Loh, Charlotte, Chen, Zhuo, Dangovski, Rumen, Soljačić, Marin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
QuanTA: Efficient High-Rank Fine-Tuning of LLMs with Quantum-Informed Tensor Adaptation
di: Chen, Zhuo, et al.
Pubblicazione: (2024)
di: Chen, Zhuo, et al.
Pubblicazione: (2024)
Active learning for photonic crystals
di: Lopez, Ryan, et al.
Pubblicazione: (2026)
di: Lopez, Ryan, et al.
Pubblicazione: (2026)
On a Novel Application of Wasserstein-Procrustes for Unsupervised Cross-Lingual Learning
di: Ramírez, Guillem, et al.
Pubblicazione: (2020)
di: Ramírez, Guillem, et al.
Pubblicazione: (2020)
Data-Informed Global Sparseness in Attention Mechanisms for Deep Neural Networks
di: Rugina, Ileana, et al.
Pubblicazione: (2020)
di: Rugina, Ileana, et al.
Pubblicazione: (2020)
QuACK: Accelerating Gradient-Based Quantum Optimization with Koopman Operator Learning
di: Luo, Di, et al.
Pubblicazione: (2022)
di: Luo, Di, et al.
Pubblicazione: (2022)
Multimodal Foundation Models for Material Property Prediction and Discovery
di: Moro, Viggo, et al.
Pubblicazione: (2023)
di: Moro, Viggo, et al.
Pubblicazione: (2023)
Predicting band gap from chemical composition: A simple learned model for a material property with atypical statistics
di: Ma, Andrew, et al.
Pubblicazione: (2025)
di: Ma, Andrew, et al.
Pubblicazione: (2025)
L$^2$M: Mutual Information Scaling Law for Long-Context Language Modeling
di: Chen, Zhuo, et al.
Pubblicazione: (2025)
di: Chen, Zhuo, et al.
Pubblicazione: (2025)
TraversalBench: Challenging Paths to Follow for Vision Language Models
di: Petrova, Clara, et al.
Pubblicazione: (2026)
di: Petrova, Clara, et al.
Pubblicazione: (2026)
Photonic probabilistic machine learning using quantum vacuum noise
di: Choi, Seou, et al.
Pubblicazione: (2024)
di: Choi, Seou, et al.
Pubblicazione: (2024)
Do Language Models Follow Occam's Razor? An Evaluation of Parsimony in Inductive and Abductive Reasoning
di: Sun, Yunxin, et al.
Pubblicazione: (2025)
di: Sun, Yunxin, et al.
Pubblicazione: (2025)
Occam's Razor and Bender and Koller's Octopus
di: Guerzhoy, Michael
Pubblicazione: (2024)
di: Guerzhoy, Michael
Pubblicazione: (2024)
AgentOccam: A Simple Yet Strong Baseline for LLM-Based Web Agents
di: Yang, Ke, et al.
Pubblicazione: (2024)
di: Yang, Ke, et al.
Pubblicazione: (2024)
T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
di: Chen, Zehui, et al.
Pubblicazione: (2023)
di: Chen, Zehui, et al.
Pubblicazione: (2023)
Language Models are Symbolic Learners in Arithmetic
di: Deng, Chunyuan, et al.
Pubblicazione: (2024)
di: Deng, Chunyuan, et al.
Pubblicazione: (2024)
Large Language Models for Single-Step and Multi-Step Flight Trajectory Prediction
di: Luo, Kaiwei, et al.
Pubblicazione: (2025)
di: Luo, Kaiwei, et al.
Pubblicazione: (2025)
On Representational Dissociation of Language and Arithmetic in Large Language Models
di: Kisako, Riku, et al.
Pubblicazione: (2025)
di: Kisako, Riku, et al.
Pubblicazione: (2025)
Arithmetic Reasoning with LLM: Prolog Generation & Permutation
di: Yang, Xiaocheng, et al.
Pubblicazione: (2024)
di: Yang, Xiaocheng, et al.
Pubblicazione: (2024)
Steering Language Models with Weight Arithmetic
di: Fierro, Constanza, et al.
Pubblicazione: (2025)
di: Fierro, Constanza, et al.
Pubblicazione: (2025)
Probing for Arithmetic Errors in Language Models
di: Sun, Yucheng, et al.
Pubblicazione: (2025)
di: Sun, Yucheng, et al.
Pubblicazione: (2025)
FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models
di: Monsefi, Amin Karimi, et al.
Pubblicazione: (2025)
di: Monsefi, Amin Karimi, et al.
Pubblicazione: (2025)
LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models
di: Hao, Shibo, et al.
Pubblicazione: (2024)
di: Hao, Shibo, et al.
Pubblicazione: (2024)
MetaGPT: Merging Large Language Models Using Model Exclusive Task Arithmetic
di: Zhou, Yuyan, et al.
Pubblicazione: (2024)
di: Zhou, Yuyan, et al.
Pubblicazione: (2024)
Unraveling Arithmetic in Large Language Models: The Role of Algebraic Structures
di: Chang, Fu-Chieh, et al.
Pubblicazione: (2024)
di: Chang, Fu-Chieh, et al.
Pubblicazione: (2024)
Interpreting and Improving Large Language Models in Arithmetic Calculation
di: Zhang, Wei, et al.
Pubblicazione: (2024)
di: Zhang, Wei, et al.
Pubblicazione: (2024)
Controlled Text Generation via Language Model Arithmetic
di: Dekoninck, Jasper, et al.
Pubblicazione: (2023)
di: Dekoninck, Jasper, et al.
Pubblicazione: (2023)
Arithmetic with Language Models: from Memorization to Computation
di: Maltoni, Davide, et al.
Pubblicazione: (2023)
di: Maltoni, Davide, et al.
Pubblicazione: (2023)
Self-training Language Models for Arithmetic Reasoning
di: Kadlčík, Marek, et al.
Pubblicazione: (2024)
di: Kadlčík, Marek, et al.
Pubblicazione: (2024)
In-Context Occam's Razor: How Transformers Prefer Simpler Hypotheses on the Fly
di: Deora, Puneesh, et al.
Pubblicazione: (2025)
di: Deora, Puneesh, et al.
Pubblicazione: (2025)
A Comparative Analysis of LLM Adaptation: SFT, LoRA, and ICL in Data-Scarce Scenarios
di: Bohnet, Bernd, et al.
Pubblicazione: (2025)
di: Bohnet, Bernd, et al.
Pubblicazione: (2025)
Instruct Large Language Models to Generate Scientific Literature Survey Step by Step
di: Lai, Yuxuan, et al.
Pubblicazione: (2024)
di: Lai, Yuxuan, et al.
Pubblicazione: (2024)
Souper-Model: How Simple Arithmetic Unlocks State-of-the-Art LLM Performance
di: Maiti, Shalini, et al.
Pubblicazione: (2025)
di: Maiti, Shalini, et al.
Pubblicazione: (2025)
Think Fast and Slow: Step-Level Cognitive Depth Adaptation for LLM Agents
di: Yang, Ruihan, et al.
Pubblicazione: (2026)
di: Yang, Ruihan, et al.
Pubblicazione: (2026)
A Careful Examination of Large Language Model Performance on Grade School Arithmetic
di: Zhang, Hugh, et al.
Pubblicazione: (2024)
di: Zhang, Hugh, et al.
Pubblicazione: (2024)
Forgetting before Learning: Utilizing Parametric Arithmetic for Knowledge Updating in Large Language Models
di: Ni, Shiwen, et al.
Pubblicazione: (2023)
di: Ni, Shiwen, et al.
Pubblicazione: (2023)
Language Models Do Hard Arithmetic Tasks Easily and Hardly Do Easy Arithmetic Tasks
di: Gambardella, Andrew, et al.
Pubblicazione: (2024)
di: Gambardella, Andrew, et al.
Pubblicazione: (2024)
Modular Arithmetic: Language Models Solve Math Digit by Digit
di: Baeumel, Tanja, et al.
Pubblicazione: (2025)
di: Baeumel, Tanja, et al.
Pubblicazione: (2025)
Leveraging Language Models and Bandit Algorithms to Drive Adoption of Battery-Electric Vehicles
di: Namikoshi, Keiichi, et al.
Pubblicazione: (2024)
di: Namikoshi, Keiichi, et al.
Pubblicazione: (2024)
Language Arithmetics: Towards Systematic Language Neuron Identification and Manipulation
di: Gurgurov, Daniil, et al.
Pubblicazione: (2025)
di: Gurgurov, Daniil, et al.
Pubblicazione: (2025)
Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis
di: Yu, Zeping, et al.
Pubblicazione: (2024)
di: Yu, Zeping, et al.
Pubblicazione: (2024)
Documenti analoghi
-
QuanTA: Efficient High-Rank Fine-Tuning of LLMs with Quantum-Informed Tensor Adaptation
di: Chen, Zhuo, et al.
Pubblicazione: (2024) -
Active learning for photonic crystals
di: Lopez, Ryan, et al.
Pubblicazione: (2026) -
On a Novel Application of Wasserstein-Procrustes for Unsupervised Cross-Lingual Learning
di: Ramírez, Guillem, et al.
Pubblicazione: (2020) -
Data-Informed Global Sparseness in Attention Mechanisms for Deep Neural Networks
di: Rugina, Ileana, et al.
Pubblicazione: (2020) -
QuACK: Accelerating Gradient-Based Quantum Optimization with Koopman Operator Learning
di: Luo, Di, et al.
Pubblicazione: (2022)