OccamLLM: Fast and Exact Language Model Arithmetic in a Single Step
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dugan, Owen, Beneto, Donato Manuel Jimenez, Loh, Charlotte, Chen, Zhuo, Dangovski, Rumen, Soljačić, Marin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
QuanTA: Efficient High-Rank Fine-Tuning of LLMs with Quantum-Informed Tensor Adaptation
von: Chen, Zhuo, et al.
Veröffentlicht: (2024)
von: Chen, Zhuo, et al.
Veröffentlicht: (2024)
Active learning for photonic crystals
von: Lopez, Ryan, et al.
Veröffentlicht: (2026)
von: Lopez, Ryan, et al.
Veröffentlicht: (2026)
On a Novel Application of Wasserstein-Procrustes for Unsupervised Cross-Lingual Learning
von: Ramírez, Guillem, et al.
Veröffentlicht: (2020)
von: Ramírez, Guillem, et al.
Veröffentlicht: (2020)
Data-Informed Global Sparseness in Attention Mechanisms for Deep Neural Networks
von: Rugina, Ileana, et al.
Veröffentlicht: (2020)
von: Rugina, Ileana, et al.
Veröffentlicht: (2020)
QuACK: Accelerating Gradient-Based Quantum Optimization with Koopman Operator Learning
von: Luo, Di, et al.
Veröffentlicht: (2022)
von: Luo, Di, et al.
Veröffentlicht: (2022)
Multimodal Foundation Models for Material Property Prediction and Discovery
von: Moro, Viggo, et al.
Veröffentlicht: (2023)
von: Moro, Viggo, et al.
Veröffentlicht: (2023)
Predicting band gap from chemical composition: A simple learned model for a material property with atypical statistics
von: Ma, Andrew, et al.
Veröffentlicht: (2025)
von: Ma, Andrew, et al.
Veröffentlicht: (2025)
L$^2$M: Mutual Information Scaling Law for Long-Context Language Modeling
von: Chen, Zhuo, et al.
Veröffentlicht: (2025)
von: Chen, Zhuo, et al.
Veröffentlicht: (2025)
TraversalBench: Challenging Paths to Follow for Vision Language Models
von: Petrova, Clara, et al.
Veröffentlicht: (2026)
von: Petrova, Clara, et al.
Veröffentlicht: (2026)
Photonic probabilistic machine learning using quantum vacuum noise
von: Choi, Seou, et al.
Veröffentlicht: (2024)
von: Choi, Seou, et al.
Veröffentlicht: (2024)
Do Language Models Follow Occam's Razor? An Evaluation of Parsimony in Inductive and Abductive Reasoning
von: Sun, Yunxin, et al.
Veröffentlicht: (2025)
von: Sun, Yunxin, et al.
Veröffentlicht: (2025)
Occam's Razor and Bender and Koller's Octopus
von: Guerzhoy, Michael
Veröffentlicht: (2024)
von: Guerzhoy, Michael
Veröffentlicht: (2024)
AgentOccam: A Simple Yet Strong Baseline for LLM-Based Web Agents
von: Yang, Ke, et al.
Veröffentlicht: (2024)
von: Yang, Ke, et al.
Veröffentlicht: (2024)
T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
von: Chen, Zehui, et al.
Veröffentlicht: (2023)
von: Chen, Zehui, et al.
Veröffentlicht: (2023)
Language Models are Symbolic Learners in Arithmetic
von: Deng, Chunyuan, et al.
Veröffentlicht: (2024)
von: Deng, Chunyuan, et al.
Veröffentlicht: (2024)
Large Language Models for Single-Step and Multi-Step Flight Trajectory Prediction
von: Luo, Kaiwei, et al.
Veröffentlicht: (2025)
von: Luo, Kaiwei, et al.
Veröffentlicht: (2025)
On Representational Dissociation of Language and Arithmetic in Large Language Models
von: Kisako, Riku, et al.
Veröffentlicht: (2025)
von: Kisako, Riku, et al.
Veröffentlicht: (2025)
Arithmetic Reasoning with LLM: Prolog Generation & Permutation
von: Yang, Xiaocheng, et al.
Veröffentlicht: (2024)
von: Yang, Xiaocheng, et al.
Veröffentlicht: (2024)
Steering Language Models with Weight Arithmetic
von: Fierro, Constanza, et al.
Veröffentlicht: (2025)
von: Fierro, Constanza, et al.
Veröffentlicht: (2025)
Probing for Arithmetic Errors in Language Models
von: Sun, Yucheng, et al.
Veröffentlicht: (2025)
von: Sun, Yucheng, et al.
Veröffentlicht: (2025)
FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2025)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2025)
LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models
von: Hao, Shibo, et al.
Veröffentlicht: (2024)
von: Hao, Shibo, et al.
Veröffentlicht: (2024)
MetaGPT: Merging Large Language Models Using Model Exclusive Task Arithmetic
von: Zhou, Yuyan, et al.
Veröffentlicht: (2024)
von: Zhou, Yuyan, et al.
Veröffentlicht: (2024)
Unraveling Arithmetic in Large Language Models: The Role of Algebraic Structures
von: Chang, Fu-Chieh, et al.
Veröffentlicht: (2024)
von: Chang, Fu-Chieh, et al.
Veröffentlicht: (2024)
Interpreting and Improving Large Language Models in Arithmetic Calculation
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
Controlled Text Generation via Language Model Arithmetic
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2023)
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2023)
Arithmetic with Language Models: from Memorization to Computation
von: Maltoni, Davide, et al.
Veröffentlicht: (2023)
von: Maltoni, Davide, et al.
Veröffentlicht: (2023)
Self-training Language Models for Arithmetic Reasoning
von: Kadlčík, Marek, et al.
Veröffentlicht: (2024)
von: Kadlčík, Marek, et al.
Veröffentlicht: (2024)
In-Context Occam's Razor: How Transformers Prefer Simpler Hypotheses on the Fly
von: Deora, Puneesh, et al.
Veröffentlicht: (2025)
von: Deora, Puneesh, et al.
Veröffentlicht: (2025)
A Comparative Analysis of LLM Adaptation: SFT, LoRA, and ICL in Data-Scarce Scenarios
von: Bohnet, Bernd, et al.
Veröffentlicht: (2025)
von: Bohnet, Bernd, et al.
Veröffentlicht: (2025)
Instruct Large Language Models to Generate Scientific Literature Survey Step by Step
von: Lai, Yuxuan, et al.
Veröffentlicht: (2024)
von: Lai, Yuxuan, et al.
Veröffentlicht: (2024)
Souper-Model: How Simple Arithmetic Unlocks State-of-the-Art LLM Performance
von: Maiti, Shalini, et al.
Veröffentlicht: (2025)
von: Maiti, Shalini, et al.
Veröffentlicht: (2025)
Think Fast and Slow: Step-Level Cognitive Depth Adaptation for LLM Agents
von: Yang, Ruihan, et al.
Veröffentlicht: (2026)
von: Yang, Ruihan, et al.
Veröffentlicht: (2026)
A Careful Examination of Large Language Model Performance on Grade School Arithmetic
von: Zhang, Hugh, et al.
Veröffentlicht: (2024)
von: Zhang, Hugh, et al.
Veröffentlicht: (2024)
Forgetting before Learning: Utilizing Parametric Arithmetic for Knowledge Updating in Large Language Models
von: Ni, Shiwen, et al.
Veröffentlicht: (2023)
von: Ni, Shiwen, et al.
Veröffentlicht: (2023)
Language Models Do Hard Arithmetic Tasks Easily and Hardly Do Easy Arithmetic Tasks
von: Gambardella, Andrew, et al.
Veröffentlicht: (2024)
von: Gambardella, Andrew, et al.
Veröffentlicht: (2024)
Modular Arithmetic: Language Models Solve Math Digit by Digit
von: Baeumel, Tanja, et al.
Veröffentlicht: (2025)
von: Baeumel, Tanja, et al.
Veröffentlicht: (2025)
Leveraging Language Models and Bandit Algorithms to Drive Adoption of Battery-Electric Vehicles
von: Namikoshi, Keiichi, et al.
Veröffentlicht: (2024)
von: Namikoshi, Keiichi, et al.
Veröffentlicht: (2024)
Language Arithmetics: Towards Systematic Language Neuron Identification and Manipulation
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis
von: Yu, Zeping, et al.
Veröffentlicht: (2024)
von: Yu, Zeping, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
QuanTA: Efficient High-Rank Fine-Tuning of LLMs with Quantum-Informed Tensor Adaptation
von: Chen, Zhuo, et al.
Veröffentlicht: (2024) -
Active learning for photonic crystals
von: Lopez, Ryan, et al.
Veröffentlicht: (2026) -
On a Novel Application of Wasserstein-Procrustes for Unsupervised Cross-Lingual Learning
von: Ramírez, Guillem, et al.
Veröffentlicht: (2020) -
Data-Informed Global Sparseness in Attention Mechanisms for Deep Neural Networks
von: Rugina, Ileana, et al.
Veröffentlicht: (2020) -
QuACK: Accelerating Gradient-Based Quantum Optimization with Koopman Operator Learning
von: Luo, Di, et al.
Veröffentlicht: (2022)