Arithmetic Without Algorithms: Language Models Solve Math With a Bag of Heuristics
Fuente:
arXiv
Guardado en:
| Autores principales: | Nikankin, Yaniv, Reusch, Anja, Mueller, Aaron, Belinkov, Yonatan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Same Task, Different Circuits: Disentangling Modality-Specific Mechanisms in VLMs
por: Nikankin, Yaniv, et al.
Publicado: (2025)
por: Nikankin, Yaniv, et al.
Publicado: (2025)
Reasoning Models Know What's Important, and Encode It in Their Activations
por: Nikankin, Yaniv, et al.
Publicado: (2026)
por: Nikankin, Yaniv, et al.
Publicado: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025)
por: Ashuach, Tomer, et al.
Publicado: (2025)
Position-aware Automatic Circuit Discovery
por: Haklay, Tal, et al.
Publicado: (2025)
por: Haklay, Tal, et al.
Publicado: (2025)
ReFACT: Updating Text-to-Image Models by Editing the Text Encoder
por: Arad, Dana, et al.
Publicado: (2023)
por: Arad, Dana, et al.
Publicado: (2023)
REVS: Unlearning Sensitive Information in Language Models via Rank Editing in the Vocabulary Space
por: Ashuach, Tomer, et al.
Publicado: (2024)
por: Ashuach, Tomer, et al.
Publicado: (2024)
Math Natural Language Inference: this should be easy!
por: de Paiva, Valeria, et al.
Publicado: (2025)
por: de Paiva, Valeria, et al.
Publicado: (2025)
Are formal and functional linguistic mechanisms dissociated in language models?
por: Hanna, Michael, et al.
Publicado: (2025)
por: Hanna, Michael, et al.
Publicado: (2025)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
por: Orgad, Hadas, et al.
Publicado: (2024)
por: Orgad, Hadas, et al.
Publicado: (2024)
Enhancing OCR for Sino-Vietnamese Language Processing via Fine-tuned PaddleOCRv5
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
Distinguishing Ignorance from Error in LLM Hallucinations
por: Simhi, Adi, et al.
Publicado: (2024)
por: Simhi, Adi, et al.
Publicado: (2024)
Constructing Benchmarks and Interventions for Combating Hallucinations in LLMs
por: Simhi, Adi, et al.
Publicado: (2024)
por: Simhi, Adi, et al.
Publicado: (2024)
Multilingual Multi-Label Emotion Classification at Scale with Synthetic Data
por: Borisov, Vadim
Publicado: (2026)
por: Borisov, Vadim
Publicado: (2026)
Transparent but Powerful: Explainability, Accuracy, and Generalizability in ADHD Detection from Social Media Data
por: Wiechmann, D., et al.
Publicado: (2024)
por: Wiechmann, D., et al.
Publicado: (2024)
Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms
por: Hanna, Michael, et al.
Publicado: (2024)
por: Hanna, Michael, et al.
Publicado: (2024)
Fast Quiet-STaR: Thinking Without Thought Tokens
por: Huang, Wei, et al.
Publicado: (2025)
por: Huang, Wei, et al.
Publicado: (2025)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
por: Schneider, Felix, et al.
Publicado: (2026)
por: Schneider, Felix, et al.
Publicado: (2026)
Pitfalls in Evaluating Interpretability Agents
por: Haklay, Tal, et al.
Publicado: (2026)
por: Haklay, Tal, et al.
Publicado: (2026)
Approaches to Semantic Textual Similarity in Slovak Language: From Algorithms to Transformers
por: Radosky, Lukas, et al.
Publicado: (2026)
por: Radosky, Lukas, et al.
Publicado: (2026)
Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer
por: Simhi, Adi, et al.
Publicado: (2025)
por: Simhi, Adi, et al.
Publicado: (2025)
Growing a Tail: Increasing Output Diversity in Large Language Models
por: Shur-Ofry, Michal, et al.
Publicado: (2024)
por: Shur-Ofry, Michal, et al.
Publicado: (2024)
Evaluating Pixel Language Models on Non-Standardized Languages
por: Muñoz-Ortiz, Alberto, et al.
Publicado: (2024)
por: Muñoz-Ortiz, Alberto, et al.
Publicado: (2024)
Logits-Constrained Framework with RoBERTa for Ancient Chinese NER
por: Hua, Wenjie, et al.
Publicado: (2025)
por: Hua, Wenjie, et al.
Publicado: (2025)
What is Wrong with Language Models that Can Not Tell a Story?
por: Yamshchikov, Ivan P., et al.
Publicado: (2022)
por: Yamshchikov, Ivan P., et al.
Publicado: (2022)
ECLAIR: Enhanced Clarification for Interactive Responses in an Enterprise AI Assistant
por: Murzaku, John, et al.
Publicado: (2025)
por: Murzaku, John, et al.
Publicado: (2025)
A Dataset for Metaphor Detection in Early Medieval Hebrew Poetry
por: Toker, Michael, et al.
Publicado: (2024)
por: Toker, Michael, et al.
Publicado: (2024)
Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness
por: Ashuach, Tomer, et al.
Publicado: (2026)
por: Ashuach, Tomer, et al.
Publicado: (2026)
The Superalignment of Superhuman Intelligence with Large Language Models
por: Huang, Minlie, et al.
Publicado: (2024)
por: Huang, Minlie, et al.
Publicado: (2024)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
por: Liu, Aiwei, et al.
Publicado: (2023)
por: Liu, Aiwei, et al.
Publicado: (2023)
Exploring State Tracking Capabilities of Large Language Models
por: Rezaee, Kiamehr, et al.
Publicado: (2025)
por: Rezaee, Kiamehr, et al.
Publicado: (2025)
A Survey of Text Watermarking in the Era of Large Language Models
por: Liu, Aiwei, et al.
Publicado: (2023)
por: Liu, Aiwei, et al.
Publicado: (2023)
Distilling Large Language Models for Efficient Clinical Information Extraction
por: Vedula, Karthik S., et al.
Publicado: (2024)
por: Vedula, Karthik S., et al.
Publicado: (2024)
Adaptive Steering and Remasking for Safe Generation in Diffusion Language Models
por: Lee, Yejin, et al.
Publicado: (2026)
por: Lee, Yejin, et al.
Publicado: (2026)
Towards Effective and Efficient Continual Pre-training of Large Language Models
por: Chen, Jie, et al.
Publicado: (2024)
por: Chen, Jie, et al.
Publicado: (2024)
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
por: Park, Seungcheol, et al.
Publicado: (2023)
por: Park, Seungcheol, et al.
Publicado: (2023)
ManagerBench: Evaluating the Safety-Pragmatism Trade-off in Autonomous LLMs
por: Simhi, Adi, et al.
Publicado: (2025)
por: Simhi, Adi, et al.
Publicado: (2025)
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
por: Park, Seungcheol, et al.
Publicado: (2025)
por: Park, Seungcheol, et al.
Publicado: (2025)
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
por: Park, Seungcheol, et al.
Publicado: (2025)
por: Park, Seungcheol, et al.
Publicado: (2025)
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
por: Bayram, M. Ali, et al.
Publicado: (2024)
por: Bayram, M. Ali, et al.
Publicado: (2024)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
Ejemplares similares
-
Same Task, Different Circuits: Disentangling Modality-Specific Mechanisms in VLMs
por: Nikankin, Yaniv, et al.
Publicado: (2025) -
Reasoning Models Know What's Important, and Encode It in Their Activations
por: Nikankin, Yaniv, et al.
Publicado: (2026) -
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025) -
Position-aware Automatic Circuit Discovery
por: Haklay, Tal, et al.
Publicado: (2025) -
ReFACT: Updating Text-to-Image Models by Editing the Text Encoder
por: Arad, Dana, et al.
Publicado: (2023)