Integrating External Tools with Large Language Models to Improve Accuracy
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Niketan, Nripesh, Batatia, Hadj |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-Model Synthetic Training for Mission-Critical Small Language Models
von: Platt, Nolan, et al.
Veröffentlicht: (2025)
von: Platt, Nolan, et al.
Veröffentlicht: (2025)
The Geometry of Persona: Disentangling Personality from Reasoning in Large Language Models
von: Wang, Zhixiang
Veröffentlicht: (2025)
von: Wang, Zhixiang
Veröffentlicht: (2025)
Towards Ontology-Enhanced Representation Learning for Large Language Models
von: Ronzano, Francesco, et al.
Veröffentlicht: (2024)
von: Ronzano, Francesco, et al.
Veröffentlicht: (2024)
Word Overuse and Alignment in Large Language Models: The Influence of Learning from Human Feedback
von: Juzek, Tom S., et al.
Veröffentlicht: (2025)
von: Juzek, Tom S., et al.
Veröffentlicht: (2025)
ProbeScale: Probing Analysis to Optimize Neural Scaling Laws for Efficient Small Language Model Inference
von: Das, Sourav
Veröffentlicht: (2026)
von: Das, Sourav
Veröffentlicht: (2026)
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
Hopscotch: Discovering and Skipping Redundancies in Language Models
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
von: Eyceoz, Mustafa, et al.
Veröffentlicht: (2025)
Improving Commonsense Bias Classification by Mitigating the Influence of Demographic Terms
von: Lee, JinKyu, et al.
Veröffentlicht: (2024)
von: Lee, JinKyu, et al.
Veröffentlicht: (2024)
A Confidence-Diversity Framework for Calibrating AI Judgement in Accessible Qualitative Coding Tasks
von: Zhao, Zhilong, et al.
Veröffentlicht: (2025)
von: Zhao, Zhilong, et al.
Veröffentlicht: (2025)
TSDS: Data Selection for Task-Specific Model Finetuning
von: Liu, Zifan, et al.
Veröffentlicht: (2024)
von: Liu, Zifan, et al.
Veröffentlicht: (2024)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
von: Yang, Yibo
Veröffentlicht: (2025)
von: Yang, Yibo
Veröffentlicht: (2025)
Causally Grounded Mechanistic Interpretability for LLMs with Faithful Natural-Language Explanations
von: Mahale, Ajay Pravin
Veröffentlicht: (2026)
von: Mahale, Ajay Pravin
Veröffentlicht: (2026)
Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs
von: Kaiser, Daniel, et al.
Veröffentlicht: (2026)
von: Kaiser, Daniel, et al.
Veröffentlicht: (2026)
DRO-InstructZero: Distributionally Robust Prompt Optimization for Large Language Models
von: Li, Yangyang
Veröffentlicht: (2025)
von: Li, Yangyang
Veröffentlicht: (2025)
No Memorization, No Detection: Output Distribution-Based Contamination Detection in Small Language Models
von: Sela, Omer
Veröffentlicht: (2026)
von: Sela, Omer
Veröffentlicht: (2026)
HEFT: A Coarse-to-Fine Hierarchy for Enhancing the Efficiency and Accuracy of Language Model Reasoning
von: Hill, Brennen
Veröffentlicht: (2025)
von: Hill, Brennen
Veröffentlicht: (2025)
Hybrid Gated Flow (HGF): Stabilizing 1.58-bit LLMs via Selective Low-Rank Correction
von: Pizzo, David Alejandro Trejo
Veröffentlicht: (2026)
von: Pizzo, David Alejandro Trejo
Veröffentlicht: (2026)
Unsolvability Ceiling in Multi-LLM Routing: An Empirical Study of Evaluation Artifacts
von: Garg, Saloni, et al.
Veröffentlicht: (2026)
von: Garg, Saloni, et al.
Veröffentlicht: (2026)
Measuring and curing reasoning rigidity: from decorative chain-of-thought to genuine faithfulness
von: Basu, Abhinaba, et al.
Veröffentlicht: (2026)
von: Basu, Abhinaba, et al.
Veröffentlicht: (2026)
Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
mHC-SSM: Manifold-Constrained Hyper-Connections for State Space Language Models with Stream-Specialized Adapters
von: Mutlu, Abdulvahap, et al.
Veröffentlicht: (2026)
von: Mutlu, Abdulvahap, et al.
Veröffentlicht: (2026)
AIPsy-Affect: A Keyword-Free Clinical Stimulus Battery for Mechanistic Interpretability of Emotion in Language Models
von: Keeman, Michael
Veröffentlicht: (2026)
von: Keeman, Michael
Veröffentlicht: (2026)
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth
von: Henry, James
Veröffentlicht: (2026)
von: Henry, James
Veröffentlicht: (2026)
Fine-tuning of Large Language Models for Constituency Parsing Using a Sequence to Sequence Approach
von: Delgado, Francisco Jose Cortes, et al.
Veröffentlicht: (2025)
von: Delgado, Francisco Jose Cortes, et al.
Veröffentlicht: (2025)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
The Trilemma of Truth in Large Language Models
von: Savcisens, Germans, et al.
Veröffentlicht: (2025)
von: Savcisens, Germans, et al.
Veröffentlicht: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference
von: Mathew, Aby Mammen
Veröffentlicht: (2026)
von: Mathew, Aby Mammen
Veröffentlicht: (2026)
Monotonicity as an Architectural Bias for Robust Language Models
von: Cooper, Patrick, et al.
Veröffentlicht: (2026)
von: Cooper, Patrick, et al.
Veröffentlicht: (2026)
Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2026)
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2026)
Sliced-Wasserstein Distribution Alignment Loss Improves the Ultra-Low-Bit Quantization of Large Language Models
von: Cao, Deyu, et al.
Veröffentlicht: (2026)
von: Cao, Deyu, et al.
Veröffentlicht: (2026)
Mechanistic Analysis of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
von: Imanov, Olaf Yunus Laitinen
Veröffentlicht: (2026)
von: Imanov, Olaf Yunus Laitinen
Veröffentlicht: (2026)
ReFactor GNNs: Revisiting Factorisation-based Models from a Message-Passing Perspective
von: Chen, Yihong, et al.
Veröffentlicht: (2022)
von: Chen, Yihong, et al.
Veröffentlicht: (2022)
Detecting Sleeper Agents in Large Language Models via Semantic Drift Analysis
von: Zanbaghi, Shahin, et al.
Veröffentlicht: (2025)
von: Zanbaghi, Shahin, et al.
Veröffentlicht: (2025)
Semantic Convergence: Investigating Shared Representations Across Scaled LLMs
von: Son, Daniel, et al.
Veröffentlicht: (2025)
von: Son, Daniel, et al.
Veröffentlicht: (2025)
lmfaoooo at SemEval-2026 Task 1: Humor Is an Audience. Preference Modeling for Constrained Humor Generation
von: Tikhonov, Alexey, et al.
Veröffentlicht: (2026)
von: Tikhonov, Alexey, et al.
Veröffentlicht: (2026)
Tool-Genesis: A Task-Driven Tool Creation Benchmark for Self-Evolving Language Agent
von: Xia, Bowei, et al.
Veröffentlicht: (2026)
von: Xia, Bowei, et al.
Veröffentlicht: (2026)
Challenges and Applications of Large Language Models: A Comparison of GPT and DeepSeek family of models
von: Sharma, Shubham, et al.
Veröffentlicht: (2025)
von: Sharma, Shubham, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multi-Model Synthetic Training for Mission-Critical Small Language Models
von: Platt, Nolan, et al.
Veröffentlicht: (2025) -
The Geometry of Persona: Disentangling Personality from Reasoning in Large Language Models
von: Wang, Zhixiang
Veröffentlicht: (2025) -
Towards Ontology-Enhanced Representation Learning for Large Language Models
von: Ronzano, Francesco, et al.
Veröffentlicht: (2024) -
Word Overuse and Alignment in Large Language Models: The Influence of Learning from Human Feedback
von: Juzek, Tom S., et al.
Veröffentlicht: (2025) -
ProbeScale: Probing Analysis to Optimize Neural Scaling Laws for Efficient Small Language Model Inference
von: Das, Sourav
Veröffentlicht: (2026)