Hammer: Robust Function-Calling for On-Device Language Models via Function Masking
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Qiqiang, Wen, Muning, Peng, Qiuying, Nie, Guanyu, Liao, Junwei, Wang, Jun, Mo, Xiaoyun, Zhou, Jiamu, Cheng, Cheng, Zhao, Yin, Zhang, Weinan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HammerBench: Fine-Grained Function-Calling Evaluation in Real Mobile Device Scenarios
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
Online-Optimized RAG for Tool Use and Function Calling
by: Pan, Yu, et al.
Published: (2025)
by: Pan, Yu, et al.
Published: (2025)
Efficient Function Orchestration for Large Language Models
by: Liu, Xiaoxia, et al.
Published: (2025)
by: Liu, Xiaoxia, et al.
Published: (2025)
Seeing is Believing: Vision-driven Non-crash Functional Bug Detection for Mobile Apps
by: Liu, Zhe, et al.
Published: (2024)
by: Liu, Zhe, et al.
Published: (2024)
Vision to Specification: Automating the Transition from Conceptual Features to Functional Requirements
by: Lian, Xiaoli, et al.
Published: (2025)
by: Lian, Xiaoli, et al.
Published: (2025)
Boosting Pointer Analysis With LLM-Enhanced Allocation Function Detection
by: Cheng, Baijun, et al.
Published: (2025)
by: Cheng, Baijun, et al.
Published: (2025)
Digging Into the Internal: Causality-Based Analysis of LLM Function Calling
by: Ji, Zhenlan, et al.
Published: (2025)
by: Ji, Zhenlan, et al.
Published: (2025)
ReF Decompile: Relabeling and Function Call Enhanced Decompile
by: Feng, Yunlong, et al.
Published: (2025)
by: Feng, Yunlong, et al.
Published: (2025)
CallNavi, A Challenge and Empirical Study on LLM Function Calling and Routing
by: Song, Yewei, et al.
Published: (2025)
by: Song, Yewei, et al.
Published: (2025)
Semantic-Enhanced Indirect Call Analysis with Large Language Models
by: Cheng, Baijun, et al.
Published: (2024)
by: Cheng, Baijun, et al.
Published: (2024)
Adaptable and Precise: Enterprise-Scenario LLM Function-Calling Capability Training Pipeline
by: Zeng, Guancheng, et al.
Published: (2024)
by: Zeng, Guancheng, et al.
Published: (2024)
ADC: Enhancing Function Calling Via Adversarial Datasets and Code Line-Level Feedback
by: Zhang, Wei, et al.
Published: (2024)
by: Zhang, Wei, et al.
Published: (2024)
GeoJSON Agents:A Multi-Agent LLM Architecture for Geospatial Analysis-Function Calling vs Code Generation
by: Luo, Qianqian, et al.
Published: (2025)
by: Luo, Qianqian, et al.
Published: (2025)
$μ$Drive: User-Controlled Autonomous Driving
by: Wang, Kun, et al.
Published: (2024)
by: Wang, Kun, et al.
Published: (2024)
EFACT: an External Function Auto-Completion Tool to Strengthen Static Binary Lifting
by: Zhang, Yilei, et al.
Published: (2024)
by: Zhang, Yilei, et al.
Published: (2024)
PredicateFix: Repairing Static Analysis Alerts with Bridging Predicates
by: Xiao, Yuan-An, et al.
Published: (2025)
by: Xiao, Yuan-An, et al.
Published: (2025)
FastFI: Enhancing API Call-Site Robustness in Microservice-Based Systems with Fault Injection
by: Tan, Yuzhen, et al.
Published: (2026)
by: Tan, Yuzhen, et al.
Published: (2026)
P3: A Policy-Driven, Pace-Adaptive, and Diversity-Promoted Framework for data pruning in LLM Training
by: Yang, Yingxuan, et al.
Published: (2024)
by: Yang, Yingxuan, et al.
Published: (2024)
SlsReuse: LLM-Powered Serverless Function Reuse
by: Wen, Jinfeng, et al.
Published: (2025)
by: Wen, Jinfeng, et al.
Published: (2025)
CodeChemist: Functional Knowledge Transfer for Low-Resource Code Generation via Test-Time Scaling
by: Wang, Kaixin, et al.
Published: (2025)
by: Wang, Kaixin, et al.
Published: (2025)
FuncDroid: Towards Inter-Functional Flows for Comprehensive Mobile App GUI Testing
by: He, Jinlong, et al.
Published: (2026)
by: He, Jinlong, et al.
Published: (2026)
A Test Suite for Efficient Robustness Evaluation of Face Recognition Systems
by: Zhang, Ruihan, et al.
Published: (2025)
by: Zhang, Ruihan, et al.
Published: (2025)
Context-Aware Functional Test Generation via Business Logic Extraction and Adaptation
by: Zhang, Yakun, et al.
Published: (2026)
by: Zhang, Yakun, et al.
Published: (2026)
BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
by: Zhuo, Terry Yue, et al.
Published: (2024)
by: Zhuo, Terry Yue, et al.
Published: (2024)
Bridging the Gap Between Domain-specific Frameworks and Multiple Hardware Devices
by: Wen, Xu, et al.
Published: (2024)
by: Wen, Xu, et al.
Published: (2024)
Function-to-Style Guidance of LLMs for Code Translation
by: Zhang, Longhui, et al.
Published: (2025)
by: Zhang, Longhui, et al.
Published: (2025)
AUITestAgent: Automatic Requirements Oriented GUI Function Testing
by: Hu, Yongxiang, et al.
Published: (2024)
by: Hu, Yongxiang, et al.
Published: (2024)
How LLMs Aid in UML Modeling: An Exploratory Study with Novice Analysts
by: Wang, Beian, et al.
Published: (2024)
by: Wang, Beian, et al.
Published: (2024)
Small Models, Big Tasks: An Exploratory Empirical Study on Small Language Models for Function Calling
by: Kavathekar, Ishan, et al.
Published: (2025)
by: Kavathekar, Ishan, et al.
Published: (2025)
Scalable and Precise Application-Centered Call Graph Construction for Python
by: Huang, Kaifeng, et al.
Published: (2023)
by: Huang, Kaifeng, et al.
Published: (2023)
MASKDROID: Robust Android Malware Detection with Masked Graph Representations
by: Zheng, Jingnan, et al.
Published: (2024)
by: Zheng, Jingnan, et al.
Published: (2024)
ToolRegistry: A Protocol-Agnostic Tool Management Library for Function-Calling LLMs
by: Ding, Peng, et al.
Published: (2025)
by: Ding, Peng, et al.
Published: (2025)
Game Rewards Vulnerabilities: Software Vulnerability Detection with Zero-Sum Game and Prototype Learning
by: Wen, Xin-Cheng, et al.
Published: (2024)
by: Wen, Xin-Cheng, et al.
Published: (2024)
EigenData: A Self-Evolving Multi-Agent Platform for Function-Calling Data Synthesis, Auditing, and Repair
by: Chen, Jiaao, et al.
Published: (2026)
by: Chen, Jiaao, et al.
Published: (2026)
CodeScore-R: An Automated Robustness Metric for Assessing the FunctionalCorrectness of Code Synthesis
by: Yang, Guang, et al.
Published: (2024)
by: Yang, Guang, et al.
Published: (2024)
SPOLRE: Semantic Preserving Object Layout Reconstruction for Image Captioning System Testing
by: Liu, Yi, et al.
Published: (2024)
by: Liu, Yi, et al.
Published: (2024)
Detect--Repair--Verify for LLM-Generated Code: A Multi-Language, Multi-Granularity Empirical Study
by: Cheng, Cheng
Published: (2026)
by: Cheng, Cheng
Published: (2026)
Detect Repair Verify for Securing LLM Generated Code: A Multi-Language Empirical Study
by: Cheng, Cheng
Published: (2026)
by: Cheng, Cheng
Published: (2026)
Unveiling Security Weaknesses in Autonomous Driving Systems: An In-Depth Empirical Study
by: Cheng, Wenyuan, et al.
Published: (2025)
by: Cheng, Wenyuan, et al.
Published: (2025)
ScaleCall -- Agentic Tool Calling at Scale for Fintech: Challenges, Methods, and Deployment Insights
by: Osuagwu, Richard, et al.
Published: (2025)
by: Osuagwu, Richard, et al.
Published: (2025)
Similar Items
-
HammerBench: Fine-Grained Function-Calling Evaluation in Real Mobile Device Scenarios
by: Wang, Jun, et al.
Published: (2024) -
Online-Optimized RAG for Tool Use and Function Calling
by: Pan, Yu, et al.
Published: (2025) -
Efficient Function Orchestration for Large Language Models
by: Liu, Xiaoxia, et al.
Published: (2025) -
Seeing is Believing: Vision-driven Non-crash Functional Bug Detection for Mobile Apps
by: Liu, Zhe, et al.
Published: (2024) -
Vision to Specification: Automating the Transition from Conceptual Features to Functional Requirements
by: Lian, Xiaoli, et al.
Published: (2025)