CmdCaliper: A Semantic-Aware Command-Line Embedding Model and Dataset for Security Research
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Sian-Yao, Yang, Cheng-Lin, Lin, Che-Yu, Huang, Chun-Ying |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Eyes-on-Me: Scalable RAG Poisoning through Transferable Attention-Steering Attractors
by: Chen, Yen-Shan, et al.
Published: (2025)
by: Chen, Yen-Shan, et al.
Published: (2025)
FinNuE: Exposing the Risks of Using BERTScore for Numerical Semantic Evaluation in Finance
by: Huang, Yu-Shiang, et al.
Published: (2025)
by: Huang, Yu-Shiang, et al.
Published: (2025)
TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories
by: Chen, Yen-Shan, et al.
Published: (2026)
by: Chen, Yen-Shan, et al.
Published: (2026)
Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
by: Zhang, Yanzhao, et al.
Published: (2025)
by: Zhang, Yanzhao, et al.
Published: (2025)
I Think, Therefore I am: Benchmarking Awareness of Large Language Models Using AwareBench
by: Li, Yuan, et al.
Published: (2024)
by: Li, Yuan, et al.
Published: (2024)
Transferable Embedding Inversion Attack: Uncovering Privacy Risks in Text Embeddings without Model Queries
by: Huang, Yu-Hsiang, et al.
Published: (2024)
by: Huang, Yu-Hsiang, et al.
Published: (2024)
Beyond Semantics: An Evidential Reasoning-Aware Multi-View Learning Framework for Trustworthy Mental Health Prediction
by: Ruan, Yucheng, et al.
Published: (2026)
by: Ruan, Yucheng, et al.
Published: (2026)
Intrusion Detection at Scale with the Assistance of a Command-line Language Model
by: Lin, Jiongliang, et al.
Published: (2024)
by: Lin, Jiongliang, et al.
Published: (2024)
SemPA: Improving Sentence Embeddings of Large Language Models through Semantic Preference Alignment
by: Chen, Ziyang, et al.
Published: (2026)
by: Chen, Ziyang, et al.
Published: (2026)
Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealing
by: Xie, Pei-Xi, et al.
Published: (2026)
by: Xie, Pei-Xi, et al.
Published: (2026)
Semantic Similarity Matching for Patent Documents Using Ensemble BERT-related Model and Novel Text Processing Method
by: Yu, Liqiang, et al.
Published: (2024)
by: Yu, Liqiang, et al.
Published: (2024)
Advancing Airport Tower Command Recognition: Integrating Squeeze-and-Excitation and Broadcasted Residual Learning
by: Lin, Yuanxi, et al.
Published: (2024)
by: Lin, Yuanxi, et al.
Published: (2024)
Improving Speech Emotion Recognition in Under-Resourced Languages via Speech-to-Speech Translation with Bootstrapping Data Selection
by: Lin, Hsi-Che, et al.
Published: (2024)
by: Lin, Hsi-Che, et al.
Published: (2024)
Language Confusion Gate: Language-Aware Decoding Through Model Self-Distillation
by: Zhang, Collin, et al.
Published: (2025)
by: Zhang, Collin, et al.
Published: (2025)
PRISM: A Framework for Producing Interpretable Political Bias Embeddings with Political-Aware Cross-Encoder
by: Sun, Yiqun, et al.
Published: (2025)
by: Sun, Yiqun, et al.
Published: (2025)
EmbedGrad: Gradient-Based Prompt Optimization in Embedding Space for Large Language Models
by: Hou, Xiaoming, et al.
Published: (2025)
by: Hou, Xiaoming, et al.
Published: (2025)
CFEVER: A Chinese Fact Extraction and VERification Dataset
by: Lin, Ying-Jia, et al.
Published: (2024)
by: Lin, Ying-Jia, et al.
Published: (2024)
Layer-Aware Task Arithmetic: Disentangling Task-Specific and Instruction-Following Knowledge
by: Chen, Yan-Lun, et al.
Published: (2025)
by: Chen, Yan-Lun, et al.
Published: (2025)
ContextCache: Context-Aware Semantic Cache for Multi-Turn Queries in Large Language Models
by: Yan, Jianxin, et al.
Published: (2025)
by: Yan, Jianxin, et al.
Published: (2025)
Let LLMs Speak Embedding Languages: Generative Text Embeddings via Iterative Contrastive Refinement
by: Tsai, Yu-Che, et al.
Published: (2025)
by: Tsai, Yu-Che, et al.
Published: (2025)
Listen and Speak Fairly: A Study on Semantic Gender Bias in Speech Integrated Large Language Models
by: Lin, Yi-Cheng, et al.
Published: (2024)
by: Lin, Yi-Cheng, et al.
Published: (2024)
Semantic Alignment across Ancient Egyptian Language Stages via Normalization-Aware Multitask Learning
by: Huang, He
Published: (2026)
by: Huang, He
Published: (2026)
PEneo: Unifying Line Extraction, Line Grouping, and Entity Linking for End-to-end Document Pair Extraction
by: Lin, Zening, et al.
Published: (2024)
by: Lin, Zening, et al.
Published: (2024)
A General Framework for Producing Interpretable Semantic Text Embeddings
by: Sun, Yiqun, et al.
Published: (2024)
by: Sun, Yiqun, et al.
Published: (2024)
The Proxy Presumption: From Semantic Embeddings to Valid Social Measures
by: Li, Baishi, et al.
Published: (2026)
by: Li, Baishi, et al.
Published: (2026)
Legal Documents Drafting with Fine-Tuned Pre-Trained Large Language Model
by: Lin, Chun-Hsien, et al.
Published: (2024)
by: Lin, Chun-Hsien, et al.
Published: (2024)
MI-Fuse: Label Fusion for Unsupervised Domain Adaptation with Closed-Source Large-Audio Language Model
by: Huang, Hsiao-Ying, et al.
Published: (2025)
by: Huang, Hsiao-Ying, et al.
Published: (2025)
The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation
by: Tsai, Yun-Shao, et al.
Published: (2026)
by: Tsai, Yun-Shao, et al.
Published: (2026)
TCDA: Thread-Constrained Discourse-Aware Modeling for Conversational Sentiment Quadruple Analysis
by: Li, Xinran, et al.
Published: (2026)
by: Li, Xinran, et al.
Published: (2026)
Remask, Don't Replace: Token-to-Mask Refinement in Diffusion Large Language Models
by: Yao, Lin
Published: (2026)
by: Yao, Lin
Published: (2026)
Decoding in Order-Agnostic Language Models: Chain-Rule Deviation and Uniform Spreading
by: Yao, Lin
Published: (2026)
by: Yao, Lin
Published: (2026)
CASE -- Condition-Aware Sentence Embeddings for Conditional Semantic Textual Similarity Measurement
by: Zhang, Gaifan, et al.
Published: (2025)
by: Zhang, Gaifan, et al.
Published: (2025)
CCI4.0: A Bilingual Pretraining Dataset for Enhancing Reasoning in Large Language Models
by: Liu, Guang, et al.
Published: (2025)
by: Liu, Guang, et al.
Published: (2025)
From the New World of Word Embeddings: A Comparative Study of Small-World Lexico-Semantic Networks in LLMs
by: Liu, Zhu, et al.
Published: (2025)
by: Liu, Zhu, et al.
Published: (2025)
Embedding-Informed Adaptive Retrieval-Augmented Generation of Large Language Models
by: Huang, Chengkai, et al.
Published: (2024)
by: Huang, Chengkai, et al.
Published: (2024)
Building Knowledge-Grounded Dialogue Systems with Graph-Based Semantic Modeling
by: Yang, Yizhe, et al.
Published: (2022)
by: Yang, Yizhe, et al.
Published: (2022)
Editing the Mind of Giants: An In-Depth Exploration of Pitfalls of Knowledge Editing in Large Language Models
by: Hsueh, Cheng-Hsun, et al.
Published: (2024)
by: Hsueh, Cheng-Hsun, et al.
Published: (2024)
CVLUE: A New Benchmark Dataset for Chinese Vision-Language Understanding Evaluation
by: Wang, Yuxuan, et al.
Published: (2024)
by: Wang, Yuxuan, et al.
Published: (2024)
KaLM-Embedding-V2: Superior Training Techniques and Data Inspire A Versatile Embedding Model
by: Zhao, Xinping, et al.
Published: (2025)
by: Zhao, Xinping, et al.
Published: (2025)
Command A: An Enterprise-Ready Large Language Model
by: Cohere, Team, et al.
Published: (2025)
by: Cohere, Team, et al.
Published: (2025)
Similar Items
-
Eyes-on-Me: Scalable RAG Poisoning through Transferable Attention-Steering Attractors
by: Chen, Yen-Shan, et al.
Published: (2025) -
FinNuE: Exposing the Risks of Using BERTScore for Numerical Semantic Evaluation in Finance
by: Huang, Yu-Shiang, et al.
Published: (2025) -
TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories
by: Chen, Yen-Shan, et al.
Published: (2026) -
Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
by: Zhang, Yanzhao, et al.
Published: (2025) -
I Think, Therefore I am: Benchmarking Awareness of Large Language Models Using AwareBench
by: Li, Yuan, et al.
Published: (2024)