Innovative tokenisation of structured data for LLM training
Fuente:
arXiv
Guardado en:
| Autores principales: | Karim, Kayvan, Batatia, Hani Ragab Hassen. Hadj |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reassessing feature-based Android malware detection in a contemporary context
por: Muzaffar, Ali, et al.
Publicado: (2023)
por: Muzaffar, Ali, et al.
Publicado: (2023)
ActDroid: An active learning framework for Android malware detection
por: Muzaffar, Ali, et al.
Publicado: (2024)
por: Muzaffar, Ali, et al.
Publicado: (2024)
Adversarial training with restricted data manipulation
por: Benfield, David, et al.
Publicado: (2025)
por: Benfield, David, et al.
Publicado: (2025)
Homomorphic WiSARDs: Efficient Weightless Neural Network training over encrypted data
por: Neumann, Leonardo, et al.
Publicado: (2024)
por: Neumann, Leonardo, et al.
Publicado: (2024)
How Feasible is Augmenting Fake Nodes with Learnable Features as a Counter-strategy against Link Stealing Attacks?
por: Mostafiz, Mir Imtiaz, et al.
Publicado: (2025)
por: Mostafiz, Mir Imtiaz, et al.
Publicado: (2025)
Backdoor Attacks on Discrete Graph Diffusion Models
por: Wang, Jiawen, et al.
Publicado: (2025)
por: Wang, Jiawen, et al.
Publicado: (2025)
LLM Dataset Inference: Did you train on my dataset?
por: Maini, Pratyush, et al.
Publicado: (2024)
por: Maini, Pratyush, et al.
Publicado: (2024)
Robust LLM safeguarding via refusal feature adversarial training
por: Yu, Lei, et al.
Publicado: (2024)
por: Yu, Lei, et al.
Publicado: (2024)
Selective Pre-training for Private Fine-tuning
por: Yu, Da, et al.
Publicado: (2023)
por: Yu, Da, et al.
Publicado: (2023)
Pre-training Differentially Private Models with Limited Public Data
por: Bu, Zhiqi, et al.
Publicado: (2024)
por: Bu, Zhiqi, et al.
Publicado: (2024)
Indiscriminate Data Poisoning Attacks on Pre-trained Feature Extractors
por: Lu, Yiwei, et al.
Publicado: (2024)
por: Lu, Yiwei, et al.
Publicado: (2024)
Augmenting Parameter-Efficient Pre-trained Language Models with Large Language Models
por: Anand, Saurabh, et al.
Publicado: (2026)
por: Anand, Saurabh, et al.
Publicado: (2026)
Privacy Backdoors: Enhancing Membership Inference through Poisoning Pre-trained Models
por: Wen, Yuxin, et al.
Publicado: (2024)
por: Wen, Yuxin, et al.
Publicado: (2024)
Mitigating Noise Detriment in Differentially Private Federated Learning with Model Pre-training
por: Jin, Huitong, et al.
Publicado: (2024)
por: Jin, Huitong, et al.
Publicado: (2024)
Accuracy Improvement in Differentially Private Logistic Regression: A Pre-training Approach
por: Hoseinpour, Mohammad, et al.
Publicado: (2023)
por: Hoseinpour, Mohammad, et al.
Publicado: (2023)
Continuous Multi-Task Pre-training for Malicious URL Detection and Webpage Classification
por: Li, Yujie, et al.
Publicado: (2024)
por: Li, Yujie, et al.
Publicado: (2024)
Rerouting LLM Routers
por: Shafran, Avital, et al.
Publicado: (2025)
por: Shafran, Avital, et al.
Publicado: (2025)
Power side-channel leakage localization through adversarial training of deep neural networks
por: Gammell, Jimmy, et al.
Publicado: (2024)
por: Gammell, Jimmy, et al.
Publicado: (2024)
Pre-trained Encoder Inference: Revealing Upstream Encoders In Downstream Machine Learning Services
por: Fu, Shaopeng, et al.
Publicado: (2024)
por: Fu, Shaopeng, et al.
Publicado: (2024)
AttackLLM: LLM-based Attack Pattern Generation for an Industrial Control System
por: Ahmed, Chuadhry Mujeeb
Publicado: (2025)
por: Ahmed, Chuadhry Mujeeb
Publicado: (2025)
Encryption-Friendly LLM Architecture
por: Rho, Donghwan, et al.
Publicado: (2024)
por: Rho, Donghwan, et al.
Publicado: (2024)
Log Probability Tracking of LLM APIs
por: Chauvin, Timothée, et al.
Publicado: (2025)
por: Chauvin, Timothée, et al.
Publicado: (2025)
Is The Watermarking Of LLM-Generated Code Robust?
por: Suresh, Tarun, et al.
Publicado: (2024)
por: Suresh, Tarun, et al.
Publicado: (2024)
Good-Enough LLM Obfuscation (GELO)
por: Belikov, Anatoly, et al.
Publicado: (2026)
por: Belikov, Anatoly, et al.
Publicado: (2026)
Secure Transfer Learning: Training Clean Models Against Backdoor in (Both) Pre-trained Encoders and Downstream Datasets
por: Zhang, Yechao, et al.
Publicado: (2025)
por: Zhang, Yechao, et al.
Publicado: (2025)
PAE MobiLLM: Privacy-Aware and Efficient LLM Fine-Tuning on the Mobile Device via Additive Side-Tuning
por: Yang, Xingke, et al.
Publicado: (2025)
por: Yang, Xingke, et al.
Publicado: (2025)
LLM-Generated Samples for Android Malware Detection
por: Rollinson, Nik, et al.
Publicado: (2025)
por: Rollinson, Nik, et al.
Publicado: (2025)
Cascade: Token-Sharded Private LLM Inference
por: Thomas, Rahul, et al.
Publicado: (2025)
por: Thomas, Rahul, et al.
Publicado: (2025)
LLM Fingerprinting via Semantically Conditioned Watermarks
por: Gloaguen, Thibaud, et al.
Publicado: (2025)
por: Gloaguen, Thibaud, et al.
Publicado: (2025)
Adversarial Contrastive Learning for LLM Quantization Attacks
por: Song, Dinghong, et al.
Publicado: (2026)
por: Song, Dinghong, et al.
Publicado: (2026)
Token-Efficient Change Detection in LLM APIs
por: Chauvin, Timothée, et al.
Publicado: (2026)
por: Chauvin, Timothée, et al.
Publicado: (2026)
Memory-Induced Tool-Drift in LLM Agents
por: Dabas, Mahavir, et al.
Publicado: (2026)
por: Dabas, Mahavir, et al.
Publicado: (2026)
Order of Magnitude Speedups for LLM Membership Inference
por: Zhang, Rongting, et al.
Publicado: (2024)
por: Zhang, Rongting, et al.
Publicado: (2024)
Proving membership in LLM pretraining data via data watermarks
por: Wei, Johnny Tian-Zheng, et al.
Publicado: (2024)
por: Wei, Johnny Tian-Zheng, et al.
Publicado: (2024)
LLM Watermarking Using Mixtures and Statistical-to-Computational Gaps
por: Abdalla, Pedro, et al.
Publicado: (2025)
por: Abdalla, Pedro, et al.
Publicado: (2025)
Fundamental Limitations in Pointwise Defences of LLM Finetuning APIs
por: Davies, Xander, et al.
Publicado: (2025)
por: Davies, Xander, et al.
Publicado: (2025)
Verifying LLM Inference to Detect Model Weight Exfiltration
por: Rinberg, Roy, et al.
Publicado: (2025)
por: Rinberg, Roy, et al.
Publicado: (2025)
Black-box Optimization of LLM Outputs by Asking for Directions
por: Zhang, Jie, et al.
Publicado: (2025)
por: Zhang, Jie, et al.
Publicado: (2025)
NeST: Neuron Selective Tuning for LLM Safety
por: Behrouzi, Sasha, et al.
Publicado: (2026)
por: Behrouzi, Sasha, et al.
Publicado: (2026)
AERO: Entropy-Guided Framework for Private LLM Inference
por: Jha, Nandan Kumar, et al.
Publicado: (2024)
por: Jha, Nandan Kumar, et al.
Publicado: (2024)
Ejemplares similares
-
Reassessing feature-based Android malware detection in a contemporary context
por: Muzaffar, Ali, et al.
Publicado: (2023) -
ActDroid: An active learning framework for Android malware detection
por: Muzaffar, Ali, et al.
Publicado: (2024) -
Adversarial training with restricted data manipulation
por: Benfield, David, et al.
Publicado: (2025) -
Homomorphic WiSARDs: Efficient Weightless Neural Network training over encrypted data
por: Neumann, Leonardo, et al.
Publicado: (2024) -
How Feasible is Augmenting Fake Nodes with Learnable Features as a Counter-strategy against Link Stealing Attacks?
por: Mostafiz, Mir Imtiaz, et al.
Publicado: (2025)