A Synthetic Dataset for Personal Attribute Inference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yukhymenko, Hanna, Staab, Robin, Vero, Mark, Vechev, Martin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Recovered in Translation: Efficient Pipeline for Automated Translation of Benchmarks and Datasets
von: Yukhymenko, Hanna, et al.
Veröffentlicht: (2026)
von: Yukhymenko, Hanna, et al.
Veröffentlicht: (2026)
Private Attribute Inference from Images with Vision-Language Models
von: Tömekçe, Batuhan, et al.
Veröffentlicht: (2024)
von: Tömekçe, Batuhan, et al.
Veröffentlicht: (2024)
Beyond Memorization: Violating Privacy Via Inference with Large Language Models
von: Staab, Robin, et al.
Veröffentlicht: (2023)
von: Staab, Robin, et al.
Veröffentlicht: (2023)
Watch your steps: Dormant Adversarial Behaviors that Activate upon LLM Finetuning
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Back to the Drawing Board for Fair Representation Learning
von: Pouget, Angéline, et al.
Veröffentlicht: (2024)
von: Pouget, Angéline, et al.
Veröffentlicht: (2024)
Widening the Gap: Exploiting LLM Quantization via Outlier Injection
von: Zhan, Xiaohua, et al.
Veröffentlicht: (2026)
von: Zhan, Xiaohua, et al.
Veröffentlicht: (2026)
Large Language Models are Advanced Anonymizers
von: Staab, Robin, et al.
Veröffentlicht: (2024)
von: Staab, Robin, et al.
Veröffentlicht: (2024)
Pay Attention to the Triggers: Constructing Backdoors That Survive Distillation
von: De Muri, Giovanni, et al.
Veröffentlicht: (2025)
von: De Muri, Giovanni, et al.
Veröffentlicht: (2025)
Mind the Gap: A Practical Attack on GGUF Quantization
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
Fewer Weights, More Problems: A Practical Attack on LLM Pruning
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
COMPL-AI Framework: A Technical Interpretation and LLM Benchmarking Suite for the EU Artificial Intelligence Act
von: Guldimann, Philipp, et al.
Veröffentlicht: (2024)
von: Guldimann, Philipp, et al.
Veröffentlicht: (2024)
Exploiting LLM Quantization
von: Egashira, Kazuki, et al.
Veröffentlicht: (2024)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2024)
Ward: Provable RAG Dataset Inference via LLM Watermarks
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024)
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024)
Delay, Plateau, or Collapse: Evaluating the Impact of Systematic Verification Error on RLVR
von: Egashira, Kazuki, et al.
Veröffentlicht: (2026)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2026)
CuTS: Customizable Tabular Synthetic Data Generation
von: Vero, Mark, et al.
Veröffentlicht: (2023)
von: Vero, Mark, et al.
Veröffentlicht: (2023)
AutoBaxBuilder: Bootstrapping Code Security Benchmarking
von: von Arx, Tobias, et al.
Veröffentlicht: (2025)
von: von Arx, Tobias, et al.
Veröffentlicht: (2025)
Every Bit, Everywhere, All at Once: A Binomial Multibit LLM Watermark
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
Watermark Stealing in Large Language Models
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024)
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024)
BrokenMath: A Benchmark for Sycophancy in Theorem Proving with LLMs
von: Petrov, Ivo, et al.
Veröffentlicht: (2025)
von: Petrov, Ivo, et al.
Veröffentlicht: (2025)
A Unified Framework for LLM Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
Discovering Spoofing Attempts on Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
Watermarking Diffusion Language Models
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
BaxBench: Can LLMs Generate Correct and Secure Backends?
von: Vero, Mark, et al.
Veröffentlicht: (2025)
von: Vero, Mark, et al.
Veröffentlicht: (2025)
Instruction Tuning for Secure Code Generation
von: He, Jingxuan, et al.
Veröffentlicht: (2024)
von: He, Jingxuan, et al.
Veröffentlicht: (2024)
Fast Training Dataset Attribution via In-Context Learning
von: Fotouhi, Milad, et al.
Veröffentlicht: (2024)
von: Fotouhi, Milad, et al.
Veröffentlicht: (2024)
Evading Data Contamination Detection for Language Models is (too) Easy
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
Towards a Method for Synthetic Generation of Persons with Aphasia Transcripts
von: Pittman, Jason M., et al.
Veröffentlicht: (2025)
von: Pittman, Jason M., et al.
Veröffentlicht: (2025)
KodCode: A Diverse, Challenging, and Verifiable Synthetic Dataset for Coding
von: Xu, Zhangchen, et al.
Veröffentlicht: (2025)
von: Xu, Zhangchen, et al.
Veröffentlicht: (2025)
BgGPT 1.0: Extending English-centric LLMs to other languages
von: Alexandrov, Anton, et al.
Veröffentlicht: (2024)
von: Alexandrov, Anton, et al.
Veröffentlicht: (2024)
CRAFT Your Dataset: Task-Specific Synthetic Dataset Generation Through Corpus Retrieval and Augmentation
von: Ziegler, Ingo, et al.
Veröffentlicht: (2024)
von: Ziegler, Ingo, et al.
Veröffentlicht: (2024)
EPSVec: Efficient and Private Synthetic Data Generation via Dataset Vectors
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
TreeCut: A Synthetic Unanswerable Math Word Problem Dataset for LLM Hallucination Evaluation
von: Ouyang, Jialin
Veröffentlicht: (2025)
von: Ouyang, Jialin
Veröffentlicht: (2025)
Is Active Persona Inference Necessary for Aligning Small Models to Personal Preferences?
von: Tang, Zilu, et al.
Veröffentlicht: (2025)
von: Tang, Zilu, et al.
Veröffentlicht: (2025)
LLM Unlearning Without an Expert Curated Dataset
von: Zhu, Xiaoyuan, et al.
Veröffentlicht: (2025)
von: Zhu, Xiaoyuan, et al.
Veröffentlicht: (2025)
AttributionBench: How Hard is Automatic Attribution Evaluation?
von: Li, Yifei, et al.
Veröffentlicht: (2024)
von: Li, Yifei, et al.
Veröffentlicht: (2024)
DSTI at LLMs4OL 2024 Task A: Intrinsic versus extrinsic knowledge for type classification
von: Akl, Hanna Abi
Veröffentlicht: (2024)
von: Akl, Hanna Abi
Veröffentlicht: (2024)
Constructing Synthetic Instruction Datasets for Improving Reasoning in Domain-Specific LLMs: A Case Study in the Japanese Financial Domain
von: Okochi, Yuma, et al.
Veröffentlicht: (2026)
von: Okochi, Yuma, et al.
Veröffentlicht: (2026)
VITAL: A New Dataset for Benchmarking Pluralistic Alignment in Healthcare
von: Shetty, Anudeex, et al.
Veröffentlicht: (2025)
von: Shetty, Anudeex, et al.
Veröffentlicht: (2025)
FinNLI: Novel Dataset for Multi-Genre Financial Natural Language Inference Benchmarking
von: Magomere, Jabez, et al.
Veröffentlicht: (2025)
von: Magomere, Jabez, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Recovered in Translation: Efficient Pipeline for Automated Translation of Benchmarks and Datasets
von: Yukhymenko, Hanna, et al.
Veröffentlicht: (2026) -
Private Attribute Inference from Images with Vision-Language Models
von: Tömekçe, Batuhan, et al.
Veröffentlicht: (2024) -
Beyond Memorization: Violating Privacy Via Inference with Large Language Models
von: Staab, Robin, et al.
Veröffentlicht: (2023) -
Watch your steps: Dormant Adversarial Behaviors that Activate upon LLM Finetuning
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025) -
Back to the Drawing Board for Fair Representation Learning
von: Pouget, Angéline, et al.
Veröffentlicht: (2024)