Saved in:
| Main Authors: | Ramesh, Krithika, Smolyak, Daniel, Zhao, Zihao, Gandhi, Nupoor, Agarwal, Ritu, Bjarnadóttir, Margrét, Field, Anjalie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.07229 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Differentially Private Synthetic Data Generation in High-Stakes Domains
by: Ramesh, Krithika, et al.
Published: (2024)
by: Ramesh, Krithika, et al.
Published: (2024)
Improving Equity in Health Modeling with GPT4-Turbo Generated Synthetic Data: A Comparative Study
by: Smolyak, Daniel, et al.
Published: (2024)
by: Smolyak, Daniel, et al.
Published: (2024)
Controlled Generation for Private Synthetic Text
by: Zhao, Zihao, et al.
Published: (2025)
by: Zhao, Zihao, et al.
Published: (2025)
Maximizing Predictive Performance for Small Subgroups: Functionally Adaptive Interaction Regularization (FAIR)
by: Smolyak, Daniel, et al.
Published: (2024)
by: Smolyak, Daniel, et al.
Published: (2024)
How do we measure privacy in text? A survey of text anonymization metrics
by: Ren, Yaxuan, et al.
Published: (2025)
by: Ren, Yaxuan, et al.
Published: (2025)
Beyond Text: Characterizing Domain Expert Needs in Document Research
by: Gururaja, Sireesh, et al.
Published: (2025)
by: Gururaja, Sireesh, et al.
Published: (2025)
The Effect of Document Selection on Query-focused Text Analysis
by: Rangreji, Sandesh S, et al.
Published: (2026)
by: Rangreji, Sandesh S, et al.
Published: (2026)
AdaProb: Efficient Machine Unlearning via Adaptive Probability
by: Zhao, Zihao, et al.
Published: (2024)
by: Zhao, Zihao, et al.
Published: (2024)
SynthEval: A Framework for Detailed Utility and Privacy Evaluation of Tabular Synthetic Data
by: Lautrup, Anton Danholt, et al.
Published: (2024)
by: Lautrup, Anton Danholt, et al.
Published: (2024)
What Makes a Good Response? An Empirical Analysis of Quality in Qualitative Interviews
by: Ivey, Jonathan, et al.
Published: (2026)
by: Ivey, Jonathan, et al.
Published: (2026)
Does Local News Stay Local?: Online Content Shifts in Sinclair-Acquired Stations
by: Wanner, Miriam, et al.
Published: (2025)
by: Wanner, Miriam, et al.
Published: (2025)
HICode: Hierarchical Inductive Coding with LLMs
by: Zhong, Mian, et al.
Published: (2025)
by: Zhong, Mian, et al.
Published: (2025)
Better Synthetic Data by Retrieving and Transforming Existing Datasets
by: Gandhi, Saumya, et al.
Published: (2024)
by: Gandhi, Saumya, et al.
Published: (2024)
Unrequited Emotions: Investigating the Gaps in Motivation and Practice in Speech Emotion Recognition Research
by: Wong, Taryn, et al.
Published: (2026)
by: Wong, Taryn, et al.
Published: (2026)
Color, Gender, and Bias: Examining the Role of Stereotyped Colors in Visualization-Driven Pay Decisions
by: Cabric, Florent, et al.
Published: (2025)
by: Cabric, Florent, et al.
Published: (2025)
CircuitSynth: Reliable Synthetic Data Generation
by: Cheng, Zehua, et al.
Published: (2026)
by: Cheng, Zehua, et al.
Published: (2026)
SyntheT2C: Generating Synthetic Data for Fine-Tuning Large Language Models on the Text2Cypher Task
by: Zhong, Ziije, et al.
Published: (2024)
by: Zhong, Ziije, et al.
Published: (2024)
Multimodal Integration of Mel Spectrograms and Text Transcripts for Enhanced Automatic Speech Recognition: Leveraging Extractive Transformer‐Based Approaches and Late Fusion Strategies
by: Sunakshi Mehra, et al.
Published: (2024)
by: Sunakshi Mehra, et al.
Published: (2024)
Using LLMs to create analytical datasets: A case study of reconstructing the historical memory of Colombia
by: Anderson, David, et al.
Published: (2025)
by: Anderson, David, et al.
Published: (2025)
CasualSynth: Generating Structurally Sound Synthetic Data
by: Cheng, Zehua, et al.
Published: (2026)
by: Cheng, Zehua, et al.
Published: (2026)
AdaptEval: Evaluating Large Language Models on Domain Adaptation for Text Summarization
by: Afzal, Anum, et al.
Published: (2024)
by: Afzal, Anum, et al.
Published: (2024)
Synth-SBDH: A Synthetic Dataset of Social and Behavioral Determinants of Health for Clinical Text
by: Mitra, Avijit, et al.
Published: (2024)
by: Mitra, Avijit, et al.
Published: (2024)
NodeSynth: Socially Aligned Synthetic Data for AI Evaluation
by: Rashid, Qazi Mamunur, et al.
Published: (2026)
by: Rashid, Qazi Mamunur, et al.
Published: (2026)
SQLStructEval: Structural Evaluation of LLM Text-to-SQL Generation
by: Zhou, Yixi, et al.
Published: (2026)
by: Zhou, Yixi, et al.
Published: (2026)
FlashEval: Towards Fast and Accurate Evaluation of Text-to-image Diffusion Generative Models
by: Zhao, Lin, et al.
Published: (2024)
by: Zhao, Lin, et al.
Published: (2024)
Digital Transformation: A Path to Economic and Societal Value
by: Ritu Agarwal
Published: (2020)
by: Ritu Agarwal
Published: (2020)
Synth-Empathy: Towards High-Quality Synthetic Empathy Data
by: Liang, Hao, et al.
Published: (2024)
by: Liang, Hao, et al.
Published: (2024)
TIDE: Training Locally Interpretable Domain Generalization Models Enables Test-time Correction
by: Agarwal, Aishwarya, et al.
Published: (2024)
by: Agarwal, Aishwarya, et al.
Published: (2024)
Statistical Mechanics of Paraparticles
by: Thakur, Nupoor, et al.
Published: (2025)
by: Thakur, Nupoor, et al.
Published: (2025)
TrialSynth: Generation of Synthetic Sequential Clinical Trial Data
by: Gao, Chufan, et al.
Published: (2024)
by: Gao, Chufan, et al.
Published: (2024)
DermaSynth: Rich Synthetic Image-Text Pairs Using Open Access Dermatology Datasets
by: Yilmaz, Abdurrahim, et al.
Published: (2025)
by: Yilmaz, Abdurrahim, et al.
Published: (2025)
SynthSAEBench: Evaluating Sparse Autoencoders on Scalable Realistic Synthetic Data
by: Chanin, David, et al.
Published: (2026)
by: Chanin, David, et al.
Published: (2026)
AudioEval: Automatic Dual-Perspective and Multi-Dimensional Evaluation of Text-to-Audio-Generation
by: Wang, Hui, et al.
Published: (2025)
by: Wang, Hui, et al.
Published: (2025)
yley123/MatSynth-Captions-Text-Prompts-for-Material-Videos: MatSynth-Captions-Text-Prompts-for-Material-Videos
by: BOWEN
Published: (2025)
by: BOWEN
Published: (2025)
MusicEval: A Generative Music Dataset with Expert Ratings for Automatic Text-to-Music Evaluation
by: Liu, Cheng, et al.
Published: (2025)
by: Liu, Cheng, et al.
Published: (2025)
SING-SQL: A Synthetic Data Generation Framework for In-Domain Text-to-SQL Translation
by: Caferoğlu, Hasan Alp, et al.
Published: (2025)
by: Caferoğlu, Hasan Alp, et al.
Published: (2025)
Locating Information Gaps and Narrative Inconsistencies Across Languages: A Case Study of LGBT People Portrayals on Wikipedia
by: Samir, Farhan, et al.
Published: (2024)
by: Samir, Farhan, et al.
Published: (2024)
Paging Dr. GPT: Extracting Information from Clinical Notes to Enhance Patient Predictions
by: Anderson, David, et al.
Published: (2025)
by: Anderson, David, et al.
Published: (2025)
BatchEval: Towards Human-like Text Evaluation
by: Yuan, Peiwen, et al.
Published: (2023)
by: Yuan, Peiwen, et al.
Published: (2023)
RepEval: Effective Text Evaluation with LLM Representation
by: Sheng, Shuqian, et al.
Published: (2024)
by: Sheng, Shuqian, et al.
Published: (2024)
Similar Items
-
Evaluating Differentially Private Synthetic Data Generation in High-Stakes Domains
by: Ramesh, Krithika, et al.
Published: (2024) -
Improving Equity in Health Modeling with GPT4-Turbo Generated Synthetic Data: A Comparative Study
by: Smolyak, Daniel, et al.
Published: (2024) -
Controlled Generation for Private Synthetic Text
by: Zhao, Zihao, et al.
Published: (2025) -
Maximizing Predictive Performance for Small Subgroups: Functionally Adaptive Interaction Regularization (FAIR)
by: Smolyak, Daniel, et al.
Published: (2024) -
How do we measure privacy in text? A survey of text anonymization metrics
by: Ren, Yaxuan, et al.
Published: (2025)