A Penalty Goes a Long Way: Measuring Lexical Diversity in Synthetic Texts Under Prompt-Influenced Length Variations
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Deshpande, Vijeta, Dasgupta, Ishita, Bhattacharya, Uttaran, Sarkhel, Somdeb, Mitra, Saayan, Rumshisky, Anna |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances
par: Narasimhaswamy, Supreeth, et autres
Publié: (2024)
par: Narasimhaswamy, Supreeth, et autres
Publié: (2024)
Atom: Efficient On-Device Video-Language Pipelines Through Modular Reuse
par: Panchal, Kunjal, et autres
Publié: (2025)
par: Panchal, Kunjal, et autres
Publié: (2025)
SKALD: Learning-Based Shot Assembly for Coherent Multi-Shot Video Creation
par: Lu, Chen Yi, et autres
Publié: (2025)
par: Lu, Chen Yi, et autres
Publié: (2025)
Diverse, not Short: A Length-Controlled Data Selection Strategy for Improving Response Diversity of Language Models
par: Deshpande, Vijeta, et autres
Publié: (2025)
par: Deshpande, Vijeta, et autres
Publié: (2025)
Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection
par: Deshpande, Vijeta, et autres
Publié: (2026)
par: Deshpande, Vijeta, et autres
Publié: (2026)
Scaling Down to Scale Up: A Guide to Parameter-Efficient Fine-Tuning
par: Lialin, Vladislav, et autres
Publié: (2023)
par: Lialin, Vladislav, et autres
Publié: (2023)
Emergent Abilities in Reduced-Scale Generative Language Models
par: Muckatira, Sherin, et autres
Publié: (2024)
par: Muckatira, Sherin, et autres
Publié: (2024)
A Pre-Training Analogue of Grokking in Language Models: Tracing Delayed Grammatical Generalization
par: Muckatira, Sherin, et autres
Publié: (2026)
par: Muckatira, Sherin, et autres
Publié: (2026)
Beyond Perplexity: A Geometric and Spectral Study of Low-Rank Pre-Training
par: Shivagunde, Namrata, et autres
Publié: (2026)
par: Shivagunde, Namrata, et autres
Publié: (2026)
GT-SVJ: Generative-Transformer-Based Self-Supervised Video Judge For Efficient Video Reward Modeling
par: Shekhar, Shivanshu, et autres
Publié: (2026)
par: Shekhar, Shivanshu, et autres
Publié: (2026)
Analysis of student understanding in short‐answer explanations to concept questions using a human‐centered AI approach
par: Harpreet Auby, et autres
Publié: (2025)
par: Harpreet Auby, et autres
Publié: (2025)
Rethinking Failure Attribution in Multi-Agent Systems: A Multi-Perspective Benchmark and Evaluation
par: In, Yeonjun, et autres
Publié: (2026)
par: In, Yeonjun, et autres
Publié: (2026)
Shape My Moves: Text-Driven Shape-Aware Synthesis of Human Motions
par: Liao, Ting-Hsuan, et autres
Publié: (2025)
par: Liao, Ting-Hsuan, et autres
Publié: (2025)
On Explaining Visual Captioning with Hybrid Markov Logic Networks
par: Shah, Monika, et autres
Publié: (2025)
par: Shah, Monika, et autres
Publié: (2025)
Disentangling Fine-Tuning from Pre-Training in Visual Captioning with Hybrid Markov Logic
par: Shah, Monika, et autres
Publié: (2025)
par: Shah, Monika, et autres
Publié: (2025)
Measuring Lexical Diversity of Synthetic Data Generated through Fine-Grained Persona Prompting
par: Kambhatla, Gauri, et autres
Publié: (2025)
par: Kambhatla, Gauri, et autres
Publié: (2025)
DanceAnyWay: Synthesizing Beat-Guided 3D Dances with Randomized Temporal Contrastive Learning
par: Bhattacharya, Aneesh, et autres
Publié: (2023)
par: Bhattacharya, Aneesh, et autres
Publié: (2023)
Playing with Words, Improving with Rewards: Training Language Models for Creative Association
par: Deshpande, Vijeta, et autres
Publié: (2026)
par: Deshpande, Vijeta, et autres
Publié: (2026)
Assessing Interactions with DNA, Gene Transfection Potential, and Cytotoxicity of Amide‐Bonded Cationic Gemini Surfactants: The Role of Chain Length
par: Homen Dahal, et autres
Publié: (2025)
par: Homen Dahal, et autres
Publié: (2025)
Verifying Relational Explanations: A Probabilistic Approach
par: Magar, Abisha Thapa, et autres
Publié: (2024)
par: Magar, Abisha Thapa, et autres
Publié: (2024)
Token-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual Information
par: Hu, Zhengmian, et autres
Publié: (2023)
par: Hu, Zhengmian, et autres
Publié: (2023)
X-Reflect: Cross-Reflection Prompting for Multimodal Recommendation
par: Lyu, Hanjia, et autres
Publié: (2024)
par: Lyu, Hanjia, et autres
Publié: (2024)
A Little Aggression Goes a Long Way
par: Krishnan, Jyothi, et autres
Publié: (2024)
par: Krishnan, Jyothi, et autres
Publié: (2024)
A Little Confidence Goes a Long Way
par: Scoville, John, et autres
Publié: (2024)
par: Scoville, John, et autres
Publié: (2024)
Analyzing the Sensitivity of Vision Language Models in Visual Question Answering
par: Shah, Monika, et autres
Publié: (2025)
par: Shah, Monika, et autres
Publié: (2025)
A Little Human Data Goes A Long Way
par: Ashok, Dhananjay, et autres
Publié: (2024)
par: Ashok, Dhananjay, et autres
Publié: (2024)
Capacity Constraints and the Multilingual Penalty for Lexical Disambiguation
par: Trott, Sean, et autres
Publié: (2026)
par: Trott, Sean, et autres
Publié: (2026)
Efficient and Robust Registration on the 3D Special Euclidean Group
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
Detecting Ambiguities to Guide Query Rewrite for Robust Conversations in Enterprise AI Assistants
par: Tanjim, Md Mehrab, et autres
Publié: (2025)
par: Tanjim, Md Mehrab, et autres
Publié: (2025)
A Faster $k$-means++ Algorithm
par: Liang, Jiehao, et autres
Publié: (2022)
par: Liang, Jiehao, et autres
Publié: (2022)
SODA: Protecting Proprietary Information in On-Device Machine Learning Models
par: Atrey, Akanksha, et autres
Publié: (2023)
par: Atrey, Akanksha, et autres
Publié: (2023)
Speech2UnifiedExpressions: Synchronous Synthesis of Co-Speech Affective Face and Body Expressions from Affordable Inputs
par: Bhattacharya, Uttaran, et autres
Publié: (2024)
par: Bhattacharya, Uttaran, et autres
Publié: (2024)
Sequential Fair Allocation With Replenishments: A Little Envy Goes An Exponentially Long Way
par: Onyeze, Chido, et autres
Publié: (2025)
par: Onyeze, Chido, et autres
Publié: (2025)
Default Effect in ESG Investment: When a Recommendation Goes a Long Way
par: Sai Sravanthi Ramadugula, et autres
Publié: (2026)
par: Sai Sravanthi Ramadugula, et autres
Publié: (2026)
A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts
par: Ge, Suyu, et autres
Publié: (2024)
par: Ge, Suyu, et autres
Publié: (2024)
Deconstructing In-Context Learning: Understanding Prompts via Corruption
par: Shivagunde, Namrata, et autres
Publié: (2024)
par: Shivagunde, Namrata, et autres
Publié: (2024)
Lexical and Statistical Analysis of Bangla Newspaper and Literature: A Corpus-Driven Study on Diversity, Readability, and NLP Adaptation
par: Bhattacharyya, Pramit, et autres
Publié: (2025)
par: Bhattacharyya, Pramit, et autres
Publié: (2025)
A Little Depth Goes a Long Way: The Expressive Power of Log-Depth Transformers
par: Merrill, William, et autres
Publié: (2025)
par: Merrill, William, et autres
Publié: (2025)
JEBS: A Fine-grained Biomedical Lexical Simplification Task
par: Xia, William, et autres
Publié: (2025)
par: Xia, William, et autres
Publié: (2025)
Lexical Variation and Change
par: Geeraerts, Dirk, et autres
Publié: (2023)
par: Geeraerts, Dirk, et autres
Publié: (2023)
Documents similaires
-
HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances
par: Narasimhaswamy, Supreeth, et autres
Publié: (2024) -
Atom: Efficient On-Device Video-Language Pipelines Through Modular Reuse
par: Panchal, Kunjal, et autres
Publié: (2025) -
SKALD: Learning-Based Shot Assembly for Coherent Multi-Shot Video Creation
par: Lu, Chen Yi, et autres
Publié: (2025) -
Diverse, not Short: A Length-Controlled Data Selection Strategy for Improving Response Diversity of Language Models
par: Deshpande, Vijeta, et autres
Publié: (2025) -
Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection
par: Deshpande, Vijeta, et autres
Publié: (2026)