Salvato in:
| Autori principali: | Shaib, Chantal, Barrow, Joe, Siu, Alexa F., Wallace, Byron C., Nenkova, Ani |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2402.18756 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Standardizing the Measurement of Text Diversity: A Tool and a Comparative Analysis of Scores
di: Shaib, Chantal, et al.
Pubblicazione: (2024)
di: Shaib, Chantal, et al.
Pubblicazione: (2024)
Who Taught You That? Tracing Teachers in Model Distillation
di: Wadhwa, Somin, et al.
Pubblicazione: (2025)
di: Wadhwa, Somin, et al.
Pubblicazione: (2025)
Detection and Measurement of Syntactic Templates in Generated Text
di: Shaib, Chantal, et al.
Pubblicazione: (2024)
di: Shaib, Chantal, et al.
Pubblicazione: (2024)
Measuring AI "Slop" in Text
di: Shaib, Chantal, et al.
Pubblicazione: (2025)
di: Shaib, Chantal, et al.
Pubblicazione: (2025)
Learning the Wrong Lessons: Syntactic-Domain Spurious Correlations in Language Models
di: Shaib, Chantal, et al.
Pubblicazione: (2025)
di: Shaib, Chantal, et al.
Pubblicazione: (2025)
Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence
di: Mo, Kaijie, et al.
Pubblicazione: (2026)
di: Mo, Kaijie, et al.
Pubblicazione: (2026)
Measuring Lexical Diversity of Synthetic Data Generated through Fine-Grained Persona Prompting
di: Kambhatla, Gauri, et al.
Pubblicazione: (2025)
di: Kambhatla, Gauri, et al.
Pubblicazione: (2025)
SQLSpace: A Representation Space for Text-to-SQL to Discover and Mitigate Robustness Gaps
di: Srikanth, Neha, et al.
Pubblicazione: (2025)
di: Srikanth, Neha, et al.
Pubblicazione: (2025)
Few-Shot Dialogue Summarization via Skeleton-Assisted Prompt Transfer in Prompt Tuning
di: Xie, Kaige, et al.
Pubblicazione: (2023)
di: Xie, Kaige, et al.
Pubblicazione: (2023)
Do Multi-Document Summarization Models Synthesize?
di: DeYoung, Jay, et al.
Pubblicazione: (2023)
di: DeYoung, Jay, et al.
Pubblicazione: (2023)
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
di: Ramprasad, Sanjana, et al.
Pubblicazione: (2024)
di: Ramprasad, Sanjana, et al.
Pubblicazione: (2024)
Turning English-centric LLMs Into Polyglots: How Much Multilinguality Is Needed?
di: Kew, Tannon, et al.
Pubblicazione: (2023)
di: Kew, Tannon, et al.
Pubblicazione: (2023)
Compared to What? Baselines and Metrics for Counterfactual Prompting
di: Yang, Zihao, et al.
Pubblicazione: (2026)
di: Yang, Zihao, et al.
Pubblicazione: (2026)
Revisiting Relation Extraction in the era of Large Language Models
di: Wadhwa, Somin, et al.
Pubblicazione: (2023)
di: Wadhwa, Somin, et al.
Pubblicazione: (2023)
MODS: Moderating a Mixture of Document Speakers to Summarize Debatable Queries in Document Collections
di: Balepur, Nishant, et al.
Pubblicazione: (2025)
di: Balepur, Nishant, et al.
Pubblicazione: (2025)
Do Automatic Factuality Metrics Measure Factuality? A Critical Evaluation
di: Ramprasad, Sanjana, et al.
Pubblicazione: (2024)
di: Ramprasad, Sanjana, et al.
Pubblicazione: (2024)
Can SAEs reveal and mitigate racial biases of LLMs in healthcare?
di: Ahsan, Hiba, et al.
Pubblicazione: (2025)
di: Ahsan, Hiba, et al.
Pubblicazione: (2025)
FactPICO: Factuality Evaluation for Plain Language Summarization of Medical Evidence
di: Joseph, Sebastian Antony, et al.
Pubblicazione: (2024)
di: Joseph, Sebastian Antony, et al.
Pubblicazione: (2024)
Chain of Logic: Rule-Based Reasoning with Large Language Models
di: Servantez, Sergio, et al.
Pubblicazione: (2024)
di: Servantez, Sergio, et al.
Pubblicazione: (2024)
Short-Context Dominance: How Much Local Context Natural Language Actually Needs?
di: Vakilian, Vala, et al.
Pubblicazione: (2025)
di: Vakilian, Vala, et al.
Pubblicazione: (2025)
How Much Context Does My Attention-Based ASR System Need?
di: Flynn, Robert, et al.
Pubblicazione: (2023)
di: Flynn, Robert, et al.
Pubblicazione: (2023)
Investigating Mysteries of CoT-Augmented Distillation
di: Wadhwa, Somin, et al.
Pubblicazione: (2024)
di: Wadhwa, Somin, et al.
Pubblicazione: (2024)
Circuit Distillation
di: Wadhwa, Somin, et al.
Pubblicazione: (2025)
di: Wadhwa, Somin, et al.
Pubblicazione: (2025)
Vector Arithmetic in Concept and Token Subspaces
di: Feucht, Sheridan, et al.
Pubblicazione: (2025)
di: Feucht, Sheridan, et al.
Pubblicazione: (2025)
Leveraging ChatGPT in Pharmacovigilance Event Extraction: An Empirical Study
di: Sun, Zhaoyue, et al.
Pubblicazione: (2024)
di: Sun, Zhaoyue, et al.
Pubblicazione: (2024)
Memory Is All You Need: Testing How Model Memory Affects LLM Performance in Annotation Tasks
di: Timoneda, Joan C., et al.
Pubblicazione: (2025)
di: Timoneda, Joan C., et al.
Pubblicazione: (2025)
How Much Does Persuasion Strategy Matter? LLM-Annotated Evidence from Charitable Donation Dialogues
di: Petrova, Tatiana, et al.
Pubblicazione: (2026)
di: Petrova, Tatiana, et al.
Pubblicazione: (2026)
How Private are Language Models in Abstractive Summarization?
di: Hughes, Anthony, et al.
Pubblicazione: (2024)
di: Hughes, Anthony, et al.
Pubblicazione: (2024)
Automatically Extracting Numerical Results from Randomized Controlled Trials with Large Language Models
di: Yun, Hye Sun, et al.
Pubblicazione: (2024)
di: Yun, Hye Sun, et al.
Pubblicazione: (2024)
Do Large Language Models Know How Much They Know?
di: Prato, Gabriele, et al.
Pubblicazione: (2025)
di: Prato, Gabriele, et al.
Pubblicazione: (2025)
The Dual-Route Model of Induction
di: Feucht, Sheridan, et al.
Pubblicazione: (2025)
di: Feucht, Sheridan, et al.
Pubblicazione: (2025)
Don't Pay Attention, PLANT It: Pretraining Attention via Learning-to-Rank
di: Roy, Debjyoti Saha, et al.
Pubblicazione: (2024)
di: Roy, Debjyoti Saha, et al.
Pubblicazione: (2024)
With Good MT There is No Need For End-to-End: A Case for Translate-then-Summarize Cross-lingual Summarization
di: Varab, Daniel, et al.
Pubblicazione: (2024)
di: Varab, Daniel, et al.
Pubblicazione: (2024)
How Much Can RAG Help the Reasoning of LLM?
di: Liu, Jingyu, et al.
Pubblicazione: (2024)
di: Liu, Jingyu, et al.
Pubblicazione: (2024)
Comparative Personalization for Multi-document Summarization
di: Li, Haoyuan, et al.
Pubblicazione: (2025)
di: Li, Haoyuan, et al.
Pubblicazione: (2025)
How Much Do LLMs Know About Chinese Zero Pronouns?
di: Li, Yifei, et al.
Pubblicazione: (2026)
di: Li, Yifei, et al.
Pubblicazione: (2026)
Learning from Natural Language Explanations for Generalizable Entity Matching
di: Wadhwa, Somin, et al.
Pubblicazione: (2024)
di: Wadhwa, Somin, et al.
Pubblicazione: (2024)
Comparative Analysis of Abstractive Summarization Models for Clinical Radiology Reports
di: Bhattacharya, Anindita, et al.
Pubblicazione: (2025)
di: Bhattacharya, Anindita, et al.
Pubblicazione: (2025)
How Much Do Circuits Tell Us? Measuring the Consistency and Specificity of Language Model Circuits
di: Li, Michael, et al.
Pubblicazione: (2026)
di: Li, Michael, et al.
Pubblicazione: (2026)
How Much is Enough? The Diminishing Returns of Tokenization Training Data
di: Reddy, Varshini, et al.
Pubblicazione: (2025)
di: Reddy, Varshini, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Standardizing the Measurement of Text Diversity: A Tool and a Comparative Analysis of Scores
di: Shaib, Chantal, et al.
Pubblicazione: (2024) -
Who Taught You That? Tracing Teachers in Model Distillation
di: Wadhwa, Somin, et al.
Pubblicazione: (2025) -
Detection and Measurement of Syntactic Templates in Generated Text
di: Shaib, Chantal, et al.
Pubblicazione: (2024) -
Measuring AI "Slop" in Text
di: Shaib, Chantal, et al.
Pubblicazione: (2025) -
Learning the Wrong Lessons: Syntactic-Domain Spurious Correlations in Language Models
di: Shaib, Chantal, et al.
Pubblicazione: (2025)