Position: Model Collapse Does Not Mean What You Think
Fuente:
arXiv
Salvato in:
| Autori principali: | Schaeffer, Rylan, Kazdan, Joshua, Arulandu, Alvan Caleb, Koyejo, Sanmi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Understanding Adversarial Transfer: Why Representation-Space Attacks Fail Where Data-Space Attacks Succeed
di: Gupta, Isha, et al.
Pubblicazione: (2025)
di: Gupta, Isha, et al.
Pubblicazione: (2025)
Collapse or Thrive? Perils and Promises of Synthetic Data in a Self-Generating World
di: Kazdan, Joshua, et al.
Pubblicazione: (2024)
di: Kazdan, Joshua, et al.
Pubblicazione: (2024)
Efficient Prediction of Pass@k Scaling in Large Language Models
di: Kazdan, Joshua, et al.
Pubblicazione: (2025)
di: Kazdan, Joshua, et al.
Pubblicazione: (2025)
Position: Machine Learning Conferences Should Establish a "Refutations and Critiques" Track
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
Consensus is Not Verification: Why Crowd Wisdom Strategies Fail for LLM Truthfulness
di: Denisov-Blanch, Yegor, et al.
Pubblicazione: (2026)
di: Denisov-Blanch, Yegor, et al.
Pubblicazione: (2026)
What Causes Polysemanticity? An Alternative Origin Story of Mixed Selectivity from Incidental Causes
di: Lecomte, Victor, et al.
Pubblicazione: (2023)
di: Lecomte, Victor, et al.
Pubblicazione: (2023)
No, of Course I Can! Deeper Fine-Tuning Attacks That Bypass Token-Level Safety Mechanisms
di: Kazdan, Joshua, et al.
Pubblicazione: (2025)
di: Kazdan, Joshua, et al.
Pubblicazione: (2025)
Scale Dependent Data Duplication
di: Kazdan, Joshua, et al.
Pubblicazione: (2026)
di: Kazdan, Joshua, et al.
Pubblicazione: (2026)
How Do Large Language Monkeys Get Their Power (Laws)?
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
Investigating Data Contamination for Pre-training Language Models
di: Jiang, Minhao, et al.
Pubblicazione: (2024)
di: Jiang, Minhao, et al.
Pubblicazione: (2024)
On Fairness of Low-Rank Adaptation of Large Models
di: Ding, Zhoujie, et al.
Pubblicazione: (2024)
di: Ding, Zhoujie, et al.
Pubblicazione: (2024)
ZIP-FIT: Embedding-Free Data Selection via Compression-Based Alignment
di: Obbad, Elyas, et al.
Pubblicazione: (2024)
di: Obbad, Elyas, et al.
Pubblicazione: (2024)
Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research
di: Cooper, A. Feder, et al.
Pubblicazione: (2024)
di: Cooper, A. Feder, et al.
Pubblicazione: (2024)
Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinants
di: Tang, Zeyu, et al.
Pubblicazione: (2025)
di: Tang, Zeyu, et al.
Pubblicazione: (2025)
Min-p, Max Exaggeration: A Critical Analysis of Min-p Sampling in Language Models
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
Beyond Scale: The Diversity Coefficient as a Data Quality Metric for Variability in Natural Language Data
di: Miranda, Brando, et al.
Pubblicazione: (2023)
di: Miranda, Brando, et al.
Pubblicazione: (2023)
Why Has Predicting Downstream Capabilities of Frontier AI Models with Scale Remained Elusive?
di: Schaeffer, Rylan, et al.
Pubblicazione: (2024)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2024)
Globalizing Fairness Attributes in Machine Learning: A Case Study on Health in Africa
di: Asiedu, Mercy Nyamewaa, et al.
Pubblicazione: (2023)
di: Asiedu, Mercy Nyamewaa, et al.
Pubblicazione: (2023)
In-Context Learning of Energy Functions
di: Schaeffer, Rylan, et al.
Pubblicazione: (2024)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2024)
Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data
di: Gerstgrasser, Matthias, et al.
Pubblicazione: (2024)
di: Gerstgrasser, Matthias, et al.
Pubblicazione: (2024)
KGGen: Extracting Knowledge Graphs from Plain Text with Language Models
di: Mo, Belinda, et al.
Pubblicazione: (2025)
di: Mo, Belinda, et al.
Pubblicazione: (2025)
Best-of-N Jailbreaking
di: Hughes, John, et al.
Pubblicazione: (2024)
di: Hughes, John, et al.
Pubblicazione: (2024)
Quantifying Variance in Evaluation Benchmarks
di: Madaan, Lovish, et al.
Pubblicazione: (2024)
di: Madaan, Lovish, et al.
Pubblicazione: (2024)
Pretraining Scaling Laws for Generative Evaluations of Language Models
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025)
Sharpe Ratio-Guided Active Learning for Preference Optimization in RLHF
di: Belakaria, Syrine, et al.
Pubblicazione: (2025)
di: Belakaria, Syrine, et al.
Pubblicazione: (2025)
Building a Domain-specific Guardrail Model in Production
di: Niknazar, Mohammad, et al.
Pubblicazione: (2024)
di: Niknazar, Mohammad, et al.
Pubblicazione: (2024)
Logits are All We Need to Adapt Closed Models
di: Hiranandani, Gaurush, et al.
Pubblicazione: (2025)
di: Hiranandani, Gaurush, et al.
Pubblicazione: (2025)
Can You Trust an LLM with Your Life-Changing Decision? An Investigation into AI High-Stakes Responses
di: Cahyono, Joshua Adrian, et al.
Pubblicazione: (2025)
di: Cahyono, Joshua Adrian, et al.
Pubblicazione: (2025)
Shaping AI's Impact on Billions of Lives
di: Cuéllar, Mariano-Florentino, et al.
Pubblicazione: (2024)
di: Cuéllar, Mariano-Florentino, et al.
Pubblicazione: (2024)
HiFA: High-fidelity Text-to-3D Generation with Advanced Diffusion Guidance
di: Zhu, Junzhe, et al.
Pubblicazione: (2023)
di: Zhu, Junzhe, et al.
Pubblicazione: (2023)
Why Do Safety Guardrails Degrade Across Languages?
di: Zhang, Max, et al.
Pubblicazione: (2026)
di: Zhang, Max, et al.
Pubblicazione: (2026)
Extracting books from production language models
di: Ahmed, Ahmed, et al.
Pubblicazione: (2026)
di: Ahmed, Ahmed, et al.
Pubblicazione: (2026)
Quantifying the Effect of Test Set Contamination on Generative Evaluations
di: Schaeffer, Rylan, et al.
Pubblicazione: (2026)
di: Schaeffer, Rylan, et al.
Pubblicazione: (2026)
Reliable and Efficient Amortized Model-based Evaluation
di: Truong, Sang, et al.
Pubblicazione: (2025)
di: Truong, Sang, et al.
Pubblicazione: (2025)
The Utility and Complexity of in- and out-of-Distribution Machine Unlearning
di: Allouah, Youssef, et al.
Pubblicazione: (2024)
di: Allouah, Youssef, et al.
Pubblicazione: (2024)
MANTA: Multi-turn Assessment for Nonhuman Thinking & Alignment
di: Lu, Allen, et al.
Pubblicazione: (2026)
di: Lu, Allen, et al.
Pubblicazione: (2026)
The Secret Agenda: LLMs Strategically Lie and Our Current Safety Tools Are Blind
di: DeLeeuw, Caleb, et al.
Pubblicazione: (2025)
di: DeLeeuw, Caleb, et al.
Pubblicazione: (2025)
From Passive to Active Reasoning: Can Large Language Models Ask the Right Questions under Incomplete Information?
di: Zhou, Zhanke, et al.
Pubblicazione: (2025)
di: Zhou, Zhanke, et al.
Pubblicazione: (2025)
Limits of trust in medical AI
di: Hatherley, Joshua
Pubblicazione: (2025)
di: Hatherley, Joshua
Pubblicazione: (2025)
Are clinicians ethically obligated to disclose their use of medical machine learning systems to patients?
di: Hatherley, Joshua
Pubblicazione: (2025)
di: Hatherley, Joshua
Pubblicazione: (2025)
Documenti analoghi
-
Understanding Adversarial Transfer: Why Representation-Space Attacks Fail Where Data-Space Attacks Succeed
di: Gupta, Isha, et al.
Pubblicazione: (2025) -
Collapse or Thrive? Perils and Promises of Synthetic Data in a Self-Generating World
di: Kazdan, Joshua, et al.
Pubblicazione: (2024) -
Efficient Prediction of Pass@k Scaling in Large Language Models
di: Kazdan, Joshua, et al.
Pubblicazione: (2025) -
Position: Machine Learning Conferences Should Establish a "Refutations and Critiques" Track
di: Schaeffer, Rylan, et al.
Pubblicazione: (2025) -
Consensus is Not Verification: Why Crowd Wisdom Strategies Fail for LLM Truthfulness
di: Denisov-Blanch, Yegor, et al.
Pubblicazione: (2026)