Understanding Hallucinations in Diffusion Models through Mode Interpolation
Fuente:
arXiv
Salvato in:
| Autori principali: | Aithal, Sumukh K, Maini, Pratyush, Lipton, Zachary C., Kolter, J. Zico |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Rethinking LLM Memorization through the Lens of Adversarial Compression
di: Schwarzschild, Avi, et al.
Pubblicazione: (2024)
di: Schwarzschild, Avi, et al.
Pubblicazione: (2024)
Scaling Laws for Data Filtering -- Data Curation cannot be Compute Agnostic
di: Goyal, Sachin, et al.
Pubblicazione: (2024)
di: Goyal, Sachin, et al.
Pubblicazione: (2024)
TOFU: A Task of Fictitious Unlearning for LLMs
di: Maini, Pratyush, et al.
Pubblicazione: (2024)
di: Maini, Pratyush, et al.
Pubblicazione: (2024)
T-MARS: Improving Visual Representations by Circumventing Text Feature Learning
di: Maini, Pratyush, et al.
Pubblicazione: (2023)
di: Maini, Pratyush, et al.
Pubblicazione: (2023)
When Should We Introduce Safety Interventions During Pretraining?
di: Sam, Dylan, et al.
Pubblicazione: (2026)
di: Sam, Dylan, et al.
Pubblicazione: (2026)
Safety Pretraining: Toward the Next Generation of Safe AI
di: Maini, Pratyush, et al.
Pubblicazione: (2025)
di: Maini, Pratyush, et al.
Pubblicazione: (2025)
One-Step Diffusion Distillation via Deep Equilibrium Models
di: Geng, Zhengyang, et al.
Pubblicazione: (2023)
di: Geng, Zhengyang, et al.
Pubblicazione: (2023)
Diffusing Differentiable Representations
di: Savani, Yash, et al.
Pubblicazione: (2024)
di: Savani, Yash, et al.
Pubblicazione: (2024)
Predicting the Performance of Black-box LLMs through Follow-up Queries
di: Sam, Dylan, et al.
Pubblicazione: (2025)
di: Sam, Dylan, et al.
Pubblicazione: (2025)
FUSE-ing Language Models: Zero-Shot Adapter Discovery for Prompt Optimization Across Tokenizers
di: Williams, Joshua Nathaniel, et al.
Pubblicazione: (2024)
di: Williams, Joshua Nathaniel, et al.
Pubblicazione: (2024)
AcceleratedLiNGAM: Learning Causal DAGs at the speed of GPUs
di: Akinwande, Victor, et al.
Pubblicazione: (2024)
di: Akinwande, Victor, et al.
Pubblicazione: (2024)
Forcing Diffuse Distributions out of Language Models
di: Zhang, Yiming, et al.
Pubblicazione: (2024)
di: Zhang, Yiming, et al.
Pubblicazione: (2024)
Why is SAM Robust to Label Noise?
di: Baek, Christina, et al.
Pubblicazione: (2024)
di: Baek, Christina, et al.
Pubblicazione: (2024)
Equilibrium Reasoners: Learning Attractors Enables Scalable Reasoning
di: Huang, Benhao, et al.
Pubblicazione: (2026)
di: Huang, Benhao, et al.
Pubblicazione: (2026)
Mimetic Initialization of MLPs
di: Trockman, Asher, et al.
Pubblicazione: (2026)
di: Trockman, Asher, et al.
Pubblicazione: (2026)
OpenUnlearning: Accelerating LLM Unlearning via Unified Benchmarking of Methods and Metrics
di: Dorna, Vineeth, et al.
Pubblicazione: (2025)
di: Dorna, Vineeth, et al.
Pubblicazione: (2025)
Understanding Augmentation-based Self-Supervised Representation Learning via RKHS Approximation and Regression
di: Zhai, Runtian, et al.
Pubblicazione: (2023)
di: Zhai, Runtian, et al.
Pubblicazione: (2023)
One-Step Diffusion Distillation through Score Implicit Matching
di: Luo, Weijian, et al.
Pubblicazione: (2024)
di: Luo, Weijian, et al.
Pubblicazione: (2024)
An Axiomatic Approach to Model-Agnostic Concept Explanations
di: Feng, Zhili, et al.
Pubblicazione: (2024)
di: Feng, Zhili, et al.
Pubblicazione: (2024)
Massive Activations in Large Language Models
di: Sun, Mingjie, et al.
Pubblicazione: (2024)
di: Sun, Mingjie, et al.
Pubblicazione: (2024)
Peeking Behind Closed Doors: Risks of LLM Evaluation by Private Data Curators
di: Bansal, Hritik, et al.
Pubblicazione: (2025)
di: Bansal, Hritik, et al.
Pubblicazione: (2025)
From Variance to Veracity: Unbundling and Mitigating Gradient Variance in Differentiable Bundle Adjustment Layers
di: Gurumurthy, Swaminathan, et al.
Pubblicazione: (2024)
di: Gurumurthy, Swaminathan, et al.
Pubblicazione: (2024)
Measuring Five-Nines Reliability: Sample-Efficient LLM Evaluation in Saturated Benchmarks
di: Kim, Eungyeup, et al.
Pubblicazione: (2026)
di: Kim, Eungyeup, et al.
Pubblicazione: (2026)
Generative Posterior Networks for Approximately Bayesian Epistemic Uncertainty Estimation
di: Roderick, Melrose, et al.
Pubblicazione: (2023)
di: Roderick, Melrose, et al.
Pubblicazione: (2023)
Evaluating Language Model Reasoning about Confidential Information
di: Sam, Dylan, et al.
Pubblicazione: (2025)
di: Sam, Dylan, et al.
Pubblicazione: (2025)
A Simple and Effective Pruning Approach for Large Language Models
di: Sun, Mingjie, et al.
Pubblicazione: (2023)
di: Sun, Mingjie, et al.
Pubblicazione: (2023)
Rethinking Distance Metrics for Counterfactual Explainability
di: Williams, Joshua Nathaniel, et al.
Pubblicazione: (2024)
di: Williams, Joshua Nathaniel, et al.
Pubblicazione: (2024)
Understanding Optimization in Deep Learning with Central Flows
di: Cohen, Jeremy M., et al.
Pubblicazione: (2024)
di: Cohen, Jeremy M., et al.
Pubblicazione: (2024)
Context-Parametric Inversion: Why Instruction Finetuning Can Worsen Context Reliance
di: Goyal, Sachin, et al.
Pubblicazione: (2024)
di: Goyal, Sachin, et al.
Pubblicazione: (2024)
The Mixing method: low-rank coordinate descent for semidefinite programming with diagonal constraints
di: Wang, Po-Wei, et al.
Pubblicazione: (2017)
di: Wang, Po-Wei, et al.
Pubblicazione: (2017)
Weight Ensembling Improves Reasoning in Language Models
di: Dang, Xingyu, et al.
Pubblicazione: (2025)
di: Dang, Xingyu, et al.
Pubblicazione: (2025)
Mimetic Initialization Helps State Space Models Learn to Recall
di: Trockman, Asher, et al.
Pubblicazione: (2024)
di: Trockman, Asher, et al.
Pubblicazione: (2024)
Accelerating Diffusion Planners in Offline RL via Reward-Aware Consistency Trajectory Distillation
di: Duan, Xintong, et al.
Pubblicazione: (2025)
di: Duan, Xintong, et al.
Pubblicazione: (2025)
STAMP Your Content: Proving Dataset Membership via Watermarked Rephrasings
di: Rastogi, Saksham, et al.
Pubblicazione: (2025)
di: Rastogi, Saksham, et al.
Pubblicazione: (2025)
Prompt Recovery for Image Generation Models: A Comparative Study of Discrete Optimizers
di: Williams, Joshua Nathaniel, et al.
Pubblicazione: (2024)
di: Williams, Joshua Nathaniel, et al.
Pubblicazione: (2024)
Memorization Sinks: Isolating Memorization during LLM Training
di: Ghosal, Gaurav R., et al.
Pubblicazione: (2025)
di: Ghosal, Gaurav R., et al.
Pubblicazione: (2025)
Predicting the Performance of Foundation Models via Agreement-on-the-Line
di: Saxena, Rahul, et al.
Pubblicazione: (2024)
di: Saxena, Rahul, et al.
Pubblicazione: (2024)
Adaptive Data Optimization: Dynamic Sample Selection with Scaling Laws
di: Jiang, Yiding, et al.
Pubblicazione: (2024)
di: Jiang, Yiding, et al.
Pubblicazione: (2024)
Consistency Models Made Easy
di: Geng, Zhengyang, et al.
Pubblicazione: (2024)
di: Geng, Zhengyang, et al.
Pubblicazione: (2024)
Looking beyond the next token
di: Thankaraj, Abitha, et al.
Pubblicazione: (2025)
di: Thankaraj, Abitha, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Rethinking LLM Memorization through the Lens of Adversarial Compression
di: Schwarzschild, Avi, et al.
Pubblicazione: (2024) -
Scaling Laws for Data Filtering -- Data Curation cannot be Compute Agnostic
di: Goyal, Sachin, et al.
Pubblicazione: (2024) -
TOFU: A Task of Fictitious Unlearning for LLMs
di: Maini, Pratyush, et al.
Pubblicazione: (2024) -
T-MARS: Improving Visual Representations by Circumventing Text Feature Learning
di: Maini, Pratyush, et al.
Pubblicazione: (2023) -
When Should We Introduce Safety Interventions During Pretraining?
di: Sam, Dylan, et al.
Pubblicazione: (2026)