Hubble: a Model Suite to Advance the Study of LLM Memorization
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Johnny Tian-Zheng, Godbole, Ameya, Khan, Mohammad Aflah, Wang, Ryan, Zhu, Xiaoyuan, Flemings, James, Kashyap, Nitya, Gummadi, Krishna P., Neiswanger, Willie, Jia, Robin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TokenSmith: Streamlining Data Editing, Search, and Inspection for Large-Scale Language Model Training and Interpretability
by: Khan, Mohammad Aflah, et al.
Published: (2025)
by: Khan, Mohammad Aflah, et al.
Published: (2025)
LLM Unlearning Without an Expert Curated Dataset
by: Zhu, Xiaoyuan, et al.
Published: (2025)
by: Zhu, Xiaoyuan, et al.
Published: (2025)
Spiking the training data to correct for test set contamination
by: Wei, Johnny Tian-Zheng, et al.
Published: (2026)
by: Wei, Johnny Tian-Zheng, et al.
Published: (2026)
Rethinking Memorization Measures and their Implications in Large Language Models
by: Ghosh, Bishwamittra, et al.
Published: (2025)
by: Ghosh, Bishwamittra, et al.
Published: (2025)
Interrogating LLM design under a fair learning doctrine
by: Wei, Johnny Tian-Zheng, et al.
Published: (2025)
by: Wei, Johnny Tian-Zheng, et al.
Published: (2025)
Verify with Caution: The Pitfalls of Relying on Imperfect Factuality Metrics
by: Godbole, Ameya, et al.
Published: (2025)
by: Godbole, Ameya, et al.
Published: (2025)
Rote Learning Considered Useful: Generalizing over Memorized Data in LLMs
by: Wu, Qinyuan, et al.
Published: (2025)
by: Wu, Qinyuan, et al.
Published: (2025)
Fractional Rotation, Full Potential? Investigating Performance and Convergence of Partial RoPE
by: Khan, Mohammad Aflah, et al.
Published: (2026)
by: Khan, Mohammad Aflah, et al.
Published: (2026)
SHRED: Retain-Set-Free Unlearning via Self-Distillation with Logit Demotion
by: Hu, Zizhao, et al.
Published: (2026)
by: Hu, Zizhao, et al.
Published: (2026)
SCENE: Self-Labeled Counterfactuals for Extrapolating to Negative Examples
by: Fu, Deqing, et al.
Published: (2023)
by: Fu, Deqing, et al.
Published: (2023)
In Agents We Trust, but Who Do Agents Trust? Latent Source Preferences Steer LLM Generations
by: Khan, Mohammad Aflah, et al.
Published: (2026)
by: Khan, Mohammad Aflah, et al.
Published: (2026)
Proving membership in LLM pretraining data via data watermarks
by: Wei, Johnny Tian-Zheng, et al.
Published: (2024)
by: Wei, Johnny Tian-Zheng, et al.
Published: (2024)
Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective
by: Ghosh, Bishwamittra, et al.
Published: (2026)
by: Ghosh, Bishwamittra, et al.
Published: (2026)
Revisiting Privacy, Utility, and Efficiency Trade-offs when Fine-Tuning Large Language Models
by: Das, Soumi, et al.
Published: (2025)
by: Das, Soumi, et al.
Published: (2025)
Understanding Memorisation in LLMs: Dynamics, Influencing Factors, and Implications
by: Speicher, Till, et al.
Published: (2024)
by: Speicher, Till, et al.
Published: (2024)
Rethinking RL for LLM Reasoning: It's Sparse Policy Selection, Not Capability Learning
by: Akgül, Ömer Faruk, et al.
Published: (2026)
by: Akgül, Ömer Faruk, et al.
Published: (2026)
Advances in Earth Abundant Metal Catalyzed Synthesis of Quinazolinones
by: Pragati Kushwaha, et al.
Published: (2026)
by: Pragati Kushwaha, et al.
Published: (2026)
From Calibration to Collaboration: LLM Uncertainty Quantification Should Be More Human-Centered
by: Devic, Siddartha, et al.
Published: (2025)
by: Devic, Siddartha, et al.
Published: (2025)
The statistical advantage of automatic NLG metrics at the system level
by: Wei, Johnny Tian-Zheng, et al.
Published: (2021)
by: Wei, Johnny Tian-Zheng, et al.
Published: (2021)
Uncertainty Quantification for Forward and Inverse Problems of PDEs via Latent Global Evolution
by: Wu, Tailin, et al.
Published: (2024)
by: Wu, Tailin, et al.
Published: (2024)
DeLLMa: Decision Making Under Uncertainty with Large Language Models
by: Liu, Ollie, et al.
Published: (2024)
by: Liu, Ollie, et al.
Published: (2024)
Probabilistic Graphical Models: A Concise Tutorial
by: Maasch, Jacqueline, et al.
Published: (2025)
by: Maasch, Jacqueline, et al.
Published: (2025)
Auditing Black-Box LLM APIs with a Rank-Based Uniformity Test
by: Zhu, Xiaoyuan, et al.
Published: (2025)
by: Zhu, Xiaoyuan, et al.
Published: (2025)
Towards Reliable Latent Knowledge Estimation in LLMs: Zero-Prompt Many-Shot Based Factual Knowledge Extraction
by: Wu, Qinyuan, et al.
Published: (2024)
by: Wu, Qinyuan, et al.
Published: (2024)
PrivacySIM: Evaluating LLM Simulation of User Privacy Behavior
by: Flemings, James, et al.
Published: (2026)
by: Flemings, James, et al.
Published: (2026)
AI-University: An LLM-based platform for instructional alignment to scientific classrooms
by: Shojaei, Mostafa Faghih, et al.
Published: (2025)
by: Shojaei, Mostafa Faghih, et al.
Published: (2025)
IsoBench: Benchmarking Multimodal Foundation Models on Isomorphic Representations
by: Fu, Deqing, et al.
Published: (2024)
by: Fu, Deqing, et al.
Published: (2024)
Operationalizing content moderation "accuracy" in the Digital Services Act
by: Wei, Johnny Tian-Zheng, et al.
Published: (2023)
by: Wei, Johnny Tian-Zheng, et al.
Published: (2023)
Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models
by: Gan, Woody Haosheng, et al.
Published: (2025)
by: Gan, Woody Haosheng, et al.
Published: (2025)
Precise Debugging Benchmark: Is Your Model Debugging or Regenerating?
by: Zhu, Wang Bill, et al.
Published: (2026)
by: Zhu, Wang Bill, et al.
Published: (2026)
Evaluating LLM-Simulated Conversations in Modeling Inconsistent and Uncollaborative Behaviors in Human Social Interaction
by: Kamoi, Ryo, et al.
Published: (2026)
by: Kamoi, Ryo, et al.
Published: (2026)
Supercharging Simulation-Based Inference for Bayesian Optimal Experimental Design
by: Klein, Samuel, et al.
Published: (2026)
by: Klein, Samuel, et al.
Published: (2026)
Euclid: Supercharging Multimodal LLMs with Synthetic High-Fidelity Visual Descriptions
by: Zhang, Jiarui, et al.
Published: (2024)
by: Zhang, Jiarui, et al.
Published: (2024)
QUENCH: Measuring the gap between Indic and Non-Indic Contextual General Reasoning in LLMs
by: Khan, Mohammad Aflah, et al.
Published: (2024)
by: Khan, Mohammad Aflah, et al.
Published: (2024)
Recite, Reconstruct, Recollect: Memorization in LMs as a Multifaceted Phenomenon
by: Prashanth, USVSN Sai, et al.
Published: (2024)
by: Prashanth, USVSN Sai, et al.
Published: (2024)
Robust Data Watermarking in Language Models by Injecting Fictitious Knowledge
by: Cui, Xinyue, et al.
Published: (2025)
by: Cui, Xinyue, et al.
Published: (2025)
Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection
by: Anwar, Abrar, et al.
Published: (2025)
by: Anwar, Abrar, et al.
Published: (2025)
LYNX: Learning Dynamic Exits for Confidence-Controlled Reasoning
by: Akgül, Ömer Faruk, et al.
Published: (2025)
by: Akgül, Ömer Faruk, et al.
Published: (2025)
Neural Nonmyopic Bayesian Optimization in Dynamic Cost Settings
by: Truong, Sang T., et al.
Published: (2026)
by: Truong, Sang T., et al.
Published: (2026)
Differentially Private Knowledge Distillation via Synthetic Text Generation
by: Flemings, James, et al.
Published: (2024)
by: Flemings, James, et al.
Published: (2024)
Similar Items
-
TokenSmith: Streamlining Data Editing, Search, and Inspection for Large-Scale Language Model Training and Interpretability
by: Khan, Mohammad Aflah, et al.
Published: (2025) -
LLM Unlearning Without an Expert Curated Dataset
by: Zhu, Xiaoyuan, et al.
Published: (2025) -
Spiking the training data to correct for test set contamination
by: Wei, Johnny Tian-Zheng, et al.
Published: (2026) -
Rethinking Memorization Measures and their Implications in Large Language Models
by: Ghosh, Bishwamittra, et al.
Published: (2025) -
Interrogating LLM design under a fair learning doctrine
by: Wei, Johnny Tian-Zheng, et al.
Published: (2025)