Salvato in:
| Autori principali: | Ghosal, Gaurav, Hashimoto, Tatsunori, Raghunathan, Aditi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2406.14785 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Language Models with Conformal Factuality Guarantees
di: Mohri, Christopher, et al.
Pubblicazione: (2024)
di: Mohri, Christopher, et al.
Pubblicazione: (2024)
Context-Parametric Inversion: Why Instruction Finetuning Can Worsen Context Reliance
di: Goyal, Sachin, et al.
Pubblicazione: (2024)
di: Goyal, Sachin, et al.
Pubblicazione: (2024)
Understanding Catastrophic Forgetting in Language Models via Implicit Inference
di: Kotha, Suhas, et al.
Pubblicazione: (2023)
di: Kotha, Suhas, et al.
Pubblicazione: (2023)
Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs
di: Zhong, Ziqian, et al.
Pubblicazione: (2025)
di: Zhong, Ziqian, et al.
Pubblicazione: (2025)
Memorization Sinks: Isolating Memorization during LLM Training
di: Ghosal, Gaurav R., et al.
Pubblicazione: (2025)
di: Ghosal, Gaurav R., et al.
Pubblicazione: (2025)
Improving Pretraining Data Using Perplexity Correlations
di: Thrush, Tristan, et al.
Pubblicazione: (2024)
di: Thrush, Tristan, et al.
Pubblicazione: (2024)
ImpossibleBench: Measuring LLMs' Propensity of Exploiting Test Cases
di: Zhong, Ziqian, et al.
Pubblicazione: (2025)
di: Zhong, Ziqian, et al.
Pubblicazione: (2025)
Self-Trained Verification for Training- and Test-Time Self-Improvement
di: Wu, Chen Henry, et al.
Pubblicazione: (2026)
di: Wu, Chen Henry, et al.
Pubblicazione: (2026)
Testing the Limits of Jailbreaking Defenses with the Purple Problem
di: Kim, Taeyoun, et al.
Pubblicazione: (2024)
di: Kim, Taeyoun, et al.
Pubblicazione: (2024)
Mitigating Bias in RAG: Controlling the Embedder
di: Kim, Taeyoun, et al.
Pubblicazione: (2025)
di: Kim, Taeyoun, et al.
Pubblicazione: (2025)
Observational Scaling Laws and the Predictability of Language Model Performance
di: Ruan, Yangjun, et al.
Pubblicazione: (2024)
di: Ruan, Yangjun, et al.
Pubblicazione: (2024)
Mode-Conditioning Unlocks Superior Test-Time Scaling
di: Wu, Chen Henry, et al.
Pubblicazione: (2025)
di: Wu, Chen Henry, et al.
Pubblicazione: (2025)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
di: Si, Chenglei, et al.
Pubblicazione: (2024)
di: Si, Chenglei, et al.
Pubblicazione: (2024)
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
di: Si, Chenglei, et al.
Pubblicazione: (2025)
di: Si, Chenglei, et al.
Pubblicazione: (2025)
Towards Reliable Latent Knowledge Estimation in LLMs: Zero-Prompt Many-Shot Based Factual Knowledge Extraction
di: Wu, Qinyuan, et al.
Pubblicazione: (2024)
di: Wu, Qinyuan, et al.
Pubblicazione: (2024)
Linguistic Calibration of Long-Form Generations
di: Band, Neil, et al.
Pubblicazione: (2024)
di: Band, Neil, et al.
Pubblicazione: (2024)
Robust Distortion-free Watermarks for Language Models
di: Kuditipudi, Rohith, et al.
Pubblicazione: (2023)
di: Kuditipudi, Rohith, et al.
Pubblicazione: (2023)
Putting It All into Context: Simplifying Agents with LCLMs
di: Jiang, Mingjian, et al.
Pubblicazione: (2025)
di: Jiang, Mingjian, et al.
Pubblicazione: (2025)
Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
di: Dubois, Yann, et al.
Pubblicazione: (2024)
di: Dubois, Yann, et al.
Pubblicazione: (2024)
On the Learnability of Watermarks for Language Models
di: Gu, Chenchen, et al.
Pubblicazione: (2023)
di: Gu, Chenchen, et al.
Pubblicazione: (2023)
Reasoning to Learn from Latent Thoughts
di: Ruan, Yangjun, et al.
Pubblicazione: (2025)
di: Ruan, Yangjun, et al.
Pubblicazione: (2025)
Repetition Improves Language Model Embeddings
di: Springer, Jacob Mitchell, et al.
Pubblicazione: (2024)
di: Springer, Jacob Mitchell, et al.
Pubblicazione: (2024)
Sharpness-Aware Pretraining Mitigates Catastrophic Forgetting
di: Watts, Ishaan, et al.
Pubblicazione: (2026)
di: Watts, Ishaan, et al.
Pubblicazione: (2026)
Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction
di: Nagarajan, Vaishnavh, et al.
Pubblicazione: (2025)
di: Nagarajan, Vaishnavh, et al.
Pubblicazione: (2025)
Understanding Contextual Recall in Transformers: How Finetuning Enables In-Context Reasoning over Pretraining Knowledge
di: Vasudeva, Bhavya, et al.
Pubblicazione: (2026)
di: Vasudeva, Bhavya, et al.
Pubblicazione: (2026)
Persuasion Tokens for Editing Factual Knowledge in LLMs
di: Youssef, Paul, et al.
Pubblicazione: (2026)
di: Youssef, Paul, et al.
Pubblicazione: (2026)
Synthetic continued pretraining
di: Yang, Zitong, et al.
Pubblicazione: (2024)
di: Yang, Zitong, et al.
Pubblicazione: (2024)
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
di: Jiang, Mingjian, et al.
Pubblicazione: (2024)
di: Jiang, Mingjian, et al.
Pubblicazione: (2024)
Agentic Adversarial QA for Improving Domain-Specific LLMs
di: Grari, Vincent, et al.
Pubblicazione: (2026)
di: Grari, Vincent, et al.
Pubblicazione: (2026)
Auditing Prompt Caching in Language Model APIs
di: Gu, Chenchen, et al.
Pubblicazione: (2025)
di: Gu, Chenchen, et al.
Pubblicazione: (2025)
Early Data Exposure Improves Robustness to Subsequent Fine-Tuning
di: Feng, Lawrence, et al.
Pubblicazione: (2026)
di: Feng, Lawrence, et al.
Pubblicazione: (2026)
AutoBencher: Towards Declarative Benchmark Construction
di: Li, Xiang Lisa, et al.
Pubblicazione: (2024)
di: Li, Xiang Lisa, et al.
Pubblicazione: (2024)
Failure Modes of LLMs for Causal Reasoning on Narratives
di: Yamin, Khurram, et al.
Pubblicazione: (2024)
di: Yamin, Khurram, et al.
Pubblicazione: (2024)
KnowLA: Enhancing Parameter-efficient Finetuning with Knowledgeable Adaptation
di: Luo, Xindi, et al.
Pubblicazione: (2024)
di: Luo, Xindi, et al.
Pubblicazione: (2024)
From Style to Facts: Mapping the Boundaries of Knowledge Injection with Finetuning
di: Zhao, Eric, et al.
Pubblicazione: (2025)
di: Zhao, Eric, et al.
Pubblicazione: (2025)
Towards Execution-Grounded Automated AI Research
di: Si, Chenglei, et al.
Pubblicazione: (2026)
di: Si, Chenglei, et al.
Pubblicazione: (2026)
Base Models Look Human To AI Detectors
di: Xu, Yixuan Even, et al.
Pubblicazione: (2026)
di: Xu, Yixuan Even, et al.
Pubblicazione: (2026)
Alternate Preference Optimization for Unlearning Factual Knowledge in Large Language Models
di: Mekala, Anmol, et al.
Pubblicazione: (2024)
di: Mekala, Anmol, et al.
Pubblicazione: (2024)
T-MARS: Improving Visual Representations by Circumventing Text Feature Learning
di: Maini, Pratyush, et al.
Pubblicazione: (2023)
di: Maini, Pratyush, et al.
Pubblicazione: (2023)
Understanding Factual Recall in Transformers via Associative Memories
di: Nichani, Eshaan, et al.
Pubblicazione: (2024)
di: Nichani, Eshaan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Language Models with Conformal Factuality Guarantees
di: Mohri, Christopher, et al.
Pubblicazione: (2024) -
Context-Parametric Inversion: Why Instruction Finetuning Can Worsen Context Reliance
di: Goyal, Sachin, et al.
Pubblicazione: (2024) -
Understanding Catastrophic Forgetting in Language Models via Implicit Inference
di: Kotha, Suhas, et al.
Pubblicazione: (2023) -
Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs
di: Zhong, Ziqian, et al.
Pubblicazione: (2025) -
Memorization Sinks: Isolating Memorization during LLM Training
di: Ghosal, Gaurav R., et al.
Pubblicazione: (2025)