Promises and Pitfalls of Generative Masked Language Modeling: Theoretical Framework and Practical Guidelines
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yuchen, Kirchmeyer, Alexandre, Mehta, Aashay, Qin, Yilong, Dadachev, Boris, Papineni, Kishore, Kumar, Sanjiv, Risteski, Andrej |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Query Complexity of Verifier-Assisted Language Generation
von: Botta, Edoardo, et al.
Veröffentlicht: (2025)
von: Botta, Edoardo, et al.
Veröffentlicht: (2025)
Fit Like You Sample: Sample-Efficient Generalized Score Matching from Fast Mixing Diffusions
von: Qin, Yilong, et al.
Veröffentlicht: (2023)
von: Qin, Yilong, et al.
Veröffentlicht: (2023)
Exploring and Improving Drafts in Blockwise Parallel Decoding
von: Kim, Taehyeon, et al.
Veröffentlicht: (2024)
von: Kim, Taehyeon, et al.
Veröffentlicht: (2024)
Large Language Models for Stemming: Promises, Pitfalls and Failures
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
CodePDE: An Inference Framework for LLM-driven PDE Solver Generation
von: Li, Shanda, et al.
Veröffentlicht: (2025)
von: Li, Shanda, et al.
Veröffentlicht: (2025)
The Promises and Pitfalls of Using Language Models to Measure Instruction Quality in Education
von: Xu, Paiheng, et al.
Veröffentlicht: (2024)
von: Xu, Paiheng, et al.
Veröffentlicht: (2024)
A 2-step Framework for Automated Literary Translation Evaluation: Its Promises and Pitfalls
von: Shafayat, Sheikh, et al.
Veröffentlicht: (2024)
von: Shafayat, Sheikh, et al.
Veröffentlicht: (2024)
Toward Understanding the Transferability of Adversarial Suffixes in Large Language Models
von: Ball, Sarah, et al.
Veröffentlicht: (2025)
von: Ball, Sarah, et al.
Veröffentlicht: (2025)
Evaluating the Promise and Pitfalls of LLMs in Hiring Decisions
von: Anzenberg, Eitan, et al.
Veröffentlicht: (2025)
von: Anzenberg, Eitan, et al.
Veröffentlicht: (2025)
A Large Language Model-based Framework for Semi-Structured Tender Document Retrieval-Augmented Generation
von: Zhao, Yilong, et al.
Veröffentlicht: (2024)
von: Zhao, Yilong, et al.
Veröffentlicht: (2024)
Benefits and Pitfalls of Reinforcement Learning for Language Model Planning: A Theoretical Perspective
von: Wang, Siwei, et al.
Veröffentlicht: (2025)
von: Wang, Siwei, et al.
Veröffentlicht: (2025)
PPLqa: An Unsupervised Information-Theoretic Quality Metric for Comparing Generative Large Language Models
von: Friedland, Gerald, et al.
Veröffentlicht: (2024)
von: Friedland, Gerald, et al.
Veröffentlicht: (2024)
The Promises and Pitfalls of LLM Annotations in Dataset Labeling: a Case Study on Media Bias Detection
von: Horych, Tomas, et al.
Veröffentlicht: (2024)
von: Horych, Tomas, et al.
Veröffentlicht: (2024)
A computational phase transition for learning-to-sample from Ising models
von: Risteski, Andrej, et al.
Veröffentlicht: (2026)
von: Risteski, Andrej, et al.
Veröffentlicht: (2026)
The Perils & Promises of Fact-checking with Large Language Models
von: Quelle, Dorian, et al.
Veröffentlicht: (2023)
von: Quelle, Dorian, et al.
Veröffentlicht: (2023)
Pitfalls of Evaluating Language Models with Open Benchmarks
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2025)
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2025)
From Guidelines to Practice: A New Paradigm for Arabic Language Model Evaluation
von: Sibaee, Serry, et al.
Veröffentlicht: (2025)
von: Sibaee, Serry, et al.
Veröffentlicht: (2025)
The tractability landscape of diffusion alignment: regularization, rewards, and computational primitives
von: Moitra, Ankur, et al.
Veröffentlicht: (2026)
von: Moitra, Ankur, et al.
Veröffentlicht: (2026)
Steering diffusion models with quadratic rewards: a fine-grained analysis
von: Moitra, Ankur, et al.
Veröffentlicht: (2026)
von: Moitra, Ankur, et al.
Veröffentlicht: (2026)
The Pitfalls of Defining Hallucination
von: van Deemter, Kees
Veröffentlicht: (2024)
von: van Deemter, Kees
Veröffentlicht: (2024)
From Explainability to Action: A Generative Operational Framework for Integrating XAI in Clinical Mental Health Screening
von: Kandala, Ratna, et al.
Veröffentlicht: (2025)
von: Kandala, Ratna, et al.
Veröffentlicht: (2025)
Masked Mixers for Language Generation and Retrieval
von: Badger, Benjamin L.
Veröffentlicht: (2024)
von: Badger, Benjamin L.
Veröffentlicht: (2024)
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining
von: Sam, Dylan, et al.
Veröffentlicht: (2025)
von: Sam, Dylan, et al.
Veröffentlicht: (2025)
The Muddy Waters of Modeling Empathy in Language: The Practical Impacts of Theoretical Constructs
von: Lahnala, Allison, et al.
Veröffentlicht: (2025)
von: Lahnala, Allison, et al.
Veröffentlicht: (2025)
Pitfalls of Scale: Investigating the Inverse Task of Redefinition in Large Language Models
von: Stringli, Elena, et al.
Veröffentlicht: (2025)
von: Stringli, Elena, et al.
Veröffentlicht: (2025)
MOTOR: Multimodal Optimal Transport via Grounded Retrieval in Medical Visual Question Answering
von: Shaaban, Mai A., et al.
Veröffentlicht: (2025)
von: Shaaban, Mai A., et al.
Veröffentlicht: (2025)
Beware of Reasoning Overconfidence: Pitfalls in the Reasoning Process for Multi-solution Tasks
von: Guan, Jiannan, et al.
Veröffentlicht: (2025)
von: Guan, Jiannan, et al.
Veröffentlicht: (2025)
Pitfalls and Outlooks in Using COMET
von: Zouhar, Vilém, et al.
Veröffentlicht: (2024)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2024)
MultiCheck: Strengthening Web Trust with Unified Multimodal Fact Verification
von: Kishore, Aditya, et al.
Veröffentlicht: (2025)
von: Kishore, Aditya, et al.
Veröffentlicht: (2025)
Enhancing Large Language Models for Clinical Decision Support by Incorporating Clinical Practice Guidelines
von: Oniani, David, et al.
Veröffentlicht: (2024)
von: Oniani, David, et al.
Veröffentlicht: (2024)
What Makes an Evaluation Useful? Common Pitfalls and Best Practices
von: Gekker, Gil, et al.
Veröffentlicht: (2025)
von: Gekker, Gil, et al.
Veröffentlicht: (2025)
Representations as Language: An Information-Theoretic Framework for Interpretability
von: Conklin, Henry, et al.
Veröffentlicht: (2024)
von: Conklin, Henry, et al.
Veröffentlicht: (2024)
CLA Guidelines on Non-Sexist Language
Veröffentlicht: (1976)
Veröffentlicht: (1976)
Refine Medical Diagnosis Using Generation Augmented Retrieval and Clinical Practice Guidelines
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
XGrammar: Flexible and Efficient Structured Generation Engine for Large Language Models
von: Dong, Yixin, et al.
Veröffentlicht: (2024)
von: Dong, Yixin, et al.
Veröffentlicht: (2024)
Editing the Mind of Giants: An In-Depth Exploration of Pitfalls of Knowledge Editing in Large Language Models
von: Hsueh, Cheng-Hsun, et al.
Veröffentlicht: (2024)
von: Hsueh, Cheng-Hsun, et al.
Veröffentlicht: (2024)
Exploration of Masked and Causal Language Modelling for Text Generation
von: Micheletti, Nicolo, et al.
Veröffentlicht: (2024)
von: Micheletti, Nicolo, et al.
Veröffentlicht: (2024)
Pelican Soup Framework: A Theoretical Framework for Language Model Capabilities
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2024)
von: Chiang, Ting-Rui, et al.
Veröffentlicht: (2024)
Deep sequence models tend to memorize geometrically; it is unclear why
von: Noroozizadeh, Shahriar, et al.
Veröffentlicht: (2025)
von: Noroozizadeh, Shahriar, et al.
Veröffentlicht: (2025)
An Interactive Framework for Profiling News Media Sources
von: Mehta, Nikhil, et al.
Veröffentlicht: (2023)
von: Mehta, Nikhil, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
On the Query Complexity of Verifier-Assisted Language Generation
von: Botta, Edoardo, et al.
Veröffentlicht: (2025) -
Fit Like You Sample: Sample-Efficient Generalized Score Matching from Fast Mixing Diffusions
von: Qin, Yilong, et al.
Veröffentlicht: (2023) -
Exploring and Improving Drafts in Blockwise Parallel Decoding
von: Kim, Taehyeon, et al.
Veröffentlicht: (2024) -
Large Language Models for Stemming: Promises, Pitfalls and Failures
von: Wang, Shuai, et al.
Veröffentlicht: (2024) -
CodePDE: An Inference Framework for LLM-driven PDE Solver Generation
von: Li, Shanda, et al.
Veröffentlicht: (2025)