To Burst or Not to Burst: Generating and Quantifying Improbable Text
Fuente:
arXiv
Saved in:
| Main Authors: | Sasse, Kuleen, Barham, Samuel, Kayi, Efsun Sarioglu, Staley, Edward W. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Controllable Hybrid Captioner for Improved Long-form Video Understanding
by: Sasse, Kuleen, et al.
Published: (2025)
by: Sasse, Kuleen, et al.
Published: (2025)
StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection
by: Alqahtani, Amal, et al.
Published: (2025)
by: Alqahtani, Amal, et al.
Published: (2025)
Making FETCH! Happen: Finding Emergent Dog Whistles Through Common Habitats
by: Sasse, Kuleen, et al.
Published: (2024)
by: Sasse, Kuleen, et al.
Published: (2024)
debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias
by: Sasse, Kuleen, et al.
Published: (2024)
by: Sasse, Kuleen, et al.
Published: (2024)
PLAID SHIRTTT for Large-Scale Streaming Dense Retrieval
by: Lawrie, Dawn, et al.
Published: (2024)
by: Lawrie, Dawn, et al.
Published: (2024)
Disease Entity Recognition and Normalization is Improved with Large Language Model Derived Synthetic Normalized Mentions
by: Sasse, Kuleen, et al.
Published: (2024)
by: Sasse, Kuleen, et al.
Published: (2024)
Sparse Autoencoder Features for Classifications and Transferability
by: Gallifant, Jack, et al.
Published: (2025)
by: Gallifant, Jack, et al.
Published: (2025)
Wait, but Tylenol is Acetaminophen... Investigating and Improving Language Models' Ability to Resist Requests for Misinformation
by: Chen, Shan, et al.
Published: (2024)
by: Chen, Shan, et al.
Published: (2024)
Improbable Bigrams Expose Vulnerabilities of Incomplete Tokens in Byte-Level Tokenizers
by: Jang, Eugene, et al.
Published: (2024)
by: Jang, Eugene, et al.
Published: (2024)
FACTUM: Mechanistic Detection of Citation Hallucination in Long-Form RAG
by: Dassen, Maxime, et al.
Published: (2026)
by: Dassen, Maxime, et al.
Published: (2026)
On the Evaluation of Machine-Generated Reports
by: Mayfield, James, et al.
Published: (2024)
by: Mayfield, James, et al.
Published: (2024)
MegaWika 2: A More Comprehensive Multilingual Collection of Articles and their Sources
by: Barham, Samuel, et al.
Published: (2025)
by: Barham, Samuel, et al.
Published: (2025)
QUDsim: Quantifying Discourse Similarities in LLM-Generated Text
by: Namuduri, Ramya, et al.
Published: (2025)
by: Namuduri, Ramya, et al.
Published: (2025)
NBF at SemEval-2025 Task 5: Light-Burst Attention Enhanced System for Multilingual Subject Recommendation
by: Islam, Baharul, et al.
Published: (2025)
by: Islam, Baharul, et al.
Published: (2025)
BurstGP: Enhancing Raw Burst Image Super Resolution with Generative Priors
by: Huo, Dong, et al.
Published: (2026)
by: Huo, Dong, et al.
Published: (2026)
JaccDiv: A Metric and Benchmark for Quantifying Diversity of Generated Marketing Text in the Music Industry
by: Afzal, Anum, et al.
Published: (2025)
by: Afzal, Anum, et al.
Published: (2025)
EditLens: Quantifying the Extent of AI Editing in Text
by: Thai, Katherine, et al.
Published: (2025)
by: Thai, Katherine, et al.
Published: (2025)
Modeling the Attack: Detecting AI-Generated Text by Quantifying Adversarial Perturbations
by: Teja, Lekkala Sai, et al.
Published: (2025)
by: Teja, Lekkala Sai, et al.
Published: (2025)
QUITE: Quantifying Uncertainty in Natural Language Text in Bayesian Reasoning Scenarios
by: Schrader, Timo Pierre, et al.
Published: (2024)
by: Schrader, Timo Pierre, et al.
Published: (2024)
Why Mean Pooling Works: Quantifying Second-Order Collapse in Text Embeddings
by: Hara, Tomomasa, et al.
Published: (2026)
by: Hara, Tomomasa, et al.
Published: (2026)
Quantifying Generalization Complexity for Large Language Models
by: Qi, Zhenting, et al.
Published: (2024)
by: Qi, Zhenting, et al.
Published: (2024)
AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
by: Lu, Ximing, et al.
Published: (2024)
by: Lu, Ximing, et al.
Published: (2024)
Quantifying Positional Biases in Text Embedding Models
by: Lee, Reagan J., et al.
Published: (2024)
by: Lee, Reagan J., et al.
Published: (2024)
Privacy Text Clustering Method Based on Burst Feature of Words
by: Xia Wu, et al.
Published: (2025)
by: Xia Wu, et al.
Published: (2025)
Exploration of Masked and Causal Language Modelling for Text Generation
by: Micheletti, Nicolo, et al.
Published: (2024)
by: Micheletti, Nicolo, et al.
Published: (2024)
Knowledge Distillation in Automated Annotation: Supervised Text Classification with LLM-Generated Training Labels
by: Pangakis, Nicholas, et al.
Published: (2024)
by: Pangakis, Nicholas, et al.
Published: (2024)
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
by: Macko, Dominik, et al.
Published: (2024)
by: Macko, Dominik, et al.
Published: (2024)
For Generated Text, Is NLI-Neutral Text the Best Text?
by: Mersinias, Michail, et al.
Published: (2023)
by: Mersinias, Michail, et al.
Published: (2023)
RAG-E: Quantifying Retriever-Generator Alignment and Failure Modes
by: Randl, Korbinian, et al.
Published: (2026)
by: Randl, Korbinian, et al.
Published: (2026)
TextMachina: Seamless Generation of Machine-Generated Text Datasets
by: Sarvazyan, Areg Mikael, et al.
Published: (2024)
by: Sarvazyan, Areg Mikael, et al.
Published: (2024)
Parameter Efficient Fine Tuning Llama 3.1 for Answering Arabic Legal Questions: A Case Study on Jordanian Laws
by: Fasha, Mohammed, et al.
Published: (2026)
by: Fasha, Mohammed, et al.
Published: (2026)
CIE: Controlling Language Model Text Generations Using Continuous Signals
by: Samuel, Vinay, et al.
Published: (2025)
by: Samuel, Vinay, et al.
Published: (2025)
Generative Voice Bursts during Phone Call
by: Ranjan, Paritosh, et al.
Published: (2025)
by: Ranjan, Paritosh, et al.
Published: (2025)
Cited Text Spans for Citation Text Generation
by: Li, Xiangci, et al.
Published: (2023)
by: Li, Xiangci, et al.
Published: (2023)
Quantifying the Gap between Understanding and Generation within Unified Multimodal Models
by: Wang, Chenlong, et al.
Published: (2026)
by: Wang, Chenlong, et al.
Published: (2026)
Argument Mining as a Text-to-Text Generation Task
by: Kawarada, Masayuki, et al.
Published: (2026)
by: Kawarada, Masayuki, et al.
Published: (2026)
Unmasking and Quantifying Racial Bias of Large Language Models in Medical Report Generation
by: Yang, Yifan, et al.
Published: (2024)
by: Yang, Yifan, et al.
Published: (2024)
A Human-in/on-the-Loop Framework for Accessible Text Generation
by: Moreno, Lourdes, et al.
Published: (2026)
by: Moreno, Lourdes, et al.
Published: (2026)
Model-Generated Pretraining Signals Improves Zero-Shot Generalization of Text-to-Text Transformers
by: Gong, Linyuan, et al.
Published: (2023)
by: Gong, Linyuan, et al.
Published: (2023)
Are We in the AI-Generated Text World Already? Quantifying and Monitoring AIGT on Social Media
by: Sun, Zhen, et al.
Published: (2024)
by: Sun, Zhen, et al.
Published: (2024)
Similar Items
-
Controllable Hybrid Captioner for Improved Long-form Video Understanding
by: Sasse, Kuleen, et al.
Published: (2025) -
StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection
by: Alqahtani, Amal, et al.
Published: (2025) -
Making FETCH! Happen: Finding Emergent Dog Whistles Through Common Habitats
by: Sasse, Kuleen, et al.
Published: (2024) -
debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias
by: Sasse, Kuleen, et al.
Published: (2024) -
PLAID SHIRTTT for Large-Scale Streaming Dense Retrieval
by: Lawrie, Dawn, et al.
Published: (2024)