Iterative Counterfactual Data Augmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Plyler, Mitchell, Chi, Min |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Transformers on Markov Data: Constant Depth Suffices
von: Rajaraman, Nived, et al.
Veröffentlicht: (2024)
von: Rajaraman, Nived, et al.
Veröffentlicht: (2024)
SENTRA: Selected-Next-Token Transformer for LLM Text Detection
von: Plyler, Mitchell, et al.
Veröffentlicht: (2025)
von: Plyler, Mitchell, et al.
Veröffentlicht: (2025)
Efficient Learned Data Compression via Dual-Stream Feature Decoupling
von: Ma, Huidong, et al.
Veröffentlicht: (2026)
von: Ma, Huidong, et al.
Veröffentlicht: (2026)
In-Context Learning with Representations: Contextual Generalization of Trained Transformers
von: Yang, Tong, et al.
Veröffentlicht: (2024)
von: Yang, Tong, et al.
Veröffentlicht: (2024)
A Rate-Distortion Framework for Summarization
von: Arda, Enes, et al.
Veröffentlicht: (2025)
von: Arda, Enes, et al.
Veröffentlicht: (2025)
Multi-turn Training with Basic Human Feedback Helps Little on LLM Reasoning
von: Liu, Qiang, et al.
Veröffentlicht: (2025)
von: Liu, Qiang, et al.
Veröffentlicht: (2025)
An Information-theoretic Multi-task Representation Learning Framework for Natural Language Understanding
von: Hu, Dou, et al.
Veröffentlicht: (2025)
von: Hu, Dou, et al.
Veröffentlicht: (2025)
What Makes the Preferred Thinking Direction for LLMs in Multiple-choice Questions?
von: Zhang, Yizhe, et al.
Veröffentlicht: (2025)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2025)
Cost-aware LLM-based Online Dataset Annotation
von: Elumar, Eray Can, et al.
Veröffentlicht: (2025)
von: Elumar, Eray Can, et al.
Veröffentlicht: (2025)
Attention with Markov: A Framework for Principled Analysis of Transformers via Markov Chains
von: Makkuva, Ashok Vardhan, et al.
Veröffentlicht: (2024)
von: Makkuva, Ashok Vardhan, et al.
Veröffentlicht: (2024)
Proposal and study of statistical features for string similarity computation and classification
von: Rodrigues, E. O., et al.
Veröffentlicht: (2026)
von: Rodrigues, E. O., et al.
Veröffentlicht: (2026)
Theoretical guarantees on the best-of-n alignment policy
von: Beirami, Ahmad, et al.
Veröffentlicht: (2024)
von: Beirami, Ahmad, et al.
Veröffentlicht: (2024)
Effective Context in Transformers: An Analysis of Fragmentation and Tokenization
von: Fesharaki, Amirmehdi Jafari, et al.
Veröffentlicht: (2026)
von: Fesharaki, Amirmehdi Jafari, et al.
Veröffentlicht: (2026)
Understanding Factual Recall in Transformers via Associative Memories
von: Nichani, Eshaan, et al.
Veröffentlicht: (2024)
von: Nichani, Eshaan, et al.
Veröffentlicht: (2024)
RateQuant: Optimal Mixed-Precision KV Cache Quantization via Rate-Distortion Theory
von: Zuo, Fei, et al.
Veröffentlicht: (2026)
von: Zuo, Fei, et al.
Veröffentlicht: (2026)
InfAlign: Inference-aware language model alignment
von: Balashankar, Ananth, et al.
Veröffentlicht: (2024)
von: Balashankar, Ananth, et al.
Veröffentlicht: (2024)
Filtering Beats Fine Tuning: A Bayesian Kalman View of In Context Learning in LLMs
von: Kiruluta, Andrew
Veröffentlicht: (2026)
von: Kiruluta, Andrew
Veröffentlicht: (2026)
Speculative Decoding Scaling Laws (SDSL): Throughput Optimization Made Simple
von: Bozorgkhoo, Amirhossein, et al.
Veröffentlicht: (2026)
von: Bozorgkhoo, Amirhossein, et al.
Veröffentlicht: (2026)
A Mathematical Theory for Learning Semantic Languages by Abstract Learners
von: Liao, Kuo-Yu, et al.
Veröffentlicht: (2024)
von: Liao, Kuo-Yu, et al.
Veröffentlicht: (2024)
An Enhanced Text Compression Approach Using Transformer-based Language Models
von: Rahman, Chowdhury Mofizur, et al.
Veröffentlicht: (2024)
von: Rahman, Chowdhury Mofizur, et al.
Veröffentlicht: (2024)
Fundamental Limits of Prompt Compression: A Rate-Distortion Framework for Black-Box Language Models
von: Nagle, Alliot, et al.
Veröffentlicht: (2024)
von: Nagle, Alliot, et al.
Veröffentlicht: (2024)
MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression
von: Elias, Noel, et al.
Veröffentlicht: (2024)
von: Elias, Noel, et al.
Veröffentlicht: (2024)
Latent Space Alignment for Semantic Channel Equalization
von: Hüttebräucker, Tomás, et al.
Veröffentlicht: (2024)
von: Hüttebräucker, Tomás, et al.
Veröffentlicht: (2024)
New Directions in Text Classification Research: Maximizing The Performance of Sentiment Classification from Limited Data
von: Agustian, Surya, et al.
Veröffentlicht: (2024)
von: Agustian, Surya, et al.
Veröffentlicht: (2024)
Information-Theoretic Generative Clustering of Documents
von: Du, Xin, et al.
Veröffentlicht: (2024)
von: Du, Xin, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Generation as Noisy In-Context Learning: A Unified Theory and Risk Bounds
von: Guo, Yang, et al.
Veröffentlicht: (2025)
von: Guo, Yang, et al.
Veröffentlicht: (2025)
Multimodal Learning Without Labeled Multimodal Data: Guarantees and Applications
von: Liang, Paul Pu, et al.
Veröffentlicht: (2023)
von: Liang, Paul Pu, et al.
Veröffentlicht: (2023)
Quantifying Logical Consistency in Transformers via Query-Key Alignment
von: Tulchinskii, Eduard, et al.
Veröffentlicht: (2025)
von: Tulchinskii, Eduard, et al.
Veröffentlicht: (2025)
Synthetic Counterfactual Labels for Efficient Conformal Counterfactual Inference
von: Farzaneh, Amirmohammad, et al.
Veröffentlicht: (2025)
von: Farzaneh, Amirmohammad, et al.
Veröffentlicht: (2025)
Theoretical Limits of Language Model Alignment
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2026)
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2026)
An Experimental Study on Data Augmentation Techniques for Named Entity Recognition on Low-Resource Domains
von: Torres, Arthur Elwing, et al.
Veröffentlicht: (2024)
von: Torres, Arthur Elwing, et al.
Veröffentlicht: (2024)
Domain Adaptation for Industrial Time-series Forecasting via Counterfactual Inference
von: Min, Chao, et al.
Veröffentlicht: (2024)
von: Min, Chao, et al.
Veröffentlicht: (2024)
iFlip: Iterative Feedback-driven Counterfactual Example Refinement
von: Wang, Yilong, et al.
Veröffentlicht: (2026)
von: Wang, Yilong, et al.
Veröffentlicht: (2026)
AT-RAG: An Adaptive RAG Model Enhancing Query Efficiency with Topic Filtering and Iterative Reasoning
von: Rezaei, Mohammad Reza, et al.
Veröffentlicht: (2024)
von: Rezaei, Mohammad Reza, et al.
Veröffentlicht: (2024)
What is Hiding in Medicine's Dark Matter? Learning with Missing Data in Medical Practices
von: Suzen, Neslihan, et al.
Veröffentlicht: (2024)
von: Suzen, Neslihan, et al.
Veröffentlicht: (2024)
A Training-free Method for LLM Text Attribution
von: Radvand, Tara, et al.
Veröffentlicht: (2025)
von: Radvand, Tara, et al.
Veröffentlicht: (2025)
Know Your Limits: Entropy Estimation Modeling for Compression and Generalization
von: Badger, Benjamin L., et al.
Veröffentlicht: (2025)
von: Badger, Benjamin L., et al.
Veröffentlicht: (2025)
Memorization-Compression Cycles Improve Generalization
von: Yu, Fangyuan
Veröffentlicht: (2025)
von: Yu, Fangyuan
Veröffentlicht: (2025)
Measuring Uncertainty in Transformer Circuits with Effective Information Consistency
von: Krasnovsky, Anatoly A.
Veröffentlicht: (2025)
von: Krasnovsky, Anatoly A.
Veröffentlicht: (2025)
SPEX: Scaling Feature Interaction Explanations for LLMs
von: Kang, Justin Singh, et al.
Veröffentlicht: (2025)
von: Kang, Justin Singh, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Transformers on Markov Data: Constant Depth Suffices
von: Rajaraman, Nived, et al.
Veröffentlicht: (2024) -
SENTRA: Selected-Next-Token Transformer for LLM Text Detection
von: Plyler, Mitchell, et al.
Veröffentlicht: (2025) -
Efficient Learned Data Compression via Dual-Stream Feature Decoupling
von: Ma, Huidong, et al.
Veröffentlicht: (2026) -
In-Context Learning with Representations: Contextual Generalization of Trained Transformers
von: Yang, Tong, et al.
Veröffentlicht: (2024) -
A Rate-Distortion Framework for Summarization
von: Arda, Enes, et al.
Veröffentlicht: (2025)