Illuminating Blind Spots of Language Models with Targeted Agent-in-the-Loop Synthetic Data
Fuente:
arXiv
Salvato in:
| Autori principali: | Lippmann, Philip, Spaan, Matthijs T. J., Yang, Jie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Positive Experience Reflection for Agents in Interactive Text Environments
di: Lippmann, Philip, et al.
Pubblicazione: (2024)
di: Lippmann, Philip, et al.
Pubblicazione: (2024)
Zero-Shot Contextual Embeddings via Offline Synthetic Corpus Generation
di: Lippmann, Philip, et al.
Pubblicazione: (2025)
di: Lippmann, Philip, et al.
Pubblicazione: (2025)
Style over Substance: Distilled Language Models Reason Via Stylistic Replication
di: Lippmann, Philip, et al.
Pubblicazione: (2025)
di: Lippmann, Philip, et al.
Pubblicazione: (2025)
Context-Informed Machine Translation of Manga using Multimodal Large Language Models
di: Lippmann, Philip, et al.
Pubblicazione: (2024)
di: Lippmann, Philip, et al.
Pubblicazione: (2024)
Temporal Blind Spots in Large Language Models
di: Wallat, Jonas, et al.
Pubblicazione: (2024)
di: Wallat, Jonas, et al.
Pubblicazione: (2024)
Simple Linguistic Inferences of Large Language Models (LLMs): Blind Spots and Blinds
di: Basmov, Victoria, et al.
Pubblicazione: (2023)
di: Basmov, Victoria, et al.
Pubblicazione: (2023)
Fluent but Unfeeling: The Emotional Blind Spots of Language Models
di: Shu, Bangzhao, et al.
Pubblicazione: (2025)
di: Shu, Bangzhao, et al.
Pubblicazione: (2025)
Linguistic Blind Spots of Large Language Models
di: Cheng, Jiali, et al.
Pubblicazione: (2025)
di: Cheng, Jiali, et al.
Pubblicazione: (2025)
Belief in the Machine: Investigating Epistemological Blind Spots of Language Models
di: Suzgun, Mirac, et al.
Pubblicazione: (2024)
di: Suzgun, Mirac, et al.
Pubblicazione: (2024)
Blind Spot Navigation in Large Language Model Reasoning with Thought Space Explorer
di: Zhang, Jinghan, et al.
Pubblicazione: (2024)
di: Zhang, Jinghan, et al.
Pubblicazione: (2024)
Red Teaming for Large Language Models At Scale: Tackling Hallucinations on Mathematics Tasks
di: Buszydlik, Aleksander, et al.
Pubblicazione: (2023)
di: Buszydlik, Aleksander, et al.
Pubblicazione: (2023)
Unchecked and Overlooked: Addressing the Checkbox Blind Spot in Large Language Models with CheckboxQA
di: Turski, Michał, et al.
Pubblicazione: (2025)
di: Turski, Michał, et al.
Pubblicazione: (2025)
Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models
di: Khan, Mohammed Safi Ur Rahman, et al.
Pubblicazione: (2026)
di: Khan, Mohammed Safi Ur Rahman, et al.
Pubblicazione: (2026)
Finding Blind Spots in Evaluator LLMs with Interpretable Checklists
di: Doddapaneni, Sumanth, et al.
Pubblicazione: (2024)
di: Doddapaneni, Sumanth, et al.
Pubblicazione: (2024)
Exploring Mathematical Extrapolation of Large Language Models with Synthetic Data
di: Li, Haolong, et al.
Pubblicazione: (2024)
di: Li, Haolong, et al.
Pubblicazione: (2024)
Linguistic Blind Spots in Clinical Decision Extraction
di: Elgaar, Mohamed, et al.
Pubblicazione: (2026)
di: Elgaar, Mohamed, et al.
Pubblicazione: (2026)
From Input Perception to Predictive Insight: Modeling Model Blind Spots Before They Become Errors
di: Mi, Maggie, et al.
Pubblicazione: (2025)
di: Mi, Maggie, et al.
Pubblicazione: (2025)
The Rarity Blind Spot: A Framework for Evaluating Statistical Reasoning in LLMs
di: Maekawa, Seiji, et al.
Pubblicazione: (2025)
di: Maekawa, Seiji, et al.
Pubblicazione: (2025)
Self-Correction Bench: Uncovering and Addressing the Self-Correction Blind Spot in Large Language Models
di: Tsui, Ken
Pubblicazione: (2025)
di: Tsui, Ken
Pubblicazione: (2025)
Non-Fluent Synthetic Target-Language Data Improve Neural Machine Translation
di: Sánchez-Cartagena, Víctor M., et al.
Pubblicazione: (2024)
di: Sánchez-Cartagena, Víctor M., et al.
Pubblicazione: (2024)
Spot the BlindSpots: Systematic Identification and Quantification of Fine-Grained LLM Biases in Contact Center Summaries
di: Mayilvaghanan, Kawin, et al.
Pubblicazione: (2025)
di: Mayilvaghanan, Kawin, et al.
Pubblicazione: (2025)
Mind the Blind Spots: A Focus-Level Evaluation Framework for LLM Reviews
di: Shin, Hyungyu, et al.
Pubblicazione: (2025)
di: Shin, Hyungyu, et al.
Pubblicazione: (2025)
Blind Spots and Biases: Exploring the Role of Annotator Cognitive Biases in NLP
di: Gautam, Sanjana, et al.
Pubblicazione: (2024)
di: Gautam, Sanjana, et al.
Pubblicazione: (2024)
Evaluating Language Models as Synthetic Data Generators
di: Kim, Seungone, et al.
Pubblicazione: (2024)
di: Kim, Seungone, et al.
Pubblicazione: (2024)
Dynamic Noise Preference Optimization: Self-Improvement of Large Language Models with Self-Synthetic Data
di: Yang, Haoyan, et al.
Pubblicazione: (2025)
di: Yang, Haoyan, et al.
Pubblicazione: (2025)
Unveiling the Flaws: Exploring Imperfections in Synthetic Data and Mitigation Strategies for Large Language Models
di: Chen, Jie, et al.
Pubblicazione: (2024)
di: Chen, Jie, et al.
Pubblicazione: (2024)
Unveiling Cultural Blind Spots: Analyzing the Limitations of mLLMs in Procedural Text Comprehension
di: Yari, Amir Hossein, et al.
Pubblicazione: (2025)
di: Yari, Amir Hossein, et al.
Pubblicazione: (2025)
Target-Aware Language Modeling via Granular Data Sampling
di: Chang, Ernie, et al.
Pubblicazione: (2024)
di: Chang, Ernie, et al.
Pubblicazione: (2024)
Scaling Laws of Synthetic Data for Language Models
di: Qin, Zeyu, et al.
Pubblicazione: (2025)
di: Qin, Zeyu, et al.
Pubblicazione: (2025)
Large Language Models are Algorithmically Blind
di: Venkatesh, Sohan, et al.
Pubblicazione: (2026)
di: Venkatesh, Sohan, et al.
Pubblicazione: (2026)
ZPD-SCA: Unveiling the Blind Spots of LLMs in Assessing Students' Cognitive Abilities
di: Dong, Wenhan, et al.
Pubblicazione: (2025)
di: Dong, Wenhan, et al.
Pubblicazione: (2025)
Neural Topic Modeling with Large Language Models in the Loop
di: Yang, Xiaohao, et al.
Pubblicazione: (2024)
di: Yang, Xiaohao, et al.
Pubblicazione: (2024)
The Autocorrelation Blind Spot: Why 42% of Turn-Level Findings in LLM Conversation Analysis May Be Spurious
di: Schessl, Ferdinand M.
Pubblicazione: (2026)
di: Schessl, Ferdinand M.
Pubblicazione: (2026)
LoopRPT: Reinforcement Pre-Training for Looped Language Models
di: Tang, Guo, et al.
Pubblicazione: (2026)
di: Tang, Guo, et al.
Pubblicazione: (2026)
Synthetic Data for any Differentiable Target
di: Thrush, Tristan, et al.
Pubblicazione: (2026)
di: Thrush, Tristan, et al.
Pubblicazione: (2026)
NagaNLP: Bootstrapping NLP for Low-Resource Nagamese Creole with Human-in-the-Loop Synthetic Data
di: Maiti, Agniva, et al.
Pubblicazione: (2025)
di: Maiti, Agniva, et al.
Pubblicazione: (2025)
Multimodal Language Models Cannot Spot Spatial Inconsistencies
di: Khangaonkar, Om, et al.
Pubblicazione: (2026)
di: Khangaonkar, Om, et al.
Pubblicazione: (2026)
Close the Loop: Synthesizing Infinite Tool-Use Data via Multi-Agent Role-Playing
di: Li, Yuwen, et al.
Pubblicazione: (2025)
di: Li, Yuwen, et al.
Pubblicazione: (2025)
On the Diversity of Synthetic Data and its Impact on Training Large Language Models
di: Chen, Hao, et al.
Pubblicazione: (2024)
di: Chen, Hao, et al.
Pubblicazione: (2024)
Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation
di: Miranda, Lester James V., et al.
Pubblicazione: (2026)
di: Miranda, Lester James V., et al.
Pubblicazione: (2026)
Documenti analoghi
-
Positive Experience Reflection for Agents in Interactive Text Environments
di: Lippmann, Philip, et al.
Pubblicazione: (2024) -
Zero-Shot Contextual Embeddings via Offline Synthetic Corpus Generation
di: Lippmann, Philip, et al.
Pubblicazione: (2025) -
Style over Substance: Distilled Language Models Reason Via Stylistic Replication
di: Lippmann, Philip, et al.
Pubblicazione: (2025) -
Context-Informed Machine Translation of Manga using Multimodal Large Language Models
di: Lippmann, Philip, et al.
Pubblicazione: (2024) -
Temporal Blind Spots in Large Language Models
di: Wallat, Jonas, et al.
Pubblicazione: (2024)