Dataset Featurization: Uncovering Natural Language Features through Unsupervised Data Reconstruction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bravansky, Michal, Kubon, Vaclav, Hariharan, Suhas, Kirk, Robert |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semantic Mastery: Enhancing LLMs with Advanced Natural Language Understanding
von: Hariharan, Mohanakrishnan
Veröffentlicht: (2025)
von: Hariharan, Mohanakrishnan
Veröffentlicht: (2025)
Deep Natural Language Feature Learning for Interpretable Prediction
von: Urrutia, Felipe, et al.
Veröffentlicht: (2023)
von: Urrutia, Felipe, et al.
Veröffentlicht: (2023)
MEQA: A Meta-Evaluation Framework for Question & Answer LLM Benchmarks
von: Veuthey, Jaime Raldua, et al.
Veröffentlicht: (2025)
von: Veuthey, Jaime Raldua, et al.
Veröffentlicht: (2025)
Uncovering Customer Issues through Topological Natural Language Analysis
von: Pi, Shu-Ting, et al.
Veröffentlicht: (2024)
von: Pi, Shu-Ting, et al.
Veröffentlicht: (2024)
Rethinking AI Cultural Alignment
von: Bravansky, Michal, et al.
Veröffentlicht: (2025)
von: Bravansky, Michal, et al.
Veröffentlicht: (2025)
LLMs know their vulnerabilities: Uncover Safety Gaps through Natural Distribution Shifts
von: Ren, Qibing, et al.
Veröffentlicht: (2024)
von: Ren, Qibing, et al.
Veröffentlicht: (2024)
Uncovering Implicit Bias in Large Language Models with Concept Learning Dataset
von: Wang, Leroy Z.
Veröffentlicht: (2025)
von: Wang, Leroy Z.
Veröffentlicht: (2025)
Unnatural Languages Are Not Bugs but Features for LLMs
von: Duan, Keyu, et al.
Veröffentlicht: (2025)
von: Duan, Keyu, et al.
Veröffentlicht: (2025)
Hate Speech Detection using Large Language Models with Data Augmentation and Feature Enhancement
von: Nge, Brian Jing Hong, et al.
Veröffentlicht: (2026)
von: Nge, Brian Jing Hong, et al.
Veröffentlicht: (2026)
Reframing Data Value for Large Language Models Through the Lens of Plausibility
von: Rammal, Mohamad Rida, et al.
Veröffentlicht: (2024)
von: Rammal, Mohamad Rida, et al.
Veröffentlicht: (2024)
Exploring the Personality Traits of LLMs through Latent Features Steering
von: Yang, Shu, et al.
Veröffentlicht: (2024)
von: Yang, Shu, et al.
Veröffentlicht: (2024)
Enhancing Antibiotic Stewardship using a Natural Language Approach for Better Feature Representation
von: Lee, Simon A., et al.
Veröffentlicht: (2024)
von: Lee, Simon A., et al.
Veröffentlicht: (2024)
Applications of Large Language Model Reasoning in Feature Generation
von: Chandra, Dharani
Veröffentlicht: (2025)
von: Chandra, Dharani
Veröffentlicht: (2025)
Advancements in eHealth Data Analytics through Natural Language Processing and Deep Learning
von: Apostol, Elena-Simona, et al.
Veröffentlicht: (2024)
von: Apostol, Elena-Simona, et al.
Veröffentlicht: (2024)
CorrSynth -- A Correlated Sampling Method for Diverse Dataset Generation from LLMs
von: Kowshik, Suhas S, et al.
Veröffentlicht: (2024)
von: Kowshik, Suhas S, et al.
Veröffentlicht: (2024)
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents
von: Costarelli, Anthony, et al.
Veröffentlicht: (2024)
von: Costarelli, Anthony, et al.
Veröffentlicht: (2024)
Synthetic Feature Augmentation Improves Generalization Performance of Language Models
von: Choudhary, Ashok, et al.
Veröffentlicht: (2025)
von: Choudhary, Ashok, et al.
Veröffentlicht: (2025)
AdParaphrase: Paraphrase Dataset for Analyzing Linguistic Features toward Generating Attractive Ad Texts
von: Murakami, Soichiro, et al.
Veröffentlicht: (2025)
von: Murakami, Soichiro, et al.
Veröffentlicht: (2025)
Grounding Synthetic Data Evaluations of Language Models in Unsupervised Document Corpora
von: Majurski, Michael, et al.
Veröffentlicht: (2025)
von: Majurski, Michael, et al.
Veröffentlicht: (2025)
Evaluating Cultural Adaptability of a Large Language Model via Simulation of Synthetic Personas
von: Kwok, Louis, et al.
Veröffentlicht: (2024)
von: Kwok, Louis, et al.
Veröffentlicht: (2024)
Uncovering the Fragility of Trustworthy LLMs through Chinese Textual Ambiguity
von: Wu, Xinwei, et al.
Veröffentlicht: (2025)
von: Wu, Xinwei, et al.
Veröffentlicht: (2025)
Causal Language Control in Multilingual Transformers via Sparse Feature Steering
von: Chou, Cheng-Ting, et al.
Veröffentlicht: (2025)
von: Chou, Cheng-Ting, et al.
Veröffentlicht: (2025)
Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models
von: Wang, Wen, et al.
Veröffentlicht: (2025)
von: Wang, Wen, et al.
Veröffentlicht: (2025)
The Death of Feature Engineering? BERT with Linguistic Features on SQuAD 2.0
von: Li, Jiawei, et al.
Veröffentlicht: (2024)
von: Li, Jiawei, et al.
Veröffentlicht: (2024)
LangFIR: Discovering Sparse Language-Specific Features from Monolingual Data for Language Steering
von: Wong, Sing Hieng, et al.
Veröffentlicht: (2026)
von: Wong, Sing Hieng, et al.
Veröffentlicht: (2026)
ODD: A Benchmark Dataset for the Natural Language Processing based Opioid Related Aberrant Behavior Detection
von: Kwon, Sunjae, et al.
Veröffentlicht: (2023)
von: Kwon, Sunjae, et al.
Veröffentlicht: (2023)
Unsupervised Elicitation of Language Models
von: Wen, Jiaxin, et al.
Veröffentlicht: (2025)
von: Wen, Jiaxin, et al.
Veröffentlicht: (2025)
Uncovering Latent Chain of Thought Vectors in Language Models
von: Zhang, Jason, et al.
Veröffentlicht: (2024)
von: Zhang, Jason, et al.
Veröffentlicht: (2024)
FeRG-LLM : Feature Engineering by Reason Generation Large Language Models
von: Ko, Jeonghyun, et al.
Veröffentlicht: (2025)
von: Ko, Jeonghyun, et al.
Veröffentlicht: (2025)
Sparse Feature Coactivation Reveals Causal Semantic Modules in Large Language Models
von: Deng, Ruixuan, et al.
Veröffentlicht: (2025)
von: Deng, Ruixuan, et al.
Veröffentlicht: (2025)
Inducing Generalization across Languages and Tasks using Featurized Low-Rank Mixtures
von: Lin, Chu-Cheng, et al.
Veröffentlicht: (2024)
von: Lin, Chu-Cheng, et al.
Veröffentlicht: (2024)
Understanding Network Behaviors through Natural Language Question-Answering
von: Xing, Mingzhe, et al.
Veröffentlicht: (2025)
von: Xing, Mingzhe, et al.
Veröffentlicht: (2025)
KatFishNet: Detecting LLM-Generated Korean Text through Linguistic Feature Analysis
von: Park, Shinwoo, et al.
Veröffentlicht: (2025)
von: Park, Shinwoo, et al.
Veröffentlicht: (2025)
Concept-aware Data Construction Improves In-context Learning of Language Models
von: Štefánik, Michal, et al.
Veröffentlicht: (2024)
von: Štefánik, Michal, et al.
Veröffentlicht: (2024)
Querying Structured Data Through Natural Language Using Language Models
von: Valentin-Micu, Hontan, et al.
Veröffentlicht: (2026)
von: Valentin-Micu, Hontan, et al.
Veröffentlicht: (2026)
Panoptic Vision-Language Feature Fields
von: Chen, Haoran, et al.
Veröffentlicht: (2023)
von: Chen, Haoran, et al.
Veröffentlicht: (2023)
On-the-fly Denoising for Data Augmentation in Natural Language Understanding
von: Fang, Tianqing, et al.
Veröffentlicht: (2022)
von: Fang, Tianqing, et al.
Veröffentlicht: (2022)
PAREDA: A Multi-Accent Speech Dataset of Natural Language Processing Research Discussions
von: Jin, Sicheng, et al.
Veröffentlicht: (2026)
von: Jin, Sicheng, et al.
Veröffentlicht: (2026)
LLM Factoscope: Uncovering LLMs' Factual Discernment through Inner States Analysis
von: He, Jinwen, et al.
Veröffentlicht: (2023)
von: He, Jinwen, et al.
Veröffentlicht: (2023)
Attention-Guided Feature Fusion (AGFF) Model for Integrating Statistical and Semantic Features in News Text Classification
von: Zare, Mohammad
Veröffentlicht: (2025)
von: Zare, Mohammad
Veröffentlicht: (2025)
Ähnliche Einträge
-
Semantic Mastery: Enhancing LLMs with Advanced Natural Language Understanding
von: Hariharan, Mohanakrishnan
Veröffentlicht: (2025) -
Deep Natural Language Feature Learning for Interpretable Prediction
von: Urrutia, Felipe, et al.
Veröffentlicht: (2023) -
MEQA: A Meta-Evaluation Framework for Question & Answer LLM Benchmarks
von: Veuthey, Jaime Raldua, et al.
Veröffentlicht: (2025) -
Uncovering Customer Issues through Topological Natural Language Analysis
von: Pi, Shu-Ting, et al.
Veröffentlicht: (2024) -
Rethinking AI Cultural Alignment
von: Bravansky, Michal, et al.
Veröffentlicht: (2025)