Style Extraction on Text Embeddings Using VAE and Parallel Dataset
Fuente:
arXiv
Saved in:
| Main Authors: | Kong, InJin, Kang, Shinyee, Park, Yuna, Kim, Sooyong, Park, Sanghyun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpiralThinker: Latent Reasoning through an Iterative Process with Text-Latent Interleaving
by: Piao, Shengmin, et al.
Published: (2025)
by: Piao, Shengmin, et al.
Published: (2025)
LitE-SQL: A Lightweight and Efficient Text-to-SQL Framework with Vector-based Schema Linking and Execution-Guided Self-Correction
by: Piao, Shengmin, et al.
Published: (2025)
by: Piao, Shengmin, et al.
Published: (2025)
TinyThinker: Distilling Reasoning through Coarse-to-Fine Knowledge Internalization with Self-Reflection
by: Piao, Shengmin, et al.
Published: (2024)
by: Piao, Shengmin, et al.
Published: (2024)
GeneralThinker: Domain-General Reasoning through Likelihood-Guided Answer-Conditioned Optimization
by: Piao, Shengmin, et al.
Published: (2026)
by: Piao, Shengmin, et al.
Published: (2026)
Thinking with Many Minds: Using Large Language Models for Multi-Perspective Problem-Solving
by: Park, Sanghyun, et al.
Published: (2025)
by: Park, Sanghyun, et al.
Published: (2025)
Detecting Redundant Health Survey Questions Using Language-agnostic BERT Sentence Embedding (LaBSE)
by: Kang, Sunghoon, et al.
Published: (2024)
by: Kang, Sunghoon, et al.
Published: (2024)
Multi-Response Preference Optimization with Augmented Ranking Dataset
by: Gwon, Hansle, et al.
Published: (2024)
by: Gwon, Hansle, et al.
Published: (2024)
Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-to-Speech Models
by: Kang, Jaehoon, et al.
Published: (2026)
by: Kang, Jaehoon, et al.
Published: (2026)
StyleDistance: Stronger Content-Independent Style Embeddings with Synthetic Parallel Examples
by: Patel, Ajay, et al.
Published: (2024)
by: Patel, Ajay, et al.
Published: (2024)
Do Music Preferences Reflect Cultural Values? A Cross-National Analysis Using Music Embedding and World Values Survey
by: Kim, Yongjae, et al.
Published: (2025)
by: Kim, Yongjae, et al.
Published: (2025)
M2S: Multi-turn to Single-turn jailbreak in Red Teaming for LLMs
by: Ha, Junwoo, et al.
Published: (2025)
by: Ha, Junwoo, et al.
Published: (2025)
AnimeScore: A Preference-Based Dataset and Framework for Evaluating Anime-Like Speech Style
by: Park, Joonyong, et al.
Published: (2026)
by: Park, Joonyong, et al.
Published: (2026)
Forecasting Future International Events: A Reliable Dataset for Text-Based Event Modeling
by: Gwak, Daehoon, et al.
Published: (2024)
by: Gwak, Daehoon, et al.
Published: (2024)
Challenging Assumptions in Learning Generic Text Style Embeddings
by: Ostheimer, Phil, et al.
Published: (2025)
by: Ostheimer, Phil, et al.
Published: (2025)
MCS-SQL: Leveraging Multiple Prompts and Multiple-Choice Selection For Text-to-SQL Generation
by: Lee, Dongjun, et al.
Published: (2024)
by: Lee, Dongjun, et al.
Published: (2024)
Multilingual Text Style Transfer: Datasets & Models for Indian Languages
by: Mukherjee, Sourabrata, et al.
Published: (2024)
by: Mukherjee, Sourabrata, et al.
Published: (2024)
Persona Extraction Through Semantic Similarity for Emotional Support Conversation Generation
by: Han, Seunghee, et al.
Published: (2024)
by: Han, Seunghee, et al.
Published: (2024)
TinyStyler: Efficient Few-Shot Text Style Transfer with Authorship Embeddings
by: Horvitz, Zachary, et al.
Published: (2024)
by: Horvitz, Zachary, et al.
Published: (2024)
RSCF: Relation-Semantics Consistent Filter for Entity Embedding of Knowledge Graph
by: Kim, Junsik, et al.
Published: (2025)
by: Kim, Junsik, et al.
Published: (2025)
InstaTrans: An Instruction-Aware Translation Framework for Non-English Instruction Datasets
by: Kim, Yungi, et al.
Published: (2024)
by: Kim, Yungi, et al.
Published: (2024)
DART: An AIGT Detector using AMR of Rephrased Text
by: Park, Hyeonchu, et al.
Published: (2024)
by: Park, Hyeonchu, et al.
Published: (2024)
mStyleDistance: Multilingual Style Embeddings and their Evaluation
by: Qiu, Justin, et al.
Published: (2025)
by: Qiu, Justin, et al.
Published: (2025)
Safe-Embed: Unveiling the Safety-Critical Knowledge of Sentence Encoders
by: Kim, Jinseok, et al.
Published: (2024)
by: Kim, Jinseok, et al.
Published: (2024)
Soft Head Selection for Injecting ICL-Derived Task Embeddings
by: Park, Jungwon, et al.
Published: (2025)
by: Park, Jungwon, et al.
Published: (2025)
Block-wise Codeword Embedding for Reliable Multi-bit Text Watermarking
by: Kim, Joeun, et al.
Published: (2026)
by: Kim, Joeun, et al.
Published: (2026)
Mitigating Semantic Leakage in Cross-lingual Embeddings via Orthogonality Constraint
by: Ki, Dayeon, et al.
Published: (2024)
by: Ki, Dayeon, et al.
Published: (2024)
Subject-level Inference for Realistic Text Anonymization Evaluation
by: Oh, Myeong Seok, et al.
Published: (2026)
by: Oh, Myeong Seok, et al.
Published: (2026)
ObjexMT: Objective Extraction and Metacognitive Calibration for LLM-as-a-Judge under Multi-Turn Jailbreaks
by: Kim, Hyunjun, et al.
Published: (2025)
by: Kim, Hyunjun, et al.
Published: (2025)
SHARE: Shared Memory-Aware Open-Domain Long-Term Dialogue Dataset Constructed from Movie Script
by: Kim, Eunwon, et al.
Published: (2024)
by: Kim, Eunwon, et al.
Published: (2024)
Automatic Extraction of Clausal Embedding Based on Large-Scale English Text Data
by: Carslaw, Iona, et al.
Published: (2025)
by: Carslaw, Iona, et al.
Published: (2025)
PET: An Annotated Dataset for Process Extraction from Natural Language Text
by: Bellan, Patrizio, et al.
Published: (2022)
by: Bellan, Patrizio, et al.
Published: (2022)
Interpretable Depression Detection from Social Media Text Using LLM-Derived Embeddings
by: Kim, Samuel, et al.
Published: (2025)
by: Kim, Samuel, et al.
Published: (2025)
One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL
by: Chae, Hyungjoo, et al.
Published: (2025)
by: Chae, Hyungjoo, et al.
Published: (2025)
SETTP: Style Extraction and Tunable Inference via Dual-level Transferable Prompt Learning
by: Jin, Chunzhen, et al.
Published: (2024)
by: Jin, Chunzhen, et al.
Published: (2024)
Zero2Text: Zero-Training Cross-Domain Inversion Attacks on Textual Embeddings
by: Kim, Doohyun, et al.
Published: (2026)
by: Kim, Doohyun, et al.
Published: (2026)
Investigating the Influence of Prompt-Specific Shortcuts in AI Generated Text Detection
by: Park, Choonghyun, et al.
Published: (2024)
by: Park, Choonghyun, et al.
Published: (2024)
Style-Specific Neurons for Steering LLMs in Text Style Transfer
by: Lai, Wen, et al.
Published: (2024)
by: Lai, Wen, et al.
Published: (2024)
Label-aware Hard Negative Sampling Strategies with Momentum Contrastive Learning for Implicit Hate Speech Detection
by: Kim, Jaehoon, et al.
Published: (2024)
by: Kim, Jaehoon, et al.
Published: (2024)
Frustratingly Simple Prompting-based Text Denoising
by: Park, Jungyeul, et al.
Published: (2024)
by: Park, Jungyeul, et al.
Published: (2024)
Enhancing Automatic Term Extraction with Large Language Models via Syntactic Retrieval
by: Chun, Yongchan, et al.
Published: (2025)
by: Chun, Yongchan, et al.
Published: (2025)
Similar Items
-
SpiralThinker: Latent Reasoning through an Iterative Process with Text-Latent Interleaving
by: Piao, Shengmin, et al.
Published: (2025) -
LitE-SQL: A Lightweight and Efficient Text-to-SQL Framework with Vector-based Schema Linking and Execution-Guided Self-Correction
by: Piao, Shengmin, et al.
Published: (2025) -
TinyThinker: Distilling Reasoning through Coarse-to-Fine Knowledge Internalization with Self-Reflection
by: Piao, Shengmin, et al.
Published: (2024) -
GeneralThinker: Domain-General Reasoning through Likelihood-Guided Answer-Conditioned Optimization
by: Piao, Shengmin, et al.
Published: (2026) -
Thinking with Many Minds: Using Large Language Models for Multi-Perspective Problem-Solving
by: Park, Sanghyun, et al.
Published: (2025)