LCTG Bench: LLM Controlled Text Generation Benchmark
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kurihara, Kentaro, Mita, Masato, Zhang, Peinan, Sasaki, Shota, Ishigami, Ryosuke, Okazaki, Naoaki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Striking Gold in Advertising: Standardization and Exploration of Ad Text Generation
von: Mita, Masato, et al.
Veröffentlicht: (2023)
von: Mita, Masato, et al.
Veröffentlicht: (2023)
AdTEC: A Unified Benchmark for Evaluating Text Quality in Search Engine Advertising
von: Zhang, Peinan, et al.
Veröffentlicht: (2024)
von: Zhang, Peinan, et al.
Veröffentlicht: (2024)
Not Eliminate but Aggregate: Post-Hoc Control over Mixture-of-Experts to Address Shortcut Shifts in Natural Language Understanding
von: Honda, Ukyo, et al.
Veröffentlicht: (2024)
von: Honda, Ukyo, et al.
Veröffentlicht: (2024)
FaithCAMERA: Construction of a Faithful Dataset for Ad Text Generation
von: Kato, Akihiko, et al.
Veröffentlicht: (2024)
von: Kato, Akihiko, et al.
Veröffentlicht: (2024)
How You Prompt Matters! Even Task-Oriented Constraints in Instructions Affect LLM-Generated Text Detection
von: Koike, Ryuto, et al.
Veröffentlicht: (2023)
von: Koike, Ryuto, et al.
Veröffentlicht: (2023)
OUTFOX: LLM-Generated Essay Detection Through In-Context Learning with Adversarially Generated Examples
von: Koike, Ryuto, et al.
Veröffentlicht: (2023)
von: Koike, Ryuto, et al.
Veröffentlicht: (2023)
A Japanese Benchmark for Evaluating Social Bias in Reasoning Based on Attribution Theory
von: Shiotani, Taihei, et al.
Veröffentlicht: (2026)
von: Shiotani, Taihei, et al.
Veröffentlicht: (2026)
SAIE Framework: Support Alone Isn't Enough -- Advancing LLM Training with Adversarial Remarks
von: Loem, Mengsay, et al.
Veröffentlicht: (2023)
von: Loem, Mengsay, et al.
Veröffentlicht: (2023)
From Interpretability to Performance: Optimizing Retrieval Heads for Long-Context Language Models
von: Ma, Youmi, et al.
Veröffentlicht: (2026)
von: Ma, Youmi, et al.
Veröffentlicht: (2026)
Knowledge of Pretrained Language Models on Surface Information of Tokens
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2024)
von: Hiraoka, Tatsuya, et al.
Veröffentlicht: (2024)
Tokenization as Finite-State Transduction
von: Cognetta, Marco, et al.
Veröffentlicht: (2024)
von: Cognetta, Marco, et al.
Veröffentlicht: (2024)
ExaGPT: Example-Based Machine-Generated Text Detection for Human Interpretability
von: Koike, Ryuto, et al.
Veröffentlicht: (2025)
von: Koike, Ryuto, et al.
Veröffentlicht: (2025)
Out-of-the-Box Conditional Text Embeddings from Large Language Models
von: Yamada, Kosuke, et al.
Veröffentlicht: (2025)
von: Yamada, Kosuke, et al.
Veröffentlicht: (2025)
LLM Output Detectability and Task Performance Can be Jointly Optimized
von: Saito, Koshiro, et al.
Veröffentlicht: (2026)
von: Saito, Koshiro, et al.
Veröffentlicht: (2026)
Building a Japanese Document-Level Relation Extraction Dataset Assisted by Cross-Lingual Transfer
von: Ma, Youmi, et al.
Veröffentlicht: (2024)
von: Ma, Youmi, et al.
Veröffentlicht: (2024)
Decoding-Free Sampling Strategies for LLM Marginalization
von: Pohl, David, et al.
Veröffentlicht: (2025)
von: Pohl, David, et al.
Veröffentlicht: (2025)
Bit-level BPE: Below the byte boundary
von: Moon, Sangwhan, et al.
Veröffentlicht: (2025)
von: Moon, Sangwhan, et al.
Veröffentlicht: (2025)
Distributional Properties of Subword Regularization
von: Cognetta, Marco, et al.
Veröffentlicht: (2024)
von: Cognetta, Marco, et al.
Veröffentlicht: (2024)
Evaluating Gender Bias of Pre-trained Language Models in Natural Language Inference by Considering All Labels
von: Anantaprayoon, Panatchakorn, et al.
Veröffentlicht: (2023)
von: Anantaprayoon, Panatchakorn, et al.
Veröffentlicht: (2023)
Solving NLP Problems through Human-System Collaboration: A Discussion-based Approach
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2023)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2023)
Social Bias Evaluation for Large Language Models Requires Prompt Variations
von: Hida, Rem, et al.
Veröffentlicht: (2024)
von: Hida, Rem, et al.
Veröffentlicht: (2024)
Machine Text Detectors are Membership Inference Attacks
von: Koike, Ryuto, et al.
Veröffentlicht: (2025)
von: Koike, Ryuto, et al.
Veröffentlicht: (2025)
JaWildText: A Benchmark for Vision-Language Models on Japanese Scene Text Understanding
von: Maeda, Koki, et al.
Veröffentlicht: (2026)
von: Maeda, Koki, et al.
Veröffentlicht: (2026)
QuantumBench: A Benchmark for Quantum Problem Solving
von: Minami, Shunya, et al.
Veröffentlicht: (2025)
von: Minami, Shunya, et al.
Veröffentlicht: (2025)
PheMT: A Phenomenon-wise Dataset for Machine Translation Robustness on User-Generated Contents
von: Fujii, Ryo, et al.
Veröffentlicht: (2020)
von: Fujii, Ryo, et al.
Veröffentlicht: (2020)
Generating Diverse and High-Quality Texts by Minimum Bayes Risk Decoding
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
Intent-Aware Self-Correction for Mitigating Social Biases in Large Language Models
von: Anantaprayoon, Panatchakorn, et al.
Veröffentlicht: (2025)
von: Anantaprayoon, Panatchakorn, et al.
Veröffentlicht: (2025)
Drifting Objectives for Refining Discrete Diffusion Language Models
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
Diffusion-State Policy Optimization for Masked Diffusion Language Models
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
von: Oba, Daisuke, et al.
Veröffentlicht: (2026)
Developmentally-plausible Working Memory Shapes a Critical Period for Language Acquisition
von: Mita, Masato, et al.
Veröffentlicht: (2025)
von: Mita, Masato, et al.
Veröffentlicht: (2025)
Large Language Models Are State-of-the-Art Evaluator for Grammatical Error Correction
von: Kobayashi, Masamune, et al.
Veröffentlicht: (2024)
von: Kobayashi, Masamune, et al.
Veröffentlicht: (2024)
Revisiting Meta-evaluation for Grammatical Error Correction
von: Kobayashi, Masamune, et al.
Veröffentlicht: (2024)
von: Kobayashi, Masamune, et al.
Veröffentlicht: (2024)
Sampling-based Pseudo-Likelihood for Membership Inference Attacks
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
Two Counterexamples to Tokenization and the Noiseless Channel
von: Cognetta, Marco, et al.
Veröffentlicht: (2024)
von: Cognetta, Marco, et al.
Veröffentlicht: (2024)
Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2024)
Aligning Tree-Search Policies with Fixed Token Budgets in Test-Time Scaling of LLMs
von: Miyamoto, Sora, et al.
Veröffentlicht: (2026)
von: Miyamoto, Sora, et al.
Veröffentlicht: (2026)
AdParaphrase v2.0: Generating Attractive Ad Texts Using a Preference-Annotated Paraphrase Dataset
von: Murakami, Soichiro, et al.
Veröffentlicht: (2025)
von: Murakami, Soichiro, et al.
Veröffentlicht: (2025)
JUBAKU: An Adversarial Benchmark for Exposing Culturally Grounded Stereotypes in Japanese LLMs
von: Shiotani, Taihei, et al.
Veröffentlicht: (2026)
von: Shiotani, Taihei, et al.
Veröffentlicht: (2026)
Continual Pre-Training for Cross-Lingual LLM Adaptation: Enhancing Japanese Language Capabilities
von: Fujii, Kazuki, et al.
Veröffentlicht: (2024)
von: Fujii, Kazuki, et al.
Veröffentlicht: (2024)
WAON: A Large-Scale Japanese Image-Text Dataset for Cultural Adaptation in Contrastive Vision-Language Models
von: Sugiura, Issa, et al.
Veröffentlicht: (2025)
von: Sugiura, Issa, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Striking Gold in Advertising: Standardization and Exploration of Ad Text Generation
von: Mita, Masato, et al.
Veröffentlicht: (2023) -
AdTEC: A Unified Benchmark for Evaluating Text Quality in Search Engine Advertising
von: Zhang, Peinan, et al.
Veröffentlicht: (2024) -
Not Eliminate but Aggregate: Post-Hoc Control over Mixture-of-Experts to Address Shortcut Shifts in Natural Language Understanding
von: Honda, Ukyo, et al.
Veröffentlicht: (2024) -
FaithCAMERA: Construction of a Faithful Dataset for Ad Text Generation
von: Kato, Akihiko, et al.
Veröffentlicht: (2024) -
How You Prompt Matters! Even Task-Oriented Constraints in Instructions Affect LLM-Generated Text Detection
von: Koike, Ryuto, et al.
Veröffentlicht: (2023)