Synthesizing Instruction-Tuning Datasets with Contrastive Decoding
Fuente:
arXiv
Guardado en:
| Autores principales: | Ichinose, Tatsuya, Ma, Youmi, Oi, Masanari, Koike, Ryuto, Okazaki, Naoaki |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Likelihood-based Mitigation of Evaluation Bias in Large Language Models
por: Oi, Masanari, et al.
Publicado: (2024)
por: Oi, Masanari, et al.
Publicado: (2024)
How You Prompt Matters! Even Task-Oriented Constraints in Instructions Affect LLM-Generated Text Detection
por: Koike, Ryuto, et al.
Publicado: (2023)
por: Koike, Ryuto, et al.
Publicado: (2023)
Building a Japanese Document-Level Relation Extraction Dataset Assisted by Cross-Lingual Transfer
por: Ma, Youmi, et al.
Publicado: (2024)
por: Ma, Youmi, et al.
Publicado: (2024)
From Interpretability to Performance: Optimizing Retrieval Heads for Long-Context Language Models
por: Ma, Youmi, et al.
Publicado: (2026)
por: Ma, Youmi, et al.
Publicado: (2026)
OUTFOX: LLM-Generated Essay Detection Through In-Context Learning with Adversarially Generated Examples
por: Koike, Ryuto, et al.
Publicado: (2023)
por: Koike, Ryuto, et al.
Publicado: (2023)
LLM Output Detectability and Task Performance Can be Jointly Optimized
por: Saito, Koshiro, et al.
Publicado: (2026)
por: Saito, Koshiro, et al.
Publicado: (2026)
From Correspondence to Actions: Human-Like Multi-Image Spatial Reasoning in Multi-modal Large Language Models
por: Oi, Masanari, et al.
Publicado: (2026)
por: Oi, Masanari, et al.
Publicado: (2026)
ExaGPT: Example-Based Machine-Generated Text Detection for Human Interpretability
por: Koike, Ryuto, et al.
Publicado: (2025)
por: Koike, Ryuto, et al.
Publicado: (2025)
Sampling-based Pseudo-Likelihood for Membership Inference Attacks
por: Kaneko, Masahiro, et al.
Publicado: (2024)
por: Kaneko, Masahiro, et al.
Publicado: (2024)
Building Instruction-Tuning Datasets from Human-Written Instructions with Open-Weight Large Language Models
por: Ma, Youmi, et al.
Publicado: (2025)
por: Ma, Youmi, et al.
Publicado: (2025)
Machine Text Detectors are Membership Inference Attacks
por: Koike, Ryuto, et al.
Publicado: (2025)
por: Koike, Ryuto, et al.
Publicado: (2025)
Knowledge of Pretrained Language Models on Surface Information of Tokens
por: Hiraoka, Tatsuya, et al.
Publicado: (2024)
por: Hiraoka, Tatsuya, et al.
Publicado: (2024)
Bit-level BPE: Below the byte boundary
por: Moon, Sangwhan, et al.
Publicado: (2025)
por: Moon, Sangwhan, et al.
Publicado: (2025)
Multi-modal, Multi-task, Multi-criteria Automatic Evaluation with Vision Language Models
por: Ohi, Masanari, et al.
Publicado: (2024)
por: Ohi, Masanari, et al.
Publicado: (2024)
QuantumBench: A Benchmark for Quantum Problem Solving
por: Minami, Shunya, et al.
Publicado: (2025)
por: Minami, Shunya, et al.
Publicado: (2025)
Tokenization as Finite-State Transduction
por: Cognetta, Marco, et al.
Publicado: (2024)
por: Cognetta, Marco, et al.
Publicado: (2024)
Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding
por: Oba, Daisuke, et al.
Publicado: (2026)
por: Oba, Daisuke, et al.
Publicado: (2026)
Decoding-Free Sampling Strategies for LLM Marginalization
por: Pohl, David, et al.
Publicado: (2025)
por: Pohl, David, et al.
Publicado: (2025)
An Analysis of BPE Vocabulary Trimming in Neural Machine Translation
por: Cognetta, Marco, et al.
Publicado: (2024)
por: Cognetta, Marco, et al.
Publicado: (2024)
WAON: A Large-Scale Japanese Image-Text Dataset for Cultural Adaptation in Contrastive Vision-Language Models
por: Sugiura, Issa, et al.
Publicado: (2025)
por: Sugiura, Issa, et al.
Publicado: (2025)
Why We Build Local Large Language Models: An Observational Analysis from 35 Japanese and Multilingual LLMs
por: Saito, Koshiro, et al.
Publicado: (2024)
por: Saito, Koshiro, et al.
Publicado: (2024)
JUBAKU: An Adversarial Benchmark for Exposing Culturally Grounded Stereotypes in Japanese LLMs
por: Shiotani, Taihei, et al.
Publicado: (2026)
por: Shiotani, Taihei, et al.
Publicado: (2026)
A Japanese Benchmark for Evaluating Social Bias in Reasoning Based on Attribution Theory
por: Shiotani, Taihei, et al.
Publicado: (2026)
por: Shiotani, Taihei, et al.
Publicado: (2026)
Distributional Properties of Subword Regularization
por: Cognetta, Marco, et al.
Publicado: (2024)
por: Cognetta, Marco, et al.
Publicado: (2024)
Evaluating Gender Bias of Pre-trained Language Models in Natural Language Inference by Considering All Labels
por: Anantaprayoon, Panatchakorn, et al.
Publicado: (2023)
por: Anantaprayoon, Panatchakorn, et al.
Publicado: (2023)
Solving NLP Problems through Human-System Collaboration: A Discussion-based Approach
por: Kaneko, Masahiro, et al.
Publicado: (2023)
por: Kaneko, Masahiro, et al.
Publicado: (2023)
Social Bias Evaluation for Large Language Models Requires Prompt Variations
por: Hida, Rem, et al.
Publicado: (2024)
por: Hida, Rem, et al.
Publicado: (2024)
SAIE Framework: Support Alone Isn't Enough -- Advancing LLM Training with Adversarial Remarks
por: Loem, Mengsay, et al.
Publicado: (2023)
por: Loem, Mengsay, et al.
Publicado: (2023)
Drifting Objectives for Refining Discrete Diffusion Language Models
por: Oba, Daisuke, et al.
Publicado: (2026)
por: Oba, Daisuke, et al.
Publicado: (2026)
Diffusion-State Policy Optimization for Masked Diffusion Language Models
por: Oba, Daisuke, et al.
Publicado: (2026)
por: Oba, Daisuke, et al.
Publicado: (2026)
Intent-Aware Self-Correction for Mitigating Social Biases in Large Language Models
por: Anantaprayoon, Panatchakorn, et al.
Publicado: (2025)
por: Anantaprayoon, Panatchakorn, et al.
Publicado: (2025)
Constructing Multimodal Datasets from Scratch for Rapid Development of a Japanese Visual Language Model
por: Sasagawa, Keito, et al.
Publicado: (2024)
por: Sasagawa, Keito, et al.
Publicado: (2024)
Autoregressive Direct Preference Optimization
por: Oi, Masanari, et al.
Publicado: (2026)
por: Oi, Masanari, et al.
Publicado: (2026)
Two Counterexamples to Tokenization and the Noiseless Channel
por: Cognetta, Marco, et al.
Publicado: (2024)
por: Cognetta, Marco, et al.
Publicado: (2024)
Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting
por: Kaneko, Masahiro, et al.
Publicado: (2024)
por: Kaneko, Masahiro, et al.
Publicado: (2024)
Aligning Tree-Search Policies with Fixed Token Budgets in Test-Time Scaling of LLMs
por: Miyamoto, Sora, et al.
Publicado: (2026)
por: Miyamoto, Sora, et al.
Publicado: (2026)
Contrastive Instruction Tuning
por: Yan, Tianyi Lorena, et al.
Publicado: (2024)
por: Yan, Tianyi Lorena, et al.
Publicado: (2024)
ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding
por: Zhong, Qihuang, et al.
Publicado: (2024)
por: Zhong, Qihuang, et al.
Publicado: (2024)
UPDESH: Synthesizing Grounded Instruction Tuning Data for 13 Indic Languages
por: Chitale, Pranjal A., et al.
Publicado: (2025)
por: Chitale, Pranjal A., et al.
Publicado: (2025)
What Makes for Good Visual Instructions? Synthesizing Complex Visual Reasoning Instructions for Visual Instruction Tuning
por: Du, Yifan, et al.
Publicado: (2023)
por: Du, Yifan, et al.
Publicado: (2023)
Ejemplares similares
-
Likelihood-based Mitigation of Evaluation Bias in Large Language Models
por: Oi, Masanari, et al.
Publicado: (2024) -
How You Prompt Matters! Even Task-Oriented Constraints in Instructions Affect LLM-Generated Text Detection
por: Koike, Ryuto, et al.
Publicado: (2023) -
Building a Japanese Document-Level Relation Extraction Dataset Assisted by Cross-Lingual Transfer
por: Ma, Youmi, et al.
Publicado: (2024) -
From Interpretability to Performance: Optimizing Retrieval Heads for Long-Context Language Models
por: Ma, Youmi, et al.
Publicado: (2026) -
OUTFOX: LLM-Generated Essay Detection Through In-Context Learning with Adversarially Generated Examples
por: Koike, Ryuto, et al.
Publicado: (2023)