IndoPref: A Multi-Domain Pairwise Preference Dataset for Indonesian
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wiyono, Vanessa Rebecca, Anugraha, David, Purwarianti, Ayu, Winata, Genta Indra |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Momento: Evaluating Persistent Memory and Reasoning with Multi-Session Agentic Conversations
von: Merin, Adril Putra, et al.
Veröffentlicht: (2026)
von: Merin, Adril Putra, et al.
Veröffentlicht: (2026)
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
von: Anugraha, David, et al.
Veröffentlicht: (2024)
von: Anugraha, David, et al.
Veröffentlicht: (2024)
MetaMetrics: Calibrating Metrics For Generation Tasks Using Human Preferences
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
What Causes Knowledge Loss in Multilingual Language Models?
von: Khelli, Maria, et al.
Veröffentlicht: (2025)
von: Khelli, Maria, et al.
Veröffentlicht: (2025)
Enhancing Natural Language Inference Performance with Knowledge Graph for COVID-19 Automated Fact-Checking in Indonesian Language
von: Muharram, Arief Purnama, et al.
Veröffentlicht: (2024)
von: Muharram, Arief Purnama, et al.
Veröffentlicht: (2024)
M4-RAG: A Massive-Scale Multilingual Multi-Cultural Multimodal RAG
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
Towards Efficient and Robust VQA-NLE Data Generation with Large Vision-Language Models
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2024)
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2024)
mR3: Multilingual Rubric-Agnostic Reward Reasoning Models
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
LinguAlchemy: Fusing Typological and Geographical Elements for Unseen Language Generalization
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
TextGames: Learning to Self-Play Text-Based Puzzle Games via Language Model Reasoning
von: Hudi, Frederikus, et al.
Veröffentlicht: (2025)
von: Hudi, Frederikus, et al.
Veröffentlicht: (2025)
T1: A Tool-Oriented Conversational Dataset for Multi-Turn Agentic Planning
von: Chakraborty, Amartya, et al.
Veröffentlicht: (2025)
von: Chakraborty, Amartya, et al.
Veröffentlicht: (2025)
R3: Robust Rubric-Agnostic Reward Models
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
SEADialogues: A Multilingual Culturally Grounded Multi-turn Dialogue Dataset on Southeast Asian Languages
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
Preference Tuning with Human Feedback on Language, Speech, and Vision Tasks: A Survey
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
Could We Have Had Better Multilingual LLMs If English Was Not the Central Language?
von: Diandaru, Ryandito, et al.
Veröffentlicht: (2024)
von: Diandaru, Ryandito, et al.
Veröffentlicht: (2024)
Pref-CTRL: Preference Driven LLM Alignment using Representation Editing
von: Ashrafi, Imranul, et al.
Veröffentlicht: (2026)
von: Ashrafi, Imranul, et al.
Veröffentlicht: (2026)
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
Can Large Language Models Understand, Reason About, and Generate Code-Switched Text?
von: Winata, Genta Indra, et al.
Veröffentlicht: (2026)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2026)
ProxyLM: Predicting Language Model Performance on Multilingual Tasks via Proxy Models
von: Anugraha, David, et al.
Veröffentlicht: (2024)
von: Anugraha, David, et al.
Veröffentlicht: (2024)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
von: Horoi, Stefan, et al.
Veröffentlicht: (2025)
PrefPO: Pairwise Preference Prompt Optimization
von: Singhal, Rahul, et al.
Veröffentlicht: (2026)
von: Singhal, Rahul, et al.
Veröffentlicht: (2026)
Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation
von: Nur'aini, Khumaisa, et al.
Veröffentlicht: (2026)
von: Nur'aini, Khumaisa, et al.
Veröffentlicht: (2026)
MINERS: Multilingual Language Models as Semantic Retrievers
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
PingPong: A Natural Benchmark for Multi-Turn Code-Switching Dialogues
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2026)
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2026)
PrefDisco: Benchmarking Proactive Personalized Reasoning
von: Li, Shuyue Stella, et al.
Veröffentlicht: (2025)
von: Li, Shuyue Stella, et al.
Veröffentlicht: (2025)
DesignPref: Capturing Personal Preferences in Visual Design Generation
von: Peng, Yi-Hao, et al.
Veröffentlicht: (2025)
von: Peng, Yi-Hao, et al.
Veröffentlicht: (2025)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2026)
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2026)
DIAL-SUMMER: A Structured Evaluation Framework of Hierarchical Errors in Dialogue Summaries
von: Ramnath, Sahana, et al.
Veröffentlicht: (2026)
von: Ramnath, Sahana, et al.
Veröffentlicht: (2026)
Benchmarking LLMs for Pairwise Causal Discovery in Biomedical and Multi-Domain Contexts
von: Anuyah, Sydney, et al.
Veröffentlicht: (2026)
von: Anuyah, Sydney, et al.
Veröffentlicht: (2026)
QLESS: A Quantized Approach for Data Valuation and Selection in Large Language Model Fine-Tuning
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
Continual Learning in Machine Speech Chain Using Gradient Episodic Memory
von: Tyndall, Geoffrey, et al.
Veröffentlicht: (2024)
von: Tyndall, Geoffrey, et al.
Veröffentlicht: (2024)
Routing with Generated Data: Annotation-Free LLM Skill Estimation and Expert Selection
von: Niu, Tianyi, et al.
Veröffentlicht: (2026)
von: Niu, Tianyi, et al.
Veröffentlicht: (2026)
Cendol: Open Instruction-tuned Generative Large Language Models for Indonesian Languages
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2024)
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2024)
PrefBench: Evaluating Zero-Shot LLM Agents in Hidden-Preference Personalized Pricing Negotiations
von: Lei, Yingjie
Veröffentlicht: (2026)
von: Lei, Yingjie
Veröffentlicht: (2026)
RainbowPO: A Unified Framework for Combining Improvements in Preference Optimization
von: Zhao, Hanyang, et al.
Veröffentlicht: (2024)
von: Zhao, Hanyang, et al.
Veröffentlicht: (2024)
Datasheets Aren't Enough: DataRubrics for Automated Quality Metrics and Accountability
von: Winata, Genta Indra, et al.
Veröffentlicht: (2025)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2025)
Crosslingual Reasoning through Test-Time Scaling
von: Yong, Zheng-Xin, et al.
Veröffentlicht: (2025)
von: Yong, Zheng-Xin, et al.
Veröffentlicht: (2025)
Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss
von: Xu, Jing, et al.
Veröffentlicht: (2023)
von: Xu, Jing, et al.
Veröffentlicht: (2023)
Linguistics Theory Meets LLM: Code-Switched Text Generation via Equivalence Constrained Large Language Models
von: Kuwanto, Garry, et al.
Veröffentlicht: (2024)
von: Kuwanto, Garry, et al.
Veröffentlicht: (2024)
DriveThru: a Document Extraction Platform and Benchmark Datasets for Indonesian Local Language Archives
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2024)
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Momento: Evaluating Persistent Memory and Reasoning with Multi-Session Agentic Conversations
von: Merin, Adril Putra, et al.
Veröffentlicht: (2026) -
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
von: Anugraha, David, et al.
Veröffentlicht: (2024) -
MetaMetrics: Calibrating Metrics For Generation Tasks Using Human Preferences
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024) -
What Causes Knowledge Loss in Multilingual Language Models?
von: Khelli, Maria, et al.
Veröffentlicht: (2025) -
Enhancing Natural Language Inference Performance with Knowledge Graph for COVID-19 Automated Fact-Checking in Indonesian Language
von: Muharram, Arief Purnama, et al.
Veröffentlicht: (2024)