DEM: Distribution Edited Model for Training with Mixed Data Distributions
Fuente:
arXiv
Guardado en:
| Autores principales: | Ram, Dhananjay, Rawal, Aditya, Hardalov, Momchil, Pappas, Nikolaos, Zha, Sheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Understanding and Improving Information Preservation in Prompt Compression for LLMs
por: Łajewska, Weronika, et al.
Publicado: (2025)
por: Łajewska, Weronika, et al.
Publicado: (2025)
Detecting Check-Worthy Claims in Political Debates, Speeches, and Interviews Using Audio Data
por: Ivanov, Petar, et al.
Publicado: (2023)
por: Ivanov, Petar, et al.
Publicado: (2023)
Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages
por: Alam, Firoj, et al.
Publicado: (2026)
por: Alam, Firoj, et al.
Publicado: (2026)
LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and Target
por: Hasan, Md Arid, et al.
Publicado: (2025)
por: Hasan, Md Arid, et al.
Publicado: (2025)
PropXplain: Can LLMs Enable Explainable Propaganda Detection?
por: Hasanain, Maram, et al.
Publicado: (2025)
por: Hasanain, Maram, et al.
Publicado: (2025)
Sinhala Transliteration: A Comparative Analysis Between Rule-based and Seq2Seq Approaches
por: De Mel, Yomal, et al.
Publicado: (2024)
por: De Mel, Yomal, et al.
Publicado: (2024)
Large Language Models for Propaganda Span Annotation
por: Hasanain, Maram, et al.
Publicado: (2023)
por: Hasanain, Maram, et al.
Publicado: (2023)
CultranAI at PalmX 2025: Data Augmentation for Cultural Knowledge Representation
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2025)
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2025)
TituLLMs: A Family of Bangla LLMs with Comprehensive Benchmarking
por: Nahin, Shahriar Kabir, et al.
Publicado: (2025)
por: Nahin, Shahriar Kabir, et al.
Publicado: (2025)
LLMeBench: A Flexible Framework for Accelerating LLMs Benchmarking
por: Dalvi, Fahim, et al.
Publicado: (2023)
por: Dalvi, Fahim, et al.
Publicado: (2023)
AraDiCE: Benchmarks for Dialectal and Cultural Capabilities in LLMs
por: Mousi, Basel, et al.
Publicado: (2024)
por: Mousi, Basel, et al.
Publicado: (2024)
NativQA Framework: Enabling LLMs and VLMs with Native, Local, and Everyday Knowledge
por: Alam, Firoj, et al.
Publicado: (2025)
por: Alam, Firoj, et al.
Publicado: (2025)
A Multiple-Fill-in-the-Blank Exam Approach for Enhancing Zero-Resource Hallucination Detection in Large Language Models
por: Munakata, Satoshi, et al.
Publicado: (2024)
por: Munakata, Satoshi, et al.
Publicado: (2024)
OASIS: A Multilingual and Multimodal Dataset for Culturally Grounded Spoken Visual QA
por: Alam, Firoj, et al.
Publicado: (2025)
por: Alam, Firoj, et al.
Publicado: (2025)
Propaganda to Hate: A Multimodal Analysis of Arabic Memes with Multi-Agent LLMs
por: Alam, Firoj, et al.
Publicado: (2024)
por: Alam, Firoj, et al.
Publicado: (2024)
LlamaLens: Specialized Multilingual LLM for Analyzing News and Social Media Content
por: Kmainasi, Mohamed Bayan, et al.
Publicado: (2024)
por: Kmainasi, Mohamed Bayan, et al.
Publicado: (2024)
LAraBench: Benchmarking Arabic AI with Large Language Models
por: Abdelali, Ahmed, et al.
Publicado: (2023)
por: Abdelali, Ahmed, et al.
Publicado: (2023)
NativQA: Multilingual Culturally-Aligned Natural Query for LLMs
por: Hasan, Md. Arid, et al.
Publicado: (2024)
por: Hasan, Md. Arid, et al.
Publicado: (2024)
Beyond MCQ: An Open-Ended Arabic Cultural QA Benchmark with Dialect Variants
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2025)
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2025)
Native vs Non-Native Language Prompting: A Comparative Analysis
por: Kmainasi, Mohamed Bayan, et al.
Publicado: (2024)
por: Kmainasi, Mohamed Bayan, et al.
Publicado: (2024)
GenAI Content Detection Task 2: AI vs. Human -- Academic Essay Authenticity Challenge
por: Chowdhury, Shammur Absar, et al.
Publicado: (2024)
por: Chowdhury, Shammur Absar, et al.
Publicado: (2024)
ThatiAR: Subjectivity Detection in Arabic News Sentences
por: Suwaileh, Reem, et al.
Publicado: (2024)
por: Suwaileh, Reem, et al.
Publicado: (2024)
DRO-InstructZero: Distributionally Robust Prompt Optimization for Large Language Models
por: Li, Yangyang
Publicado: (2025)
por: Li, Yangyang
Publicado: (2025)
Adaptive Engram Memory System for Indonesian Language Model: Generative AI Based on TOBA LM for Batak and Minang Language
por: Situngkir, Hokky, et al.
Publicado: (2026)
por: Situngkir, Hokky, et al.
Publicado: (2026)
Exploiting Web Search Tools of AI Agents for Data Exfiltration
por: Rall, Dennis, et al.
Publicado: (2025)
por: Rall, Dennis, et al.
Publicado: (2025)
Enhancing Trust in LLMs: Algorithms for Comparing and Interpreting LLMs
por: Brown, Nik Bear
Publicado: (2024)
por: Brown, Nik Bear
Publicado: (2024)
Uncovering Uncertainty in Transformer Inference
por: Brothers, Greyson, et al.
Publicado: (2024)
por: Brothers, Greyson, et al.
Publicado: (2024)
Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs
por: Kaiser, Daniel, et al.
Publicado: (2026)
por: Kaiser, Daniel, et al.
Publicado: (2026)
Tokenizations for Austronesian Language Models: study on languages in Indonesia Archipelago
por: Lumbantobing, Andhika Bernard, et al.
Publicado: (2026)
por: Lumbantobing, Andhika Bernard, et al.
Publicado: (2026)
Approaching I/O-optimality for Approximate Attention
por: Papp, Pál András, et al.
Publicado: (2026)
por: Papp, Pál András, et al.
Publicado: (2026)
Multi-Task Instruction Tuning via Data Scheduling for Low-Resource Arabic AudioLLMs
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2026)
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2026)
SAGED: A Holistic Bias-Benchmarking Pipeline for Language Models with Customisable Fairness Calibration
por: Guan, Xin, et al.
Publicado: (2024)
por: Guan, Xin, et al.
Publicado: (2024)
Separate Before You Compress: The WWHO Tokenization Architecture
por: Darshana, Kusal
Publicado: (2026)
por: Darshana, Kusal
Publicado: (2026)
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
por: Palit, Sayon, et al.
Publicado: (2025)
por: Palit, Sayon, et al.
Publicado: (2025)
ArAIEval Shared Task: Propagandistic Techniques Detection in Unimodal and Multimodal Arabic Content
por: Hasanain, Maram, et al.
Publicado: (2024)
por: Hasanain, Maram, et al.
Publicado: (2024)
MENASpeechBank: A Reference Voice Bank with Persona-Conditioned Multi-Turn Conversations for AudioLLMs
por: Ali, Zien Sheikh, et al.
Publicado: (2026)
por: Ali, Zien Sheikh, et al.
Publicado: (2026)
Reference-Guided Verdict: LLMs-as-Judges in Automatic Evaluation of Free-Form QA
por: Badshah, Sher, et al.
Publicado: (2024)
por: Badshah, Sher, et al.
Publicado: (2024)
Train-Attention: Meta-Learning Where to Focus in Continual Knowledge Learning
por: Seo, Yeongbin, et al.
Publicado: (2024)
por: Seo, Yeongbin, et al.
Publicado: (2024)
Drift and selection in LLM text ecosystems
por: Riis, Søren
Publicado: (2026)
por: Riis, Søren
Publicado: (2026)
R-Genie: Reasoning-Guided Generative Image Editing
por: Zhang, Dong, et al.
Publicado: (2025)
por: Zhang, Dong, et al.
Publicado: (2025)
Ejemplares similares
-
Understanding and Improving Information Preservation in Prompt Compression for LLMs
por: Łajewska, Weronika, et al.
Publicado: (2025) -
Detecting Check-Worthy Claims in Political Debates, Speeches, and Interviews Using Audio Data
por: Ivanov, Petar, et al.
Publicado: (2023) -
Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages
por: Alam, Firoj, et al.
Publicado: (2026) -
LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and Target
por: Hasan, Md Arid, et al.
Publicado: (2025) -
PropXplain: Can LLMs Enable Explainable Propaganda Detection?
por: Hasanain, Maram, et al.
Publicado: (2025)