Align and Shine: Building High-Quality Sentence-Aligned Corpora for Multilingual Text Simplification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hilasaca, Kenji, Khallaf, Nouran, Sharoff, Serge |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KAConvText: Novel Approach to Burmese Sentence Classification using Kolmogorov-Arnold Convolution
von: Thu, Ye Kyaw, et al.
Veröffentlicht: (2025)
von: Thu, Ye Kyaw, et al.
Veröffentlicht: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
PRPO: Aligning Process Reward with Outcome Reward in Policy Optimization
von: Ding, Ruiyi, et al.
Veröffentlicht: (2026)
von: Ding, Ruiyi, et al.
Veröffentlicht: (2026)
Induce, Align, Predict: Zero-Shot Stance Detection via Cognitive Inductive Reasoning
von: Zhang, Bowen, et al.
Veröffentlicht: (2025)
von: Zhang, Bowen, et al.
Veröffentlicht: (2025)
Aligning LLMs on a Budget: Inference-Time Alignment with Heuristic Reward Models
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
Automated CAD Modeling Sequence Generation from Text Descriptions via Transformer-Based Large Language Models
von: Liao, Jianxing, et al.
Veröffentlicht: (2025)
von: Liao, Jianxing, et al.
Veröffentlicht: (2025)
HyperPersona: A Multi-Level Hypergraph Framework for Text-Based Automatic Personality Prediction
von: Heydari, Sina, et al.
Veröffentlicht: (2026)
von: Heydari, Sina, et al.
Veröffentlicht: (2026)
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind
von: Srishty, Sharmin Sultana, et al.
Veröffentlicht: (2026)
von: Srishty, Sharmin Sultana, et al.
Veröffentlicht: (2026)
Qtok: A Comprehensive Framework for Evaluating Multilingual Tokenizer Quality in Large Language Models
von: Chelombitko, Iaroslav, et al.
Veröffentlicht: (2024)
von: Chelombitko, Iaroslav, et al.
Veröffentlicht: (2024)
Distilling Self-Consistency into Verbal Confidence: A Pre-Registered Negative Result and Post-Hoc Rescue on Gemma 3 4B
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
Exemplar Retrieval Without Overhypothesis Induction: Limits of Distributional Sequence Learning in Early Word Learning
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
ImmigrationQA: A Source-Grounded Dataset and Small-Model Adaptation for U.S. Immigration Law
von: Shportun, Nazarii
Veröffentlicht: (2026)
von: Shportun, Nazarii
Veröffentlicht: (2026)
Intention Collapse: Intention-Level Metrics for Reasoning in Language Models
von: Vera, Patricio
Veröffentlicht: (2026)
von: Vera, Patricio
Veröffentlicht: (2026)
Whether, Not Which: Mechanistic Interpretability Reveals Dissociable Affect Reception and Emotion Categorization in LLMs
von: Keeman, Michael
Veröffentlicht: (2026)
von: Keeman, Michael
Veröffentlicht: (2026)
When Persuasion Overrides Truth in Multi-Agent LLM Debates: Introducing a Confidence-Weighted Persuasion Override Rate (CW-POR)
von: Agarwal, Mahak, et al.
Veröffentlicht: (2025)
von: Agarwal, Mahak, et al.
Veröffentlicht: (2025)
Why Models Know But Don't Say: Chain-of-Thought Faithfulness Divergence Between Thinking Tokens and Answers in Open-Weight Reasoning Models
von: Young, Richard J.
Veröffentlicht: (2026)
von: Young, Richard J.
Veröffentlicht: (2026)
The Pragmatic Persona: Discovering LLM Persona through Bridging Inference
von: Yang, Jisoo, et al.
Veröffentlicht: (2026)
von: Yang, Jisoo, et al.
Veröffentlicht: (2026)
A Hierarchical Error Framework for Reliable Automated Coding in Communication Research: Applications to Health and Political Communication
von: Zhao, Zhilong, et al.
Veröffentlicht: (2025)
von: Zhao, Zhilong, et al.
Veröffentlicht: (2025)
RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences
von: Zhou, Yangyang, et al.
Veröffentlicht: (2026)
von: Zhou, Yangyang, et al.
Veröffentlicht: (2026)
Assessing Large Language Models on Islamic Legal Reasoning: Evidence from Inheritance Law Evaluation
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2025)
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2025)
UrduBench: An Urdu Reasoning Benchmark using Contextually Ensembled Translations with Human-in-the-Loop
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2026)
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2026)
Can AI Read Between The Lines? Benchmarking LLMs On Financial Nuance
von: Kubica, Dominick, et al.
Veröffentlicht: (2025)
von: Kubica, Dominick, et al.
Veröffentlicht: (2025)
Eyla: Toward an Identity-Anchored LLM Architecture with Integrated Biological Priors -- Vision, Implementation Attempt, and Lessons from AI-Assisted Development
von: Aditto, Arif
Veröffentlicht: (2026)
von: Aditto, Arif
Veröffentlicht: (2026)
Truth as a Compression Artifact in Language Model Training
von: Krestnikov, Konstantin
Veröffentlicht: (2026)
von: Krestnikov, Konstantin
Veröffentlicht: (2026)
SECURA: Sigmoid-Enhanced CUR Decomposition with Uninterrupted Retention and Low-Rank Adaptation in Large Language Models
von: Zhang, Yuxuan
Veröffentlicht: (2025)
von: Zhang, Yuxuan
Veröffentlicht: (2025)
Machine Unlearning for Masked Diffusion Language Models
von: Lee, Georu, et al.
Veröffentlicht: (2026)
von: Lee, Georu, et al.
Veröffentlicht: (2026)
Human Values in a Single Sentence: Moral Presence, Hierarchies, and Transformer Ensembles on the Schwartz Continuum
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
von: Yeste, Víctor, et al.
Veröffentlicht: (2026)
JURY-RL: Votes Propose, Proofs Dispose for Label-Free RLVR
von: Chen, Xinjie, et al.
Veröffentlicht: (2026)
von: Chen, Xinjie, et al.
Veröffentlicht: (2026)
BitCal-TTS: Bit-Calibrated Test-Time Scaling for Quantized Reasoning Models
von: Patarlapalli, Sai Babu, et al.
Veröffentlicht: (2026)
von: Patarlapalli, Sai Babu, et al.
Veröffentlicht: (2026)
Planning vs Reasoning: Ablations to Test Capabilities of LoRA layers
von: Redkar, Neel
Veröffentlicht: (2024)
von: Redkar, Neel
Veröffentlicht: (2024)
CoE: Collaborative Entropy for Uncertainty Quantification in Agentic Multi-LLM Systems
von: Sun, Kangkang, et al.
Veröffentlicht: (2026)
von: Sun, Kangkang, et al.
Veröffentlicht: (2026)
Curveball Steering: The Right Direction To Steer Isn't Always Linear
von: Raval, Shivam, et al.
Veröffentlicht: (2026)
von: Raval, Shivam, et al.
Veröffentlicht: (2026)
Tiny QA Benchmark++: Ultra-Lightweight, Synthetic Multilingual Dataset Generation & Smoke-Tests for Continuous LLM Evaluation
von: Koc, Vincent
Veröffentlicht: (2025)
von: Koc, Vincent
Veröffentlicht: (2025)
On the Influence of Discourse Relations in Persuasive Texts
von: Turk, Nawar, et al.
Veröffentlicht: (2025)
von: Turk, Nawar, et al.
Veröffentlicht: (2025)
Alif: Advancing Urdu Large Language Models via Multilingual Synthetic Data Distillation
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2025)
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2025)
Structured Prompt Optimization Meets Reinforcement Learning for Global and Local Interpretability over Complex Text
von: Zhou, Tianyang, et al.
Veröffentlicht: (2026)
von: Zhou, Tianyang, et al.
Veröffentlicht: (2026)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
von: Gwak, Jiho, et al.
Veröffentlicht: (2025)
von: Gwak, Jiho, et al.
Veröffentlicht: (2025)
MyoSem: Aligning Electromyography to Natural-Language Action Semantics for Hand Action Understanding
von: Wang, Chiyue, et al.
Veröffentlicht: (2026)
von: Wang, Chiyue, et al.
Veröffentlicht: (2026)
EvoIdeator: Evolving Scientific Ideas through Checklist-Grounded Reinforcement Learning
von: Sauter, Andreas, et al.
Veröffentlicht: (2026)
von: Sauter, Andreas, et al.
Veröffentlicht: (2026)
MeMo: Towards Language Models with Associative Memory Mechanisms
von: Zanzotto, Fabio Massimo, et al.
Veröffentlicht: (2025)
von: Zanzotto, Fabio Massimo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
KAConvText: Novel Approach to Burmese Sentence Classification using Kolmogorov-Arnold Convolution
von: Thu, Ye Kyaw, et al.
Veröffentlicht: (2025) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025) -
PRPO: Aligning Process Reward with Outcome Reward in Policy Optimization
von: Ding, Ruiyi, et al.
Veröffentlicht: (2026) -
Induce, Align, Predict: Zero-Shot Stance Detection via Cognitive Inductive Reasoning
von: Zhang, Bowen, et al.
Veröffentlicht: (2025) -
Aligning LLMs on a Budget: Inference-Time Alignment with Heuristic Reward Models
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)