C-Mining: Unsupervised Discovery of Seeds for Cultural Data Synthesis via Geometric Misalignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zeng, Pufan, Liu, Yilun, Dai, Mingchen, Piao, Mengyao, Zhao, Chunguang, Miao, Lingqi, Tao, Shimin, Meng, Weibin, He, Minggui, Liu, Chenxin, Qin, Zhenzhen, Zhang, Li, Ma, Hongxia, Chen, Boxing, Wei, Daimeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
M-DaQ: Retrieving Samples with Multilingual Diversity and Quality for Instruction Fine-Tuning Datasets
von: Zhao, Chunguang, et al.
Veröffentlicht: (2025)
von: Zhao, Chunguang, et al.
Veröffentlicht: (2025)
MIDB: Multilingual Instruction Data Booster for Enhancing Cultural Equality in Multilingual Instruction Synthesis
von: Liu, Yilun, et al.
Veröffentlicht: (2025)
von: Liu, Yilun, et al.
Veröffentlicht: (2025)
RationAnomaly: Log Anomaly Detection with Rationality via Chain-of-Thought and Reinforcement Learning
von: Xu, Song, et al.
Veröffentlicht: (2025)
von: Xu, Song, et al.
Veröffentlicht: (2025)
The GaoYao Benchmark: A Comprehensive Framework for Evaluating Multilingual and Multicultural Abilities of Large Language Models
von: Liu, Yilun, et al.
Veröffentlicht: (2026)
von: Liu, Yilun, et al.
Veröffentlicht: (2026)
Chart Specification: Structural Representations for Incentivizing VLM Reasoning in Chart-to-Code Generation
von: He, Minggui, et al.
Veröffentlicht: (2026)
von: He, Minggui, et al.
Veröffentlicht: (2026)
ELSPR: Evaluator LLM Training Data Self-Purification on Non-Transitive Preferences via Tournament Graph Reconstruction
von: Yu, Yan, et al.
Veröffentlicht: (2025)
von: Yu, Yan, et al.
Veröffentlicht: (2025)
R1-T1: Fully Incentivizing Translation Capability in LLMs via Reasoning Learning
von: He, Minggui, et al.
Veröffentlicht: (2025)
von: He, Minggui, et al.
Veröffentlicht: (2025)
R-Log: Incentivizing Log Analysis Capability in LLMs via Reasoning-based Reinforcement Learning
von: Liu, Yilun, et al.
Veröffentlicht: (2025)
von: Liu, Yilun, et al.
Veröffentlicht: (2025)
LogLM: From Task-based to Instruction-based Automated Log Analysis
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
Adapting Large Language Models to Log Analysis with Interpretable Domain Knowledge
von: Ji, Yuhe, et al.
Veröffentlicht: (2024)
von: Ji, Yuhe, et al.
Veröffentlicht: (2024)
CultureScope: A Dimensional Lens for Probing Cultural Understanding in LLMs
von: Zhang, Jinghao, et al.
Veröffentlicht: (2025)
von: Zhang, Jinghao, et al.
Veröffentlicht: (2025)
Clustering and Ranking: Diversity-preserved Instruction Selection through Expert-aligned Quality Estimation
von: Ge, Yuan, et al.
Veröffentlicht: (2024)
von: Ge, Yuan, et al.
Veröffentlicht: (2024)
Taming Text-to-Image Synthesis for Novices: User-centric Prompt Generation via Multi-turn Guidance
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
An Economic Analysis of Electricity Consumption and Electric Vehicle Adoption in California from 2010 to 2021
von: Qi, Pufan
Veröffentlicht: (2025)
von: Qi, Pufan
Veröffentlicht: (2025)
Improving LLM-based Document-level Machine Translation with Multi-Knowledge Fusion
von: Liu, Bin, et al.
Veröffentlicht: (2025)
von: Liu, Bin, et al.
Veröffentlicht: (2025)
UCorrect: An Unsupervised Framework for Automatic Speech Recognition Error Correction
von: Guo, Jiaxin, et al.
Veröffentlicht: (2024)
von: Guo, Jiaxin, et al.
Veröffentlicht: (2024)
From Handcrafted Features to LLMs: A Brief Survey for Machine Translation Quality Estimation
von: Zhao, Haofei, et al.
Veröffentlicht: (2024)
von: Zhao, Haofei, et al.
Veröffentlicht: (2024)
An Investigation into Value Misalignment in LLM-Generated Texts for Cultural Heritage
von: Bu, Fan, et al.
Veröffentlicht: (2025)
von: Bu, Fan, et al.
Veröffentlicht: (2025)
From Part to Whole: 3D Generative World Model with an Adaptive Structural Hierarchy
von: Du, Bi'an, et al.
Veröffentlicht: (2026)
von: Du, Bi'an, et al.
Veröffentlicht: (2026)
Interpretable Online Log Analysis Using Large Language Models with Prompt Strategies
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
Mitigating Misalignment Contagion by Steering with Implicit Traits
von: Chang, Maria, et al.
Veröffentlicht: (2026)
von: Chang, Maria, et al.
Veröffentlicht: (2026)
A latticed total K-theory
von: An, Qingnan, et al.
Veröffentlicht: (2024)
von: An, Qingnan, et al.
Veröffentlicht: (2024)
Two Intermediate Translations Are Better Than One: Fine-tuning LLMs for Document-level Translation Refinement
von: Dong, Yichen, et al.
Veröffentlicht: (2025)
von: Dong, Yichen, et al.
Veröffentlicht: (2025)
The Axion Helical Misalignment Mechanism
von: Chao, Wei, et al.
Veröffentlicht: (2026)
von: Chao, Wei, et al.
Veröffentlicht: (2026)
Geometry and Perception Guided Gaussians for Multiview-consistent 3D Generation from a Single Image
von: Li, Pufan, et al.
Veröffentlicht: (2025)
von: Li, Pufan, et al.
Veröffentlicht: (2025)
Dynamic Capabilities and Circular Economy Innovations for Achieving NetZero in Mining
von: Vivian Osei, et al.
Veröffentlicht: (2026)
von: Vivian Osei, et al.
Veröffentlicht: (2026)
Flow Structure and Manning Coefficient in Open‐Channel Flows With Staggered Tall‐Short Vegetation
von: Liu Yameng, et al.
Veröffentlicht: (2026)
von: Liu Yameng, et al.
Veröffentlicht: (2026)
Seed-Guided Topic Discovery with Out-of-Vocabulary Seeds
von: Zhang, Yu, et al.
Veröffentlicht: (2022)
von: Zhang, Yu, et al.
Veröffentlicht: (2022)
Geometrical Cross-Attention and Nonvoid Voxelization for Efficient 3D Medical Image Segmentation
von: Yuan, Chenxin, et al.
Veröffentlicht: (2026)
von: Yuan, Chenxin, et al.
Veröffentlicht: (2026)
An Improved Morphology‐Aware Model for Predicting the Settling Velocity of Oncomelania hupensis
von: Lingqi Yi, et al.
Veröffentlicht: (2026)
von: Lingqi Yi, et al.
Veröffentlicht: (2026)
Deep Taxonomic Networks for Unsupervised Hierarchical Prototype Discovery
von: Wang, Zekun, et al.
Veröffentlicht: (2025)
von: Wang, Zekun, et al.
Veröffentlicht: (2025)
Random Forest–Based Coal Mine Roof Displacement Prediction and Application
von: Hongxia Li, et al.
Veröffentlicht: (2025)
von: Hongxia Li, et al.
Veröffentlicht: (2025)
Loong: A Human-Like Long Document Translation Agent with Observe-and-Act Adaptive Context Selection
von: Wang, Yutong, et al.
Veröffentlicht: (2026)
von: Wang, Yutong, et al.
Veröffentlicht: (2026)
Seeing Hate Differently: Hate Subspace Modeling for Culture-Aware Hate Speech Detection
von: Cai, Weibin, et al.
Veröffentlicht: (2025)
von: Cai, Weibin, et al.
Veröffentlicht: (2025)
PLATZ transcription factors and their emerging roles in plant responses to environmental stresses.
von: Zhang, Hongxia, et al.
Veröffentlicht: (2025)
von: Zhang, Hongxia, et al.
Veröffentlicht: (2025)
Unsupervised Concept Discovery Mitigates Spurious Correlations
von: Arefin, Md Rifat, et al.
Veröffentlicht: (2024)
von: Arefin, Md Rifat, et al.
Veröffentlicht: (2024)
CoachLM: Automatic Instruction Revisions Improve the Data Quality in LLM Instruction Tuning
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
Unlocking Fine-Grained Translation Quality Estimation in LRMs through Synergistically Evolving Implicit and Explicit Reasoning
von: Dang, Renfei, et al.
Veröffentlicht: (2026)
von: Dang, Renfei, et al.
Veröffentlicht: (2026)
DeMPT: Decoding-enhanced Multi-phase Prompt Tuning for Making LLMs Be Better Context-aware Translators
von: Lyu, Xinglin, et al.
Veröffentlicht: (2024)
von: Lyu, Xinglin, et al.
Veröffentlicht: (2024)
Cross-Preference Learning for Sentence-Level and Context-Aware Machine Translation
von: Li, Ying, et al.
Veröffentlicht: (2026)
von: Li, Ying, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
M-DaQ: Retrieving Samples with Multilingual Diversity and Quality for Instruction Fine-Tuning Datasets
von: Zhao, Chunguang, et al.
Veröffentlicht: (2025) -
MIDB: Multilingual Instruction Data Booster for Enhancing Cultural Equality in Multilingual Instruction Synthesis
von: Liu, Yilun, et al.
Veröffentlicht: (2025) -
RationAnomaly: Log Anomaly Detection with Rationality via Chain-of-Thought and Reinforcement Learning
von: Xu, Song, et al.
Veröffentlicht: (2025) -
The GaoYao Benchmark: A Comprehensive Framework for Evaluating Multilingual and Multicultural Abilities of Large Language Models
von: Liu, Yilun, et al.
Veröffentlicht: (2026) -
Chart Specification: Structural Representations for Incentivizing VLM Reasoning in Chart-to-Code Generation
von: He, Minggui, et al.
Veröffentlicht: (2026)