Guardado en:
| Autores principales: | Zhang, Zheyu, Yang, Shuo, Kasneci, Gjergji |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2605.31494 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Not All Features Deserve Attention: Graph-Guided Dependency Learning for Tabular Data Generation with Language Models
por: Zhang, Zheyu, et al.
Publicado: (2025)
por: Zhang, Zheyu, et al.
Publicado: (2025)
RAZOR: Sharpening Knowledge by Cutting Bias with Unsupervised Text Rewriting
por: Yang, Shuo, et al.
Publicado: (2024)
por: Yang, Shuo, et al.
Publicado: (2024)
CURE: Controlled Unlearning for Robust Embeddings -- Mitigating Conceptual Shortcuts in Pre-Trained Language Models
por: Kocak, Aysenur, et al.
Publicado: (2025)
por: Kocak, Aysenur, et al.
Publicado: (2025)
Understanding Knowledge Drift in LLMs through Misinformation
por: Fastowski, Alina, et al.
Publicado: (2024)
por: Fastowski, Alina, et al.
Publicado: (2024)
Position: Uncertainty Quantification Needs Reassessment for Large-language Model Agents
por: Kirchhof, Michael, et al.
Publicado: (2025)
por: Kirchhof, Michael, et al.
Publicado: (2025)
Doubling Your Data in Minutes: Ultra-fast Tabular Data Generation via LLM-Induced Dependency Graphs
por: Yang, Shuo, et al.
Publicado: (2025)
por: Yang, Shuo, et al.
Publicado: (2025)
Gender Bias in Explainability: Investigating Performance Disparity in Post-hoc Methods
por: Dhaini, Mahdi, et al.
Publicado: (2025)
por: Dhaini, Mahdi, et al.
Publicado: (2025)
SAGE: Sparse Adaptive Guidance for Dependency-Aware Tabular Data Generation
por: Yang, Shuo, et al.
Publicado: (2026)
por: Yang, Shuo, et al.
Publicado: (2026)
Active Tabular Augmentation via Policy-Guided Diffusion Inpainting
por: Zhang, Zheyu, et al.
Publicado: (2026)
por: Zhang, Zheyu, et al.
Publicado: (2026)
EvalxNLP: A Framework for Benchmarking Post-Hoc Explainability Methods on NLP Models
por: Dhaini, Mahdi, et al.
Publicado: (2025)
por: Dhaini, Mahdi, et al.
Publicado: (2025)
Emergent Abilities in Large Language Models: A Survey
por: Berti, Leonardo, et al.
Publicado: (2025)
por: Berti, Leonardo, et al.
Publicado: (2025)
Is Crowdsourcing Breaking Your Bank? Cost-Effective Fine-Tuning of Pre-trained Language Models with Proximal Policy Optimization
por: Yang, Shuo, et al.
Publicado: (2024)
por: Yang, Shuo, et al.
Publicado: (2024)
Where Paths Split: Localized, Calibrated Control of Moral Reasoning in Large Language Models
por: Yuan, Chenchen, et al.
Publicado: (2026)
por: Yuan, Chenchen, et al.
Publicado: (2026)
Probabilistic Aggregation and Targeted Embedding Optimization for Collective Moral Reasoning in Large Language Models
por: Yuan, Chenchen, et al.
Publicado: (2025)
por: Yuan, Chenchen, et al.
Publicado: (2025)
Attention Mechanisms Don't Learn Additive Models: Rethinking Feature Importance for Transformers
por: Leemann, Tobias, et al.
Publicado: (2024)
por: Leemann, Tobias, et al.
Publicado: (2024)
Enriching Tabular Data with Contextual LLM Embeddings: A Comprehensive Ablation Study for Ensemble Classifiers
por: Kasneci, Gjergji, et al.
Publicado: (2024)
por: Kasneci, Gjergji, et al.
Publicado: (2024)
SCISSOR: Mitigating Semantic Bias through Cluster-Aware Siamese Networks for Robust Classification
por: Yang, Shuo, et al.
Publicado: (2025)
por: Yang, Shuo, et al.
Publicado: (2025)
Entry Dependent Expert Selection in Distributed Gaussian Processes Using Multilabel Classification
por: Jalali, Hamed, et al.
Publicado: (2022)
por: Jalali, Hamed, et al.
Publicado: (2022)
Grokking in the Wild: Data Augmentation for Real-World Multi-Hop Reasoning with Transformers
por: Abramov, Roman, et al.
Publicado: (2025)
por: Abramov, Roman, et al.
Publicado: (2025)
Alternating Reinforcement Learning for Rubric-Based Reward Modeling in Non-Verifiable LLM Post-Training
por: Xu, Ran, et al.
Publicado: (2026)
por: Xu, Ran, et al.
Publicado: (2026)
From Confidence to Collapse in LLM Factual Robustness
por: Fastowski, Alina, et al.
Publicado: (2025)
por: Fastowski, Alina, et al.
Publicado: (2025)
Post-Training Sparse Attention with Double Sparsity
por: Yang, Shuo, et al.
Publicado: (2024)
por: Yang, Shuo, et al.
Publicado: (2024)
RewardHarness: Self-Evolving Agentic Post-Training
por: Zhang, Yuxuan, et al.
Publicado: (2026)
por: Zhang, Yuxuan, et al.
Publicado: (2026)
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
por: Liu, Zihan, et al.
Publicado: (2024)
por: Liu, Zihan, et al.
Publicado: (2024)
Moral Lenses, Political Coordinates: Towards Ideological Positioning of Morally Conditioned LLMs
por: Yuan, Chenchen, et al.
Publicado: (2026)
por: Yuan, Chenchen, et al.
Publicado: (2026)
Enhancing Fairness through Reweighting: A Path to Attain the Sufficiency Rule
por: Zhao, Xuan, et al.
Publicado: (2024)
por: Zhao, Xuan, et al.
Publicado: (2024)
RUBRIC-ARROW: Alternating Pointwise Rubric Reward Modeling for LLM Post-training in Non-verifiable Domains
por: Jiang, Haoxiang, et al.
Publicado: (2026)
por: Jiang, Haoxiang, et al.
Publicado: (2026)
P-TA: Using Proximal Policy Optimization to Enhance Tabular Data Augmentation via Large Language Models
por: Yang, Shuo, et al.
Publicado: (2024)
por: Yang, Shuo, et al.
Publicado: (2024)
Prompt Curriculum Learning for Efficient LLM Post-Training
por: Gao, Zhaolin, et al.
Publicado: (2025)
por: Gao, Zhaolin, et al.
Publicado: (2025)
Rethinking Local Learning: A Cheaper and Faster Recipe for LLM Post-Training
por: Shi, Hengyu, et al.
Publicado: (2026)
por: Shi, Hengyu, et al.
Publicado: (2026)
GRAM-R$^2$: Self-Training Generative Foundation Reward Models for Reward Reasoning
por: Wang, Chenglong, et al.
Publicado: (2025)
por: Wang, Chenglong, et al.
Publicado: (2025)
TLOB: A Novel Transformer Model with Dual Attention for Price Trend Prediction with Limit Order Book Data
por: Berti, Leonardo, et al.
Publicado: (2025)
por: Berti, Leonardo, et al.
Publicado: (2025)
DFPO: Scaling Value Modeling via Distributional Flow towards Robust and Generalizable LLM Post-Training
por: Zhu, Dingwei, et al.
Publicado: (2026)
por: Zhu, Dingwei, et al.
Publicado: (2026)
On Designing Effective RL Reward at Training Time for LLM Reasoning
por: Gao, Jiaxuan, et al.
Publicado: (2024)
por: Gao, Jiaxuan, et al.
Publicado: (2024)
Knowledge Graph-Assisted LLM Post-Training for Enhanced Legal Reasoning
por: Song, Dezhao, et al.
Publicado: (2026)
por: Song, Dezhao, et al.
Publicado: (2026)
Multimodal Behavioral Patterns Analysis with Eye-Tracking and LLM-Based Reasoning
por: Guo, Dongyang, et al.
Publicado: (2025)
por: Guo, Dongyang, et al.
Publicado: (2025)
Pre-Trained Policy Discriminators are General Reward Models
por: Dou, Shihan, et al.
Publicado: (2025)
por: Dou, Shihan, et al.
Publicado: (2025)
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
por: Huang, Wei, et al.
Publicado: (2024)
por: Huang, Wei, et al.
Publicado: (2024)
Reinforcement Unlearning via Group Relative Policy Optimization
por: Zaradoukas, Efstratios, et al.
Publicado: (2026)
por: Zaradoukas, Efstratios, et al.
Publicado: (2026)
Graph Inverse Style Transfer for Counterfactual Explainability
por: Prenkaj, Bardh, et al.
Publicado: (2025)
por: Prenkaj, Bardh, et al.
Publicado: (2025)
Ejemplares similares
-
Not All Features Deserve Attention: Graph-Guided Dependency Learning for Tabular Data Generation with Language Models
por: Zhang, Zheyu, et al.
Publicado: (2025) -
RAZOR: Sharpening Knowledge by Cutting Bias with Unsupervised Text Rewriting
por: Yang, Shuo, et al.
Publicado: (2024) -
CURE: Controlled Unlearning for Robust Embeddings -- Mitigating Conceptual Shortcuts in Pre-Trained Language Models
por: Kocak, Aysenur, et al.
Publicado: (2025) -
Understanding Knowledge Drift in LLMs through Misinformation
por: Fastowski, Alina, et al.
Publicado: (2024) -
Position: Uncertainty Quantification Needs Reassessment for Large-language Model Agents
por: Kirchhof, Michael, et al.
Publicado: (2025)