PEARL: Towards Permutation-Resilient LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Liang, Shen, Li, Deng, Yang, Zhao, Xiaoyan, Liang, Bin, Wong, Kam-Fai |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Instance Relation Learning Network with Label Knowledge Propagation for Few-shot Multi-label Intent Detection
di: Zhao, Shiman, et al.
Pubblicazione: (2025)
di: Zhao, Shiman, et al.
Pubblicazione: (2025)
PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning
di: Chang, Qikai, et al.
Pubblicazione: (2026)
di: Chang, Qikai, et al.
Pubblicazione: (2026)
Permute-and-Flip: An optimally stable and watermarkable decoder for LLMs
di: Zhao, Xuandong, et al.
Pubblicazione: (2024)
di: Zhao, Xuandong, et al.
Pubblicazione: (2024)
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
di: Li, Miaomiao, et al.
Pubblicazione: (2025)
di: Li, Miaomiao, et al.
Pubblicazione: (2025)
WHERE and WHICH: Iterative Debate for Biomedical Synthetic Data Augmentation
di: Zhao, Zhengyi, et al.
Pubblicazione: (2025)
di: Zhao, Zhengyi, et al.
Pubblicazione: (2025)
Paying Attention to Facts: Quantifying the Knowledge Capacity of Attention Layers
di: Wong, Liang Ze
Pubblicazione: (2025)
di: Wong, Liang Ze
Pubblicazione: (2025)
Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning
di: Chen, Liang, et al.
Pubblicazione: (2025)
di: Chen, Liang, et al.
Pubblicazione: (2025)
Towards Better Multi-head Attention via Channel-wise Sample Permutation
di: Yuan, Shen, et al.
Pubblicazione: (2024)
di: Yuan, Shen, et al.
Pubblicazione: (2024)
Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning
di: Chen, Liang, et al.
Pubblicazione: (2025)
di: Chen, Liang, et al.
Pubblicazione: (2025)
Sentence Bag Graph Formulation for Biomedical Distant Supervision Relation Extraction
di: Zhang, Hao, et al.
Pubblicazione: (2023)
di: Zhang, Hao, et al.
Pubblicazione: (2023)
FReM: A Flexible Reasoning Mechanism for Balancing Quick and Slow Thinking in Long-Context Question Answering
di: Zhao, Zhengyi, et al.
Pubblicazione: (2025)
di: Zhao, Zhengyi, et al.
Pubblicazione: (2025)
IndiVec: An Exploration of Leveraging Large Language Models for Media Bias Detection with Fine-Grained Bias Indicators
di: Lin, Luyang, et al.
Pubblicazione: (2024)
di: Lin, Luyang, et al.
Pubblicazione: (2024)
Sorbet: A Neuromorphic Hardware-Compatible Transformer-Based Spiking Language Model
di: Tang, Kaiwen, et al.
Pubblicazione: (2024)
di: Tang, Kaiwen, et al.
Pubblicazione: (2024)
Learnable Permutation for Structured Sparsity on Transformer Models
di: Li, Zekai, et al.
Pubblicazione: (2026)
di: Li, Zekai, et al.
Pubblicazione: (2026)
Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning
di: Shen, Junhao, et al.
Pubblicazione: (2026)
di: Shen, Junhao, et al.
Pubblicazione: (2026)
Towards Infinite-Long Prefix in Transformer
di: Liang, Yingyu, et al.
Pubblicazione: (2024)
di: Liang, Yingyu, et al.
Pubblicazione: (2024)
Towards Automated Kernel Generation in the Era of LLMs
di: Yu, Yang, et al.
Pubblicazione: (2026)
di: Yu, Yang, et al.
Pubblicazione: (2026)
WatME: Towards Lossless Watermarking Through Lexical Redundancy
di: Chen, Liang, et al.
Pubblicazione: (2023)
di: Chen, Liang, et al.
Pubblicazione: (2023)
Position-Aware Parameter Efficient Fine-Tuning Approach for Reducing Positional Bias in LLMs
di: Zhang, Zheng, et al.
Pubblicazione: (2024)
di: Zhang, Zheng, et al.
Pubblicazione: (2024)
Can Multimodal LLMs Perform Time Series Anomaly Detection?
di: Xu, Xiongxiao, et al.
Pubblicazione: (2025)
di: Xu, Xiongxiao, et al.
Pubblicazione: (2025)
Benchmarking and Understanding Compositional Relational Reasoning of LLMs
di: Ni, Ruikang, et al.
Pubblicazione: (2024)
di: Ni, Ruikang, et al.
Pubblicazione: (2024)
Late-to-Early Training: LET LLMs Learn Earlier, So Faster and Better
di: Zhao, Ji, et al.
Pubblicazione: (2026)
di: Zhao, Ji, et al.
Pubblicazione: (2026)
RateQuant: Optimal Mixed-Precision KV Cache Quantization via Rate-Distortion Theory
di: Zuo, Fei, et al.
Pubblicazione: (2026)
di: Zuo, Fei, et al.
Pubblicazione: (2026)
CPC-CMS: Cognitive Pairwise Comparison Classification Model Selection Framework for Document-level Sentiment Analysis
di: Li, Jianfei, et al.
Pubblicazione: (2025)
di: Li, Jianfei, et al.
Pubblicazione: (2025)
PoSE: Efficient Context Window Extension of LLMs via Positional Skip-wise Training
di: Zhu, Dawei, et al.
Pubblicazione: (2023)
di: Zhu, Dawei, et al.
Pubblicazione: (2023)
Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling
di: Ren, Liliang, et al.
Pubblicazione: (2024)
di: Ren, Liliang, et al.
Pubblicazione: (2024)
NorMuon: Making Muon more efficient and scalable
di: Li, Zichong, et al.
Pubblicazione: (2025)
di: Li, Zichong, et al.
Pubblicazione: (2025)
PEARL: Prototype-Enhanced Alignment for Label-Efficient Representation Learning with Deployment-Driven Insights from Digital Governance Communication Systems
di: Zhang, Ruiyu, et al.
Pubblicazione: (2026)
di: Zhang, Ruiyu, et al.
Pubblicazione: (2026)
Triplets Better Than Pairs: Towards Stable and Effective Self-Play Fine-Tuning for LLMs
di: Wang, Yibo, et al.
Pubblicazione: (2026)
di: Wang, Yibo, et al.
Pubblicazione: (2026)
ReSURE: Regularizing Supervision Unreliability for Multi-turn Dialogue Fine-tuning
di: Du, Yiming, et al.
Pubblicazione: (2025)
di: Du, Yiming, et al.
Pubblicazione: (2025)
Multi-modal Stance Detection: New Datasets and Model
di: Liang, Bin, et al.
Pubblicazione: (2024)
di: Liang, Bin, et al.
Pubblicazione: (2024)
Toward a Theory of Tokenization in LLMs
di: Rajaraman, Nived, et al.
Pubblicazione: (2024)
di: Rajaraman, Nived, et al.
Pubblicazione: (2024)
RLFR: Extending Reinforcement Learning for LLMs with Flow Environment
di: Zhang, Jinghao, et al.
Pubblicazione: (2025)
di: Zhang, Jinghao, et al.
Pubblicazione: (2025)
Latent-GRPO: Group Relative Policy Optimization for Latent Reasoning
di: Deng, Jingcheng, et al.
Pubblicazione: (2026)
di: Deng, Jingcheng, et al.
Pubblicazione: (2026)
GEMQ: Global Expert-Level Mixed-Precision Quantization for MoE LLMs
di: Deng, Jianing, et al.
Pubblicazione: (2026)
di: Deng, Jianing, et al.
Pubblicazione: (2026)
AutoBencher: Towards Declarative Benchmark Construction
di: Li, Xiang Lisa, et al.
Pubblicazione: (2024)
di: Li, Xiang Lisa, et al.
Pubblicazione: (2024)
Why Are Web AI Agents More Vulnerable Than Standalone LLMs? A Security Analysis
di: Chiang, Jeffrey Yang Fan, et al.
Pubblicazione: (2025)
di: Chiang, Jeffrey Yang Fan, et al.
Pubblicazione: (2025)
Retrieval Backward Attention without Additional Training: Enhance Embeddings of Large Language Models via Repetition
di: Duan, Yifei, et al.
Pubblicazione: (2025)
di: Duan, Yifei, et al.
Pubblicazione: (2025)
Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs
di: Zhao, Siyan, et al.
Pubblicazione: (2025)
di: Zhao, Siyan, et al.
Pubblicazione: (2025)
Quantifying In-Context Reasoning Effects and Memorization Effects in LLMs
di: Lou, Siyu, et al.
Pubblicazione: (2024)
di: Lou, Siyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Instance Relation Learning Network with Label Knowledge Propagation for Few-shot Multi-label Intent Detection
di: Zhao, Shiman, et al.
Pubblicazione: (2025) -
PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning
di: Chang, Qikai, et al.
Pubblicazione: (2026) -
Permute-and-Flip: An optimally stable and watermarkable decoder for LLMs
di: Zhao, Xuandong, et al.
Pubblicazione: (2024) -
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
di: Li, Miaomiao, et al.
Pubblicazione: (2025) -
WHERE and WHICH: Iterative Debate for Biomedical Synthetic Data Augmentation
di: Zhao, Zhengyi, et al.
Pubblicazione: (2025)