ARF-RLHF: Adaptive Reward-Following for RLHF through Emotion-Driven Self-Supervision and Trace-Biased Dynamic Optimization
Fuente:
arXiv
Salvato in:
| Autore principale: | Zhang, YuXuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Natural Language Processing: A Comprehensive Practical Guide from Tokenisation to RLHF
di: Arabov, Mullosharaf K.
Pubblicazione: (2026)
di: Arabov, Mullosharaf K.
Pubblicazione: (2026)
Graded Transformers
di: Shaska Sr, Tony
Pubblicazione: (2025)
di: Shaska Sr, Tony
Pubblicazione: (2025)
IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following
di: Sun, Mingrui, et al.
Pubblicazione: (2026)
di: Sun, Mingrui, et al.
Pubblicazione: (2026)
When Are Two RLHF Objectives the Same?
di: Gaikwad, Madhava
Pubblicazione: (2025)
di: Gaikwad, Madhava
Pubblicazione: (2025)
Tool-Genesis: A Task-Driven Tool Creation Benchmark for Self-Evolving Language Agent
di: Xia, Bowei, et al.
Pubblicazione: (2026)
di: Xia, Bowei, et al.
Pubblicazione: (2026)
NeuronSpark: A Spiking Neural Network Language Model with Selective State Space Dynamics
di: Tang, Zhengzheng
Pubblicazione: (2026)
di: Tang, Zhengzheng
Pubblicazione: (2026)
How much do LLMs learn from negative examples?
di: Hamdan, Shadi, et al.
Pubblicazione: (2025)
di: Hamdan, Shadi, et al.
Pubblicazione: (2025)
Communicative Agents for Slideshow Storytelling Video Generation based on LLMs
di: Fan, Jingxing, et al.
Pubblicazione: (2025)
di: Fan, Jingxing, et al.
Pubblicazione: (2025)
When Does Content-Based Routing Work? Representation Requirements for Selective Attention in Hybrid Sequence Models
di: Basu, Abhinaba
Pubblicazione: (2026)
di: Basu, Abhinaba
Pubblicazione: (2026)
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
di: Kim, Heejun, et al.
Pubblicazione: (2026)
di: Kim, Heejun, et al.
Pubblicazione: (2026)
MAcPNN: Mutual Assisted Learning on Data Streams with Temporal Dependence
di: Giannini, Federico, et al.
Pubblicazione: (2026)
di: Giannini, Federico, et al.
Pubblicazione: (2026)
NOTAI.AI: Explainable Detection of Machine-Generated Text via Curvature and Feature Attribution
di: Breneur, Oleksandr Marchenko, et al.
Pubblicazione: (2026)
di: Breneur, Oleksandr Marchenko, et al.
Pubblicazione: (2026)
Rethinking the Multilingual Reasoning Gap with Layer Swap
di: Lasbordes, Maxence, et al.
Pubblicazione: (2026)
di: Lasbordes, Maxence, et al.
Pubblicazione: (2026)
Query-Aware Flow Diffusion for Graph-Based RAG with Retrieval Guarantees
di: Zhou, Zhuoping, et al.
Pubblicazione: (2026)
di: Zhou, Zhuoping, et al.
Pubblicazione: (2026)
Murphys Laws of AI Alignment: Why the Gap Always Wins
di: Gaikwad, Madhava
Pubblicazione: (2025)
di: Gaikwad, Madhava
Pubblicazione: (2025)
Harnessing non-adversarial robustness in large language models
di: Zhou, Qinghua, et al.
Pubblicazione: (2026)
di: Zhou, Qinghua, et al.
Pubblicazione: (2026)
AIPsy-Affect: A Keyword-Free Clinical Stimulus Battery for Mechanistic Interpretability of Emotion in Language Models
di: Keeman, Michael
Pubblicazione: (2026)
di: Keeman, Michael
Pubblicazione: (2026)
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
di: Da, Longchao, et al.
Pubblicazione: (2025)
di: Da, Longchao, et al.
Pubblicazione: (2025)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
di: Chen, Tiejin, et al.
Pubblicazione: (2026)
di: Chen, Tiejin, et al.
Pubblicazione: (2026)
Think Thrice Before You Speak: Dual knowledge-enhanced Theory-of-Mind Reasoning for Persuasive Agents
di: Ma, Minghui, et al.
Pubblicazione: (2026)
di: Ma, Minghui, et al.
Pubblicazione: (2026)
SAGE: Sign-Adaptive Gradient for Memory-Efficient LLM Optimization
di: Lee, Wooin, et al.
Pubblicazione: (2026)
di: Lee, Wooin, et al.
Pubblicazione: (2026)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
di: Yang, Yibo
Pubblicazione: (2025)
di: Yang, Yibo
Pubblicazione: (2025)
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
di: Mitchell, Rupert, et al.
Pubblicazione: (2025)
di: Mitchell, Rupert, et al.
Pubblicazione: (2025)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
di: Jia, Xiao
Pubblicazione: (2026)
di: Jia, Xiao
Pubblicazione: (2026)
Thinking Machines: Mathematical Reasoning in the Age of LLMs
di: Asperti, Andrea, et al.
Pubblicazione: (2025)
di: Asperti, Andrea, et al.
Pubblicazione: (2025)
Constitution or Collapse? Exploring Constitutional AI with Llama 3-8B
di: Zhang, Xue
Pubblicazione: (2025)
di: Zhang, Xue
Pubblicazione: (2025)
XAutoLM: Efficient Fine-Tuning of Language Models via Meta-Learning and AutoML
di: Estevanell-Valladares, Ernesto L., et al.
Pubblicazione: (2025)
di: Estevanell-Valladares, Ernesto L., et al.
Pubblicazione: (2025)
BitSkip: An Empirical Analysis of Quantization and Early Exit Composition in Transformers
di: Bhuvaneswaran, Ramshankar, et al.
Pubblicazione: (2025)
di: Bhuvaneswaran, Ramshankar, et al.
Pubblicazione: (2025)
Fine-tuning of Large Language Models for Constituency Parsing Using a Sequence to Sequence Approach
di: Delgado, Francisco Jose Cortes, et al.
Pubblicazione: (2025)
di: Delgado, Francisco Jose Cortes, et al.
Pubblicazione: (2025)
Doğal Dil İşlemede Tokenizasyon Standartları ve Ölçümü: Türkçe Üzerinden Büyük Dil Modellerinin Karşılaştırmalı Analizi
di: Bayram, M. Ali, et al.
Pubblicazione: (2025)
di: Bayram, M. Ali, et al.
Pubblicazione: (2025)
Büyük Dil Modelleri için TR-MMLU Benchmarkı: Performans Değerlendirmesi, Zorluklar ve İyileştirme Fırsatları
di: Bayram, M. Ali, et al.
Pubblicazione: (2025)
di: Bayram, M. Ali, et al.
Pubblicazione: (2025)
Targeted Lexical Injection: Unlocking Latent Cross-Lingual Alignment in Lugha-Llama via Early-Layer LoRA Fine-Tuning
di: Ngugi, Stanley
Pubblicazione: (2025)
di: Ngugi, Stanley
Pubblicazione: (2025)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
di: Nieth, Björn, et al.
Pubblicazione: (2026)
di: Nieth, Björn, et al.
Pubblicazione: (2026)
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
di: Luijkx, Jelle, et al.
Pubblicazione: (2025)
di: Luijkx, Jelle, et al.
Pubblicazione: (2025)
AI Model for Predicting Binding Affinity of Antidiabetic Compounds Targeting PPAR
di: Aman, La Ode, et al.
Pubblicazione: (2024)
di: Aman, La Ode, et al.
Pubblicazione: (2024)
Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems
di: Berman, Shmuel, et al.
Pubblicazione: (2024)
di: Berman, Shmuel, et al.
Pubblicazione: (2024)
Feature Selection Based on Reinforcement Learning and Hazard State Classification for Magnetic Adhesion Wall-Climbing Robots
di: Ma, Zhen, et al.
Pubblicazione: (2025)
di: Ma, Zhen, et al.
Pubblicazione: (2025)
Theoretical Analysis of Positional Encodings in Transformer Models: Impact on Expressiveness and Generalization
di: Li, Yin
Pubblicazione: (2025)
di: Li, Yin
Pubblicazione: (2025)
Order-Robust Class Incremental Learning: Graph-Driven Dynamic Similarity Grouping
di: Lai, Guannan, et al.
Pubblicazione: (2025)
di: Lai, Guannan, et al.
Pubblicazione: (2025)
A Comparative Study of Feature Selection in Tsetlin Machines
di: Halenka, Vojtech, et al.
Pubblicazione: (2025)
di: Halenka, Vojtech, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Natural Language Processing: A Comprehensive Practical Guide from Tokenisation to RLHF
di: Arabov, Mullosharaf K.
Pubblicazione: (2026) -
Graded Transformers
di: Shaska Sr, Tony
Pubblicazione: (2025) -
IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following
di: Sun, Mingrui, et al.
Pubblicazione: (2026) -
When Are Two RLHF Objectives the Same?
di: Gaikwad, Madhava
Pubblicazione: (2025) -
Tool-Genesis: A Task-Driven Tool Creation Benchmark for Self-Evolving Language Agent
di: Xia, Bowei, et al.
Pubblicazione: (2026)