Humanline: Online Alignment as Perceptual Loss
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Sijia, Muennighoff, Niklas, Ethayarajh, Kawin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
KTO: Model Alignment as Prospect Theoretic Optimization
di: Ethayarajh, Kawin, et al.
Pubblicazione: (2024)
di: Ethayarajh, Kawin, et al.
Pubblicazione: (2024)
Mecha-nudges for Machines
di: Frey, Giulio, et al.
Pubblicazione: (2026)
di: Frey, Giulio, et al.
Pubblicazione: (2026)
Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
di: Ethayarajh, Kawin, et al.
Pubblicazione: (2021)
di: Ethayarajh, Kawin, et al.
Pubblicazione: (2021)
CP Loss: Channel-wise Perceptual Loss for Time Series Forecasting
di: Zha, Yaohua, et al.
Pubblicazione: (2026)
di: Zha, Yaohua, et al.
Pubblicazione: (2026)
The Scandinavian Embedding Benchmarks: Comprehensive Assessment of Multilingual and Monolingual Text Embedding
di: Enevoldsen, Kenneth, et al.
Pubblicazione: (2024)
di: Enevoldsen, Kenneth, et al.
Pubblicazione: (2024)
LeakBoost: Perceptual-Loss-Based Membership Inference Attack
di: Taub, Amit Kravchik, et al.
Pubblicazione: (2026)
di: Taub, Amit Kravchik, et al.
Pubblicazione: (2026)
Diffusion Model with Perceptual Loss
di: Lin, Shanchuan, et al.
Pubblicazione: (2023)
di: Lin, Shanchuan, et al.
Pubblicazione: (2023)
Whose Alignment? Comparing LLM Process Alignment Across Diverse Organizational Decision Contexts
di: Weller, Niklas, et al.
Pubblicazione: (2026)
di: Weller, Niklas, et al.
Pubblicazione: (2026)
LLM-as-an-Interviewer: Beyond Static Testing Through Dynamic LLM Evaluation
di: Kim, Eunsu, et al.
Pubblicazione: (2024)
di: Kim, Eunsu, et al.
Pubblicazione: (2024)
Spot the BlindSpots: Systematic Identification and Quantification of Fine-Grained LLM Biases in Contact Center Summaries
di: Mayilvaghanan, Kawin, et al.
Pubblicazione: (2025)
di: Mayilvaghanan, Kawin, et al.
Pubblicazione: (2025)
C-Pack: Packed Resources For General Chinese Embeddings
di: Xiao, Shitao, et al.
Pubblicazione: (2023)
di: Xiao, Shitao, et al.
Pubblicazione: (2023)
Synthesizing Images on Perceptual Boundaries of ANNs for Uncovering and Manipulating Human Perceptual Variability
di: Wei, Chen, et al.
Pubblicazione: (2025)
di: Wei, Chen, et al.
Pubblicazione: (2025)
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement
di: Ayalew, Tewodros, et al.
Pubblicazione: (2024)
di: Ayalew, Tewodros, et al.
Pubblicazione: (2024)
RegMix: Data Mixture as Regression for Language Model Pre-training
di: Liu, Qian, et al.
Pubblicazione: (2024)
di: Liu, Qian, et al.
Pubblicazione: (2024)
Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies
di: Tao, Chaofan, et al.
Pubblicazione: (2024)
di: Tao, Chaofan, et al.
Pubblicazione: (2024)
DMA: Online RAG Alignment with Human Feedback
di: Bai, Yu, et al.
Pubblicazione: (2025)
di: Bai, Yu, et al.
Pubblicazione: (2025)
Anchor Points: Benchmarking Models with Much Fewer Examples
di: Vivek, Rajan, et al.
Pubblicazione: (2023)
di: Vivek, Rajan, et al.
Pubblicazione: (2023)
Safety-Preserving PTQ via Contrastive Alignment Loss
di: Wee, Sunghyun, et al.
Pubblicazione: (2025)
di: Wee, Sunghyun, et al.
Pubblicazione: (2025)
Astraios: Parameter-Efficient Instruction Tuning Code Large Language Models
di: Zhuo, Terry Yue, et al.
Pubblicazione: (2024)
di: Zhuo, Terry Yue, et al.
Pubblicazione: (2024)
Data Checklist: On Unit-Testing Datasets with Usable Information
di: Zhang, Heidi C., et al.
Pubblicazione: (2024)
di: Zhang, Heidi C., et al.
Pubblicazione: (2024)
Perceptual Influence: Improving the Perceptual Loss Design for Low-Dose CT Enhancement
di: Viana, Gabriel A., et al.
Pubblicazione: (2025)
di: Viana, Gabriel A., et al.
Pubblicazione: (2025)
Generative Representational Instruction Tuning
di: Muennighoff, Niklas, et al.
Pubblicazione: (2024)
di: Muennighoff, Niklas, et al.
Pubblicazione: (2024)
Towards Unifying Perceptual Reasoning and Logical Reasoning
di: Kido, Hiroyuki
Pubblicazione: (2022)
di: Kido, Hiroyuki
Pubblicazione: (2022)
OctoPack: Instruction Tuning Code Large Language Models
di: Muennighoff, Niklas, et al.
Pubblicazione: (2023)
di: Muennighoff, Niklas, et al.
Pubblicazione: (2023)
A Principled Loss Function for Direct Language Model Alignment
di: Tan, Yuandong
Pubblicazione: (2025)
di: Tan, Yuandong
Pubblicazione: (2025)
A Cortically Inspired Architecture for Modular Perceptual AI
di: Luthra, Prerna
Pubblicazione: (2026)
di: Luthra, Prerna
Pubblicazione: (2026)
Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
di: Zhang, Shenao, et al.
Pubblicazione: (2024)
di: Zhang, Shenao, et al.
Pubblicazione: (2024)
Steering Frozen LLMs: Adaptive Social Alignment via Online Prompt Routing
di: Zhang, Zeyu, et al.
Pubblicazione: (2026)
di: Zhang, Zeyu, et al.
Pubblicazione: (2026)
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning
di: Bhattarai, Manish, et al.
Pubblicazione: (2024)
di: Bhattarai, Manish, et al.
Pubblicazione: (2024)
Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations
di: Singh, Simardeep, et al.
Pubblicazione: (2026)
di: Singh, Simardeep, et al.
Pubblicazione: (2026)
Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
Human Alignment of Large Language Models through Online Preference Optimisation
di: Calandriello, Daniele, et al.
Pubblicazione: (2024)
di: Calandriello, Daniele, et al.
Pubblicazione: (2024)
Harmonizing Multi-Objective LLM Unlearning via Unified Domain Representation and Bidirectional Logit Distillation
di: Zhong, Yisheng, et al.
Pubblicazione: (2026)
di: Zhong, Yisheng, et al.
Pubblicazione: (2026)
TCEval: Using Thermal Comfort to Assess Cognitive and Perceptual Abilities of AI
di: Li, Jingming
Pubblicazione: (2025)
di: Li, Jingming
Pubblicazione: (2025)
rSIM: Incentivizing Reasoning Capabilities of LLMs via Reinforced Strategy Injection
di: Chen, Sijia, et al.
Pubblicazione: (2025)
di: Chen, Sijia, et al.
Pubblicazione: (2025)
Dimensions of Vulnerability in Visual Working Memory: An AI-Driven Approach to Perceptual Comparison
di: Cao, Yuang, et al.
Pubblicazione: (2025)
di: Cao, Yuang, et al.
Pubblicazione: (2025)
Predicting and Enhancing the Fairness of DNNs with the Curvature of Perceptual Manifolds
di: Ma, Yanbiao, et al.
Pubblicazione: (2023)
di: Ma, Yanbiao, et al.
Pubblicazione: (2023)
A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?
di: Zhang, Qiyuan, et al.
Pubblicazione: (2025)
di: Zhang, Qiyuan, et al.
Pubblicazione: (2025)
iCLP: Large Language Model Reasoning with Implicit Cognition Latent Planning
di: Chen, Sijia, et al.
Pubblicazione: (2025)
di: Chen, Sijia, et al.
Pubblicazione: (2025)
Debate as Optimization: Adaptive Conformal Prediction and Diverse Retrieval for Event Extraction
di: Wang, Sijia, et al.
Pubblicazione: (2024)
di: Wang, Sijia, et al.
Pubblicazione: (2024)
Documenti analoghi
-
KTO: Model Alignment as Prospect Theoretic Optimization
di: Ethayarajh, Kawin, et al.
Pubblicazione: (2024) -
Mecha-nudges for Machines
di: Frey, Giulio, et al.
Pubblicazione: (2026) -
Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
di: Ethayarajh, Kawin, et al.
Pubblicazione: (2021) -
CP Loss: Channel-wise Perceptual Loss for Time Series Forecasting
di: Zha, Yaohua, et al.
Pubblicazione: (2026) -
The Scandinavian Embedding Benchmarks: Comprehensive Assessment of Multilingual and Monolingual Text Embedding
di: Enevoldsen, Kenneth, et al.
Pubblicazione: (2024)