Humanline: Online Alignment as Perceptual Loss
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Sijia, Muennighoff, Niklas, Ethayarajh, Kawin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KTO: Model Alignment as Prospect Theoretic Optimization
von: Ethayarajh, Kawin, et al.
Veröffentlicht: (2024)
von: Ethayarajh, Kawin, et al.
Veröffentlicht: (2024)
Mecha-nudges for Machines
von: Frey, Giulio, et al.
Veröffentlicht: (2026)
von: Frey, Giulio, et al.
Veröffentlicht: (2026)
Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
von: Ethayarajh, Kawin, et al.
Veröffentlicht: (2021)
von: Ethayarajh, Kawin, et al.
Veröffentlicht: (2021)
CP Loss: Channel-wise Perceptual Loss for Time Series Forecasting
von: Zha, Yaohua, et al.
Veröffentlicht: (2026)
von: Zha, Yaohua, et al.
Veröffentlicht: (2026)
The Scandinavian Embedding Benchmarks: Comprehensive Assessment of Multilingual and Monolingual Text Embedding
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2024)
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2024)
LeakBoost: Perceptual-Loss-Based Membership Inference Attack
von: Taub, Amit Kravchik, et al.
Veröffentlicht: (2026)
von: Taub, Amit Kravchik, et al.
Veröffentlicht: (2026)
Diffusion Model with Perceptual Loss
von: Lin, Shanchuan, et al.
Veröffentlicht: (2023)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2023)
Whose Alignment? Comparing LLM Process Alignment Across Diverse Organizational Decision Contexts
von: Weller, Niklas, et al.
Veröffentlicht: (2026)
von: Weller, Niklas, et al.
Veröffentlicht: (2026)
LLM-as-an-Interviewer: Beyond Static Testing Through Dynamic LLM Evaluation
von: Kim, Eunsu, et al.
Veröffentlicht: (2024)
von: Kim, Eunsu, et al.
Veröffentlicht: (2024)
Spot the BlindSpots: Systematic Identification and Quantification of Fine-Grained LLM Biases in Contact Center Summaries
von: Mayilvaghanan, Kawin, et al.
Veröffentlicht: (2025)
von: Mayilvaghanan, Kawin, et al.
Veröffentlicht: (2025)
C-Pack: Packed Resources For General Chinese Embeddings
von: Xiao, Shitao, et al.
Veröffentlicht: (2023)
von: Xiao, Shitao, et al.
Veröffentlicht: (2023)
Synthesizing Images on Perceptual Boundaries of ANNs for Uncovering and Manipulating Human Perceptual Variability
von: Wei, Chen, et al.
Veröffentlicht: (2025)
von: Wei, Chen, et al.
Veröffentlicht: (2025)
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement
von: Ayalew, Tewodros, et al.
Veröffentlicht: (2024)
von: Ayalew, Tewodros, et al.
Veröffentlicht: (2024)
RegMix: Data Mixture as Regression for Language Model Pre-training
von: Liu, Qian, et al.
Veröffentlicht: (2024)
von: Liu, Qian, et al.
Veröffentlicht: (2024)
Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies
von: Tao, Chaofan, et al.
Veröffentlicht: (2024)
von: Tao, Chaofan, et al.
Veröffentlicht: (2024)
DMA: Online RAG Alignment with Human Feedback
von: Bai, Yu, et al.
Veröffentlicht: (2025)
von: Bai, Yu, et al.
Veröffentlicht: (2025)
Anchor Points: Benchmarking Models with Much Fewer Examples
von: Vivek, Rajan, et al.
Veröffentlicht: (2023)
von: Vivek, Rajan, et al.
Veröffentlicht: (2023)
Safety-Preserving PTQ via Contrastive Alignment Loss
von: Wee, Sunghyun, et al.
Veröffentlicht: (2025)
von: Wee, Sunghyun, et al.
Veröffentlicht: (2025)
Astraios: Parameter-Efficient Instruction Tuning Code Large Language Models
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2024)
von: Zhuo, Terry Yue, et al.
Veröffentlicht: (2024)
Data Checklist: On Unit-Testing Datasets with Usable Information
von: Zhang, Heidi C., et al.
Veröffentlicht: (2024)
von: Zhang, Heidi C., et al.
Veröffentlicht: (2024)
Perceptual Influence: Improving the Perceptual Loss Design for Low-Dose CT Enhancement
von: Viana, Gabriel A., et al.
Veröffentlicht: (2025)
von: Viana, Gabriel A., et al.
Veröffentlicht: (2025)
Generative Representational Instruction Tuning
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
Towards Unifying Perceptual Reasoning and Logical Reasoning
von: Kido, Hiroyuki
Veröffentlicht: (2022)
von: Kido, Hiroyuki
Veröffentlicht: (2022)
OctoPack: Instruction Tuning Code Large Language Models
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2023)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2023)
A Principled Loss Function for Direct Language Model Alignment
von: Tan, Yuandong
Veröffentlicht: (2025)
von: Tan, Yuandong
Veröffentlicht: (2025)
A Cortically Inspired Architecture for Modular Perceptual AI
von: Luthra, Prerna
Veröffentlicht: (2026)
von: Luthra, Prerna
Veröffentlicht: (2026)
Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
von: Zhang, Shenao, et al.
Veröffentlicht: (2024)
von: Zhang, Shenao, et al.
Veröffentlicht: (2024)
Steering Frozen LLMs: Adaptive Social Alignment via Online Prompt Routing
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning
von: Bhattarai, Manish, et al.
Veröffentlicht: (2024)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2024)
Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations
von: Singh, Simardeep, et al.
Veröffentlicht: (2026)
von: Singh, Simardeep, et al.
Veröffentlicht: (2026)
Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms
von: Liu, Jingyi, et al.
Veröffentlicht: (2026)
von: Liu, Jingyi, et al.
Veröffentlicht: (2026)
Human Alignment of Large Language Models through Online Preference Optimisation
von: Calandriello, Daniele, et al.
Veröffentlicht: (2024)
von: Calandriello, Daniele, et al.
Veröffentlicht: (2024)
Harmonizing Multi-Objective LLM Unlearning via Unified Domain Representation and Bidirectional Logit Distillation
von: Zhong, Yisheng, et al.
Veröffentlicht: (2026)
von: Zhong, Yisheng, et al.
Veröffentlicht: (2026)
TCEval: Using Thermal Comfort to Assess Cognitive and Perceptual Abilities of AI
von: Li, Jingming
Veröffentlicht: (2025)
von: Li, Jingming
Veröffentlicht: (2025)
rSIM: Incentivizing Reasoning Capabilities of LLMs via Reinforced Strategy Injection
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
Dimensions of Vulnerability in Visual Working Memory: An AI-Driven Approach to Perceptual Comparison
von: Cao, Yuang, et al.
Veröffentlicht: (2025)
von: Cao, Yuang, et al.
Veröffentlicht: (2025)
Predicting and Enhancing the Fairness of DNNs with the Curvature of Perceptual Manifolds
von: Ma, Yanbiao, et al.
Veröffentlicht: (2023)
von: Ma, Yanbiao, et al.
Veröffentlicht: (2023)
A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
iCLP: Large Language Model Reasoning with Implicit Cognition Latent Planning
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
Debate as Optimization: Adaptive Conformal Prediction and Diverse Retrieval for Event Extraction
von: Wang, Sijia, et al.
Veröffentlicht: (2024)
von: Wang, Sijia, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
KTO: Model Alignment as Prospect Theoretic Optimization
von: Ethayarajh, Kawin, et al.
Veröffentlicht: (2024) -
Mecha-nudges for Machines
von: Frey, Giulio, et al.
Veröffentlicht: (2026) -
Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
von: Ethayarajh, Kawin, et al.
Veröffentlicht: (2021) -
CP Loss: Channel-wise Perceptual Loss for Time Series Forecasting
von: Zha, Yaohua, et al.
Veröffentlicht: (2026) -
The Scandinavian Embedding Benchmarks: Comprehensive Assessment of Multilingual and Monolingual Text Embedding
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2024)