Problem Solving Through Human-AI Preference-Based Cooperation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dutta, Subhabrata, Kaufmann, Timo, Glavaš, Goran, Habernal, Ivan, Kersting, Kristian, Kreuter, Frauke, Mezini, Mira, Gurevych, Iryna, Hüllermeier, Eyke, Schuetze, Hinrich |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ObscuraCoder: Powering Efficient Code LM Pre-Training Via Obfuscation Grounding
von: Paul, Indraneil, et al.
Veröffentlicht: (2025)
von: Paul, Indraneil, et al.
Veröffentlicht: (2025)
Expert Preference-based Evaluation of Automated Related Work Generation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2025)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2025)
OCALM: Object-Centric Assessment with Language Models
von: Kaufmann, Timo, et al.
Veröffentlicht: (2024)
von: Kaufmann, Timo, et al.
Veröffentlicht: (2024)
Themis: Training Robust Multilingual Code Reward Models for Flexible Multi-Criteria Scoring
von: Paul, Indraneil, et al.
Veröffentlicht: (2026)
von: Paul, Indraneil, et al.
Veröffentlicht: (2026)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
To share or not to share: What risks would laypeople accept to give sensitive data to differentially-private NLP systems?
von: Weiss, Christopher, et al.
Veröffentlicht: (2023)
von: Weiss, Christopher, et al.
Veröffentlicht: (2023)
ResponseRank: Data-Efficient Reward Modeling through Preference Strength Learning
von: Kaufmann, Timo, et al.
Veröffentlicht: (2025)
von: Kaufmann, Timo, et al.
Veröffentlicht: (2025)
Subjective Code Preferences in Experts and Large Language Models
von: Mokhova, Anna, et al.
Veröffentlicht: (2026)
von: Mokhova, Anna, et al.
Veröffentlicht: (2026)
Is Your Prompt Safe? Investigating Prompt Injection Attacks Against Open-Source LLMs
von: Wang, Jiawen, et al.
Veröffentlicht: (2025)
von: Wang, Jiawen, et al.
Veröffentlicht: (2025)
Inverse Constitutional AI: Compressing Preferences into Principles
von: Findeis, Arduin, et al.
Veröffentlicht: (2024)
von: Findeis, Arduin, et al.
Veröffentlicht: (2024)
Factual Self-Awareness in Language Models: Representation, Robustness, and Scaling
von: Tamoyan, Hovhannes, et al.
Veröffentlicht: (2025)
von: Tamoyan, Hovhannes, et al.
Veröffentlicht: (2025)
Hierarchical Latent Structures in Data Generation Process Unify Mechanistic Phenomena across Scale
von: Rohweder, Jonas, et al.
Veröffentlicht: (2026)
von: Rohweder, Jonas, et al.
Veröffentlicht: (2026)
Reward Modeling for Scientific Writing Evaluation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2026)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2026)
Modular Sentence Encoders: Separating Language Specialization from Cross-Lingual Alignment
von: Huang, Yongxin, et al.
Veröffentlicht: (2024)
von: Huang, Yongxin, et al.
Veröffentlicht: (2024)
A Survey of Reinforcement Learning from Human Feedback
von: Kaufmann, Timo, et al.
Veröffentlicht: (2023)
von: Kaufmann, Timo, et al.
Veröffentlicht: (2023)
Feedback Forensics: A Toolkit to Measure AI Personality
von: Findeis, Arduin, et al.
Veröffentlicht: (2025)
von: Findeis, Arduin, et al.
Veröffentlicht: (2025)
Patches of Nonlinearity: Instruction Vectors in Large Language Models
von: Bigoulaeva, Irina, et al.
Veröffentlicht: (2026)
von: Bigoulaeva, Irina, et al.
Veröffentlicht: (2026)
Calibrated Preference Learning: The Case of Label Ranking
von: Thies, Santo M. A. R., et al.
Veröffentlicht: (2026)
von: Thies, Santo M. A. R., et al.
Veröffentlicht: (2026)
Adaptable Hindsight Experience Replay for Search-Based Learning
von: Vazaios, Alexandros, et al.
Veröffentlicht: (2025)
von: Vazaios, Alexandros, et al.
Veröffentlicht: (2025)
Co-FactChecker: A Framework for Human-AI Collaborative Claim Verification Using Large Reasoning Models
von: Sahnan, Dhruv, et al.
Veröffentlicht: (2026)
von: Sahnan, Dhruv, et al.
Veröffentlicht: (2026)
Amplifying Exploration in Monte-Carlo Tree Search by Focusing on the Unknown
von: Derstroff, Cedric, et al.
Veröffentlicht: (2024)
von: Derstroff, Cedric, et al.
Veröffentlicht: (2024)
Illusion or Algorithm? Investigating Memorization, Emergence, and Symbolic Processing in In-Context Learning
von: Niu, Jingcheng, et al.
Veröffentlicht: (2025)
von: Niu, Jingcheng, et al.
Veröffentlicht: (2025)
Linear Opinion Pooling for Uncertainty Quantification on Graphs
von: Damke, Clemens, et al.
Veröffentlicht: (2024)
von: Damke, Clemens, et al.
Veröffentlicht: (2024)
Distribution Matching for Graph Quantification Under Structural Covariate Shift
von: Damke, Clemens, et al.
Veröffentlicht: (2025)
von: Damke, Clemens, et al.
Veröffentlicht: (2025)
Adjusted Count Quantification Learning on Graphs
von: Damke, Clemens, et al.
Veröffentlicht: (2025)
von: Damke, Clemens, et al.
Veröffentlicht: (2025)
CUQ-GNN: Committee-based Graph Uncertainty Quantification using Posterior Networks
von: Damke, Clemens, et al.
Veröffentlicht: (2024)
von: Damke, Clemens, et al.
Veröffentlicht: (2024)
Aleatoric and Epistemic Uncertainty Measures for Ordinal Classification through Binary Reduction
von: Haas, Stefan, et al.
Veröffentlicht: (2025)
von: Haas, Stefan, et al.
Veröffentlicht: (2025)
Mitigating Label Noise through Data Ambiguation
von: Lienen, Julian, et al.
Veröffentlicht: (2023)
von: Lienen, Julian, et al.
Veröffentlicht: (2023)
Credal Prediction based on Relative Likelihood
von: Löhr, Timo, et al.
Veröffentlicht: (2025)
von: Löhr, Timo, et al.
Veröffentlicht: (2025)
STRICTA: Structured Reasoning in Critical Text Assessment for Peer Review and Beyond
von: Dycke, Nils, et al.
Veröffentlicht: (2024)
von: Dycke, Nils, et al.
Veröffentlicht: (2024)
Random Forest Calibration
von: Shaker, Mohammad Hossein, et al.
Veröffentlicht: (2025)
von: Shaker, Mohammad Hossein, et al.
Veröffentlicht: (2025)
A calibration test for evaluating set-based epistemic uncertainty representations
von: Jürgens, Mira, et al.
Veröffentlicht: (2025)
von: Jürgens, Mira, et al.
Veröffentlicht: (2025)
Is Epistemic Uncertainty Faithfully Represented by Evidential Deep Learning Methods?
von: Jürgens, Mira, et al.
Veröffentlicht: (2024)
von: Jürgens, Mira, et al.
Veröffentlicht: (2024)
From Problem-Solving to Teaching Problem-Solving: Aligning LLMs with Pedagogy using Reinforcement Learning
von: Dinucu-Jianu, David, et al.
Veröffentlicht: (2025)
von: Dinucu-Jianu, David, et al.
Veröffentlicht: (2025)
Unifying Feature-Based Explanations with Functional ANOVA and Cooperative Game Theory
von: Fumagalli, Fabian, et al.
Veröffentlicht: (2024)
von: Fumagalli, Fabian, et al.
Veröffentlicht: (2024)
Efficient Credal Prediction through Decalibration
von: Hofman, Paul, et al.
Veröffentlicht: (2026)
von: Hofman, Paul, et al.
Veröffentlicht: (2026)
Reading Between the Tokens: Improving Preference Predictions through Mechanistic Forecasting
von: Ball, Sarah, et al.
Veröffentlicht: (2026)
von: Ball, Sarah, et al.
Veröffentlicht: (2026)
How reparametrization trick broke differentially-private text representation learning
von: Habernal, Ivan
Veröffentlicht: (2022)
von: Habernal, Ivan
Veröffentlicht: (2022)
Chapter 7 Digital Trace Data
von: Keusch, Florian, et al.
Veröffentlicht: (2021)
von: Keusch, Florian, et al.
Veröffentlicht: (2021)
Quantifying Aleatoric and Epistemic Uncertainty with Proper Scoring Rules
von: Hofman, Paul, et al.
Veröffentlicht: (2024)
von: Hofman, Paul, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ObscuraCoder: Powering Efficient Code LM Pre-Training Via Obfuscation Grounding
von: Paul, Indraneil, et al.
Veröffentlicht: (2025) -
Expert Preference-based Evaluation of Automated Related Work Generation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2025) -
OCALM: Object-Centric Assessment with Language Models
von: Kaufmann, Timo, et al.
Veröffentlicht: (2024) -
Themis: Training Robust Multilingual Code Reward Models for Flexible Multi-Criteria Scoring
von: Paul, Indraneil, et al.
Veröffentlicht: (2026) -
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)