Harnessing Preference Optimisation in Protein LMs for Hit Maturation in Cell Therapy
Fuente:
arXiv
Saved in:
| Main Authors: | Janocha, Katarzyna, Ling, Annabel, Godson, Alice, Lampi, Yulia, Bornschein, Simon, Hammerla, Nils Y. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Forecasting Application Counts in Talent Acquisition Platforms: Harnessing Multimodal Signals using LMs
by: Kabir, Md Ahsanul, et al.
Published: (2024)
by: Kabir, Md Ahsanul, et al.
Published: (2024)
Reflect-RL: Two-Player Online RL Fine-Tuning for LMs
by: Zhou, Runlong, et al.
Published: (2024)
by: Zhou, Runlong, et al.
Published: (2024)
Don't Retrain, Align: Adapting Autoregressive LMs to Diffusion LMs via Representation Alignment
by: Peng, Fred Zhangzhi, et al.
Published: (2026)
by: Peng, Fred Zhangzhi, et al.
Published: (2026)
Statistical Convergence of Spherical First Hitting Diffusion Models
by: Bienewald, Simon, et al.
Published: (2026)
by: Bienewald, Simon, et al.
Published: (2026)
Transformers for Supervised Online Continual Learning
by: Bornschein, Jorg, et al.
Published: (2024)
by: Bornschein, Jorg, et al.
Published: (2024)
Modeling Adoptive Cell Therapy in Bladder Cancer from Sparse Biological Data using PINNs
by: Olumoyin, Kayode, et al.
Published: (2025)
by: Olumoyin, Kayode, et al.
Published: (2025)
Compositional preference models for aligning LMs
by: Go, Dongyoung, et al.
Published: (2023)
by: Go, Dongyoung, et al.
Published: (2023)
Enabling Approximate Joint Sampling in Diffusion LMs
by: Bansal, Parikshit, et al.
Published: (2025)
by: Bansal, Parikshit, et al.
Published: (2025)
Denoising Autoregressive Representation Learning
by: Li, Yazhe, et al.
Published: (2024)
by: Li, Yazhe, et al.
Published: (2024)
Evaluating Representations with Readout Model Switching
by: Li, Yazhe, et al.
Published: (2023)
by: Li, Yazhe, et al.
Published: (2023)
Preference Learning for AI Alignment: a Causal Perspective
by: Kobalczyk, Katarzyna, et al.
Published: (2025)
by: Kobalczyk, Katarzyna, et al.
Published: (2025)
Training Bilingual LMs with Data Constraints in the Targeted Language
by: Seto, Skyler, et al.
Published: (2024)
by: Seto, Skyler, et al.
Published: (2024)
Entropy-Aligned Decoding of LMs for Better Writing and Reasoning
by: Ahmed, Kareem, et al.
Published: (2026)
by: Ahmed, Kareem, et al.
Published: (2026)
Scalable Quantum Optimisation using HADOF: Hamiltonian Auto-Decomposition Optimisation Framework
by: Sankar, Namasi G, et al.
Published: (2025)
by: Sankar, Namasi G, et al.
Published: (2025)
Human Alignment of Large Language Models through Online Preference Optimisation
by: Calandriello, Daniele, et al.
Published: (2024)
by: Calandriello, Daniele, et al.
Published: (2024)
Preference Guided Iterated Pareto Referent Optimisation for Accessible Route Planning
by: Speziali, Paolo, et al.
Published: (2026)
by: Speziali, Paolo, et al.
Published: (2026)
From Static Structures to Ensembles: Studying and Harnessing Protein Structure Tokenization
by: Liu, Zijing, et al.
Published: (2025)
by: Liu, Zijing, et al.
Published: (2025)
Learning How to Ask: Querying LMs with Mixtures of Soft Prompts
by: Qin, Guanghui, et al.
Published: (2021)
by: Qin, Guanghui, et al.
Published: (2021)
DiffuMamba: High-Throughput Diffusion LMs with Mamba Backbone
by: Singh, Vaibhav, et al.
Published: (2025)
by: Singh, Vaibhav, et al.
Published: (2025)
Diffusion LMs Can Approximate Optimal Infilling Lengths Implicitly
by: Liu, Hengchang, et al.
Published: (2026)
by: Liu, Hengchang, et al.
Published: (2026)
Smaug: Fixing Failure Modes of Preference Optimisation with DPO-Positive
by: Pal, Arka, et al.
Published: (2024)
by: Pal, Arka, et al.
Published: (2024)
DisorderUnetLM: Validating ProteinUnet for efficient protein intrinsic disorder prediction
by: Kotowski, Krzysztof, et al.
Published: (2024)
by: Kotowski, Krzysztof, et al.
Published: (2024)
ProteinOPD: Towards Effective and Efficient Preference Alignment for Protein Design
by: Zhang, Yulin, et al.
Published: (2026)
by: Zhang, Yulin, et al.
Published: (2026)
g-DPO: Scalable Preference Optimization for Protein Language Models
by: Ferragu, Constance, et al.
Published: (2025)
by: Ferragu, Constance, et al.
Published: (2025)
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
by: Zhang, Dylan, et al.
Published: (2024)
by: Zhang, Dylan, et al.
Published: (2024)
Harnessing Collective Structure Knowledge in Data Augmentation for Graph Neural Networks
by: Ma, Rongrong, et al.
Published: (2024)
by: Ma, Rongrong, et al.
Published: (2024)
Hitting Time Isomorphism for Multi-Stage Planning with Foundation Policies
by: Boock, Magnus Victor, et al.
Published: (2026)
by: Boock, Magnus Victor, et al.
Published: (2026)
Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
by: Damani, Mehul, et al.
Published: (2025)
by: Damani, Mehul, et al.
Published: (2025)
Grounded Token Initialization for New Vocabulary in LMs for Generative Recommendation
by: Chen, Daiwei, et al.
Published: (2026)
by: Chen, Daiwei, et al.
Published: (2026)
Insights Into the Inner Workings of Transformer Models for Protein Function Prediction
by: Wenzel, Markus, et al.
Published: (2023)
by: Wenzel, Markus, et al.
Published: (2023)
Many Needles in a Haystack: Active Hit Discovery for Perturbation Experiments
by: Rubbi, Andrea, et al.
Published: (2026)
by: Rubbi, Andrea, et al.
Published: (2026)
When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs
by: Jeong, Soyeong, et al.
Published: (2025)
by: Jeong, Soyeong, et al.
Published: (2025)
Frame-Level Internal Tool Use for Temporal Grounding in Audio LMs
by: An, Joesph, et al.
Published: (2026)
by: An, Joesph, et al.
Published: (2026)
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
by: Schmied, Thomas, et al.
Published: (2025)
by: Schmied, Thomas, et al.
Published: (2025)
Path Integral Optimiser: Global Optimisation via Neural Schrödinger-Föllmer Diffusion
by: McGuinness, Max, et al.
Published: (2025)
by: McGuinness, Max, et al.
Published: (2025)
Addressing the Ecological Fallacy in Larger LMs with Human Context
by: Soni, Nikita, et al.
Published: (2026)
by: Soni, Nikita, et al.
Published: (2026)
Learning Model Parameter Dynamics in a Combination Therapy for Bladder Cancer from Sparse Biological Data
by: Olumoyin, Kayode, et al.
Published: (2025)
by: Olumoyin, Kayode, et al.
Published: (2025)
MUStReason: A Benchmark for Diagnosing Pragmatic Reasoning in Video-LMs for Multimodal Sarcasm Detection
by: Saha, Anisha, et al.
Published: (2025)
by: Saha, Anisha, et al.
Published: (2025)
Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF
by: Xie, Tengyang, et al.
Published: (2024)
by: Xie, Tengyang, et al.
Published: (2024)
TGDPO: Harnessing Token-Level Reward Guidance for Enhancing Direct Preference Optimization
by: Zhu, Mingkang, et al.
Published: (2025)
by: Zhu, Mingkang, et al.
Published: (2025)
Similar Items
-
Forecasting Application Counts in Talent Acquisition Platforms: Harnessing Multimodal Signals using LMs
by: Kabir, Md Ahsanul, et al.
Published: (2024) -
Reflect-RL: Two-Player Online RL Fine-Tuning for LMs
by: Zhou, Runlong, et al.
Published: (2024) -
Don't Retrain, Align: Adapting Autoregressive LMs to Diffusion LMs via Representation Alignment
by: Peng, Fred Zhangzhi, et al.
Published: (2026) -
Statistical Convergence of Spherical First Hitting Diffusion Models
by: Bienewald, Simon, et al.
Published: (2026) -
Transformers for Supervised Online Continual Learning
by: Bornschein, Jorg, et al.
Published: (2024)