Propensity Inference: Environmental Contributors to LLM Behaviour
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Järviniemi, Olli, Makins, Oliver, Merizian, Jacob, Kirk, Robert, Millwood, Ben |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Uncovering Deceptive Tendencies in Language Models: A Simulated Company AI Assistant
von: Järviniemi, Olli, et al.
Veröffentlicht: (2024)
von: Järviniemi, Olli, et al.
Veröffentlicht: (2024)
Subversion via Focal Points: Investigating Collusion in LLM Monitoring
von: Järviniemi, Olli
Veröffentlicht: (2025)
von: Järviniemi, Olli
Veröffentlicht: (2025)
Vocabulary Dropout for Curriculum Diversity in LLM Co-Evolution
von: Dineen, Jacob, et al.
Veröffentlicht: (2026)
von: Dineen, Jacob, et al.
Veröffentlicht: (2026)
SafeCtrl-RL: Inference-Time Adaptive Behaviour Control for LLM Dialogue via RL-Driven Prompt Optimisation
von: Orme, Michael, et al.
Veröffentlicht: (2026)
von: Orme, Michael, et al.
Veröffentlicht: (2026)
The Knowledge-Behaviour Disconnect in LLM-based Chatbots
von: Broersen, Jan
Veröffentlicht: (2025)
von: Broersen, Jan
Veröffentlicht: (2025)
XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models
von: Röttger, Paul, et al.
Veröffentlicht: (2023)
von: Röttger, Paul, et al.
Veröffentlicht: (2023)
Investigating Non-Transitivity in LLM-as-a-Judge
von: Xu, Yi, et al.
Veröffentlicht: (2025)
von: Xu, Yi, et al.
Veröffentlicht: (2025)
UK AISI Alignment Evaluation Case-Study
von: Souly, Alexandra, et al.
Veröffentlicht: (2026)
von: Souly, Alexandra, et al.
Veröffentlicht: (2026)
CHIMERA: Compact Synthetic Data for Generalizable LLM Reasoning
von: Zhu, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhu, Xinyu, et al.
Veröffentlicht: (2026)
Communication Compression for Tensor Parallel LLM Inference
von: Hansen-Palmus, Jan, et al.
Veröffentlicht: (2024)
von: Hansen-Palmus, Jan, et al.
Veröffentlicht: (2024)
PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing
von: Żurawicki, Krzysztof, et al.
Veröffentlicht: (2026)
von: Żurawicki, Krzysztof, et al.
Veröffentlicht: (2026)
STACK: Adversarial Attacks on LLM Safeguard Pipelines
von: McKenzie, Ian R., et al.
Veröffentlicht: (2025)
von: McKenzie, Ian R., et al.
Veröffentlicht: (2025)
Next Token Knowledge Tracing: Exploiting Pretrained LLM Representations to Decode Student Behaviour
von: Norris, Max, et al.
Veröffentlicht: (2025)
von: Norris, Max, et al.
Veröffentlicht: (2025)
Refusal Steering: Fine-grained Control over LLM Refusal Behaviour for Sensitive Topics
von: García-Ferrero, Iker, et al.
Veröffentlicht: (2025)
von: García-Ferrero, Iker, et al.
Veröffentlicht: (2025)
AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents
von: Naik, Akshat, et al.
Veröffentlicht: (2025)
von: Naik, Akshat, et al.
Veröffentlicht: (2025)
Dataset Featurization: Uncovering Natural Language Features through Unsupervised Data Reconstruction
von: Bravansky, Michal, et al.
Veröffentlicht: (2025)
von: Bravansky, Michal, et al.
Veröffentlicht: (2025)
NITRO: LLM Inference on Intel Laptop NPUs
von: Fei, Anthony, et al.
Veröffentlicht: (2024)
von: Fei, Anthony, et al.
Veröffentlicht: (2024)
Simulation, Modelling and Classification of Wiki Contributors: Spotting The Good, The Bad, and The Ugly
von: Méndez, Silvia García, et al.
Veröffentlicht: (2024)
von: Méndez, Silvia García, et al.
Veröffentlicht: (2024)
Understanding the Effects of RLHF on LLM Generalisation and Diversity
von: Kirk, Robert, et al.
Veröffentlicht: (2023)
von: Kirk, Robert, et al.
Veröffentlicht: (2023)
LLM Inference Unveiled: Survey and Roofline Model Insights
von: Yuan, Zhihang, et al.
Veröffentlicht: (2024)
von: Yuan, Zhihang, et al.
Veröffentlicht: (2024)
Scaling LLM Inference with Optimized Sample Compute Allocation
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
LLM-based Translation Inference with Iterative Bilingual Understanding
von: Chen, Andong, et al.
Veröffentlicht: (2024)
von: Chen, Andong, et al.
Veröffentlicht: (2024)
Plausibility Vaccine: Injecting LLM Knowledge for Event Plausibility
von: Chmura, Jacob, et al.
Veröffentlicht: (2025)
von: Chmura, Jacob, et al.
Veröffentlicht: (2025)
GroundAct: Can LLM Agents Ground Actions in Environmental States?
von: Wang, Zixuan, et al.
Veröffentlicht: (2025)
von: Wang, Zixuan, et al.
Veröffentlicht: (2025)
CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference
von: Liu, Dong, et al.
Veröffentlicht: (2025)
von: Liu, Dong, et al.
Veröffentlicht: (2025)
Codebook-Injected Dialogue Segmentation for Multi-Utterance Constructs Annotation: LLM-Assisted and Gold-Label-Free Evaluation
von: Lee, Jinsook, et al.
Veröffentlicht: (2026)
von: Lee, Jinsook, et al.
Veröffentlicht: (2026)
EvoP: Robust LLM Inference via Evolutionary Pruning
von: Wu, Shangyu, et al.
Veröffentlicht: (2025)
von: Wu, Shangyu, et al.
Veröffentlicht: (2025)
Reward Model Overoptimisation in Iterated RLHF
von: Wolf, Lorenz, et al.
Veröffentlicht: (2025)
von: Wolf, Lorenz, et al.
Veröffentlicht: (2025)
Reformulating KV Cache Eviction Problem for Long-Context LLM Inference
von: Mai, Tho, et al.
Veröffentlicht: (2026)
von: Mai, Tho, et al.
Veröffentlicht: (2026)
XC-Cache: Cross-Attending to Cached Context for Efficient LLM Inference
von: Monteiro, João, et al.
Veröffentlicht: (2024)
von: Monteiro, João, et al.
Veröffentlicht: (2024)
FLRC: Fine-grained Low-Rank Compressor for Efficient LLM Inference
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference
von: Ouyang, Haojie, et al.
Veröffentlicht: (2025)
von: Ouyang, Haojie, et al.
Veröffentlicht: (2025)
Amphista: Bi-directional Multi-head Decoding for Accelerating LLM Inference
von: Li, Zeping, et al.
Veröffentlicht: (2024)
von: Li, Zeping, et al.
Veröffentlicht: (2024)
AutoManual: Constructing Instruction Manuals by LLM Agents via Interactive Environmental Learning
von: Chen, Minghao, et al.
Veröffentlicht: (2024)
von: Chen, Minghao, et al.
Veröffentlicht: (2024)
Datarus-R1: An Adaptive Multi-Step Reasoning LLM for Automated Data Analysis
von: Chaliah, Ayoub Ben, et al.
Veröffentlicht: (2025)
von: Chaliah, Ayoub Ben, et al.
Veröffentlicht: (2025)
Faithfulness metric fusion: Improving the evaluation of LLM trustworthiness across domains
von: Malin, Ben, et al.
Veröffentlicht: (2025)
von: Malin, Ben, et al.
Veröffentlicht: (2025)
Exploring the Knowledge Mismatch Hypothesis: Hallucination Propensity in Small Models Fine-tuned on Data from Larger Models
von: Wee, Phil, et al.
Veröffentlicht: (2024)
von: Wee, Phil, et al.
Veröffentlicht: (2024)
AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference
von: He, Zhuomin, et al.
Veröffentlicht: (2025)
von: He, Zhuomin, et al.
Veröffentlicht: (2025)
Beyond Behavioural Trade-Offs: Mechanistic Tracing of Pain-Pleasure Decisions in an LLM
von: Bianco, Francesca, et al.
Veröffentlicht: (2026)
von: Bianco, Francesca, et al.
Veröffentlicht: (2026)
Performance Characterization of Expert Router for Scalable LLM Inference
von: Pichlmeier, Josef, et al.
Veröffentlicht: (2024)
von: Pichlmeier, Josef, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Uncovering Deceptive Tendencies in Language Models: A Simulated Company AI Assistant
von: Järviniemi, Olli, et al.
Veröffentlicht: (2024) -
Subversion via Focal Points: Investigating Collusion in LLM Monitoring
von: Järviniemi, Olli
Veröffentlicht: (2025) -
Vocabulary Dropout for Curriculum Diversity in LLM Co-Evolution
von: Dineen, Jacob, et al.
Veröffentlicht: (2026) -
SafeCtrl-RL: Inference-Time Adaptive Behaviour Control for LLM Dialogue via RL-Driven Prompt Optimisation
von: Orme, Michael, et al.
Veröffentlicht: (2026) -
The Knowledge-Behaviour Disconnect in LLM-based Chatbots
von: Broersen, Jan
Veröffentlicht: (2025)