From Simulation to Enaction: Post-trained language models recognize and react to their own generations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | G., Asvin, Lindsey, Jack |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Post-training makes large language models less human-like
von: Binz, Marcel, et al.
Veröffentlicht: (2026)
von: Binz, Marcel, et al.
Veröffentlicht: (2026)
Stylometry recognizes human and LLM-generated texts in short samples
von: Przystalski, Karol, et al.
Veröffentlicht: (2025)
von: Przystalski, Karol, et al.
Veröffentlicht: (2025)
Lightweight reranking for language model generations
von: Jain, Siddhartha, et al.
Veröffentlicht: (2023)
von: Jain, Siddhartha, et al.
Veröffentlicht: (2023)
Diffusion on language model encodings for protein sequence generation
von: Meshchaninov, Viacheslav, et al.
Veröffentlicht: (2024)
von: Meshchaninov, Viacheslav, et al.
Veröffentlicht: (2024)
The Impact of Post-training on Data Contamination
von: Kocyigit, Muhammed Yusuf, et al.
Veröffentlicht: (2026)
von: Kocyigit, Muhammed Yusuf, et al.
Veröffentlicht: (2026)
Aleph-Alpha-GermanWeb: Improving German-language LLM pre-training with model-based data curation and synthetic data generation
von: Burns, Thomas F, et al.
Veröffentlicht: (2025)
von: Burns, Thomas F, et al.
Veröffentlicht: (2025)
Agentic reinforcement learning empowers next-generation chemical language models for molecular design and synthesis
von: Li, Hao, et al.
Veröffentlicht: (2026)
von: Li, Hao, et al.
Veröffentlicht: (2026)
Data filtering methods for training language models
von: Shevchenko, Egor, et al.
Veröffentlicht: (2026)
von: Shevchenko, Egor, et al.
Veröffentlicht: (2026)
Pre-trained knowledge elevates large language models beyond traditional chemical reaction optimizers
von: MacKnight, Robert, et al.
Veröffentlicht: (2025)
von: MacKnight, Robert, et al.
Veröffentlicht: (2025)
Prior-informed optimization of treatment recommendation via bandit algorithms trained on large language model-processed historical records
von: Nessari, Saman, et al.
Veröffentlicht: (2025)
von: Nessari, Saman, et al.
Veröffentlicht: (2025)
Towards the generation of hierarchical attack models from cybersecurity vulnerabilities using language models
von: Sowka, Kacper, et al.
Veröffentlicht: (2024)
von: Sowka, Kacper, et al.
Veröffentlicht: (2024)
Improving training time and GPU utilization in geo-distributed language model training
von: Palak, et al.
Veröffentlicht: (2024)
von: Palak, et al.
Veröffentlicht: (2024)
Aviary: training language agents on challenging scientific tasks
von: Narayanan, Siddharth, et al.
Veröffentlicht: (2024)
von: Narayanan, Siddharth, et al.
Veröffentlicht: (2024)
Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena
von: Luo, Haipeng, et al.
Veröffentlicht: (2024)
von: Luo, Haipeng, et al.
Veröffentlicht: (2024)
AI-AI Bias: large language models favor communications generated by large language models
von: Laurito, Walter, et al.
Veröffentlicht: (2024)
von: Laurito, Walter, et al.
Veröffentlicht: (2024)
Auditing language models for hidden objectives
von: Marks, Samuel, et al.
Veröffentlicht: (2025)
von: Marks, Samuel, et al.
Veröffentlicht: (2025)
AutoOR: Scalably Post-training LLMs to Autoformalize Operations Research Problems
von: Motwani, Sumeet Ramesh, et al.
Veröffentlicht: (2026)
von: Motwani, Sumeet Ramesh, et al.
Veröffentlicht: (2026)
BoA: Attention-aware Post-training Quantization without Backpropagation
von: Kim, Junhan, et al.
Veröffentlicht: (2024)
von: Kim, Junhan, et al.
Veröffentlicht: (2024)
It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amplified by the language of the prompt
von: Bladon, Stuart, et al.
Veröffentlicht: (2026)
von: Bladon, Stuart, et al.
Veröffentlicht: (2026)
Robust training of implicit generative models for multivariate and heavy-tailed distributions with an invariant statistical loss
von: de Frutos, José Manuel, et al.
Veröffentlicht: (2024)
von: de Frutos, José Manuel, et al.
Veröffentlicht: (2024)
Prioritized Replay for RL Post-training
von: Fatemi, Mehdi
Veröffentlicht: (2026)
von: Fatemi, Mehdi
Veröffentlicht: (2026)
The promising potential of vision language models for the generation of textual weather forecasts
von: Steele, Edward C. C., et al.
Veröffentlicht: (2025)
von: Steele, Edward C. C., et al.
Veröffentlicht: (2025)
Beneficial Reasoning Behaviors in Agentic Search and Effective Post-training to Obtain Them
von: Jin, Jiahe, et al.
Veröffentlicht: (2025)
von: Jin, Jiahe, et al.
Veröffentlicht: (2025)
AdamS: Momentum Itself Can Be A Normalizer for LLM Pretraining and Post-training
von: Zhang, Huishuai, et al.
Veröffentlicht: (2025)
von: Zhang, Huishuai, et al.
Veröffentlicht: (2025)
Continual Quantization-Aware Pre-Training: When to transition from 16-bit to 1.58-bit pre-training for BitNet language models?
von: Nielsen, Jacob, et al.
Veröffentlicht: (2025)
von: Nielsen, Jacob, et al.
Veröffentlicht: (2025)
Effective internal language model training and fusion for factorized transducer model
von: Guo, Jinxi, et al.
Veröffentlicht: (2024)
von: Guo, Jinxi, et al.
Veröffentlicht: (2024)
What happens when generative AI models train recursively on each others' outputs?
von: Vu, Hung Anh, et al.
Veröffentlicht: (2025)
von: Vu, Hung Anh, et al.
Veröffentlicht: (2025)
How predictable is language model benchmark performance?
von: Owen, David
Veröffentlicht: (2024)
von: Owen, David
Veröffentlicht: (2024)
Alignment faking in large language models
von: Greenblatt, Ryan, et al.
Veröffentlicht: (2024)
von: Greenblatt, Ryan, et al.
Veröffentlicht: (2024)
Automatic Pair Construction for Contrastive Post-training
von: Xu, Canwen, et al.
Veröffentlicht: (2023)
von: Xu, Canwen, et al.
Veröffentlicht: (2023)
Early-stopping for Transformer model training
von: He, Jing, et al.
Veröffentlicht: (2025)
von: He, Jing, et al.
Veröffentlicht: (2025)
Efficiency optimization of large-scale language models based on deep learning in natural language processing tasks
von: Mei, Taiyuan, et al.
Veröffentlicht: (2024)
von: Mei, Taiyuan, et al.
Veröffentlicht: (2024)
Quantifying construct validity in large language model evaluations
von: Kearns, Ryan Othniel
Veröffentlicht: (2026)
von: Kearns, Ryan Othniel
Veröffentlicht: (2026)
Applying sparse autoencoders to unlearn knowledge in language models
von: Farrell, Eoin, et al.
Veröffentlicht: (2024)
von: Farrell, Eoin, et al.
Veröffentlicht: (2024)
FoldToken2: Learning compact, invariant and generative protein structure language
von: Gao, Zhangyang, et al.
Veröffentlicht: (2024)
von: Gao, Zhangyang, et al.
Veröffentlicht: (2024)
PBP: Post-training Backdoor Purification for Malware Classifiers
von: Nguyen, Dung Thuy, et al.
Veröffentlicht: (2024)
von: Nguyen, Dung Thuy, et al.
Veröffentlicht: (2024)
Post-training for Efficient Communication via Convention Formation
von: Hua, Yilun, et al.
Veröffentlicht: (2025)
von: Hua, Yilun, et al.
Veröffentlicht: (2025)
Efficient Post-training Quantization with FP8 Formats
von: Shen, Haihao, et al.
Veröffentlicht: (2023)
von: Shen, Haihao, et al.
Veröffentlicht: (2023)
Are we still able to recognize pearls? Machine-driven peer review and the risk to creativity: An explainable RAG-XAI detection framework with markers extraction
von: Văduva, Alin-Gabriel, et al.
Veröffentlicht: (2026)
von: Văduva, Alin-Gabriel, et al.
Veröffentlicht: (2026)
Why Pass@k Optimization Can Degrade Pass@1: Prompt Interference in LLM Post-training
von: Barakat, Anas, et al.
Veröffentlicht: (2026)
von: Barakat, Anas, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Post-training makes large language models less human-like
von: Binz, Marcel, et al.
Veröffentlicht: (2026) -
Stylometry recognizes human and LLM-generated texts in short samples
von: Przystalski, Karol, et al.
Veröffentlicht: (2025) -
Lightweight reranking for language model generations
von: Jain, Siddhartha, et al.
Veröffentlicht: (2023) -
Diffusion on language model encodings for protein sequence generation
von: Meshchaninov, Viacheslav, et al.
Veröffentlicht: (2024) -
The Impact of Post-training on Data Contamination
von: Kocyigit, Muhammed Yusuf, et al.
Veröffentlicht: (2026)