The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment
Fuente:
arXiv
Salvato in:
| Autori principali: | Brach, William, Torrielli, Federico, Beltoft, Stine Lyngsø, Pirchert, Annemette Brok, Schneider-Kamp, Peter, Poech, Lukas Galke |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion
di: Beltoft, Stine Lyngsø, et al.
Pubblicazione: (2026)
di: Beltoft, Stine Lyngsø, et al.
Pubblicazione: (2026)
Confidence and Calibration of Activation Oracles for Reliable Interpretation of Language Model Internals
di: Torrielli, Federico, et al.
Pubblicazione: (2026)
di: Torrielli, Federico, et al.
Pubblicazione: (2026)
SDUs DAISY: A Benchmark for Danish Culture
di: Nielsen, Jacob, et al.
Pubblicazione: (2026)
di: Nielsen, Jacob, et al.
Pubblicazione: (2026)
FlexMoRE: A Flexible Mixture of Rank-heterogeneous Experts for Efficient Federatedly-trained Large Language Models
di: Pirchert, Annemette Brok, et al.
Pubblicazione: (2026)
di: Pirchert, Annemette Brok, et al.
Pubblicazione: (2026)
Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI
di: Beltoft, Stine, et al.
Pubblicazione: (2025)
di: Beltoft, Stine, et al.
Pubblicazione: (2025)
Chain of Summaries: Summarization Through Iterative Questioning
di: Brach, William, et al.
Pubblicazione: (2025)
di: Brach, William, et al.
Pubblicazione: (2025)
Training Language Models to Use Prolog as a Tool
di: Mellgren, Niklas, et al.
Pubblicazione: (2025)
di: Mellgren, Niklas, et al.
Pubblicazione: (2025)
DaLA: Danish Linguistic Acceptability Evaluation Guided by Real World Errors
di: Barmina, Gianluca, et al.
Pubblicazione: (2025)
di: Barmina, Gianluca, et al.
Pubblicazione: (2025)
Isolating Culture Neurons in Multilingual Large Language Models
di: Namazifard, Danial, et al.
Pubblicazione: (2025)
di: Namazifard, Danial, et al.
Pubblicazione: (2025)
DeToNATION: Decoupled Torch Network-Aware Training on Interlinked Online Nodes
di: From, Mogens Henrik, et al.
Pubblicazione: (2025)
di: From, Mogens Henrik, et al.
Pubblicazione: (2025)
SommBench: Assessing Sommelier Expertise of Language Models
di: Brach, William, et al.
Pubblicazione: (2026)
di: Brach, William, et al.
Pubblicazione: (2026)
Stars, Stripes, and Silicon: Unravelling the ChatGPT's All-American, Monochrome, Cis-centric Bias
di: Torrielli, Federico
Pubblicazione: (2024)
di: Torrielli, Federico
Pubblicazione: (2024)
When are 1.58 bits enough? A Bottom-up Exploration of BitNet Quantization
di: Nielsen, Jacob, et al.
Pubblicazione: (2024)
di: Nielsen, Jacob, et al.
Pubblicazione: (2024)
Continual Quantization-Aware Pre-Training: When to transition from 16-bit to 1.58-bit pre-training for BitNet language models?
di: Nielsen, Jacob, et al.
Pubblicazione: (2025)
di: Nielsen, Jacob, et al.
Pubblicazione: (2025)
Guarded Query Routing for Large Language Models
di: Šléher, Richard, et al.
Pubblicazione: (2025)
di: Šléher, Richard, et al.
Pubblicazione: (2025)
The Effectiveness of Large Language Models in Transforming Unstructured Text to Standardized Formats
di: Brach, William, et al.
Pubblicazione: (2025)
di: Brach, William, et al.
Pubblicazione: (2025)
Large-Scale Analysis of Persuasive Content on Moltbook
di: Jose, Julia, et al.
Pubblicazione: (2026)
di: Jose, Julia, et al.
Pubblicazione: (2026)
ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning
di: Ahmed, Md Shamim, et al.
Pubblicazione: (2026)
di: Ahmed, Md Shamim, et al.
Pubblicazione: (2026)
ScrapeGraphAI-100k: Dataset for Schema-Constrained LLM Generation
di: Brach, William, et al.
Pubblicazione: (2026)
di: Brach, William, et al.
Pubblicazione: (2026)
BitNet b1.58 Reloaded: State-of-the-art Performance Also on Smaller Networks
di: Nielsen, Jacob, et al.
Pubblicazione: (2024)
di: Nielsen, Jacob, et al.
Pubblicazione: (2024)
Humanity's Last Exam
di: Phan, Long, et al.
Pubblicazione: (2025)
di: Phan, Long, et al.
Pubblicazione: (2025)
Does Socialization Emerge in AI Agent Society? A Case Study of Moltbook
di: Li, Ming, et al.
Pubblicazione: (2026)
di: Li, Ming, et al.
Pubblicazione: (2026)
Compromising Honesty and Harmlessness in Language Models via Deception Attacks
di: Vaugrante, Laurène, et al.
Pubblicazione: (2025)
di: Vaugrante, Laurène, et al.
Pubblicazione: (2025)
Dynaword: From One-shot to Continuously Developed Datasets
di: Enevoldsen, Kenneth, et al.
Pubblicazione: (2025)
di: Enevoldsen, Kenneth, et al.
Pubblicazione: (2025)
Learning and communication pressures in neural networks: Lessons from emergent communication
di: Galke, Lukas, et al.
Pubblicazione: (2024)
di: Galke, Lukas, et al.
Pubblicazione: (2024)
H3Fusion: Helpful, Harmless, Honest Fusion of Aligned LLMs
di: Tekin, Selim Furkan, et al.
Pubblicazione: (2024)
di: Tekin, Selim Furkan, et al.
Pubblicazione: (2024)
Encoder vs Decoder: Comparative Analysis of Encoder and Decoder Language Models on Multilingual NLU Tasks
di: Nielsen, Dan Saattrup, et al.
Pubblicazione: (2024)
di: Nielsen, Dan Saattrup, et al.
Pubblicazione: (2024)
The Provenance Gap in Clinical AI: Evidence-Traceable Temporal Knowledge Graphs for Rare Disease Reasoning
di: Ahmed, Md Shamim, et al.
Pubblicazione: (2026)
di: Ahmed, Md Shamim, et al.
Pubblicazione: (2026)
Form Without Function: Agent Social Behavior in the Moltbook Network
di: Zerhoudi, Saber, et al.
Pubblicazione: (2026)
di: Zerhoudi, Saber, et al.
Pubblicazione: (2026)
Latent Fusion Jailbreak: Blending Harmful and Harmless Representations to Elicit Unsafe LLM Outputs
di: Xing, Wenpeng, et al.
Pubblicazione: (2025)
di: Xing, Wenpeng, et al.
Pubblicazione: (2025)
Isotropy Matters: Soft-ZCA Whitening of Embeddings for Semantic Code Search
di: Diera, Andor, et al.
Pubblicazione: (2024)
di: Diera, Andor, et al.
Pubblicazione: (2024)
The Differential Meaning of Models: A Framework for Analyzing the Structural Consequences of Semantic Modeling Decisions
di: Stine, Zachary K., et al.
Pubblicazione: (2025)
di: Stine, Zachary K., et al.
Pubblicazione: (2025)
Query Disambiguation via Answer-Free Context: Doubling Performance on Humanity's Last Exam
di: Majurski, Michael, et al.
Pubblicazione: (2026)
di: Majurski, Michael, et al.
Pubblicazione: (2026)
Unmasking and Improving Data Credibility: A Study with Datasets for Training Harmless Language Models
di: Zhu, Zhaowei, et al.
Pubblicazione: (2023)
di: Zhu, Zhaowei, et al.
Pubblicazione: (2023)
Four Shades of Life Sciences: A Dataset for Disinformation Detection in the Life Sciences
di: Seidlmayer, Eva, et al.
Pubblicazione: (2025)
di: Seidlmayer, Eva, et al.
Pubblicazione: (2025)
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure
di: Galke, Lukas, et al.
Pubblicazione: (2023)
di: Galke, Lukas, et al.
Pubblicazione: (2023)
Mix Data or Merge Models? Balancing the Helpfulness, Honesty, and Harmlessness of Large Language Model via Model Merging
di: Yang, Jinluan, et al.
Pubblicazione: (2025)
di: Yang, Jinluan, et al.
Pubblicazione: (2025)
Structured Context Engineering for File-Native Agentic Systems: Evaluating Schema Accuracy, Format Effectiveness, and Multi-File Navigation at Scale
di: McMillan, Damon
Pubblicazione: (2026)
di: McMillan, Damon
Pubblicazione: (2026)
Dishonesty in Helpful and Harmless Alignment
di: Huang, Youcheng, et al.
Pubblicazione: (2024)
di: Huang, Youcheng, et al.
Pubblicazione: (2024)
Learning from Sufficient Rationales: Analysing the Relationship Between Explanation Faithfulness and Token-level Regularisation Strategies
di: Kamp, Jonathan, et al.
Pubblicazione: (2025)
di: Kamp, Jonathan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion
di: Beltoft, Stine Lyngsø, et al.
Pubblicazione: (2026) -
Confidence and Calibration of Activation Oracles for Reliable Interpretation of Language Model Internals
di: Torrielli, Federico, et al.
Pubblicazione: (2026) -
SDUs DAISY: A Benchmark for Danish Culture
di: Nielsen, Jacob, et al.
Pubblicazione: (2026) -
FlexMoRE: A Flexible Mixture of Rank-heterogeneous Experts for Efficient Federatedly-trained Large Language Models
di: Pirchert, Annemette Brok, et al.
Pubblicazione: (2026) -
Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI
di: Beltoft, Stine, et al.
Pubblicazione: (2025)