Scaling from 8B to 14B Yields No Meaningful Improvement in Biomimetic Prompt Following: A Paired Comparison Across 3 Model Families and 35 Configurations
Fuente:
Zenodo
Saved in:
| Main Author: | COYAUD, Denis |
|---|---|
| Format: | Recurso digital |
| Language: | English |
| Published: |
Zenodo
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompt Engineering: a methodology for optimizing interactions with AI-Language Models in the field of engineering
by: Juan David Velásquez-Henao
Published: (2023)
by: Juan David Velásquez-Henao
Published: (2023)
The Invisible Chaperone: The Secret World of System Prompts
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Ep. 598: Audio Engineering as Prompt Engineering: Better Sound, Better AI
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Ep. 1086: Why AI Can't Stop Talking About Second Order Effects
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
The Transformer Trinity: Why Three Architectures Rule AI
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Ep. 1111: The Architecture of Intelligence: Beyond the Transformer
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Ep. 651: Decoding the Blueprint: An Expert Guide to AI Model Cards
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Ep. 111: Beyond Transformers: Solving the AI Memory Crisis
by: Rosehill, Daniel, et al.
Published: (2025)
by: Rosehill, Daniel, et al.
Published: (2025)
Ep. 1080: Beyond the Prompt: Mapping the Future of Claude Opus
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Ep. 103: The Future of Coding: Is Your Brain Wired for AI?
by: Rosehill, Daniel, et al.
Published: (2025)
by: Rosehill, Daniel, et al.
Published: (2025)
The More You Tell It, The Less It Sees: Anchoring Bias in Vision-Language Models
by: Dubey, Mradul
Published: (2026)
by: Dubey, Mradul
Published: (2026)
Ep. 170: The Heavy Metal of Machine Learning: Inside PyTorch
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Why AI Can't Simulate Extreme Decision-Making
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
LLM Token Estimation Benchmarks: Tokenizer Efficiency and Cost Analysis Across 17 Large Language Models
by: Khare, Mohit
Published: (2026)
by: Khare, Mohit
Published: (2026)
Ep. 23: AI's Blind Spot: Data, Bias & Common Crawl
by: Rosehill, Daniel, et al.
Published: (2025)
by: Rosehill, Daniel, et al.
Published: (2025)
Ep. 713: The AI Cyber Frontier: Israel as a Global Testing Ground
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Ep. 476: Beyond the Plateau: AI-Powered Language Mastery in 2026
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Perturbing LLM Attractors, Intentionally: The Thermodynamics of Human-AI Interaction
by: Pourdavood, Parham
Published: (2025)
by: Pourdavood, Parham
Published: (2025)
When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings
by: Kolb, Christian
Published: (2026)
by: Kolb, Christian
Published: (2026)
Local Large Language Models in R with Ollama
by: Schweinberger, Martin
Published: (2026)
by: Schweinberger, Martin
Published: (2026)
Ep. 1108: Beyond the Emoji: How Hugging Face Conquered AI
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Reliability Inference Drives Cue Extraction in Large Language Models Consuming External Reasoning Traces
by: HIDEKI
Published: (2026)
by: HIDEKI
Published: (2026)
Benchmark run results by Abhinav Gorantla, on benchmark context Tuning PC v3
by: Abhinav Gorantla
Published: (2026)
by: Abhinav Gorantla
Published: (2026)
Benchmark run results by Ertugrul Coban, on benchmark context Tuning PC v2
by: Ertugrul Coban
Published: (2025)
by: Ertugrul Coban
Published: (2025)
Benchmark run results by Abhinav Gorantla, on benchmark context CB-StaticDiscovery v1
by: Abhinav Gorantla
Published: (2025)
by: Abhinav Gorantla
Published: (2025)
Benchmark run results by Abhinav Gorantla, on benchmark context Benchmark: VAR-LiNGAM, PCMCIplus v3
by: Abhinav Gorantla
Published: (2025)
by: Abhinav Gorantla
Published: (2025)
Benchmark run results by Pratanu Mandal, on benchmark context Tutorial: Static Causal Discovery (Scenario 3) v1
by: Pratanu Mandal
Published: (2026)
by: Pratanu Mandal
Published: (2026)
Benchmark run results by Pratanu Mandal, on benchmark context Tuning PC v3
by: Pratanu Mandal
Published: (2025)
by: Pratanu Mandal
Published: (2025)
Benchmark run results by Pratanu Mandal, on benchmark context Tuning PC v3
by: Pratanu Mandal
Published: (2026)
by: Pratanu Mandal
Published: (2026)
Benchmark run results by Ertugrul Coban, on benchmark context Tuning PC v3
by: Ertugrul Coban
Published: (2025)
by: Ertugrul Coban
Published: (2025)
Benchmark run results by Abhinav Gorantla, on benchmark context Tutorial: Static Causal Discovery (Scenario 3) v1
by: Abhinav Gorantla
Published: (2025)
by: Abhinav Gorantla
Published: (2025)
Benchmark run results by Shu Wan, on benchmark context PC Hyperparameter Tuning v2
by: Shu Wan
Published: (2025)
by: Shu Wan
Published: (2025)
Beyond Buttons: Is the Admin Dashboard Dead?
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
Theatrical Compliance: A Failure Mode in Large Language Models
by: Nowickij (Navitski), Kirill Vladimirovich
Published: (2026)
by: Nowickij (Navitski), Kirill Vladimirovich
Published: (2026)
Ep. 869: Why Tiny Digital Savants Are Outperforming God-Models
by: Rosehill, Daniel, et al.
Published: (2026)
by: Rosehill, Daniel, et al.
Published: (2026)
The Absurdist's Guide to AI Probing: How I Learned to Stop Worrying and Love the Nonsense
by: Walton, Mathew
Published: (2026)
by: Walton, Mathew
Published: (2026)
REAL-AI-Benchmark: Real-World Reasoning and Physical-AI Benchmark Suite
by: Ivković, Jovan
Published: (2026)
by: Ivković, Jovan
Published: (2026)
Design and Construction of a Snake-Like Robot Implementing Rectilinear and Sidewinding Gait Motions
by: Jairo José Marín Arciniegas
Published: (2023)
by: Jairo José Marín Arciniegas
Published: (2023)
Fortifying NLP models - dataset + code
by: Ferdinan, Teddy, et al.
Published: (2025)
by: Ferdinan, Teddy, et al.
Published: (2025)
31. DATASET COMPLETO DE EVALUACIONES CRUZADAS RFC-EVAL-001 – 6 SISTEMAS DE IA (ENERO 2026).
by: Bernal Díaz, Víctor Cristóbal
Published: (2026)
by: Bernal Díaz, Víctor Cristóbal
Published: (2026)
Similar Items
-
Prompt Engineering: a methodology for optimizing interactions with AI-Language Models in the field of engineering
by: Juan David Velásquez-Henao
Published: (2023) -
The Invisible Chaperone: The Secret World of System Prompts
by: Rosehill, Daniel, et al.
Published: (2026) -
Ep. 598: Audio Engineering as Prompt Engineering: Better Sound, Better AI
by: Rosehill, Daniel, et al.
Published: (2026) -
Ep. 1086: Why AI Can't Stop Talking About Second Order Effects
by: Rosehill, Daniel, et al.
Published: (2026) -
The Transformer Trinity: Why Three Architectures Rule AI
by: Rosehill, Daniel, et al.
Published: (2026)