LLMs and the Human Condition
Fuente:
arXiv
Guardado en:
| Autor principal: | Wallis, Peter |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Syntactic Blind Spots: How Misalignment Leads to LLMs Mathematical Errors
por: Williamson, Dane, et al.
Publicado: (2025)
por: Williamson, Dane, et al.
Publicado: (2025)
ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration
por: Deng, Minghang, et al.
Publicado: (2025)
por: Deng, Minghang, et al.
Publicado: (2025)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
por: Du, Bangde, et al.
Publicado: (2025)
por: Du, Bangde, et al.
Publicado: (2025)
ChemPro: A Progressive Chemistry Benchmark for Large Language Models
por: Baranwal, Aaditya, et al.
Publicado: (2026)
por: Baranwal, Aaditya, et al.
Publicado: (2026)
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
por: Venkatasubramanian, Venkat, et al.
Publicado: (2024)
por: Venkatasubramanian, Venkat, et al.
Publicado: (2024)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
por: Khanna, Danush, et al.
Publicado: (2025)
por: Khanna, Danush, et al.
Publicado: (2025)
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
por: Blair-Stanek, Andrew, et al.
Publicado: (2023)
por: Blair-Stanek, Andrew, et al.
Publicado: (2023)
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
por: Teixeira, Tiago, et al.
Publicado: (2026)
por: Teixeira, Tiago, et al.
Publicado: (2026)
Automated Circuit Interpretation via Probe Prompting
por: Birardi, Giuseppe
Publicado: (2025)
por: Birardi, Giuseppe
Publicado: (2025)
CopySpec: Accelerating LLMs with Speculative Copy-and-Paste Without Compromising Quality
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2025)
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2025)
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2024)
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2024)
Aspect-Based Sentiment Analysis for Future Tourism Experiences: A BERT-MoE Framework for Persian User Reviews
por: Taskooh, Hamidreza Kazemi, et al.
Publicado: (2026)
por: Taskooh, Hamidreza Kazemi, et al.
Publicado: (2026)
The Invisible Coalition Partner: How LLMs Vote When Democracy Gets Concrete
por: Barmettler, Joel
Publicado: (2026)
por: Barmettler, Joel
Publicado: (2026)
A Survey of Text and Speech Resources for Hausa and Fongbe: Availability, Quality, and Gaps for NLP Development
por: Adjovi, Mahounan Pericles, et al.
Publicado: (2026)
por: Adjovi, Mahounan Pericles, et al.
Publicado: (2026)
Pareto-Optimized Open-Source LLMs for Healthcare via Context Retrieval
por: Bayarri-Planas, Jordi, et al.
Publicado: (2024)
por: Bayarri-Planas, Jordi, et al.
Publicado: (2024)
LLMs Aren't Human: A Critical Perspective on LLM Personality
por: Zierahn, Kim, et al.
Publicado: (2026)
por: Zierahn, Kim, et al.
Publicado: (2026)
Discovering Differences in Strategic Behavior Between Humans and LLMs
por: Wang, Caroline, et al.
Publicado: (2026)
por: Wang, Caroline, et al.
Publicado: (2026)
Change Is the Only Constant: Dynamic LLM Slicing based on Layer Redundancy
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2024)
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2024)
ALISON: Fast and Effective Stylometric Authorship Obfuscation
por: Xing, Eric, et al.
Publicado: (2024)
por: Xing, Eric, et al.
Publicado: (2024)
ACCORD: Closing the Commonsense Measurability Gap
por: Roewer-Després, François, et al.
Publicado: (2024)
por: Roewer-Després, François, et al.
Publicado: (2024)
Enhancing Transformer RNNs with Multiple Temporal Perspectives
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2024)
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2024)
ConciseRL: Conciseness-Guided Reinforcement Learning for Efficient Reasoning Models
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2025)
por: Dumitru, Razvan-Gabriel, et al.
Publicado: (2025)
Key-Value Means: Transformers with Expandable Block-Recurrent Compressed Memory
por: Goldstein, Daniel, et al.
Publicado: (2026)
por: Goldstein, Daniel, et al.
Publicado: (2026)
Chatbots put to the test in math and logic problems: A preliminary comparison and assessment of ChatGPT-3.5, ChatGPT-4, and Google Bard
por: Plevris, Vagelis, et al.
Publicado: (2023)
por: Plevris, Vagelis, et al.
Publicado: (2023)
Correcting Gradient-Based Circuit Localization via Interaction-Aware Backpropagation
por: Edin, Joakim, et al.
Publicado: (2025)
por: Edin, Joakim, et al.
Publicado: (2025)
Prompt Tuned Embedding Classification for Multi-Label Industry Sector Allocation
por: Buchner, Valentin Leonhard, et al.
Publicado: (2023)
por: Buchner, Valentin Leonhard, et al.
Publicado: (2023)
NRR-Core: Non-Resolution Reasoning as a Computational Framework for Contextual Identity and Ambiguity Preservation
por: Saito, Kei
Publicado: (2025)
por: Saito, Kei
Publicado: (2025)
Benchmarking quantized LLaMa-based models on the Brazilian Secondary School Exam
por: Santos, Matheus L. O., et al.
Publicado: (2023)
por: Santos, Matheus L. O., et al.
Publicado: (2023)
NRR-Phi: Text-to-State Mapping for Ambiguity Preservation in LLM Inference
por: Saito, Kei
Publicado: (2026)
por: Saito, Kei
Publicado: (2026)
Next Token Prediction Is a Dead End for Creativity
por: Olatunji, Ibukun, et al.
Publicado: (2025)
por: Olatunji, Ibukun, et al.
Publicado: (2025)
RWKV-7 "Goose" with Expressive Dynamic State Evolution
por: Peng, Bo, et al.
Publicado: (2025)
por: Peng, Bo, et al.
Publicado: (2025)
BabyReasoningBench: Generating Developmentally-Inspired Reasoning Tasks for Evaluating Baby Language Models
por: Dhole, Kaustubh D.
Publicado: (2026)
por: Dhole, Kaustubh D.
Publicado: (2026)
Reference-Guided Verdict: LLMs-as-Judges in Automatic Evaluation of Free-Form QA
por: Badshah, Sher, et al.
Publicado: (2024)
por: Badshah, Sher, et al.
Publicado: (2024)
Diverse LLMs or Diverse Question Interpretations? That is the Ensembling Question
por: Rosales, Rafael, et al.
Publicado: (2025)
por: Rosales, Rafael, et al.
Publicado: (2025)
Graph Language Models
por: Plenz, Moritz, et al.
Publicado: (2024)
por: Plenz, Moritz, et al.
Publicado: (2024)
Perturbation Dose Responses in Recursive LLM Loops: Raw Switching, Stochastic Floors, and Persistent Escape under Append, Replace, and Dialog Updates
por: Kaplanski, Pawel
Publicado: (2026)
por: Kaplanski, Pawel
Publicado: (2026)
Incentives or Ontology? A Structural Rebuttal to OpenAI's Hallucination Thesis
por: Ackermann, Richard, et al.
Publicado: (2025)
por: Ackermann, Richard, et al.
Publicado: (2025)
Behavioural vs. Representational Systematicity in End-to-End Models: An Opinionated Survey
por: Vegner, Ivan, et al.
Publicado: (2025)
por: Vegner, Ivan, et al.
Publicado: (2025)
The Drill-Down and Fabricate Test (DDFT): A Protocol for Measuring Epistemic Robustness in Language Models
por: Baxi, Rahul
Publicado: (2025)
por: Baxi, Rahul
Publicado: (2025)
Beyond Recall: Behavioral Specification as an Interpretive Layer for AI Personalization
por: Gulaya, Aarik
Publicado: (2026)
por: Gulaya, Aarik
Publicado: (2026)
Ejemplares similares
-
Syntactic Blind Spots: How Misalignment Leads to LLMs Mathematical Errors
por: Williamson, Dane, et al.
Publicado: (2025) -
ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration
por: Deng, Minghang, et al.
Publicado: (2025) -
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
por: Du, Bangde, et al.
Publicado: (2025) -
ChemPro: A Progressive Chemistry Benchmark for Large Language Models
por: Baranwal, Aaditya, et al.
Publicado: (2026) -
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
por: Venkatasubramanian, Venkat, et al.
Publicado: (2024)