Mathematical reasoning and the computer
Fuente:
arXiv
Guardado en:
| Autor principal: | Buzzard, Kevin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Script-Based Dialog Policy Planning for LLM-Powered Conversational Agents: A Basic Architecture for an "AI Therapist"
por: Wasenmüller, Robert, et al.
Publicado: (2024)
por: Wasenmüller, Robert, et al.
Publicado: (2024)
humancompatible.detect: a Python Toolkit for Detecting Bias in AI Models
por: Matilla, German M., et al.
Publicado: (2025)
por: Matilla, German M., et al.
Publicado: (2025)
Toward a Dynamic Stackelberg Game-Theoretic Framework for Agentic AI Defense Against LLM Jailbreaking
por: Han, Zhengye, et al.
Publicado: (2025)
por: Han, Zhengye, et al.
Publicado: (2025)
Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading
por: Sadjoli, Nicholas, et al.
Publicado: (2026)
por: Sadjoli, Nicholas, et al.
Publicado: (2026)
Creativity in the Age of AI: Rethinking the Role of Intentional Agency
por: Pearson, James S., et al.
Publicado: (2026)
por: Pearson, James S., et al.
Publicado: (2026)
Modeling Clinical Concern Trajectories in Language Model Agents
por: Subaharan, Sukesh, et al.
Publicado: (2026)
por: Subaharan, Sukesh, et al.
Publicado: (2026)
How well can a large language model explain business processes as perceived by users?
por: Fahland, Dirk, et al.
Publicado: (2024)
por: Fahland, Dirk, et al.
Publicado: (2024)
BernGraph: Probabilistic Graph Neural Networks for EHR-based Medication Recommendations
por: Piao, Xihao, et al.
Publicado: (2024)
por: Piao, Xihao, et al.
Publicado: (2024)
Benchmarking PNW Model for MedMNIST to 100% Accuracy
por: Deng, Bo
Publicado: (2026)
por: Deng, Bo
Publicado: (2026)
What is the $\textit{intrinsic}$ dimension of your binary data? -- and how to compute it quickly
por: Hanika, Tom, et al.
Publicado: (2024)
por: Hanika, Tom, et al.
Publicado: (2024)
Quantifying Behavioral Dissimilarity Between Mathematical Expressions
por: Mežnar, Sebastian, et al.
Publicado: (2024)
por: Mežnar, Sebastian, et al.
Publicado: (2024)
seqme: a Python library for evaluating biological sequence design
por: Møller-Larsen, Rasmus, et al.
Publicado: (2025)
por: Møller-Larsen, Rasmus, et al.
Publicado: (2025)
The Hidden Costs of AI: A Review of Energy, E-Waste, and Inequality in Model Development
por: Winsta, Jenis
Publicado: (2025)
por: Winsta, Jenis
Publicado: (2025)
Adaptive Orchestration for Large-Scale Inference on Heterogeneous Accelerator Systems Balancing Cost, Performance, and Resilience
por: Biran, Yahav, et al.
Publicado: (2025)
por: Biran, Yahav, et al.
Publicado: (2025)
Tape: A Cellular Automata Benchmark for Evaluating Rule-Shift Generalization in Reinforcement Learning
por: Pan, Enze
Publicado: (2026)
por: Pan, Enze
Publicado: (2026)
RASP-Tuner: Retrieval-Augmented Soft Prompts for Context-Aware Black-Box Optimization in Non-Stationary Environments
por: Pan, Enze
Publicado: (2026)
por: Pan, Enze
Publicado: (2026)
Resilient Federated Chain: Transforming Blockchain Consensus into an Active Defense Layer for Federated Learning
por: García-Márquez, Mario, et al.
Publicado: (2026)
por: García-Márquez, Mario, et al.
Publicado: (2026)
A Theoretical Analysis of Soft-Label vs Hard-Label Training in Neural Networks
por: Mandal, Saptarshi, et al.
Publicado: (2024)
por: Mandal, Saptarshi, et al.
Publicado: (2024)
Leveraging Diversity in Online Interactions
por: Osman, Nardine, et al.
Publicado: (2023)
por: Osman, Nardine, et al.
Publicado: (2023)
ConSensus: Multi-Agent Collaboration for Multimodal Sensing
por: Yoon, Hyungjun, et al.
Publicado: (2026)
por: Yoon, Hyungjun, et al.
Publicado: (2026)
Fast, close, non-singular and property-preserving approximations of entropic measures
por: Horenko, Illia, et al.
Publicado: (2025)
por: Horenko, Illia, et al.
Publicado: (2025)
A computational framework for human values
por: Osman, Nardine, et al.
Publicado: (2023)
por: Osman, Nardine, et al.
Publicado: (2023)
Feature Relevancy, Necessity and Usefulness: Complexity and Algorithms
por: Capdevielle, Tomás, et al.
Publicado: (2025)
por: Capdevielle, Tomás, et al.
Publicado: (2025)
Understanding Knowledge Transferability for Transfer Learning: A Survey
por: Wang, Haohua, et al.
Publicado: (2025)
por: Wang, Haohua, et al.
Publicado: (2025)
A Taxonomy of Omnicidal Futures Involving Artificial Intelligence
por: Critch, Andrew, et al.
Publicado: (2025)
por: Critch, Andrew, et al.
Publicado: (2025)
Beyond the Org Chart: AI and the Transformation of Invisible Work
por: Rosenthal, Stephanie, et al.
Publicado: (2026)
por: Rosenthal, Stephanie, et al.
Publicado: (2026)
A Theoretical Framework for Adaptive Utility-Weighted Benchmarking
por: Waggoner, Philip
Publicado: (2026)
por: Waggoner, Philip
Publicado: (2026)
From Language Models to Practical Self-Improving Computer Agents
por: Sheng, Alex
Publicado: (2024)
por: Sheng, Alex
Publicado: (2024)
From Euler to Today: Universal Mathematical Fallibility A Large-Scale Computational Analysis of Errors in ArXiv Papers
por: Rivin, Igor
Publicado: (2025)
por: Rivin, Igor
Publicado: (2025)
Abductive explanations of classifiers under constraints: Complexity and properties
por: Cooper, Martin, et al.
Publicado: (2024)
por: Cooper, Martin, et al.
Publicado: (2024)
Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring
por: Aksoy, Sinan G., et al.
Publicado: (2026)
por: Aksoy, Sinan G., et al.
Publicado: (2026)
Can Out-of-Distribution Evaluations Uncover Reliance on Shortcuts? A Case Study in Question Answering
por: Štefánik, Michal, et al.
Publicado: (2025)
por: Štefánik, Michal, et al.
Publicado: (2025)
Acceptable Use Policies for Foundation Models
por: Klyman, Kevin
Publicado: (2024)
por: Klyman, Kevin
Publicado: (2024)
Authenticated Delegation and Authorized AI Agents
por: South, Tobin, et al.
Publicado: (2025)
por: South, Tobin, et al.
Publicado: (2025)
March Madness Tournament Predictions Model: A Mathematical Modeling Approach
por: McIver, Christian, et al.
Publicado: (2025)
por: McIver, Christian, et al.
Publicado: (2025)
Neural Concept Verifier: Scaling Prover-Verifier Games via Concept Encodings
por: Turan, Berkant, et al.
Publicado: (2025)
por: Turan, Berkant, et al.
Publicado: (2025)
Challenges and Future Directions in Agentic Reverse Engineering Systems
por: Radey, Salem, et al.
Publicado: (2026)
por: Radey, Salem, et al.
Publicado: (2026)
AI Governance InternationaL Evaluation Index (AGILE Index) 2024
por: Zeng, Yi, et al.
Publicado: (2025)
por: Zeng, Yi, et al.
Publicado: (2025)
Top-Theta Attention: Sparsifying Transformers by Compensated Thresholding
por: Berestizshevsky, Konstantin, et al.
Publicado: (2025)
por: Berestizshevsky, Konstantin, et al.
Publicado: (2025)
Reducing Instability in Synthetic Data Evaluation with a Super-Metric in MalDataGen
por: da Silva, Anna Luiza Gomes, et al.
Publicado: (2025)
por: da Silva, Anna Luiza Gomes, et al.
Publicado: (2025)
Ejemplares similares
-
Script-Based Dialog Policy Planning for LLM-Powered Conversational Agents: A Basic Architecture for an "AI Therapist"
por: Wasenmüller, Robert, et al.
Publicado: (2024) -
humancompatible.detect: a Python Toolkit for Detecting Bias in AI Models
por: Matilla, German M., et al.
Publicado: (2025) -
Toward a Dynamic Stackelberg Game-Theoretic Framework for Agentic AI Defense Against LLM Jailbreaking
por: Han, Zhengye, et al.
Publicado: (2025) -
Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading
por: Sadjoli, Nicholas, et al.
Publicado: (2026) -
Creativity in the Age of AI: Rethinking the Role of Intentional Agency
por: Pearson, James S., et al.
Publicado: (2026)