Benchmark run results by Shu Wan, on benchmark context PC Hyperparameter Tuning v2
Fuente:
Zenodo
Gespeichert in:
| 1. Verfasser: | Shu Wan |
|---|---|
| Format: | Recurso digital |
| Veröffentlicht: |
Zenodo
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Benchmark run results by Ertugrul Coban, on benchmark context Tuning PC v2
von: Ertugrul Coban
Veröffentlicht: (2025)
von: Ertugrul Coban
Veröffentlicht: (2025)
Benchmark run results by Pratanu Mandal, on benchmark context Tuning PC v3
von: Pratanu Mandal
Veröffentlicht: (2025)
von: Pratanu Mandal
Veröffentlicht: (2025)
Benchmark run results by Ertugrul Coban, on benchmark context Tuning PC v3
von: Ertugrul Coban
Veröffentlicht: (2025)
von: Ertugrul Coban
Veröffentlicht: (2025)
Benchmark run results by Abhinav Gorantla, on benchmark context Tuning PC v3
von: Abhinav Gorantla
Veröffentlicht: (2026)
von: Abhinav Gorantla
Veröffentlicht: (2026)
Benchmark run results by Pratanu Mandal, on benchmark context Tuning PC v3
von: Pratanu Mandal
Veröffentlicht: (2026)
von: Pratanu Mandal
Veröffentlicht: (2026)
Benchmark run results by Abhinav Gorantla, on benchmark context CB-StaticDiscovery v1
von: Abhinav Gorantla
Veröffentlicht: (2025)
von: Abhinav Gorantla
Veröffentlicht: (2025)
Benchmark run results by Abhinav Gorantla, on benchmark context Benchmark: VAR-LiNGAM, PCMCIplus v3
von: Abhinav Gorantla
Veröffentlicht: (2025)
von: Abhinav Gorantla
Veröffentlicht: (2025)
Benchmark run results by Abhinav Gorantla, on benchmark context Tutorial: Static Causal Discovery (Scenario 3) v1
von: Abhinav Gorantla
Veröffentlicht: (2025)
von: Abhinav Gorantla
Veröffentlicht: (2025)
Benchmark run results by Pratanu Mandal, on benchmark context Tutorial: Static Causal Discovery (Scenario 3) v1
von: Pratanu Mandal
Veröffentlicht: (2026)
von: Pratanu Mandal
Veröffentlicht: (2026)
REAL-AI-Benchmark: Real-World Reasoning and Physical-AI Benchmark Suite
von: Ivković, Jovan
Veröffentlicht: (2026)
von: Ivković, Jovan
Veröffentlicht: (2026)
Introduction of an Evaluation Tool to Predict the Probability of Success of Companies: The Innovativeness, Capabilities and Potential Model (ICP)
von: Michael Lewrick
Veröffentlicht: (2009)
von: Michael Lewrick
Veröffentlicht: (2009)
Public Comment on NIST AI 800-2: Anthropomorphic Construct Projection in AI Benchmark Evaluation
von: Sophia, Franny Philos
Veröffentlicht: (2026)
von: Sophia, Franny Philos
Veröffentlicht: (2026)
31. DATASET COMPLETO DE EVALUACIONES CRUZADAS RFC-EVAL-001 – 6 SISTEMAS DE IA (ENERO 2026).
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
TSB: A Time-Saved Benchmark for AI Systems — Measuring Net Productivity Impact Across Knowledge Work
von: Shalom Lijo, Solomon
Veröffentlicht: (2026)
von: Shalom Lijo, Solomon
Veröffentlicht: (2026)
AGI Certification Framework: A Multi-Dimensional Evaluation Standard for Measuring AI Understanding
von: Head, Hank
Veröffentlicht: (2026)
von: Head, Hank
Veröffentlicht: (2026)
Oracle Difficulty Decomposed: Four Independent Mechanisms Explain 95%+ of Benchmark Variance
von: Sanchez, Bryan
Veröffentlicht: (2026)
von: Sanchez, Bryan
Veröffentlicht: (2026)
REFERENCE MODEL FOR EVALUATING THE EFFECTIVENESS OF INCLUSIVE EDUCATION POLICIES
von: Souza Melo, Ítala Raquel, et al.
Veröffentlicht: (2024)
von: Souza Melo, Ítala Raquel, et al.
Veröffentlicht: (2024)
Modular Ebbinghaus Benchmark for LLMs and Human Participants
von: Cohen, Yann
Veröffentlicht: (2026)
von: Cohen, Yann
Veröffentlicht: (2026)
How Far Does the Trolley Problem Go in AI Ethics Evaluation? Limits of a Canonical Benchmark and the Risks of Its Misuse
von: mizutani, aya
Veröffentlicht: (2026)
von: mizutani, aya
Veröffentlicht: (2026)
MedMNIST-NAS-Bench: A Tabular Neural Architecture Search Benchmark on MedMNIST v2
von: Wang, Wei (William)
Veröffentlicht: (2026)
von: Wang, Wei (William)
Veröffentlicht: (2026)
Benchmarking como instrumento dirigido al cliente
von: Emerson Franco de Abreu
Veröffentlicht: (2006)
von: Emerson Franco de Abreu
Veröffentlicht: (2006)
Metacognition Benchmark: Evaluating Confidence Calibration and Sycophancy Resistance in Clinical AI
von: Khan, Nabeera
Veröffentlicht: (2026)
von: Khan, Nabeera
Veröffentlicht: (2026)
30. TEORÍA DE LA POTENCIALIDAD CONSCIENTE (TPC): BENCHMARK DE CAPACIDADES COGNITIVAS EN IA - APLICACIÓN DEL PROTOCOLO RFC-EVAL-001. RESULTADOS COMPLETOS DE EVALUACIÓN CRUZADA CIEGA ENTRE 6 IAS COMERCIALES.
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
30. TEORÍA DE LA POTENCIALIDAD CONSCIENTE (TPC): BENCHMARK DE CAPACIDADES COGNITIVAS EN IA - APLICACIÓN DEL PROTOCOLO RFC-EVAL-001 V1.1. RESULTADOS COMPLETOS DE EVALUACIÓN CRUZADA CIEGA ENTRE 6 IAS COMERCIALES.
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
von: Bernal Díaz, Víctor Cristóbal
Veröffentlicht: (2026)
Dataset for the study of two-wheeler seepage behavior in dense mixed traffic
von: <Hidden>
Veröffentlicht: (2026)
von: <Hidden>
Veröffentlicht: (2026)
afdb_clusters v1.0: AlphaFold-derived structure-based dataset for benchmarking MSA tools
von: Zielezinski, Andrzej, et al.
Veröffentlicht: (2025)
von: Zielezinski, Andrzej, et al.
Veröffentlicht: (2025)
A Proposal of Indicators and Policy Framework for Innovation Benchmark in Europe
von: Juan Vicente García Manjón
Veröffentlicht: (2010)
von: Juan Vicente García Manjón
Veröffentlicht: (2010)
Propuesta metodológica para la aplicación del benchmarking internacional en la evaluación de la calidad de la educación superior virtual
von: Renata Marciniak
Veröffentlicht: (2015)
von: Renata Marciniak
Veröffentlicht: (2015)
NextStat Replication Bundle: replication-rerun-prod-doi-18542624
von: NextStat Contributors
Veröffentlicht: (2026)
von: NextStat Contributors
Veröffentlicht: (2026)
Desde el atractivo poder de los datos de PISA a las desilusiones del benchmarking. ¿Desafío a la evaluación de los sistemas educativos?
von: Marie Duru-Bellat
Veröffentlicht: (2013)
von: Marie Duru-Bellat
Veröffentlicht: (2013)
LLM Token Estimation Benchmarks: Tokenizer Efficiency and Cost Analysis Across 17 Large Language Models
von: Khare, Mohit
Veröffentlicht: (2026)
von: Khare, Mohit
Veröffentlicht: (2026)
Benchmarking for International Competitiveness: Lessons for Public Policy
von: Densil A. Williams
Veröffentlicht: (2010)
von: Densil A. Williams
Veröffentlicht: (2010)
Scaling from 8B to 14B Yields No Meaningful Improvement in Biomimetic Prompt Following: A Paired Comparison Across 3 Model Families and 35 Configurations
von: COYAUD, Denis
Veröffentlicht: (2026)
von: COYAUD, Denis
Veröffentlicht: (2026)
OAEI 2005 benchmark tests (bench22)
von: Euzenat, Jérôme
Veröffentlicht: (2005)
von: Euzenat, Jérôme
Veröffentlicht: (2005)
Internal benchmarking efficiency assessment in a steel company using the DEA network
von: Josiane Vogt
Veröffentlicht: (2025)
von: Josiane Vogt
Veröffentlicht: (2025)
Estimating energy savings and greenhouse gas emission reductions from energy rating and disclosure policies
von: Fátima Uriarte Cáceres
Veröffentlicht: (2018)
von: Fátima Uriarte Cáceres
Veröffentlicht: (2018)
Agenda prospectiva de investigación de la cadena productiva de la panela y su agroindustria
von: DIEGO HERNANDO FLÓREZ MARTÍNEZ
Veröffentlicht: (2013)
von: DIEGO HERNANDO FLÓREZ MARTÍNEZ
Veröffentlicht: (2013)
COMPARISON BETWEEN BRAZIL AND CANADA AS REGARDS COMPETITIVENESS IN CONIFEROUS SAWN TIMBER PRODUCTION
von: Alexandre Nascimento de Almeida
Veröffentlicht: (2013)
von: Alexandre Nascimento de Almeida
Veröffentlicht: (2013)
O subsetor de edificações da construção civil no Brasil: uma análise comparativa em relação à União Europeia e aos Estados Unidos
von: Luiz Carlos Brasil de Brito Mello
Veröffentlicht: (2009)
von: Luiz Carlos Brasil de Brito Mello
Veröffentlicht: (2009)
Demostración de la aplicación del Modelo global de referencia para las tasas de cesárea (C-Model) y la Clasificación de Robson en la estimación y la caracterización del exceso de cesáreas institucionales
von: John Jairo Zuleta-Tobón
Veröffentlicht: (2021)
von: John Jairo Zuleta-Tobón
Veröffentlicht: (2021)
Ähnliche Einträge
-
Benchmark run results by Ertugrul Coban, on benchmark context Tuning PC v2
von: Ertugrul Coban
Veröffentlicht: (2025) -
Benchmark run results by Pratanu Mandal, on benchmark context Tuning PC v3
von: Pratanu Mandal
Veröffentlicht: (2025) -
Benchmark run results by Ertugrul Coban, on benchmark context Tuning PC v3
von: Ertugrul Coban
Veröffentlicht: (2025) -
Benchmark run results by Abhinav Gorantla, on benchmark context Tuning PC v3
von: Abhinav Gorantla
Veröffentlicht: (2026) -
Benchmark run results by Pratanu Mandal, on benchmark context Tuning PC v3
von: Pratanu Mandal
Veröffentlicht: (2026)