AILuminate: Introducing v1.0 of the AI Risk and Reliability Benchmark from MLCommons
Fuente:
arXiv
Salvato in:
Documenti analoghi
MoPEFT: A Mixture-of-PEFTs for the Segment Anything Model
di: Sahay, Rajat, et al.
Pubblicazione: (2024)
di: Sahay, Rajat, et al.
Pubblicazione: (2024)
IMAS: A Comprehensive Agentic Approach to Rural Healthcare Delivery
di: Gangavarapu, Agasthya, et al.
Pubblicazione: (2024)
di: Gangavarapu, Agasthya, et al.
Pubblicazione: (2024)
LLM Robustness Leaderboard v1 --Technical report
di: Lefebvre, Pierre Peigné -, et al.
Pubblicazione: (2025)
di: Lefebvre, Pierre Peigné -, et al.
Pubblicazione: (2025)
Introducing L2M3, A Multilingual Medical Large Language Model to Advance Health Equity in Low-Resource Regions
di: Gangavarapu, Agasthya
Pubblicazione: (2024)
di: Gangavarapu, Agasthya
Pubblicazione: (2024)
Enhancing Guardrails for Safe and Secure Healthcare AI
di: Gangavarapu, Ananya
Pubblicazione: (2024)
di: Gangavarapu, Ananya
Pubblicazione: (2024)
Action Shapley: A Training Data Selection Metric for World Model in Reinforcement Learning
di: Ghosh, Rajat, et al.
Pubblicazione: (2026)
di: Ghosh, Rajat, et al.
Pubblicazione: (2026)
Context Lineage Assurance for Non-Human Identities in Critical Multi-Agent Systems
di: Malkapuram, Sumana, et al.
Pubblicazione: (2025)
di: Malkapuram, Sumana, et al.
Pubblicazione: (2025)
Quantitative Systems Pharmacology Modeling Amid the Rise of Agentic AI
di: James Lu, et al.
Pubblicazione: (2026)
di: James Lu, et al.
Pubblicazione: (2026)
Efficient Alignment of Large Language Models via Data Sampling
di: Khera, Amrit, et al.
Pubblicazione: (2024)
di: Khera, Amrit, et al.
Pubblicazione: (2024)
CPP-UT-Bench: Can LLMs Write Complex Unit Tests in C++?
di: Bhargava, Vaishnavi, et al.
Pubblicazione: (2024)
di: Bhargava, Vaishnavi, et al.
Pubblicazione: (2024)
Chromium Neutralization Report
di: Dwivedi, Rajat
Pubblicazione: (2026)
di: Dwivedi, Rajat
Pubblicazione: (2026)
DisasterQA: A Benchmark for Assessing the performance of LLMs in Disaster Response
di: Rawat, Rajat
Pubblicazione: (2024)
di: Rawat, Rajat
Pubblicazione: (2024)
An Upper Bound on the Linear Turán Number of $k$-Crowns
di: Adak, Rajat
Pubblicazione: (2026)
di: Adak, Rajat
Pubblicazione: (2026)
Agentic AI-Driven Technical Troubleshooting for Enterprise Systems: A Novel Weighted Retrieval-Augmented Generation Paradigm
di: Khanda, Rajat
Pubblicazione: (2024)
di: Khanda, Rajat
Pubblicazione: (2024)
FLOW PHYSICS OF 3-BLADED STRAIGHT CHORD H- DARRIEUS WIND TURBINE
di: Rajat Gupta
Pubblicazione: (2013)
di: Rajat Gupta
Pubblicazione: (2013)
CMET: Clustering guided METric for quantifying embedding quality
di: Ghosh, Sourav, et al.
Pubblicazione: (2025)
di: Ghosh, Sourav, et al.
Pubblicazione: (2025)
Forward-Cooperation-Backward (FCB) learning in a Multi-Encoding Uni-Decoding neural network architecture
di: Dutta, Prasun, et al.
Pubblicazione: (2025)
di: Dutta, Prasun, et al.
Pubblicazione: (2025)
Synthesis of Two Stereoisomers of a Natural Cyclotetrapeptide and Four Stereoisomeric Diheteropeptins Through Late‐Stage Scaffold Diversification
di: Rajat Ghosh, et al.
Pubblicazione: (2025)
di: Rajat Ghosh, et al.
Pubblicazione: (2025)
RANGER -- Repository-Level Agent for Graph-Enhanced Retrieval
di: Shah, Pratik, et al.
Pubblicazione: (2025)
di: Shah, Pratik, et al.
Pubblicazione: (2025)
CR-Bench: Evaluating the Real-World Utility of AI Code Review Agents
di: Pereira, Kristen, et al.
Pubblicazione: (2026)
di: Pereira, Kristen, et al.
Pubblicazione: (2026)
A Multi-Agent Framework for Stateful Inference-Time Search
di: Lalan, Arshika, et al.
Pubblicazione: (2025)
di: Lalan, Arshika, et al.
Pubblicazione: (2025)
What Drives Compositional Generalization? The Importance of Continuous Training Objectives in Visual Generative Models
di: Farid, Karim, et al.
Pubblicazione: (2025)
di: Farid, Karim, et al.
Pubblicazione: (2025)
Important metabolic diseases of ruminants: Aetiology, occurrences, prevention and control measures
di: Dr. Rajat Buragohain
Pubblicazione: (2025)
di: Dr. Rajat Buragohain
Pubblicazione: (2025)
Metastability, chaos and spectrum tomography for Bose-Hubbard rings and chains
di: Rajat, et al.
Pubblicazione: (2026)
di: Rajat, et al.
Pubblicazione: (2026)
Exploring Quantum Materials & Applications: A Review
di: Goyal, Rajat Kumar
Pubblicazione: (2024)
di: Goyal, Rajat Kumar
Pubblicazione: (2024)
Leveraging PointNet and PointNet++ for Lyft Point Cloud Classification Challenge
di: Doshi, Rajat K.
Pubblicazione: (2024)
di: Doshi, Rajat K.
Pubblicazione: (2024)
Divine Speech in Human Words: Thomistic Engagements with Scripture, Emmanuel Durand, OP, with Matthew K.Minerd (ed.), The Catholic University of America Press, 2022 (ISBN 978‐0‐8132‐3536‐3), xvi + 462 pp., hb $65
di: Rajat Denzil Acharya
Pubblicazione: (2024)
di: Rajat Denzil Acharya
Pubblicazione: (2024)
Organometallic Ru(III) Catalysts for α‐Alkylation of Carbonyl Compounds using Alcohols: Mechanistic Insights via Detection of Key Intermediates
di: Sain Singh, et al.
Pubblicazione: (2024)
di: Sain Singh, et al.
Pubblicazione: (2024)
Generalizable Blood Cell Detection via Unified Dataset and Faster R-CNN
di: Sahay, Siddharth
Pubblicazione: (2025)
di: Sahay, Siddharth
Pubblicazione: (2025)
REPENSAR LA PROFUNDIZACIÓN FINANCIERA: ESTABILIDAD Y CRECIMIENTO EN LOS MERCADOS EMERGENTES
di: Ratna Sahay
Pubblicazione: (2015)
di: Ratna Sahay
Pubblicazione: (2015)
A note on the Möbius uncertainty principle for posets
di: Sahay, Anurag
Pubblicazione: (2026)
di: Sahay, Anurag
Pubblicazione: (2026)
Bounds on Linear Turán Number for Trees
di: Adak, Rajat, et al.
Pubblicazione: (2026)
di: Adak, Rajat, et al.
Pubblicazione: (2026)
Grammar Boosting: A New Technique for Proving Lower Bounds for Computation over Compressed Data
di: De, Rajat, et al.
Pubblicazione: (2023)
di: De, Rajat, et al.
Pubblicazione: (2023)
Can the Rookies Cut the Tough Cookie? Exploring the Use of LLMs for SQL Equivalence Checking
di: Singh, Rajat, et al.
Pubblicazione: (2024)
di: Singh, Rajat, et al.
Pubblicazione: (2024)
Word Break on SLP-Compressed Texts
di: De, Rajat, et al.
Pubblicazione: (2025)
di: De, Rajat, et al.
Pubblicazione: (2025)
Safe Distributed Control of Multi-Robot Systems with Communication Delays
di: Ballotta, Luca, et al.
Pubblicazione: (2024)
di: Ballotta, Luca, et al.
Pubblicazione: (2024)
Evaluation of residual sodium carbonate (RSC) and sodium adsorption ratio (SAR) in fresh water and laundry grey water for irrigation usage
di: Rajat Khapra, et al.
Pubblicazione: (2024)
di: Rajat Khapra, et al.
Pubblicazione: (2024)
Hallucinations and Truth: A Comprehensive Accuracy Evaluation of RAG, LoRA and DoRA
di: Baqar, Mohammad, et al.
Pubblicazione: (2025)
di: Baqar, Mohammad, et al.
Pubblicazione: (2025)
A segment of Euler product associated to a certain Dirichlet series
di: Gupta, Rajat, et al.
Pubblicazione: (2024)
di: Gupta, Rajat, et al.
Pubblicazione: (2024)
The Future of Software Testing: AI-Powered Test Case Generation and Validation
di: Baqar, Mohammad, et al.
Pubblicazione: (2024)
di: Baqar, Mohammad, et al.
Pubblicazione: (2024)
Documenti analoghi
-
MoPEFT: A Mixture-of-PEFTs for the Segment Anything Model
di: Sahay, Rajat, et al.
Pubblicazione: (2024) -
IMAS: A Comprehensive Agentic Approach to Rural Healthcare Delivery
di: Gangavarapu, Agasthya, et al.
Pubblicazione: (2024) -
LLM Robustness Leaderboard v1 --Technical report
di: Lefebvre, Pierre Peigné -, et al.
Pubblicazione: (2025) -
Introducing L2M3, A Multilingual Medical Large Language Model to Advance Health Equity in Low-Resource Regions
di: Gangavarapu, Agasthya
Pubblicazione: (2024) -
Enhancing Guardrails for Safe and Secure Healthcare AI
di: Gangavarapu, Ananya
Pubblicazione: (2024)