FactoryBench: Evaluating Industrial Machine Understanding
Fuente:
arXiv
Salvato in:
| Autori principali: | Merzouki, Yanis, Izquierdo, Coral, Ignuta-Ciuncanu, Matei, Gomez-Bracamonte, Marcos, Maggioni, Riccardo, Lombardi, Alessandro, Mazzoleni, Camilla, Martelli, Federico, Gunther, Balazs, Petersen, Jonas, Petersen, Philipp |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
HADA: Human-AI Agent Decision Alignment Architecture
di: Pitkäranta, Tapio, et al.
Pubblicazione: (2025)
di: Pitkäranta, Tapio, et al.
Pubblicazione: (2025)
Breaking the Chains of Probability: Neutrosophic Logic as a New Framework for Epistemic Uncertainty in Large Language Models
di: Leyva-Vázquez, Maikel Yelandi, et al.
Pubblicazione: (2026)
di: Leyva-Vázquez, Maikel Yelandi, et al.
Pubblicazione: (2026)
Where did you get that? Towards Summarization Attribution for Analysts
di: B, Violet, et al.
Pubblicazione: (2025)
di: B, Violet, et al.
Pubblicazione: (2025)
Practical Code RAG at Scale: Task-Aware Retrieval Design Choices under Compute Budgets
di: Galimzyanov, Timur, et al.
Pubblicazione: (2025)
di: Galimzyanov, Timur, et al.
Pubblicazione: (2025)
Accelerating Monte-Carlo Tree Search with Optimized Posterior Policies
di: Frankston, Keith, et al.
Pubblicazione: (2026)
di: Frankston, Keith, et al.
Pubblicazione: (2026)
Optimizing Hospital Capacity During Pandemics: A Dual-Component Framework for Strategic Patient Relocation
di: Tabatabaee, Sadaf, et al.
Pubblicazione: (2026)
di: Tabatabaee, Sadaf, et al.
Pubblicazione: (2026)
Fine-Tuning Integrity for Modern Neural Networks: Structured Drift Proofs via Norm, Rank, and Sparsity Certificates
di: Shang, Zhenhang, et al.
Pubblicazione: (2026)
di: Shang, Zhenhang, et al.
Pubblicazione: (2026)
Can AI Outperform Human Experts in Creating Social Media Creatives?
di: Park, Eunkyung, et al.
Pubblicazione: (2024)
di: Park, Eunkyung, et al.
Pubblicazione: (2024)
AEM: Adaptive Entropy Modulation for Multi-Turn Agentic Reinforcement Learning
di: Zhao, Haotian, et al.
Pubblicazione: (2026)
di: Zhao, Haotian, et al.
Pubblicazione: (2026)
Autonomous motion in changing environment, fibrations and reaction mechanisms
di: Farber, Michael, et al.
Pubblicazione: (2025)
di: Farber, Michael, et al.
Pubblicazione: (2025)
Self-Improving Multilingual Long Reasoning via Translation-Reasoning Integrated Training
di: Liu, Junxiao, et al.
Pubblicazione: (2026)
di: Liu, Junxiao, et al.
Pubblicazione: (2026)
LLMTrace: A Corpus for Classification and Fine-Grained Localization of AI-Written Text
di: Tolstykh, Irina, et al.
Pubblicazione: (2025)
di: Tolstykh, Irina, et al.
Pubblicazione: (2025)
Relation Extraction Capabilities of LLMs on Clinical Text: A Bilingual Evaluation for English and Turkish
di: Aidynkyzy, Aidana, et al.
Pubblicazione: (2026)
di: Aidynkyzy, Aidana, et al.
Pubblicazione: (2026)
TIGQA:An Expert Annotated Question Answering Dataset in Tigrinya
di: Teklehaymanot, Hailay, et al.
Pubblicazione: (2024)
di: Teklehaymanot, Hailay, et al.
Pubblicazione: (2024)
Over-Squashing in Graph Neural Networks: A Comprehensive survey
di: Akansha, Singh
Pubblicazione: (2023)
di: Akansha, Singh
Pubblicazione: (2023)
CENTS: Generating synthetic electricity consumption time series for rare and unseen scenarios
di: Fuest, Michael, et al.
Pubblicazione: (2025)
di: Fuest, Michael, et al.
Pubblicazione: (2025)
Optimal Transport-Guided Safety in Temporal Difference Reinforcement Learning
di: Shahrooei, Zahra, et al.
Pubblicazione: (2025)
di: Shahrooei, Zahra, et al.
Pubblicazione: (2025)
CapTune: Adapting Non-Speech Captions With Anchored Generative Models
di: Huang, Jeremy Zhengqi, et al.
Pubblicazione: (2025)
di: Huang, Jeremy Zhengqi, et al.
Pubblicazione: (2025)
Semantic Web and Software Agents -- A Forgotten Wave of Artificial Intelligence?
di: Pitkäranta, Tapio, et al.
Pubblicazione: (2025)
di: Pitkäranta, Tapio, et al.
Pubblicazione: (2025)
GraphCompNet: A Position-Aware Model for Predicting and Compensating Shape Deviations in 3D Printing
di: Lee, Juheon, et al.
Pubblicazione: (2025)
di: Lee, Juheon, et al.
Pubblicazione: (2025)
Distractor Generation in Multiple-Choice Tasks: A Survey of Methods, Datasets, and Evaluation
di: Alhazmi, Elaf, et al.
Pubblicazione: (2024)
di: Alhazmi, Elaf, et al.
Pubblicazione: (2024)
A Human-in/on-the-Loop Framework for Accessible Text Generation
di: Moreno, Lourdes, et al.
Pubblicazione: (2026)
di: Moreno, Lourdes, et al.
Pubblicazione: (2026)
Semantic Content Determines Algorithmic Performance
di: Ríos-García, Martiño, et al.
Pubblicazione: (2026)
di: Ríos-García, Martiño, et al.
Pubblicazione: (2026)
Beyond Tokens in Language Models: Interpreting Activations through Text Genre Chunks
di: Benito-Rodriguez, Éloïse, et al.
Pubblicazione: (2025)
di: Benito-Rodriguez, Éloïse, et al.
Pubblicazione: (2025)
The NordDRG AI Benchmark for Large Language Models
di: Pitkäranta, Tapio
Pubblicazione: (2025)
di: Pitkäranta, Tapio
Pubblicazione: (2025)
WebGameBench: Requirement-to-Application Evaluation for Coding Agents via Browser-Native Games
di: Zhang, Wenyu, et al.
Pubblicazione: (2026)
di: Zhang, Wenyu, et al.
Pubblicazione: (2026)
A General Control Method for Human-Robot Integration
di: Feder, Maddalena, et al.
Pubblicazione: (2024)
di: Feder, Maddalena, et al.
Pubblicazione: (2024)
Developing ChatGPT for Biology and Medicine: A Complete Review of Biomedical Question Answering
di: Li, Qing, et al.
Pubblicazione: (2024)
di: Li, Qing, et al.
Pubblicazione: (2024)
The Exploration of Neural Collapse under Imbalanced Data
di: Liu, Haixia
Pubblicazione: (2024)
di: Liu, Haixia
Pubblicazione: (2024)
Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels
di: Zhao, Yaqi, et al.
Pubblicazione: (2026)
di: Zhao, Yaqi, et al.
Pubblicazione: (2026)
Cycling Race Time Prediction: A Personalized Machine Learning Approach Using Route Topology and Training Load
di: Moreno, Francisco Aguilera
Pubblicazione: (2026)
di: Moreno, Francisco Aguilera
Pubblicazione: (2026)
LLMs: A Game-Changer for Software Engineers?
di: Haque, Md Asraful
Pubblicazione: (2024)
di: Haque, Md Asraful
Pubblicazione: (2024)
RA: A machine based rational agent, Part 2, Preliminary test
di: Pantelis, G.
Pubblicazione: (2024)
di: Pantelis, G.
Pubblicazione: (2024)
RA: A machine based rational agent, Part 1
di: Pantelis, G.
Pubblicazione: (2024)
di: Pantelis, G.
Pubblicazione: (2024)
Towards Foundation Models for Relational Databases with Language Models and Graph Neural Networks
di: Wu, Jingcheng, et al.
Pubblicazione: (2026)
di: Wu, Jingcheng, et al.
Pubblicazione: (2026)
Towards Objective Gastrointestinal Auscultation: Automated Segmentation and Annotation of Bowel Sound Patterns
di: Mansour, Zahra, et al.
Pubblicazione: (2026)
di: Mansour, Zahra, et al.
Pubblicazione: (2026)
Benchmarking machine learning for bowel sound pattern classification from tabular features to pretrained models
di: Mansour, Zahra, et al.
Pubblicazione: (2025)
di: Mansour, Zahra, et al.
Pubblicazione: (2025)
Factored Diffusion Policies:Compositionally Generalized Robot Control with a Single Score Network
di: Mitra, Sayan, et al.
Pubblicazione: (2026)
di: Mitra, Sayan, et al.
Pubblicazione: (2026)
Progress and Opportunities of Foundation Models in Bioinformatics
di: Li, Qing, et al.
Pubblicazione: (2024)
di: Li, Qing, et al.
Pubblicazione: (2024)
Cognitive bias in LLM reasoning compromises interpretation of clinical oncology notes
di: Kenaston, Matthew W., et al.
Pubblicazione: (2025)
di: Kenaston, Matthew W., et al.
Pubblicazione: (2025)
Documenti analoghi
-
HADA: Human-AI Agent Decision Alignment Architecture
di: Pitkäranta, Tapio, et al.
Pubblicazione: (2025) -
Breaking the Chains of Probability: Neutrosophic Logic as a New Framework for Epistemic Uncertainty in Large Language Models
di: Leyva-Vázquez, Maikel Yelandi, et al.
Pubblicazione: (2026) -
Where did you get that? Towards Summarization Attribution for Analysts
di: B, Violet, et al.
Pubblicazione: (2025) -
Practical Code RAG at Scale: Task-Aware Retrieval Design Choices under Compute Budgets
di: Galimzyanov, Timur, et al.
Pubblicazione: (2025) -
Accelerating Monte-Carlo Tree Search with Optimized Posterior Policies
di: Frankston, Keith, et al.
Pubblicazione: (2026)