SciIntegrity-Bench: A Benchmark for Evaluating Academic Integrity in AI Scientist Systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Zonglin, Liu, Xingtong, Xu, Xinyan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CI-Bench: Benchmarking Contextual Integrity of AI Assistants on Synthetic Data
di: Cheng, Zhao, et al.
Pubblicazione: (2024)
di: Cheng, Zhao, et al.
Pubblicazione: (2024)
MPCI-Bench: A Benchmark for Multimodal Pairwise Contextual Integrity Evaluation of Language Model Agents
di: Wang, Shouju, et al.
Pubblicazione: (2026)
di: Wang, Shouju, et al.
Pubblicazione: (2026)
SciIntBench: Measuring LLM Compliance with Research Integrity Norms Under Adversarial Framing
di: Meguimtsop, Almene De Meran, et al.
Pubblicazione: (2026)
di: Meguimtsop, Almene De Meran, et al.
Pubblicazione: (2026)
AI-generated Essays: Characteristics and Implications on Automated Scoring and Academic Integrity
di: Zhong, Yang, et al.
Pubblicazione: (2024)
di: Zhong, Yang, et al.
Pubblicazione: (2024)
RewardHackingAgents: Benchmarking Evaluation Integrity for LLM ML-Engineering Agents
di: Atinafu, Yonas, et al.
Pubblicazione: (2026)
di: Atinafu, Yonas, et al.
Pubblicazione: (2026)
Stylometry Analysis of Human and Machine Text for Academic Integrity
di: Albaqami, Hezam, et al.
Pubblicazione: (2026)
di: Albaqami, Hezam, et al.
Pubblicazione: (2026)
SciCode: A Research Coding Benchmark Curated by Scientists
di: Tian, Minyang, et al.
Pubblicazione: (2024)
di: Tian, Minyang, et al.
Pubblicazione: (2024)
Use of AI Tools: Guidelines to Maintain Academic Integrity in Computing Colleges
di: El-boghdadi, Hatem M., et al.
Pubblicazione: (2026)
di: El-boghdadi, Hatem M., et al.
Pubblicazione: (2026)
Behavioral Integrity Verification for AI Agent Skills
di: Wu, Yuhao, et al.
Pubblicazione: (2026)
di: Wu, Yuhao, et al.
Pubblicazione: (2026)
Detecting AI-Generated Text in Educational Content: Leveraging Machine Learning and Explainable AI for Academic Integrity
di: Najjar, Ayat A., et al.
Pubblicazione: (2025)
di: Najjar, Ayat A., et al.
Pubblicazione: (2025)
Semantic Integrity Constraints: Declarative Guardrails for AI-Augmented Data Processing Systems
di: Lee, Alexander W., et al.
Pubblicazione: (2025)
di: Lee, Alexander W., et al.
Pubblicazione: (2025)
Semantic Reward Collapse and the Preservation of Epistemic Integrity in Adaptive AI Systems
di: Parris, William
Pubblicazione: (2026)
di: Parris, William
Pubblicazione: (2026)
IntegrityAI at GenAI Detection Task 2: Detecting Machine-Generated Academic Essays in English and Arabic Using ELECTRA and Stylometry
di: AL-Smadi, Mohammad
Pubblicazione: (2025)
di: AL-Smadi, Mohammad
Pubblicazione: (2025)
Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression
di: Liu, Xiang, et al.
Pubblicazione: (2025)
di: Liu, Xiang, et al.
Pubblicazione: (2025)
Operationalizing Contextual Integrity in Privacy-Conscious Assistants
di: Ghalebikesabi, Sahra, et al.
Pubblicazione: (2024)
di: Ghalebikesabi, Sahra, et al.
Pubblicazione: (2024)
Meta-Sealing: A Revolutionizing Integrity Assurance Protocol for Transparent, Tamper-Proof, and Trustworthy AI System
di: Krishnamoorthy, Mahesh Vaijainthymala
Pubblicazione: (2024)
di: Krishnamoorthy, Mahesh Vaijainthymala
Pubblicazione: (2024)
Advancing AI with Integrity: Ethical Challenges and Solutions in Neural Machine Translation
di: Kimera, Richard, et al.
Pubblicazione: (2024)
di: Kimera, Richard, et al.
Pubblicazione: (2024)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
di: Jia, Xiao
Pubblicazione: (2026)
di: Jia, Xiao
Pubblicazione: (2026)
What's Pulling the Strings? Evaluating Integrity and Attribution in AI Training and Inference through Concept Shift
di: Chang, Jiamin, et al.
Pubblicazione: (2025)
di: Chang, Jiamin, et al.
Pubblicazione: (2025)
Do Vision-Language Models Respect Contextual Integrity in Location Disclosure?
di: Yang, Ruixin, et al.
Pubblicazione: (2026)
di: Yang, Ruixin, et al.
Pubblicazione: (2026)
AI Integrity: A New Paradigm for Verifiable AI Governance
di: Lee, Seulki
Pubblicazione: (2026)
di: Lee, Seulki
Pubblicazione: (2026)
SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents
di: Ai, Kuangshi, et al.
Pubblicazione: (2026)
di: Ai, Kuangshi, et al.
Pubblicazione: (2026)
Modular Speaker Architecture: A Framework for Sustaining Responsibility and Contextual Integrity in Multi-Agent AI Communication
di: Toh, Khe-Han, et al.
Pubblicazione: (2025)
di: Toh, Khe-Han, et al.
Pubblicazione: (2025)
Structural Enforcement of Goal Integrity in AI Agents via Separation-of-Powers Architecture
di: Xiang, Rong
Pubblicazione: (2026)
di: Xiang, Rong
Pubblicazione: (2026)
Revisiting Software Engineering Education in the Era of Large Language Models: A Curriculum Adaptation and Academic Integrity Framework
di: Degerli, Mustafa
Pubblicazione: (2026)
di: Degerli, Mustafa
Pubblicazione: (2026)
HumanStudy-Bench: Towards AI Agent Design for Participant Simulation
di: Liu, Xuan, et al.
Pubblicazione: (2026)
di: Liu, Xuan, et al.
Pubblicazione: (2026)
ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry
di: Xu, Tianze, et al.
Pubblicazione: (2025)
di: Xu, Tianze, et al.
Pubblicazione: (2025)
DataSciBench: An LLM Agent Benchmark for Data Science
di: Zhang, Dan, et al.
Pubblicazione: (2025)
di: Zhang, Dan, et al.
Pubblicazione: (2025)
Survey on Plagiarism Detection in Large Language Models: The Impact of ChatGPT and Gemini on Academic Integrity
di: Pudasaini, Shushanta, et al.
Pubblicazione: (2024)
di: Pudasaini, Shushanta, et al.
Pubblicazione: (2024)
Epistemic Integrity in Large Language Models
di: Ghafouri, Bijean, et al.
Pubblicazione: (2024)
di: Ghafouri, Bijean, et al.
Pubblicazione: (2024)
A Survey of AI Scientists
di: Tie, Guiyao, et al.
Pubblicazione: (2025)
di: Tie, Guiyao, et al.
Pubblicazione: (2025)
AirQualityBench: A Realistic Evaluation Benchmark for Global Air Quality Forecasting
di: Xu, Xing, et al.
Pubblicazione: (2026)
di: Xu, Xing, et al.
Pubblicazione: (2026)
From Blind Solvers to Logical Thinkers: Benchmarking LLMs' Logical Integrity on Faulty Mathematical Problems
di: Rahman, A M Muntasir, et al.
Pubblicazione: (2024)
di: Rahman, A M Muntasir, et al.
Pubblicazione: (2024)
MatSciBench: Benchmarking the Reasoning Ability of Large Language Models in Materials Science
di: Zhang, Junkai, et al.
Pubblicazione: (2025)
di: Zhang, Junkai, et al.
Pubblicazione: (2025)
Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking
di: You, Ziyang, et al.
Pubblicazione: (2026)
di: You, Ziyang, et al.
Pubblicazione: (2026)
Maintaining Journalistic Integrity in the Digital Age: A Comprehensive NLP Framework for Evaluating Online News Content
di: Bojic, Ljubisa, et al.
Pubblicazione: (2024)
di: Bojic, Ljubisa, et al.
Pubblicazione: (2024)
CSR-Bench: A Benchmark for Evaluating the Cross-modal Safety and Reliability of MLLMs
di: Liu, Yuxuan, et al.
Pubblicazione: (2026)
di: Liu, Yuxuan, et al.
Pubblicazione: (2026)
TAI3: Testing Agent Integrity in Interpreting User Intent
di: Feng, Shiwei, et al.
Pubblicazione: (2025)
di: Feng, Shiwei, et al.
Pubblicazione: (2025)
Zer0n: An AI-Assisted Vulnerability Discovery and Blockchain-Backed Integrity Framework
di: Parmar, Harshil, et al.
Pubblicazione: (2026)
di: Parmar, Harshil, et al.
Pubblicazione: (2026)
SciVideoBench: Benchmarking Scientific Video Reasoning in Large Multimodal Models
di: Deng, Andong, et al.
Pubblicazione: (2025)
di: Deng, Andong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CI-Bench: Benchmarking Contextual Integrity of AI Assistants on Synthetic Data
di: Cheng, Zhao, et al.
Pubblicazione: (2024) -
MPCI-Bench: A Benchmark for Multimodal Pairwise Contextual Integrity Evaluation of Language Model Agents
di: Wang, Shouju, et al.
Pubblicazione: (2026) -
SciIntBench: Measuring LLM Compliance with Research Integrity Norms Under Adversarial Framing
di: Meguimtsop, Almene De Meran, et al.
Pubblicazione: (2026) -
AI-generated Essays: Characteristics and Implications on Automated Scoring and Academic Integrity
di: Zhong, Yang, et al.
Pubblicazione: (2024) -
RewardHackingAgents: Benchmarking Evaluation Integrity for LLM ML-Engineering Agents
di: Atinafu, Yonas, et al.
Pubblicazione: (2026)