Metrics and evaluations for computational and sustainable AI efficiency
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Hongyuan, Liu, Xinyang, Hu, Guosheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration
von: Tu, Dezhan, et al.
Veröffentlicht: (2024)
von: Tu, Dezhan, et al.
Veröffentlicht: (2024)
PipeFusion: Patch-level Pipeline Parallelism for Diffusion Transformers Inference
von: Fang, Jiarui, et al.
Veröffentlicht: (2024)
von: Fang, Jiarui, et al.
Veröffentlicht: (2024)
Twill: Scheduling Compound AI Systems on Heterogeneous Mobile Edge Platforms
von: Taufique, Zain, et al.
Veröffentlicht: (2025)
von: Taufique, Zain, et al.
Veröffentlicht: (2025)
Investigation of Energy-efficient AI Model Architectures and Compression Techniques for "Green" Fetal Brain Segmentation
von: Mazurek, Szymon, et al.
Veröffentlicht: (2024)
von: Mazurek, Szymon, et al.
Veröffentlicht: (2024)
AI for Service: Proactive Assistance with AI Glasses
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device
von: Wu, Yushu, et al.
Veröffentlicht: (2024)
von: Wu, Yushu, et al.
Veröffentlicht: (2024)
FastCache: Fast Caching for Diffusion Transformer Through Learnable Linear Approximation
von: Liu, Dong, et al.
Veröffentlicht: (2025)
von: Liu, Dong, et al.
Veröffentlicht: (2025)
Fast-SEnSeI: Lightweight Sensor-Independent Cloud Masking for On-board Multispectral Sensors
von: Kněžík, Jan, et al.
Veröffentlicht: (2025)
von: Kněžík, Jan, et al.
Veröffentlicht: (2025)
Zero-Shot, But at What Cost? Unveiling the Hidden Overhead of MILS's LLM-CLIP Framework for Image Captioning
von: Benhammou, Yassir, et al.
Veröffentlicht: (2025)
von: Benhammou, Yassir, et al.
Veröffentlicht: (2025)
Spatially-Aware Evaluation of Segmentation Uncertainty
von: Zeevi, Tal, et al.
Veröffentlicht: (2025)
von: Zeevi, Tal, et al.
Veröffentlicht: (2025)
VLM in a flash: I/O-Efficient Sparsification of Vision-Language Model via Neuron Chunking
von: Yang, Kichang, et al.
Veröffentlicht: (2025)
von: Yang, Kichang, et al.
Veröffentlicht: (2025)
Multi-domain performance analysis with scores tailored to user preferences
von: Piérard, Sébastien, et al.
Veröffentlicht: (2025)
von: Piérard, Sébastien, et al.
Veröffentlicht: (2025)
SpargeAttention: Accurate and Training-free Sparse Attention Accelerating Any Model Inference
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
What Is the Optimal Ranking Score Between Precision and Recall? We Can Always Find It and It Is Rarely $F_1$
von: Piérard, Sébastien, et al.
Veröffentlicht: (2025)
von: Piérard, Sébastien, et al.
Veröffentlicht: (2025)
Revisiting 16-bit Neural Network Training: A Practical Approach for Resource-Limited Learning
von: Yun, Juyoung, et al.
Veröffentlicht: (2023)
von: Yun, Juyoung, et al.
Veröffentlicht: (2023)
A Closer Look at Data Augmentation Strategies for Finetuning-Based Low/Few-Shot Object Detection
von: Li, Vladislav, et al.
Veröffentlicht: (2024)
von: Li, Vladislav, et al.
Veröffentlicht: (2024)
Dual-Signal Adaptive KV-Cache Optimization for Long-Form Video Understanding in Vision-Language Models
von: Sai, Vishnu, et al.
Veröffentlicht: (2026)
von: Sai, Vishnu, et al.
Veröffentlicht: (2026)
Class Confidence Aware Reweighting for Long Tailed Learning
von: Jagati, Brainard Philemon, et al.
Veröffentlicht: (2026)
von: Jagati, Brainard Philemon, et al.
Veröffentlicht: (2026)
Learning Human-Perceived Fakeness in AI-Generated Videos via Multimodal LLMs
von: Fu, Xingyu, et al.
Veröffentlicht: (2025)
von: Fu, Xingyu, et al.
Veröffentlicht: (2025)
EXPERT: An Explainable Image Captioning Evaluation Metric with Structured Explanations
von: Kim, Hyunjong, et al.
Veröffentlicht: (2025)
von: Kim, Hyunjong, et al.
Veröffentlicht: (2025)
TeleEval-OS: Performance evaluations of large language models for operations scheduling
von: Wang, Yanyan, et al.
Veröffentlicht: (2025)
von: Wang, Yanyan, et al.
Veröffentlicht: (2025)
SMILE: A Composite Lexical-Semantic Metric for Question-Answering Evaluation
von: Kendre, Shrikant, et al.
Veröffentlicht: (2025)
von: Kendre, Shrikant, et al.
Veröffentlicht: (2025)
DENEB: A Hallucination-Robust Automatic Evaluation Metric for Image Captioning
von: Matsuda, Kazuki, et al.
Veröffentlicht: (2024)
von: Matsuda, Kazuki, et al.
Veröffentlicht: (2024)
Certainly Uncertain: A Benchmark and Metric for Multimodal Epistemic and Aleatoric Awareness
von: Chandu, Khyathi Raghavi, et al.
Veröffentlicht: (2024)
von: Chandu, Khyathi Raghavi, et al.
Veröffentlicht: (2024)
Polos: Multimodal Metric Learning from Human Feedback for Image Captioning
von: Wada, Yuiga, et al.
Veröffentlicht: (2024)
von: Wada, Yuiga, et al.
Veröffentlicht: (2024)
Shifting AI Efficiency From Model-Centric to Data-Centric Compression
von: Liu, Xuyang, et al.
Veröffentlicht: (2025)
von: Liu, Xuyang, et al.
Veröffentlicht: (2025)
Vibe-Eval: A hard evaluation suite for measuring progress of multimodal language models
von: Padlewski, Piotr, et al.
Veröffentlicht: (2024)
von: Padlewski, Piotr, et al.
Veröffentlicht: (2024)
Sample-efficient Integration of New Modalities into Large Language Models
von: İnce, Osman Batur, et al.
Veröffentlicht: (2025)
von: İnce, Osman Batur, et al.
Veröffentlicht: (2025)
CulturalFrames: Assessing Cultural Expectation Alignment in Text-to-Image Models and Evaluation Metrics
von: Nayak, Shravan, et al.
Veröffentlicht: (2025)
von: Nayak, Shravan, et al.
Veröffentlicht: (2025)
CRIMSON: A Clinically-Grounded LLM-Based Metric for Generative Radiology Report Evaluation
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2026)
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2026)
AI Meets Brain: Memory Systems from Cognitive Neuroscience to Autonomous Agents
von: Liang, Jiafeng, et al.
Veröffentlicht: (2025)
von: Liang, Jiafeng, et al.
Veröffentlicht: (2025)
VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI
von: Cheng, Sijie, et al.
Veröffentlicht: (2024)
von: Cheng, Sijie, et al.
Veröffentlicht: (2024)
FLEUR: An Explainable Reference-Free Evaluation Metric for Image Captioning Using a Large Multimodal Model
von: Lee, Yebin, et al.
Veröffentlicht: (2024)
von: Lee, Yebin, et al.
Veröffentlicht: (2024)
G-VEval: A Versatile Metric for Evaluating Image and Video Captions Using GPT-4o
von: Tong, Tony Cheng, et al.
Veröffentlicht: (2024)
von: Tong, Tony Cheng, et al.
Veröffentlicht: (2024)
Euclid: Supercharging Multimodal LLMs with Synthetic High-Fidelity Visual Descriptions
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
RandLoRA: Full-rank parameter-efficient fine-tuning of large models
von: Albert, Paul, et al.
Veröffentlicht: (2025)
von: Albert, Paul, et al.
Veröffentlicht: (2025)
Energy-Aware LLMs: A step towards sustainable AI for downstream applications
von: Tran, Nguyen Phuc, et al.
Veröffentlicht: (2025)
von: Tran, Nguyen Phuc, et al.
Veröffentlicht: (2025)
Who Evaluates the Evaluations? Objectively Scoring Text-to-Image Prompt Coherence Metrics with T2IScoreScore (TS2)
von: Saxon, Michael, et al.
Veröffentlicht: (2024)
von: Saxon, Michael, et al.
Veröffentlicht: (2024)
OraPO: Oracle-educated Reinforcement Learning for Data-efficient and Factual Radiology Report Generation
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2025)
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2025)
Attention-based transformer models for image captioning across languages: An in-depth survey and evaluation
von: Albadarneh, Israa A., et al.
Veröffentlicht: (2025)
von: Albadarneh, Israa A., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration
von: Tu, Dezhan, et al.
Veröffentlicht: (2024) -
PipeFusion: Patch-level Pipeline Parallelism for Diffusion Transformers Inference
von: Fang, Jiarui, et al.
Veröffentlicht: (2024) -
Twill: Scheduling Compound AI Systems on Heterogeneous Mobile Edge Platforms
von: Taufique, Zain, et al.
Veröffentlicht: (2025) -
Investigation of Energy-efficient AI Model Architectures and Compression Techniques for "Green" Fetal Brain Segmentation
von: Mazurek, Szymon, et al.
Veröffentlicht: (2024) -
AI for Service: Proactive Assistance with AI Glasses
von: Wen, Zichen, et al.
Veröffentlicht: (2025)