Cost-Effective Attention Mechanisms for Low Resource Settings: Necessity & Sufficiency of Linear Transformations
Fuente:
arXiv
Salvato in:
| Autori principali: | Hosseini, Peyman, Hosseini, Mehran, Castro, Ignacio, Purver, Matthew |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PDFMathTranslate: Scientific Document Translation Preserving Layouts
di: Ouyang, Rongxin, et al.
Pubblicazione: (2025)
di: Ouyang, Rongxin, et al.
Pubblicazione: (2025)
AI-based Drone Assisted Human Rescue in Disaster Environments: Challenges and Opportunities
di: Papyan, Narek, et al.
Pubblicazione: (2024)
di: Papyan, Narek, et al.
Pubblicazione: (2024)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
di: Patel, Hitesh Laxmichand, et al.
Pubblicazione: (2025)
di: Patel, Hitesh Laxmichand, et al.
Pubblicazione: (2025)
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
di: Yim, Wen-wai, et al.
Pubblicazione: (2025)
di: Yim, Wen-wai, et al.
Pubblicazione: (2025)
Unpacking Hateful Memes: Presupposed Context and False Claims
di: Cai, Weibin, et al.
Pubblicazione: (2025)
di: Cai, Weibin, et al.
Pubblicazione: (2025)
Reference-Guided Verdict: LLMs-as-Judges in Automatic Evaluation of Free-Form QA
di: Badshah, Sher, et al.
Pubblicazione: (2024)
di: Badshah, Sher, et al.
Pubblicazione: (2024)
A Computational Approach to Modeling Conversational Systems: Analyzing Large-Scale Quasi-Patterned Dialogue Flows
di: Ammar, Mohamed Achref Ben, et al.
Pubblicazione: (2025)
di: Ammar, Mohamed Achref Ben, et al.
Pubblicazione: (2025)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
di: Huo, Dongjie, et al.
Pubblicazione: (2026)
di: Huo, Dongjie, et al.
Pubblicazione: (2026)
Large Language Models Report Subjective Experience Under Self-Referential Processing
di: Berg, Cameron, et al.
Pubblicazione: (2025)
di: Berg, Cameron, et al.
Pubblicazione: (2025)
Developing Acoustic Models for Automatic Speech Recognition in Swedish
di: Salvi, Giampiero
Pubblicazione: (2024)
di: Salvi, Giampiero
Pubblicazione: (2024)
K-MetBench: A Multi-Dimensional Benchmark for Fine-Grained Evaluation of Expert Reasoning, Locality, and Multimodality in Meteorology
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
VGA: Vision GUI Assistant -- Minimizing Hallucinations through Image-Centric Fine-Tuning
di: Meng, Ziyang, et al.
Pubblicazione: (2024)
di: Meng, Ziyang, et al.
Pubblicazione: (2024)
Beyond RNNs: Benchmarking Attention-Based Image Captioning Models
di: Yanambakkam, Hemanth Teja, et al.
Pubblicazione: (2025)
di: Yanambakkam, Hemanth Teja, et al.
Pubblicazione: (2025)
Multi-Agent Object Detection Framework Based on Raspberry Pi YOLO Detector and Slack-Ollama Natural Language Interface
di: Kalušev, Vladimir, et al.
Pubblicazione: (2026)
di: Kalušev, Vladimir, et al.
Pubblicazione: (2026)
Chatbots put to the test in math and logic problems: A preliminary comparison and assessment of ChatGPT-3.5, ChatGPT-4, and Google Bard
di: Plevris, Vagelis, et al.
Pubblicazione: (2023)
di: Plevris, Vagelis, et al.
Pubblicazione: (2023)
Prompt Tuned Embedding Classification for Multi-Label Industry Sector Allocation
di: Buchner, Valentin Leonhard, et al.
Pubblicazione: (2023)
di: Buchner, Valentin Leonhard, et al.
Pubblicazione: (2023)
Enhancing OCR for Sino-Vietnamese Language Processing via Fine-tuned PaddleOCRv5
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
GIM: Evaluating models via tasks that integrate multiple cognitive domains
di: Patel, Rohit, et al.
Pubblicazione: (2026)
di: Patel, Rohit, et al.
Pubblicazione: (2026)
Context-Dependent Affordance Computation in Vision-Language Models
di: Farzulla, Murad
Pubblicazione: (2026)
di: Farzulla, Murad
Pubblicazione: (2026)
TowerVision: Understanding and Improving Multilinguality in Vision-Language Models
di: Viveiros, André G., et al.
Pubblicazione: (2025)
di: Viveiros, André G., et al.
Pubblicazione: (2025)
LegalGuardian: A Privacy-Preserving Framework for Secure Integration of Large Language Models in Legal Practice
di: Demir, M. Mikail, et al.
Pubblicazione: (2025)
di: Demir, M. Mikail, et al.
Pubblicazione: (2025)
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation
di: Qi, Dekang, et al.
Pubblicazione: (2026)
di: Qi, Dekang, et al.
Pubblicazione: (2026)
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
di: Sarkar, Nilesh, et al.
Pubblicazione: (2026)
di: Sarkar, Nilesh, et al.
Pubblicazione: (2026)
FrameFusion: Combining Similarity and Importance for Video Token Reduction on Large Vision Language Models
di: Fu, Tianyu, et al.
Pubblicazione: (2024)
di: Fu, Tianyu, et al.
Pubblicazione: (2024)
The MSR-Video to Text Dataset with Clean Annotations
di: Chen, Haoran, et al.
Pubblicazione: (2021)
di: Chen, Haoran, et al.
Pubblicazione: (2021)
Mubeen AI: A Specialized Arabic Language Model for Heritage Preservation and User Intent Understanding
di: Aljafari, Mohammed, et al.
Pubblicazione: (2025)
di: Aljafari, Mohammed, et al.
Pubblicazione: (2025)
Vibe-Creation: The Epistemology of Human-AI Emergent Cognition
di: Levin, Ilya
Pubblicazione: (2026)
di: Levin, Ilya
Pubblicazione: (2026)
Data and AI governance: Promoting equity, ethics, and fairness in large language models
di: Abhishek, Alok, et al.
Pubblicazione: (2025)
di: Abhishek, Alok, et al.
Pubblicazione: (2025)
SHARP: Social Harm Analysis via Risk Profiles for Measuring Inequities in Large Language Models
di: Abhishek, Alok, et al.
Pubblicazione: (2026)
di: Abhishek, Alok, et al.
Pubblicazione: (2026)
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
di: Abhishek, Alok, et al.
Pubblicazione: (2025)
di: Abhishek, Alok, et al.
Pubblicazione: (2025)
Reasoning Promotes Robustness in Theory of Mind Tasks
di: de Haan, Ian B., et al.
Pubblicazione: (2026)
di: de Haan, Ian B., et al.
Pubblicazione: (2026)
CLEV: LLM-Based Evaluation Through Lightweight Efficient Voting for Free-Form Question-Answering
di: Badshah, Sher, et al.
Pubblicazione: (2025)
di: Badshah, Sher, et al.
Pubblicazione: (2025)
JAM: Controllable and Responsible Text Generation via Causal Reasoning and Latent Vector Manipulation
di: Huang, Yingbing, et al.
Pubblicazione: (2025)
di: Huang, Yingbing, et al.
Pubblicazione: (2025)
A Survey on Vision-Language-Action Models for Embodied AI
di: Ma, Yueen, et al.
Pubblicazione: (2024)
di: Ma, Yueen, et al.
Pubblicazione: (2024)
Does CLIP perceive art the same way we do?
di: Asperti, Andrea, et al.
Pubblicazione: (2025)
di: Asperti, Andrea, et al.
Pubblicazione: (2025)
ReaderLM-v2: Small Language Model for HTML to Markdown and JSON
di: Wang, Feng, et al.
Pubblicazione: (2025)
di: Wang, Feng, et al.
Pubblicazione: (2025)
RCI: A Score for Evaluating Global and Local Reasoning in Multimodal Benchmarks
di: Agarwal, Amit, et al.
Pubblicazione: (2025)
di: Agarwal, Amit, et al.
Pubblicazione: (2025)
Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception
di: Bajpai, Ashutosh, et al.
Pubblicazione: (2026)
di: Bajpai, Ashutosh, et al.
Pubblicazione: (2026)
SUGARCREPE++ Dataset: Vision-Language Model Sensitivity to Semantic and Lexical Alterations
di: Dumpala, Sri Harsha, et al.
Pubblicazione: (2024)
di: Dumpala, Sri Harsha, et al.
Pubblicazione: (2024)
Documenti analoghi
-
PDFMathTranslate: Scientific Document Translation Preserving Layouts
di: Ouyang, Rongxin, et al.
Pubblicazione: (2025) -
AI-based Drone Assisted Human Rescue in Disaster Environments: Challenges and Opportunities
di: Papyan, Narek, et al.
Pubblicazione: (2024) -
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
di: Patel, Hitesh Laxmichand, et al.
Pubblicazione: (2025) -
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
di: Yim, Wen-wai, et al.
Pubblicazione: (2025) -
Unpacking Hateful Memes: Presupposed Context and False Claims
di: Cai, Weibin, et al.
Pubblicazione: (2025)