Voting-based Multimodal Automatic Deception Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Touma, Lana, Horani, Mohammad Al, Tailouni, Manar, Dahabiah, Anas, Jallad, Khloud Al |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SyriSign: A Parallel Corpus for Arabic Text to Syrian Arabic Sign Language Translation
di: Khalil, Mohammad Amer, et al.
Pubblicazione: (2026)
di: Khalil, Mohammad Amer, et al.
Pubblicazione: (2026)
Deciphering Emotions in Children Storybooks: A Comparative Analysis of Multimodal LLMs in Educational Applications
di: Asseri, Bushra, et al.
Pubblicazione: (2025)
di: Asseri, Bushra, et al.
Pubblicazione: (2025)
Arabic Little STT: Arabic Children Speech Recognition Dataset
di: Alkadri, Mouhand, et al.
Pubblicazione: (2025)
di: Alkadri, Mouhand, et al.
Pubblicazione: (2025)
ArEEG_Chars: Dataset for Envisioned Speech Recognition using EEG for Arabic Characters
di: Darwish, Hazem, et al.
Pubblicazione: (2024)
di: Darwish, Hazem, et al.
Pubblicazione: (2024)
ArEEG_Words: Dataset for Envisioned Speech Recognition using EEG for Arabic Words
di: Darwish, Hazem, et al.
Pubblicazione: (2024)
di: Darwish, Hazem, et al.
Pubblicazione: (2024)
Survey of NLU Benchmarks Diagnosing Linguistic Phenomena: Why not Standardize Diagnostics Benchmarks?
di: Jallad, Khloud AL, et al.
Pubblicazione: (2025)
di: Jallad, Khloud AL, et al.
Pubblicazione: (2025)
CAF-Mamba: Mamba-Based Cross-Modal Adaptive Attention Fusion for Multimodal Depression Detection
di: Zhou, Bowen, et al.
Pubblicazione: (2026)
di: Zhou, Bowen, et al.
Pubblicazione: (2026)
Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs
di: You, Keen, et al.
Pubblicazione: (2024)
di: You, Keen, et al.
Pubblicazione: (2024)
A Picture is Worth a Thousand (Correct) Captions: A Vision-Guided Judge-Corrector System for Multimodal Machine Translation
di: Betala, Siddharth, et al.
Pubblicazione: (2025)
di: Betala, Siddharth, et al.
Pubblicazione: (2025)
Learning Multimodal Cues of Children's Uncertainty
di: Cheng, Qi, et al.
Pubblicazione: (2024)
di: Cheng, Qi, et al.
Pubblicazione: (2024)
Measuring Agreeableness Bias in Multimodal Models
di: Lim, Jaehyuk, et al.
Pubblicazione: (2024)
di: Lim, Jaehyuk, et al.
Pubblicazione: (2024)
Abjad-Kids: An Arabic Speech Classification Dataset for Primary Education
di: Snoubara, Abdul Aziz, et al.
Pubblicazione: (2026)
di: Snoubara, Abdul Aziz, et al.
Pubblicazione: (2026)
Learning 6-DoF Fine-grained Grasp Detection Based on Part Affordance Grounding
di: Song, Yaoxian, et al.
Pubblicazione: (2023)
di: Song, Yaoxian, et al.
Pubblicazione: (2023)
mEBAL: A Multimodal Database for Eye Blink Detection and Attention Level Estimation
di: Daza, Roberto, et al.
Pubblicazione: (2020)
di: Daza, Roberto, et al.
Pubblicazione: (2020)
ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows
di: Sun, Qiushi, et al.
Pubblicazione: (2025)
di: Sun, Qiushi, et al.
Pubblicazione: (2025)
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT
di: Gao, Nan, et al.
Pubblicazione: (2023)
di: Gao, Nan, et al.
Pubblicazione: (2023)
Long-Term Ad Memorability: Understanding & Generating Memorable Ads
di: SI, Harini, et al.
Pubblicazione: (2023)
di: SI, Harini, et al.
Pubblicazione: (2023)
A Review on Large Language Models for Visual Analytics
di: Agarwal, Navya Sonal, et al.
Pubblicazione: (2025)
di: Agarwal, Navya Sonal, et al.
Pubblicazione: (2025)
Morae: Proactively Pausing UI Agents for User Choices
di: Peng, Yi-Hao, et al.
Pubblicazione: (2025)
di: Peng, Yi-Hao, et al.
Pubblicazione: (2025)
What Color Scheme is More Effective in Assisting Readers to Locate Information in a Color-Coded Article?
di: Ng, Ho Yin, et al.
Pubblicazione: (2024)
di: Ng, Ho Yin, et al.
Pubblicazione: (2024)
UI-E2I-Synth: Advancing GUI Grounding with Large-Scale Instruction Synthesis
di: Liu, Xinyi, et al.
Pubblicazione: (2025)
di: Liu, Xinyi, et al.
Pubblicazione: (2025)
GPT-5 Model Corrected GPT-4V's Chart Reading Errors, Not Prompting
di: Yang, Kaichun, et al.
Pubblicazione: (2025)
di: Yang, Kaichun, et al.
Pubblicazione: (2025)
GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
di: Luo, Run, et al.
Pubblicazione: (2025)
di: Luo, Run, et al.
Pubblicazione: (2025)
CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
di: Verma, Arnav, et al.
Pubblicazione: (2025)
di: Verma, Arnav, et al.
Pubblicazione: (2025)
UIClip: A Data-driven Model for Assessing User Interface Design
di: Wu, Jason, et al.
Pubblicazione: (2024)
di: Wu, Jason, et al.
Pubblicazione: (2024)
True (VIS) Lies: Analyzing How Generative AI Recognizes Intentionality, Rhetoric, and Misleadingness in Visualization Lies
di: Blasilli, Graziano, et al.
Pubblicazione: (2026)
di: Blasilli, Graziano, et al.
Pubblicazione: (2026)
OS-ATLAS: A Foundation Action Model for Generalist GUI Agents
di: Wu, Zhiyong, et al.
Pubblicazione: (2024)
di: Wu, Zhiyong, et al.
Pubblicazione: (2024)
Computer-Use Agents as Judges for Generative User Interface
di: Lin, Kevin Qinghong, et al.
Pubblicazione: (2025)
di: Lin, Kevin Qinghong, et al.
Pubblicazione: (2025)
ScreenQA: Large-Scale Question-Answer Pairs over Mobile App Screenshots
di: Hsiao, Yu-Chung, et al.
Pubblicazione: (2022)
di: Hsiao, Yu-Chung, et al.
Pubblicazione: (2022)
Fool Me Once? Contrasting Textual and Visual Explanations in a Clinical Decision-Support Setting
di: Kayser, Maxime, et al.
Pubblicazione: (2024)
di: Kayser, Maxime, et al.
Pubblicazione: (2024)
What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric
di: Kerkouri, Mohamed Amine, et al.
Pubblicazione: (2026)
di: Kerkouri, Mohamed Amine, et al.
Pubblicazione: (2026)
VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents
di: Mazumdar, Amrita, et al.
Pubblicazione: (2026)
di: Mazumdar, Amrita, et al.
Pubblicazione: (2026)
SpatialViz-Bench: A Cognitively-Grounded Benchmark for Diagnosing Spatial Visualization in MLLMs
di: Wang, Siting, et al.
Pubblicazione: (2025)
di: Wang, Siting, et al.
Pubblicazione: (2025)
SiMing-Bench: Evaluating Procedural Correctness from Continuous Interactions in Clinical Skill Videos
di: Huang, Xiyang, et al.
Pubblicazione: (2026)
di: Huang, Xiyang, et al.
Pubblicazione: (2026)
EvoDiagram: Agentic Editable Diagram Creation via Design Expertise Evolution
di: Wang, Tianfu, et al.
Pubblicazione: (2026)
di: Wang, Tianfu, et al.
Pubblicazione: (2026)
InterFeedback: Unveiling Interactive Intelligence of Large Multimodal Models via Human Feedback
di: Zhao, Henry Hengyuan, et al.
Pubblicazione: (2025)
di: Zhao, Henry Hengyuan, et al.
Pubblicazione: (2025)
Analyzing Persona Effects in Generated Explanations from Multimodal LLM Agents in Urban Perception
di: da Silva, Neemias, et al.
Pubblicazione: (2026)
di: da Silva, Neemias, et al.
Pubblicazione: (2026)
OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web
di: Kapoor, Raghav, et al.
Pubblicazione: (2024)
di: Kapoor, Raghav, et al.
Pubblicazione: (2024)
How Good (Or Bad) Are LLMs at Detecting Misleading Visualizations?
di: Lo, Leo Yu-Ho, et al.
Pubblicazione: (2024)
di: Lo, Leo Yu-Ho, et al.
Pubblicazione: (2024)
E-ANT: A Large-Scale Dataset for Efficient Automatic GUI NavigaTion
di: Wang, Ke, et al.
Pubblicazione: (2024)
di: Wang, Ke, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SyriSign: A Parallel Corpus for Arabic Text to Syrian Arabic Sign Language Translation
di: Khalil, Mohammad Amer, et al.
Pubblicazione: (2026) -
Deciphering Emotions in Children Storybooks: A Comparative Analysis of Multimodal LLMs in Educational Applications
di: Asseri, Bushra, et al.
Pubblicazione: (2025) -
Arabic Little STT: Arabic Children Speech Recognition Dataset
di: Alkadri, Mouhand, et al.
Pubblicazione: (2025) -
ArEEG_Chars: Dataset for Envisioned Speech Recognition using EEG for Arabic Characters
di: Darwish, Hazem, et al.
Pubblicazione: (2024) -
ArEEG_Words: Dataset for Envisioned Speech Recognition using EEG for Arabic Words
di: Darwish, Hazem, et al.
Pubblicazione: (2024)