OpenLifelogQA: An Open-Ended Multi-Modal Lifelog Question-Answering Dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tran, Quang-Linh, Le, Hoang-Bao, Diep, Tuong-Nghiem, Nguyen, Binh, Jones, Gareth J. F., Gurrin, Cathal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The State-of-the-Art in Lifelog Retrieval: A Review of Progress at the ACM Lifelog Search Challenge Workshop 2022-24
von: Tran, Allie, et al.
Veröffentlicht: (2025)
von: Tran, Allie, et al.
Veröffentlicht: (2025)
lifeXplore at the Lifelog Search Challenge 2020
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025)
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025)
lifeXplore at the Lifelog Search Challenge 2021
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025)
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025)
Lifelogging As An Extreme Form of Personal Information Management -- What Lessons To Learn
von: Tran, Ly-Duyen, et al.
Veröffentlicht: (2024)
von: Tran, Ly-Duyen, et al.
Veröffentlicht: (2024)
Visual Lifelog Retrieval through Captioning-Enhanced Interpretation
von: Shih, Yu-Fei, et al.
Veröffentlicht: (2025)
von: Shih, Yu-Fei, et al.
Veröffentlicht: (2025)
Quizzard@INOVA Challenge 2025 -- Track A: Plug-and-Play Technique in Interleaved Multi-Image Model
von: Cuong, Dinh Viet, et al.
Veröffentlicht: (2025)
von: Cuong, Dinh Viet, et al.
Veröffentlicht: (2025)
LSC-ADL: An Activity of Daily Living (ADL)-Annotated Lifelog Dataset Generated via Semi-Automatic Clustering
von: Ho-Le, Minh-Quan, et al.
Veröffentlicht: (2025)
von: Ho-Le, Minh-Quan, et al.
Veröffentlicht: (2025)
The CASTLE 2024 Dataset: Advancing the Art of Multimodal Understanding
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
FIGROTD: A Friendly-to-Handle Dataset for Image Guided Retrieval with Optional Text
von: Le, Hoang-Bao, et al.
Veröffentlicht: (2025)
von: Le, Hoang-Bao, et al.
Veröffentlicht: (2025)
Results of the 2025 Video Browser Showdown
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
Results of the 2024 Video Browser Showdown
von: Rossetto, Luca, et al.
Veröffentlicht: (2024)
von: Rossetto, Luca, et al.
Veröffentlicht: (2024)
Jamendo-QA: A Large-Scale Music Question Answering Dataset
von: Koh, Junyoung, et al.
Veröffentlicht: (2025)
von: Koh, Junyoung, et al.
Veröffentlicht: (2025)
Towards Signboard-Oriented Visual Question Answering: ViSignVQA Dataset, Method and Benchmark
von: Nguyen, Hieu Minh, et al.
Veröffentlicht: (2025)
von: Nguyen, Hieu Minh, et al.
Veröffentlicht: (2025)
UNION: A Lightweight Target Representation for Efficient Zero-Shot Image-Guided Retrieval with Optional Textual Queries
von: Le, Hoang-Bao, et al.
Veröffentlicht: (2025)
von: Le, Hoang-Bao, et al.
Veröffentlicht: (2025)
DMC$^3$: Dual-Modal Counterfactual Contrastive Construction for Egocentric Video Question Answering
von: Zou, Jiayi, et al.
Veröffentlicht: (2025)
von: Zou, Jiayi, et al.
Veröffentlicht: (2025)
Jamendo-MT-QA: A Benchmark for Multi-Track Comparative Music Question Answering
von: Koh, Junyoung, et al.
Veröffentlicht: (2026)
von: Koh, Junyoung, et al.
Veröffentlicht: (2026)
FedMAC: Tackling Partial-Modality Missing in Federated Learning with Cross-Modal Aggregation and Contrastive Regularization
von: Nguyen, Manh Duong, et al.
Veröffentlicht: (2024)
von: Nguyen, Manh Duong, et al.
Veröffentlicht: (2024)
SRA: Semantic Relation-Aware Flowchart Question Answering
von: Li, Xinyu, et al.
Veröffentlicht: (2026)
von: Li, Xinyu, et al.
Veröffentlicht: (2026)
Ada2I: Enhancing Modality Balance for Multimodal Conversational Emotion Recognition
von: Nguyen, Cam-Van Thi, et al.
Veröffentlicht: (2024)
von: Nguyen, Cam-Van Thi, et al.
Veröffentlicht: (2024)
Integrated Semantic and Temporal Alignment for Interactive Video Retrieval
von: Luu, Thanh-Danh, et al.
Veröffentlicht: (2025)
von: Luu, Thanh-Danh, et al.
Veröffentlicht: (2025)
BDIQA: A New Dataset for Video Question Answering to Explore Cognitive Reasoning through Theory of Mind
von: Mao, Yuanyuan, et al.
Veröffentlicht: (2024)
von: Mao, Yuanyuan, et al.
Veröffentlicht: (2024)
MHier-RAG: Multi-Modal RAG for Visual-Rich Document Question-Answering via Hierarchical and Multi-Granularity Reasoning
von: Gong, Ziyu, et al.
Veröffentlicht: (2025)
von: Gong, Ziyu, et al.
Veröffentlicht: (2025)
Structure-Aware Residual-Center Representation for Self-Supervised Open-Set 3D Cross-Modal Retrieval
von: Xu, Yang, et al.
Veröffentlicht: (2024)
von: Xu, Yang, et al.
Veröffentlicht: (2024)
Proposing Smart System for Detecting and Monitoring Vehicle Using Multiobject Multicamera Tracking
von: Phat Nguyen Huu, et al.
Veröffentlicht: (2024)
von: Phat Nguyen Huu, et al.
Veröffentlicht: (2024)
Learning Trimodal Relation for Audio-Visual Question Answering with Missing Modality
von: Park, Kyu Ri, et al.
Veröffentlicht: (2024)
von: Park, Kyu Ri, et al.
Veröffentlicht: (2024)
Ges-QA: A Multidimensional Quality Assessment Dataset for Audio-to-3D Gesture Generation
von: Gao, Zhilin, et al.
Veröffentlicht: (2025)
von: Gao, Zhilin, et al.
Veröffentlicht: (2025)
CUS-QA: Local-Knowledge-Oriented Open-Ended Question Answering Dataset
von: Libovický, Jindřich, et al.
Veröffentlicht: (2025)
von: Libovický, Jindřich, et al.
Veröffentlicht: (2025)
Fact-Checking at Scale: Multimodal AI for Authenticity and Context Verification in Online Media
von: Phan, Van-Hoang, et al.
Veröffentlicht: (2025)
von: Phan, Van-Hoang, et al.
Veröffentlicht: (2025)
CinePile: A Long Video Question Answering Dataset and Benchmark
von: Rawal, Ruchit, et al.
Veröffentlicht: (2024)
von: Rawal, Ruchit, et al.
Veröffentlicht: (2024)
EMID: An Emotional Aligned Dataset in Audio-Visual Modality
von: Zou, Jialing, et al.
Veröffentlicht: (2023)
von: Zou, Jialing, et al.
Veröffentlicht: (2023)
Towards Open-Vocabulary Video Semantic Segmentation
von: Li, Xinhao, et al.
Veröffentlicht: (2024)
von: Li, Xinhao, et al.
Veröffentlicht: (2024)
Memory-Centric Embodied Question Answering
von: Zhai, Mingliang, et al.
Veröffentlicht: (2025)
von: Zhai, Mingliang, et al.
Veröffentlicht: (2025)
Can Current Detectors Catch Face-to-Voice Deepfake Attacks?
von: Nguyen, Nguyen Linh Bao, et al.
Veröffentlicht: (2025)
von: Nguyen, Nguyen Linh Bao, et al.
Veröffentlicht: (2025)
KAN-Based Fusion of Dual-Domain for Audio-Driven Facial Landmarks Generation
von: Vo-Thanh, Hoang-Son, et al.
Veröffentlicht: (2024)
von: Vo-Thanh, Hoang-Son, et al.
Veröffentlicht: (2024)
Modeling the Impacts of Swipe Delay on User Quality of Experience in Short Video Streaming
von: Nguyen, Duc V., et al.
Veröffentlicht: (2026)
von: Nguyen, Duc V., et al.
Veröffentlicht: (2026)
MagicLens: Self-Supervised Image Retrieval with Open-Ended Instructions
von: Zhang, Kai, et al.
Veröffentlicht: (2024)
von: Zhang, Kai, et al.
Veröffentlicht: (2024)
EMO100DB: An Open Dataset of Improvised Songs with Emotion Data
von: Hwang, Daeun, et al.
Veröffentlicht: (2025)
von: Hwang, Daeun, et al.
Veröffentlicht: (2025)
E-FreeM2: Efficient Training-Free Multi-Scale and Cross-Modal News Verification via MLLMs
von: Phan, Van-Hoang, et al.
Veröffentlicht: (2025)
von: Phan, Van-Hoang, et al.
Veröffentlicht: (2025)
TimeLogic Challenge @ CVPR 2026: Strong MLLMs Meet Evidence-Seeking Agents for Temporal-Logic Video Question Answering
von: Xu, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Xu, Zhaoyang, et al.
Veröffentlicht: (2026)
MUDI: A Multimodal Biomedical Dataset for Understanding Pharmacodynamic Drug-Drug Interactions
von: Ngo, Tung-Lam, et al.
Veröffentlicht: (2025)
von: Ngo, Tung-Lam, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The State-of-the-Art in Lifelog Retrieval: A Review of Progress at the ACM Lifelog Search Challenge Workshop 2022-24
von: Tran, Allie, et al.
Veröffentlicht: (2025) -
lifeXplore at the Lifelog Search Challenge 2020
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025) -
lifeXplore at the Lifelog Search Challenge 2021
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025) -
Lifelogging As An Extreme Form of Personal Information Management -- What Lessons To Learn
von: Tran, Ly-Duyen, et al.
Veröffentlicht: (2024) -
Visual Lifelog Retrieval through Captioning-Enhanced Interpretation
von: Shih, Yu-Fei, et al.
Veröffentlicht: (2025)