Salvato in:
| Autori principali: | Rašajski, Nemanja, Trivedi, Chintan, Makantasis, Konstantinos, Liapis, Antonios, Yannakakis, Georgios N. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2402.01335 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GameVibe: A Multimodal Affective Game Corpus
di: Barthet, Matthew, et al.
Pubblicazione: (2024)
di: Barthet, Matthew, et al.
Pubblicazione: (2024)
Across-Game Engagement Modelling via Few-Shot Learning
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
Can Large Language Models Capture Video Game Engagement?
di: Melhart, David, et al.
Pubblicazione: (2025)
di: Melhart, David, et al.
Pubblicazione: (2025)
Dynamic Quality-Diversity Search
di: Gallotta, Roberto, et al.
Pubblicazione: (2024)
di: Gallotta, Roberto, et al.
Pubblicazione: (2024)
FREYR: A Framework for Recognizing and Executing Your Requests
di: Gallotta, Roberto, et al.
Pubblicazione: (2025)
di: Gallotta, Roberto, et al.
Pubblicazione: (2025)
Large Language Models and Games: A Survey and Roadmap
di: Gallotta, Roberto, et al.
Pubblicazione: (2024)
di: Gallotta, Roberto, et al.
Pubblicazione: (2024)
MTR-VP: Towards End-to-End Trajectory Planning through Context-Driven Image Encoding and Multiple Trajectory Prediction
di: Keskar, Maitrayee, et al.
Pubblicazione: (2025)
di: Keskar, Maitrayee, et al.
Pubblicazione: (2025)
Enhancing Object Detection with Privileged Information: A Model-Agnostic Teacher-Student Approach
di: Bartolo, Matthias, et al.
Pubblicazione: (2026)
di: Bartolo, Matthias, et al.
Pubblicazione: (2026)
The Procedural Content Generation Benchmark: An Open-source Testbed for Generative Challenges in Games
di: Khalifa, Ahmed, et al.
Pubblicazione: (2025)
di: Khalifa, Ahmed, et al.
Pubblicazione: (2025)
mAVE: A Watermark for Joint Audio-Visual Generation Models
di: Si, Luyang, et al.
Pubblicazione: (2026)
di: Si, Luyang, et al.
Pubblicazione: (2026)
Affectively Framework: Towards Human-like Affect-Based Agents
di: Barthet, Matthew, et al.
Pubblicazione: (2024)
di: Barthet, Matthew, et al.
Pubblicazione: (2024)
Few-shot Semantic Encoding and Decoding for Video Surveillance
di: Cheng, Baoping, et al.
Pubblicazione: (2025)
di: Cheng, Baoping, et al.
Pubblicazione: (2025)
VideoGameQA-Bench: Evaluating Vision-Language Models for Video Game Quality Assurance
di: Taesiri, Mohammad Reza, et al.
Pubblicazione: (2025)
di: Taesiri, Mohammad Reza, et al.
Pubblicazione: (2025)
An Integrated Framework for Multi-Granular Explanation of Video Summarization
di: Tsigos, Konstantinos, et al.
Pubblicazione: (2024)
di: Tsigos, Konstantinos, et al.
Pubblicazione: (2024)
GameGen-X: Interactive Open-world Game Video Generation
di: Che, Haoxuan, et al.
Pubblicazione: (2024)
di: Che, Haoxuan, et al.
Pubblicazione: (2024)
CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering
di: Bhosale, Mahesh, et al.
Pubblicazione: (2026)
di: Bhosale, Mahesh, et al.
Pubblicazione: (2026)
PAS: A Training-Free Stabilizer for Temporal Encoding in Video LLMs
di: Sun, Bowen, et al.
Pubblicazione: (2025)
di: Sun, Bowen, et al.
Pubblicazione: (2025)
How Much 3D Do Video Foundation Models Encode?
di: Huang, Zixuan, et al.
Pubblicazione: (2025)
di: Huang, Zixuan, et al.
Pubblicazione: (2025)
Encoding and Controlling Global Semantics for Long-form Video Question Answering
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
STANCE: Motion Coherent Video Generation Via Sparse-to-Dense Anchored Encoding
di: Chen, Zhifei, et al.
Pubblicazione: (2025)
di: Chen, Zhifei, et al.
Pubblicazione: (2025)
The More, the Merrier: Contrastive Fusion for Higher-Order Multimodal Alignment
di: Koutoupis, Stefanos, et al.
Pubblicazione: (2025)
di: Koutoupis, Stefanos, et al.
Pubblicazione: (2025)
MoCA-Video: Motion-Aware Concept Alignment for Consistent Video Editing
di: Zhang, Tong, et al.
Pubblicazione: (2025)
di: Zhang, Tong, et al.
Pubblicazione: (2025)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
di: Lai, Yixuan, et al.
Pubblicazione: (2026)
di: Lai, Yixuan, et al.
Pubblicazione: (2026)
REMAP: Regularized Matching and Partial Alignment of Video Embeddings
di: Chandra, Soumyadeep, et al.
Pubblicazione: (2025)
di: Chandra, Soumyadeep, et al.
Pubblicazione: (2025)
MAP-Elites with Transverse Assessment for Multimodal Problems in Creative Domains
di: Zammit, Marvin, et al.
Pubblicazione: (2024)
di: Zammit, Marvin, et al.
Pubblicazione: (2024)
Reasoning over the Behaviour of Objects in Video-Clips for Adverb-Type Recognition
di: Seshadri, Amrit Diggavi, et al.
Pubblicazione: (2023)
di: Seshadri, Amrit Diggavi, et al.
Pubblicazione: (2023)
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
di: Gopalkrishnan, Akshay, et al.
Pubblicazione: (2024)
di: Gopalkrishnan, Akshay, et al.
Pubblicazione: (2024)
Uncertainty-Guided Self-Questioning and Answering for Video-Language Alignment
di: Chen, Jin, et al.
Pubblicazione: (2024)
di: Chen, Jin, et al.
Pubblicazione: (2024)
PIPE: Physics-Informed Position Encoding for Alignment of Satellite Images and Time Series
di: Li, Haobo, et al.
Pubblicazione: (2025)
di: Li, Haobo, et al.
Pubblicazione: (2025)
Graph Alignment via Dual-Pass Spectral Encoding and Latent Space Communication
di: Behmanesh, Maysam, et al.
Pubblicazione: (2025)
di: Behmanesh, Maysam, et al.
Pubblicazione: (2025)
Sense Less, Generate More: Pre-training LiDAR Perception with Masked Autoencoders for Ultra-Efficient 3D Sensing
di: Tayebati, Sina, et al.
Pubblicazione: (2024)
di: Tayebati, Sina, et al.
Pubblicazione: (2024)
Video2Game: Real-time, Interactive, Realistic and Browser-Compatible Environment from a Single Video
di: Xia, Hongchi, et al.
Pubblicazione: (2024)
di: Xia, Hongchi, et al.
Pubblicazione: (2024)
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
di: Lee, SuBeen, et al.
Pubblicazione: (2025)
di: Lee, SuBeen, et al.
Pubblicazione: (2025)
V-LynX: Token Interface Alignment for Video+X LLMs
di: Park, Jungin, et al.
Pubblicazione: (2026)
di: Park, Jungin, et al.
Pubblicazione: (2026)
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding
di: Kim, Namho, et al.
Pubblicazione: (2025)
di: Kim, Namho, et al.
Pubblicazione: (2025)
Learning Using Privileged Information for Litter Detection
di: Bartolo, Matthias, et al.
Pubblicazione: (2025)
di: Bartolo, Matthias, et al.
Pubblicazione: (2025)
Comment-aided Video-Language Alignment via Contrastive Pre-training for Short-form Video Humor Detection
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
VideoLights: Feature Refinement and Cross-Task Alignment Transformer for Joint Video Highlight Detection and Moment Retrieval
di: Paul, Dhiman, et al.
Pubblicazione: (2024)
di: Paul, Dhiman, et al.
Pubblicazione: (2024)
Rethinking Weakly-supervised Video Temporal Grounding From a Game Perspective
di: Fang, Xiang, et al.
Pubblicazione: (2026)
di: Fang, Xiang, et al.
Pubblicazione: (2026)
A Hybrid Co-Finetuning Approach for Visual Bug Detection in Video Games
di: Yi, Faliu, et al.
Pubblicazione: (2025)
di: Yi, Faliu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GameVibe: A Multimodal Affective Game Corpus
di: Barthet, Matthew, et al.
Pubblicazione: (2024) -
Across-Game Engagement Modelling via Few-Shot Learning
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024) -
Can Large Language Models Capture Video Game Engagement?
di: Melhart, David, et al.
Pubblicazione: (2025) -
Dynamic Quality-Diversity Search
di: Gallotta, Roberto, et al.
Pubblicazione: (2024) -
FREYR: A Framework for Recognizing and Executing Your Requests
di: Gallotta, Roberto, et al.
Pubblicazione: (2025)