Guardado en:
| Autores principales: | Rašajski, Nemanja, Trivedi, Chintan, Makantasis, Konstantinos, Liapis, Antonios, Yannakakis, Georgios N. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2402.01335 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GameVibe: A Multimodal Affective Game Corpus
por: Barthet, Matthew, et al.
Publicado: (2024)
por: Barthet, Matthew, et al.
Publicado: (2024)
Across-Game Engagement Modelling via Few-Shot Learning
por: Pinitas, Kosmas, et al.
Publicado: (2024)
por: Pinitas, Kosmas, et al.
Publicado: (2024)
Can Large Language Models Capture Video Game Engagement?
por: Melhart, David, et al.
Publicado: (2025)
por: Melhart, David, et al.
Publicado: (2025)
Dynamic Quality-Diversity Search
por: Gallotta, Roberto, et al.
Publicado: (2024)
por: Gallotta, Roberto, et al.
Publicado: (2024)
FREYR: A Framework for Recognizing and Executing Your Requests
por: Gallotta, Roberto, et al.
Publicado: (2025)
por: Gallotta, Roberto, et al.
Publicado: (2025)
Large Language Models and Games: A Survey and Roadmap
por: Gallotta, Roberto, et al.
Publicado: (2024)
por: Gallotta, Roberto, et al.
Publicado: (2024)
MTR-VP: Towards End-to-End Trajectory Planning through Context-Driven Image Encoding and Multiple Trajectory Prediction
por: Keskar, Maitrayee, et al.
Publicado: (2025)
por: Keskar, Maitrayee, et al.
Publicado: (2025)
Enhancing Object Detection with Privileged Information: A Model-Agnostic Teacher-Student Approach
por: Bartolo, Matthias, et al.
Publicado: (2026)
por: Bartolo, Matthias, et al.
Publicado: (2026)
The Procedural Content Generation Benchmark: An Open-source Testbed for Generative Challenges in Games
por: Khalifa, Ahmed, et al.
Publicado: (2025)
por: Khalifa, Ahmed, et al.
Publicado: (2025)
mAVE: A Watermark for Joint Audio-Visual Generation Models
por: Si, Luyang, et al.
Publicado: (2026)
por: Si, Luyang, et al.
Publicado: (2026)
Affectively Framework: Towards Human-like Affect-Based Agents
por: Barthet, Matthew, et al.
Publicado: (2024)
por: Barthet, Matthew, et al.
Publicado: (2024)
Few-shot Semantic Encoding and Decoding for Video Surveillance
por: Cheng, Baoping, et al.
Publicado: (2025)
por: Cheng, Baoping, et al.
Publicado: (2025)
VideoGameQA-Bench: Evaluating Vision-Language Models for Video Game Quality Assurance
por: Taesiri, Mohammad Reza, et al.
Publicado: (2025)
por: Taesiri, Mohammad Reza, et al.
Publicado: (2025)
An Integrated Framework for Multi-Granular Explanation of Video Summarization
por: Tsigos, Konstantinos, et al.
Publicado: (2024)
por: Tsigos, Konstantinos, et al.
Publicado: (2024)
GameGen-X: Interactive Open-world Game Video Generation
por: Che, Haoxuan, et al.
Publicado: (2024)
por: Che, Haoxuan, et al.
Publicado: (2024)
CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering
por: Bhosale, Mahesh, et al.
Publicado: (2026)
por: Bhosale, Mahesh, et al.
Publicado: (2026)
PAS: A Training-Free Stabilizer for Temporal Encoding in Video LLMs
por: Sun, Bowen, et al.
Publicado: (2025)
por: Sun, Bowen, et al.
Publicado: (2025)
How Much 3D Do Video Foundation Models Encode?
por: Huang, Zixuan, et al.
Publicado: (2025)
por: Huang, Zixuan, et al.
Publicado: (2025)
Encoding and Controlling Global Semantics for Long-form Video Question Answering
por: Nguyen, Thong Thanh, et al.
Publicado: (2024)
por: Nguyen, Thong Thanh, et al.
Publicado: (2024)
STANCE: Motion Coherent Video Generation Via Sparse-to-Dense Anchored Encoding
por: Chen, Zhifei, et al.
Publicado: (2025)
por: Chen, Zhifei, et al.
Publicado: (2025)
The More, the Merrier: Contrastive Fusion for Higher-Order Multimodal Alignment
por: Koutoupis, Stefanos, et al.
Publicado: (2025)
por: Koutoupis, Stefanos, et al.
Publicado: (2025)
MoCA-Video: Motion-Aware Concept Alignment for Consistent Video Editing
por: Zhang, Tong, et al.
Publicado: (2025)
por: Zhang, Tong, et al.
Publicado: (2025)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
por: Lai, Yixuan, et al.
Publicado: (2026)
por: Lai, Yixuan, et al.
Publicado: (2026)
REMAP: Regularized Matching and Partial Alignment of Video Embeddings
por: Chandra, Soumyadeep, et al.
Publicado: (2025)
por: Chandra, Soumyadeep, et al.
Publicado: (2025)
MAP-Elites with Transverse Assessment for Multimodal Problems in Creative Domains
por: Zammit, Marvin, et al.
Publicado: (2024)
por: Zammit, Marvin, et al.
Publicado: (2024)
Reasoning over the Behaviour of Objects in Video-Clips for Adverb-Type Recognition
por: Seshadri, Amrit Diggavi, et al.
Publicado: (2023)
por: Seshadri, Amrit Diggavi, et al.
Publicado: (2023)
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
por: Gopalkrishnan, Akshay, et al.
Publicado: (2024)
por: Gopalkrishnan, Akshay, et al.
Publicado: (2024)
Uncertainty-Guided Self-Questioning and Answering for Video-Language Alignment
por: Chen, Jin, et al.
Publicado: (2024)
por: Chen, Jin, et al.
Publicado: (2024)
PIPE: Physics-Informed Position Encoding for Alignment of Satellite Images and Time Series
por: Li, Haobo, et al.
Publicado: (2025)
por: Li, Haobo, et al.
Publicado: (2025)
Graph Alignment via Dual-Pass Spectral Encoding and Latent Space Communication
por: Behmanesh, Maysam, et al.
Publicado: (2025)
por: Behmanesh, Maysam, et al.
Publicado: (2025)
Sense Less, Generate More: Pre-training LiDAR Perception with Masked Autoencoders for Ultra-Efficient 3D Sensing
por: Tayebati, Sina, et al.
Publicado: (2024)
por: Tayebati, Sina, et al.
Publicado: (2024)
Video2Game: Real-time, Interactive, Realistic and Browser-Compatible Environment from a Single Video
por: Xia, Hongchi, et al.
Publicado: (2024)
por: Xia, Hongchi, et al.
Publicado: (2024)
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
por: Lee, SuBeen, et al.
Publicado: (2025)
por: Lee, SuBeen, et al.
Publicado: (2025)
V-LynX: Token Interface Alignment for Video+X LLMs
por: Park, Jungin, et al.
Publicado: (2026)
por: Park, Jungin, et al.
Publicado: (2026)
Multimodal Alignment with Cross-Attentive GRUs for Fine-Grained Video Understanding
por: Kim, Namho, et al.
Publicado: (2025)
por: Kim, Namho, et al.
Publicado: (2025)
Learning Using Privileged Information for Litter Detection
por: Bartolo, Matthias, et al.
Publicado: (2025)
por: Bartolo, Matthias, et al.
Publicado: (2025)
Comment-aided Video-Language Alignment via Contrastive Pre-training for Short-form Video Humor Detection
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
VideoLights: Feature Refinement and Cross-Task Alignment Transformer for Joint Video Highlight Detection and Moment Retrieval
por: Paul, Dhiman, et al.
Publicado: (2024)
por: Paul, Dhiman, et al.
Publicado: (2024)
Rethinking Weakly-supervised Video Temporal Grounding From a Game Perspective
por: Fang, Xiang, et al.
Publicado: (2026)
por: Fang, Xiang, et al.
Publicado: (2026)
A Hybrid Co-Finetuning Approach for Visual Bug Detection in Video Games
por: Yi, Faliu, et al.
Publicado: (2025)
por: Yi, Faliu, et al.
Publicado: (2025)
Ejemplares similares
-
GameVibe: A Multimodal Affective Game Corpus
por: Barthet, Matthew, et al.
Publicado: (2024) -
Across-Game Engagement Modelling via Few-Shot Learning
por: Pinitas, Kosmas, et al.
Publicado: (2024) -
Can Large Language Models Capture Video Game Engagement?
por: Melhart, David, et al.
Publicado: (2025) -
Dynamic Quality-Diversity Search
por: Gallotta, Roberto, et al.
Publicado: (2024) -
FREYR: A Framework for Recognizing and Executing Your Requests
por: Gallotta, Roberto, et al.
Publicado: (2025)