TPIFM: A Task-Aware Model for Evaluating Perceptual Interaction Fluency in Remote AR Collaboration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Jiarun, Wan, Ninghao, Yang, Fuzheng, Lin, Weisi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Perception to Cognition: How Latency Affects Interaction Fluency and Social Presence in VR Conferencing
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
Dynamic Multimodal Expression Generation for LLM-Driven Pedagogical Agents: From User Experience Perspective
von: Wan, Ninghao, et al.
Veröffentlicht: (2026)
von: Wan, Ninghao, et al.
Veröffentlicht: (2026)
Latency Effects on Multi-Dimensional QoE in Networked VR Whiteboards
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
Perceptual Quality Assessment of Octree-RAHT Encoded 3D Point Clouds
von: Duan, Dongshuai, et al.
Veröffentlicht: (2024)
von: Duan, Dongshuai, et al.
Veröffentlicht: (2024)
RBFIM: Perceptual Quality Assessment for Compressed Point Clouds Using Radial Basis Function Interpolation
von: Chen, Zhang, et al.
Veröffentlicht: (2025)
von: Chen, Zhang, et al.
Veröffentlicht: (2025)
Rendering-Oriented 3D Point Cloud Attribute Compression using Sparse Tensor-based Transformer
von: Huo, Xiao, et al.
Veröffentlicht: (2024)
von: Huo, Xiao, et al.
Veröffentlicht: (2024)
Harmony-Aware Music-driven Motion Synthesis with Perceptual Constraint on UGC Datasets
von: Wu, Xinyi, et al.
Veröffentlicht: (2025)
von: Wu, Xinyi, et al.
Veröffentlicht: (2025)
Feature Coding in the Era of Large Models: Dataset, Test Conditions, and Benchmark
von: Gao, Changsheng, et al.
Veröffentlicht: (2024)
von: Gao, Changsheng, et al.
Veröffentlicht: (2024)
Multimodal Interaction Modeling via Self-Supervised Multi-Task Learning for Review Helpfulness Prediction
von: Gong, HongLin, et al.
Veröffentlicht: (2024)
von: Gong, HongLin, et al.
Veröffentlicht: (2024)
PC-JND: Subjective Study and Dataset on Just Noticeable Difference for Point Clouds in 6DoF Virtual Reality
von: Fan, Chunling, et al.
Veröffentlicht: (2025)
von: Fan, Chunling, et al.
Veröffentlicht: (2025)
Perceptual-oriented Learned Image Compression with Dynamic Kernel
von: Fu, Nianxiang, et al.
Veröffentlicht: (2024)
von: Fu, Nianxiang, et al.
Veröffentlicht: (2024)
MM-InstructEval: Zero-Shot Evaluation of (Multimodal) Large Language Models on Multimodal Reasoning Tasks
von: Yang, Xiaocui, et al.
Veröffentlicht: (2024)
von: Yang, Xiaocui, et al.
Veröffentlicht: (2024)
SFQA: A Comprehensive Perceptual Quality Assessment Dataset for Singing Face Generation
von: Gao, Zhilin, et al.
Veröffentlicht: (2026)
von: Gao, Zhilin, et al.
Veröffentlicht: (2026)
Dynamic Interaction-Aware and Causality-Disentangled Framework for Multimodal Sentiment Analysis
von: Dong, Guangyuan, et al.
Veröffentlicht: (2026)
von: Dong, Guangyuan, et al.
Veröffentlicht: (2026)
Task Presentation and Human Perception in Interactive Video Retrieval
von: Willis, Nina, et al.
Veröffentlicht: (2024)
von: Willis, Nina, et al.
Veröffentlicht: (2024)
VSpeechLM: A Visual Speech Language Model for Visual Text-to-Speech Task
von: Wang, Yuyue, et al.
Veröffentlicht: (2025)
von: Wang, Yuyue, et al.
Veröffentlicht: (2025)
PP-Motion: Physical-Perceptual Fidelity Evaluation for Human Motion Generation
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
DT-UFC: Universal Large Model Feature Coding via Peaky-to-Balanced Distribution Transformation
von: Gao, Changsheng, et al.
Veröffentlicht: (2025)
von: Gao, Changsheng, et al.
Veröffentlicht: (2025)
Design and Development of Laughter Recognition System Based on Multimodal Fusion and Deep Learning
von: Zhao, Fuzheng, et al.
Veröffentlicht: (2024)
von: Zhao, Fuzheng, et al.
Veröffentlicht: (2024)
TOP:A New Target-Audience Oriented Content Paraphrase Task
von: Lin, Boda, et al.
Veröffentlicht: (2024)
von: Lin, Boda, et al.
Veröffentlicht: (2024)
Beyond Text: Multimodal Jailbreaking of Vision-Language and Audio Models through Perceptually Simple Transformations
von: Kumar, Divyanshu, et al.
Veröffentlicht: (2025)
von: Kumar, Divyanshu, et al.
Veröffentlicht: (2025)
Adaptive Wireless Image Semantic Transmission and Over-The-Air Testing
von: Ding, Jiarun, et al.
Veröffentlicht: (2024)
von: Ding, Jiarun, et al.
Veröffentlicht: (2024)
ConvBench: A Multi-Turn Conversation Evaluation Benchmark with Hierarchical Capability for Large Vision-Language Models
von: Liu, Shuo, et al.
Veröffentlicht: (2024)
von: Liu, Shuo, et al.
Veröffentlicht: (2024)
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
von: Xu, Junhao, et al.
Veröffentlicht: (2025)
von: Xu, Junhao, et al.
Veröffentlicht: (2025)
Challenging Dataset and Multi-modal Gated Mixture of Experts Model for Remote Sensing Copy-Move Forgery Understanding
von: Zhang, Ze, et al.
Veröffentlicht: (2025)
von: Zhang, Ze, et al.
Veröffentlicht: (2025)
ESVQA: Perceptual Quality Assessment of Egocentric Spatial Videos
von: Zhu, Xilei, et al.
Veröffentlicht: (2024)
von: Zhu, Xilei, et al.
Veröffentlicht: (2024)
FakeBench: Probing Explainable Fake Image Detection via Large Multimodal Models
von: Li, Yixuan, et al.
Veröffentlicht: (2024)
von: Li, Yixuan, et al.
Veröffentlicht: (2024)
Scaling Audio-Visual Quality Assessment Dataset via Crowdsourcing
von: Yang, Renyu, et al.
Veröffentlicht: (2026)
von: Yang, Renyu, et al.
Veröffentlicht: (2026)
CDIO: Cross-Domain Inference Optimization with Resource Preference Prediction for Edge-Cloud Collaboration
von: Yang, Zheming, et al.
Veröffentlicht: (2025)
von: Yang, Zheming, et al.
Veröffentlicht: (2025)
LMM4Edit: Benchmarking and Evaluating Multimodal Image Editing with LMMs
von: Xu, Zitong, et al.
Veröffentlicht: (2025)
von: Xu, Zitong, et al.
Veröffentlicht: (2025)
Learning Perceptual Representations for Gaming NR-VQA with Multi-Task FR Signals
von: Chen, Yu-Chih, et al.
Veröffentlicht: (2026)
von: Chen, Yu-Chih, et al.
Veröffentlicht: (2026)
EmotionTalk: An Interactive Chinese Multimodal Emotion Dataset With Rich Annotations
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
von: Sun, Haoqin, et al.
Veröffentlicht: (2025)
Wireless Multi-User Interactive Virtual Reality in Metaverse with Edge-Device Collaborative Computing
von: Xu, Caolu, et al.
Veröffentlicht: (2024)
von: Xu, Caolu, et al.
Veröffentlicht: (2024)
Multi-task Just Recognizable Difference for Video Coding for Machines: Database, Model, and Coding Application
von: Liu, Junqi, et al.
Veröffentlicht: (2026)
von: Liu, Junqi, et al.
Veröffentlicht: (2026)
MMTB: Evaluating Terminal Agents on Multimedia-File Tasks
von: Heo, Chiyeong, et al.
Veröffentlicht: (2026)
von: Heo, Chiyeong, et al.
Veröffentlicht: (2026)
Q-Adapt: Adapting LMM for Visual Quality Assessment with Progressive Instruction Tuning
von: Lu, Yiting, et al.
Veröffentlicht: (2025)
von: Lu, Yiting, et al.
Veröffentlicht: (2025)
Evaluating Magic Leap 2 Tool Tracking for AR Sensor Guidance in Industrial Inspections
von: Masuhr, Christian, et al.
Veröffentlicht: (2025)
von: Masuhr, Christian, et al.
Veröffentlicht: (2025)
TUNA: Comprehensive Fine-grained Temporal Understanding Evaluation on Dense Dynamic Videos
von: Kong, Fanheng, et al.
Veröffentlicht: (2025)
von: Kong, Fanheng, et al.
Veröffentlicht: (2025)
Startup Delay Aware Short Video Ordering: Problem, Model, and A Reinforcement Learning based Algorithm
von: Gao, Zhipeng, et al.
Veröffentlicht: (2024)
von: Gao, Zhipeng, et al.
Veröffentlicht: (2024)
PRISM-XR: Empowering Privacy-Aware XR Collaboration with Multimodal Large Language Models
von: Chen, Jiangong, et al.
Veröffentlicht: (2026)
von: Chen, Jiangong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
From Perception to Cognition: How Latency Affects Interaction Fluency and Social Presence in VR Conferencing
von: Song, Jiarun, et al.
Veröffentlicht: (2026) -
Dynamic Multimodal Expression Generation for LLM-Driven Pedagogical Agents: From User Experience Perspective
von: Wan, Ninghao, et al.
Veröffentlicht: (2026) -
Latency Effects on Multi-Dimensional QoE in Networked VR Whiteboards
von: Song, Jiarun, et al.
Veröffentlicht: (2026) -
Perceptual Quality Assessment of Octree-RAHT Encoded 3D Point Clouds
von: Duan, Dongshuai, et al.
Veröffentlicht: (2024) -
RBFIM: Perceptual Quality Assessment for Compressed Point Clouds Using Radial Basis Function Interpolation
von: Chen, Zhang, et al.
Veröffentlicht: (2025)