Interdisciplinary Translations: Sensory Perception as a Universal Language
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kang, Xindi, Huang, Xuanyang, Song, Mingdong, Guljajeva, Varvara, Kuchera-Morin, JoAnn |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generative AI-enabled Mobile Tactical Multimedia Networks: Distribution, Generation, and Perception
von: Xu, Minrui, et al.
Veröffentlicht: (2024)
von: Xu, Minrui, et al.
Veröffentlicht: (2024)
Diverse Sign Language Translation
von: Shen, Xin, et al.
Veröffentlicht: (2024)
von: Shen, Xin, et al.
Veröffentlicht: (2024)
Task Presentation and Human Perception in Interactive Video Retrieval
von: Willis, Nina, et al.
Veröffentlicht: (2024)
von: Willis, Nina, et al.
Veröffentlicht: (2024)
Perception-Aware Video Semantic Communication
von: Huang, Yinhuan, et al.
Veröffentlicht: (2026)
von: Huang, Yinhuan, et al.
Veröffentlicht: (2026)
QoS-QoE Translation with Large Language Model
von: Yu, Yingjie, et al.
Veröffentlicht: (2026)
von: Yu, Yingjie, et al.
Veröffentlicht: (2026)
ALMol: Aligned Language-Molecule Translation LLMs through Offline Preference Contrastive Optimisation
von: Gkoumas, Dimitris
Veröffentlicht: (2024)
von: Gkoumas, Dimitris
Veröffentlicht: (2024)
Situational Agency: The Framework for Designing Behavior in Agent-based art
von: Huang, Ary-Yue, et al.
Veröffentlicht: (2025)
von: Huang, Ary-Yue, et al.
Veröffentlicht: (2025)
VSpeechLM: A Visual Speech Language Model for Visual Text-to-Speech Task
von: Wang, Yuyue, et al.
Veröffentlicht: (2025)
von: Wang, Yuyue, et al.
Veröffentlicht: (2025)
Hierarchical Refinement of Universal Multimodal Attacks on Vision-Language Models
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2026)
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2026)
EidetiCom: A Cross-modal Brain-Computer Semantic Communication Paradigm for Decoding Visual Perception
von: Zheng, Linfeng, et al.
Veröffentlicht: (2024)
von: Zheng, Linfeng, et al.
Veröffentlicht: (2024)
Universal Organizer of SAM for Unsupervised Semantic Segmentation
von: Li, Tingting, et al.
Veröffentlicht: (2024)
von: Li, Tingting, et al.
Veröffentlicht: (2024)
Rethinking Bjøntegaard Delta for Compression Efficiency Evaluation: Are We Calculating It Precisely and Reliably?
von: Hang, Xinyu, et al.
Veröffentlicht: (2024)
von: Hang, Xinyu, et al.
Veröffentlicht: (2024)
Spatial Orchestra: Locomotion Music Instruments through Spatial Exploration
von: Kim, You-Jin, et al.
Veröffentlicht: (2025)
von: Kim, You-Jin, et al.
Veröffentlicht: (2025)
From Perception to Cognition: How Latency Affects Interaction Fluency and Social Presence in VR Conferencing
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
Virtual Social Immersive Multi-Sensory E-Commerce
von: Dubey, Alpana, et al.
Veröffentlicht: (2025)
von: Dubey, Alpana, et al.
Veröffentlicht: (2025)
Listen, Pause, and Reason: Toward Perception-Grounded Hybrid Reasoning for Audio Understanding
von: Wang, Jieyi, et al.
Veröffentlicht: (2026)
von: Wang, Jieyi, et al.
Veröffentlicht: (2026)
COPA: Efficient Vision-Language Pre-training Through Collaborative Object- and Patch-Text Alignment
von: Jiang, Chaoya, et al.
Veröffentlicht: (2023)
von: Jiang, Chaoya, et al.
Veröffentlicht: (2023)
Retracted: The Development Strategy of the Multimedia Fusion Mode of Big Data Technology in Japanese Translation Teaching
von: Advances in Multimedia
Veröffentlicht: (2024)
von: Advances in Multimedia
Veröffentlicht: (2024)
RFNNS: Robust Fixed Neural Network Steganography with Universal Text-to-Image Models
von: Cheng, Yu, et al.
Veröffentlicht: (2025)
von: Cheng, Yu, et al.
Veröffentlicht: (2025)
A Survey of Multi-sensor Fusion Perception for Embodied AI: Background, Methods, Challenges and Prospects
von: Ruan, Shulan, et al.
Veröffentlicht: (2025)
von: Ruan, Shulan, et al.
Veröffentlicht: (2025)
Recognizing Everything from All Modalities at Once: Grounded Multimodal Universal Information Extraction
von: Zhang, Meishan, et al.
Veröffentlicht: (2024)
von: Zhang, Meishan, et al.
Veröffentlicht: (2024)
Seeing Sarcasm Through Different Eyes: Analyzing Multimodal Sarcasm Perception in Large Vision-Language Models
von: Chen, Junjie, et al.
Veröffentlicht: (2025)
von: Chen, Junjie, et al.
Veröffentlicht: (2025)
The Dream Within Huang Long Cave: AI-Driven Interactive Narrative for Family Storytelling and Emotional Reflection
von: Huang, Jiayang, et al.
Veröffentlicht: (2025)
von: Huang, Jiayang, et al.
Veröffentlicht: (2025)
Visions of Destruction: Exploring a Potential of Generative AI in Interactive Art
von: Sola, Mar Canet, et al.
Veröffentlicht: (2024)
von: Sola, Mar Canet, et al.
Veröffentlicht: (2024)
Learning Shared Sentiment Prototypes for Adaptive Multimodal Sentiment Analysis
von: Su, Chen, et al.
Veröffentlicht: (2026)
von: Su, Chen, et al.
Veröffentlicht: (2026)
Latency Effects on Multi-Dimensional QoE in Networked VR Whiteboards
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
Multimodal Emotion Recognition with Large Language Models
von: Zhang, Hongrui, et al.
Veröffentlicht: (2026)
von: Zhang, Hongrui, et al.
Veröffentlicht: (2026)
Visions Of Destruction: Exploring Human Impact on Nature by Navigating the Latent Space of a Diffusion Model via Gaze
von: Sola, Mar Canet, et al.
Veröffentlicht: (2023)
von: Sola, Mar Canet, et al.
Veröffentlicht: (2023)
SCI-Reason: A Dataset with Chain-of-Thought Rationales for Complex Multimodal Reasoning in Academic Areas
von: Ma, Chenghao, et al.
Veröffentlicht: (2025)
von: Ma, Chenghao, et al.
Veröffentlicht: (2025)
Large Language Models (LLMs): Deployment, Tokenomics and Sustainability
von: Dong, Haiwei, et al.
Veröffentlicht: (2024)
von: Dong, Haiwei, et al.
Veröffentlicht: (2024)
TPIFM: A Task-Aware Model for Evaluating Perceptual Interaction Fluency in Remote AR Collaboration
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
LapisGS: Layered Progressive 3D Gaussian Splatting for Adaptive Streaming
von: Shi, Yuang, et al.
Veröffentlicht: (2024)
von: Shi, Yuang, et al.
Veröffentlicht: (2024)
Automatically Generating High-Precision Simulated Road Networking in Traffic Scenario
von: Xie, Liang, et al.
Veröffentlicht: (2025)
von: Xie, Liang, et al.
Veröffentlicht: (2025)
Enabling American Sign Language Communication Under Low Data Rates
von: Santhalingam, Panneer Selvam, et al.
Veröffentlicht: (2025)
von: Santhalingam, Panneer Selvam, et al.
Veröffentlicht: (2025)
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
von: Xu, Junhao, et al.
Veröffentlicht: (2025)
von: Xu, Junhao, et al.
Veröffentlicht: (2025)
ITEACH-Net: Inverted Teacher-studEnt seArCH Network for Emotion Recognition in Conversation
von: Sun, Haiyang, et al.
Veröffentlicht: (2023)
von: Sun, Haiyang, et al.
Veröffentlicht: (2023)
A Multi-modal Fusion Network for Terrain Perception Based on Illumination Aware
von: Wang, Rui, et al.
Veröffentlicht: (2025)
von: Wang, Rui, et al.
Veröffentlicht: (2025)
Revisiting Vision-Language Features Adaptation and Inconsistency for Social Media Popularity Prediction
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
Fine-grained Knowledge Graph-driven Video-Language Learning for Action Recognition
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
Language-oriented Semantic Communication for Image Transmission with Fine-Tuned Diffusion Model
von: Wei, Xinfeng, et al.
Veröffentlicht: (2024)
von: Wei, Xinfeng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Generative AI-enabled Mobile Tactical Multimedia Networks: Distribution, Generation, and Perception
von: Xu, Minrui, et al.
Veröffentlicht: (2024) -
Diverse Sign Language Translation
von: Shen, Xin, et al.
Veröffentlicht: (2024) -
Task Presentation and Human Perception in Interactive Video Retrieval
von: Willis, Nina, et al.
Veröffentlicht: (2024) -
Perception-Aware Video Semantic Communication
von: Huang, Yinhuan, et al.
Veröffentlicht: (2026) -
QoS-QoE Translation with Large Language Model
von: Yu, Yingjie, et al.
Veröffentlicht: (2026)