Sensorium Arc: AI Agent System for Oceanic Data Exploration and Interactive Eco-Art
Fuente:
arXiv
Salvato in:
| Autori principali: | Bissell, Noah, Paley, Ethan, Harrison, Joshua, Calil, Juliano, Lee, Myungin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Back to Basics: Revisiting ASR in the Age of Voice Agents
di: Tay, Geeyang, et al.
Pubblicazione: (2026)
di: Tay, Geeyang, et al.
Pubblicazione: (2026)
The Dream Within Huang Long Cave: AI-Driven Interactive Narrative for Family Storytelling and Emotional Reflection
di: Huang, Jiayang, et al.
Pubblicazione: (2025)
di: Huang, Jiayang, et al.
Pubblicazione: (2025)
Human-Data Interaction, Exploration, and Visualization in the AI Era: Challenges and Opportunities
di: Fekete, Jean-Daniel, et al.
Pubblicazione: (2026)
di: Fekete, Jean-Daniel, et al.
Pubblicazione: (2026)
LLMER: Crafting Interactive Extended Reality Worlds with JSON Data Generated by Large Language Models
di: Chen, Jiangong, et al.
Pubblicazione: (2025)
di: Chen, Jiangong, et al.
Pubblicazione: (2025)
MMTB: Evaluating Terminal Agents on Multimedia-File Tasks
di: Heo, Chiyeong, et al.
Pubblicazione: (2026)
di: Heo, Chiyeong, et al.
Pubblicazione: (2026)
MuDoC: An Interactive Multimodal Document-grounded Conversational AI System
di: Taneja, Karan, et al.
Pubblicazione: (2025)
di: Taneja, Karan, et al.
Pubblicazione: (2025)
PETLP: A Privacy-by-Design Pipeline for Social Media Data in AI Research
di: Oh, Nick, et al.
Pubblicazione: (2025)
di: Oh, Nick, et al.
Pubblicazione: (2025)
Modeling Human Responses to Multimodal AI Content
di: Shen, Zhiqi, et al.
Pubblicazione: (2025)
di: Shen, Zhiqi, et al.
Pubblicazione: (2025)
AniME: Adaptive Multi-Agent Planning for Long Animation Generation
di: Zhang, Lisai, et al.
Pubblicazione: (2025)
di: Zhang, Lisai, et al.
Pubblicazione: (2025)
Contestable Multi-Agent Debate with Arena-based Argumentative Computation for Multimedia Verification
di: Nguyen, Truong Thanh Hung, et al.
Pubblicazione: (2026)
di: Nguyen, Truong Thanh Hung, et al.
Pubblicazione: (2026)
MetaDesigner: Advancing Artistic Typography Through AI-Driven, User-Centric, and Multilingual WordArt Synthesis
di: He, Jun-Yan, et al.
Pubblicazione: (2024)
di: He, Jun-Yan, et al.
Pubblicazione: (2024)
A Survey on Multimodal Benchmarks: In the Era of Large AI Models
di: Li, Lin, et al.
Pubblicazione: (2024)
di: Li, Lin, et al.
Pubblicazione: (2024)
MindFuse: Towards GenAI Explainability in Marketing Strategy Co-Creation
di: Farseev, Aleksandr, et al.
Pubblicazione: (2025)
di: Farseev, Aleksandr, et al.
Pubblicazione: (2025)
Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs
di: Chen, Jianhao, et al.
Pubblicazione: (2026)
di: Chen, Jianhao, et al.
Pubblicazione: (2026)
Multi-Agent System for AI-Assisted Extraction of Narrative Arcs in TV Series
di: Balestri, Roberto, et al.
Pubblicazione: (2025)
di: Balestri, Roberto, et al.
Pubblicazione: (2025)
A Survey of Multi-sensor Fusion Perception for Embodied AI: Background, Methods, Challenges and Prospects
di: Ruan, Shulan, et al.
Pubblicazione: (2025)
di: Ruan, Shulan, et al.
Pubblicazione: (2025)
End-to-End Learning-based Video Streaming Enhancement Pipeline: A Generative AI Approach
di: Artioli, Emanuele, et al.
Pubblicazione: (2025)
di: Artioli, Emanuele, et al.
Pubblicazione: (2025)
Synthesizing Sentiment-Controlled Feedback For Multimodal Text and Image Data
di: Kumar, Puneet, et al.
Pubblicazione: (2024)
di: Kumar, Puneet, et al.
Pubblicazione: (2024)
QMAVIS: Long Video-Audio Understanding using Fusion of Large Multimodal Models
di: Lin, Zixing, et al.
Pubblicazione: (2026)
di: Lin, Zixing, et al.
Pubblicazione: (2026)
Speak the Art: A Direct Speech to Image Generation Framework
di: Saeed, Mariam, et al.
Pubblicazione: (2025)
di: Saeed, Mariam, et al.
Pubblicazione: (2025)
Large Language Model Based Multi-Agent System Augmented Complex Event Processing Pipeline for Internet of Multimedia Things
di: Zeeshan, Talha, et al.
Pubblicazione: (2025)
di: Zeeshan, Talha, et al.
Pubblicazione: (2025)
Stage-Adaptive Reliability Modeling for Continuous Valence-Arousal Estimation
di: Lee, Yubeen, et al.
Pubblicazione: (2026)
di: Lee, Yubeen, et al.
Pubblicazione: (2026)
Livia: An Emotion-Aware AR Companion Powered by Modular AI Agents and Progressive Memory Compression
di: Xi, Rui, et al.
Pubblicazione: (2025)
di: Xi, Rui, et al.
Pubblicazione: (2025)
AI-Integrated Decision Support System for Real-Time Market Growth Forecasting and Multi-Source Content Diffusion Analytics
di: Yin, Ziqing, et al.
Pubblicazione: (2025)
di: Yin, Ziqing, et al.
Pubblicazione: (2025)
SynthGuard: An Open Platform for Detecting AI-Generated Multimedia with Multimodal LLMs
di: Desai, Shail, et al.
Pubblicazione: (2025)
di: Desai, Shail, et al.
Pubblicazione: (2025)
AesopAgent: Agent-driven Evolutionary System on Story-to-Video Production
di: Wang, Jiuniu, et al.
Pubblicazione: (2024)
di: Wang, Jiuniu, et al.
Pubblicazione: (2024)
Harnessing Self-Supervised Features for Art Classification
di: Melis, Federico, et al.
Pubblicazione: (2026)
di: Melis, Federico, et al.
Pubblicazione: (2026)
Adaptive Multi-Agent Reasoning for Text-to-Video Retrieval
di: Wu, Jiaxin, et al.
Pubblicazione: (2025)
di: Wu, Jiaxin, et al.
Pubblicazione: (2025)
Multi-Agent Game Generation and Evaluation via Audio-Visual Recordings
di: Jolicoeur-Martineau, Alexia
Pubblicazione: (2025)
di: Jolicoeur-Martineau, Alexia
Pubblicazione: (2025)
FairStream: Fair Multimedia Streaming Benchmark for Reinforcement Learning Agents
di: Weil, Jannis, et al.
Pubblicazione: (2024)
di: Weil, Jannis, et al.
Pubblicazione: (2024)
Detecting Multimedia Generated by Large AI Models: A Survey
di: Lin, Li, et al.
Pubblicazione: (2024)
di: Lin, Li, et al.
Pubblicazione: (2024)
Real-Time Mobile Video Analytics for Pre-arrival Emergency Medical Services
di: Jin, Liuyi, et al.
Pubblicazione: (2025)
di: Jin, Liuyi, et al.
Pubblicazione: (2025)
Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
di: Lin, Yuxiang, et al.
Pubblicazione: (2025)
di: Lin, Yuxiang, et al.
Pubblicazione: (2025)
A Multi-modal Fusion Network for Terrain Perception Based on Illumination Aware
di: Wang, Rui, et al.
Pubblicazione: (2025)
di: Wang, Rui, et al.
Pubblicazione: (2025)
Wireless Video Semantic Communication with Decoupled Diffusion Multi-frame Compensation
di: Xie, Bingyan, et al.
Pubblicazione: (2025)
di: Xie, Bingyan, et al.
Pubblicazione: (2025)
When Harmful Content Gets Camouflaged: Unveiling Perception Failure of LVLMs with CamHarmTI
di: Li, Yanhui, et al.
Pubblicazione: (2025)
di: Li, Yanhui, et al.
Pubblicazione: (2025)
ALDEN: Reinforcement Learning for Active Navigation and Evidence Gathering in Long Documents
di: Yang, Tianyu, et al.
Pubblicazione: (2025)
di: Yang, Tianyu, et al.
Pubblicazione: (2025)
MaLoRA: Gated Modality LoRA for Key-Space Alignment in Multimodal LLM Fine-Tuning
di: Zheng, Xinhan, et al.
Pubblicazione: (2025)
di: Zheng, Xinhan, et al.
Pubblicazione: (2025)
Attention of a Kiss: Exploring Attention Maps in Video Diffusion for XAIxArts
di: Cole, Adam, et al.
Pubblicazione: (2025)
di: Cole, Adam, et al.
Pubblicazione: (2025)
Semantic Communication-Enabled Cloud-Edge-End-collaborative Metaverse Services Architecure
di: Li, Yuxuan, et al.
Pubblicazione: (2025)
di: Li, Yuxuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Back to Basics: Revisiting ASR in the Age of Voice Agents
di: Tay, Geeyang, et al.
Pubblicazione: (2026) -
The Dream Within Huang Long Cave: AI-Driven Interactive Narrative for Family Storytelling and Emotional Reflection
di: Huang, Jiayang, et al.
Pubblicazione: (2025) -
Human-Data Interaction, Exploration, and Visualization in the AI Era: Challenges and Opportunities
di: Fekete, Jean-Daniel, et al.
Pubblicazione: (2026) -
LLMER: Crafting Interactive Extended Reality Worlds with JSON Data Generated by Large Language Models
di: Chen, Jiangong, et al.
Pubblicazione: (2025) -
MMTB: Evaluating Terminal Agents on Multimedia-File Tasks
di: Heo, Chiyeong, et al.
Pubblicazione: (2026)