Saved in:
| Main Authors: | Taniguchi, Takara, Shimizu, Ryohei, Vo, Duc Minh, Izumi, Kota, Yang, Shiqi, Suzuki, Teppei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.06063 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OnomatoGen: Onomatopoeia Generation with the Alpha-Channel in Manga
by: Taniguchi, Takara, et al.
Published: (2025)
by: Taniguchi, Takara, et al.
Published: (2025)
Learning Gaussian Data Augmentation in Feature Space for One-shot Object Detection in Manga
by: Taniguchi, Takara, et al.
Published: (2024)
by: Taniguchi, Takara, et al.
Published: (2024)
ShieldVLM: Safeguarding the Multimodal Implicit Toxicity via Deliberative Reasoning with LVLMs
by: Cui, Shiyao, et al.
Published: (2025)
by: Cui, Shiyao, et al.
Published: (2025)
QMedShield: A Novel Quantum Chaos-based Image Encryption Scheme for Secure Medical Image Storage in the Cloud
by: Rajan, Arun Amaithi, et al.
Published: (2024)
by: Rajan, Arun Amaithi, et al.
Published: (2024)
Avoiding Quality Saturation in UGC Compression Using Denoised References
by: Xiong, Xin, et al.
Published: (2025)
by: Xiong, Xin, et al.
Published: (2025)
Short-Form Video Viewing Behavior Analysis and Multi-Step Viewing Time Prediction
by: Yen, Vu Thi Hai, et al.
Published: (2026)
by: Yen, Vu Thi Hai, et al.
Published: (2026)
Fact-Checking at Scale: Multimodal AI for Authenticity and Context Verification in Online Media
by: Phan, Van-Hoang, et al.
Published: (2025)
by: Phan, Van-Hoang, et al.
Published: (2025)
EidetiCom: A Cross-modal Brain-Computer Semantic Communication Paradigm for Decoding Visual Perception
by: Zheng, Linfeng, et al.
Published: (2024)
by: Zheng, Linfeng, et al.
Published: (2024)
Detecting Content Rating Violations in Android Applications: A Vision-Language Approach
by: Denipitiyage, D., et al.
Published: (2025)
by: Denipitiyage, D., et al.
Published: (2025)
Modeling the Impacts of Swipe Delay on User Quality of Experience in Short Video Streaming
by: Nguyen, Duc V., et al.
Published: (2026)
by: Nguyen, Duc V., et al.
Published: (2026)
Voxel-GS: Quantized Scaffold Gaussian Splatting Compression with Run-Length Coding
by: Fu, Chunyang, et al.
Published: (2025)
by: Fu, Chunyang, et al.
Published: (2025)
FreeEnhance: Tuning-Free Image Enhancement via Content-Consistent Noising-and-Denoising Process
by: Luo, Yang, et al.
Published: (2024)
by: Luo, Yang, et al.
Published: (2024)
On Parallelism in Music and Language: A Perspective from Symbol Emergence Systems based on Probabilistic Generative Models
by: Taniguchi, Tadahiro
Published: (2025)
by: Taniguchi, Tadahiro
Published: (2025)
A Subjective Quality Evaluation of 3D Mesh with Dynamic Level of Detail in Virtual Reality
by: Nguyen, Duc, et al.
Published: (2024)
by: Nguyen, Duc, et al.
Published: (2024)
Competitive Learning for Achieving Content-specific Filters in Video Coding for Machines
by: Zhang, Honglei, et al.
Published: (2024)
by: Zhang, Honglei, et al.
Published: (2024)
Subjective Quality Assessment of Dynamic 3D Meshes in Virtual Reality Environment
by: Nguyen, Duc V., et al.
Published: (2026)
by: Nguyen, Duc V., et al.
Published: (2026)
TMDC: A Two-Stage Modality Denoising and Complementation Framework for Multimodal Sentiment Analysis with Missing and Noisy Modalities
by: Zhuang, Yan, et al.
Published: (2025)
by: Zhuang, Yan, et al.
Published: (2025)
Improving the Efficiency of VVC using Partitioning of Reference Frames
by: Qureshi, Kamran, et al.
Published: (2025)
by: Qureshi, Kamran, et al.
Published: (2025)
Multi-Reference Generative Face Video Compression with Contrastive Learning
by: Konuko, Goluck, et al.
Published: (2024)
by: Konuko, Goluck, et al.
Published: (2024)
Identity-Driven Multimedia Forgery Detection via Reference Assistance
by: Xu, Junhao, et al.
Published: (2024)
by: Xu, Junhao, et al.
Published: (2024)
ABO: Abandon Bayer Filter for Adaptive Edge Offloading in Responsive Augmented Reality
by: Han, Yongxuan, et al.
Published: (2025)
by: Han, Yongxuan, et al.
Published: (2025)
Ada2I: Enhancing Modality Balance for Multimodal Conversational Emotion Recognition
by: Nguyen, Cam-Van Thi, et al.
Published: (2024)
by: Nguyen, Cam-Van Thi, et al.
Published: (2024)
Scalable On-the-fly Transcoding for Adaptive Streaming of Dynamic Point Clouds
by: Rudolph, Michael, et al.
Published: (2026)
by: Rudolph, Michael, et al.
Published: (2026)
TOP:A New Target-Audience Oriented Content Paraphrase Task
by: Lin, Boda, et al.
Published: (2024)
by: Lin, Boda, et al.
Published: (2024)
Volume Tracking Based Reference Mesh Extraction for Time-Varying Mesh Compression
by: Chen, Guodong, et al.
Published: (2024)
by: Chen, Guodong, et al.
Published: (2024)
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing
by: Wang, Yisu, et al.
Published: (2025)
by: Wang, Yisu, et al.
Published: (2025)
Self-Training Boosted Multi-Factor Matching Network for Composed Image Retrieval
by: Wen, Haokun, et al.
Published: (2023)
by: Wen, Haokun, et al.
Published: (2023)
Content-Adaptive Rate-Quality Curve Prediction Model in Media Processing System
by: Yin, Shibo, et al.
Published: (2024)
by: Yin, Shibo, et al.
Published: (2024)
The Future is Meta: Metadata, Formats and Perspectives towards Interactive and Personalized AV Content
by: Weller, Alexander, et al.
Published: (2024)
by: Weller, Alexander, et al.
Published: (2024)
Fully Automatic Content-Aware Tiling Pipeline for Pathology Whole Slide Images
by: Jabar, Falah, et al.
Published: (2024)
by: Jabar, Falah, et al.
Published: (2024)
MAR3: Multi-Agent Recognition, Reasoning, and Reflection for Reference Audio-Visual Segmentation
by: Zhao, Yuan, et al.
Published: (2026)
by: Zhao, Yuan, et al.
Published: (2026)
Smart Fitting Room: A One-stop Framework for Matching-aware Virtual Try-on
by: Yu, Mingzhe, et al.
Published: (2024)
by: Yu, Mingzhe, et al.
Published: (2024)
A Large-scale Dataset with Behavior, Attributes, and Content of Mobile Short-video Platform
by: Shang, Yu, et al.
Published: (2025)
by: Shang, Yu, et al.
Published: (2025)
3DMambaIPF: A State Space Model for Iterative Point Cloud Filtering via Differentiable Rendering
by: Zhou, Qingyuan, et al.
Published: (2024)
by: Zhou, Qingyuan, et al.
Published: (2024)
Efficient and Accurate Image Provenance Analysis: A Scalable Pipeline for Large-scale Images
by: Lai, Jiewei, et al.
Published: (2025)
by: Lai, Jiewei, et al.
Published: (2025)
The State-of-the-Art in Lifelog Retrieval: A Review of Progress at the ACM Lifelog Search Challenge Workshop 2022-24
by: Tran, Allie, et al.
Published: (2025)
by: Tran, Allie, et al.
Published: (2025)
Think before You Leap: Content-Aware Low-Cost Edge-Assisted Video Semantic Segmentation
by: Yan, Mingxuan, et al.
Published: (2024)
by: Yan, Mingxuan, et al.
Published: (2024)
A Distribution Matching Approach to Neural Piano Transcription with Optimal Transport
by: Wei, Weixing, et al.
Published: (2026)
by: Wei, Weixing, et al.
Published: (2026)
UNQA: Unified No-Reference Quality Assessment for Audio, Image, Video, and Audio-Visual Content
by: Cao, Yuqin, et al.
Published: (2024)
by: Cao, Yuqin, et al.
Published: (2024)
Training-and-Prompt-Free General Painterly Harmonization via Zero-Shot Disentenglement on Style and Content References
by: Hsiao, Teng-Fang, et al.
Published: (2024)
by: Hsiao, Teng-Fang, et al.
Published: (2024)
Similar Items
-
OnomatoGen: Onomatopoeia Generation with the Alpha-Channel in Manga
by: Taniguchi, Takara, et al.
Published: (2025) -
Learning Gaussian Data Augmentation in Feature Space for One-shot Object Detection in Manga
by: Taniguchi, Takara, et al.
Published: (2024) -
ShieldVLM: Safeguarding the Multimodal Implicit Toxicity via Deliberative Reasoning with LVLMs
by: Cui, Shiyao, et al.
Published: (2025) -
QMedShield: A Novel Quantum Chaos-based Image Encryption Scheme for Secure Medical Image Storage in the Cloud
by: Rajan, Arun Amaithi, et al.
Published: (2024) -
Avoiding Quality Saturation in UGC Compression Using Denoised References
by: Xiong, Xin, et al.
Published: (2025)