Salvato in:
| Autori principali: | Zhang, Liyun, Liu, Fengkai, Sha, Xuanmeng, Wang, Bowen, Liu, Hong, Lian, Zheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2508.10393 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
QuMAB: Query-based Multi-Annotator Behavior Modeling with Reliability under Sparse Labels
di: Zhang, Liyun, et al.
Pubblicazione: (2025)
di: Zhang, Liyun, et al.
Pubblicazione: (2025)
AcoustEmo: Open-Vocabulary Emotion Reasoning via Utterance-Aware Acoustic Q-Former
di: Zhang, Liyun, et al.
Pubblicazione: (2026)
di: Zhang, Liyun, et al.
Pubblicazione: (2026)
SimLabel: Similarity-Weighted Iterative Framework for Multi-annotator Learning with Missing Annotations
di: Zhang, Liyun, et al.
Pubblicazione: (2025)
di: Zhang, Liyun, et al.
Pubblicazione: (2025)
3DFacePolicy: Audio-Driven 3D Facial Animation Based on Action Control
di: Sha, Xuanmeng, et al.
Pubblicazione: (2024)
di: Sha, Xuanmeng, et al.
Pubblicazione: (2024)
QuMATL: Query-based Multi-annotator Tendency Learning
di: Zhang, Liyun, et al.
Pubblicazione: (2025)
di: Zhang, Liyun, et al.
Pubblicazione: (2025)
3DGesPolicy: Phoneme-Aware Holistic Co-Speech Gesture Generation Based on Action Control
di: Sha, Xuanmeng, et al.
Pubblicazione: (2026)
di: Sha, Xuanmeng, et al.
Pubblicazione: (2026)
Harmony: A Unified Framework for Modality Incremental Learning
di: Song, Yaguang, et al.
Pubblicazione: (2025)
di: Song, Yaguang, et al.
Pubblicazione: (2025)
LayerT2V: A Unified Multi-Layer Video Generation Framework
di: Li, Guangzhao, et al.
Pubblicazione: (2025)
di: Li, Guangzhao, et al.
Pubblicazione: (2025)
HeLo: Heterogeneous Multi-Modal Fusion with Label Correlation for Emotion Distribution Learning
di: Zheng, Chuhang, et al.
Pubblicazione: (2025)
di: Zheng, Chuhang, et al.
Pubblicazione: (2025)
GTA-HDR: A Large-Scale Synthetic Dataset for HDR Image Reconstruction
di: Barua, Hrishav Bakul, et al.
Pubblicazione: (2024)
di: Barua, Hrishav Bakul, et al.
Pubblicazione: (2024)
SABR: A Stable Adaptive Bitrate Framework Using Behavior Cloning Pretraining and Reinforcement Learning Fine-Tuning
di: Luo, Pengcheng, et al.
Pubblicazione: (2025)
di: Luo, Pengcheng, et al.
Pubblicazione: (2025)
PhysHDR: When Lighting Meets Materials and Scene Geometry in HDR Reconstruction
di: Barua, Hrishav Bakul, et al.
Pubblicazione: (2025)
di: Barua, Hrishav Bakul, et al.
Pubblicazione: (2025)
HistoHDR-Net: Histogram Equalization for Single LDR to HDR Image Translation
di: Barua, Hrishav Bakul, et al.
Pubblicazione: (2024)
di: Barua, Hrishav Bakul, et al.
Pubblicazione: (2024)
FISHER: A Foundation Model for Multi-Modal Industrial Signal Comprehensive Representation
di: Fan, Pingyi, et al.
Pubblicazione: (2025)
di: Fan, Pingyi, et al.
Pubblicazione: (2025)
MixEval-X: Any-to-Any Evaluations from Real-World Data Mixtures
di: Ni, Jinjie, et al.
Pubblicazione: (2024)
di: Ni, Jinjie, et al.
Pubblicazione: (2024)
Sample then Identify: A General Framework for Risk Control and Assessment in Multimodal Large Language Models
di: Wang, Qingni, et al.
Pubblicazione: (2024)
di: Wang, Qingni, et al.
Pubblicazione: (2024)
Arondight: Red Teaming Large Vision Language Models with Auto-generated Multi-modal Jailbreak Prompts
di: Liu, Yi, et al.
Pubblicazione: (2024)
di: Liu, Yi, et al.
Pubblicazione: (2024)
Cross-Space Synergy: A Unified Framework for Multimodal Emotion Recognition in Conversation
di: Lyu, Xiaosen, et al.
Pubblicazione: (2025)
di: Lyu, Xiaosen, et al.
Pubblicazione: (2025)
Detecting Multimedia Generated by Large AI Models: A Survey
di: Lin, Li, et al.
Pubblicazione: (2024)
di: Lin, Li, et al.
Pubblicazione: (2024)
Continuity, Piecewise Corrections, and Functor Models in Function-Based Learning
di: Harby, John
Pubblicazione: (2026)
di: Harby, John
Pubblicazione: (2026)
MultiMedEdit: A Scenario-Aware Benchmark for Evaluating Knowledge Editing in Medical VQA
di: Wen, Shengtao, et al.
Pubblicazione: (2025)
di: Wen, Shengtao, et al.
Pubblicazione: (2025)
AI-Integrated Decision Support System for Real-Time Market Growth Forecasting and Multi-Source Content Diffusion Analytics
di: Yin, Ziqing, et al.
Pubblicazione: (2025)
di: Yin, Ziqing, et al.
Pubblicazione: (2025)
OmniMER: Auxiliary-Enhanced LLM Adaptation for Indonesian Multimodal Emotion Recognition
di: Yan, Xueming, et al.
Pubblicazione: (2025)
di: Yan, Xueming, et al.
Pubblicazione: (2025)
Unveiling Covert Toxicity in Multimodal Data via Toxicity Association Graphs: A Graph-Based Metric and Interpretable Detection Framework
di: Wu, Guanzong, et al.
Pubblicazione: (2026)
di: Wu, Guanzong, et al.
Pubblicazione: (2026)
Seeing in the Dark: A Teacher-Student Framework for Dark Video Action Recognition via Knowledge Distillation and Contrastive Learning
di: Dass, Sharana Dharshikgan Suresh, et al.
Pubblicazione: (2025)
di: Dass, Sharana Dharshikgan Suresh, et al.
Pubblicazione: (2025)
CFAE Framework Specification — ABO (Archetypal Behavior Ontology): Predictive Priors & Active Inference Seeding Layer - version: v0.1
di: Brown, Cameron
Pubblicazione: (2026)
di: Brown, Cameron
Pubblicazione: (2026)
Time-RA: Towards Time Series Reasoning for Anomaly Diagnosis with LLM Feedback
di: Yang, Yiyuan, et al.
Pubblicazione: (2025)
di: Yang, Yiyuan, et al.
Pubblicazione: (2025)
A review on Machine Learning based User-Centric Multimedia Streaming Techniques
di: Ghosh, Monalisa, et al.
Pubblicazione: (2024)
di: Ghosh, Monalisa, et al.
Pubblicazione: (2024)
Decoding the Hook: A Multimodal LLM Framework for Analyzing the Hooking Period of Video Ads
di: Zhang, Kunpeng, et al.
Pubblicazione: (2026)
di: Zhang, Kunpeng, et al.
Pubblicazione: (2026)
Identifying Multi-modal Knowledge Neurons in Pretrained Transformers via Two-stage Filtering
di: Sato, Yugen, et al.
Pubblicazione: (2025)
di: Sato, Yugen, et al.
Pubblicazione: (2025)
The Path to Open Innovation: Peer-Review Under Fire (PRUF)
di: Billions, Ava, et al.
Pubblicazione: (2025)
di: Billions, Ava, et al.
Pubblicazione: (2025)
Uncertainty-Penalized Direct Preference Optimization
di: Houliston, Sam, et al.
Pubblicazione: (2024)
di: Houliston, Sam, et al.
Pubblicazione: (2024)
Reliability Quantification of Deep Reinforcement Learning-based Control
di: Yoshioka, Hitoshi, et al.
Pubblicazione: (2023)
di: Yoshioka, Hitoshi, et al.
Pubblicazione: (2023)
EquiAV: Leveraging Equivariance for Audio-Visual Contrastive Learning
di: Kim, Jongsuk, et al.
Pubblicazione: (2024)
di: Kim, Jongsuk, et al.
Pubblicazione: (2024)
Enhancing Modality Representation and Alignment for Multimodal Cold-start Active Learning
di: Shen, Meng, et al.
Pubblicazione: (2024)
di: Shen, Meng, et al.
Pubblicazione: (2024)
Rethinking Prompting Strategies for Multi-Label Recognition with Partial Annotations
di: Rawlekar, Samyak, et al.
Pubblicazione: (2024)
di: Rawlekar, Samyak, et al.
Pubblicazione: (2024)
MotionCtrl: A Unified and Flexible Motion Controller for Video Generation
di: Wang, Zhouxia, et al.
Pubblicazione: (2023)
di: Wang, Zhouxia, et al.
Pubblicazione: (2023)
Unified Hallucination Detection for Multimodal Large Language Models
di: Chen, Xiang, et al.
Pubblicazione: (2024)
di: Chen, Xiang, et al.
Pubblicazione: (2024)
Doctor Sun: A Bilingual Multimodal Large Language Model for Biomedical AI
di: Xue, Dong, et al.
Pubblicazione: (2025)
di: Xue, Dong, et al.
Pubblicazione: (2025)
Adaptive Social Metaverse Streaming based on Federated Multi-Agent Deep Reinforcement Learning
di: Long, Zijian, et al.
Pubblicazione: (2025)
di: Long, Zijian, et al.
Pubblicazione: (2025)
Documenti analoghi
-
QuMAB: Query-based Multi-Annotator Behavior Modeling with Reliability under Sparse Labels
di: Zhang, Liyun, et al.
Pubblicazione: (2025) -
AcoustEmo: Open-Vocabulary Emotion Reasoning via Utterance-Aware Acoustic Q-Former
di: Zhang, Liyun, et al.
Pubblicazione: (2026) -
SimLabel: Similarity-Weighted Iterative Framework for Multi-annotator Learning with Missing Annotations
di: Zhang, Liyun, et al.
Pubblicazione: (2025) -
3DFacePolicy: Audio-Driven 3D Facial Animation Based on Action Control
di: Sha, Xuanmeng, et al.
Pubblicazione: (2024) -
QuMATL: Query-based Multi-annotator Tendency Learning
di: Zhang, Liyun, et al.
Pubblicazione: (2025)