Learning an Ensemble Token from Task-driven Priors in Facial Analysis
Fuente:
arXiv
Guardado en:
| Autores principales: | Seo, Sunyong, Kim, Semin, Lee, Jongha |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Data Augmentation For Small Object using Fast AutoAugment
por: Yoon, DaeEun, et al.
Publicado: (2025)
por: Yoon, DaeEun, et al.
Publicado: (2025)
Color Universal Design Neural Network for the Color Vision Deficiencies
por: Seo, Sunyong, et al.
Publicado: (2025)
por: Seo, Sunyong, et al.
Publicado: (2025)
TabFlash: Efficient Table Understanding with Progressive Question Conditioning and Token Focusing
por: Kim, Jongha, et al.
Publicado: (2025)
por: Kim, Jongha, et al.
Publicado: (2025)
Exploiting Diffusion Prior for Task-driven Image Restoration
por: Kim, Jaeha, et al.
Publicado: (2025)
por: Kim, Jaeha, et al.
Publicado: (2025)
Full-scale Representation Guided Network for Retinal Vessel Segmentation
por: Seo, Sunyong, et al.
Publicado: (2025)
por: Seo, Sunyong, et al.
Publicado: (2025)
VidChain: Chain-of-Tasks with Metric-based Direct Preference Optimization for Dense Video Captioning
por: Lee, Ji Soo, et al.
Publicado: (2025)
por: Lee, Ji Soo, et al.
Publicado: (2025)
DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning
por: Choi, Joonmyung, et al.
Publicado: (2026)
por: Choi, Joonmyung, et al.
Publicado: (2026)
NeRFFaceSpeech: One-shot Audio-driven 3D Talking Head Synthesis via Generative Prior
por: Kim, Gihoon, et al.
Publicado: (2024)
por: Kim, Gihoon, et al.
Publicado: (2024)
Implementation of a Skin Lesion Detection System for Managing Children with Atopic Dermatitis Based on Ensemble Learning
por: Jeon, Soobin, et al.
Publicado: (2025)
por: Jeon, Soobin, et al.
Publicado: (2025)
Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
por: Kim, Jongha, et al.
Publicado: (2024)
por: Kim, Jongha, et al.
Publicado: (2024)
Bridging the gap to real-world language-grounded visual concept learning
por: Jung, Whie, et al.
Publicado: (2025)
por: Jung, Whie, et al.
Publicado: (2025)
Learning a Delighting Prior for Facial Appearance Capture in the Wild
por: Han, Yuxuan, et al.
Publicado: (2026)
por: Han, Yuxuan, et al.
Publicado: (2026)
Relevance-aware Multi-context Contrastive Decoding for Retrieval-augmented Visual Question Answering
por: Kim, Jongha, et al.
Publicado: (2026)
por: Kim, Jongha, et al.
Publicado: (2026)
Discrete Facial Encoding: : A Framework for Data-driven Facial Display Discovery
por: Tran, Minh, et al.
Publicado: (2025)
por: Tran, Minh, et al.
Publicado: (2025)
RealTalk: Real-time and Realistic Audio-driven Face Generation with 3D Facial Prior-guided Identity Alignment Network
por: Ji, Xiaozhong, et al.
Publicado: (2024)
por: Ji, Xiaozhong, et al.
Publicado: (2024)
Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model
por: Kim, Dongwon, et al.
Publicado: (2026)
por: Kim, Dongwon, et al.
Publicado: (2026)
Chameleon: A Data-Efficient Generalist for Dense Visual Prediction in the Wild
por: Kim, Donggyun, et al.
Publicado: (2024)
por: Kim, Donggyun, et al.
Publicado: (2024)
Multi-Task Multi-Modal Self-Supervised Learning for Facial Expression Recognition
por: Halawa, Marah, et al.
Publicado: (2024)
por: Halawa, Marah, et al.
Publicado: (2024)
4D Facial Expression Diffusion Model
por: Zou, Kaifeng, et al.
Publicado: (2023)
por: Zou, Kaifeng, et al.
Publicado: (2023)
Expressive Speech-driven Facial Animation with controllable emotions
por: Chen, Yutong, et al.
Publicado: (2023)
por: Chen, Yutong, et al.
Publicado: (2023)
VQTalker: Towards Multilingual Talking Avatars through Facial Motion Tokenization
por: Liu, Tao, et al.
Publicado: (2024)
por: Liu, Tao, et al.
Publicado: (2024)
ERASE: Eliminating Redundant Visual Tokens via Adaptive Two-Stage Token Pruning
por: Lee, Yuna, et al.
Publicado: (2026)
por: Lee, Yuna, et al.
Publicado: (2026)
TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens
por: Zhao, Qingcheng, et al.
Publicado: (2026)
por: Zhao, Qingcheng, et al.
Publicado: (2026)
FRIDAY: Mitigating Unintentional Facial Identity in Deepfake Detectors Guided by Facial Recognizers
por: Kim, Younhun, et al.
Publicado: (2024)
por: Kim, Younhun, et al.
Publicado: (2024)
Analysis of Bias in Deep Learning Facial Beauty Regressors
por: Hamel, Chandon, et al.
Publicado: (2025)
por: Hamel, Chandon, et al.
Publicado: (2025)
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens
por: Lew, Jaihyun, et al.
Publicado: (2024)
por: Lew, Jaihyun, et al.
Publicado: (2024)
Navigating Label Ambiguity for Facial Expression Recognition in the Wild
por: Lee, JunGyu, et al.
Publicado: (2025)
por: Lee, JunGyu, et al.
Publicado: (2025)
Prior-based Objective Inference Mining Potential Uncertainty for Facial Expression Recognition
por: Liu, Hanwei, et al.
Publicado: (2024)
por: Liu, Hanwei, et al.
Publicado: (2024)
Polyglot: Multilingual Style Preserving Speech-Driven Facial Animation
por: Nocentini, Federico, et al.
Publicado: (2026)
por: Nocentini, Federico, et al.
Publicado: (2026)
Deep Learning Based Facial Retargeting Using Local Patches
por: Choi, Yeonsoo, et al.
Publicado: (2026)
por: Choi, Yeonsoo, et al.
Publicado: (2026)
Leveraging 3D Geometric Priors in 2D Rotation Symmetry Detection
por: Seo, Ahyun, et al.
Publicado: (2025)
por: Seo, Ahyun, et al.
Publicado: (2025)
Facial Appearance Capture at Home with Patch-Level Reflectance Prior
por: Han, Yuxuan, et al.
Publicado: (2025)
por: Han, Yuxuan, et al.
Publicado: (2025)
PropFly: Learning to Propagate via On-the-Fly Supervision from Pre-trained Video Diffusion Models
por: Seo, Wonyong, et al.
Publicado: (2026)
por: Seo, Wonyong, et al.
Publicado: (2026)
Masked Autoregressive Model for Weather Forecasting
por: Kim, Doyi, et al.
Publicado: (2024)
por: Kim, Doyi, et al.
Publicado: (2024)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
por: Chung, Jiwan, et al.
Publicado: (2025)
por: Chung, Jiwan, et al.
Publicado: (2025)
Multimodal Representation Learning Techniques for Comprehensive Facial State Analysis
por: Zheng, Kaiwen, et al.
Publicado: (2025)
por: Zheng, Kaiwen, et al.
Publicado: (2025)
RA-Touch: Retrieval-Augmented Touch Understanding with Enriched Visual Data
por: Cho, Yoorhim, et al.
Publicado: (2025)
por: Cho, Yoorhim, et al.
Publicado: (2025)
SynergyNet: Fusing Generative Priors and State-Space Models for Facial Beauty Prediction
por: Boukhari, Djamel Eddine
Publicado: (2025)
por: Boukhari, Djamel Eddine
Publicado: (2025)
Facial-R1: Aligning Reasoning and Recognition for Facial Emotion Analysis
por: Wu, Jiulong, et al.
Publicado: (2025)
por: Wu, Jiulong, et al.
Publicado: (2025)
LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior
por: Wang, Hanyu, et al.
Publicado: (2024)
por: Wang, Hanyu, et al.
Publicado: (2024)
Ejemplares similares
-
Data Augmentation For Small Object using Fast AutoAugment
por: Yoon, DaeEun, et al.
Publicado: (2025) -
Color Universal Design Neural Network for the Color Vision Deficiencies
por: Seo, Sunyong, et al.
Publicado: (2025) -
TabFlash: Efficient Table Understanding with Progressive Question Conditioning and Token Focusing
por: Kim, Jongha, et al.
Publicado: (2025) -
Exploiting Diffusion Prior for Task-driven Image Restoration
por: Kim, Jaeha, et al.
Publicado: (2025) -
Full-scale Representation Guided Network for Retinal Vessel Segmentation
por: Seo, Sunyong, et al.
Publicado: (2025)