Salvato in:
| Autori principali: | Chen, Sichen, Zhang, Yingyi, Huang, Siming, Yi, Ran, Fan, Ke, Zhang, Ruixin, Chen, Peixian, Wang, Jun, Ding, Shouhong, Ma, Lizhuang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2404.03518 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SDPose: Exploiting Diffusion Priors for Out-of-Domain and Robust Pose Estimation
di: Liang, Shuang, et al.
Pubblicazione: (2025)
di: Liang, Shuang, et al.
Pubblicazione: (2025)
Test-Time Domain Generalization for Face Anti-Spoofing
di: Zhou, Qianyu, et al.
Pubblicazione: (2024)
di: Zhou, Qianyu, et al.
Pubblicazione: (2024)
Switchable Token-Specific Codebook Quantization For Face Image Compression
di: Wang, Yongbo, et al.
Pubblicazione: (2025)
di: Wang, Yongbo, et al.
Pubblicazione: (2025)
PoseAnything: Universal Pose-guided Video Generation with Part-aware Temporal Coherence
di: Wang, Ruiyan, et al.
Pubblicazione: (2025)
di: Wang, Ruiyan, et al.
Pubblicazione: (2025)
From Enhancement to Understanding: Build a Generalized Bridge for Low-light Vision via Semantically Consistent Unsupervised Fine-tuning
di: Wang, Sen, et al.
Pubblicazione: (2025)
di: Wang, Sen, et al.
Pubblicazione: (2025)
IAR2: Improving Autoregressive Visual Generation with Semantic-Detail Associated Token Prediction
di: Yi, Ran, et al.
Pubblicazione: (2025)
di: Yi, Ran, et al.
Pubblicazione: (2025)
Complementarity-Supervised Spectral-Band Routing for Multimodal Emotion Recognition
di: Huang, Zhexian, et al.
Pubblicazione: (2026)
di: Huang, Zhexian, et al.
Pubblicazione: (2026)
Reconstructing Topology-Consistent Face Mesh by Volume Rendering from Multi-View Images
di: Wang, Yating, et al.
Pubblicazione: (2024)
di: Wang, Yating, et al.
Pubblicazione: (2024)
Streaming Looking Ahead with Token-level Self-reward
di: Zhang, Hongming, et al.
Pubblicazione: (2025)
di: Zhang, Hongming, et al.
Pubblicazione: (2025)
AdR-Gaussian: Accelerating Gaussian Splatting with Adaptive Radius
di: Wang, Xinzhe, et al.
Pubblicazione: (2024)
di: Wang, Xinzhe, et al.
Pubblicazione: (2024)
SuperMat: Physically Consistent PBR Material Estimation at Interactive Rates
di: Hong, Yijia, et al.
Pubblicazione: (2024)
di: Hong, Yijia, et al.
Pubblicazione: (2024)
Improving Autoregressive Visual Generation with Cluster-Oriented Token Prediction
di: Hu, Teng, et al.
Pubblicazione: (2025)
di: Hu, Teng, et al.
Pubblicazione: (2025)
InstanceV: Instance-Level Video Generation
di: Chen, Yuheng, et al.
Pubblicazione: (2025)
di: Chen, Yuheng, et al.
Pubblicazione: (2025)
D2Pruner: Debiased Importance and Structural Diversity for MLLM Token Pruning
di: Zhang, Evelyn, et al.
Pubblicazione: (2025)
di: Zhang, Evelyn, et al.
Pubblicazione: (2025)
Causality-aware Graph Aggregation Weight Estimator for Popularity Debiasing in Top-K Recommendation
di: Que, Yue, et al.
Pubblicazione: (2025)
di: Que, Yue, et al.
Pubblicazione: (2025)
SAFE: Slow and Fast Parameter-Efficient Tuning for Continual Learning with Pre-Trained Models
di: Zhao, Linglan, et al.
Pubblicazione: (2024)
di: Zhao, Linglan, et al.
Pubblicazione: (2024)
Cross-Domain Knowledge Distillation for Low-Resolution Human Pose Estimation
di: Gu, Zejun, et al.
Pubblicazione: (2024)
di: Gu, Zejun, et al.
Pubblicazione: (2024)
Fuse Before Transfer: Knowledge Fusion for Heterogeneous Distillation
di: Li, Guopeng, et al.
Pubblicazione: (2024)
di: Li, Guopeng, et al.
Pubblicazione: (2024)
SaRA: High-Efficient Diffusion Model Fine-tuning with Progressive Sparse Low-Rank Adaptation
di: Hu, Teng, et al.
Pubblicazione: (2024)
di: Hu, Teng, et al.
Pubblicazione: (2024)
Towards Rationale-Answer Alignment of LVLMs via Self-Rationale Calibration
di: Wu, Yuanchen, et al.
Pubblicazione: (2025)
di: Wu, Yuanchen, et al.
Pubblicazione: (2025)
GloTok: Global Perspective Tokenizer for Image Reconstruction and Generation
di: Zhao, Xuan, et al.
Pubblicazione: (2025)
di: Zhao, Xuan, et al.
Pubblicazione: (2025)
DiffusionFake: Enhancing Generalization in Deepfake Detection via Guided Stable Diffusion
di: Sun, Ke, et al.
Pubblicazione: (2024)
di: Sun, Ke, et al.
Pubblicazione: (2024)
Textual Decomposition Then Sub-motion-space Scattering for Open-Vocabulary Motion Generation
di: Fan, Ke, et al.
Pubblicazione: (2024)
di: Fan, Ke, et al.
Pubblicazione: (2024)
Collaborative Face Experts Fusion in Video Generation: Boosting Identity Consistency Across Large Face Poses
di: Wang, Yuji, et al.
Pubblicazione: (2025)
di: Wang, Yuji, et al.
Pubblicazione: (2025)
Navigating the Emotion Tree: Hierarchical Hyperbolic RAG for Multimodal Emotion Recognition
di: Wang, Zeheng, et al.
Pubblicazione: (2026)
di: Wang, Zeheng, et al.
Pubblicazione: (2026)
PCIE_Pose Solution for EgoExo4D Pose and Proficiency Estimation Challenge
di: Chen, Feng, et al.
Pubblicazione: (2025)
di: Chen, Feng, et al.
Pubblicazione: (2025)
Fixed Anchors Are Not Enough: Dynamic Retrieval and Persistent Homology for Dataset Distillation
di: Li, Muquan, et al.
Pubblicazione: (2026)
di: Li, Muquan, et al.
Pubblicazione: (2026)
Long-VITA: Scaling Large Multi-modal Models to 1 Million Tokens with Leading Short-Context Accuracy
di: Shen, Yunhang, et al.
Pubblicazione: (2025)
di: Shen, Yunhang, et al.
Pubblicazione: (2025)
VISA: Group-wise Visual Token Selection and Aggregation via Graph Summarization for Efficient MLLMs Inference
di: Jiang, Pengfei, et al.
Pubblicazione: (2025)
di: Jiang, Pengfei, et al.
Pubblicazione: (2025)
EyeSeg: An Uncertainty-Aware Eye Segmentation Framework for AR/VR
di: Peng, Zhengyuan, et al.
Pubblicazione: (2025)
di: Peng, Zhengyuan, et al.
Pubblicazione: (2025)
SCJD: Sparse Correlation and Joint Distillation for Efficient 3D Human Pose Estimation
di: Chen, Weihong, et al.
Pubblicazione: (2025)
di: Chen, Weihong, et al.
Pubblicazione: (2025)
CleanPose: Category-Level Object Pose Estimation via Causal Learning and Knowledge Distillation
di: Lin, Xiao, et al.
Pubblicazione: (2025)
di: Lin, Xiao, et al.
Pubblicazione: (2025)
LaRE$^2$: Latent Reconstruction Error Based Method for Diffusion-Generated Image Detection
di: Luo, Yunpeng, et al.
Pubblicazione: (2024)
di: Luo, Yunpeng, et al.
Pubblicazione: (2024)
SQL-R1: Training Natural Language to SQL Reasoning Model By Reinforcement Learning
di: Ma, Peixian, et al.
Pubblicazione: (2025)
di: Ma, Peixian, et al.
Pubblicazione: (2025)
FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis
di: Fan, Ke, et al.
Pubblicazione: (2024)
di: Fan, Ke, et al.
Pubblicazione: (2024)
The Fruits of Opportunism: Noncompliance and the Evolution of China's Supplemental Education Industry by Le Lin, Chicago, IL, The University of Chicago Press, 2022, 244 pp.
di: Yingyi Ma
Pubblicazione: (2024)
di: Yingyi Ma
Pubblicazione: (2024)
Mind Your Margin and Boundary: Are Your Distilled Datasets Truly Robust?
di: Li, Muquan, et al.
Pubblicazione: (2026)
di: Li, Muquan, et al.
Pubblicazione: (2026)
Improved AdaBoost for Virtual Reality Experience Prediction Based on Long Short-Term Memory Network
di: Fan, Wenhan, et al.
Pubblicazione: (2024)
di: Fan, Wenhan, et al.
Pubblicazione: (2024)
One-for-More: Continual Diffusion Model for Anomaly Detection
di: Li, Xiaofan, et al.
Pubblicazione: (2025)
di: Li, Xiaofan, et al.
Pubblicazione: (2025)
MV-Adapter: Multi-view Consistent Image Generation Made Easy
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SDPose: Exploiting Diffusion Priors for Out-of-Domain and Robust Pose Estimation
di: Liang, Shuang, et al.
Pubblicazione: (2025) -
Test-Time Domain Generalization for Face Anti-Spoofing
di: Zhou, Qianyu, et al.
Pubblicazione: (2024) -
Switchable Token-Specific Codebook Quantization For Face Image Compression
di: Wang, Yongbo, et al.
Pubblicazione: (2025) -
PoseAnything: Universal Pose-guided Video Generation with Part-aware Temporal Coherence
di: Wang, Ruiyan, et al.
Pubblicazione: (2025) -
From Enhancement to Understanding: Build a Generalized Bridge for Low-light Vision via Semantically Consistent Unsupervised Fine-tuning
di: Wang, Sen, et al.
Pubblicazione: (2025)