Guardado en:
| Autores principales: | Lin, Qing, Zhang, Jingfeng, Ong, Yew-Soon, Zhang, Mengmi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2403.08255 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Lightweight and Accurate Multi-View Stereo with Confidence-Aware Diffusion Model
por: Wang, Fangjinhua, et al.
Publicado: (2025)
por: Wang, Fangjinhua, et al.
Publicado: (2025)
Learning to See Through a Baby's Eyes: Early Visual Diets Enable Robust Visual Intelligence in Humans and Machines
por: Cai, Yusen, et al.
Publicado: (2025)
por: Cai, Yusen, et al.
Publicado: (2025)
Fine-grained Abnormality Prompt Learning for Zero-shot Anomaly Detection
por: Zhu, Jiawen, et al.
Publicado: (2024)
por: Zhu, Jiawen, et al.
Publicado: (2024)
EmoEdit: Evoking Emotions through Image Manipulation
por: Yang, Jingyuan, et al.
Publicado: (2024)
por: Yang, Jingyuan, et al.
Publicado: (2024)
Agentic Spatio-Temporal Grounding via Collaborative Reasoning
por: Zhao, Heng, et al.
Publicado: (2026)
por: Zhao, Heng, et al.
Publicado: (2026)
Hierarchically Robust Zero-shot Vision-language Models
por: Dong, Junhao, et al.
Publicado: (2026)
por: Dong, Junhao, et al.
Publicado: (2026)
Prototype Optimization with Neural ODE for Few-Shot Learning
por: Zhang, Baoquan, et al.
Publicado: (2024)
por: Zhang, Baoquan, et al.
Publicado: (2024)
Few-shot NeRF by Adaptive Rendering Loss Regularization
por: Xu, Qingshan, et al.
Publicado: (2024)
por: Xu, Qingshan, et al.
Publicado: (2024)
Deterministic-to-Stochastic Diverse Latent Feature Mapping for Human Motion Synthesis
por: Hua, Yu, et al.
Publicado: (2025)
por: Hua, Yu, et al.
Publicado: (2025)
SiamNAS: Siamese Surrogate Model for Dominance Relation Prediction in Multi-objective Neural Architecture Search
por: Zhou, Yuyang, et al.
Publicado: (2025)
por: Zhou, Yuyang, et al.
Publicado: (2025)
Precise-Physics Driven Text-to-3D Generation
por: Xu, Qingshan, et al.
Publicado: (2024)
por: Xu, Qingshan, et al.
Publicado: (2024)
Learning to Perceive "Where": Spatial Pretext Tasks for Robust Self-Supervised Learning
por: Shen, Yang, et al.
Publicado: (2026)
por: Shen, Yang, et al.
Publicado: (2026)
MosaicFusion: Diffusion Models as Data Augmenters for Large Vocabulary Instance Segmentation
por: Xie, Jiahao, et al.
Publicado: (2023)
por: Xie, Jiahao, et al.
Publicado: (2023)
Enhancing Adversarial Robustness via Uncertainty-Aware Distributional Adversarial Training
por: Dong, Junhao, et al.
Publicado: (2024)
por: Dong, Junhao, et al.
Publicado: (2024)
Video Set Distillation: Information Diversification and Temporal Densification
por: Zhao, Yinjie, et al.
Publicado: (2024)
por: Zhao, Yinjie, et al.
Publicado: (2024)
Pushing Rendering Boundaries: Hard Gaussian Splatting
por: Xu, Qingshan, et al.
Publicado: (2024)
por: Xu, Qingshan, et al.
Publicado: (2024)
Possibilistic Predictive Uncertainty for Deep Learning
por: Ni, Yao, et al.
Publicado: (2026)
por: Ni, Yao, et al.
Publicado: (2026)
Flow Snapshot Neurons in Action: Deep Neural Networks Generalize to Biological Motion Perception
por: Han, Shuangpeng, et al.
Publicado: (2024)
por: Han, Shuangpeng, et al.
Publicado: (2024)
Pose Prior Learner: Unsupervised Categorical Prior Learning for Pose Estimation
por: Wang, Ziyu, et al.
Publicado: (2024)
por: Wang, Ziyu, et al.
Publicado: (2024)
PRISM: Progressive Reasoning through Iterative Slot Memory for Vision
por: Wang, Ziyu, et al.
Publicado: (2026)
por: Wang, Ziyu, et al.
Publicado: (2026)
Adaptive Visual Scene Understanding: Incremental Scene Graph Generation
por: Khandelwal, Naitik, et al.
Publicado: (2023)
por: Khandelwal, Naitik, et al.
Publicado: (2023)
LLM-to-Phy3D: Physically Conform Online 3D Object Generation with LLMs
por: Wong, Melvin, et al.
Publicado: (2025)
por: Wong, Melvin, et al.
Publicado: (2025)
EmoAttack: Emotion-to-Image Diffusion Models for Emotional Backdoor Generation
por: Wei, Tianyu, et al.
Publicado: (2024)
por: Wei, Tianyu, et al.
Publicado: (2024)
Pix2Fact: When Vision Is Not Enough -- Benchmarking Fine-Grained VQA with Web Verification on High-Resolution Real-World Scenes
por: Jiang, Yifan, et al.
Publicado: (2026)
por: Jiang, Yifan, et al.
Publicado: (2026)
Generative AI-based Prompt Evolution Engineering Design Optimization With Vision-Language Model
por: Wong, Melvin, et al.
Publicado: (2024)
por: Wong, Melvin, et al.
Publicado: (2024)
Peering into the Unknown: Active View Selection with Neural Uncertainty Maps for 3D Reconstruction
por: Zhang, Zhengquan, et al.
Publicado: (2025)
por: Zhang, Zhengquan, et al.
Publicado: (2025)
A Plug-and-Play Multi-Criteria Guidance for Diverse In-Betweening Human Motion Generation
por: Yu, Hua, et al.
Publicado: (2025)
por: Yu, Hua, et al.
Publicado: (2025)
Dynamic-Aware Video Distillation: Optimizing Temporal Resolution Based on Video Semantics
por: Zhao, Yinjie, et al.
Publicado: (2025)
por: Zhao, Yinjie, et al.
Publicado: (2025)
MMT-ARD: Multimodal Multi-Teacher Adversarial Distillation for Robust Vision-Language Models
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
Unforgettable Lessons from Forgettable Images: Intra-Class Memorability Matters in Computer Vision
por: Jing, Jie, et al.
Publicado: (2024)
por: Jing, Jie, et al.
Publicado: (2024)
EEmo-Logic: A Unified Dataset and Multi-Stage Framework for Comprehensive Image-Evoked Emotion Assessment
por: Gao, Lancheng, et al.
Publicado: (2026)
por: Gao, Lancheng, et al.
Publicado: (2026)
Hard-Label Black-Box Attacks on 3D Point Clouds
por: Liu, Daizong, et al.
Publicado: (2024)
por: Liu, Daizong, et al.
Publicado: (2024)
EmoDiffusion: Enhancing Emotional 3D Facial Animation with Latent Diffusion Models
por: Zhang, Yixuan, et al.
Publicado: (2025)
por: Zhang, Yixuan, et al.
Publicado: (2025)
NeuSpring: Neural Spring Fields for Reconstruction and Simulation of Deformable Objects from Videos
por: Xu, Qingshan, et al.
Publicado: (2025)
por: Xu, Qingshan, et al.
Publicado: (2025)
ResDiT: Evoking the Intrinsic Resolution Scalability in Diffusion Transformers
por: Ma, Yiyang, et al.
Publicado: (2025)
por: Ma, Yiyang, et al.
Publicado: (2025)
LLM2TEA: An Agentic AI Designer for Discovery with Generative Evolutionary Multitasking
por: Wong, Melvin, et al.
Publicado: (2024)
por: Wong, Melvin, et al.
Publicado: (2024)
Seeing Through Uncertainty: A Free-Energy Approach for Real-Time Perceptual Adaptation in Robust Visual Navigation
por: Piriyajitakonkij, Maytus, et al.
Publicado: (2024)
por: Piriyajitakonkij, Maytus, et al.
Publicado: (2024)
Preserving Image Properties Through Initializations in Diffusion Models
por: Zhang, Jeffrey, et al.
Publicado: (2024)
por: Zhang, Jeffrey, et al.
Publicado: (2024)
Show Me: Unifying Instructional Image and Video Generation with Diffusion Models
por: Pu, Yujiang, et al.
Publicado: (2025)
por: Pu, Yujiang, et al.
Publicado: (2025)
Unveiling the Tapestry: the Interplay of Generalization and Forgetting in Continual Learning
por: Shi, Zenglin, et al.
Publicado: (2022)
por: Shi, Zenglin, et al.
Publicado: (2022)
Ejemplares similares
-
Lightweight and Accurate Multi-View Stereo with Confidence-Aware Diffusion Model
por: Wang, Fangjinhua, et al.
Publicado: (2025) -
Learning to See Through a Baby's Eyes: Early Visual Diets Enable Robust Visual Intelligence in Humans and Machines
por: Cai, Yusen, et al.
Publicado: (2025) -
Fine-grained Abnormality Prompt Learning for Zero-shot Anomaly Detection
por: Zhu, Jiawen, et al.
Publicado: (2024) -
EmoEdit: Evoking Emotions through Image Manipulation
por: Yang, Jingyuan, et al.
Publicado: (2024) -
Agentic Spatio-Temporal Grounding via Collaborative Reasoning
por: Zhao, Heng, et al.
Publicado: (2026)