DNA Family: Boosting Weight-Sharing NAS with Block-Wise Supervisions
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Guangrun, Li, Changlin, Yuan, Liuchun, Peng, Jiefeng, Xian, Xiaoyu, Liang, Xiaodan, Chang, Xiaojun, Lin, Liang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Efficient Training of Large Vision Models via Advanced Automated Progressive Learning
por: Li, Changlin, et al.
Publicado: (2024)
por: Li, Changlin, et al.
Publicado: (2024)
ACD: Direct Conditional Control for Video Diffusion Models via Attention Supervision
por: Li, Weiqi, et al.
Publicado: (2025)
por: Li, Weiqi, et al.
Publicado: (2025)
MixReorg: Cross-Modal Mixed Patch Reorganization is a Good Mask Learner for Open-World Semantic Segmentation
por: Cai, Kaixin, et al.
Publicado: (2023)
por: Cai, Kaixin, et al.
Publicado: (2023)
NeRF-VPT: Learning Novel View Representations with Neural Radiance Fields via View Prompt Tuning
por: Chen, Linsheng, et al.
Publicado: (2024)
por: Chen, Linsheng, et al.
Publicado: (2024)
AlignMiF: Geometry-Aligned Multimodal Implicit Field for LiDAR-Camera Joint Synthesis
por: Tang, Tao, et al.
Publicado: (2024)
por: Tang, Tao, et al.
Publicado: (2024)
SWAP-NAS: Sample-Wise Activation Patterns for Ultra-fast NAS
por: Peng, Yameng, et al.
Publicado: (2024)
por: Peng, Yameng, et al.
Publicado: (2024)
Contrastive Learning with Counterfactual Explanations for Radiology Report Generation
por: Li, Mingjie, et al.
Publicado: (2024)
por: Li, Mingjie, et al.
Publicado: (2024)
GS: Generative Segmentation via Label Diffusion
por: Chen, Yuhao, et al.
Publicado: (2025)
por: Chen, Yuhao, et al.
Publicado: (2025)
HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics
por: Li, Weiqi, et al.
Publicado: (2025)
por: Li, Weiqi, et al.
Publicado: (2025)
Making Large Language Models Better Planners with Reasoning-Decision Alignment
por: Huang, Zhijian, et al.
Publicado: (2024)
por: Huang, Zhijian, et al.
Publicado: (2024)
Human-Centric Open-Future Task Discovery: Formulation, Benchmark, and Scalable Tree-Based Search
por: Song, Zijian, et al.
Publicado: (2025)
por: Song, Zijian, et al.
Publicado: (2025)
RealignDiff: Boosting Text-to-Image Diffusion Model with Coarse-to-fine Semantic Re-alignment
por: Jiang, Zutao, et al.
Publicado: (2023)
por: Jiang, Zutao, et al.
Publicado: (2023)
MLP Can Be A Good Transformer Learner
por: Lin, Sihao, et al.
Publicado: (2024)
por: Lin, Sihao, et al.
Publicado: (2024)
Knowledge Distillation via the Target-aware Transformer
por: Lin, Sihao, et al.
Publicado: (2022)
por: Lin, Sihao, et al.
Publicado: (2022)
Physical Autoregressive Model for Robotic Manipulation without Action Pretraining
por: Song, Zijian, et al.
Publicado: (2025)
por: Song, Zijian, et al.
Publicado: (2025)
One Model for All: Unified Try-On and Try-Off in Any Pose via LLM-Inspired Bidirectional Tweedie Diffusion
por: Liu, Jinxi, et al.
Publicado: (2025)
por: Liu, Jinxi, et al.
Publicado: (2025)
Ground-R1: Incentivizing Grounded Visual Reasoning via Reinforcement Learning
por: Cao, Meng, et al.
Publicado: (2025)
por: Cao, Meng, et al.
Publicado: (2025)
Sitcom-Crafter: A Plot-Driven Human Motion Generation System in 3D Scenes
por: Chen, Jianqi, et al.
Publicado: (2024)
por: Chen, Jianqi, et al.
Publicado: (2024)
Predicting Genetic Mutation from Whole Slide Images via Biomedical-Linguistic Knowledge Enhanced Multi-label Classification
por: Huang, Gexin, et al.
Publicado: (2024)
por: Huang, Gexin, et al.
Publicado: (2024)
Learning Spatial-Temporal Coherent Correlations for Speech-Preserving Facial Expression Manipulation
por: Chen, Tianshui, et al.
Publicado: (2026)
por: Chen, Tianshui, et al.
Publicado: (2026)
SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks
por: Song, Zijian, et al.
Publicado: (2025)
por: Song, Zijian, et al.
Publicado: (2025)
ShapeBoost: Boosting Human Shape Estimation with Part-Based Parameterization and Clothing-Preserving Augmentation
por: Bian, Siyuan, et al.
Publicado: (2024)
por: Bian, Siyuan, et al.
Publicado: (2024)
Multi-View People Detection in Large Scenes via Supervised View-Wise Contribution Weighting
por: Zhang, Qi, et al.
Publicado: (2024)
por: Zhang, Qi, et al.
Publicado: (2024)
Flexiffusion: Training-Free Segment-Wise Neural Architecture Search for Efficient Diffusion Models
por: Huang, Hongtao, et al.
Publicado: (2025)
por: Huang, Hongtao, et al.
Publicado: (2025)
MirrorDiffusion: Stabilizing Diffusion Process in Zero-shot Image Translation by Prompts Redescription and Beyond
por: Lin, Yupei, et al.
Publicado: (2024)
por: Lin, Yupei, et al.
Publicado: (2024)
NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning
por: Lin, Bingqian, et al.
Publicado: (2024)
por: Lin, Bingqian, et al.
Publicado: (2024)
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation
por: Hao, Haihong, et al.
Publicado: (2025)
por: Hao, Haihong, et al.
Publicado: (2025)
ActionSink: Toward Precise Robot Manipulation with Dynamic Integration of Action Flow
por: Guo, Shanshan, et al.
Publicado: (2025)
por: Guo, Shanshan, et al.
Publicado: (2025)
Aquarius: A Family of Industry-Level Video Generation Models for Marketing Scenarios
por: Shi, Huafeng, et al.
Publicado: (2025)
por: Shi, Huafeng, et al.
Publicado: (2025)
DFVO: Learning Darkness-free Visible and Infrared Image Disentanglement and Fusion All at Once
por: Zhou, Qi, et al.
Publicado: (2025)
por: Zhou, Qi, et al.
Publicado: (2025)
Geometry aware 3D generation from in-the-wild images in ImageNet
por: Shen, Qijia, et al.
Publicado: (2024)
por: Shen, Qijia, et al.
Publicado: (2024)
WildVidFit: Video Virtual Try-On in the Wild via Image-Based Controlled Diffusion Models
por: He, Zijian, et al.
Publicado: (2024)
por: He, Zijian, et al.
Publicado: (2024)
TopoNAS: Boosting Search Efficiency of Gradient-based NAS via Topological Simplification
por: Zhao, Danpei, et al.
Publicado: (2024)
por: Zhao, Danpei, et al.
Publicado: (2024)
PhyBlock: A Progressive Benchmark for Physical Understanding and Planning via 3D Block Assembly
por: Ma, Liang, et al.
Publicado: (2025)
por: Ma, Liang, et al.
Publicado: (2025)
SIRST-5K: Exploring Massive Negatives Synthesis with Self-supervised Learning for Robust Infrared Small Target Detection
por: Lu, Yahao, et al.
Publicado: (2024)
por: Lu, Yahao, et al.
Publicado: (2024)
Diff-Mosaic: Augmenting Realistic Representations in Infrared Small Target Detection via Diffusion Prior
por: Shi, Yukai, et al.
Publicado: (2024)
por: Shi, Yukai, et al.
Publicado: (2024)
Implicit Geometry Representations for Vision-and-Language Navigation from Web Videos
por: Han, Mingfei, et al.
Publicado: (2026)
por: Han, Mingfei, et al.
Publicado: (2026)
AnimateZoo: Zero-shot Video Generation of Cross-Species Animation via Subject Alignment
por: Xu, Yuanfeng, et al.
Publicado: (2024)
por: Xu, Yuanfeng, et al.
Publicado: (2024)
VTON 360: High-Fidelity Virtual Try-On from Any Viewing Direction
por: He, Zijian, et al.
Publicado: (2025)
por: He, Zijian, et al.
Publicado: (2025)
Boosting Self-Supervised Tracking with Contextual Prompts and Noise Learning
por: Zheng, Yaozong, et al.
Publicado: (2026)
por: Zheng, Yaozong, et al.
Publicado: (2026)
Ejemplares similares
-
Efficient Training of Large Vision Models via Advanced Automated Progressive Learning
por: Li, Changlin, et al.
Publicado: (2024) -
ACD: Direct Conditional Control for Video Diffusion Models via Attention Supervision
por: Li, Weiqi, et al.
Publicado: (2025) -
MixReorg: Cross-Modal Mixed Patch Reorganization is a Good Mask Learner for Open-World Semantic Segmentation
por: Cai, Kaixin, et al.
Publicado: (2023) -
NeRF-VPT: Learning Novel View Representations with Neural Radiance Fields via View Prompt Tuning
por: Chen, Linsheng, et al.
Publicado: (2024) -
AlignMiF: Geometry-Aligned Multimodal Implicit Field for LiDAR-Camera Joint Synthesis
por: Tang, Tao, et al.
Publicado: (2024)