Neural Lineage
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Runpeng, Wang, Xinchao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Encapsulating Knowledge in One Prompt
von: Li, Qi, et al.
Veröffentlicht: (2024)
von: Li, Qi, et al.
Veröffentlicht: (2024)
Dimple: Discrete Diffusion Multimodal Large Language Model with Parallel Decoding
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
Attention Prompting on Image for Large Vision-Language Models
von: Yu, Runpeng, et al.
Veröffentlicht: (2024)
von: Yu, Runpeng, et al.
Veröffentlicht: (2024)
Introducing Visual Perception Token into Multimodal Large Language Model
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
Vid-SME: Membership Inference Attacks against Large Video Understanding Models
von: Li, Qi, et al.
Veröffentlicht: (2025)
von: Li, Qi, et al.
Veröffentlicht: (2025)
NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language Priors
von: Ren, Lingfeng, et al.
Veröffentlicht: (2026)
von: Ren, Lingfeng, et al.
Veröffentlicht: (2026)
Neural Metamorphosis
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Hash3D: Training-free Acceleration for 3D Generation
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Language Model as Visual Explainer
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Compositional Video Generation as Flow Equalization
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Sponge Tool Attack: Stealthy Denial-of-Efficiency against Tool-Augmented Agentic Reasoning
von: Li, Qi, et al.
Veröffentlicht: (2026)
von: Li, Qi, et al.
Veröffentlicht: (2026)
ViMU: Benchmarking Video Metaphorical Understanding
von: Li, Qi, et al.
Veröffentlicht: (2026)
von: Li, Qi, et al.
Veröffentlicht: (2026)
Teddy: Efficient Large-Scale Dataset Distillation via Taylor-Approximated Matching
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
Ultra-Resolution Adaptation with Ease
von: Yu, Ruonan, et al.
Veröffentlicht: (2025)
von: Yu, Ruonan, et al.
Veröffentlicht: (2025)
PE3R: Perception-Efficient 3D Reconstruction
von: Hu, Jie, et al.
Veröffentlicht: (2025)
von: Hu, Jie, et al.
Veröffentlicht: (2025)
Unsegment Anything by Simulating Deformation
von: Lu, Jiahao, et al.
Veröffentlicht: (2024)
von: Lu, Jiahao, et al.
Veröffentlicht: (2024)
Relation Rectification in Diffusion Model
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
von: Wu, Yinwei, et al.
Veröffentlicht: (2024)
CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
Efficient Gaussian Splatting for Monocular Dynamic Scene Rendering via Sparse Time-Variant Attribute Modeling
von: Kong, Hanyang, et al.
Veröffentlicht: (2025)
von: Kong, Hanyang, et al.
Veröffentlicht: (2025)
CoDA: From Text-to-Image Diffusion Models to Training-Free Dataset Distillation
von: Zhou, Letian, et al.
Veröffentlicht: (2025)
von: Zhou, Letian, et al.
Veröffentlicht: (2025)
MambaOut: Do We Really Need Mamba for Vision?
von: Yu, Weihao, et al.
Veröffentlicht: (2024)
von: Yu, Weihao, et al.
Veröffentlicht: (2024)
Ungeneralizable Examples
von: Ye, Jingwen, et al.
Veröffentlicht: (2024)
von: Ye, Jingwen, et al.
Veröffentlicht: (2024)
Heavy Labels Out! Dataset Distillation with Label Space Lightening
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
ViFeEdit: A Video-Free Tuner of Your Video Diffusion Transformer
von: Yu, Ruonan, et al.
Veröffentlicht: (2026)
von: Yu, Ruonan, et al.
Veröffentlicht: (2026)
Top-Down Compression: Revisit Efficient Vision Token Projection for Visual Instruction Tuning
von: li, Bonan, et al.
Veröffentlicht: (2025)
von: li, Bonan, et al.
Veröffentlicht: (2025)
Domain-Adaptive 2D Human Pose Estimation via Dual Teachers in Extremely Low-Light Conditions
von: Ai, Yihao, et al.
Veröffentlicht: (2024)
von: Ai, Yihao, et al.
Veröffentlicht: (2024)
Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
Focus on Neighbors and Know the Whole: Towards Consistent Dense Multiview Text-to-Image Generator for 3D Creation
von: Li, Bonan, et al.
Veröffentlicht: (2024)
von: Li, Bonan, et al.
Veröffentlicht: (2024)
Control and Realism: Best of Both Worlds in Layout-to-Image without Training
von: Li, Bonan, et al.
Veröffentlicht: (2025)
von: Li, Bonan, et al.
Veröffentlicht: (2025)
1000+ FPS 4D Gaussian Splatting for Dynamic Scene Rendering
von: Yuan, Yuheng, et al.
Veröffentlicht: (2025)
von: Yuan, Yuheng, et al.
Veröffentlicht: (2025)
Anatomy of a Lie: A Multi-Stage Diagnostic Framework for Tracing Hallucinations in Vision-Language Models
von: Xiong, Lexiang, et al.
Veröffentlicht: (2026)
von: Xiong, Lexiang, et al.
Veröffentlicht: (2026)
SlimSAM: 0.1% Data Makes Segment Anything Slim
von: Chen, Zigeng, et al.
Veröffentlicht: (2023)
von: Chen, Zigeng, et al.
Veröffentlicht: (2023)
Flash Sculptor: Modular 3D Worlds from Objects
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
Distilled Datamodel with Reverse Gradient Matching
von: Ye, Jingwen, et al.
Veröffentlicht: (2024)
von: Ye, Jingwen, et al.
Veröffentlicht: (2024)
LinFusion: 1 GPU, 1 Minute, 16K Image
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
Hypergraph-State Collaborative Reasoning for Multi-Object Tracking
von: Song, Zikai, et al.
Veröffentlicht: (2026)
von: Song, Zikai, et al.
Veröffentlicht: (2026)
DreamDrone: Text-to-Image Diffusion Models are Zero-shot Perpetual View Generators
von: Kong, Hanyang, et al.
Veröffentlicht: (2023)
von: Kong, Hanyang, et al.
Veröffentlicht: (2023)
Test3R: Learning to Reconstruct 3D at Test Time
von: Yuan, Yuheng, et al.
Veröffentlicht: (2025)
von: Yuan, Yuheng, et al.
Veröffentlicht: (2025)
One-shot Federated Learning via Synthetic Distiller-Distillate Communication
von: Zhang, Junyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Junyuan, et al.
Veröffentlicht: (2024)
Remix-DiT: Mixing Diffusion Transformers for Multi-Expert Denoising
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Encapsulating Knowledge in One Prompt
von: Li, Qi, et al.
Veröffentlicht: (2024) -
Dimple: Discrete Diffusion Multimodal Large Language Model with Parallel Decoding
von: Yu, Runpeng, et al.
Veröffentlicht: (2025) -
Attention Prompting on Image for Large Vision-Language Models
von: Yu, Runpeng, et al.
Veröffentlicht: (2024) -
Introducing Visual Perception Token into Multimodal Large Language Model
von: Yu, Runpeng, et al.
Veröffentlicht: (2025) -
Vid-SME: Membership Inference Attacks against Large Video Understanding Models
von: Li, Qi, et al.
Veröffentlicht: (2025)