LiWi: Layering in the Wild
Fuente:
arXiv
Saved in:
| Main Authors: | He, Yu, Li, Fang, Tong, Haoyang, Ma, Lichen, Shan, Xinyuan, Fu, Jingling, Chen, Dong, Liu, Luohang, Huang, Junshi, Li, Yan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FrequencyBooster: Full-Frequency Modeling for High-Fidelity Pixel Diffusion
by: Ma, Lichen, et al.
Published: (2026)
by: Ma, Lichen, et al.
Published: (2026)
HyperDiT: Hyper-Connected Transformers for High-Fidelity Pixel-Space Diffusion
by: He, Yu, et al.
Published: (2026)
by: He, Yu, et al.
Published: (2026)
Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition
by: He, Yu, et al.
Published: (2026)
by: He, Yu, et al.
Published: (2026)
Dynamic-TreeRPO: Breaking the Independent Trajectory Bottleneck with Structured Sampling
by: Fu, Xiaolong, et al.
Published: (2025)
by: Fu, Xiaolong, et al.
Published: (2025)
RePainter: Empowering E-commerce Object Removal via Spatial-matting Reinforcement Learning
by: Guo, Zipeng, et al.
Published: (2025)
by: Guo, Zipeng, et al.
Published: (2025)
UM-Text: A Unified Multimodal Model for Image Understanding and Visual Text Editing
by: Ma, Lichen, et al.
Published: (2026)
by: Ma, Lichen, et al.
Published: (2026)
Enhancing Quantization-Aware Training on Edge Devices via Relative Entropy Coreset Selection and Cascaded Layer Correction
by: Tong, Yujia, et al.
Published: (2025)
by: Tong, Yujia, et al.
Published: (2025)
Diffusion-RWKV: Scaling RWKV-Like Architectures for Diffusion Models
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
Scaling Diffusion Transformers to 16 Billion Parameters
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
Dimba: Transformer-Mamba Diffusion Models
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
Scalable Diffusion Models with State Space Backbone
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
Generalizing WiFi Gesture Recognition via Large-Model-Aware Semantic Distillation and Alignment
by: Cui, Feng-Qi, et al.
Published: (2025)
by: Cui, Feng-Qi, et al.
Published: (2025)
Towards Visual Query Segmentation in the Wild
by: Fan, Bing, et al.
Published: (2026)
by: Fan, Bing, et al.
Published: (2026)
Visual Attention Drifts,but Anchors Hold:Mitigating Hallucination in Multimodal Large Language Models via Cross-Layer Visual Anchors
by: Yang, Chengxu, et al.
Published: (2026)
by: Yang, Chengxu, et al.
Published: (2026)
WildGHand: Learning Anti-Perturbation Gaussian Hand Avatars from Monocular In-the-Wild Videos
by: Li, Hanhui, et al.
Published: (2026)
by: Li, Hanhui, et al.
Published: (2026)
Scaling Up to Excellence: Practicing Model Scaling for Photo-Realistic Image Restoration In the Wild
by: Yu, Fanghua, et al.
Published: (2024)
by: Yu, Fanghua, et al.
Published: (2024)
ReWiTe: Realistic Wide-angle and Telephoto Dual Camera Fusion Dataset via Beam Splitter Camera Rig
by: Peng, Chunli, et al.
Published: (2024)
by: Peng, Chunli, et al.
Published: (2024)
OAT: Object-Level Attention Transformer for Gaze Scanpath Prediction
by: Fang, Yini, et al.
Published: (2024)
by: Fang, Yini, et al.
Published: (2024)
Accelerating Text-to-Image Editing via Cache-Enabled Sparse Diffusion Inference
by: Yu, Zihao, et al.
Published: (2023)
by: Yu, Zihao, et al.
Published: (2023)
Convolution Meets LoRA: Parameter Efficient Finetuning for Segment Anything Model
by: Zhong, Zihan, et al.
Published: (2024)
by: Zhong, Zihan, et al.
Published: (2024)
Cascading Refinement Video Denoising with Uncertainty Adaptivity
by: Yu, Xinyuan
Published: (2024)
by: Yu, Xinyuan
Published: (2024)
Poisoning Prompt-Guided Sampling in Video Large Language Models
by: Cao, Yuxin, et al.
Published: (2025)
by: Cao, Yuxin, et al.
Published: (2025)
OmniAID: Decoupling Semantic and Artifacts for Universal AI-Generated Image Detection in the Wild
by: Guo, Yuncheng, et al.
Published: (2025)
by: Guo, Yuncheng, et al.
Published: (2025)
InterSyn: Interleaved Learning for Dynamic Motion Synthesis in the Wild
by: Ma, Yiyi, et al.
Published: (2025)
by: Ma, Yiyi, et al.
Published: (2025)
All in One Framework for Multimodal Re-identification in the Wild
by: Li, He, et al.
Published: (2024)
by: Li, He, et al.
Published: (2024)
JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing
by: Wang, Qili, et al.
Published: (2025)
by: Wang, Qili, et al.
Published: (2025)
WildVidFit: Video Virtual Try-On in the Wild via Image-Based Controlled Diffusion Models
by: He, Zijian, et al.
Published: (2024)
by: He, Zijian, et al.
Published: (2024)
FocalCount: Towards Class-Count Imbalance in Class-Agnostic Counting
by: Zhu, Huilin, et al.
Published: (2025)
by: Zhu, Huilin, et al.
Published: (2025)
TopoPoint: Enhance Topology Reasoning via Endpoint Detection in Autonomous Driving
by: Fu, Yanping, et al.
Published: (2025)
by: Fu, Yanping, et al.
Published: (2025)
FLUX that Plays Music
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
Expanding Zero-Shot Object Counting with Rich Prompts
by: Zhu, Huilin, et al.
Published: (2025)
by: Zhu, Huilin, et al.
Published: (2025)
Cinemo: Consistent and Controllable Image Animation with Motion Diffusion Models
by: Ma, Xin, et al.
Published: (2024)
by: Ma, Xin, et al.
Published: (2024)
A-JEPA: Joint-Embedding Predictive Architecture Can Listen
by: Fei, Zhengcong, et al.
Published: (2023)
by: Fei, Zhengcong, et al.
Published: (2023)
AdaPose: Towards Cross-Site Device-Free Human Pose Estimation with Commodity WiFi
by: Zhou, Yunjiao, et al.
Published: (2023)
by: Zhou, Yunjiao, et al.
Published: (2023)
Reconstructing In-the-Wild Open-Vocabulary Human-Object Interactions
by: Wen, Boran, et al.
Published: (2025)
by: Wen, Boran, et al.
Published: (2025)
HAD: Hierarchical Asymmetric Distillation to Bridge Spatio-Temporal Gaps in Event-Based Object Tracking
by: Deng, Yao, et al.
Published: (2025)
by: Deng, Yao, et al.
Published: (2025)
TrackVLA: Embodied Visual Tracking in the Wild
by: Wang, Shaoan, et al.
Published: (2025)
by: Wang, Shaoan, et al.
Published: (2025)
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder
by: Ma, Lichen, et al.
Published: (2024)
by: Ma, Lichen, et al.
Published: (2024)
4DRaL: Bridging 4D Radar with LiDAR for Place Recognition using Knowledge Distillation
by: Huang, Ningyuan, et al.
Published: (2026)
by: Huang, Ningyuan, et al.
Published: (2026)
Wi-Spike: A Low-power WiFi Human Multi-action Recognition Model with Spiking Neural Networks
by: Zhang, Nengbo, et al.
Published: (2026)
by: Zhang, Nengbo, et al.
Published: (2026)
Similar Items
-
FrequencyBooster: Full-Frequency Modeling for High-Fidelity Pixel Diffusion
by: Ma, Lichen, et al.
Published: (2026) -
HyperDiT: Hyper-Connected Transformers for High-Fidelity Pixel-Space Diffusion
by: He, Yu, et al.
Published: (2026) -
Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition
by: He, Yu, et al.
Published: (2026) -
Dynamic-TreeRPO: Breaking the Independent Trajectory Bottleneck with Structured Sampling
by: Fu, Xiaolong, et al.
Published: (2025) -
RePainter: Empowering E-commerce Object Removal via Spatial-matting Reinforcement Learning
by: Guo, Zipeng, et al.
Published: (2025)