Lotus-2: Advancing Geometric Dense Prediction with Powerful Image Generative Model
Fuente:
arXiv
Saved in:
| Main Authors: | He, Jing, Li, Haodong, Sheng, Mingzhi, Chen, Ying-Cong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction
by: He, Jing, et al.
Published: (2024)
by: He, Jing, et al.
Published: (2024)
DisEnvisioner: Disentangled and Enriched Visual Prompt for Customized Image Generation
by: He, Jing, et al.
Published: (2024)
by: He, Jing, et al.
Published: (2024)
Spatial Lifting for Dense Prediction
by: Xu, Mingzhi, et al.
Published: (2025)
by: Xu, Mingzhi, et al.
Published: (2025)
Bi-TTA: Bidirectional Test-Time Adapter for Remote Physiological Measurement
by: Li, Haodong, et al.
Published: (2024)
by: Li, Haodong, et al.
Published: (2024)
FlexAM: Flexible Appearance-Motion Decomposition for Versatile Video Generation Control
by: Sheng, Mingzhi, et al.
Published: (2026)
by: Sheng, Mingzhi, et al.
Published: (2026)
Frequency Dynamic Convolution for Dense Image Prediction
by: Chen, Linwei, et al.
Published: (2025)
by: Chen, Linwei, et al.
Published: (2025)
DA$^{2}$: Depth Anything in Any Direction
by: Li, Haodong, et al.
Published: (2025)
by: Li, Haodong, et al.
Published: (2025)
Go with Your Gut: Scaling Confidence for Autoregressive Image Generation
by: Chen, Harold Haodong, et al.
Published: (2025)
by: Chen, Harold Haodong, et al.
Published: (2025)
Advancing high-fidelity 3D and Texture Generation with 2.5D latents
by: Yang, Xin, et al.
Published: (2025)
by: Yang, Xin, et al.
Published: (2025)
Show, Don't Tell: Morphing Latent Reasoning into Image Generation
by: Chen, Harold Haodong, et al.
Published: (2026)
by: Chen, Harold Haodong, et al.
Published: (2026)
DVD: Deterministic Video Depth Estimation with Generative Priors
by: Zhang, Hongfei, et al.
Published: (2026)
by: Zhang, Hongfei, et al.
Published: (2026)
Unified Dense Prediction of Video Diffusion
by: Yang, Lehan, et al.
Published: (2025)
by: Yang, Lehan, et al.
Published: (2025)
DenseSR: Image Shadow Removal as Dense Prediction
by: Lin, Yu-Fan, et al.
Published: (2025)
by: Lin, Yu-Fan, et al.
Published: (2025)
S-VAM: Shortcut Video-Action Model by Self-Distilling Geometric and Semantic Foresight
by: Yan, Haodong, et al.
Published: (2026)
by: Yan, Haodong, et al.
Published: (2026)
Dense-Face: Personalized Face Generation Model via Dense Annotation Prediction
by: Guo, Xiao, et al.
Published: (2024)
by: Guo, Xiao, et al.
Published: (2024)
Frequency-aware Feature Fusion for Dense Image Prediction
by: Chen, Linwei, et al.
Published: (2024)
by: Chen, Linwei, et al.
Published: (2024)
Adversarial Patch Generation for Visual-Infrared Dense Prediction Tasks via Joint Position-Color Optimization
by: Li, He, et al.
Published: (2026)
by: Li, He, et al.
Published: (2026)
Exploiting Diffusion Prior for Generalizable Dense Prediction
by: Lee, Hsin-Ying, et al.
Published: (2023)
by: Lee, Hsin-Ying, et al.
Published: (2023)
AttenST: A Training-Free Attention-Driven Style Transfer Framework with Pre-Trained Diffusion Models
by: Huang, Bo, et al.
Published: (2025)
by: Huang, Bo, et al.
Published: (2025)
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
by: Zhu, Hanxin, et al.
Published: (2026)
by: Zhu, Hanxin, et al.
Published: (2026)
SyntheOcc: Synthesize Geometric-Controlled Street View Images through 3D Semantic MPIs
by: Li, Leheng, et al.
Published: (2024)
by: Li, Leheng, et al.
Published: (2024)
On the Robustness of Object Detection Models on Aerial Images
by: He, Haodong, et al.
Published: (2023)
by: He, Haodong, et al.
Published: (2023)
BiDense: Binarization for Dense Prediction
by: Yin, Rui, et al.
Published: (2024)
by: Yin, Rui, et al.
Published: (2024)
CLIPSelf: Vision Transformer Distills Itself for Open-Vocabulary Dense Prediction
by: Wu, Size, et al.
Published: (2023)
by: Wu, Size, et al.
Published: (2023)
RS-Mamba for Large Remote Sensing Image Dense Prediction
by: Zhao, Sijie, et al.
Published: (2024)
by: Zhao, Sijie, et al.
Published: (2024)
Sat2City: 3D City Generation from A Single Satellite Image with Cascaded Latent Diffusion
by: Hua, Tongyan, et al.
Published: (2025)
by: Hua, Tongyan, et al.
Published: (2025)
Frequency-Dynamic Attention Modulation for Dense Prediction
by: Chen, Linwei, et al.
Published: (2025)
by: Chen, Linwei, et al.
Published: (2025)
DiffCalib: Reformulating Monocular Camera Calibration as Diffusion-Based Dense Incident Map Generation
by: He, Xiankang, et al.
Published: (2024)
by: He, Xiankang, et al.
Published: (2024)
A Unified Image-Dense Annotation Generation Model for Underwater Scenes
by: Lin, Hongkai, et al.
Published: (2025)
by: Lin, Hongkai, et al.
Published: (2025)
OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlaps
by: Li, Bingnan, et al.
Published: (2025)
by: Li, Bingnan, et al.
Published: (2025)
RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data
by: Chen, Harold Haodong, et al.
Published: (2026)
by: Chen, Harold Haodong, et al.
Published: (2026)
Federated Distillation for Whole Slide Image via Gaussian-Mixture Feature Alignment and Curriculum Integration
by: Jing, Luru, et al.
Published: (2026)
by: Jing, Luru, et al.
Published: (2026)
PRJ: Perception-Retrieval-Judgement for Generated Images
by: Fu, Qiang, et al.
Published: (2025)
by: Fu, Qiang, et al.
Published: (2025)
GeoTikzBridge: Advancing Multimodal Code Generation for Geometric Perception and Reasoning
by: Sun, Jiayin, et al.
Published: (2026)
by: Sun, Jiayin, et al.
Published: (2026)
Sparse2Dense: A Keypoint-driven Generative Framework for Human Video Compression and Vertex Prediction
by: Chen, Bolin, et al.
Published: (2025)
by: Chen, Bolin, et al.
Published: (2025)
Beyond ViT Tokens: Masked-Diffusion Pretrained Convolutional Pathology Foundation Model for Cell-Level Dense Prediction
by: Chen, Weiming, et al.
Published: (2026)
by: Chen, Weiming, et al.
Published: (2026)
GOBench: Benchmarking Geometric Optics Generation and Understanding of MLLMs
by: Zhu, Xiaorong, et al.
Published: (2025)
by: Zhu, Xiaorong, et al.
Published: (2025)
2DGS-Room: Seed-Guided 2D Gaussian Splatting with Geometric Constrains for High-Fidelity Indoor Scene Reconstruction
by: Zhang, Wanting, et al.
Published: (2024)
by: Zhang, Wanting, et al.
Published: (2024)
SGEdit: Bridging LLM with Text2Image Generative Model for Scene Graph-based Image Editing
by: Zhang, Zhiyuan, et al.
Published: (2024)
by: Zhang, Zhiyuan, et al.
Published: (2024)
Adv3D: Generating 3D Adversarial Examples for 3D Object Detection in Driving Scenarios with NeRF
by: Li, Leheng, et al.
Published: (2023)
by: Li, Leheng, et al.
Published: (2023)
Similar Items
-
Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction
by: He, Jing, et al.
Published: (2024) -
DisEnvisioner: Disentangled and Enriched Visual Prompt for Customized Image Generation
by: He, Jing, et al.
Published: (2024) -
Spatial Lifting for Dense Prediction
by: Xu, Mingzhi, et al.
Published: (2025) -
Bi-TTA: Bidirectional Test-Time Adapter for Remote Physiological Measurement
by: Li, Haodong, et al.
Published: (2024) -
FlexAM: Flexible Appearance-Motion Decomposition for Versatile Video Generation Control
by: Sheng, Mingzhi, et al.
Published: (2026)