Learning Triangular Distribution in Visual World
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Ping, Zhang, Xingpeng, Zhou, Chengtao, Fan, Dichao, Tu, Peng, Zhang, Le, Qian, Yanlin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Progressive Multi-task Anti-Noise Learning and Distilling Frameworks for Fine-grained Vehicle Recognition
by: Liu, Dichao
Published: (2024)
by: Liu, Dichao
Published: (2024)
Optimizing for the Shortest Path in Denoising Diffusion Model
by: Chen, Ping, et al.
Published: (2025)
by: Chen, Ping, et al.
Published: (2025)
Beyond Geometry: Artistic Disparity Synthesis for Immersive 2D-to-3D
by: Chen, Ping, et al.
Published: (2026)
by: Chen, Ping, et al.
Published: (2026)
Open-World Dynamic Prompt and Continual Visual Representation Learning
by: Kim, Youngeun, et al.
Published: (2024)
by: Kim, Youngeun, et al.
Published: (2024)
Learning with Adaptive Prototype Manifolds for Out-of-Distribution Detection
by: Peng, Ningkang, et al.
Published: (2026)
by: Peng, Ningkang, et al.
Published: (2026)
SADL: An Effective In-Context Learning Method for Compositional Visual QA
by: Dang, Long Hoang, et al.
Published: (2024)
by: Dang, Long Hoang, et al.
Published: (2024)
Test-time Distribution Learning Adapter for Cross-modal Visual Reasoning
by: Zhang, Yi, et al.
Published: (2024)
by: Zhang, Yi, et al.
Published: (2024)
Towards Visual Query Localization in the 3D World
by: Peng, Liang, et al.
Published: (2026)
by: Peng, Liang, et al.
Published: (2026)
Soft Tail-dropping for Adaptive Visual Tokenization
by: Chen, Zeyuan, et al.
Published: (2026)
by: Chen, Zeyuan, et al.
Published: (2026)
TokenCarve: Information-Preserving Visual Token Compression in Multimodal Large Language Models
by: Tan, Xudong, et al.
Published: (2025)
by: Tan, Xudong, et al.
Published: (2025)
CauSight: Learning to Supersense for Visual Causal Discovery
by: Zhang, Yize, et al.
Published: (2025)
by: Zhang, Yize, et al.
Published: (2025)
Zero-Shot CFC: Fast Real-World Image Denoising based on Cross-Frequency Consistency
by: Jiang, Yanlin, et al.
Published: (2025)
by: Jiang, Yanlin, et al.
Published: (2025)
DORAEMON: A Unified Library for Visual Object Modeling and Representation Learning at Scale
by: Du, Ke, et al.
Published: (2025)
by: Du, Ke, et al.
Published: (2025)
PlayerOne: Egocentric World Simulator
by: Tu, Yuanpeng, et al.
Published: (2025)
by: Tu, Yuanpeng, et al.
Published: (2025)
Assessing and Learning Alignment of Unimodal Vision and Language Models
by: Zhang, Le, et al.
Published: (2024)
by: Zhang, Le, et al.
Published: (2024)
LinVideo: A Post-Training Framework towards O(n) Attention in Efficient Video Generation
by: Huang, Yushi, et al.
Published: (2025)
by: Huang, Yushi, et al.
Published: (2025)
Learning Weakly Supervised Audio-Visual Violence Detection in Hyperbolic Space
by: Peng, Xiaogang, et al.
Published: (2023)
by: Peng, Xiaogang, et al.
Published: (2023)
DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT
by: Hu, Xiaotao, et al.
Published: (2024)
by: Hu, Xiaotao, et al.
Published: (2024)
Semantic Feature Decomposition based Semantic Communication System of Images with Large-scale Visual Generation Models
by: Fan, Senran, et al.
Published: (2024)
by: Fan, Senran, et al.
Published: (2024)
Unified Multi-Modal Image Synthesis for Missing Modality Imputation
by: Zhang, Yue, et al.
Published: (2023)
by: Zhang, Yue, et al.
Published: (2023)
ControlLoc: Physical-World Hijacking Attack on Visual Perception in Autonomous Driving
by: Ma, Chen, et al.
Published: (2024)
by: Ma, Chen, et al.
Published: (2024)
Mani-GS: Gaussian Splatting Manipulation with Triangular Mesh
by: Gao, Xiangjun, et al.
Published: (2024)
by: Gao, Xiangjun, et al.
Published: (2024)
Distributed Zero-Shot Learning for Visual Recognition
by: Chen, Zhi, et al.
Published: (2025)
by: Chen, Zhi, et al.
Published: (2025)
VSCode: General Visual Salient and Camouflaged Object Detection with 2D Prompt Learning
by: Luo, Ziyang, et al.
Published: (2023)
by: Luo, Ziyang, et al.
Published: (2023)
PacTure: Efficient PBR Texture Generation on Packed Views with Visual Autoregressive Models
by: Fei, Fan, et al.
Published: (2025)
by: Fei, Fan, et al.
Published: (2025)
Joint Geometric and Trajectory Consistency Learning for One-Step Real-World Super-Resolution
by: Deng, Chengyan, et al.
Published: (2026)
by: Deng, Chengyan, et al.
Published: (2026)
How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning
by: Yang, Qian, et al.
Published: (2026)
by: Yang, Qian, et al.
Published: (2026)
Beyond Quantity: Distribution-Aware Labeling for Visual Grounding
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
Deep Learning in Image Classification: Evaluating VGG19's Performance on Complex Visual Data
by: He, Weijie, et al.
Published: (2024)
by: He, Weijie, et al.
Published: (2024)
DreamWorld: Unified World Modeling in Video Generation
by: Tan, Boming, et al.
Published: (2026)
by: Tan, Boming, et al.
Published: (2026)
Efficient High-Resolution Visual Representation Learning with State Space Model for Human Pose Estimation
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
The Devil is in the Few Shots: Iterative Visual Knowledge Completion for Few-shot Learning
by: Li, Yaohui, et al.
Published: (2024)
by: Li, Yaohui, et al.
Published: (2024)
Hierarchical Visual Categories Modeling: A Joint Representation Learning and Density Estimation Framework for Out-of-Distribution Detection
by: Li, Jinglun, et al.
Published: (2024)
by: Li, Jinglun, et al.
Published: (2024)
The Role of World Models in Shaping Autonomous Driving: A Comprehensive Survey
by: Tu, Sifan, et al.
Published: (2025)
by: Tu, Sifan, et al.
Published: (2025)
Semantically Guided Dynamic Visual Prototype Refinement for Compositional Zero-Shot Learning
by: Peng, Zhong, et al.
Published: (2025)
by: Peng, Zhong, et al.
Published: (2025)
Uni-HOI:A Unified framework for Learning the Joint distribution of Text and Human-Object Interaction
by: Zhang, Mengfei, et al.
Published: (2026)
by: Zhang, Mengfei, et al.
Published: (2026)
MV-CoRe: Multimodal Visual-Conceptual Reasoning for Complex Visual Question Answering
by: Peng, Jingwei, et al.
Published: (2025)
by: Peng, Jingwei, et al.
Published: (2025)
VQ-VA World: Towards High-Quality Visual Question-Visual Answering
by: Gou, Chenhui, et al.
Published: (2025)
by: Gou, Chenhui, et al.
Published: (2025)
When Pedestrian Detection Meets Multi-Modal Learning: Generalist Model and Benchmark Dataset
by: Zhang, Yi, et al.
Published: (2024)
by: Zhang, Yi, et al.
Published: (2024)
ORB-SfMLearner: ORB-Guided Self-supervised Visual Odometry with Selective Online Adaptation
by: Jin, Yanlin, et al.
Published: (2024)
by: Jin, Yanlin, et al.
Published: (2024)
Similar Items
-
Progressive Multi-task Anti-Noise Learning and Distilling Frameworks for Fine-grained Vehicle Recognition
by: Liu, Dichao
Published: (2024) -
Optimizing for the Shortest Path in Denoising Diffusion Model
by: Chen, Ping, et al.
Published: (2025) -
Beyond Geometry: Artistic Disparity Synthesis for Immersive 2D-to-3D
by: Chen, Ping, et al.
Published: (2026) -
Open-World Dynamic Prompt and Continual Visual Representation Learning
by: Kim, Youngeun, et al.
Published: (2024) -
Learning with Adaptive Prototype Manifolds for Out-of-Distribution Detection
by: Peng, Ningkang, et al.
Published: (2026)