Learning Generalized and Flexible Trajectory Models from Omni-Semantic Supervision
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Yuanshao, Yu, James Jianqiao, Zhao, Xiangyu, Han, Xiao, Liu, Qidong, Wei, Xuetao, Liang, Yuxuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ControlTraj: Controllable Trajectory Generation with Topology-Constrained Diffusion Model
by: Zhu, Yuanshao, et al.
Published: (2024)
by: Zhu, Yuanshao, et al.
Published: (2024)
UniTraj: Learning a Universal Trajectory Foundation Model from Billion-Scale Worldwide Traces
by: Zhu, Yuanshao, et al.
Published: (2024)
by: Zhu, Yuanshao, et al.
Published: (2024)
InsTraj: Instructing Diffusion Models with Travel Intentions to Generate Real-world Trajectories
by: Zhu, Yuanshao, et al.
Published: (2026)
by: Zhu, Yuanshao, et al.
Published: (2026)
OmniSegmentor: A Flexible Multi-Modal Learning Framework for Semantic Segmentation
by: Yin, Bo-Wen, et al.
Published: (2025)
by: Yin, Bo-Wen, et al.
Published: (2025)
MetaSeg: Content-Aware Meta-Net for Omni-Supervised Semantic Segmentation
by: Jiang, Shenwang, et al.
Published: (2024)
by: Jiang, Shenwang, et al.
Published: (2024)
Frozen CLIP: A Strong Backbone for Weakly Supervised Semantic Segmentation
by: Zhang, Bingfeng, et al.
Published: (2024)
by: Zhang, Bingfeng, et al.
Published: (2024)
GeoRouter: Dynamic Paradigm Routing for Worldwide Image Geolocalization
by: Jia, Pengyue, et al.
Published: (2026)
by: Jia, Pengyue, et al.
Published: (2026)
Omni$^2$: Unifying Omnidirectional Image Generation and Editing in an Omni Model
by: Yang, Liu, et al.
Published: (2025)
by: Yang, Liu, et al.
Published: (2025)
Context Unrolling in Omni Models
by: Yang, Ceyuan, et al.
Published: (2026)
by: Yang, Ceyuan, et al.
Published: (2026)
Convolutional Networks as Extremely Small Foundation Models: Visual Prompting and Theoretical Perspective
by: Wangni, Jianqiao
Published: (2024)
by: Wangni, Jianqiao
Published: (2024)
OMUDA: Omni-level Masking for Unsupervised Domain Adaptation in Semantic Segmentation
by: Ou, Yang, et al.
Published: (2025)
by: Ou, Yang, et al.
Published: (2025)
OmniVDiff: Omni Controllable Video Diffusion for Generation and Understanding
by: Xi, Dianbing, et al.
Published: (2025)
by: Xi, Dianbing, et al.
Published: (2025)
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation
by: Liao, Chao, et al.
Published: (2025)
by: Liao, Chao, et al.
Published: (2025)
Boosting Fine-Grained Urban Flow Inference via Lightweight Architecture and Focalized Optimization
by: Zhu, Yuanshao, et al.
Published: (2025)
by: Zhu, Yuanshao, et al.
Published: (2025)
R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning
by: Zhao, Jiaxing, et al.
Published: (2025)
by: Zhao, Jiaxing, et al.
Published: (2025)
Swarm Intelligence in Geo-Localization: A Multi-Agent Large Vision-Language Model Collaborative Framework
by: Han, Xiao, et al.
Published: (2024)
by: Han, Xiao, et al.
Published: (2024)
OmniPart: Part-Aware 3D Generation with Semantic Decoupling and Structural Cohesion
by: Yang, Yunhan, et al.
Published: (2025)
by: Yang, Yunhan, et al.
Published: (2025)
X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again
by: Geng, Zigang, et al.
Published: (2025)
by: Geng, Zigang, et al.
Published: (2025)
OmniSVG: A Unified Scalable Vector Graphics Generation Model
by: Yang, Yiying, et al.
Published: (2025)
by: Yang, Yiying, et al.
Published: (2025)
OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation
by: Zhang, Guohui, et al.
Published: (2026)
by: Zhang, Guohui, et al.
Published: (2026)
OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning
by: Zhao, Shifang, et al.
Published: (2025)
by: Zhao, Shifang, et al.
Published: (2025)
SegMAN: Omni-scale Context Modeling with State Space Models and Local Attention for Semantic Segmentation
by: Fu, Yunxiang, et al.
Published: (2024)
by: Fu, Yunxiang, et al.
Published: (2024)
Semi-Supervised Learning for Visual Bird's Eye View Semantic Segmentation
by: Zhu, Junyu, et al.
Published: (2023)
by: Zhu, Junyu, et al.
Published: (2023)
Visual-Augmented Dynamic Semantic Prototype for Generative Zero-Shot Learning
by: Hou, Wenjin, et al.
Published: (2024)
by: Hou, Wenjin, et al.
Published: (2024)
Grid: Omni Visual Generation
by: Wan, Cong, et al.
Published: (2024)
by: Wan, Cong, et al.
Published: (2024)
Omni2Sound: Towards Unified Video-Text-to-Audio Generation
by: Dai, Yusheng, et al.
Published: (2026)
by: Dai, Yusheng, et al.
Published: (2026)
Boundary-Refined Prototype Generation: A General End-to-End Paradigm for Semi-Supervised Semantic Segmentation
by: Dong, Junhao, et al.
Published: (2023)
by: Dong, Junhao, et al.
Published: (2023)
OmniMotion-X: Versatile Multimodal Whole-Body Motion Generation
by: Xu, Guowei, et al.
Published: (2025)
by: Xu, Guowei, et al.
Published: (2025)
OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision
by: Wei, Cong, et al.
Published: (2024)
by: Wei, Cong, et al.
Published: (2024)
OmniVaT: Single Domain Generalization for Multimodal Visual-Tactile Learning
by: Qiu, Liuxiang, et al.
Published: (2026)
by: Qiu, Liuxiang, et al.
Published: (2026)
From Easy to Hard: Progressive Active Learning Framework for Infrared Small Target Detection with Single Point Supervision
by: Yu, Chuang, et al.
Published: (2024)
by: Yu, Chuang, et al.
Published: (2024)
Omni-Supervised Motion Editing: Balancing Change and Invariance through Positive-Negative Learning
by: Shi, Zhenwu, et al.
Published: (2026)
by: Shi, Zhenwu, et al.
Published: (2026)
G3: An Effective and Adaptive Framework for Worldwide Geolocalization Using Large Multi-Modality Models
by: Jia, Pengyue, et al.
Published: (2024)
by: Jia, Pengyue, et al.
Published: (2024)
Learning Semantic Directions for Feature Augmentation in Domain-Generalized Medical Segmentation
by: Wang, Yingkai, et al.
Published: (2025)
by: Wang, Yingkai, et al.
Published: (2025)
OmniCLIP: Adapting CLIP for Video Recognition with Spatial-Temporal Omni-Scale Feature Learning
by: Liu, Mushui, et al.
Published: (2024)
by: Liu, Mushui, et al.
Published: (2024)
Hierarchical Semi-Supervised Active Learning for Remote Sensing
by: Huang, Wei, et al.
Published: (2025)
by: Huang, Wei, et al.
Published: (2025)
InteractiveOmni: A Unified Omni-modal Model for Audio-Visual Multi-turn Dialogue
by: Tong, Wenwen, et al.
Published: (2025)
by: Tong, Wenwen, et al.
Published: (2025)
PixelDINO: Semi-Supervised Semantic Segmentation for Detecting Permafrost Disturbances
by: Heidler, Konrad, et al.
Published: (2024)
by: Heidler, Konrad, et al.
Published: (2024)
OmniAID: Decoupling Semantic and Artifacts for Universal AI-Generated Image Detection in the Wild
by: Guo, Yuncheng, et al.
Published: (2025)
by: Guo, Yuncheng, et al.
Published: (2025)
Omni-AD: Learning to Reconstruct Global and Local Features for Multi-class Anomaly Detection
by: Quan, Jiajie, et al.
Published: (2025)
by: Quan, Jiajie, et al.
Published: (2025)
Similar Items
-
ControlTraj: Controllable Trajectory Generation with Topology-Constrained Diffusion Model
by: Zhu, Yuanshao, et al.
Published: (2024) -
UniTraj: Learning a Universal Trajectory Foundation Model from Billion-Scale Worldwide Traces
by: Zhu, Yuanshao, et al.
Published: (2024) -
InsTraj: Instructing Diffusion Models with Travel Intentions to Generate Real-world Trajectories
by: Zhu, Yuanshao, et al.
Published: (2026) -
OmniSegmentor: A Flexible Multi-Modal Learning Framework for Semantic Segmentation
by: Yin, Bo-Wen, et al.
Published: (2025) -
MetaSeg: Content-Aware Meta-Net for Omni-Supervised Semantic Segmentation
by: Jiang, Shenwang, et al.
Published: (2024)