Open-world Instance Segmentation: Top-down Learning with Bottom-up Supervision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kalluri, Tarun, Wang, Weiyao, Wang, Heng, Chandraker, Manmohan, Torresani, Lorenzo, Tran, Du |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
UDA-Bench: Revisiting Common Assumptions in Unsupervised Domain Adaptation Using a Standardized Framework
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
Robust Disaster Assessment from Aerial Imagery Using Text-to-Image Synthetic Data
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
Unifying Top-down and Bottom-up Scanpath Prediction Using Transformers
von: Yang, Zhibo, et al.
Veröffentlicht: (2023)
von: Yang, Zhibo, et al.
Veröffentlicht: (2023)
Tuned Contrastive Learning
von: Animesh, Chaitanya, et al.
Veröffentlicht: (2023)
von: Animesh, Chaitanya, et al.
Veröffentlicht: (2023)
PhyCo: Learning Controllable Physical Priors for Generative Motion
von: Narayanan, Sriram, et al.
Veröffentlicht: (2026)
von: Narayanan, Sriram, et al.
Veröffentlicht: (2026)
Progressive Token Length Scaling in Transformer Encoders for Efficient Universal Segmentation
von: Aich, Abhishek, et al.
Veröffentlicht: (2024)
von: Aich, Abhishek, et al.
Veröffentlicht: (2024)
Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion
von: Li, Haodong, et al.
Veröffentlicht: (2026)
von: Li, Haodong, et al.
Veröffentlicht: (2026)
LLM-Assist: Enhancing Closed-Loop Planning with Language-Based Reasoning
von: Sharan, S P, et al.
Veröffentlicht: (2023)
von: Sharan, S P, et al.
Veröffentlicht: (2023)
RAD-LAD: Rule and Language Grounded Autonomous Driving in Real-Time
von: Ghosh, Anurag, et al.
Veröffentlicht: (2026)
von: Ghosh, Anurag, et al.
Veröffentlicht: (2026)
Semantic Compositions Enhance Vision-Language Contrastive Learning
von: Aladago, Maxwell, et al.
Veröffentlicht: (2024)
von: Aladago, Maxwell, et al.
Veröffentlicht: (2024)
Locally Orderless Images for Optimization in Differentiable Rendering
von: Mehta, Ishit, et al.
Veröffentlicht: (2025)
von: Mehta, Ishit, et al.
Veröffentlicht: (2025)
Tracking Skiers from the Top to the Bottom
von: Dunnhofer, Matteo, et al.
Veröffentlicht: (2023)
von: Dunnhofer, Matteo, et al.
Veröffentlicht: (2023)
OE3DIS: Open-Ended 3D Point Cloud Instance Segmentation
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2024)
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2024)
Materialist: Physically Based Editing Using Single-Image Inverse Rendering
von: Wang, Lezhong, et al.
Veröffentlicht: (2025)
von: Wang, Lezhong, et al.
Veröffentlicht: (2025)
Top-down Activity Representation Learning for Video Question Answering
von: Wang, Yanan, et al.
Veröffentlicht: (2024)
von: Wang, Yanan, et al.
Veröffentlicht: (2024)
CAST: Contrastive Adaptation and Distillation for Semi-Supervised Instance Segmentation
von: Taghavi, Pardis, et al.
Veröffentlicht: (2025)
von: Taghavi, Pardis, et al.
Veröffentlicht: (2025)
MARIS: Marine Open-Vocabulary Instance Segmentation with Geometric Enhancement and Semantic Alignment
von: Li, Bingyu, et al.
Veröffentlicht: (2025)
von: Li, Bingyu, et al.
Veröffentlicht: (2025)
FMaMIL: Frequency-Driven Mamba Multi-Instance Learning for Weakly Supervised Lesion Segmentation in Medical Images
von: Cheng, Hangbei, et al.
Veröffentlicht: (2025)
von: Cheng, Hangbei, et al.
Veröffentlicht: (2025)
Step Differences in Instructional Video
von: Nagarajan, Tushar, et al.
Veröffentlicht: (2024)
von: Nagarajan, Tushar, et al.
Veröffentlicht: (2024)
Instantaneous Perception of Moving Objects in 3D
von: Liu, Di, et al.
Veröffentlicht: (2024)
von: Liu, Di, et al.
Veröffentlicht: (2024)
Open-Vocabulary Segmentation with Unpaired Mask-Text Supervision
von: Wang, Zhaoqing, et al.
Veröffentlicht: (2024)
von: Wang, Zhaoqing, et al.
Veröffentlicht: (2024)
VITED: Video Temporal Evidence Distillation
von: Lu, Yujie, et al.
Veröffentlicht: (2025)
von: Lu, Yujie, et al.
Veröffentlicht: (2025)
HorizonForge: Driving Scene Editing with Any Trajectories and Any Vehicles
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
Image-Specific Adaptation of Transformer Encoders for Compute-Efficient Segmentation
von: Yao, Manyi, et al.
Veröffentlicht: (2024)
von: Yao, Manyi, et al.
Veröffentlicht: (2024)
OW-VISCapTor: Abstractors for Open-World Video Instance Segmentation and Captioning
von: Choudhuri, Anwesa, et al.
Veröffentlicht: (2024)
von: Choudhuri, Anwesa, et al.
Veröffentlicht: (2024)
Foveated Instance Segmentation
von: Zeng, Hongyi, et al.
Veröffentlicht: (2025)
von: Zeng, Hongyi, et al.
Veröffentlicht: (2025)
AIDE: An Automatic Data Engine for Object Detection in Autonomous Driving
von: Liang, Mingfu, et al.
Veröffentlicht: (2024)
von: Liang, Mingfu, et al.
Veröffentlicht: (2024)
Instance Consistency Regularization for Semi-Supervised 3D Instance Segmentation
von: Wu, Yizheng, et al.
Veröffentlicht: (2024)
von: Wu, Yizheng, et al.
Veröffentlicht: (2024)
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation
von: Barsellotti, Luca, et al.
Veröffentlicht: (2024)
von: Barsellotti, Luca, et al.
Veröffentlicht: (2024)
Complete Instances Mining for Weakly Supervised Instance Segmentation
von: Li, Zecheng, et al.
Veröffentlicht: (2024)
von: Li, Zecheng, et al.
Veröffentlicht: (2024)
Self-Training Large Language Models for Improved Visual Program Synthesis With Visual Reinforcement
von: Khan, Zaid, et al.
Veröffentlicht: (2024)
von: Khan, Zaid, et al.
Veröffentlicht: (2024)
Self-Supervision Enhances Instance-based Multiple Instance Learning Methods in Digital Pathology: A Benchmark Study
von: Mammadov, Ali, et al.
Veröffentlicht: (2025)
von: Mammadov, Ali, et al.
Veröffentlicht: (2025)
SEINE: Structure Encoding and Interaction Network for Nuclei Instance Segmentation
von: Zhang, Ye, et al.
Veröffentlicht: (2024)
von: Zhang, Ye, et al.
Veröffentlicht: (2024)
Unsupervised Instance Segmentation with Superpixels
von: Hoang, Cuong Manh
Veröffentlicht: (2025)
von: Hoang, Cuong Manh
Veröffentlicht: (2025)
Self-Supervised Multiple Instance Learning for Acute Myeloid Leukemia Classification
von: Kazeminia, Salome, et al.
Veröffentlicht: (2024)
von: Kazeminia, Salome, et al.
Veröffentlicht: (2024)
Instance-Aware Robust Consistency Regularization for Semi-Supervised Nuclei Instance Segmentation
von: Lin, Zenan, et al.
Veröffentlicht: (2025)
von: Lin, Zenan, et al.
Veröffentlicht: (2025)
OV-MAP : Open-Vocabulary Zero-Shot 3D Instance Segmentation Map for Robots
von: Kim, Juno, et al.
Veröffentlicht: (2025)
von: Kim, Juno, et al.
Veröffentlicht: (2025)
RECIPE: Procedural Planning via Grounding in Instructional Video
von: Seminara, Luigi, et al.
Veröffentlicht: (2026)
von: Seminara, Luigi, et al.
Veröffentlicht: (2026)
Extreme Point Supervised Instance Segmentation
von: Lee, Hyeonjun, et al.
Veröffentlicht: (2024)
von: Lee, Hyeonjun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024) -
UDA-Bench: Revisiting Common Assumptions in Unsupervised Domain Adaptation Using a Standardized Framework
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024) -
Robust Disaster Assessment from Aerial Imagery Using Text-to-Image Synthetic Data
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024) -
Unifying Top-down and Bottom-up Scanpath Prediction Using Transformers
von: Yang, Zhibo, et al.
Veröffentlicht: (2023) -
Tuned Contrastive Learning
von: Animesh, Chaitanya, et al.
Veröffentlicht: (2023)