HRSAM: Efficient Interactive Segmentation in High-Resolution Images
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, You, Lai, Wenbin, Ji, Jiayi, Cao, Liujuan, Zhang, Shengchuan, Ji, Rongrong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Inter2Former: Dynamic Hybrid Attention for Efficient High-Precision Interactive
von: Huang, You, et al.
Veröffentlicht: (2025)
von: Huang, You, et al.
Veröffentlicht: (2025)
Evolving, Not Training: Zero-Shot Reasoning Segmentation via Evolutionary Prompting
von: Ye, Kai, et al.
Veröffentlicht: (2025)
von: Ye, Kai, et al.
Veröffentlicht: (2025)
FocSAM: Delving Deeply into Focused Objects in Segmenting Anything
von: Huang, You, et al.
Veröffentlicht: (2024)
von: Huang, You, et al.
Veröffentlicht: (2024)
HRSeg: High-Resolution Visual Perception and Enhancement for Reasoning Segmentation
von: Lin, Weihuang, et al.
Veröffentlicht: (2025)
von: Lin, Weihuang, et al.
Veröffentlicht: (2025)
Evolving High-Quality Rendering and Reconstruction in a Unified Framework with Contribution-Adaptive Regularization
von: Shen, You, et al.
Veröffentlicht: (2025)
von: Shen, You, et al.
Veröffentlicht: (2025)
Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
CutDiffusion: A Simple, Fast, Cheap, and Strong Diffusion Extrapolation Method
von: Lin, Mingbao, et al.
Veröffentlicht: (2024)
von: Lin, Mingbao, et al.
Veröffentlicht: (2024)
What You Perceive Is What You Conceive: A Cognition-Inspired Framework for Open Vocabulary Image Segmentation
von: Lin, Jianghang, et al.
Veröffentlicht: (2025)
von: Lin, Jianghang, et al.
Veröffentlicht: (2025)
UniPTS: A Unified Framework for Proficient Post-Training Sparsity
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
von: Xie, Jingjing, et al.
Veröffentlicht: (2024)
LightMotion: A Light and Tuning-free Method for Simulating Camera Motion in Video Generation
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
INF-LLaVA: Dual-perspective Perception for High-Resolution Multimodal Large Language Model
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
Pseudo-Label Quality Decoupling and Correction for Semi-Supervised Instance Segmentation
von: Lin, Jianghang, et al.
Veröffentlicht: (2025)
von: Lin, Jianghang, et al.
Veröffentlicht: (2025)
DS$^2$Net: Detail-Semantic Deep Supervision Network for Medical Image Segmentation
von: Huang, Zhaohong, et al.
Veröffentlicht: (2025)
von: Huang, Zhaohong, et al.
Veröffentlicht: (2025)
Dual3D: Efficient and Consistent Text-to-3D Generation with Dual-mode Multi-view Latent Diffusion
von: Li, Xinyang, et al.
Veröffentlicht: (2024)
von: Li, Xinyang, et al.
Veröffentlicht: (2024)
AnySR: Realizing Image Super-Resolution as Any-Scale, Any-Resource
von: Zhan, Wengyi, et al.
Veröffentlicht: (2024)
von: Zhan, Wengyi, et al.
Veröffentlicht: (2024)
FastVGGT: Training-Free Acceleration of Visual Geometry Transformer
von: Shen, You, et al.
Veröffentlicht: (2025)
von: Shen, You, et al.
Veröffentlicht: (2025)
Discover, Segment, and Select: A Progressive Mechanism for Zero-shot Camouflaged Object Segmentation
von: Yang, Yilong, et al.
Veröffentlicht: (2026)
von: Yang, Yilong, et al.
Veröffentlicht: (2026)
M4-BLIP: Advancing Multi-Modal Media Manipulation Detection through Face-Enhanced Local Analysis
von: Wu, Hang, et al.
Veröffentlicht: (2025)
von: Wu, Hang, et al.
Veröffentlicht: (2025)
An Efficient and Mixed Heterogeneous Model for Image Restoration
von: Gu, Yubin, et al.
Veröffentlicht: (2025)
von: Gu, Yubin, et al.
Veröffentlicht: (2025)
S$^2$Teacher: Step-by-step Teacher for Sparsely Annotated Oriented Object Detection
von: Lin, Yu, et al.
Veröffentlicht: (2025)
von: Lin, Yu, et al.
Veröffentlicht: (2025)
Director3D: Real-world Camera Trajectory and 3D Scene Generation from Text
von: Li, Xinyang, et al.
Veröffentlicht: (2024)
von: Li, Xinyang, et al.
Veröffentlicht: (2024)
PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation
von: Ke, Shuyan, et al.
Veröffentlicht: (2026)
von: Ke, Shuyan, et al.
Veröffentlicht: (2026)
Generate Aligned Anomaly: Region-Guided Few-Shot Anomaly Image-Mask Pair Synthesis for Industrial Inspection
von: Lu, Yilin, et al.
Veröffentlicht: (2025)
von: Lu, Yilin, et al.
Veröffentlicht: (2025)
GOI: Find 3D Gaussians of Interest with an Optimizable Open-vocabulary Semantic-space Hyperplane
von: Qu, Yansong, et al.
Veröffentlicht: (2024)
von: Qu, Yansong, et al.
Veröffentlicht: (2024)
AccDiffusion v2: Towards More Accurate Higher-Resolution Diffusion Extrapolation
von: Lin, Zhihang, et al.
Veröffentlicht: (2024)
von: Lin, Zhihang, et al.
Veröffentlicht: (2024)
Purifying, Labeling, and Utilizing: A High-Quality Pipeline for Small Object Detection
von: Wang, Siwei, et al.
Veröffentlicht: (2025)
von: Wang, Siwei, et al.
Veröffentlicht: (2025)
I2EBench: A Comprehensive Benchmark for Instruction-based Image Editing
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
von: Ma, Yiwei, et al.
Veröffentlicht: (2024)
CamoTeacher: Dual-Rotation Consistency Learning for Semi-Supervised Camouflaged Object Detection
von: Lai, Xunfa, et al.
Veröffentlicht: (2024)
von: Lai, Xunfa, et al.
Veröffentlicht: (2024)
Drag Your Gaussian: Effective Drag-Based Editing with Score Distillation for 3D Gaussian Splatting
von: Qu, Yansong, et al.
Veröffentlicht: (2025)
von: Qu, Yansong, et al.
Veröffentlicht: (2025)
Learning Image Demoireing from Unpaired Real Data
von: Zhong, Yunshan, et al.
Veröffentlicht: (2024)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2024)
IPDN: Image-enhanced Prompt Decoding Network for 3D Referring Expression Segmentation
von: Chen, Qi, et al.
Veröffentlicht: (2025)
von: Chen, Qi, et al.
Veröffentlicht: (2025)
An Efficient Aerial Image Detection with Variable Receptive Fields
von: Wenbin, Liu
Veröffentlicht: (2025)
von: Wenbin, Liu
Veröffentlicht: (2025)
MIHBench: Benchmarking and Mitigating Multi-Image Hallucinations in Multimodal Large Language Models
von: Li, Jiale, et al.
Veröffentlicht: (2025)
von: Li, Jiale, et al.
Veröffentlicht: (2025)
Rotated Multi-Scale Interaction Network for Referring Remote Sensing Image Segmentation
von: Liu, Sihan, et al.
Veröffentlicht: (2023)
von: Liu, Sihan, et al.
Veröffentlicht: (2023)
Referring Industrial Anomaly Segmentation
von: Yue, Pengfei, et al.
Veröffentlicht: (2026)
von: Yue, Pengfei, et al.
Veröffentlicht: (2026)
Depth-Guided Semi-Supervised Instance Segmentation
von: Chen, Xin, et al.
Veröffentlicht: (2024)
von: Chen, Xin, et al.
Veröffentlicht: (2024)
3D-DRES: Detailed 3D Referring Expression Segmentation
von: Chen, Qi, et al.
Veröffentlicht: (2026)
von: Chen, Qi, et al.
Veröffentlicht: (2026)
BadSR: Stealthy Label Backdoor Attacks on Image Super-Resolution
von: Guo, Ji, et al.
Veröffentlicht: (2025)
von: Guo, Ji, et al.
Veröffentlicht: (2025)
Efficient Transformer for High Resolution Image Motion Deblurring
von: Akmaral, Amanturdieva, et al.
Veröffentlicht: (2025)
von: Akmaral, Amanturdieva, et al.
Veröffentlicht: (2025)
Active-SAOOD: Active Sparsely Annotated Oriented Object Detection in Remote Sensing Images
von: Lin, Yu, et al.
Veröffentlicht: (2026)
von: Lin, Yu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Inter2Former: Dynamic Hybrid Attention for Efficient High-Precision Interactive
von: Huang, You, et al.
Veröffentlicht: (2025) -
Evolving, Not Training: Zero-Shot Reasoning Segmentation via Evolutionary Prompting
von: Ye, Kai, et al.
Veröffentlicht: (2025) -
FocSAM: Delving Deeply into Focused Objects in Segmenting Anything
von: Huang, You, et al.
Veröffentlicht: (2024) -
HRSeg: High-Resolution Visual Perception and Enhancement for Reasoning Segmentation
von: Lin, Weihuang, et al.
Veröffentlicht: (2025) -
Evolving High-Quality Rendering and Reconstruction in a Unified Framework with Contribution-Adaptive Regularization
von: Shen, You, et al.
Veröffentlicht: (2025)