SonoSelect: Efficient Ultrasound Perception via Active Probe Exploration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yixin, Hou, Yunzhong, Li, Longqi, Qin, Zhenyue, Liu, Yang, Yao, Yue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HandCraft: Anatomically Correct Restoration of Malformed Hands in Diffusion Generated Images
von: Qin, Zhenyue, et al.
Veröffentlicht: (2024)
von: Qin, Zhenyue, et al.
Veröffentlicht: (2024)
Learning Camera Movement Control from Real-World Drone Videos
von: Hou, Yunzhong, et al.
Veröffentlicht: (2024)
von: Hou, Yunzhong, et al.
Veröffentlicht: (2024)
Decision-Driven Semantic Object Exploration for Legged Robots via Confidence-Calibrated Perception and Topological Subgoal Selection
von: Zhao, Guoyang, et al.
Veröffentlicht: (2025)
von: Zhao, Guoyang, et al.
Veröffentlicht: (2025)
Authentic Emotion Mapping: Benchmarking Facial Expressions in Real News
von: Zhang, Qixuan, et al.
Veröffentlicht: (2024)
von: Zhang, Qixuan, et al.
Veröffentlicht: (2024)
ActFormer: Scalable Collaborative Perception via Active Queries
von: Huang, Suozhi, et al.
Veröffentlicht: (2024)
von: Huang, Suozhi, et al.
Veröffentlicht: (2024)
MP5: A Multi-modal Open-ended Embodied System in Minecraft via Active Perception
von: Qin, Yiran, et al.
Veröffentlicht: (2023)
von: Qin, Yiran, et al.
Veröffentlicht: (2023)
HeatV2X: Scalable Heterogeneous Collaborative Perception via Efficient Alignment and Interaction
von: Zhao, Yueran, et al.
Veröffentlicht: (2025)
von: Zhao, Yueran, et al.
Veröffentlicht: (2025)
Active Visual Perception: Opportunities and Challenges
von: Li, Yian, et al.
Veröffentlicht: (2025)
von: Li, Yian, et al.
Veröffentlicht: (2025)
Mind the Rarities: Can Rare Skin Diseases Be Reliably Diagnosed via Diagnostic Reasoning?
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
Visual Prompting in LLMs for Enhancing Emotion Recognition
von: Zhang, Qixuan, et al.
Veröffentlicht: (2024)
von: Zhang, Qixuan, et al.
Veröffentlicht: (2024)
ADAptation: Reconstruction-based Unsupervised Active Learning for Breast Ultrasound Diagnosis
von: Duan, Yaofei, et al.
Veröffentlicht: (2025)
von: Duan, Yaofei, et al.
Veröffentlicht: (2025)
Towards High-Fidelity CAD Generation via LLM-Driven Program Generation and Text-Based B-Rep Primitive Grounding
von: Li, Jiahao, et al.
Veröffentlicht: (2026)
von: Li, Jiahao, et al.
Veröffentlicht: (2026)
Evaluating Time Awareness and Cross-modal Active Perception of Large Models via 4D Escape Room Task
von: Dong, Yurui, et al.
Veröffentlicht: (2026)
von: Dong, Yurui, et al.
Veröffentlicht: (2026)
Perceive, Verify and Understand Long Video: Multi-Granular Perception and Active Verification via Interactive Agents
von: Li, Jiahua, et al.
Veröffentlicht: (2025)
von: Li, Jiahua, et al.
Veröffentlicht: (2025)
JAQ: Joint Efficient Architecture Design and Low-Bit Quantization with Hardware-Software Co-Exploration
von: Wang, Mingzi, et al.
Veröffentlicht: (2025)
von: Wang, Mingzi, et al.
Veröffentlicht: (2025)
SCSA: Exploring the Synergistic Effects Between Spatial and Channel Attention
von: Si, Yunzhong, et al.
Veröffentlicht: (2024)
von: Si, Yunzhong, et al.
Veröffentlicht: (2024)
ReCAD: Reinforcement Learning Enhanced Parametric CAD Model Generation with Vision-Language Models
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
UMind-VL: A Generalist Ultrasound Vision-Language Model for Unified Grounded Perception and Comprehensive Interpretation
von: Chen, Dengbo, et al.
Veröffentlicht: (2025)
von: Chen, Dengbo, et al.
Veröffentlicht: (2025)
Pursuing Minimal Sufficiency in Spatial Reasoning
von: Guo, Yejie, et al.
Veröffentlicht: (2025)
von: Guo, Yejie, et al.
Veröffentlicht: (2025)
Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming
von: Zhou, Yue, et al.
Veröffentlicht: (2026)
von: Zhou, Yue, et al.
Veröffentlicht: (2026)
ActiView: Evaluating Active Perception Ability for Multimodal Large Language Models
von: Wang, Ziyue, et al.
Veröffentlicht: (2024)
von: Wang, Ziyue, et al.
Veröffentlicht: (2024)
Plane Geometry Problem Solving with Multi-modal Reasoning: A Survey
von: Cho, Seunghyuk, et al.
Veröffentlicht: (2025)
von: Cho, Seunghyuk, et al.
Veröffentlicht: (2025)
GeoDANO: Geometric VLM with Domain Agnostic Vision Encoder
von: Cho, Seunghyuk, et al.
Veröffentlicht: (2025)
von: Cho, Seunghyuk, et al.
Veröffentlicht: (2025)
FailureAtlas:Mapping the Failure Landscape of T2I Models via Active Exploration
von: Chen, Muxi, et al.
Veröffentlicht: (2025)
von: Chen, Muxi, et al.
Veröffentlicht: (2025)
ComGS: Efficient 3D Object-Scene Composition via Surface Octahedral Probes
von: Gao, Jian, et al.
Veröffentlicht: (2025)
von: Gao, Jian, et al.
Veröffentlicht: (2025)
Spatial-VLN: Zero-Shot Vision-and-Language Navigation With Explicit Spatial Perception and Exploration
von: Yue, Lu, et al.
Veröffentlicht: (2026)
von: Yue, Lu, et al.
Veröffentlicht: (2026)
Effective Training Data Synthesis for Improving MLLM Chart Understanding
von: Yang, Yuwei, et al.
Veröffentlicht: (2025)
von: Yang, Yuwei, et al.
Veröffentlicht: (2025)
ESAM++: Efficient Online 3D Perception on the Edge
von: Liu, Qin, et al.
Veröffentlicht: (2026)
von: Liu, Qin, et al.
Veröffentlicht: (2026)
Learn 3D VQA Better with Active Selection and Reannotation
von: Zhou, Shengli, et al.
Veröffentlicht: (2025)
von: Zhou, Shengli, et al.
Veröffentlicht: (2025)
Move to Understand a 3D Scene: Bridging Visual Grounding and Exploration for Efficient and Versatile Embodied Navigation
von: Zhu, Ziyu, et al.
Veröffentlicht: (2025)
von: Zhu, Ziyu, et al.
Veröffentlicht: (2025)
Active-O3: Empowering Multimodal Large Language Models with Active Perception via GRPO
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
PolarMAE: Efficient Fetal Ultrasound Pre-training via Semantic Screening and Polar-Guided Masking
von: Lv, Meng, et al.
Veröffentlicht: (2026)
von: Lv, Meng, et al.
Veröffentlicht: (2026)
ESA: Annotation-Efficient Active Learning for Semantic Segmentation
von: Ge, Jinchao, et al.
Veröffentlicht: (2024)
von: Ge, Jinchao, et al.
Veröffentlicht: (2024)
Seek-CAD: A Self-refined Generative Modeling for 3D Parametric CAD Using Local Inference via DeepSeek
von: Li, Xueyang, et al.
Veröffentlicht: (2025)
von: Li, Xueyang, et al.
Veröffentlicht: (2025)
Scaling Laws for Deepfake Detection
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
TASL-Net: Tri-Attention Selective Learning Network for Intelligent Diagnosis of Bimodal Ultrasound Video
von: Zhao, Chengqian, et al.
Veröffentlicht: (2024)
von: Zhao, Chengqian, et al.
Veröffentlicht: (2024)
Annotation-Efficient Polyp Segmentation via Active Learning
von: Huang, Duojun, et al.
Veröffentlicht: (2024)
von: Huang, Duojun, et al.
Veröffentlicht: (2024)
GeoExplorer: Active Geo-localization with Curiosity-Driven Exploration
von: Mi, Li, et al.
Veröffentlicht: (2025)
von: Mi, Li, et al.
Veröffentlicht: (2025)
ZoomEarth: Active Perception for Ultra-High-Resolution Geospatial Vision-Language Tasks
von: Liu, Ruixun, et al.
Veröffentlicht: (2025)
von: Liu, Ruixun, et al.
Veröffentlicht: (2025)
GeoVista: Visually Grounded Active Perception for Ultra-High-Resolution Remote Sensing Understanding
von: Zhu, Jiashun, et al.
Veröffentlicht: (2026)
von: Zhu, Jiashun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
HandCraft: Anatomically Correct Restoration of Malformed Hands in Diffusion Generated Images
von: Qin, Zhenyue, et al.
Veröffentlicht: (2024) -
Learning Camera Movement Control from Real-World Drone Videos
von: Hou, Yunzhong, et al.
Veröffentlicht: (2024) -
Decision-Driven Semantic Object Exploration for Legged Robots via Confidence-Calibrated Perception and Topological Subgoal Selection
von: Zhao, Guoyang, et al.
Veröffentlicht: (2025) -
Authentic Emotion Mapping: Benchmarking Facial Expressions in Real News
von: Zhang, Qixuan, et al.
Veröffentlicht: (2024) -
ActFormer: Scalable Collaborative Perception via Active Queries
von: Huang, Suozhi, et al.
Veröffentlicht: (2024)