PinPoint: Prompting with Informative Interior Points
Fuente:
arXiv
Saved in:
| Main Authors: | Sadeghi, Pouya, He, Shawn, Vela, Pedro Pablo Guerrero, Thomas, C., Wong, Alex, Rambhatla, Sirisha |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PinPoint: Evaluation of Composed Image Retrieval with Explicit Negatives, Multi-Image Queries, and Paraphrase Testing
by: Mahadev, Rohan, et al.
Published: (2026)
by: Mahadev, Rohan, et al.
Published: (2026)
PinPoint3D: Fine-Grained 3D Part Segmentation from a Few Clicks
by: Zhang, Bojun, et al.
Published: (2025)
by: Zhang, Bojun, et al.
Published: (2025)
Zero-Shot Object Re-Identification in Egocentric Kitchen Videos via Multi-Stage SAM3 Feature Fusion
by: Klepachevskyi, Dmytro, et al.
Published: (2026)
by: Klepachevskyi, Dmytro, et al.
Published: (2026)
LOCATEdit: Graph Laplacian Optimized Cross Attention for Localized Text-Guided Image Editing
by: Soni, Achint, et al.
Published: (2025)
by: Soni, Achint, et al.
Published: (2025)
PointAD: Comprehending 3D Anomalies from Points and Pixels for Zero-shot 3D Anomaly Detection
by: Zhou, Qihang, et al.
Published: (2024)
by: Zhou, Qihang, et al.
Published: (2024)
Talking Points: Describing and Localizing Pixels
by: Rusanovsky, Matan, et al.
Published: (2025)
by: Rusanovsky, Matan, et al.
Published: (2025)
Avatar4D: Synthesizing Domain-Specific 4D Humans for Real-World Pose Estimation
by: Bright, Jerrin, et al.
Published: (2025)
by: Bright, Jerrin, et al.
Published: (2025)
Domain-Guided Masked Autoencoders for Unique Player Identification
by: Balaji, Bavesh, et al.
Published: (2024)
by: Balaji, Bavesh, et al.
Published: (2024)
PointLLM: Empowering Large Language Models to Understand Point Clouds
by: Xu, Runsen, et al.
Published: (2023)
by: Xu, Runsen, et al.
Published: (2023)
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding
by: Sun, Qi, et al.
Published: (2024)
by: Sun, Qi, et al.
Published: (2024)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
by: Chung, Jiwan, et al.
Published: (2025)
by: Chung, Jiwan, et al.
Published: (2025)
Masked Point-Entity Contrast for Open-Vocabulary 3D Scene Understanding
by: Wang, Yan, et al.
Published: (2025)
by: Wang, Yan, et al.
Published: (2025)
More Text, Less Point: Towards 3D Data-Efficient Point-Language Understanding
by: Tang, Yuan, et al.
Published: (2024)
by: Tang, Yuan, et al.
Published: (2024)
Read Anywhere Pointed: Layout-aware GUI Screen Reading with Tree-of-Lens Grounding
by: Fan, Yue, et al.
Published: (2024)
by: Fan, Yue, et al.
Published: (2024)
FlairGPT: Repurposing LLMs for Interior Designs
by: Littlefair, Gabrielle, et al.
Published: (2025)
by: Littlefair, Gabrielle, et al.
Published: (2025)
Phoneme-Level Visual Speech Recognition via Point-Visual Fusion and Language Model Reconstruction
by: Teng, Matthew Kit Khinn, et al.
Published: (2025)
by: Teng, Matthew Kit Khinn, et al.
Published: (2025)
UPP: Unified Point-Level Prompting for Robust Point Cloud Analysis
by: Ai, Zixiang, et al.
Published: (2025)
by: Ai, Zixiang, et al.
Published: (2025)
Co-Layout: LLM-driven Co-optimization for Interior Layout
by: Xiang, Chucheng, et al.
Published: (2025)
by: Xiang, Chucheng, et al.
Published: (2025)
OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models
by: Yu, Wenwen, et al.
Published: (2025)
by: Yu, Wenwen, et al.
Published: (2025)
PromptMRG: Diagnosis-Driven Prompts for Medical Report Generation
by: Jin, Haibo, et al.
Published: (2023)
by: Jin, Haibo, et al.
Published: (2023)
The Devil is in the Prompts: Retrieval-Augmented Prompt Optimization for Text-to-Video Generation
by: Gao, Bingjie, et al.
Published: (2025)
by: Gao, Bingjie, et al.
Published: (2025)
Joint Extraction Matters: Prompt-Based Visual Question Answering for Multi-Field Document Information Extraction
by: Loem, Mengsay, et al.
Published: (2025)
by: Loem, Mengsay, et al.
Published: (2025)
LADR: Locality-Aware Dynamic Rescue for Efficient Text-to-Image Generation with Diffusion Large Language Models
by: Wang, Chenglin, et al.
Published: (2026)
by: Wang, Chenglin, et al.
Published: (2026)
Point-PRC: A Prompt Learning Based Regulation Framework for Generalizable Point Cloud Analysis
by: Sun, Hongyu, et al.
Published: (2024)
by: Sun, Hongyu, et al.
Published: (2024)
Prompting4Debugging: Red-Teaming Text-to-Image Diffusion Models by Finding Problematic Prompts
by: Chin, Zhi-Yi, et al.
Published: (2023)
by: Chin, Zhi-Yi, et al.
Published: (2023)
LLM-PCGC: Large Language Model-based Point Cloud Geometry Compression
by: Ye, Yuqi, et al.
Published: (2024)
by: Ye, Yuqi, et al.
Published: (2024)
Optimizing Prompts for Text-to-Image Generation
by: Hao, Yaru, et al.
Published: (2022)
by: Hao, Yaru, et al.
Published: (2022)
Semantic-aware SAM for Point-Prompted Instance Segmentation
by: Wei, Zhaoyang, et al.
Published: (2023)
by: Wei, Zhaoyang, et al.
Published: (2023)
Point Prompting: Counterfactual Tracking with Video Diffusion Models
by: Shrivastava, Ayush, et al.
Published: (2025)
by: Shrivastava, Ayush, et al.
Published: (2025)
Guiding Medical Vision-Language Models with Explicit Visual Prompts: Framework Design and Comprehensive Exploration of Prompt Variations
by: Zhu, Kangyu, et al.
Published: (2025)
by: Zhu, Kangyu, et al.
Published: (2025)
Pipe Reconstruction from Point Cloud Data
by: Alex, Antje, et al.
Published: (2025)
by: Alex, Antje, et al.
Published: (2025)
Text Prompt Injection of Vision Language Models
by: Zhu, Ruizhe
Published: (2025)
by: Zhu, Ruizhe
Published: (2025)
Fast Prompt Alignment for Text-to-Image Generation
by: Mrini, Khalil, et al.
Published: (2024)
by: Mrini, Khalil, et al.
Published: (2024)
Dual-branch Prompting for Multimodal Machine Translation
by: Wang, Jie, et al.
Published: (2025)
by: Wang, Jie, et al.
Published: (2025)
Adaptive Point-Prompt Tuning: Fine-Tuning Heterogeneous Foundation Models for 3D Point Cloud Analysis
by: Li, Mengke, et al.
Published: (2025)
by: Li, Mengke, et al.
Published: (2025)
Depth as Points: Center Point-based Depth Estimation
by: Tu, Zhiheng, et al.
Published: (2025)
by: Tu, Zhiheng, et al.
Published: (2025)
Prompt Highlighter: Interactive Control for Multi-Modal LLMs
by: Zhang, Yuechen, et al.
Published: (2023)
by: Zhang, Yuechen, et al.
Published: (2023)
Generalizable Prompt Learning of CLIP: A Brief Overview
by: Cui, Fangming, et al.
Published: (2025)
by: Cui, Fangming, et al.
Published: (2025)
Universal Prompt Optimizer for Safe Text-to-Image Generation
by: Wu, Zongyu, et al.
Published: (2024)
by: Wu, Zongyu, et al.
Published: (2024)
Agri-CPJ: A Training-Free Explainable Framework for Agricultural Pest Diagnosis Using Caption-Prompt-Judge and LLM-as-a-Judge
by: Zhang, Wentao, et al.
Published: (2026)
by: Zhang, Wentao, et al.
Published: (2026)
Similar Items
-
PinPoint: Evaluation of Composed Image Retrieval with Explicit Negatives, Multi-Image Queries, and Paraphrase Testing
by: Mahadev, Rohan, et al.
Published: (2026) -
PinPoint3D: Fine-Grained 3D Part Segmentation from a Few Clicks
by: Zhang, Bojun, et al.
Published: (2025) -
Zero-Shot Object Re-Identification in Egocentric Kitchen Videos via Multi-Stage SAM3 Feature Fusion
by: Klepachevskyi, Dmytro, et al.
Published: (2026) -
LOCATEdit: Graph Laplacian Optimized Cross Attention for Localized Text-Guided Image Editing
by: Soni, Achint, et al.
Published: (2025) -
PointAD: Comprehending 3D Anomalies from Points and Pixels for Zero-shot 3D Anomaly Detection
by: Zhou, Qihang, et al.
Published: (2024)