UniFine: A Unified and Fine-grained Approach for Zero-shot Vision-Language Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Zhecan, Sun, Rui, You, Haoxuan, Codella, Noel, Chang, Kai-Wei, Chang, Shih-Fu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
IdealGPT: Iteratively Decomposing Vision and Language Reasoning via Large Language Models
von: You, Haoxuan, et al.
Veröffentlicht: (2023)
von: You, Haoxuan, et al.
Veröffentlicht: (2023)
FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding
von: Faure, Gueter Josmy, et al.
Veröffentlicht: (2026)
von: Faure, Gueter Josmy, et al.
Veröffentlicht: (2026)
Detecting Multimodal Situations with Insufficient Context and Abstaining from Baseless Predictions
von: Liu, Junzhang, et al.
Veröffentlicht: (2024)
von: Liu, Junzhang, et al.
Veröffentlicht: (2024)
Understanding Zero-shot Rare Word Recognition Improvements Through LLM Integration
von: Wang, Haoxuan
Veröffentlicht: (2025)
von: Wang, Haoxuan
Veröffentlicht: (2025)
JourneyBench: A Challenging One-Stop Vision-Language Understanding Benchmark of Generated Images
von: Wang, Zhecan, et al.
Veröffentlicht: (2024)
von: Wang, Zhecan, et al.
Veröffentlicht: (2024)
Large Margin Prototypical Network for Few-shot Relation Classification with Fine-grained Features
von: Fan, Miao, et al.
Veröffentlicht: (2024)
von: Fan, Miao, et al.
Veröffentlicht: (2024)
KPEval: Towards Fine-Grained Semantic-Based Keyphrase Evaluation
von: Wu, Di, et al.
Veröffentlicht: (2023)
von: Wu, Di, et al.
Veröffentlicht: (2023)
Distilling Fine-grained Sentiment Understanding from Large Language Models
von: Zhang, Yice, et al.
Veröffentlicht: (2024)
von: Zhang, Yice, et al.
Veröffentlicht: (2024)
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding
von: Sun, Qi, et al.
Veröffentlicht: (2024)
von: Sun, Qi, et al.
Veröffentlicht: (2024)
Understanding Fine-grained Distortions in Reports of Scientific Findings
von: Wührl, Amelie, et al.
Veröffentlicht: (2024)
von: Wührl, Amelie, et al.
Veröffentlicht: (2024)
FLRC: Fine-grained Low-Rank Compressor for Efficient LLM Inference
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
Rephrase and Contrast: Fine-Tuning Language Models for Enhanced Understanding of Communication and Computer Networks
von: Wang, Liujianfu, et al.
Veröffentlicht: (2024)
von: Wang, Liujianfu, et al.
Veröffentlicht: (2024)
FGAIF: Aligning Large Vision-Language Models with Fine-grained AI Feedback
von: Jing, Liqiang, et al.
Veröffentlicht: (2024)
von: Jing, Liqiang, et al.
Veröffentlicht: (2024)
CompAlign: Improving Compositional Text-to-Image Generation with a Complex Benchmark and Fine-Grained Feedback
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
Navigating the Nuances: A Fine-grained Evaluation of Vision-Language Navigation
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
ChartHal: A Fine-grained Framework Evaluating Hallucination of Large Vision Language Models in Chart Understanding
von: Wang, Xingqi, et al.
Veröffentlicht: (2025)
von: Wang, Xingqi, et al.
Veröffentlicht: (2025)
IFShip: Interpretable Fine-grained Ship Classification with Domain Knowledge-Enhanced Vision-Language Models
von: Guo, Mingning, et al.
Veröffentlicht: (2024)
von: Guo, Mingning, et al.
Veröffentlicht: (2024)
Fine-grained Hallucination Detection and Editing for Language Models
von: Mishra, Abhika, et al.
Veröffentlicht: (2024)
von: Mishra, Abhika, et al.
Veröffentlicht: (2024)
MM-MATH: Advancing Multimodal Math Evaluation with Process Evaluation and Fine-grained Classification
von: Sun, Kai, et al.
Veröffentlicht: (2024)
von: Sun, Kai, et al.
Veröffentlicht: (2024)
UniEDU: A Unified Language and Vision Assistant for Education Applications
von: Chu, Zhendong, et al.
Veröffentlicht: (2025)
von: Chu, Zhendong, et al.
Veröffentlicht: (2025)
Lyrics: Boosting Fine-grained Language-Vision Alignment and Comprehension via Semantic-aware Visual Objects
von: Lu, Junyu, et al.
Veröffentlicht: (2023)
von: Lu, Junyu, et al.
Veröffentlicht: (2023)
Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs
von: Gu, Yuzhe, et al.
Veröffentlicht: (2025)
von: Gu, Yuzhe, et al.
Veröffentlicht: (2025)
Taming CLIP for Fine-grained and Structured Visual Understanding of Museum Exhibits
von: Balauca, Ada-Astrid, et al.
Veröffentlicht: (2024)
von: Balauca, Ada-Astrid, et al.
Veröffentlicht: (2024)
Improving Fine-grained Visual Understanding in VLMs through Text-Only Training
von: Choi, Dasol, et al.
Veröffentlicht: (2024)
von: Choi, Dasol, et al.
Veröffentlicht: (2024)
Hal-Eval: A Universal and Fine-grained Hallucination Evaluation Framework for Large Vision Language Models
von: Jiang, Chaoya, et al.
Veröffentlicht: (2024)
von: Jiang, Chaoya, et al.
Veröffentlicht: (2024)
BlockPruner: Fine-grained Pruning for Large Language Models
von: Zhong, Longguang, et al.
Veröffentlicht: (2024)
von: Zhong, Longguang, et al.
Veröffentlicht: (2024)
SciPrompt: Knowledge-augmented Prompting for Fine-grained Categorization of Scientific Topics
von: You, Zhiwen, et al.
Veröffentlicht: (2024)
von: You, Zhiwen, et al.
Veröffentlicht: (2024)
Fine-Refine: Iterative Fine-grained Refinement for Mitigating Dialogue Hallucination
von: Chen, Xiangyan, et al.
Veröffentlicht: (2026)
von: Chen, Xiangyan, et al.
Veröffentlicht: (2026)
Attribute Controlled Fine-tuning for Large Language Models: A Case Study on Detoxification
von: Meng, Tao, et al.
Veröffentlicht: (2024)
von: Meng, Tao, et al.
Veröffentlicht: (2024)
CroPrompt: Cross-task Interactive Prompting for Zero-shot Spoken Language Understanding
von: Qin, Libo, et al.
Veröffentlicht: (2024)
von: Qin, Libo, et al.
Veröffentlicht: (2024)
Systematic Analysis for Pretrained Language Model Priming for Parameter-Efficient Fine-tuning
von: Huang, Shih-Cheng, et al.
Veröffentlicht: (2022)
von: Huang, Shih-Cheng, et al.
Veröffentlicht: (2022)
UniSumEval: Towards Unified, Fine-Grained, Multi-Dimensional Summarization Evaluation for LLMs
von: Lee, Yuho, et al.
Veröffentlicht: (2024)
von: Lee, Yuho, et al.
Veröffentlicht: (2024)
FineDialFact: A benchmark for Fine-grained Dialogue Fact Verification
von: Chen, Xiangyan, et al.
Veröffentlicht: (2025)
von: Chen, Xiangyan, et al.
Veröffentlicht: (2025)
Multimodal Fine-grained Context Interaction Graph Modeling for Conversational Speech Synthesis
von: Jia, Zhenqi, et al.
Veröffentlicht: (2025)
von: Jia, Zhenqi, et al.
Veröffentlicht: (2025)
Enhancing Fine-Grained Image Classifications via Cascaded Vision Language Models
von: Wei, Canshi
Veröffentlicht: (2024)
von: Wei, Canshi
Veröffentlicht: (2024)
Fine-grained Gender Control in Machine Translation with Large Language Models
von: Lee, Minwoo, et al.
Veröffentlicht: (2024)
von: Lee, Minwoo, et al.
Veröffentlicht: (2024)
UniEdit: A Unified Knowledge Editing Benchmark for Large Language Models
von: Chen, Qizhou, et al.
Veröffentlicht: (2025)
von: Chen, Qizhou, et al.
Veröffentlicht: (2025)
UniVIE: A Unified Label Space Approach to Visual Information Extraction from Form-like Documents
von: Hu, Kai, et al.
Veröffentlicht: (2024)
von: Hu, Kai, et al.
Veröffentlicht: (2024)
UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs
von: Jiang, Houcheng, et al.
Veröffentlicht: (2026)
von: Jiang, Houcheng, et al.
Veröffentlicht: (2026)
TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models
von: Cai, Mu, et al.
Veröffentlicht: (2024)
von: Cai, Mu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
IdealGPT: Iteratively Decomposing Vision and Language Reasoning via Large Language Models
von: You, Haoxuan, et al.
Veröffentlicht: (2023) -
FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding
von: Faure, Gueter Josmy, et al.
Veröffentlicht: (2026) -
Detecting Multimodal Situations with Insufficient Context and Abstaining from Baseless Predictions
von: Liu, Junzhang, et al.
Veröffentlicht: (2024) -
Understanding Zero-shot Rare Word Recognition Improvements Through LLM Integration
von: Wang, Haoxuan
Veröffentlicht: (2025) -
JourneyBench: A Challenging One-Stop Vision-Language Understanding Benchmark of Generated Images
von: Wang, Zhecan, et al.
Veröffentlicht: (2024)