Is Nano Banana Pro a Low-Level Vision All-Rounder? A Comprehensive Evaluation on 14 Tasks and 40 Datasets
Fuente:
arXiv
Saved in:
| Main Authors: | Zuo, Jialong, Deng, Haoyou, Zhou, Hanyu, Zhu, Jiaxin, Zhang, Yicheng, Zhang, Yiwei, Yan, Yongxin, Huang, Kaixing, Chen, Weisen, Deng, Yongtai, Jin, Rui, Sang, Nong, Gao, Changxin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VideoLucy: Deep Memory Backtracking for Long Video Understanding
by: Zuo, Jialong, et al.
Published: (2025)
by: Zuo, Jialong, et al.
Published: (2025)
ReID5o: Achieving Omni Multi-modal Person Re-identification in a Single Model
by: Zuo, Jialong, et al.
Published: (2025)
by: Zuo, Jialong, et al.
Published: (2025)
DenseGRPO: From Sparse to Dense Reward for Flow Matching Model Alignment
by: Deng, Haoyou, et al.
Published: (2026)
by: Deng, Haoyou, et al.
Published: (2026)
Partial Forward Blocking: A Novel Data Pruning Paradigm for Lossless Training Acceleration
by: Wu, Dongyue, et al.
Published: (2025)
by: Wu, Dongyue, et al.
Published: (2025)
Cross-video Identity Correlating for Person Re-identification Pre-training
by: Zuo, Jialong, et al.
Published: (2024)
by: Zuo, Jialong, et al.
Published: (2024)
PLIP: Language-Image Pre-training for Person Representation Learning
by: Zuo, Jialong, et al.
Published: (2023)
by: Zuo, Jialong, et al.
Published: (2023)
UFineBench: Towards Text-based Person Retrieval with Ultra-fine Granularity
by: Zuo, Jialong, et al.
Published: (2023)
by: Zuo, Jialong, et al.
Published: (2023)
Learning Unpaired Image Dehazing with Physics-based Rehazy Generation
by: Deng, Haoyou, et al.
Published: (2025)
by: Deng, Haoyou, et al.
Published: (2025)
Learning to Tell Apart: Weakly Supervised Video Anomaly Detection via Disentangled Semantic Alignment
by: Yin, Wenti, et al.
Published: (2025)
by: Yin, Wenti, et al.
Published: (2025)
HR-Pro: Point-supervised Temporal Action Localization via Hierarchical Reliability Propagation
by: Zhang, Huaxin, et al.
Published: (2023)
by: Zhang, Huaxin, et al.
Published: (2023)
Spatial Cascaded Clustering and Weighted Memory for Unsupervised Person Re-identification
by: Hong, Jiahao, et al.
Published: (2024)
by: Hong, Jiahao, et al.
Published: (2024)
High-resolution Photo Enhancement in Real-time: A Laplacian Pyramid Network
by: Zhang, Feng, et al.
Published: (2025)
by: Zhang, Feng, et al.
Published: (2025)
Holmes-VAU: Towards Long-term Video Anomaly Understanding at Any Granularity
by: Zhang, Huaxin, et al.
Published: (2024)
by: Zhang, Huaxin, et al.
Published: (2024)
Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM
by: Zhang, Huaxin, et al.
Published: (2024)
by: Zhang, Huaxin, et al.
Published: (2024)
Banana100: Breaking NR-IQA Metrics by 100 Iterative Image Replications with Nano Banana Pro
by: Tang, Kenan, et al.
Published: (2026)
by: Tang, Kenan, et al.
Published: (2026)
Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models
by: Suo, Wei, et al.
Published: (2024)
by: Suo, Wei, et al.
Published: (2024)
REPAIR: Rank Correlation and Noisy Pair Half-replacing with Memory for Noisy Correspondence
by: Zheng, Ruochen, et al.
Published: (2024)
by: Zheng, Ruochen, et al.
Published: (2024)
DFIMat: Decoupled Flexible Interactive Matting in Multi-Person Scenarios
by: Jiao, Siyi, et al.
Published: (2024)
by: Jiao, Siyi, et al.
Published: (2024)
Object-Aware Video Matting with Cross-Frame Guidance
by: Zhang, Huayu, et al.
Published: (2025)
by: Zhang, Huayu, et al.
Published: (2025)
Cohn--Vossen-Type Inequalities for Three-Manifolds and Locally Conformally Flat Manifolds
by: Deng, Jialong
Published: (2026)
by: Deng, Jialong
Published: (2026)
Scalar Curvature in Dimension 4
by: Deng, Jialong
Published: (2025)
by: Deng, Jialong
Published: (2025)
OmniClone: Engineering a Robust, All-Rounder Whole-Body Humanoid Teleoperation System
by: Li, Yixuan, et al.
Published: (2026)
by: Li, Yixuan, et al.
Published: (2026)
Structural Pruning via Spatial-aware Information Redundancy for Semantic Segmentation
by: Wu, Dongyue, et al.
Published: (2024)
by: Wu, Dongyue, et al.
Published: (2024)
An All‐Rounder for NIR‐II Phototheranostics: Well‐Tailored 1064 nm‐Excitable Molecule for Photothermal Combating of Orthotopic Breast Cancer
by: Dingyuan Yan, et al.
Published: (2024)
by: Dingyuan Yan, et al.
Published: (2024)
An All‐Rounder for NIR‐II Phototheranostics: Well‐Tailored 1064 nm‐Excitable Molecule for Photothermal Combating of Orthotopic Breast Cancer
by: Dingyuan Yan, et al.
Published: (2024)
by: Dingyuan Yan, et al.
Published: (2024)
CLIP-guided Prototype Modulating for Few-shot Action Recognition
by: Wang, Xiang, et al.
Published: (2023)
by: Wang, Xiang, et al.
Published: (2023)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
by: Wang, Xiang, et al.
Published: (2025)
by: Wang, Xiang, et al.
Published: (2025)
Discovering Expert-Level Nash Equilibrium Algorithms with Large Language Models
by: Li, Hanyu, et al.
Published: (2025)
by: Li, Hanyu, et al.
Published: (2025)
The Enumerative Geometry and Arithmetic of Banana Nano-Manifolds
by: Bryan, Jim, et al.
Published: (2024)
by: Bryan, Jim, et al.
Published: (2024)
Lookup Table meets Local Laplacian Filter: Pyramid Reconstruction Network for Tone Mapping
by: Zhang, Feng, et al.
Published: (2023)
by: Zhang, Feng, et al.
Published: (2023)
SCTNet: Single-Branch CNN with Transformer Semantic Information for Real-Time Segmentation
by: Xu, Zhengze, et al.
Published: (2023)
by: Xu, Zhengze, et al.
Published: (2023)
Open-Vocabulary Semantic Segmentation with Image Embedding Balancing
by: Shan, Xiangheng, et al.
Published: (2024)
by: Shan, Xiangheng, et al.
Published: (2024)
Adaptive Prototype Replay for Class Incremental Semantic Segmentation
by: Zhu, Guilin, et al.
Published: (2024)
by: Zhu, Guilin, et al.
Published: (2024)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
by: Wang, Xiang, et al.
Published: (2024)
by: Wang, Xiang, et al.
Published: (2024)
Taming Consistency Distillation for Accelerated Human Image Animation
by: Wang, Xiang, et al.
Published: (2025)
by: Wang, Xiang, et al.
Published: (2025)
MP-Mat: A 3D-and-Instance-Aware Human Matting and Editing Framework with Multiplane Representation
by: Jiao, Siyi, et al.
Published: (2025)
by: Jiao, Siyi, et al.
Published: (2025)
A Safety Report on GPT-5.2, Gemini 3 Pro, Qwen3-VL, Grok 4.1 Fast, Nano Banana Pro, and Seedream 4.5
by: Ma, Xingjun, et al.
Published: (2026)
by: Ma, Xingjun, et al.
Published: (2026)
Reward Guidance for Reinforcement Learning Tasks Based on Large Language Models: The LMGT Framework
by: Deng, Yongxin, et al.
Published: (2024)
by: Deng, Yongxin, et al.
Published: (2024)
Decomposition of general grain boundaries
by: Wan, Wei, et al.
Published: (2025)
by: Wan, Wei, et al.
Published: (2025)
Will AI Trade? A Computational Inversion of the No-Trade Theorem
by: Li, Hanyu, et al.
Published: (2025)
by: Li, Hanyu, et al.
Published: (2025)
Similar Items
-
VideoLucy: Deep Memory Backtracking for Long Video Understanding
by: Zuo, Jialong, et al.
Published: (2025) -
ReID5o: Achieving Omni Multi-modal Person Re-identification in a Single Model
by: Zuo, Jialong, et al.
Published: (2025) -
DenseGRPO: From Sparse to Dense Reward for Flow Matching Model Alignment
by: Deng, Haoyou, et al.
Published: (2026) -
Partial Forward Blocking: A Novel Data Pruning Paradigm for Lossless Training Acceleration
by: Wu, Dongyue, et al.
Published: (2025) -
Cross-video Identity Correlating for Person Re-identification Pre-training
by: Zuo, Jialong, et al.
Published: (2024)