Insect-Foundation: A Foundation Model and Large Multimodal Dataset for Vision-Language Insect Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Truong, Thanh-Dat, Nguyen, Hoang-Quan, Nguyen, Xuan-Bac, Dowling, Ashley, Li, Xin, Luu, Khoa |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Insect-Foundation: A Foundation Model and Large-scale 1M Dataset for Visual Insect Understanding
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2023)
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2023)
BRAIN: Bias-Mitigation Continual Learning Approach to Vision-Brain Understanding
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2025)
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2025)
Multi-view Action Recognition via Directed Gromov-Wasserstein Discrepancy
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2024)
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2024)
ED-SAM: An Efficient Diffusion Sampling Approach to Domain Generalization in Vision-Language Foundation Models
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2024)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2024)
Cross-view Action Recognition Understanding From Exocentric to Egocentric Perspective
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
BIMA: Bijective Maximum Likelihood Learning Approach to Hallucination Prediction and Mitigation in Large Vision-Language Models
von: Tran, Huu-Thien, et al.
Veröffentlicht: (2025)
von: Tran, Huu-Thien, et al.
Veröffentlicht: (2025)
Hierarchical Quantum Control Gates for Functional MRI Understanding
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
Quantum-Brain: Quantum-Inspired Neural Network Approach to Vision-Brain Understanding
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2024)
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2024)
Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation
von: Nguyen, Huu Tien, et al.
Veröffentlicht: (2025)
von: Nguyen, Huu Tien, et al.
Veröffentlicht: (2025)
Quantum Visual Feature Encoding Revisited
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
Brainformer: Mimic Human Visual Brain Functions to Machine Vision Models via fMRI
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2023)
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2023)
Directed-Tokens: A Robust Multi-Modality Alignment Approach to Large Language-Vision Models
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2025)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2025)
MANGO: Multimodal Attention-based Normalizing Flow Approach to Fusion Learning
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2025)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2025)
COBRA: A Continual Learning Approach to Vision-Brain Understanding
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
Quantum Vision Clustering
von: Nguyen, Xuan Bac, et al.
Veröffentlicht: (2023)
von: Nguyen, Xuan Bac, et al.
Veröffentlicht: (2023)
A Novel Dataset for Video-Based Neurodivergent Classification Leveraging Extra-Stimulatory Behavior
von: Serna-Aguilera, Manuel, et al.
Veröffentlicht: (2024)
von: Serna-Aguilera, Manuel, et al.
Veröffentlicht: (2024)
QuPAINT: Physics-Aware Instruction Tuning Approach to Quantum Material Discovery
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2026)
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2026)
OpenQlaw: An Agentic AI Assistant for Analysis of 2D Quantum Materials
von: Pandey, Sankalp, et al.
Veröffentlicht: (2026)
von: Pandey, Sankalp, et al.
Veröffentlicht: (2026)
QMoE: A Quantum Mixture of Experts Framework for Scalable Quantum Neural Networks
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2025)
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2025)
$ϕ$-DPO: Fairness Direct Preference Optimization Approach to Continual Learning in Large Multimodal Models
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2026)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2026)
QClusformer: A Quantum Transformer-based Framework for Unsupervised Visual Clustering
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
FALCON: Fairness Learning via Contrastive Attention Approach to Continual Semantic Scene Understanding
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
$φ$-Adapt: A Physics-Informed Adaptation Learning Approach to 2D Quantum Material Discovery
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2025)
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2025)
QLAM: A Quantum Long-Attention Memory Approach to Long-Sequence Token Modeling
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2026)
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2026)
Towards Robust and Fair Vision Learning in Open-World Environments
von: Truong, Thanh-Dat
Veröffentlicht: (2024)
von: Truong, Thanh-Dat
Veröffentlicht: (2024)
CLIFF: Continual Learning for Incremental Flake Features in 2D Material Identification
von: Pandey, Sankalp, et al.
Veröffentlicht: (2025)
von: Pandey, Sankalp, et al.
Veröffentlicht: (2025)
BRACTIVE: A Brain Activation Approach to Human Visual Brain Learning
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2024)
HIG: Hierarchical Interlacement Graph Approach to Scene Graph Generation in Video Understanding
von: Nguyen, Trong-Thuan, et al.
Veröffentlicht: (2023)
von: Nguyen, Trong-Thuan, et al.
Veröffentlicht: (2023)
EAGLE: Efficient Adaptive Geometry-based Learning in Cross-view Understanding
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2024)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2024)
Deep-Wide Learning Assistance for Insect Pest Classification
von: Nguyen, Toan, et al.
Veröffentlicht: (2024)
von: Nguyen, Toan, et al.
Veröffentlicht: (2024)
CONDA: Continual Unsupervised Domain Adaptation Learning in Visual Perception for Self-Driving Cars
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2022)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2022)
Seeing Through the Tool: A Controlled Benchmark for Occlusion Robustness in Foundation Segmentation Models
von: Ho, Nhan, et al.
Veröffentlicht: (2026)
von: Ho, Nhan, et al.
Veröffentlicht: (2026)
Temporal-Oriented Recipe for Transferring Large Vision-Language Model to Video Understanding
von: Nguyen, Thong, et al.
Veröffentlicht: (2025)
von: Nguyen, Thong, et al.
Veröffentlicht: (2025)
LeafNet: A Large-Scale Dataset and Comprehensive Benchmark for Foundational Vision-Language Understanding of Plant Diseases
von: Quoc, Khang Nguyen, et al.
Veröffentlicht: (2026)
von: Quoc, Khang Nguyen, et al.
Veröffentlicht: (2026)
TinyGiantVLM: A Lightweight Vision-Language Architecture for Spatial Reasoning under Resource Constraints
von: Ly, Vinh-Thuan, et al.
Veröffentlicht: (2025)
von: Ly, Vinh-Thuan, et al.
Veröffentlicht: (2025)
DEFEND: A Large-scale 1M Dataset and Foundation Model for Tobacco Addiction Prevention
von: Chappa, Naga VS Raviteja, et al.
Veröffentlicht: (2025)
von: Chappa, Naga VS Raviteja, et al.
Veröffentlicht: (2025)
ViOCRVQA: Novel Benchmark Dataset and Vision Reader for Visual Question Answering by Understanding Vietnamese Text in Images
von: Pham, Huy Quang, et al.
Veröffentlicht: (2024)
von: Pham, Huy Quang, et al.
Veröffentlicht: (2024)
Sampling Foundational Transformer: A Theoretical Perspective
von: Nguyen, Viet Anh, et al.
Veröffentlicht: (2024)
von: Nguyen, Viet Anh, et al.
Veröffentlicht: (2024)
Guiding Noisy Label Conditional Diffusion Models with Score-based Discriminator Correction
von: Cong, Dat Nguyen, et al.
Veröffentlicht: (2025)
von: Cong, Dat Nguyen, et al.
Veröffentlicht: (2025)
Towards multi-modal forgery representation learning for AI-generated video detection and localization
von: Le, Dat, et al.
Veröffentlicht: (2026)
von: Le, Dat, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Insect-Foundation: A Foundation Model and Large-scale 1M Dataset for Visual Insect Understanding
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2023) -
BRAIN: Bias-Mitigation Continual Learning Approach to Vision-Brain Understanding
von: Nguyen, Xuan-Bac, et al.
Veröffentlicht: (2025) -
Multi-view Action Recognition via Directed Gromov-Wasserstein Discrepancy
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2024) -
ED-SAM: An Efficient Diffusion Sampling Approach to Domain Generalization in Vision-Language Foundation Models
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2024) -
Cross-view Action Recognition Understanding From Exocentric to Egocentric Perspective
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)