TaxaAdapter: Vision Taxonomy Models are Key to Fine-grained Image Generation over the Tree of Life
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Khurana, Mridul, Monsefi, Amin Karimi, Lee, Justin, Sawhney, Medha, Carlyn, David, Chae, Julia, Gu, Jianyang, Ramnath, Rajiv, Beery, Sara, Chao, Wei-Lun, Karpatne, Anuj, Zhang, Cheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TaxaDiffusion: Progressively Trained Diffusion Model for Fine-Grained Species Generation
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2025)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2025)
Beyond Loss Guidance: Using PDE Residuals as Spectral Attention in Diffusion Neural Operators
von: Sawhney, Medha, et al.
Veröffentlicht: (2025)
von: Sawhney, Medha, et al.
Veröffentlicht: (2025)
SeamCam: Quantifying Seamless Camouflage via Multi-Cue Visual Detectability
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026)
Controlla: Learning Controllability via Graph-Constrained Latent Geometry
von: Murthy, Jamuna S., et al.
Veröffentlicht: (2026)
von: Murthy, Jamuna S., et al.
Veröffentlicht: (2026)
Frequency-Guided Masking for Enhanced Vision Self-Supervised Learning
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
Finer-Personalization Rank: Fine-Grained Retrieval Examines Identity Preservation for Personalized Generation
von: Kilrain, Connor, et al.
Veröffentlicht: (2025)
von: Kilrain, Connor, et al.
Veröffentlicht: (2025)
DiReCT: Disentangled Regularization of Contrastive Trajectories for Physics-Refined Video Generation
von: Meyarian, Abolfazl, et al.
Veröffentlicht: (2026)
von: Meyarian, Abolfazl, et al.
Veröffentlicht: (2026)
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
CrashFormer: A Multimodal Architecture to Predict the Risk of Crash
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
KnobGen: Controlling the Sophistication of Artwork in Sketch-Based Diffusion Models
von: Navard, Pouyan, et al.
Veröffentlicht: (2024)
von: Navard, Pouyan, et al.
Veröffentlicht: (2024)
A Unified Framework for Forward and Inverse Problems in Subsurface Imaging using Latent Space Translations
von: Gupta, Naveen, et al.
Veröffentlicht: (2024)
von: Gupta, Naveen, et al.
Veröffentlicht: (2024)
Open World Scene Graph Generation using Vision Language Models
von: Dutta, Amartya, et al.
Veröffentlicht: (2025)
von: Dutta, Amartya, et al.
Veröffentlicht: (2025)
Motion Enhanced Multi‐Level Tracker (MEMTrack): A Deep Learning‐Based Approach to Microrobot Tracking in Dense and Low‐Contrast Environments
von: Medha Sawhney, et al.
Veröffentlicht: (2024)
von: Medha Sawhney, et al.
Veröffentlicht: (2024)
Masked LoGoNet: Fast and Accurate 3D Image Analysis for Medical Domain
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
Efficient and Interpretable Information Retrieval for Product Question Answering with Heterogeneous Data
von: Biswas, Biplob, et al.
Veröffentlicht: (2024)
von: Biswas, Biplob, et al.
Veröffentlicht: (2024)
Knowledge-guided Machine Learning: Current Trends and Future Prospects
von: Karpatne, Anuj, et al.
Veröffentlicht: (2024)
von: Karpatne, Anuj, et al.
Veröffentlicht: (2024)
Hierarchical Conditioning of Diffusion Models Using Tree-of-Life for Studying Species Evolution
von: Khurana, Mridul, et al.
Veröffentlicht: (2024)
von: Khurana, Mridul, et al.
Veröffentlicht: (2024)
A Simple Interpretable Transformer for Fine-Grained Image Classification and Analysis
von: Paul, Dipanjyoti, et al.
Veröffentlicht: (2023)
von: Paul, Dipanjyoti, et al.
Veröffentlicht: (2023)
Pairwise Matching of Intermediate Representations for Fine-grained Explainability
von: Shrack, Lauren, et al.
Veröffentlicht: (2025)
von: Shrack, Lauren, et al.
Veröffentlicht: (2025)
Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis
von: Chowdhury, Arpita, et al.
Veröffentlicht: (2025)
von: Chowdhury, Arpita, et al.
Veröffentlicht: (2025)
Recent Advances in Traffic Accident Analysis and Prediction: A Comprehensive Review of Machine Learning Techniques
von: Behboudi, Noushin, et al.
Veröffentlicht: (2024)
von: Behboudi, Noushin, et al.
Veröffentlicht: (2024)
VisTaxa: Developing a Taxonomy of Historical Visualizations
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
Strong Libraries, Strong Scores
von: Gray, Carlyn
Veröffentlicht: (2006)
von: Gray, Carlyn
Veröffentlicht: (2006)
FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2025)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2025)
Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026)
VLM4Bio: A Benchmark Dataset to Evaluate Pretrained Vision-Language Models for Trait Discovery from Biological Images
von: Maruf, M., et al.
Veröffentlicht: (2024)
von: Maruf, M., et al.
Veröffentlicht: (2024)
A Taxonomy for Data Contamination in Large Language Models
von: Palavalli, Medha, et al.
Veröffentlicht: (2024)
von: Palavalli, Medha, et al.
Veröffentlicht: (2024)
Finer-CAM: Spotting the Difference Reveals Finer Details for Visual Explanation
von: Zhang, Ziheng, et al.
Veröffentlicht: (2025)
von: Zhang, Ziheng, et al.
Veröffentlicht: (2025)
Scalable Deep Metric Learning on Attributed Graphs
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
InsightBuild: LLM-Powered Causal Reasoning in Smart Building Systems
von: Neogi, Pinaki Prasad Guha, et al.
Veröffentlicht: (2025)
von: Neogi, Pinaki Prasad Guha, et al.
Veröffentlicht: (2025)
PIAD-SRNN: Physics-Informed Adaptive Decomposition in State-Space RNN
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2024)
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2024)
ALCo-FM: Adaptive Long-Context Foundation Model for Accident Prediction
von: Neogi, Pinaki Prasad Guha, et al.
Veröffentlicht: (2025)
von: Neogi, Pinaki Prasad Guha, et al.
Veröffentlicht: (2025)
Federated Contrastive Learning of Graph-Level Representations
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
DACA-GRPO: Denoising-Aware Credit Assignment for Reinforcement Learning in Diffusion Language Models
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026)
AnimeAdapter: Fine-grained and Consistent Zero-shot Anime Character Generation
von: Han, Yixuan
Veröffentlicht: (2026)
von: Han, Yixuan
Veröffentlicht: (2026)
Investigating a Model-Agnostic and Imputation-Free Approach for Irregularly-Sampled Multivariate Time-Series Modeling
von: Neog, Abhilash, et al.
Veröffentlicht: (2025)
von: Neog, Abhilash, et al.
Veröffentlicht: (2025)
Personalized Representation from Personalized Generation
von: Sundaram, Shobhita, et al.
Veröffentlicht: (2024)
von: Sundaram, Shobhita, et al.
Veröffentlicht: (2024)
ARIAL: An Agentic Framework for Document VQA with Precise Answer Localization
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2025)
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2025)
EPIC: Efficient Pruning for Inference on Constrained Devices
von: Das, Subhransu, et al.
Veröffentlicht: (2025)
von: Das, Subhransu, et al.
Veröffentlicht: (2025)
Scalable Deep Graph Clustering with Random-walk based Self-supervised Learning
von: Li, Xiang, et al.
Veröffentlicht: (2021)
von: Li, Xiang, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
TaxaDiffusion: Progressively Trained Diffusion Model for Fine-Grained Species Generation
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2025) -
Beyond Loss Guidance: Using PDE Residuals as Spectral Attention in Diffusion Neural Operators
von: Sawhney, Medha, et al.
Veröffentlicht: (2025) -
SeamCam: Quantifying Seamless Camouflage via Multi-Cue Visual Detectability
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2026) -
Controlla: Learning Controllability via Graph-Constrained Latent Geometry
von: Murthy, Jamuna S., et al.
Veröffentlicht: (2026) -
Frequency-Guided Masking for Enhanced Vision Self-Supervised Learning
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)