IRGen: Generative Modeling for Image Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yidan, Zhang, Ting, Chen, Dong, Wang, Yujing, Chen, Qi, Xie, Xing, Sun, Hao, Deng, Weiwei, Zhang, Qi, Yang, Fan, Yang, Mao, Liao, Qingmin, Wang, Jingdong, Guo, Baining |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ScIRGen: Synthesize Realistic and Large-Scale RAG Dataset for Scientific Research
by: Lin, Junyong, et al.
Published: (2025)
by: Lin, Junyong, et al.
Published: (2025)
ASI++: Towards Distributionally Balanced End-to-End Generative Retrieval
by: Liu, Yuxuan, et al.
Published: (2024)
by: Liu, Yuxuan, et al.
Published: (2024)
VASA-3D: Lifelike Audio-Driven Gaussian Head Avatars from a Single Image
by: Xu, Sicheng, et al.
Published: (2025)
by: Xu, Sicheng, et al.
Published: (2025)
RodinHD: High-Fidelity 3D Avatar Generation with Diffusion Models
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
VolumeDiffusion: Flexible Text-to-3D Generation with Efficient Volumetric Encoder
by: Tang, Zhicong, et al.
Published: (2023)
by: Tang, Zhicong, et al.
Published: (2023)
Research on Graph-Retrieval Augmented Generation Based on Historical Text Knowledge Graphs
by: Fan, Yang, et al.
Published: (2025)
by: Fan, Yang, et al.
Published: (2025)
Gaussian Variation Field Diffusion for High-fidelity Video-to-4D Synthesis
by: Zhang, Bowen, et al.
Published: (2025)
by: Zhang, Bowen, et al.
Published: (2025)
GeAR: Generation Augmented Retrieval
by: Liu, Haoyu, et al.
Published: (2025)
by: Liu, Haoyu, et al.
Published: (2025)
Aligning Vision Models with Human Aesthetics in Retrieval: Benchmarks and Algorithms
by: Zhang, Miaosen, et al.
Published: (2024)
by: Zhang, Miaosen, et al.
Published: (2024)
Benchmark Shadows: Data Alignment, Parameter Footprints, and Generalization in Large Language Models
by: Zou, Hongjian, et al.
Published: (2026)
by: Zou, Hongjian, et al.
Published: (2026)
Mamba Retriever: Utilizing Mamba for Effective and Efficient Dense Retrieval
by: Zhang, Hanqi, et al.
Published: (2024)
by: Zhang, Hanqi, et al.
Published: (2024)
CCA: Collaborative Competitive Agents for Image Editing
by: Hang, Tiankai, et al.
Published: (2024)
by: Hang, Tiankai, et al.
Published: (2024)
CLIP-based Synergistic Knowledge Transfer for Text-based Person Retrieval
by: Liu, Yating, et al.
Published: (2023)
by: Liu, Yating, et al.
Published: (2023)
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs
by: Xu, Sicheng, et al.
Published: (2026)
by: Xu, Sicheng, et al.
Published: (2026)
Shadow of Kerr black hole surrounded by a cloud of strings in Rastall gravity and constraints from M87*
by: Sun, Qi, et al.
Published: (2024)
by: Sun, Qi, et al.
Published: (2024)
Circular-ribbon flares and the related activities
by: Zhang, Qingmin
Published: (2024)
by: Zhang, Qingmin
Published: (2024)
BDC-Occ: Binarized Deep Convolution Unit For Binarized Occupancy Network
by: Zhang, Zongkai, et al.
Published: (2024)
by: Zhang, Zongkai, et al.
Published: (2024)
Cover Picture: Putting Hybrid Nanomaterials to Work for Biomedical Applications (Angew. Chem. Int. Ed. 16/2024)
by: Dong Zhang, et al.
Published: (2024)
by: Dong Zhang, et al.
Published: (2024)
Titelbild: Putting Hybrid Nanomaterials to Work for Biomedical Applications (Angew. Chem. 16/2024)
by: Dong Zhang, et al.
Published: (2024)
by: Dong Zhang, et al.
Published: (2024)
Putting Hybrid Nanomaterials to Work for Biomedical Applications
by: Dong Zhang, et al.
Published: (2024)
by: Dong Zhang, et al.
Published: (2024)
Improving Retrieval-Augmented Generation through Multi-Agent Reinforcement Learning
by: Chen, Yiqun, et al.
Published: (2025)
by: Chen, Yiqun, et al.
Published: (2025)
GaussianCube: A Structured and Explicit Radiance Representation for 3D Generative Modeling
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
Rich Semantic Knowledge Enhanced Large Language Models for Few-shot Chinese Spell Checking
by: Dong, Ming, et al.
Published: (2024)
by: Dong, Ming, et al.
Published: (2024)
Bone Marrow Mesenchymal Stem Cells Improve Cognitive Impairment Induced by Neuropathic Pain Through Blood CXCL12 / CXCR4 Axis in Male Mice
by: Kai Sun, et al.
Published: (2026)
by: Kai Sun, et al.
Published: (2026)
RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval
by: Liu, Di, et al.
Published: (2024)
by: Liu, Di, et al.
Published: (2024)
Diffusion Models without Classifier-free Guidance
by: Tang, Zhicong, et al.
Published: (2025)
by: Tang, Zhicong, et al.
Published: (2025)
HyperBlocker: Accelerating Rule-based Blocking in Entity Resolution using GPUs
by: Zhu, Xiaoke, et al.
Published: (2024)
by: Zhu, Xiaoke, et al.
Published: (2024)
ViTime: Foundation Model for Time Series Forecasting Powered by Vision Intelligence
by: Yang, Luoxiao, et al.
Published: (2024)
by: Yang, Luoxiao, et al.
Published: (2024)
DiffVein: A Unified Diffusion Network for Finger Vein Segmentation and Authentication
by: Liu, Yanjun, et al.
Published: (2024)
by: Liu, Yanjun, et al.
Published: (2024)
GraphicsDreamer: Image to 3D Generation with Physical Consistency
by: Chen, Pei, et al.
Published: (2024)
by: Chen, Pei, et al.
Published: (2024)
MageBench: Bridging Large Multimodal Models to Agents
by: Zhang, Miaosen, et al.
Published: (2024)
by: Zhang, Miaosen, et al.
Published: (2024)
Soliton resolution, asymptotic stability and Painlevé transcendents in the combined Wadati-Konno-Ichikawa and short-pulse equation
by: Zhang, Yidan, et al.
Published: (2025)
by: Zhang, Yidan, et al.
Published: (2025)
MTL-LoRA: Low-Rank Adaptation for Multi-Task Learning
by: Yang, Yaming, et al.
Published: (2024)
by: Yang, Yaming, et al.
Published: (2024)
MobileManiBench: Simplifying Model Verification for Mobile Manipulation
by: Wang, Wenbo, et al.
Published: (2026)
by: Wang, Wenbo, et al.
Published: (2026)
Pose Magic: Efficient and Temporally Consistent Human Pose Estimation with a Hybrid Mamba-GCN Network
by: Zhang, Xinyi, et al.
Published: (2024)
by: Zhang, Xinyi, et al.
Published: (2024)
Efficient Feature Aggregation and Scale-Aware Regression for Monocular 3D Object Detection
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
UniAlignment: Semantic Alignment for Unified Image Generation, Understanding, Manipulation and Perception
by: Song, Xinyang, et al.
Published: (2025)
by: Song, Xinyang, et al.
Published: (2025)
Query-Kontext: An Unified Multimodal Model for Image Generation and Editing
by: Song, Yuxin, et al.
Published: (2025)
by: Song, Yuxin, et al.
Published: (2025)
E5-V: Universal Embeddings with Multimodal Large Language Models
by: Jiang, Ting, et al.
Published: (2024)
by: Jiang, Ting, et al.
Published: (2024)
LinearRAG: Linear Graph Retrieval Augmented Generation on Large-scale Corpora
by: Zhuang, Luyao, et al.
Published: (2025)
by: Zhuang, Luyao, et al.
Published: (2025)
Similar Items
-
ScIRGen: Synthesize Realistic and Large-Scale RAG Dataset for Scientific Research
by: Lin, Junyong, et al.
Published: (2025) -
ASI++: Towards Distributionally Balanced End-to-End Generative Retrieval
by: Liu, Yuxuan, et al.
Published: (2024) -
VASA-3D: Lifelike Audio-Driven Gaussian Head Avatars from a Single Image
by: Xu, Sicheng, et al.
Published: (2025) -
RodinHD: High-Fidelity 3D Avatar Generation with Diffusion Models
by: Zhang, Bowen, et al.
Published: (2024) -
VolumeDiffusion: Flexible Text-to-3D Generation with Efficient Volumetric Encoder
by: Tang, Zhicong, et al.
Published: (2023)