$\texttt{InfoHier}$: Hierarchical Information Extraction via Encoding and Embedding
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Tianru, Ju, Li, Singh, Prashant, Toor, Salman |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Embedding-based Retrieval in Multimodal Content Moderation
by: Liang, Hanzhong, et al.
Published: (2025)
by: Liang, Hanzhong, et al.
Published: (2025)
Dual Pose-invariant Embeddings: Learning Category and Object-specific Discriminative Representations for Recognition and Retrieval
by: Sarkar, Rohan, et al.
Published: (2024)
by: Sarkar, Rohan, et al.
Published: (2024)
Character-based Outfit Generation with Vision-augmented Style Extraction via LLMs
by: Forouzandehmehr, Najmeh, et al.
Published: (2024)
by: Forouzandehmehr, Najmeh, et al.
Published: (2024)
Online Learning via Memory: Retrieval-Augmented Detector Adaptation
by: Jian, Yanan, et al.
Published: (2024)
by: Jian, Yanan, et al.
Published: (2024)
CoopHash: Cooperative Learning of Multipurpose Descriptor and Contrastive Pair Generator via Variational MCMC Teaching for Supervised Image Hashing
by: Doan, Khoa D., et al.
Published: (2022)
by: Doan, Khoa D., et al.
Published: (2022)
GBSK: Skeleton Clustering via Granular-ball Computing and Multi-Sampling for Large-Scale Data
by: Chen, Yewang, et al.
Published: (2025)
by: Chen, Yewang, et al.
Published: (2025)
MOON Embedding: Multimodal Representation Learning for E-commerce Search Advertising
by: Fu, Chenghan, et al.
Published: (2025)
by: Fu, Chenghan, et al.
Published: (2025)
Image Hashing via Cross-View Code Alignment in the Age of Foundation Models
by: Moummad, Ilyass, et al.
Published: (2025)
by: Moummad, Ilyass, et al.
Published: (2025)
LotusFilter: Fast Diverse Nearest Neighbor Search via a Learned Cutoff Table
by: Matsui, Yusuke
Published: (2025)
by: Matsui, Yusuke
Published: (2025)
AutoPP: Towards Automated Product Poster Generation and Optimization
by: Fan, Jiahao, et al.
Published: (2025)
by: Fan, Jiahao, et al.
Published: (2025)
SOLAR: SVD-Optimized Lifelong Attention for Recommendation
by: Zhang, Chenghao, et al.
Published: (2026)
by: Zhang, Chenghao, et al.
Published: (2026)
Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting
by: Maurya, Amritansh, et al.
Published: (2026)
by: Maurya, Amritansh, et al.
Published: (2026)
Uni-Layout: Integrating Human Feedback in Unified Layout Generation and Evaluation
by: Lu, Shuo, et al.
Published: (2025)
by: Lu, Shuo, et al.
Published: (2025)
NanoVDR: Distilling a 2B Vision-Language Retriever into a 70M Text-Only Encoder for Visual Document Retrieval
by: Liu, Zhuchenyang, et al.
Published: (2026)
by: Liu, Zhuchenyang, et al.
Published: (2026)
Maximal Matching Matters: Preventing Representation Collapse for Robust Cross-Modal Retrieval
by: Alomari, Hani, et al.
Published: (2025)
by: Alomari, Hani, et al.
Published: (2025)
Multi-event Video-Text Retrieval
by: Zhang, Gengyuan, et al.
Published: (2023)
by: Zhang, Gengyuan, et al.
Published: (2023)
Towards Universal Video Retrieval: Generalizing Video Embedding via Synthesized Multimodal Pyramid Curriculum
by: Guo, Zhuoning, et al.
Published: (2025)
by: Guo, Zhuoning, et al.
Published: (2025)
MuseChat: A Conversational Music Recommendation System for Videos
by: Dong, Zhikang, et al.
Published: (2023)
by: Dong, Zhikang, et al.
Published: (2023)
LD-DETR: Loop Decoder DEtection TRansformer for Video Moment Retrieval and Highlight Detection
by: Zhao, Pengcheng, et al.
Published: (2025)
by: Zhao, Pengcheng, et al.
Published: (2025)
Compact Hypercube Embeddings for Fast Text-based Wildlife Observation Retrieval
by: Moummad, Ilyass, et al.
Published: (2026)
by: Moummad, Ilyass, et al.
Published: (2026)
CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval
by: Xu, Yifan, et al.
Published: (2024)
by: Xu, Yifan, et al.
Published: (2024)
When & How to Write for Personalized Demand-aware Query Rewriting in Video Search
by: cheng, Cheng, et al.
Published: (2025)
by: cheng, Cheng, et al.
Published: (2025)
AdaTask: A Task-aware Adaptive Learning Rate Approach to Multi-task Learning
by: Yang, Enneng, et al.
Published: (2022)
by: Yang, Enneng, et al.
Published: (2022)
VQPP: Video Query Performance Prediction Benchmark
by: Lutu, Adrian Catalin, et al.
Published: (2026)
by: Lutu, Adrian Catalin, et al.
Published: (2026)
Hierarchy-of-Visual-Words: a Learning-based Approach for Trademark Image Retrieval
by: Lourenço, Vítor N., et al.
Published: (2019)
by: Lourenço, Vítor N., et al.
Published: (2019)
Transfer Learning and Mixup for Fine-Grained Few-Shot Fungi Classification
by: Tam, Jason Kahei, et al.
Published: (2025)
by: Tam, Jason Kahei, et al.
Published: (2025)
Tile-Based ViT Inference with Visual-Cluster Priors for Zero-Shot Multi-Species Plant Identification
by: Gustineli, Murilo, et al.
Published: (2025)
by: Gustineli, Murilo, et al.
Published: (2025)
Improving Visual Recommendation on E-commerce Platforms Using Vision-Language Models
by: Yada, Yuki, et al.
Published: (2025)
by: Yada, Yuki, et al.
Published: (2025)
PixRec: Leveraging Visual Context for Next-Item Prediction in Sequential Recommendation
by: Chakrabarty, Sayak, et al.
Published: (2026)
by: Chakrabarty, Sayak, et al.
Published: (2026)
FOR: Finetuning for Object Level Open Vocabulary Image Retrieval
by: Levi, Hila, et al.
Published: (2024)
by: Levi, Hila, et al.
Published: (2024)
Image Outlier Detection Without Training using RANSAC
by: Tsai, Chen-Han, et al.
Published: (2023)
by: Tsai, Chen-Han, et al.
Published: (2023)
Towards Resource-Efficient Streaming of Large-Scale Medical Image Datasets for Deep Learning
by: Kulkarni, Pranav, et al.
Published: (2023)
by: Kulkarni, Pranav, et al.
Published: (2023)
Generalized Contrastive Learning for Multi-Modal Retrieval and Ranking
by: Zhu, Tianyu, et al.
Published: (2024)
by: Zhu, Tianyu, et al.
Published: (2024)
Decoupled Training: Return of Frustratingly Easy Multi-Domain Learning
by: Wang, Ximei, et al.
Published: (2023)
by: Wang, Ximei, et al.
Published: (2023)
PICS: Pipeline for Image Captioning and Search
by: Rosario, Grant, et al.
Published: (2024)
by: Rosario, Grant, et al.
Published: (2024)
Patent Figure Classification using Large Vision-language Models
by: Awale, Sushil, et al.
Published: (2025)
by: Awale, Sushil, et al.
Published: (2025)
Proceedings of the 6th International Workshop on Reading Music Systems
by: Calvo-Zaragoza, Jorge, et al.
Published: (2024)
by: Calvo-Zaragoza, Jorge, et al.
Published: (2024)
Metric Compatible Training for Online Backfilling in Large-Scale Retrieval
by: Seo, Seonguk, et al.
Published: (2023)
by: Seo, Seonguk, et al.
Published: (2023)
A Dataset and Framework for Learning State-invariant Object Representations
by: Sarkar, Rohan, et al.
Published: (2024)
by: Sarkar, Rohan, et al.
Published: (2024)
iRAG: Advancing RAG for Videos with an Incremental Approach
by: Arefeen, Md Adnan, et al.
Published: (2024)
by: Arefeen, Md Adnan, et al.
Published: (2024)
Similar Items
-
Embedding-based Retrieval in Multimodal Content Moderation
by: Liang, Hanzhong, et al.
Published: (2025) -
Dual Pose-invariant Embeddings: Learning Category and Object-specific Discriminative Representations for Recognition and Retrieval
by: Sarkar, Rohan, et al.
Published: (2024) -
Character-based Outfit Generation with Vision-augmented Style Extraction via LLMs
by: Forouzandehmehr, Najmeh, et al.
Published: (2024) -
Online Learning via Memory: Retrieval-Augmented Detector Adaptation
by: Jian, Yanan, et al.
Published: (2024) -
CoopHash: Cooperative Learning of Multipurpose Descriptor and Contrastive Pair Generator via Variational MCMC Teaching for Supervised Image Hashing
by: Doan, Khoa D., et al.
Published: (2022)