DataCube: A Video Retrieval Platform via Natural Language Semantic Profiling
Fuente:
arXiv
Saved in:
| Main Authors: | Ju, Yiming, Zhao, Hanyu, Ma, Quanyue, Hao, Donglin, Wu, Chengwei, Li, Ming, Wang, Songjing, Pan, Tengfei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond IID: Optimizing Instruction Learning from the Perspective of Instruction Interaction and Dependency
by: Zhao, Hanyu, et al.
Published: (2024)
by: Zhao, Hanyu, et al.
Published: (2024)
CI-VID: A Coherent Interleaved Text-Video Dataset
by: Ju, Yiming, et al.
Published: (2025)
by: Ju, Yiming, et al.
Published: (2025)
Scaling Towards the Information Boundary of Instruction Sets: The Infinity Instruct Subject Technical Report
by: Du, Li, et al.
Published: (2025)
by: Du, Li, et al.
Published: (2025)
Accelerate Scaling of LLM Finetuning via Quantifying the Coverage and Depth of Instruction Set
by: Wu, Chengwei, et al.
Published: (2025)
by: Wu, Chengwei, et al.
Published: (2025)
CCI3.0-HQ: a large-scale Chinese dataset of high quality designed for pre-training large language models
by: Wang, Liangdong, et al.
Published: (2024)
by: Wang, Liangdong, et al.
Published: (2024)
Training Data for Large Language Model
by: Ju, Yiming, et al.
Published: (2024)
by: Ju, Yiming, et al.
Published: (2024)
Real Face Video Animation Platform
by: Chen, Xiaokai, et al.
Published: (2024)
by: Chen, Xiaokai, et al.
Published: (2024)
Evaluating the Semantic Profiling Abilities of LLMs for Natural Language Utterances in Data Visualization
by: Bako, Hannah K., et al.
Published: (2024)
by: Bako, Hannah K., et al.
Published: (2024)
CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
by: Wang, Feiyang, et al.
Published: (2025)
by: Wang, Feiyang, et al.
Published: (2025)
Figure 8 from: Yang X, Zhu Y, Duan S, Wu X, Zhao C (2025) Morphology and multigene phylogeny revealed four new species of Geastrum (Geastrales, Basidiomycota) from China. MycoKeys 113: 73-100. https://doi.org/10.3897/mycokeys.113.139672
by: Yang, Xin, et al.
Published: (2025)
by: Yang, Xin, et al.
Published: (2025)
SLFNet: Generating Semantic Logic Forms from Natural Language Using Semantic Probability Graphs
by: Wu, Hao, et al.
Published: (2024)
by: Wu, Hao, et al.
Published: (2024)
LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval
by: Ning, Zhenyu, et al.
Published: (2025)
by: Ning, Zhenyu, et al.
Published: (2025)
Video2LoRA: Unified Semantic-Controlled Video Generation via Per-Reference-Video LoRA
by: Wu, Zexi, et al.
Published: (2026)
by: Wu, Zexi, et al.
Published: (2026)
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing Scenes
by: Jiang, Jiajun, et al.
Published: (2025)
by: Jiang, Jiajun, et al.
Published: (2025)
CiteRadar: A Citation Intelligence Platform for Researcher Profiling and Geographic Visualization
by: Niu, Chenxu, et al.
Published: (2026)
by: Niu, Chenxu, et al.
Published: (2026)
VideoAfford: Grounding 3D Affordance from Human-Object-Interaction Videos via Multimodal Large Language Model
by: Wang, Hanqing, et al.
Published: (2026)
by: Wang, Hanqing, et al.
Published: (2026)
Self-Supervised Representation Learning with ID-Content Modality Alignment for Sequential Recommendation
by: Zhou, Donglin, et al.
Published: (2025)
by: Zhou, Donglin, et al.
Published: (2025)
AquilaMoE: Efficient Training for MoE Models with Scale-Up and Scale-Out Strategies
by: Zhang, Bo-Wen, et al.
Published: (2024)
by: Zhang, Bo-Wen, et al.
Published: (2024)
PromptEVC: Controllable Emotional Voice Conversion with Natural Language Prompts
by: Qi, Tianhua, et al.
Published: (2025)
by: Qi, Tianhua, et al.
Published: (2025)
An LMM for Efficient Video Understanding via Reinforced Compression of Video Cubes
by: Qi, Ji, et al.
Published: (2025)
by: Qi, Ji, et al.
Published: (2025)
CubeGraph: Efficient Retrieval-Augmented Generation for Spatial and Temporal Data
by: Yang, Mingyu, et al.
Published: (2026)
by: Yang, Mingyu, et al.
Published: (2026)
Scalable In-Context Learning on Tabular Data via Retrieval-Augmented Large Language Models
by: Wen, Xumeng, et al.
Published: (2025)
by: Wen, Xumeng, et al.
Published: (2025)
MQRLD: A Multimodal Data Retrieval Platform with Query-aware Feature Representation and Learned Index Based on Data Lake
by: Sheng, Ming, et al.
Published: (2024)
by: Sheng, Ming, et al.
Published: (2024)
One-Shot Pose-Driving Face Animation Platform
by: Feng, He, et al.
Published: (2024)
by: Feng, He, et al.
Published: (2024)
Natural Language Query to Configuration for Retrieval Agents
by: Pan, Melissa Z., et al.
Published: (2026)
by: Pan, Melissa Z., et al.
Published: (2026)
SSR: Enhancing Depth Perception in Vision-Language Models via Rationale-Guided Spatial Reasoning
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
OmniRetriever: Any-to-Any Audio-Video-Text Retrieval via Fusion-as-Teacher Distillation
by: Liu, Yunze, et al.
Published: (2026)
by: Liu, Yunze, et al.
Published: (2026)
Advancing super‐resolution stimulated emission depletion microscopy with luminescent lanthanide nanocrystals
by: Yaqi Chen, et al.
Published: (2026)
by: Yaqi Chen, et al.
Published: (2026)
MoCA: Identity-Preserving Text-to-Video Generation via Mixture of Cross Attention
by: Xie, Qi, et al.
Published: (2025)
by: Xie, Qi, et al.
Published: (2025)
Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval
by: Chen, Tao, et al.
Published: (2025)
by: Chen, Tao, et al.
Published: (2025)
The Semantic Scholar Open Data Platform
by: Kinney, Rodney, et al.
Published: (2023)
by: Kinney, Rodney, et al.
Published: (2023)
Natural Language-Driven Viewpoint Navigation for Volume Exploration via Semantic Block Representation
by: Zhao, Xuan, et al.
Published: (2025)
by: Zhao, Xuan, et al.
Published: (2025)
Efficient Imputation for Patch-based Missing Single-cell Data via Cluster-regularized Optimal Transport
by: Liu, Yuyu, et al.
Published: (2026)
by: Liu, Yuyu, et al.
Published: (2026)
Cocoon: Semantic Table Profiling Using Large Language Models
by: Huang, Zezhou, et al.
Published: (2024)
by: Huang, Zezhou, et al.
Published: (2024)
Enhancing Collaborative Semantics of Language Model-Driven Recommendations via Graph-Aware Learning
by: Guan, Zhong, et al.
Published: (2024)
by: Guan, Zhong, et al.
Published: (2024)
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
FAC$^2$E: Better Understanding Large Language Model Capabilities by Dissociating Language and Cognition
by: Wang, Xiaoqiang, et al.
Published: (2024)
by: Wang, Xiaoqiang, et al.
Published: (2024)
VDCook:DIY video data cook your MLLMs
by: Wu, Chengwei
Published: (2026)
by: Wu, Chengwei
Published: (2026)
PiTe: Pixel-Temporal Alignment for Large Video-Language Model
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
A Plug-and-Play Natural Language Rewriter for Natural Language to SQL
by: Ma, Peixian, et al.
Published: (2024)
by: Ma, Peixian, et al.
Published: (2024)
Similar Items
-
Beyond IID: Optimizing Instruction Learning from the Perspective of Instruction Interaction and Dependency
by: Zhao, Hanyu, et al.
Published: (2024) -
CI-VID: A Coherent Interleaved Text-Video Dataset
by: Ju, Yiming, et al.
Published: (2025) -
Scaling Towards the Information Boundary of Instruction Sets: The Infinity Instruct Subject Technical Report
by: Du, Li, et al.
Published: (2025) -
Accelerate Scaling of LLM Finetuning via Quantifying the Coverage and Depth of Instruction Set
by: Wu, Chengwei, et al.
Published: (2025) -
CCI3.0-HQ: a large-scale Chinese dataset of high quality designed for pre-training large language models
by: Wang, Liangdong, et al.
Published: (2024)