ReSpec: Relevance and Specificity Grounded Online Filtering for Learning on Video-Text Data Streams
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Chris Dongjoo, Moon, Jihwan, Moon, Sangwoo, Yun, Heeseung, Lee, Sihaeng, Kembhavi, Aniruddha, Lee, Soonyoung, Kim, Gunhee, Lee, Sangho, Clark, Christopher |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sample Selection via Contrastive Fragmentation for Noisy Label Regression
von: Kim, Chris Dongjoo, et al.
Veröffentlicht: (2025)
von: Kim, Chris Dongjoo, et al.
Veröffentlicht: (2025)
MASH-VLM: Mitigating Action-Scene Hallucination in Video-LLMs through Disentangled Spatial-Temporal Representations
von: Bae, Kyungho, et al.
Veröffentlicht: (2025)
von: Bae, Kyungho, et al.
Veröffentlicht: (2025)
Can Language Models Laugh at YouTube Short-form Videos?
von: Ko, Dayoon, et al.
Veröffentlicht: (2023)
von: Ko, Dayoon, et al.
Veröffentlicht: (2023)
ReSpec: Towards Optimizing Speculative Decoding in Reinforcement Learning Systems
von: Chen, Qiaoling, et al.
Veröffentlicht: (2025)
von: Chen, Qiaoling, et al.
Veröffentlicht: (2025)
Bi-directional Contextual Attention for 3D Dense Captioning
von: Kim, Minjung, et al.
Veröffentlicht: (2024)
von: Kim, Minjung, et al.
Veröffentlicht: (2024)
How to Move Your Dragon: Text-to-Motion Synthesis for Large-Vocabulary Objects
von: Lee, Wonkwang, et al.
Veröffentlicht: (2025)
von: Lee, Wonkwang, et al.
Veröffentlicht: (2025)
ViSAGe: Video-to-Spatial Audio Generation
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
See It All: Contextualized Late Aggregation for 3D Dense Captioning
von: Kim, Minjung, et al.
Veröffentlicht: (2024)
von: Kim, Minjung, et al.
Veröffentlicht: (2024)
Can LLMs Deceive CLIP? Benchmarking Adversarial Compositionality of Pre-trained Multimodal Representation via Text Updates
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2025)
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2025)
One Diffusion to Generate Them All
von: Le, Duong H., et al.
Veröffentlicht: (2024)
von: Le, Duong H., et al.
Veröffentlicht: (2024)
When Meta-Learning Meets Online and Continual Learning: A Survey
von: Son, Jaehyeon, et al.
Veröffentlicht: (2023)
von: Son, Jaehyeon, et al.
Veröffentlicht: (2023)
ChatEXAONEPath: An Expert-level Multimodal Large Language Model for Histopathology Using Whole Slide Images
von: Kim, Sangwook, et al.
Veröffentlicht: (2025)
von: Kim, Sangwook, et al.
Veröffentlicht: (2025)
EdiText: Controllable Coarse-to-Fine Text Editing with Diffusion Language Models
von: Lee, Che Hyun, et al.
Veröffentlicht: (2025)
von: Lee, Che Hyun, et al.
Veröffentlicht: (2025)
Pri4R: Learning World Dynamics for Vision-Language-Action Models with Privileged 4D Representation
von: Kim, Jisoo, et al.
Veröffentlicht: (2026)
von: Kim, Jisoo, et al.
Veröffentlicht: (2026)
Experiences of Generation Z Nurses Adapting to Work in a Tertiary Hospital: A Grounded Theory Study
von: Youngji Moon, et al.
Veröffentlicht: (2024)
von: Youngji Moon, et al.
Veröffentlicht: (2024)
Towards Diverse Evaluation of Class Incremental Learning: A Representation Learning Perspective
von: Cha, Sungmin, et al.
Veröffentlicht: (2022)
von: Cha, Sungmin, et al.
Veröffentlicht: (2022)
Emerging Photon Jets in the Hadronic Calorimeter: A Novel Signature of Neutral Long-Lived Particles at the LHC
von: Kim, Jinheung, et al.
Veröffentlicht: (2025)
von: Kim, Jinheung, et al.
Veröffentlicht: (2025)
SAGE: Training Smart Any-Horizon Agents for Long Video Reasoning with Reinforcement Learning
von: Jain, Jitesh, et al.
Veröffentlicht: (2025)
von: Jain, Jitesh, et al.
Veröffentlicht: (2025)
Holistic Unlearning Benchmark: A Multi-Faceted Evaluation for Text-to-Image Diffusion Model Unlearning
von: Moon, Saemi, et al.
Veröffentlicht: (2024)
von: Moon, Saemi, et al.
Veröffentlicht: (2024)
Autonomous Robotic Radio Source Localization via a Novel Gaussian Mixture Filtering Approach
von: Kim, Sukkeun, et al.
Veröffentlicht: (2025)
von: Kim, Sukkeun, et al.
Veröffentlicht: (2025)
Gaze Beyond the Frame: Forecasting Egocentric 3D Visual Span
von: Yun, Heeseung, et al.
Veröffentlicht: (2025)
von: Yun, Heeseung, et al.
Veröffentlicht: (2025)
VoiceTailor: Lightweight Plug-In Adapter for Diffusion-Based Personalized Text-to-Speech
von: Kim, Heeseung, et al.
Veröffentlicht: (2024)
von: Kim, Heeseung, et al.
Veröffentlicht: (2024)
Microwave Quantum Illumination with Optical Memory and Single-Mode Phase-Conjugate Receiver
von: Jeon, Sangwoo, et al.
Veröffentlicht: (2024)
von: Jeon, Sangwoo, et al.
Veröffentlicht: (2024)
Single‐Mode Phase‐Conjugate Receiver for Microwave Quantum Illumination with a Lossy Optical Memory
von: Sangwoo Jeon, et al.
Veröffentlicht: (2025)
von: Sangwoo Jeon, et al.
Veröffentlicht: (2025)
Study on the Spinning Behavior and Nanoweb Fabrication of Polypropylene ( PP ) Nanofibers Under Different Hot Air and Voltage Conditions in a Melt Electrospinning/Melt Blown Hybrid System
von: Eunji Moon, et al.
Veröffentlicht: (2025)
von: Eunji Moon, et al.
Veröffentlicht: (2025)
Real‐Time Characterization of Polypropylene Jet Dynamics in Melt Electrospinning Using Laser Diffraction and Image Analysis
von: Jihwan Lim, et al.
Veröffentlicht: (2025)
von: Jihwan Lim, et al.
Veröffentlicht: (2025)
Positional functionalizations of metal–organic frameworks through invasive ligand exchange and additory MOF‐on‐MOF strategies: A review
von: Daeyeon Lee, et al.
Veröffentlicht: (2024)
von: Daeyeon Lee, et al.
Veröffentlicht: (2024)
Solid‐Phase Synthesis of 1,3,5,6‐Tetra‐Substituted 3,5‐Dihydroimidazo[4,5‐c][1,2]Thiazin‐4(1H)‐One 2,2‐Dioxide Derivatives
von: Jimin Moon, et al.
Veröffentlicht: (2025)
von: Jimin Moon, et al.
Veröffentlicht: (2025)
From Generation to Attribution: Music AI Agent Architectures for the Post-Streaming Era
von: Kim, Wonil, et al.
Veröffentlicht: (2025)
von: Kim, Wonil, et al.
Veröffentlicht: (2025)
Finding NeMo: Negative-mined Mosaic Augmentation for Referring Image Segmentation
von: Ha, Seongsu, et al.
Veröffentlicht: (2024)
von: Ha, Seongsu, et al.
Veröffentlicht: (2024)
Recasting Continual Learning as Sequence Modeling
von: Lee, Soochan, et al.
Veröffentlicht: (2023)
von: Lee, Soochan, et al.
Veröffentlicht: (2023)
Compositional Conservatism: A Transductive Approach in Offline Reinforcement Learning
von: Song, Yeda, et al.
Veröffentlicht: (2024)
von: Song, Yeda, et al.
Veröffentlicht: (2024)
Distilling Reinforcement Learning Algorithms for In-Context Model-Based Planning
von: Son, Jaehyeon, et al.
Veröffentlicht: (2025)
von: Son, Jaehyeon, et al.
Veröffentlicht: (2025)
Factors Associated With Long‐Term Visual Field Variability in Patients With Normal‐Tension Glaucoma
von: Seunghee Ha, et al.
Veröffentlicht: (2026)
von: Seunghee Ha, et al.
Veröffentlicht: (2026)
P‐75: Extreme Low Power a‐InGaZnO TFT Scan Driver with Extra Clock Signal Modulation
von: Hyunwoo Kim, et al.
Veröffentlicht: (2024)
von: Hyunwoo Kim, et al.
Veröffentlicht: (2024)
Mask2Flow-TSE: Two-Stage Target Speaker Extraction with Masking and Flow Matching
von: Moon, Junwon, et al.
Veröffentlicht: (2026)
von: Moon, Junwon, et al.
Veröffentlicht: (2026)
X-CANIDS: Signal-Aware Explainable Intrusion Detection System for Controller Area Network-Based In-Vehicle Network
von: Jeong, Seonghoon, et al.
Veröffentlicht: (2023)
von: Jeong, Seonghoon, et al.
Veröffentlicht: (2023)
β‐Cell dysfunction in diabetes: The role of neuroplastin
von: Moon‐Kyu Lee
Veröffentlicht: (2025)
von: Moon‐Kyu Lee
Veröffentlicht: (2025)
Analysis of U.S. Government Printing Office Documents Depository Survey Items. Follow-Up Study Report.
von: Lee, Tae Moon
Veröffentlicht: (1979)
von: Lee, Tae Moon
Veröffentlicht: (1979)
A Novel Design Strategy for Benzo[ b ]Thiophene Substituted Multiple Resonance Type Blue Fluorescent Dopant With Maximized Efficiency and Super‐Narrow FWHM
von: Dayeon Lee, et al.
Veröffentlicht: (2026)
von: Dayeon Lee, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Sample Selection via Contrastive Fragmentation for Noisy Label Regression
von: Kim, Chris Dongjoo, et al.
Veröffentlicht: (2025) -
MASH-VLM: Mitigating Action-Scene Hallucination in Video-LLMs through Disentangled Spatial-Temporal Representations
von: Bae, Kyungho, et al.
Veröffentlicht: (2025) -
Can Language Models Laugh at YouTube Short-form Videos?
von: Ko, Dayoon, et al.
Veröffentlicht: (2023) -
ReSpec: Towards Optimizing Speculative Decoding in Reinforcement Learning Systems
von: Chen, Qiaoling, et al.
Veröffentlicht: (2025) -
Bi-directional Contextual Attention for 3D Dense Captioning
von: Kim, Minjung, et al.
Veröffentlicht: (2024)