ReSpec: Relevance and Specificity Grounded Online Filtering for Learning on Video-Text Data Streams

Kim, Chris Dongjoo, Moon, Jihwan, Moon, Sangwoo, Yun, Heeseung, Lee, Sihaeng, Kembhavi, Aniruddha, Lee, Soonyoung, Kim, Gunhee, Lee, Sangho, Clark, Christopher

Apr-22-2025–arXiv.org Artificial Intelligence

The rapid growth of video-text data presents challenges in storage and computation during training. Online learning, which processes streaming data in real-time, offers a promising solution to these issues while also allowing swift adaptations in scenarios demanding real-time responsiveness. One strategy to enhance the efficiency and effectiveness of learning involves identifying and prioritizing data that enhances performance on target downstream tasks. We propose Relevance and Specificity-based online filtering framework (ReSpec) that selects data based on four criteria: (i) modality alignment for clean data, (ii) task relevance for target focused data, (iii) specificity for informative and detailed data, and (iv) efficiency for low-latency processing. Relevance is determined by the probabilistic alignment of incoming data with downstream tasks, while specificity employs the distance to a root embedding representing the least specific data as an efficient proxy for informativeness. By establishing reference points from target task data, ReSpec filters incoming data in real-time, eliminating the need for extensive storage and compute. Evaluating on large-scale datasets WebVid2M and VideoCC3M, ReSpec attains state-of-the-art performance on five zeroshot video retrieval tasks, using as little as 5% of the data while incurring minimal compute. The source code is available at https://github.com/cdjkim/ReSpec.

large language model, machine learning, real time system, (19 more...)

arXiv.org Artificial Intelligence

Apr-22-2025

arXiv.org PDF

Add feedback

Genre:
- Research Report (1.00)

Industry:
- Education > Educational Setting > Online (0.50)

Technology:
- Information Technology
  - Architecture > Real Time Systems (1.00)
  - Data Science (0.93)
  - Artificial Intelligence
    - Machine Learning (1.00)
    - Natural Language > Large Language Model (0.46)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found