IEEE CSICC 2025 Oral
RAG-Driven Video QA with Adaptive Chunking: A Bilingual Educational Dataset
Arshia Hemmat, Mohammad Hassan Heydari, Kianoosh Vadaei, Melika Shirian, Afsaneh Fatemi

CLIP-SSIM adaptive chunking for video retrieval, with EduViQA-Alpha, a bilingual Persian and English educational VideoQA dataset.
A longer write-up of this work is on its way. Until then the paper and the code are the best place to look.