Research topic

Multimodal AI Research

This topic page groups Naser Ezzati-Jivan research papers related to multimodal ai. Each linked record provides the paper's problem, method, findings, limitations, keywords, and authoritative source links.

Related search terms: multimodal ai

2 papers in this topic, ordered newest first. The detailed paper records contain the evidence-grounded methods, tools, datasets, findings, and citation guidance.

Selected papers

2025 · 2025 International Conference on Artificial Intelligence, Computer, Data Sciences and Applications (ACDSA)

AI Video Retrieval: A Semantic Search & Timestamp Alignment System

Hridoy Rahman, Naser Ezzati-Jivan, Blessing Ogbuokiri

The paper implements a timestamp-aware multimodal video-retrieval pipeline that joins speech transcription, sampled-frame captioning, text embeddings, and approximate-nearest-neighbor search.

Keywords: video retrieval · semantic search · timestamp alignment · AI video search · ACDSA 2025

Read the detailed paper record · · Authoritative source