Latent Wander: an Alternative Interface for Interactive and Serendipitous Discovery of Large AV Archives

PROCEEDINGS OF THE 5TH WORKSHOP ON THE ANALYSIS, UNDERSTANDING AND PROMOTION OF HERITAGE CONTENTS, SUMAC 2023(2023)

引用 0|浏览0
暂无评分
摘要
Audiovisual (AV) archives are invaluable for holistically preserving the past. Unlike other forms, AV archives can be difficult to explore. This is not only because of its complex modality and sheer volume but also the lack of appropriate interfaces beyond keyword search. The recent rise in text-to-video retrieval tasks in computer science opens the gate to accessing AV content more naturally and semantically, able to map natural language descriptive sentences to matching videos. However, applications of this model are rarely seen. The contribution of this work is threefold. First, working with RTS (Television Suisse Romande), we identified the key blockers in a real archive for implementing such models. We built a functioning pipeline for encoding raw archive videos to the text-to-video feature vectors. Second, we designed and verified a method to encode and retrieve videos using emotionally abundant descriptions not supported in the original model. Third, we proposed an initial prototype for immersive and interactive exploration of AV archives in a latent space based on the previously mentioned encoding of videos.
更多
查看译文
关键词
Audiovisual archive,computational archival science,experimental museology,text-to-video retrieval,latent space
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要