Towards Engineering a Web-Scale Multimedia Service: A Case Study Using Spark
- Gylfi Þór Guðmundsson,
- Laurent Amsaleg,
- ,
- Michael J. Franklin
- Reykjavík University,
- Institut de recherche en informatique et systèmes aléatoires,
- The University of Chicago
Publikation:
Konference artikel i Proceeding eller bog/rapport kapitel
Konferencebidrag i proceedings
Peer-reviewOpen Access
Resume
Computing power has now become abundant with multi-core machines, grids and clouds, but it remains a challenge to harness the available power and move towards gracefully handling web-scale datasets. Several researchers have used automatically distributed computing frameworks, notably Hadoop and Spark, for processing multimedia material, but mostly using small collections on small clusters. In this paper, we describe the engineering process for a prototype of a (near) web-scale multimedia service using the Spark framework running on the AWS cloud service. We present experimental results using up to 43 billion SIFT feature vectors from the public YFCC 100M collection, making this the largest high-dimensional feature vector collection reported in the literature. The design of the prototype and performance results demonstrate both the flexibility and scalability of the Spark framework for implementing multimedia services.
Publikation information
Produktionstype
Publikation:
Konference artikel i Proceeding eller bog/rapport kapitel
Konferencebidrag i proceedings
Peer-reviewOriginalsprog
EngelskSider fra-til (Antal sider)
Sider 1-12Publikationsmilepæle
- Udgivet - 06/2017
Publikationsstatus
Udgivet - 06/2017
Udgivelsessted
Taipei, TaiwanForlag
Association for Computing Machinery, USAISBN (Trykt)
978-1-4503-5002-0Publication IDs
- Scopus: 85025671251
Titel på værtspublikation
Proceedings of the ACM Multimedia Systems Conference (MMSys)Metrikker
PlumX
Hentninger
11
Citationer
14
Adgang til dokumenter
Indsendt manuskript, 575.96 KB
Forlagets udgivne version
