Building the howto100m Video Corpus
Data Skeptic19 Elo 2019

Building the howto100m Video Corpus

Video annotation is an expensive and time-consuming process. As a consequence, the available video datasets are useful but small. The availability of machine transcribed explainer videos offers a unique opportunity to rapidly develop a useful, if dirty, corpus of videos that are "self annotating", as hosts explain the actions they are taking on the screen.

This episode is a discussion of the HowTo100m dataset - a project which has assembled a video corpus of 136M video clips with captions covering 23k activities.

Related Links

The paper will be presented at ICCV 2019

@antoine77340

Antoine on Github

Antoine's homepage

Jaksot(589)

Suosittua kategoriassa Tiede

rss-mita-tulisi-tietaa
tiedekulma-podcast
hippokrateen-vastaanotolla
rss-lihavuudesta-podcast
rss-poliisin-mieli
utelias-mieli
sotataidon-ytimessa
docemilia
filocast-filosofian-perusteet
mielipaivakirja
rss-totta-vai-tuubaa
rss-duodecim-lehti
rss-radplus
radio-antro
rss-ammamafia
rss-astetta-parempi-elama-podcast
rss-tiedetta-vai-tarinaa
rss-ilmasto-kriisissa
rss-ihmisen-aani
rss-tervetta-skeptisyytta