Data Provenance and Reproducibility with Pachyderm
Data Skeptic3 Feb 2017

Data Provenance and Reproducibility with Pachyderm

Versioning isn't just for source code. Being able to track changes to data is critical for answering questions about data provenance, quality, and reproducibility. Daniel Whitenack joins me this week to talk about these concepts and share his work on Pachyderm. Pachyderm is an open source containerized data lake.

During the show, Daniel mentioned the Gopher Data Science github repo as a great resource for any data scientists interested in the Go language. Although we didn't mention it, Daniel also did an interesting analysis on the 2016 world chess championship that complements our recent episode on chess well. You can find that post here

Supplemental music is Lee Rosevere's Let's Start at the Beginning.

Thanks to Periscope Data for sponsoring this episode. More about them at periscopedata.com/skeptics

Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(601)

Populärt inom Vetenskap

allt-du-velat-veta
dumma-manniskor
p3-dystopia
rss-ufobortom-rimligt-tvivel
sexet
rss-vetenskapsradion
medicinvetarna
ufo-sverige
rss-vetenskapsradion-2
svd-nyhetsartiklar
hacka-livet
det-morka-psyket
kapitalet-en-podd-om-ekonomi
halsorevolutionen
paranormalt-med-caroline-giertz
ufo-sverige-2
rss-klotet
ideer-som-forandrar-varlden
pojkmottagningen
bildningspodden