In this episode we're joined by Youen Chéné and Aurélien Vandel from Saagie who talk to us about their experiences deploying Spark Streaming workloads in production (based on their Dataworks Summit talk), what worked well, what didn't and what they'd recommend you might want to do if you follow in their footsteps.   Enjoy!

00:00 Recent events

Dave

Big Data Videos

http://www.kdnuggets.com/2017/05/top-recent-big-data-videos-youtube.html https://www.youtube.com/watch?v=RQ9czRAdmMs https://www.youtube.com/watch?v=hsoKlE67rTw

Jhon

InsightOut: The role of Apache Atlas in the open metadata ecosystem

http://www.ibmbigdatahub.com/blog/insightout-role-apache-atlas-open-metadata-ecosystem https://www.youtube.com/watch?v=yQvmoDtGgbo Apache Atlas API Version 2

https://atlas.incubator.apache.org/api/v2/index.html

Cloud giants 'ran out' of fast GPUs for AI boffins

https://www.theregister.co.uk/2017/05/22/cloud_providers_ai_researchers/

Benchmark: Sub-Second Analytics with Apache Hive and Druid

https://hortonworks.com/blog/sub-second-analytics-hive-druid/

26:00 Spark Streaming and Suicidal Tendencies

https://dataworkssummit.com/munich-2017/sessions/spark-streaming-and-suicidal-tendencies/

Video: https://www.youtube.com/watch?v=Us8kizlbJtc Slides: https://www.slideshare.net/HadoopSummit/spark-streaming-and-suicidal-tendencies

Youen Chéné, CTO @Saagie

https://www.linkedin.com/in/youenchene/

Aurélien Vandel, Data Engineer

https://www.linkedin.com/in/aur%C3%A9lien-vandel-060b5b8a/

01:11:17 End

Please use the Contact Form on this blog or our twitter feed to send us your questions, or to suggest future episode topics you would like us to cover.