In this episode we're joined by Youen Chéné and Aurélien Vandel from Saagie who talk to us about their experiences deploying Spark Streaming workloads in production (based on their Dataworks Summit talk), what worked well, what didn't and what they'd recommend you might want to do if you follow in their footsteps. Enjoy!
00:00 Recent events
Dave
Big Data Videos
http://www.kdnuggets.com/2017/05/top-recent-big-data-videos-youtube.html https://www.youtube.com/watch?v=RQ9czRAdmMs https://www.youtube.com/watch?v=hsoKlE67rTw
Jhon
InsightOut: The role of Apache Atlas in the open metadata ecosystem
http://www.ibmbigdatahub.com/blog/insightout-role-apache-atlas-open-metadata-ecosystem https://www.youtube.com/watch?v=yQvmoDtGgbo Apache Atlas API Version 2
https://atlas.incubator.apache.org/api/v2/index.html
Cloud giants 'ran out' of fast GPUs for AI boffins
https://www.theregister.co.uk/2017/05/22/cloud_providers_ai_researchers/
Benchmark: Sub-Second Analytics with Apache Hive and Druid
https://hortonworks.com/blog/sub-second-analytics-hive-druid/
26:00 Spark Streaming and Suicidal Tendencies
https://dataworkssummit.com/munich-2017/sessions/spark-streaming-and-suicidal-tendencies/
Video: https://www.youtube.com/watch?v=Us8kizlbJtc Slides: https://www.slideshare.net/HadoopSummit/spark-streaming-and-suicidal-tendencies
Youen Chéné, CTO @Saagie
https://www.linkedin.com/in/youenchene/
Aurélien Vandel, Data Engineer
https://www.linkedin.com/in/aur%C3%A9lien-vandel-060b5b8a/
01:11:17 End
Please use the Contact Form on this blog or our twitter feed to send us your questions, or to suggest future episode topics you would like us to cover.