Josh Yudaken | Building domain specific databases using Python for prototypes during take off
474 views · Published 24 August 2016 · 33:19 · Indexed 20 September 2026
Channel: PyData · 2016 · Science & Technology
PyData SF 2016 Smyte is blocking fraudsters, spammers, scammers and harassers through analysis of web and mobile event data. There are hundreds of new database options cropping up, and while most of them are great they all make a range of different design choices which may/may not align with what you need. At Smyte we're solving infrastructure-heavy problems with a tiny engineering team. We've had to build custom databases in order to efficiently solve problems, but don't have the luxury of a ton of development time. In this talk I'll go through how we used Kafka & Python to implement a prototype Sliding HyperLogLog server. It held up for the six months we needed to, and some early decisions made it trivial to port the server to our new efficient C++ & RocksDB stack. 00:00 Welcome! 00:10 Help us add time stamps or captions to this video! See the description for details. Want to help add timestamps to our YouTube videos to help with discoverability? Find out more here: https://github.com/numfocus/YouTubeVideoTimestamps
More from this channel
-
24:46
Dino Viehland & Raymond Laghaeian: Jupyter Notebooks and ML Model Operationalization
-
38:29
Simon Byrne - Julia for data analysis
-
49:10
Rui Miguel Forte - The CV: A Data Scientist's View
-
50:23
PyData London 2016 Lightning Talks and Closing Address
-
38:04
Peter Wang | Keynote: Python for Pythonistas
-
39:16
Taposh Roy, Austin Powell | A hybrid approach to model randomness and fuzziness
-
41:39
Chris Fregly - High Performance Distributed Tensorflow
-
1:28:19
Tricks, tips and topics in Text Analysis - Bhargav Srinivasa Desikan