fbpx
How I Fight Slavery – Eric Schles ODSC Boston 2015
http://bit.ly/EricSchlesODSCTalk In this talk I will be covering a set to techniques I’ve used to track and find instances of human trafficking on the internet. I’ll be going over web scraping, entity recognition, some techniques for text comparison, and data storage concerns. The tool that I will be explaining... Read more
Kaggle The Home of Data Science – Anthony Goldbloom ODSC Boston 2015
Kaggle The Home of Data Science from odsc Keynote Presenter Bio Anthony Goldbloom is the founder and CEO of Kaggle. In 2011 & 2012, Forbes Magazine named Anthony as one of the 30 under 30 in technology, in 2013 the MIT Tech Review named him one of top 35... Read more
Intro to Text Mining Using tm, openNLP and topicmodels – Ted Kwartler ODSC Boston 2015
Intro to Text Mining Using tm, openNLP and topicmodels from odsc You will learn how modern customer service organizations use data to understand important customer attributes and how R is used for workforce optimization. Topics include real world examples of how R is used in large scale operations to... Read more
The Art of Data Science – Josh Wills ODSC Boston 2015
The Art of Data Science from odsc Keynote Presenter Bio Josh Wills is Cloudera’s Senior Director of Data Science, working with customers and engineers to develop Hadoop-based solutions across a wide-range of industries. He is the founder and VP of the Apache Crunch project for creating optimized MapReduce pipelines... Read more
Jumping to Conclusions – Richard Robehr Bijjani ODSC Boston 2015
Jumping to Conclusions from odsc Data Science is the study of the extraction of knowledge from data. What if we extract partial or inaccurate knowledge? This illusion of knowledge would lead us to make wrong decisions, with sometimes disastrous consequences such as in the case of medical diagnosis, security... Read more
Machine Learning Based Personalization Using Uplift Analytics: Examples and Applications – Victor Lo ODSC Boston 2015
Uplift Modeling Workshop from odsc Traditional randomized experiments allow us to determine the overall causal impact of a treatment program (e.g. marketing, medical, social, education, political). Uplift modeling (also known as true lift, net lift, incremental lift) takes a further step to identify individuals who are truly positively influenced... Read more
Data Science 101 – Todd Cioffi ODSC Boston 2015
Data Science 101 from odsc Curious about Data Science? Self-taught on some aspects, but missing the big picture? Well, you’ve got to start somewhere and this session is the place to do it. This session will cover, at a layman’s level, some of the basic concepts of Data Science.... Read more
Can We Automate Predictive Analytics – Thomas Dinsmore ODSC Boston 2015
Can We Automate Predictive Analytics from odsc Recent news about the pending shortage of data scientists prompts speculation about automation: will machines replace human analysts? We propose a model of automation, and briefly review progress in automated machine learning over the past twenty years. Summarizing the current state of... Read more
Learning to Love Bayesian Statistics – Allen Downey ODSC Boston 2015
http://tinyurl.com/lovebayes Bayesian statistical methods provide powerful tools for answering questions and making decisions. For example, the result of Bayesian analysis is a set of values and probabilties that can be fed directly into a cost-benefit analysis, which is not possible with conventional statistics. But there are several barriers to... Read more
Predictive Modeling Workshop – Max Kuhn ODSC Boston 2015
Predictive Modeling Workshop from odsc The workshop is an overview of creating predictive models using R. An example data set will be used to demonstrate a typical workflow: data splitting, pre-processing, model tuning and evaluation. Several R packages will be shown along with the caret package which provides a... Read more