DataWeek Conference + Expo has ended
Back To Schedule
Wednesday, September 17 • 4:00pm - 4:40pm
TALK: Data Science at Scale! Creating Complex Machine Scoring Applications with Cascading Pattern

Sign up or log in to save this to your schedule, view media, leave feedback and see who's attending!

Cascading Pattern is an open source project that takes models trained in popular analytics frameworks, such as R, SAS, Microstrategy, etc., and runs them at scale on Apache Hadoop. With Pattern, developers can use a Java API to create complex machine learning applications, such as recommenders or fraud detection. Pattern effectively lowers the barrier of adoption to Apache Hadoop for developers because developers can use existing skillsets to immediately begin building these complex applications.

This machine-learning library works by translating PMML – an established XML standard for predictive model markup – into data workflows based on the Cascading API in Java. PMML models can be run in a pre-defined JAR file with no coding required. PMML can also be combined with other flows based on ANSI SQL (Cascading Lingual), Scala (Scalding), Clojure (Cascalog), etc. Multiple companies have collaborated to implement parallelized algorithms: Random Forest, Logistic Regression, K-Means, Hierarchical Clustering, etc. Benefits include greatly reduced development costs and less licensing issues at scale – while leveraging a combination of Apache Hadoop clusters, existing intellectual property in predictive models, and the core competencies of analytics staff. 

In this presentation, Concurrent, Inc.’s Supreet Oberoi, will provide sample code that will show applications using predictive models built in SAS and R, such as anti-fraud classifiers. Additionally, Supreet will compare variations of models for enterprise-class customer experiments.


Supreet Oberoi

Vice President of Field Engineering, Concurrent, Inc.
Supreet Oberoi is a hands-on, entrepreneurial, technology leader with over two decades of experience in successfully developing transformative information technologies, and working in leadership roles at Concurrent Inc., American Express, Oracle, Microsoft and many privately held... Read More →

Wednesday September 17, 2014 4:00pm - 4:40pm PDT
Hotel Kabuki

Attendees (0)