Data ScientistWashington DC or Remote Full TimeResponsibilities:Modeling complex problems, discovering insights and identifying opportunities through the use of statistical, algorithmic, mining and visualization techniquesParticipating in the areas of architecture, design, implementation, and testingProposing innovative ways to look at problems by using data mining approaches on the set of information availableDesigning experiments, testing hypotheses, and building modelsConducting advanced data analysis and designing highly complex algorithmApplying advanced statistical and predictive modeling techniques to build, maintain, and improve on multiple real-time decision systemsQualifications:At least 5 years experience in Java, Spring and MySQL (or any relational database) and PythonExperience with databases (including NoSQL)Experience in machine learning frameworks and librariesSupervised and Unsupervised learningMachine learning concepts and techniques: Regularization, Boosting, Random Forests, Decision Trees, Bayesian models, Neural networks, Support Vector Machines (SVM)Experience with the whole ETL data cycle (extract, validate, transform, clean, aggregate, audit, archive)Computer Science or Mathematics or Physics degreeExcellent communication and analytical skillsWillingness to work hard (50 hrs per week)Very good EnglishNice To Have (Not Required):Experience with Apache SparkNatural Language Processing (tokenization, tagging, sentiment analysis, entity recognition, summarization)R programming language