Job Details
Key Responsibilities
- Work with stakeholders within the organization and more importantly with our customers to identify opportunities for leveraging customer?s machine data to nuggets of information.
- Mine and analyze data from machine log data to drive optimization and improvement of product development, reduce failures, proactively identify failures, and predict trends or Anomalies.
- Assess the effectiveness and accuracy of new data sources and data gathering techniques.
- Develop custom data models and algorithms to apply to data sets.
- Use predictive modeling to increase and optimize customer experiences, revenue generation, ad targeting and other business outcomes.
- Develop company A/B testing framework and test model quality.
- Coordinate with different functional teams to implement models and monitor outcomes.
- Develop processes and tools to monitor and analyze model performance and data accuracy.
Our Ideal Candidate:
- Strong problem solving skills with an emphasis on product development.
- We?re looking for someone with 5-7 years of experience manipulating data sets and building statistical models, has a Master?s or PHD in Statistics Mathematics, Computer Science or another quantitative field.
- Experience using statistical computer languages (R, Spark MLLib, Python, SLQ, etc.) to manipulate data and draw insights from large data sets.
- Knowledge of a variety of machine learning techniques (clustering, decision tree learning, artificial neural networks, etc.) and their real-world advantages/drawbacks.
- Knowledge of advanced statistical techniques and concepts (regression, properties of distributions, statistical tests and proper usage, etc.) and experience with applications.
- Excellent written and verbal communication skills for coordinating across teams.
- A drive to learn and master new technologies and techniques.
- Coding knowledge and experience with several languages: C, C++, Java, Scala, JavaScript, etc.
- Knowledge and experience in statistical and data mining techniques: GLM/Regression, Random Forest, Boosting, Trees, text mining, social network analysis, etc.
- Experience querying databases and using statistical computer languages: R, Python, SLQ, etc.
- Experience visualizing/presenting data for stakeholders using: Tableau, D3, ggplot, etc.