During this course you will:
- Identify practical problems which can be solved with machine learning
- Build, tune and apply linear models with Spark MLLib
- Understand methods of text processing
- Fit decision trees and boost them with ensemble learning
- Construct your own recommender system.
As a practical assignment, you will
- build and apply linear models for classification and regression tasks;
- learn how to work with texts;
- automatically construct decision trees and improve their performance with ensemble learning;
- finally, you will build your own recommender system!
With these skills, you will be able to tackle many practical machine learning tasks.
We provide the tools, you choose the place of application to make this world of machines more intelligent.
Special thanks to:
- Prof. Mikhail Roytberg, APT dept., MIPT, who was the initial reviewer of the project, the supervisor and mentor of half of the BigData team. He was the one, who helped to get this show on the road.
- Oleg Sukhoroslov (PhD, Senior Researcher at IITP RAS), who has been teaching MapReduce, Hadoop and friends since 2008. Now he is leading the infrastructure team.
- Oleg Ivchenko (PhD student APT dept., MIPT), Pavel Akhtyamov (MSc. student at APT dept., MIPT) and Vladimir Kuznetsov (Assistant at P.G. Demidov Yaroslavl State University), superbrains who have developed and now maintain the infrastructure used for practical assignments in this course.
- Asya Roitberg, Eugene Baulin, Marina Sudarikova. These people never sleep to babysit this course day and night, to make your learning experience productive, smooth and exciting.