Chevron Left
返回到 Big Data Analysis: Hive, Spark SQL, DataFrames and GraphFrames

學生對 Yandex 提供的 Big Data Analysis: Hive, Spark SQL, DataFrames and GraphFrames 的評價和反饋

161 個評分
36 條評論


No doubt working with huge data volumes is hard, but to move a mountain, you have to deal with a lot of small stones. But why strain yourself? Using Mapreduce and Spark you tackle the issue partially, thus leaving some space for high-level tools. Stop struggling to make your big data workflow productive and efficient, make use of the tools we are offering you. This course will teach you how to: - Warehouse your data efficiently using Hive, Spark SQL and Spark DataFframes. - Work with large graphs, such as social graphs or networks. - Optimize your Spark applications for maximum performance. Precisely, you will master your knowledge in: - Writing and executing Hive & Spark SQL queries; - Reasoning how the queries are translated into actual execution primitives (be it MapReduce jobs or Spark transformations); - Organizing your data in Hive to optimize disk space usage and execution times; - Constructing Spark DataFrames and using them to write ad-hoc analytical jobs easily; - Processing large graphs with Spark GraphFrames; - Debugging, profiling and optimizing Spark application performance. Still in doubt? Check this out. Become a data ninja by taking this course! Special thanks to: - Prof. Mikhail Roytberg, APT dept., MIPT, who was the initial reviewer of the project, the supervisor and mentor of half of the BigData team. He was the one, who helped to get this show on the road. - Oleg Sukhoroslov (PhD, Senior Researcher at IITP RAS), who has been teaching MapReduce, Hadoop and friends since 2008. Now he is leading the infrastructure team. - Oleg Ivchenko (PhD student APT dept., MIPT), Pavel Akhtyamov (MSc. student at APT dept., MIPT) and Vladimir Kuznetsov (Assistant at P.G. Demidov Yaroslavl State University), superbrains who have developed and now maintain the infrastructure used for practical assignments in this course. - Asya Roitberg, Eugene Baulin, Marina Sudarikova. These people never sleep to babysit this course day and night, to make your learning experience productive, smooth and exciting....



Nov 13, 2018

content of the course is remarkable and the way they explained concepts is very lucid. I just want to give suggestions please give link to the data set they are using for illustrating the concepts.


Feb 03, 2018

I wish I could give more rating than 5 :). Excellent course. Thanks so much for such an excellent course. All the instructors are great.


26 - Big Data Analysis: Hive, Spark SQL, DataFrames and GraphFrames 的 36 個評論(共 36 個)

創建者 Viacheslav I

Dec 09, 2017

Quite a good course, but weeks about graphs I found somewhat poor. Apart from that you learn basics of Hive, details of DataFrame API of Spark and how it internally works.

創建者 Kiselyov A

May 23, 2019

Nice course, but the impression about practical tasks is really awful. The tasks are ok, but grading system is too buggy

創建者 Aldrin

Jul 30, 2019

Good course...Deep dive into optimization of spark for production environment.Also interesting graph implementations

創建者 Luis M A P

Mar 27, 2019

Good. Please fix assignments explanations. i.e In week 5.

創建者 Adarsh G

Dec 29, 2019

Wonderful course for new learners. Thanks !!!

創建者 Дюкарев В В

Sep 16, 2019

This is 2nd course of specialization and its better than 1st, but big big problems with outer grader system is still there :(

Another minus is that there is too much theory (graph algorithms especially) and too less practice, i think 1/10. You know theory is forgotten very fast.

Nevertheless its good, valuable course for review of technologies

創建者 Evgenii K

May 15, 2018

Grading system it terrible, it hadn't work for a week summary, no help from staff on Coursera forum, only Slack channel could help. Theoretical material is quite good.


Aug 26, 2019

Few Environment not supporting frequently.

創建者 Григорьева М М

Nov 04, 2019

Didn't like the course for several reasons:

1. Mentors seem to leave the course. Do not expect any feedback from them.

2. Notebooks are broken for months so you can't perform assignments without docker.

3. Weeks are not consistent, the material quality differs enormally from week to week.

創建者 Haithem H

May 13, 2020

There is many issues in LTI grader. At the begining, I have been attracted by the content, and as far as I moved on the grader has pushed every thing down

創建者 Lontsi s c

Apr 30, 2020

Le cours n'est pas bien structuré