Chevron Left
Back to Introduction to Big Data

Learner Reviews & Feedback for Introduction to Big Data by University of California San Diego

10,693 ratings

About the Course

Interested in increasing your knowledge of the Big Data landscape? This course is for those new to data science and interested in understanding why the Big Data Era has come to be. It is for those who want to become conversant with the terminology and the core concepts behind big data problems, applications, and systems. It is for those who want to start thinking about how Big Data might be useful in their business or career. It provides an introduction to one of the most common frameworks, Hadoop, that has made big data analysis easier and more accessible -- increasing the potential for data to transform our world! At the end of this course, you will be able to: * Describe the Big Data landscape including examples of real world big data problems including the three key sources of Big Data: people, organizations, and sensors. * Explain the V’s of Big Data (volume, velocity, variety, veracity, valence, and value) and why each impacts data collection, monitoring, storage, analysis and reporting. * Get value out of Big Data by using a 5-step process to structure your analysis. * Identify what are and what are not big data problems and be able to recast big data problems as data science questions. * Provide an explanation of the architectural components and programming models used for scalable big data analysis. * Summarize the features and value of core Hadoop stack components including the YARN resource and job management system, the HDFS file system and the MapReduce programming model. * Install and run a program using Hadoop! This course is for those new to data science. No prior programming experience is needed, although the ability to install applications and utilize a virtual machine is necessary to complete the hands-on assignments. Hardware Requirements: (A) Quad Core Processor (VT-x or AMD-V support recommended), 64-bit; (B) 8 GB RAM; (C) 20 GB disk free. How to find your hardware information: (Windows): Open System by clicking the Start button, right-clicking Computer, and then clicking Properties; (Mac): Open Overview by clicking on the Apple menu and clicking “About This Mac.” Most computers with 8 GB RAM purchased in the last 3 years will meet the minimum requirements.You will need a high speed internet connection because you will be downloading files up to 4 Gb in size. Software Requirements: This course relies on several open-source software tools, including Apache Hadoop. All required software can be downloaded and installed free of charge. Software requirements include: Windows 7+, Mac OS X 10.10+, Ubuntu 14.04+ or CentOS 6+ VirtualBox 5+....

Top reviews


Aug 11, 2021

I love the course. It goes deep into the foundations, and then finishes up with an actual lab where you learn by practice. I greatly benefited from it and feel I have achieved a milestone in big data.


Sep 8, 2019

I love the course. It goes deep into the foundations, and then finishes up with an actual lab where you learn by practice. I greatly benefited from it and feel I have achieved a milestone in big data.

Filter by:

1726 - 1750 of 2,444 Reviews for Introduction to Big Data


Feb 14, 2019





By Kusam C M R

Dec 17, 2018


By Ronald S

Jan 16, 2017


By Enrique G R

Jun 30, 2016


By Gella A P

Sep 14, 2022


By Mohamed E

Jun 4, 2021


By Yaakoubi A

Oct 11, 2020


By Pavithra M

May 24, 2020


By Hongyang Y

Feb 7, 2018


By Marilyn V

Oct 17, 2016



Jan 14, 2021


By Adil A

Jun 9, 2020


By Basawaraj P

Sep 20, 2018


By Rajeev D

Aug 23, 2018


By Divya A P

Jul 8, 2018


By Janet S

Feb 5, 2018


By Farana A

Oct 9, 2017


By Kavita T

Aug 2, 2017



Jul 31, 2017


By João P S S

Jul 5, 2017


By Rohan B

Jun 22, 2017



Jan 8, 2017


By Suad A

Dec 13, 2016


By 현아 이

Aug 9, 2016


By Martin D

Aug 12, 2020

For a newcomer like me in this subject, it was an interesting learning, after this first course I am even more motivated to continue and to complete this specialization. As English is not my first language, I could still note some errors in the subtitle script.

I got difficulties with the last Quizz (the 2 questions on Hadoop) since the "out" directory already exist (created throughout the step by step videos and lectures) I could'nt process with the alice.txt question, the system can't overwrite the existing files and exixting directory. I had to delete first the out directory (thanks google with linux commands I don't know) with hadoop fs -rm -r out command. I had to repeat the same for the question #2 (Shakespeare text). Please udpate your instructions to clean the place prior to start an other exercice.

Thank you