Big Data Integration and Processing
The Big Data Integration and Processing course is for those new to data science. Completion of Intro to Big Data is recommended. No prior programming experience is needed, although the ability to install applications and utilize a virtual machine is necessary to complete the hands-on assignments. Refer to the specialization technical requirements for complete hardware and software specifications.
With hardware requirements: Quad-Core Processor (VT-x or AMD-V support recommended), 64-bit; 8GB RAM; 20 GB disk free. You will need a high-speed internet connection because you will be downloading files up to 4 Gb in size. This course relates to several open-source software tools, including Apache Hadoop. All required software can be downloaded and installed free of charge (except for data charges from your internet provider). Software requirements include: Windows 7+, Mac OS X 10.10+, Ubuntu 14.04+ or CentOS 6+ VirtualBox 5+.
At the end of the course, you will be able to:
- Retrieve data from example database and big data management systems
- Describe the connections between data management operations and the big data processing patterns needed to utilize them in large-scale analytical applications
- Identify when a big data problem needs data integration
- Execute simple big data integration and processing on Hadoop and Spark platforms
This course offers:
- Flexible deadlines: Reset deadlines based on your availability.
- 100% online
- Get a Certificate when you finish
- Beginner level
- Approximately 18 hours to complete
- Subtitles: Arabic, French, Portuguese (European), Italian, Vietnamese, Korean, German, Russian, English, Spanish
Rating: 4.4/5
Participants: 67,301
Enroll here: https://www.coursera.org/learn/big-data-integration-processing