TrueschoTruescho
All Courses
Big Data Integration and Processing
Coursera
Course
Unknown

Big Data Integration and Processing

University of California San Diego

Learn to retrieve data from databases and big data management systems and understand data management operations and processing patterns for large-scale analytics.

Unknown7 weeksEnglish82,632 enrolled

About this Course

At the end of the course, you will be able to: Retrieve data from example database and big data management systems Describe the connections between data management operations and the big data processing patterns needed to utilize them in large-scale analytical applications Identify when a big data problem needs data integration Execute simple big data integration and processing on Hadoop and Spark platforms This course is for those new to data science. Completion of Intro to Big Data is recommended. No prior programming experience is needed, although the ability to install applications and utilize a virtual machine is necessary to complete the hands-on assignments. Refer to the specialization technical requirements for complete hardware and software specifications. Hardware Requirements: (A) Quad Core Processor (VT-x or AMD-V support recommended), 64-bit; (B) 8 GB RAM; (C) 20 GB disk free. How to find your hardware information: (Windows): Open System by clicking the Start button, right-clicking Computer, and then clicking Properties; (Mac): Open Overview by clicking on the Apple menu and clicking “About This Mac.” Most computers with 8 GB RAM purchased in the last 3 years will meet the minimum requirements.You will need a high speed internet connection because you will be downloading files up to 4 Gb in size. Software Requirements: This course relies on several open-source software tools, including Apache Hadoop. All required software can be downloaded and installed free of charge (except for data charges from your internet provider). Software requirements include: Windows 7+, Mac OS X 10.10+, Ubuntu 14.04+ or CentOS 6+ VirtualBox 5+

What You'll Learn

  • Retrieve data from databases and big data management systems
  • Describe connections between data management operations and big data processing patterns
  • Identify when data integration is needed for big data problems
  • Execute big data integration and processing with Hadoop and Spark

Prerequisites

  • Basic computer and internet skills
  • Ability to read English instructions and complete short practice activities

Instructors

I

Ilkay Altintas

Chief Data Science Officer

A

Amarnath Gupta

Director, Advanced Query Processing Lab

Topics

Data Analysis
Data Science
Big Data
SQL
Data Processing
Database Systems
Pandas (Python Package)
Data Pipelines
Data Integration
NoSQL

Course Info

PlatformCoursera
LevelUnknown
PacingUnknown
PriceFree

Skills

تحليل البيانات
علم البيانات
البيانات الكبيرة
SQL
معالجة البيانات
أنظمة قواعد البيانات
بايثون بانداز
قنوات البيانات
Data Integration
NoSQL

Start Learning Now