CSCI 372
Big Data Management and Analytics
New York Institute of Technology · UGRD · Fall 2026
Catalog description
Organizations today are generating massive amounts of data that are too large and unstructured to fit in relational databases. Existing platforms such as the Apache Hadoop allow for distributed and enterprise systems processing, of large data sets across clusters of computers, using the MapReduce programming model. In this course, students will gain an in-depth understanding of how MapReduce and Distributed File Systems work. In addition, they will be able to author Hadoop-based MapReduce applications in Java and use Hadoop subprojects Hive and Pig to build powerful data processing applications. Industry systems, such as IBM InfoSphere BigInsights and IBM InfoSphere Streams will be studied. Classroom Hours - Laboratory and/or Studio Hours – Course Credits: 3-0-3
Sections
Current meeting, instructor, credit, and enrollment details
001
Availability not recently verified- Days & times
- No scheduled meeting time
- Meeting dates
- —
- Location
- —
- Instructor
- Staff