Publication | Open Access
Multi-dimensional geospatial data mining in a distributed environment using MapReduce
31
Citations
40
References
2019
Year
Cluster ComputingApache Hadoop EcosystemEngineeringMachine LearningMap-reduceDistributed Data AnalyticsSocial SciencesUnsupervised Machine LearningGeographic Information SystemsImage AnalysisData ScienceData MiningPattern RecognitionMachine Learning TechniquesSpatial Data ManagementDistributed EnvironmentCartographyGeographyKnowledge DiscoveryComputer ScienceHyperspectral ImagingData ClassificationRemote SensingGeospatial DataMassive Data ProcessingBig Data
Data mining and machine learning techniques for processing raster data consider a single spectral band of data at a time. The individual results are combined to obtain the final output. The essence of related multi-spectral information is lost when the bands are considered independently. The proposed platform is based on Apache Hadoop ecosystem and supports performing analysis on large amounts of multispectral raster data using MapReduce. A novel technique of transforming the spectral space to the geometrical space is also proposed. The technique allows to consider multiple bands coherently. The results of clustering 106 pixels for multiband imagery with widely used GIS software have been tested and other machine learning methods are planned to be incorporated in the platform. The platform is scalable to support tens of spectral bands. The results from our platform were found to be better and are also available faster due to application of distributed processing.
| Year | Citations | |
|---|---|---|
Page 1
Page 1