Publication | Closed Access
HadoopDB in action
62
Citations
7
References
2010
Year
Unknown Venue
Cluster ComputingEngineeringCloud DatabaseMap-reduceSemantic WebData ScienceDatabase SupportData-intensive PlatformData IntegrationParallel ComputingData ManagementEmbedded DatabaseParallel DatabasesCloud ComputingBusiness DataParallel ProgrammingFlexible ArchitectureMassive Data ProcessingBig Data
HadoopDB is a hybrid of MapReduce and DBMS technologies, designed to meet the growing demand of analyzing massive datasets on very large clusters of machines. Our previous work has shown that HadoopDB approaches parallel databases in performance and still yields the scalability and fault tolerance of MapReduce-based systems. In this demonstration, we focus on HadoopDB's flexible architecture and versatility with two real world application scenarios: a semantic web data application for protein sequence analysis and a business data warehousing application based on TPC-H. The demonstration offers a thorough walk-through of how to easily build applications on top of HadoopDB.
| Year | Citations | |
|---|---|---|
Page 1
Page 1