Databases are not only tools for data intensive science but are also subject to
active research. The exponentially growing data volumes make scale-out parallelism
necessary. Even by running data analysis on clusters of machines, existing algorithms
cannot always be scaled up to the problems which requires new approaches. Learn
about how to use distributed system to process large amounts of data and how to tweak
existing relational databases into parallel data warehouses.