lucasdang311 / shark

Hive on Spark

Home Page:http://shark.cs.berkeley.edu/

Geek Repo:Geek Repo

Github PK Tool:Github PK Tool

Shark (Hive on Spark)

Shark is a large-scale data warehouse system for Spark designed to be compatible with Apache Hive. It can answer Hive QL queries up to 100 times faster than Hive without modification to the existing data nor queries. Shark supports Hive's query language, metastore, serialization formats, and user-defined functions.

Shark 0.8.0 requires:

  • Scala 2.9.3
  • Hive 0.9
  • Spark 0.8.x

For current documentation, see the Shark Project Wiki

About

Hive on Spark

http://shark.cs.berkeley.edu/

License:Apache License 2.0