Skip to content

msaber1/Impala

This branch is 4052 commits ahead of, 2 commits behind KarthikTunga/impala:master.

Folders and files

NameName
Last commit message
Last commit date

Latest commit

Bharath VissapragadaInternal Jenkins
Bharath Vissapragada
and
Internal Jenkins
Dec 29, 2015
552636b · Dec 29, 2015
Dec 22, 2015
Dec 22, 2015
Dec 15, 2015
Dec 19, 2015
May 20, 2015
Dec 29, 2015
Dec 17, 2015
Jun 1, 2012
Dec 22, 2015
Jan 8, 2014
Dec 29, 2015
Dec 19, 2015
Sep 16, 2015
Nov 5, 2015
Dec 22, 2015
Dec 25, 2015
May 8, 2014
Jul 2, 2014
Mar 23, 2015
Dec 17, 2015

Repository files navigation

Welcome to Impala

Lightning-fast, distributed SQL queries for petabytes of data stored in Apache Hadoop clusters.

Impala is a modern, massively-distributed, massively-parallel, C++ query engine that lets you analyze, transform and combine data from a variety of data sources:

  • Best of breed performance and scalability.
  • Support for data stored in HDFS, Apache HBase and Amazon S3.
  • Wide analytic SQL support, including window functions and subqueries.
  • On-the-fly code generation using LLVM to generate CPU-efficient code tailored specifically to each individual query.
  • Support for the most commonly-used Hadoop file formats, including the Apache Parquet (incubating) project.
  • Apache-licensed, 100% open source.

More about Impala

To learn more about Impala as a business user, or to try Impala live or in a VM, please visit the Impala homepage.

If you are interested in contributing to Impala as a developer, or learning more about Impala's internals and architecture, visit the Impala wiki.

About

Real-time Query for Hadoop

Resources

License

Stars

Watchers

Forks

Packages

No packages published

Languages

  • C++ 54.4%
  • Java 25.9%
  • Python 13.2%
  • Thrift 1.7%
  • C 1.4%
  • Shell 1.2%
  • Other 2.2%