Help developers discover the tools you use. Get visibility for your team's tech choices and contribute to the community's knowledge.
Impala is a modern, open source, MPP SQL query engine for Apache Hadoop. Impala is shipped by Cloudera, MapR, and Amazon. With Impala, you can query data, whether stored in HDFS or Apache HBase – including SELECT, JOIN, and aggregate functions – in real time. | It is a stand-alone platform for data observability and monitoring. It ensures data quality by automatically monitoring your data stack for anomalies, alerting you of issues, and analyzing the root cause. |
Do BI-style Queries on Hadoop;Unify Your Infrastructure;Implement Quickly;Count on Enterprise-class Security;Retain Freedom from Lock-in;Expand the Hadoop User-verse | Profile & setup basic data quality monitors in minutes;
No-code, stand-alone full web application;
Automated, built-in scheduler & jobs;
Integrate & monitor your entire data stack |
Statistics | |
GitHub Stars 34 | GitHub Stars 326 |
GitHub Forks 33 | GitHub Forks 34 |
Stacks 145 | Stacks 1 |
Followers 301 | Followers 6 |
Votes 18 | Votes 0 |
Pros & Cons | |
Pros
| No community feedback yet |
Integrations | |

Grafana is a general purpose dashboard and graph composer. It's focused on providing rich ways to visualize time series metrics, mainly though graphs but supports other ways to visualize data through a pluggable panel architecture. It currently has rich support for for Graphite, InfluxDB and OpenTSDB. But supports other data sources via plugins.

Kibana is an open source (Apache Licensed), browser based analytics and search dashboard for Elasticsearch. Kibana is a snap to setup and start using. Kibana strives to be easy to get started with, while also being flexible and powerful, just like Elasticsearch.

Prometheus is a systems and service monitoring system. It collects metrics from configured targets at given intervals, evaluates rule expressions, displays the results, and can trigger alerts if some condition is observed to be true.

Spark is a fast and general processing engine compatible with Hadoop data. It can run in Hadoop clusters through YARN or Spark's standalone mode, and it can process data in HDFS, HBase, Cassandra, Hive, and any Hadoop InputFormat. It is designed to perform both batch processing (similar to MapReduce) and new workloads like streaming, interactive queries, and machine learning.

Nagios is a host/service/network monitoring program written in C and released under the GNU General Public License.

Netdata collects metrics per second & presents them in low-latency dashboards. It's designed to run on all of your physical & virtual servers, cloud deployments, Kubernetes clusters & edge/IoT devices, to monitor systems, containers & apps

Zabbix is a mature and effortless enterprise-class open source monitoring solution for network monitoring and application monitoring of millions of metrics.

Distributed SQL Query Engine for Big Data

Sensu is the future-proof solution for multi-cloud monitoring at scale. The Sensu monitoring event pipeline empowers businesses to automate their monitoring workflows and gain deep visibility into their multi-cloud environments.

Amazon Athena is an interactive query service that makes it easy to analyze data in Amazon S3 using standard SQL. Athena is serverless, so there is no infrastructure to manage, and you pay only for the queries that you run.