All Projects → Flume → Similar Projects or Alternatives

380 Open source projects that are alternatives of or similar to Flume

A complete (distributed) BigData stack, running in containers

Stars: ✭ 14 (-99.36%)

Mutual labels: big-data, flume

大数据入门指南 ⭐

Stars: ✭ 10,991 (+399.59%)

Mutual labels: big-data, flume

Hazelcast Nodejs Client

Hazelcast IMDG Node.js Client

Stars: ✭ 124 (-94.36%)

Mutual labels: big-data

Sparkling Graph

SparklingGraph provides easy to use set of features that will give you ability to proces large scala graphs using Spark and GraphX.

Stars: ✭ 139 (-93.68%)

Mutual labels: big-data

Asakusa Framework

Stars: ✭ 114 (-94.82%)

Mutual labels: big-data

Feature Store for Machine Learning

Stars: ✭ 2,576 (+17.09%)

Mutual labels: big-data

🐳 big data study

Stars: ✭ 141 (-93.59%)

Mutual labels: big-data

HDFS Shell is a HDFS manipulation tool to work with functions integrated in Hadoop DFS

Stars: ✭ 117 (-94.68%)

Mutual labels: big-data

Tools and Docs on the Azure Data Science Virtual Machine (http://aka.ms/dsvm)

Stars: ✭ 153 (-93.05%)

Mutual labels: big-data

Distributed Big Data Orchestration Service

Stars: ✭ 1,544 (-29.82%)

Mutual labels: big-data

Attic Apex Malhar

Mirror of Apache Apex malhar

Stars: ✭ 131 (-94.05%)

Mutual labels: big-data

Tennis Crystal Ball

Ultimate Tennis Statistics and Tennis Crystal Ball - Tennis Big Data Analysis and Prediction

Stars: ✭ 107 (-95.14%)

Mutual labels: big-data

Mirror of Apache Tajo

Stars: ✭ 128 (-94.18%)

Mutual labels: big-data

Mirror of Apache Metamodel

Stars: ✭ 143 (-93.5%)

Mutual labels: big-data

High-performance Terrain and Hydrology Analysis

Stars: ✭ 127 (-94.23%)

Mutual labels: big-data

A Clojure dataframe library that runs on Spark

Stars: ✭ 152 (-93.09%)

Mutual labels: big-data

Scala Spark Tutorial

Project for James' Apache Spark with Scala course

Stars: ✭ 121 (-94.5%)

Mutual labels: big-data

Big Data Toolkit for the JVM

Stars: ✭ 140 (-93.64%)

Mutual labels: big-data

CMAK is a tool for managing Apache Kafka clusters

Stars: ✭ 10,544 (+379.27%)

Mutual labels: big-data

GeoTrellis for PySpark

Stars: ✭ 167 (-92.41%)

Mutual labels: big-data

repo for code published on pythondata.com

Stars: ✭ 113 (-94.86%)

Mutual labels: big-data

Spark On Lambda

Apache Spark on AWS Lambda

Stars: ✭ 137 (-93.77%)

Mutual labels: big-data

Spark R Notebooks

R on Apache Spark (SparkR) tutorials for Big Data analysis and Machine Learning as IPython / Jupyter notebooks

Stars: ✭ 109 (-95.05%)

Mutual labels: big-data

Spark With Python

Fundamentals of Spark with Python (using PySpark), code examples

Stars: ✭ 150 (-93.18%)

Mutual labels: big-data

A framework for rapid reporting API development; with out of the box support for high cardinality dimension lookups with druid.

Stars: ✭ 101 (-95.41%)

Mutual labels: big-data

Calcite Avatica

Mirror of Apache Calcite - Avatica

Stars: ✭ 130 (-94.09%)

Mutual labels: big-data

Graph Sampling is a python package containing various approaches which samples the original graph according to different sample sizes.

Stars: ✭ 99 (-95.5%)

Mutual labels: big-data

Couchdb Documentation

Apache CouchDB Documentation

Stars: ✭ 128 (-94.18%)

Mutual labels: big-data

A visual ETL development and debugging tool for big data

Stars: ✭ 144 (-93.45%)

Mutual labels: big-data

Samples and Docs for Azure Data Lake Store and Analytics

Stars: ✭ 128 (-94.18%)

Mutual labels: big-data

The official home of the Presto distributed SQL query engine for big data

Stars: ✭ 12,957 (+488.95%)

Mutual labels: big-data

Griffon Data Science Virtual Machine

Stars: ✭ 128 (-94.18%)

Mutual labels: big-data

Apache Storm 官方文档中文版

Stars: ✭ 142 (-93.55%)

Mutual labels: big-data

🐳 Tool to automate data quality checks on data pipelines

Stars: ✭ 123 (-94.41%)

Mutual labels: big-data

Attic Predictionio

PredictionIO, a machine learning server for developers and ML engineers.

Stars: ✭ 12,522 (+469.18%)

Mutual labels: big-data

自动化配置报表平台。演示地址http://58.87.112.247/report 账号 visitor密码123456

Stars: ✭ 123 (-94.41%)

Mutual labels: big-data

Belajarpython.com

Open Source Indonesian Python Programming Tutorial Site

Stars: ✭ 141 (-93.59%)

Mutual labels: big-data

The Signal Metadata Format Specification

Stars: ✭ 120 (-94.55%)

Mutual labels: big-data

Julia binding for Apache Spark

Stars: ✭ 153 (-93.05%)

Mutual labels: big-data

Apache Drill is a distributed MPP query layer for self describing data

Stars: ✭ 1,619 (-26.41%)

Mutual labels: big-data

Hazelcast Go Client

Hazelcast IMDG Go Client

Stars: ✭ 140 (-93.64%)

Mutual labels: big-data

Amazon S3 Find And Forget

Amazon S3 Find and Forget is a solution to handle data erasure requests from data lakes stored on Amazon S3, for example, pursuant to the European General Data Protection Regulation (GDPR)

Stars: ✭ 115 (-94.77%)

Mutual labels: big-data

Distributed, Versioned, Image-oriented Dataservice

Stars: ✭ 174 (-92.09%)

Mutual labels: big-data

📊 📋 Dashboards using YAML or JSON files

Stars: ✭ 1,511 (-31.32%)

Mutual labels: big-data

A search engine which can hold 100 trillion lines of log data.

Stars: ✭ 1,793 (-18.5%)

Mutual labels: big-data

Mirror of Apache Ambari

Stars: ✭ 1,576 (-28.36%)

Mutual labels: big-data

Easily make RESTful web services for time series reporting with Big Data analytics engines like Druid and SQL Databases.

Stars: ✭ 151 (-93.14%)

Mutual labels: big-data

Two-day workshop that covers how to use R to interact databases and Spark

Stars: ✭ 110 (-95%)

Mutual labels: big-data

The Accelerator is a tool for fast and reproducible processing of large amounts of data.

Stars: ✭ 137 (-93.77%)

Mutual labels: big-data

Attic Predictionio Sdk Java

PredictionIO Java SDK

Stars: ✭ 107 (-95.14%)

Mutual labels: big-data

Keyvi - the key value index. It is an in-memory FST-based data structure highly optimized for size and lookup performance.

Stars: ✭ 161 (-92.68%)

Mutual labels: big-data

Mysql perf analyzer

MySQL performance monitoring and analysis.

Stars: ✭ 1,423 (-35.32%)

Mutual labels: big-data

Open Source Handbook

⭐️ Open source projects for all skill levels

Stars: ✭ 131 (-94.05%)

Mutual labels: big-data

Explore high-dimensional datasets and how your algo handles specific regions.

Stars: ✭ 100 (-95.45%)

Mutual labels: big-data

Simple windows desktop application for viewing & querying Apache Parquet files

Stars: ✭ 145 (-93.41%)

Mutual labels: big-data

Samza Hello Samza

Mirror of Apache Samza

Stars: ✭ 99 (-95.5%)

Mutual labels: big-data

Mirror of Apache Hama

Stars: ✭ 129 (-94.14%)

Mutual labels: big-data

Bigdata Playground

A complete example of a big data application using : Kubernetes (kops/aws), Apache Spark SQL/Streaming/MLib, Apache Flink, Scala, Python, Apache Kafka, Apache Hbase, Apache Parquet, Apache Avro, Apache Storm, Twitter Api, MongoDB, NodeJS, Angular, GraphQL

Stars: ✭ 177 (-91.95%)

Mutual labels: big-data

Keyvi - a key value index that powers Cliqz search engine. It is an in-memory FST-based data structure highly optimized for size and lookup performance.

Stars: ✭ 171 (-92.23%)

Mutual labels: big-data

Apache Fluo

Stars: ✭ 159 (-92.77%)

Mutual labels: big-data

1-60 of 380 similar projects