Helical Insight software is world’s first Open Source Business Intelligence framework which helps you to make sense out of your data and make well informed decisions.

Stars: ✭ 214 (+167.5%)

Mutual labels: big-data

Fili

Easily make RESTful web services for time series reporting with Big Data analytics engines like Druid and SQL Databases.

Stars: ✭ 151 (+88.75%)

Mutual labels: big-data

Parquetviewer

Simple windows desktop application for viewing & querying Apache Parquet files

Stars: ✭ 145 (+81.25%)

Mutual labels: big-data

Attic Predictionio Sdk Python

PredictionIO Python SDK

Stars: ✭ 196 (+145%)

Mutual labels: big-data

Hydrograph

A visual ETL development and debugging tool for big data

Stars: ✭ 144 (+80%)

Mutual labels: big-data

Storm Doc Zh

Apache Storm 官方文档中文版

Stars: ✭ 142 (+77.5%)

Mutual labels: big-data

Data Accelerator

Data Accelerator for Apache Spark simplifies onboarding to Streaming of Big Data. It offers a rich, easy to use experience to help with creation, editing and management of Spark jobs on Azure HDInsights or Databricks while enabling the full power of the Spark engine.

Stars: ✭ 247 (+208.75%)

Mutual labels: big-data

Eland

Python Client and Toolkit for DataFrames, Big Data, Machine Learning and ETL in Elasticsearch

Stars: ✭ 235 (+193.75%)

Mutual labels: big-data

Data Science Live Book

An open source book to learn data science, data analysis and machine learning, suitable for all ages!

Stars: ✭ 193 (+141.25%)

Mutual labels: big-data

Belajarpython.com

Open Source Indonesian Python Programming Tutorial Site

Stars: ✭ 141 (+76.25%)

Mutual labels: big-data

Hazelcast Go Client

Hazelcast IMDG Go Client

Stars: ✭ 140 (+75%)

Mutual labels: big-data

Geopyspark

GeoTrellis for PySpark

Stars: ✭ 167 (+108.75%)

Mutual labels: big-data

Gimel

Big Data Processing Framework - Unified Data API or SQL on Any Storage

Stars: ✭ 216 (+170%)

Mutual labels: big-data

Fluo

Apache Fluo

Stars: ✭ 159 (+98.75%)

Mutual labels: big-data

Trafodion

Apache Trafodion

Stars: ✭ 242 (+202.5%)

Mutual labels: big-data

Geni

A Clojure dataframe library that runs on Spark

Stars: ✭ 152 (+90%)

Mutual labels: big-data

Sparkrdma

RDMA accelerated, high-performance, scalable and efficient ShuffleManager plugin for Apache Spark

Stars: ✭ 215 (+168.75%)

Mutual labels: big-data

Datasciencevm

Tools and Docs on the Azure Data Science Virtual Machine (http://aka.ms/dsvm)

Stars: ✭ 153 (+91.25%)

Mutual labels: big-data

Cboard

An easy to use, self-service open BI reporting and BI dashboard platform.

Stars: ✭ 2,795 (+3393.75%)

Mutual labels: big-data

Spark With Python

Fundamentals of Spark with Python (using PySpark), code examples

Stars: ✭ 150 (+87.5%)

Mutual labels: big-data

Calcite

Apache Calcite

Stars: ✭ 2,816 (+3420%)

Mutual labels: big-data

100daysofmlcode

My journey to learn and grow in the domain of Machine Learning and Artificial Intelligence by performing the #100DaysofMLCode Challenge.

Stars: ✭ 146 (+82.5%)

Mutual labels: big-data

Selinon

An advanced distributed task flow management on top of Celery

Stars: ✭ 237 (+196.25%)

Mutual labels: big-data

Metamodel

Mirror of Apache Metamodel

Stars: ✭ 143 (+78.75%)

Mutual labels: big-data

Couchdb Docker

Semi-official Apache CouchDB Docker images

Stars: ✭ 194 (+142.5%)

Mutual labels: big-data

Big Data Study

🐳 big data study

Stars: ✭ 141 (+76.25%)

Mutual labels: big-data

Koalas

Koalas: pandas API on Apache Spark

Stars: ✭ 3,044 (+3705%)

Mutual labels: big-data

Eel Sdk

Big Data Toolkit for the JVM

Stars: ✭ 140 (+75%)

Mutual labels: big-data

Attic Predictionio Sdk Ruby

PredictionIO Ruby SDK

Stars: ✭ 192 (+140%)

Mutual labels: big-data

Books

整理一些书籍 ,包含 C&C++ 、git 、Java、Keras 、Linux 、NLP 、Python 、Scala 、TensorFlow 、大数据、推荐系统、数据库、数据挖掘、机器学习、深度学习、算法等。

Stars: ✭ 222 (+177.5%)

Mutual labels: big-data

Sparkling Graph

SparklingGraph provides easy to use set of features that will give you ability to proces large scala graphs using Spark and GraphX.

Stars: ✭ 139 (+73.75%)

Mutual labels: big-data

Gun

An open source cybersecurity protocol for syncing decentralized graph data.

Stars: ✭ 15,172 (+18865%)

Mutual labels: big-data

Poseidon

A search engine which can hold 100 trillion lines of log data.

Stars: ✭ 1,793 (+2141.25%)

Mutual labels: big-data

Spark On Lambda

Apache Spark on AWS Lambda

Stars: ✭ 137 (+71.25%)

Mutual labels: big-data

Presto Go Client

A Presto client for the Go programming language.

Stars: ✭ 183 (+128.75%)

Mutual labels: big-data

Accelerator

The Accelerator is a tool for fast and reproducible processing of large amounts of data.

Stars: ✭ 137 (+71.25%)

Mutual labels: big-data

Attic Apex Malhar

Mirror of Apache Apex malhar

Stars: ✭ 131 (+63.75%)

Mutual labels: big-data

Hyperspace

An open source indexing subsystem that brings index-based query acceleration to Apache Spark™ and big data workloads.

Stars: ✭ 246 (+207.5%)

Mutual labels: big-data

Lite Virtual List

Virtual list component library supporting waterfall flow based on vue

Stars: ✭ 223 (+178.75%)

Mutual labels: big-data

Flume

Mirror of Apache Flume

Stars: ✭ 2,200 (+2650%)

Mutual labels: big-data

Open Source Handbook

⭐️ Open source projects for all skill levels

Stars: ✭ 131 (+63.75%)

Mutual labels: big-data

Calcite Avatica

Mirror of Apache Calcite - Avatica

Stars: ✭ 130 (+62.5%)

Mutual labels: big-data

Bigdata Playground

A complete example of a big data application using : Kubernetes (kops/aws), Apache Spark SQL/Streaming/MLib, Apache Flink, Scala, Python, Apache Kafka, Apache Hbase, Apache Parquet, Apache Avro, Apache Storm, Twitter Api, MongoDB, NodeJS, Angular, GraphQL

Stars: ✭ 177 (+121.25%)

Mutual labels: big-data

Hama

Mirror of Apache Hama

Stars: ✭ 129 (+61.25%)

Mutual labels: big-data

Gaffer

A large-scale entity and relation database supporting aggregation of properties

Stars: ✭ 1,642 (+1952.5%)

Mutual labels: big-data

1-60 of 369 similar projects

›

next*5