A complete example of a big data application using : Kubernetes (kops/aws), Apache Spark SQL/Streaming/MLib, Apache Flink, Scala, Python, Apache Kafka, Apache Hbase, Apache Parquet, Apache Avro, Apache Storm, Twitter Api, MongoDB, NodeJS, Angular, GraphQL

Stars: ✭ 177 (+342.5%)

Mutual labels: big-data

predictionio-sdk-ruby

PredictionIO Ruby SDK

Stars: ✭ 192 (+380%)

Mutual labels: big-data

Keyvi

Keyvi - a key value index that powers Cliqz search engine. It is an in-memory FST-based data structure highly optimized for size and lookup performance.

Stars: ✭ 171 (+327.5%)

Mutual labels: big-data

Selinon

An advanced distributed task flow management on top of Celery

Stars: ✭ 237 (+492.5%)

Mutual labels: big-data

Geopyspark

GeoTrellis for PySpark

Stars: ✭ 167 (+317.5%)

Mutual labels: big-data

CytoPy

A data-centric flow/mass cytometry automated analysis framework

Stars: ✭ 27 (-32.5%)

Mutual labels: document-database

Fluo

Apache Fluo

Stars: ✭ 159 (+297.5%)

Mutual labels: big-data

Books

整理一些书籍 ,包含 C&C++ 、git 、Java、Keras 、Linux 、NLP 、Python 、Scala 、TensorFlow 、大数据、推荐系统、数据库、数据挖掘、机器学习、深度学习、算法等。

Stars: ✭ 222 (+455%)

Mutual labels: big-data

Geni

A Clojure dataframe library that runs on Spark

Stars: ✭ 152 (+280%)

Mutual labels: big-data

NoSQLDataEngineering

NoSQL Data Engineering

Stars: ✭ 25 (-37.5%)

Mutual labels: document-database

Datasciencevm

Tools and Docs on the Azure Data Science Virtual Machine (http://aka.ms/dsvm)

Stars: ✭ 153 (+282.5%)

Mutual labels: big-data

Nakedtensor

Bare bone examples of machine learning in TensorFlow

Stars: ✭ 2,443 (+6007.5%)

Mutual labels: big-data

Spark With Python

Fundamentals of Spark with Python (using PySpark), code examples

Stars: ✭ 150 (+275%)

Mutual labels: big-data

Koalas

Koalas: pandas API on Apache Spark

Stars: ✭ 3,044 (+7510%)

Mutual labels: big-data

100daysofmlcode

My journey to learn and grow in the domain of Machine Learning and Artificial Intelligence by performing the #100DaysofMLCode Challenge.

Stars: ✭ 146 (+265%)

Mutual labels: big-data

Gimel

Big Data Processing Framework - Unified Data API or SQL on Any Storage

Stars: ✭ 216 (+440%)

Mutual labels: big-data

Metamodel

Mirror of Apache Metamodel

Stars: ✭ 143 (+257.5%)

Mutual labels: big-data

acousticbrainz-server

The server components for the AcousticBrainz project

Stars: ✭ 128 (+220%)

Mutual labels: big-data

Sparkrdma

RDMA accelerated, high-performance, scalable and efficient ShuffleManager plugin for Apache Spark

Stars: ✭ 215 (+437.5%)

Mutual labels: big-data

Big Data Study

🐳 big data study

Stars: ✭ 141 (+252.5%)

Mutual labels: big-data

Belajarpython.com

Open Source Indonesian Python Programming Tutorial Site

Stars: ✭ 141 (+252.5%)

Mutual labels: big-data

Eel Sdk

Big Data Toolkit for the JVM

Stars: ✭ 140 (+250%)

Mutual labels: big-data

Cboard

An easy to use, self-service open BI reporting and BI dashboard platform.

Stars: ✭ 2,795 (+6887.5%)

Mutual labels: big-data

Helicalinsight

Helical Insight software is world’s first Open Source Business Intelligence framework which helps you to make sense out of your data and make well informed decisions.

Stars: ✭ 214 (+435%)

Mutual labels: big-data

Hazelcast Go Client

Hazelcast IMDG Go Client

Stars: ✭ 140 (+250%)

Mutual labels: big-data

Sparkling Graph

SparklingGraph provides easy to use set of features that will give you ability to proces large scala graphs using Spark and GraphX.

Stars: ✭ 139 (+247.5%)

Mutual labels: big-data

Calcite

Apache Calcite