Data Accelerator for Apache Spark simplifies onboarding to Streaming of Big Data. It offers a rich, easy to use experience to help with creation, editing and management of Spark jobs on Azure HDInsights or Databricks while enabling the full power of the Spark engine.

Stars: ✭ 247 (+586.11%)

Mutual labels: big-data

Gun

An open source cybersecurity protocol for syncing decentralized graph data.

Stars: ✭ 15,172 (+42044.44%)

Mutual labels: big-data

masc

Microsoft's contributions for Spark with Apache Accumulo

Stars: ✭ 20 (-44.44%)

Mutual labels: big-data

Dvid

Distributed, Versioned, Image-oriented Dataservice

Stars: ✭ 174 (+383.33%)

Mutual labels: big-data

Kafka Ui

Open-Source Web GUI for Apache Kafka Management

Stars: ✭ 230 (+538.89%)

Mutual labels: big-data

Keyvi

Keyvi - the key value index. It is an in-memory FST-based data structure highly optimized for size and lookup performance.

Stars: ✭ 161 (+347.22%)

Mutual labels: big-data

bagri

XML/Document DB on top of distributed cache

Stars: ✭ 40 (+11.11%)

Mutual labels: big-data

Spark.jl

Julia binding for Apache Spark

Stars: ✭ 153 (+325%)

Mutual labels: big-data

Lite Virtual List

Virtual list component library supporting waterfall flow based on vue

Stars: ✭ 223 (+519.44%)

Mutual labels: big-data

Parquetviewer

Simple windows desktop application for viewing & querying Apache Parquet files

Stars: ✭ 145 (+302.78%)

Mutual labels: big-data

Clickhouse

ClickHouse® is a free analytics DBMS for big data

Stars: ✭ 21,089 (+58480.56%)

Mutual labels: big-data

Storm Doc Zh

Apache Storm 官方文档中文版

Stars: ✭ 142 (+294.44%)

Mutual labels: big-data

Awkward 0.x

Manipulate arrays of complex data structures as easily as Numpy.

Stars: ✭ 216 (+500%)

Mutual labels: big-data

Calcite

Apache Calcite

Stars: ✭ 2,816 (+7722.22%)

Mutual labels: big-data

Eel Sdk

Big Data Toolkit for the JVM

Stars: ✭ 140 (+288.89%)

Mutual labels: big-data

Cboard

An easy to use, self-service open BI reporting and BI dashboard platform.

Stars: ✭ 2,795 (+7663.89%)

Mutual labels: big-data

Couchdb Docker

Semi-official Apache CouchDB Docker images

Stars: ✭ 194 (+438.89%)

Mutual labels: big-data

predictionio-sdk-ruby

PredictionIO Ruby SDK

Stars: ✭ 192 (+433.33%)

Mutual labels: big-data

Attic Predictionio Sdk Ruby

PredictionIO Ruby SDK

Stars: ✭ 192 (+433.33%)

Mutual labels: big-data

Hyperspace

An open source indexing subsystem that brings index-based query acceleration to Apache Spark™ and big data workloads.

Stars: ✭ 246 (+583.33%)

Mutual labels: big-data

Presto Go Client

A Presto client for the Go programming language.

Stars: ✭ 183 (+408.33%)

Mutual labels: big-data

sketches

HyperLogLog and other probabilistic data structures for mining in data streams

Stars: ✭ 15 (-58.33%)

Mutual labels: sketches

Bigdata Playground

A complete example of a big data application using : Kubernetes (kops/aws), Apache Spark SQL/Streaming/MLib, Apache Flink, Scala, Python, Apache Kafka, Apache Hbase, Apache Parquet, Apache Avro, Apache Storm, Twitter Api, MongoDB, NodeJS, Angular, GraphQL

Stars: ✭ 177 (+391.67%)

Mutual labels: big-data

Trafodion

Apache Trafodion

Stars: ✭ 242 (+572.22%)

Mutual labels: big-data

Keyvi

Keyvi - a key value index that powers Cliqz search engine. It is an in-memory FST-based data structure highly optimized for size and lookup performance.

Stars: ✭ 171 (+375%)

Mutual labels: big-data

acousticbrainz-server

The server components for the AcousticBrainz project

Stars: ✭ 128 (+255.56%)

Mutual labels: big-data

Geopyspark

GeoTrellis for PySpark

Stars: ✭ 167 (+363.89%)

Mutual labels: big-data

Selinon

An advanced distributed task flow management on top of Celery

Stars: ✭ 237 (+558.33%)

Mutual labels: big-data

Fluo

Apache Fluo

Stars: ✭ 159 (+341.67%)

Mutual labels: big-data

incubator-tez

Mirror of Apache Tez (Incubating)

Stars: ✭ 60 (+66.67%)

Mutual labels: big-data

Geni

A Clojure dataframe library that runs on Spark

Stars: ✭ 152 (+322.22%)

Mutual labels: big-data

Books

整理一些书籍 ,包含 C&C++ 、git 、Java、Keras 、Linux 、NLP 、Python 、Scala 、TensorFlow 、大数据、推荐系统、数据库、数据挖掘、机器学习、深度学习、算法等。

Stars: ✭ 222 (+516.67%)

Mutual labels: big-data

Datasciencevm

Tools and Docs on the Azure Data Science Virtual Machine (http://aka.ms/dsvm)

Stars: ✭ 153 (+325%)

Mutual labels: big-data

predictionio-template-recommender

PredictionIO Recommendation Engine Template (Scala-based parallelized engine)

Stars: ✭ 80 (+122.22%)

Mutual labels: big-data

Spark With Python

Fundamentals of Spark with Python (using PySpark), code examples

Stars: ✭ 150 (+316.67%)

Mutual labels: big-data

Nakedtensor

Bare bone examples of machine learning in TensorFlow

Stars: ✭ 2,443 (+6686.11%)

Mutual labels: big-data

100daysofmlcode

My journey to learn and grow in the domain of Machine Learning and Artificial Intelligence by performing the #100DaysofMLCode Challenge.

Stars: ✭ 146 (+305.56%)

Mutual labels: big-data

Social-Network-Analysis-in-Python

Social Network Facebook Analysis (Python, Networkx)

Stars: ✭ 26 (-27.78%)

Mutual labels: big-data

Metamodel

Mirror of Apache Metamodel

Stars: ✭ 143 (+297.22%)

Mutual labels: big-data

Gimel

Big Data Processing Framework - Unified Data API or SQL on Any Storage

Stars: ✭ 216 (+500%)

Mutual labels: big-data

Big Data Study

🐳 big data study

Stars: ✭ 141 (+291.67%)

Mutual labels: big-data

Koalas

Koalas: pandas API on Apache Spark

Stars: ✭ 3,044 (+8355.56%)

Mutual labels: big-data

Sparkrdma

RDMA accelerated, high-performance, scalable and efficient ShuffleManager plugin for Apache Spark

Stars: ✭ 215 (+497.22%)

Mutual labels: big-data

accumulo-testing

Apache Accumulo Testing

Stars: ✭ 14 (-61.11%)

Mutual labels: big-data

STAWM

Code for the paper 'A Biologically Inspired Visual Working Memory for Deep Networks'

Stars: ✭ 21 (-41.67%)

Mutual labels: sketches

TT Tech Space

TT Tech Research Notes

Stars: ✭ 21 (-41.67%)

Mutual labels: big-data

Vue Virtual Scroll List

⚡️A vue component support big amount data list with high render performance and efficient.

Stars: ✭ 3,201 (+8791.67%)

Mutual labels: big-data

Helicalinsight

Helical Insight software is world’s first Open Source Business Intelligence framework which helps you to make sense out of your data and make well informed decisions.

Stars: ✭ 214 (+494.44%)

Mutual labels: big-data

1-60 of 383 similar projects

›

next*5