All Projects → masc → Similar Projects or Alternatives

596 Open source projects that are alternatives of or similar to masc

couchdb-pkg
Apache CouchDB Packaging support files
Stars: ✭ 24 (+20%)
Mutual labels:  big-data, apache
accumulo-testing
Apache Accumulo Testing
Stars: ✭ 14 (-30%)
Mutual labels:  big-data, accumulo
Hive
Apache Hive
Stars: ✭ 4,031 (+20055%)
Mutual labels:  big-data, apache
Couchdb Docker
Semi-official Apache CouchDB Docker images
Stars: ✭ 194 (+870%)
Mutual labels:  big-data, apache
accumulo-docker
Apache Accumulo Docker
Stars: ✭ 17 (-15%)
Mutual labels:  big-data, accumulo
Spark With Python
Fundamentals of Spark with Python (using PySpark), code examples
Stars: ✭ 150 (+650%)
Mutual labels:  big-data, apache
hadoop-data-ingestion-tool
OLAP and ETL of Big Data
Stars: ✭ 17 (-15%)
Mutual labels:  big-data, apache
nifi
Deploy a secured, clustered, auto-scaling NiFi service in AWS.
Stars: ✭ 37 (+85%)
Mutual labels:  big-data, apache
Tez
Apache Tez
Stars: ✭ 313 (+1465%)
Mutual labels:  big-data, apache
Gaffer
A large-scale entity and relation database supporting aggregation of properties
Stars: ✭ 1,642 (+8110%)
Mutual labels:  big-data, accumulo
Bigdata Playground
A complete example of a big data application using : Kubernetes (kops/aws), Apache Spark SQL/Streaming/MLib, Apache Flink, Scala, Python, Apache Kafka, Apache Hbase, Apache Parquet, Apache Avro, Apache Storm, Twitter Api, MongoDB, NodeJS, Angular, GraphQL
Stars: ✭ 177 (+785%)
Mutual labels:  big-data
Presto Go Client
A Presto client for the Go programming language.
Stars: ✭ 183 (+815%)
Mutual labels:  big-data
Selinon
An advanced distributed task flow management on top of Celery
Stars: ✭ 237 (+1085%)
Mutual labels:  big-data
Koalas
Koalas: pandas API on Apache Spark
Stars: ✭ 3,044 (+15120%)
Mutual labels:  big-data
Keyvi
Keyvi - a key value index that powers Cliqz search engine. It is an in-memory FST-based data structure highly optimized for size and lookup performance.
Stars: ✭ 171 (+755%)
Mutual labels:  big-data
Books
整理一些书籍 ,包含 C&C++ 、git 、Java、Keras 、Linux 、NLP 、Python 、Scala 、TensorFlow 、大数据 、推荐系统、数据库、数据挖掘 、机器学习 、深度学习 、算法等。
Stars: ✭ 222 (+1010%)
Mutual labels:  big-data
Geopyspark
GeoTrellis for PySpark
Stars: ✭ 167 (+735%)
Mutual labels:  big-data
Fluo
Apache Fluo
Stars: ✭ 159 (+695%)
Mutual labels:  big-data
Nakedtensor
Bare bone examples of machine learning in TensorFlow
Stars: ✭ 2,443 (+12115%)
Mutual labels:  big-data
Geni
A Clojure dataframe library that runs on Spark
Stars: ✭ 152 (+660%)
Mutual labels:  big-data
Datasciencevm
Tools and Docs on the Azure Data Science Virtual Machine (http://aka.ms/dsvm)
Stars: ✭ 153 (+665%)
Mutual labels:  big-data
dotenvy
Speed up your production sites by ditching .env for key/value variable pairs as Apache, Nginx, and shell equivalents
Stars: ✭ 31 (+55%)
Mutual labels:  apache
Cboard
An easy to use, self-service open BI reporting and BI dashboard platform.
Stars: ✭ 2,795 (+13875%)
Mutual labels:  big-data
Gimel
Big Data Processing Framework - Unified Data API or SQL on Any Storage
Stars: ✭ 216 (+980%)
Mutual labels:  big-data
Fili
Easily make RESTful web services for time series reporting with Big Data analytics engines like Druid and SQL Databases.
Stars: ✭ 151 (+655%)
Mutual labels:  big-data
100daysofmlcode
My journey to learn and grow in the domain of Machine Learning and Artificial Intelligence by performing the #100DaysofMLCode Challenge.
Stars: ✭ 146 (+630%)
Mutual labels:  big-data
Gun
An open source cybersecurity protocol for syncing decentralized graph data.
Stars: ✭ 15,172 (+75760%)
Mutual labels:  big-data
Kafka Ui
Open-Source Web GUI for Apache Kafka Management
Stars: ✭ 230 (+1050%)
Mutual labels:  big-data
Flume
Mirror of Apache Flume
Stars: ✭ 2,200 (+10900%)
Mutual labels:  big-data
Clickhouse
ClickHouse® is a free analytics DBMS for big data
Stars: ✭ 21,089 (+105345%)
Mutual labels:  big-data
Dvid
Distributed, Versioned, Image-oriented Dataservice
Stars: ✭ 174 (+770%)
Mutual labels:  big-data
Eland
Python Client and Toolkit for DataFrames, Big Data, Machine Learning and ETL in Elasticsearch
Stars: ✭ 235 (+1075%)
Mutual labels:  big-data
Attic Predictionio
PredictionIO, a machine learning server for developers and ML engineers.
Stars: ✭ 12,522 (+62510%)
Mutual labels:  big-data
predictionio-template-recommender
PredictionIO Recommendation Engine Template (Scala-based parallelized engine)
Stars: ✭ 80 (+300%)
Mutual labels:  big-data
Keyvi
Keyvi - the key value index. It is an in-memory FST-based data structure highly optimized for size and lookup performance.
Stars: ✭ 161 (+705%)
Mutual labels:  big-data
Lite Virtual List
Virtual list component library supporting waterfall flow based on vue
Stars: ✭ 223 (+1015%)
Mutual labels:  big-data
Presto
The official home of the Presto distributed SQL query engine for big data
Stars: ✭ 12,957 (+64685%)
Mutual labels:  big-data
Vue Virtual Scroll List
⚡️A vue component support big amount data list with high render performance and efficient.
Stars: ✭ 3,201 (+15905%)
Mutual labels:  big-data
Spark.jl
Julia binding for Apache Spark
Stars: ✭ 153 (+665%)
Mutual labels:  big-data
Usql
U-SQL Examples and Issue Tracking
Stars: ✭ 221 (+1005%)
Mutual labels:  big-data
Detecting-Malicious-URL-Machine-Learning
No description or website provided.
Stars: ✭ 47 (+135%)
Mutual labels:  big-data
Sparkrdma
RDMA accelerated, high-performance, scalable and efficient ShuffleManager plugin for Apache Spark
Stars: ✭ 215 (+975%)
Mutual labels:  big-data
Metamodel
Mirror of Apache Metamodel
Stars: ✭ 143 (+615%)
Mutual labels:  big-data
Parquetviewer
Simple windows desktop application for viewing & querying Apache Parquet files
Stars: ✭ 145 (+625%)
Mutual labels:  big-data
Awkward 0.x
Manipulate arrays of complex data structures as easily as Numpy.
Stars: ✭ 216 (+980%)
Mutual labels:  big-data
Hydrograph
A visual ETL development and debugging tool for big data
Stars: ✭ 144 (+620%)
Mutual labels:  big-data
Data Accelerator
Data Accelerator for Apache Spark simplifies onboarding to Streaming of Big Data. It offers a rich, easy to use experience to help with creation, editing and management of Spark jobs on Azure HDInsights or Databricks while enabling the full power of the Spark engine.
Stars: ✭ 247 (+1135%)
Mutual labels:  big-data
Storm Doc Zh
Apache Storm 官方文档中文版
Stars: ✭ 142 (+610%)
Mutual labels:  big-data
Helicalinsight
Helical Insight software is world’s first Open Source Business Intelligence framework which helps you to make sense out of your data and make well informed decisions.
Stars: ✭ 214 (+970%)
Mutual labels:  big-data
Big Data Study
🐳 big data study
Stars: ✭ 141 (+605%)
Mutual labels:  big-data
Belajarpython.com
Open Source Indonesian Python Programming Tutorial Site
Stars: ✭ 141 (+605%)
Mutual labels:  big-data
openwhisk-package-kafka
Apache OpenWhisk package for communicating with Kafka or Message Hub
Stars: ✭ 35 (+75%)
Mutual labels:  apache
Hyperspace
An open source indexing subsystem that brings index-based query acceleration to Apache Spark™ and big data workloads.
Stars: ✭ 246 (+1130%)
Mutual labels:  big-data
Calcite
Apache Calcite
Stars: ✭ 2,816 (+13980%)
Mutual labels:  big-data
Eel Sdk
Big Data Toolkit for the JVM
Stars: ✭ 140 (+600%)
Mutual labels:  big-data
Hazelcast Go Client
Hazelcast IMDG Go Client
Stars: ✭ 140 (+600%)
Mutual labels:  big-data
Attic Predictionio Sdk Python
PredictionIO Python SDK
Stars: ✭ 196 (+880%)
Mutual labels:  big-data
Sparkling Graph
SparklingGraph provides easy to use set of features that will give you ability to proces large scala graphs using Spark and GraphX.
Stars: ✭ 139 (+595%)
Mutual labels:  big-data
Aws Etl Orchestrator
A serverless architecture for orchestrating ETL jobs in arbitrarily-complex workflows using AWS Step Functions and AWS Lambda.
Stars: ✭ 245 (+1125%)
Mutual labels:  big-data
Poseidon
A search engine which can hold 100 trillion lines of log data.
Stars: ✭ 1,793 (+8865%)
Mutual labels:  big-data
1-60 of 596 similar projects