HydrographA visual ETL development and debugging tool for big data
Stars: ✭ 144 (-93.45%)
ReefMirror of Apache REEF
Stars: ✭ 92 (-95.82%)
AzuredatalakeSamples and Docs for Azure Data Lake Store and Analytics
Stars: ✭ 128 (-94.18%)
Bitcoin Value Predictor[NOT MAINTAINED] Predicting Bit coin price using Time series analysis and sentiment analysis of tweets on bitcoin
Stars: ✭ 91 (-95.86%)
PrestoThe official home of the Presto distributed SQL query engine for big data
Stars: ✭ 12,957 (+488.95%)
Griffon VmGriffon Data Science Virtual Machine
Stars: ✭ 128 (-94.18%)
PanoptesA Global Scale Network Telemetry Ecosystem
Stars: ✭ 80 (-96.36%)
IotdbApache IoTDB
Stars: ✭ 1,221 (-44.5%)
Mobydq🐳 Tool to automate data quality checks on data pipelines
Stars: ✭ 123 (-94.41%)
Attic PredictionioPredictionIO, a machine learning server for developers and ML engineers.
Stars: ✭ 12,522 (+469.18%)
CookbookThe Data Engineering Cookbook
Stars: ✭ 9,829 (+346.77%)
Report自动化配置报表平台。演示地址http://58.87.112.247/report 账号 visitor密码123456
Stars: ✭ 123 (-94.41%)
BookkeeperApache Bookkeeper
Stars: ✭ 1,178 (-46.45%)
Belajarpython.comOpen Source Indonesian Python Programming Tutorial Site
Stars: ✭ 141 (-93.59%)
SigmfThe Signal Metadata Format Specification
Stars: ✭ 120 (-94.55%)
Countly Sdk CordovaCountly Product Analytics SDK for Cordova, Icenium and Phonegap
Stars: ✭ 69 (-96.86%)
Spark.jlJulia binding for Apache Spark
Stars: ✭ 153 (-93.05%)
DrillApache Drill is a distributed MPP query layer for self describing data
Stars: ✭ 1,619 (-26.41%)
RsparklingRSparkling: Use H2O Sparkling Water from R (Spark + R + Machine Learning)
Stars: ✭ 65 (-97.05%)
Amazon S3 Find And ForgetAmazon S3 Find and Forget is a solution to handle data erasure requests from data lakes stored on Amazon S3, for example, pursuant to the European General Data Protection Regulation (GDPR)
Stars: ✭ 115 (-94.77%)
NabhashAn extremely fast Non-crypto-safe AES Based Hash algorithm for Big Data
Stars: ✭ 62 (-97.18%)
DvidDistributed, Versioned, Image-oriented Dataservice
Stars: ✭ 174 (-92.09%)
Attic LensMirror of Apache Lens
Stars: ✭ 58 (-97.36%)
Just Dashboard📊 📋 Dashboards using YAML or JSON files
Stars: ✭ 1,511 (-31.32%)
PoseidonA search engine which can hold 100 trillion lines of log data.
Stars: ✭ 1,793 (-18.5%)
Lifion KinesisA native Node.js producer and consumer library for Amazon Kinesis Data Streams
Stars: ✭ 54 (-97.55%)
AmbariMirror of Apache Ambari
Stars: ✭ 1,576 (-28.36%)
OodtMirror of Apache OODT
Stars: ✭ 52 (-97.64%)
FiliEasily make RESTful web services for time series reporting with Big Data analytics engines like Druid and SQL Databases.
Stars: ✭ 151 (-93.14%)
TrckQuery engine for TrailDB
Stars: ✭ 48 (-97.82%)
BigdataclassTwo-day workshop that covers how to use R to interact databases and Spark
Stars: ✭ 110 (-95%)
MoosefsMooseFS – Open Source, Petabyte, Fault-Tolerant, Highly Performing, Scalable Network Distributed File System (Software-Defined Storage)
Stars: ✭ 1,025 (-53.41%)
AcceleratorThe Accelerator is a tool for fast and reproducible processing of large amounts of data.
Stars: ✭ 137 (-93.77%)
AttacaRobust, distributed version control for large files.
Stars: ✭ 41 (-98.14%)
KeyviKeyvi - the key value index. It is an in-memory FST-based data structure highly optimized for size and lookup performance.
Stars: ✭ 161 (-92.68%)
MetricsMeasure behavior of Java applications
Stars: ✭ 35 (-98.41%)
SkymapHigh-throughput gene to knowledge mapping through massive integration of public sequencing data.
Stars: ✭ 29 (-98.68%)
Awesome ScalabilityThe Patterns of Scalable, Reliable, and Performant Large-Scale Systems
Stars: ✭ 36,688 (+1567.64%)
VizukaExplore high-dimensional datasets and how your algo handles specific regions.
Stars: ✭ 100 (-95.45%)
K8s Ingress ClaimAn admission control policy that safeguards against accidental duplicate claiming of Hosts/Domains.
Stars: ✭ 14 (-99.36%)
ParquetviewerSimple windows desktop application for viewing & querying Apache Parquet files
Stars: ✭ 145 (-93.41%)
KuduMirror of Apache Kudu
Stars: ✭ 1,360 (-38.18%)
HamaMirror of Apache Hama
Stars: ✭ 129 (-94.14%)
Bigdata PlaygroundA complete example of a big data application using : Kubernetes (kops/aws), Apache Spark SQL/Streaming/MLib, Apache Flink, Scala, Python, Apache Kafka, Apache Hbase, Apache Parquet, Apache Avro, Apache Storm, Twitter Api, MongoDB, NodeJS, Angular, GraphQL
Stars: ✭ 177 (-91.95%)
KeyviKeyvi - a key value index that powers Cliqz search engine. It is an in-memory FST-based data structure highly optimized for size and lookup performance.
Stars: ✭ 171 (-92.23%)
FluoApache Fluo
Stars: ✭ 159 (-92.77%)
100daysofmlcodeMy journey to learn and grow in the domain of Machine Learning and Artificial Intelligence by performing the #100DaysofMLCode Challenge.
Stars: ✭ 146 (-93.36%)
GafferA large-scale entity and relation database supporting aggregation of properties
Stars: ✭ 1,642 (-25.36%)