Uproot3ROOT I/O in pure Python and NumPy.
Stars: ✭ 312 (+1100%)
dxramA distributed in-memory key-value storage for billions of small objects.
Stars: ✭ 25 (-3.85%)
Awkward 0.xManipulate arrays of complex data structures as easily as Numpy.
Stars: ✭ 216 (+730.77%)
Reddit DetectivePlay detective on Reddit: Discover political disinformation campaigns, secret influencers and more
Stars: ✭ 129 (+396.15%)
tsplib95Library for working with TSPLIB files.
Stars: ✭ 48 (+84.62%)
autThe Archives Unleashed Toolkit is an open-source toolkit for analyzing web archives.
Stars: ✭ 111 (+326.92%)
Uproot4ROOT I/O in pure Python and NumPy.
Stars: ✭ 80 (+207.69%)
GimelBig Data Processing Framework - Unified Data API or SQL on Any Storage
Stars: ✭ 216 (+730.77%)
NakedtensorBare bone examples of machine learning in TensorFlow
Stars: ✭ 2,443 (+9296.15%)
causefolioFor the community, by the community.
Stars: ✭ 44 (+69.23%)
seo-audits-toolkitSEO & Security Audit for Websites. Lighthouse & Security Headers crawler, Sitemap/Keywords/Images Extractor, Summarizer, etc ...
Stars: ✭ 311 (+1096.15%)
SparkrdmaRDMA accelerated, high-performance, scalable and efficient ShuffleManager plugin for Apache Spark
Stars: ✭ 215 (+726.92%)
rhinoAgile Sandbox for analyzing Windows, Linux and macOS malware and execution behaviors
Stars: ✭ 49 (+88.46%)
CalciteApache Calcite
Stars: ✭ 2,816 (+10730.77%)
Couchdb DockerSemi-official Apache CouchDB Docker images
Stars: ✭ 194 (+646.15%)
KoalasKoalas: pandas API on Apache Spark
Stars: ✭ 3,044 (+11607.69%)
Presto Go ClientA Presto client for the Go programming language.
Stars: ✭ 183 (+603.85%)
CboardAn easy to use, self-service open BI reporting and BI dashboard platform.
Stars: ✭ 2,795 (+10650%)
Bigdata PlaygroundA complete example of a big data application using : Kubernetes (kops/aws), Apache Spark SQL/Streaming/MLib, Apache Flink, Scala, Python, Apache Kafka, Apache Hbase, Apache Parquet, Apache Avro, Apache Storm, Twitter Api, MongoDB, NodeJS, Angular, GraphQL
Stars: ✭ 177 (+580.77%)
KeyviKeyvi - a key value index that powers Cliqz search engine. It is an in-memory FST-based data structure highly optimized for size and lookup performance.
Stars: ✭ 171 (+557.69%)
Lite Virtual ListVirtual list component library supporting waterfall flow based on vue
Stars: ✭ 223 (+757.69%)
UsqlU-SQL Examples and Issue Tracking
Stars: ✭ 221 (+750%)
shell-historyVisualize your shell usage with Highcharts!
Stars: ✭ 100 (+284.62%)
HyperspaceAn open source indexing subsystem that brings index-based query acceleration to Apache Spark™ and big data workloads.
Stars: ✭ 246 (+846.15%)
GeopysparkGeoTrellis for PySpark
Stars: ✭ 167 (+542.31%)
spinmobRapid and flexible acquisition, analysis, fitting, and plotting in Python. Designed for scientific laboratories.
Stars: ✭ 34 (+30.77%)
HelicalinsightHelical Insight software is world’s first Open Source Business Intelligence framework which helps you to make sense out of your data and make well informed decisions.
Stars: ✭ 214 (+723.08%)
ClickhouseClickHouse® is a free analytics DBMS for big data
Stars: ✭ 21,089 (+81011.54%)
Data Science Live BookAn open source book to learn data science, data analysis and machine learning, suitable for all ages!
Stars: ✭ 193 (+642.31%)
GunAn open source cybersecurity protocol for syncing decentralized graph data.
Stars: ✭ 15,172 (+58253.85%)
Vue Virtual Scroll List⚡️A vue component support big amount data list with high render performance and efficient.
Stars: ✭ 3,201 (+12211.54%)
FlumeMirror of Apache Flume
Stars: ✭ 2,200 (+8361.54%)
phisherpriceAll In One Pentesting Tool For Recon & Auditing , Phone Number Lookup , Header , SSH Scan , SSL/TLS Scan & Much More.
Stars: ✭ 38 (+46.15%)
DvidDistributed, Versioned, Image-oriented Dataservice
Stars: ✭ 174 (+569.23%)
Data AcceleratorData Accelerator for Apache Spark simplifies onboarding to Streaming of Big Data. It offers a rich, easy to use experience to help with creation, editing and management of Spark jobs on Azure HDInsights or Databricks while enabling the full power of the Spark engine.
Stars: ✭ 247 (+850%)
Attic PredictionioPredictionIO, a machine learning server for developers and ML engineers.
Stars: ✭ 12,522 (+48061.54%)
J2NJava-like Components for .NET
Stars: ✭ 37 (+42.31%)
KeyviKeyvi - the key value index. It is an in-memory FST-based data structure highly optimized for size and lookup performance.
Stars: ✭ 161 (+519.23%)
Aws Etl OrchestratorA serverless architecture for orchestrating ETL jobs in arbitrarily-complex workflows using AWS Step Functions and AWS Lambda.
Stars: ✭ 245 (+842.31%)
FluoApache Fluo
Stars: ✭ 159 (+511.54%)
PrestoThe official home of the Presto distributed SQL query engine for big data
Stars: ✭ 12,957 (+49734.62%)
d3-force-graphForce-directed graph using D3-force and WebGL, support massive data rendering and custom style.
Stars: ✭ 74 (+184.62%)
TrafodionApache Trafodion
Stars: ✭ 242 (+830.77%)
GeniA Clojure dataframe library that runs on Spark
Stars: ✭ 152 (+484.62%)
Spark.jlJulia binding for Apache Spark
Stars: ✭ 153 (+488.46%)
Kafka UiOpen-Source Web GUI for Apache Kafka Management
Stars: ✭ 230 (+784.62%)
DatasciencevmTools and Docs on the Azure Data Science Virtual Machine (http://aka.ms/dsvm)
Stars: ✭ 153 (+488.46%)
FiliEasily make RESTful web services for time series reporting with Big Data analytics engines like Druid and SQL Databases.
Stars: ✭ 151 (+480.77%)