大数据
大数据处理、分析与可视化
共 315 个源码项目kareldb
A Relational Database Backed by Apache Kafka
phoebus
Phoebus is a distributed framework for large scale graph processing written in Erlang.
Decider
Flexible and Extensible Machine Learning in Ruby
zoie
realtime search/indexing system
freshlytics
Open source privacy-friendly analytics
Conjecture
Scalable Machine Learning in Scalding
spark-gotchas
Spark Gotchas. A subjective compilation of the Apache Spark tips and tricks
apex-core
Mirror of Apache Apex core
spindle
Next-generation web analytics processing with Scala, Spark, and Parquet.
sparrow
Sparrow scheduling platform (U.C. Berkeley).
mist
Serverless proxy for Spark cluster
incubator-hivemall
Mirror of Apache Hivemall (incubating)
ksql
The database purpose-built for stream processing applications.
packetpig
Packetpig - Open Source Big Data Security Analytics
tigon
High Throughput Real-time Stream Processing Framework
kyoto
Kyoto Tycoon key-value store (and the underlying Kyoto Cabinet library)
squall
A streaming / online query processing / analytics engine based on Apache Storm
HiveRunner
An Open Source unit test framework for Hive queries based on JUnit 4 and 5
streamflow
StreamFlow™ is a stream processing tool designed to help build and monitor processing workflows.
parkour
Hadoop MapReduce in idiomatic Clojure.
scramjet
Public tracker for Scramjet Cloud Platform, a platform that bring data from many environments together.
incubator-samoa
Mirror of Apache Samoa (Incubating)
hadoopy
Python MapReduce library written in Cython. Visit us in #hadoopy on freenode. See the link below for documentation and tutorials.
Flotilla
Automated message queue orchestration for scaled-up benchmarking.
第 9 / 14 页,共 315 个项目
