AMPLab at UC Berkeley (@amplab)

Top repositories

1

shark

Development in Shark has been ended.
Scala
994
star
2

SparkNet

Distributed Neural Networks for Spark
Scala
605
star
3

keystone

Simplifying robust end-to-end machine learning on Apache Spark.
Scala
468
star
4

spark-ec2

Scripts used to setup a Spark cluster on EC2
Python
390
star
5

graphx

Former GraphX development repository. GraphX has been merged into Apache Spark; please submit pull requests there.
Scala
353
star
6

snap

Scalable Nucleotide Alignment Program -- a fast and accurate read aligner for high-throughput sequencing data
C++
279
star
7

succinct

Enabling queries on compressed data.
Java
278
star
8

docker-scripts

Dockerfiles and scripts for Spark and Shark Docker images
Shell
259
star
9

spark-indexedrdd

An efficient updatable key-value store for Apache Spark
Scala
249
star
10

datascience-sp14

Repository for data science course Spring 14
Shell
182
star
11

MLI

An API for Distributed Machine Learning
Scala
154
star
12

training

Training materials for Strata, AMP Camp, etc
Scala
150
star
13

drizzle-spark

Drizzle integration with Apache Spark
Scala
120
star
14

carat

Carat: Collaborative Energy Debugging
Java
114
star
15

velox-modelserver

Scala
110
star
16

benchmark

Large scale query engine benchmark
Python
99
star
17

ml-matrix

Distributed Matrix Library
Scala
70
star
18

ampcrowd

A RESTful web service that runs microtasks across multiple crowds, provides quality control techniques, and is easily extensible.
Python
51
star
19

smash

Benchmarking toolkit for variant calling
Python
46
star
20

training-scripts

Scripts to launch cluster used for Strata
Python
33
star
21

ernest

Code for Ernest
Python
32
star
22

cyclades

Cyclades
C++
28
star
23

succinct-cpp

Succinct C++
C++
24
star
24

ampcamp

scripts used for ampcamp
Python
16
star
25

zipg

A Memory-efficient Graph Store for Interactive Queries
Java
12
star
26

orchestra

Fine-Grained Distributed Computing
Python
11
star
27

iolap

Scala
11
star
28

cs262a-fall2016

HTML
9
star
29

sprint

Sprint Transformations for RegEx queries
C++
8
star
30

keystone-example

A example skeleton for an application built on top of KeystoneML
Shell
8
star
31

mlsys

An open source survey of the emerging systems for large-scale machine learning.
CSS
4
star
32

clipper-v0

Rust
3
star
33

ray-core

Experiments for the Ray backend
C++
3
star
34

Buggypedia

Objective-C
3
star
35

sparse-covariance-inverse

1
star
36

siren-release

Public version of the SiRen project
Scala
1
star
37

keystone-integration-tests

Integration Tests for KeystoneML
Shell
1
star
38

ampcrowd-client-py

A python client for using the AMPCrowd service
Python
1
star