Top Rating
- Top Contributors
  Discover the Top Open Source contributors by country or by language
- Interviews
  Discover real stories from Open Source developers
Discover

Discover your Favorite Language
Discover the top trending repositories and projects on Github. Explore the latest trends in your preferred languages.

Groovy

Objective-C

Julia

Go

MATLAB

Elixir

Perl

Java

More Languages
Awesome

Awesome repositories
Discover the most awesome repositories and projects of your favorite languages. Inspired by the Awesome-* lists trend in GitHub.

Scala

Nix

Go

Julia

JavaScript

PHP

Python

Rust

More Languages
By Country

Rankings by Country
Discover the community of talented open source contributors in each country.

🇳🇺 Niue

🇹🇲 Turkmenistan

🇸🇯 Svalbard and Jan Mayen

🇸🇳 Senegal

🇮🇶 Iraq

🇬🇦 Gabon

🇰🇭 Cambodia

🇷🇴 Romania

All Countries Compare Countries

pbloem/former

Stars
1,037
Rank 44,435 (Top 0.9 %)
Language
Python
License
MIT License
Created over 5 years ago
Updated 6 months ago

pbloem/former

pbloem

There are no reviews yet. Be the first to send feedback to the community and the maintainers!

Simple transformer implementation from scratch in pytorch.

former

Simple transformer implementation from scratch in pytorch. See http://peterbloem.nl/blog/transformers for an in-depth explanation.

Limitations

The models implemented here are designed to show the simplicity of transformer models and self-attention. As such they will not scale as far as the bigger transformers. For that you'll need a number of tricks that complicate the code (see the blog post for details).

All models in the repository consist of a single stack of transformer blocks (that is, no encoder/decoder structures). It turns out that this simple configuration often works best.

Installation and use

First, download or clone the repository. Then, in the directory that contains setup.py, run

pip install -e .

The switch -e ensures that when you edit the code, the installed packaged is also changed. This means that you can, for instance, add print statements to the code to see how it works.

Then, from the same directory, run:

python experiments/classify.py

This will run a simple classification experiment on the IMDb dataset.

Hyperparameters are passed as command line arguments. The defaults should work well. The classification data is automatically downloaded, and the wikipedia data is included in the repository.

Requirements

Python 3.6+ is required. The pip command above should install all required packages. You may also need pip install future depending on the exact python version.

conda environment

The file environment.yml describes a complete conda environment with all dependencies. After cloning or downloading the project, you create the environment as follows:

conda env create -f environment.yml --name former
conda activate former

language-models

Keras implementations of three language models: character-level RNN, word-level RNN and Sentence VAE (Bowman, Vilnis et al 2016).

pca-book

Source files for a book on Principal component analysis

pixel-models

Pytorch implementations of the PixelCNN (va Oord et al. 2016) and PixelVAE (Gulrajani et al. 2016) models

kgbench-data

A set of benchmark repositories for node classification on knowledge graphs. To use, see https://github.com/pbloem/kgbench-loader

blog

Example code to accompany blog posts

Jupyter Notebook

motive

A proof-of-concept library for motif analysis using MDL techniques.

peterbloem.nl

Repository for personal website

Lilian

Machine Learning Toolkit

machine-learning

Worksheets, sample code and homework for the course Machine Learning at the VU University Amsterdam.

Jupyter Notebook

attn-book

Work in progress. Book on attention and embeddings.

Lilian-experimental

Any research work in progress

ifsem

Proof-of-concept implementation of the IFS-EM algorithm.

gated-rgcn

Proof-of-concept gated RGCN.

score

Jupyter Notebook

Publications

Anything scientific I'm writing, whether intended for publication or not.

orca

A JAVA port of the ORCA algorithm by Tomaz Hocevar

cleartxt

A simple website (cleartxt.info) designed to quickly inform website owners about the problem of cleartext passwords.

Misc

Miscellaneous projects

Jupyter Notebook

leitner

Proof of concept flash card system

voynich-experiments

motive-cls

Motif classificastion experiment

oblique

Oblique strategy cards for academia.

dyna

Dynamic knowledge graph embeddings

krraftwerk

embed

Implementation of basic KG embedding methods. Meant both as an example implementation and as a testbed for simple improvements.