There are no reviews yet. Be the first to send feedback to the community and the maintainers!
tika-python
Tika-Python is a Python binding to the Apache Tikaβ’ REST services allowing Tika to be called natively in the Python community.MLwithTensorFlow2ed
Code for Machine Learning with TensorFlow: 2nd Edition Published by Manning Publicationstika-similarity
Tika-Similarity uses the Tika-Python package (Python port of Apache Tika) to compute file similarity based on Metadata features.imagecat
ImageCat is an Apache OODT RADIX application that uses Apache Solr, Apache Tika and Apache OODT to ingest 10s of millions of files (images,but could be extended to other files) in place, and to extract metadata and OCR information from those files/images using Tika and Tesseract OCR.lucene-geo-gazetteer
Uses Apache Lucene, OpenNLP and geonames and extracts locations from text and geocodes them.nutch-python
Nutch-Python is a Python binding to the Apache Nutchβ’ REST services allowing Nutch to be called natively in the Python community. β Editetllib
This is the ETL lib package. It provides an API to munge and prepare JSON, TSV and other data using Apache Tika and JSON parsing/loading for ETL via Apache OODT (or other libs) into Apache Solr.solrcene
Spatial Branch of Apache Solrtrec-dd-polar
A dataset downloaded from the deep and scientific web across three major Polar data centers for use in research.shangridocs
Document exploration tooldrat
The Distributed Release Audit Tool (DRAT) for code analysis and verification.politics-hacking
Scripts to process & analyze web data regarding politics.apachestuff
DCGAN-AnimeFaces
NSFDataVizHackathon-2014
DCGAN-Dog-Generator
disco
Data Intensive Software Connectorsdeeplearning-udacity
Chris's assignments from DeepLearning class on udacity.ctakesparser-utils
grobidparser-resources
bigtranslate
An Apache OODT, Apache Tika, and Apache Solr based system to automatically take large TSV file datasets, and to translate them from one language to another. Built and inspired by the DARPA XDATA Employment dataset.geotopicparser-utils
memex-autonomy
HyspIRI
ace
Automated Concept Extraction from Search Enginesapple
Automatic precondition, convert and publish remote sending data to the ESGF.memex-weapons
earthcube
oodt-pushpull-plugins
smartcontracts
labkey-dumper
maars-search
videocat
Love Open Source and this site? Check out how you can help us