The easy way to get, use and share data

Neurocommons text mining pilot

About

The complete dataset is composed of a set of smaller datasets. Each download is in one of two formats: (1) WARC or (2) tar.gz. You can read about the WARC format by following this link to the mailing list. The tar.gz format is a tarred and gzipped file containing triples given in the N-Triples syntax.