.
FRDCSA | git codebases | grobid
Homepage

[Project image]
grobid

Jump to: Project Description

Project Description

grobid is a machine learning library for extracting, parsing and re-structuring raw documents such as PDF into structured TEI-encoded documents with a particular focus on technical and scientific publications. First developments started in 2008 as a hobby. In 2011 the tool has been made available in open source. Work on grobid has been steady as side project since the beginning and is expected to continue until at least 2020 :)


This page is part of the FWeb package.
Last updated Sat Oct 26 17:00:20 EDT 2019 .