bookmark

Apache Tika - Apache Tika


Description

Apache Tika is a toolkit for detecting and extracting metadata and structured text content from various documents using existing parser libraries. For more information about Tika, please see the list of supported document formats and the available documentation . You can find the latest release on the download page . See the Getting Started guide for instructions on how to start using Tika.

Tika is a subproject of Apache Lucene . Lucene is a project of the Apache Software Foundation .

Preview

Tags

Users

  • @pitman

Comments and Reviews