% This file is GENERATED. Check the tools on https://nocomplexity.com/ or my github.com/nocomplexity to use it too! Its FOSS.
:width: 200px
:align: center
The Apache Tika™ toolkit detects and extracts metadata and text from over a thousand different file types (such as PPT, XLS, and PDF). All of these file types can be parsed through a single interface, making Tika useful for search engine indexing, content analysis, translation, and much more.
Home page for this solution: https://tika.apache.org/
| Key | Value |
|---|---|
| Name | tika |
| Description | The Apache Tika toolkit detects and extracts metadata and text from over a thousand different file types (such as PPT, XLS, and PDF). |
| License | Apache License 2.0 |
| Programming Language | Java |
| Created | 2009-05-21 |
| Last update | 2025-05-01 |
| Github Stars | 2948 |
| Project Home Page | https://tika.apache.org/ |
| Code Repository | https://github.com/apache/tika |
| OpenSSF Scorecard | Report |
Note:
-
Created date is date that repro is created on Github.com.
-
Last update: Last update of repository on Github found on {sub-ref}
today. -
Do not attach a wrong value to github stars. Its a vanity metric! Stars count are misleading and don't indicate if the SBB is high-quality or very popular.
% This file is GENERATED. Check the tools on https://nocomplexity.com/ or my github.com/nocomplexity to use it too! Its FOSS.