jackrabbit-oak-issues mailing list archives

Site index · List index
Message view « Date » · « Thread »
Top « Date » · « Thread »
From "Chetan Mehrotra (JIRA)" <j...@apache.org>
Subject [jira] [Commented] (OAK-5048) Upgrade to latest Tika version
Date Mon, 03 Jul 2017 09:18:01 GMT

    [ https://issues.apache.org/jira/browse/OAK-5048?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=16072160#comment-16072160

Chetan Mehrotra commented on OAK-5048:

Following are the sizes post change

|oak-run|44M|44M| Embeds tika-core and tika-parsers|
|oak-lucene|5.5 M|5.5 M| No embed|
|oak-solr-core|155K|155K| No embed|
|oak-examples/standalone|72M|99M| Embeds whole Tika stuff|
|oak-examples/webapp|53M|78M| Embeds whole Tika stuff|

> Upgrade to latest Tika version
> ------------------------------
>                 Key: OAK-5048
>                 URL: https://issues.apache.org/jira/browse/OAK-5048
>             Project: Jackrabbit Oak
>          Issue Type: Improvement
>          Components: lucene
>            Reporter: Tommaso Teofili
>            Assignee: Chetan Mehrotra
>             Fix For: 1.8
> Oak Lucene indes is currently using Tika 1.5 version while current latest release of
Apache Tika is 1.14, I think there're lots of "interesting" bugs fixed, and possibly improvements
(performance, more accurate text extraction, etc.) we could get at almost 0 cost by just bumping
the version number.

This message was sent by Atlassian JIRA

View raw message