[
https://issues.apache.org/jira/browse/OAK-5048?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=16072160#comment-16072160
]
Chetan Mehrotra edited comment on OAK-5048 at 7/3/17 9:21 AM:
--------------------------------------------------------------
Following are the sizes post change
||module||1.5 (current used) ||1.15 (latest and proposed)||Remarks||
|oak-run|44M|44M| Embeds tika-core and tika-parsers|
|oak-lucene|5.5 M|5.5 M| No embed|
|oak-solr-core|155K|155K| No embed|
|oak-examples/standalone|72M|99M| Embeds whole Tika stuff|
|oak-examples/webapp|53M|78M| Embeds whole Tika stuff|
was (Author: chetanm):
Following are the sizes post change
||module||1.5||1.15||Remarks||
|oak-run|44M|44M| Embeds tika-core and tika-parsers|
|oak-lucene|5.5 M|5.5 M| No embed|
|oak-solr-core|155K|155K| No embed|
|oak-examples/standalone|72M|99M| Embeds whole Tika stuff|
|oak-examples/webapp|53M|78M| Embeds whole Tika stuff|
> Upgrade to latest Tika version
> ------------------------------
>
> Key: OAK-5048
> URL: https://issues.apache.org/jira/browse/OAK-5048
> Project: Jackrabbit Oak
> Issue Type: Improvement
> Components: lucene
> Reporter: Tommaso Teofili
> Assignee: Chetan Mehrotra
> Fix For: 1.8
>
>
> Oak Lucene indes is currently using Tika 1.5 version while current latest
> release of Apache Tika is 1.14, I think there're lots of "interesting" bugs
> fixed, and possibly improvements (performance, more accurate text extraction,
> etc.) we could get at almost 0 cost by just bumping the version number.
--
This message was sent by Atlassian JIRA
(v6.4.14#64029)