[ 
https://issues.apache.org/jira/browse/LUCENE-6993?page=com.atlassian.jira.plugin.system.issuetabpanels:all-tabpanel
 ]

Mike Drob updated LUCENE-6993:
------------------------------
    Attachment: LUCENE-6993.patch

Attaching a patch against trunk that updates the TLD Macro file and the 
UAX29URLEmailTokenizerImpl.

When running {{ant jflex}} I had to increase the amount of heap space available 
due to the increased number of TLDs, not sure if this will result in a negative 
impact to the rest of the build.

The {{.an}} and {{.tp}} domains were removed from the list, 
{{random.text.with.urls}} was updated accordingly.

> Update TLDs to latest list
> --------------------------
>
>                 Key: LUCENE-6993
>                 URL: https://issues.apache.org/jira/browse/LUCENE-6993
>             Project: Lucene - Core
>          Issue Type: Improvement
>          Components: modules/analysis
>            Reporter: Mike Drob
>         Attachments: LUCENE-6993.patch
>
>
> We did this once before in LUCENE-5357, but it might be time to update the 
> list of TLDs again. Comparing our old list with a new list indicates 800+ new 
> domains, so it would be nice to include them.



--
This message was sent by Atlassian JIRA
(v6.3.4#6332)

---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]

Reply via email to