This is an automated email from the ASF dual-hosted git repository.
krickert pushed a change to branch OPENNLP-1850-3-dl
in repository https://gitbox.apache.org/repos/asf/opennlp.git
omit 125162bd8 OPENNLP-1850 DL hardening: IllegalArgumentException null
contract and null-element guards
omit 4515f61dd OPENNLP-1850 Follow AlignedText.normalized() -> CharSequence
in NameFinderDL
omit 9aec5a624 OPENNLP-1850 Review nits: extract testable DL guards;
merge-copy; capitalize msgs; migration note
omit 9b783fe06 OPENNLP-1850 Make mergeOverlappingSpans O(n log n) (dl)
omit 97d86eb19 OPENNLP-1850 Reject non-finite logits in softmax, not just
NaN (dl)
omit fa87d9887 OPENNLP-1850 Fully-qualify TokenNameFinder javadoc links in
NameFinderDL
omit 6f2dd62a0 OPENNLP-1850 Fail loud on corrupt document-classification
model output
omit 200d05219 OPENNLP-1850 Fail fast on null finder input; fix the GPU
eval test options
omit b213e99ba OPENNLP-1850 Harden fail-loud paths in the DL components
omit 35e8e64b9 OPENNLP-1850 Add real-model chunk-boundary eval tests; drop
dead label constants
omit d28988c47 OPENNLP-1850 Resolve overlapping chunk spans and compose the
input alignment
omit 93c141ad3 OPENNLP-1850 Add OffsetMappingNameFinder capability
interface and a findInOriginal end-to-end test
omit 9e89e779d OPENNLP-1850 Offset-safe, Unicode-aware input normalization
in the DL components
omit 99d108c6c OPENNLP-1850 Review nits: rename
searchAnalyzer->matchingAnalyzer; drop 'search' framing in profile docs
omit 4d67305f4 OPENNLP-1850 Review nits: add Turkish profile; derive
coverage from the enum (profiles)
omit 32eac82de OPENNLP-1850 Resolve Norwegian nb/nn to the Norwegian
profile (profiles)
omit a6ed01e4d OPENNLP-1850 Per-language NormalizationProfile registry (2c)
omit 1b1f99fcd OPENNLP-1850 Term model hardening: fail loud on a
contract-violating Lemmatizer
omit df04c058b OPENNLP-1850 Review nits: TermAnalyzer javadoc references
matchingAnalyzer()
omit 214c94253 OPENNLP-1850 Review nits: rename dashes()->dash(); LEMMA
doc+test; soften forward-link (Term)
omit cad0e810b OPENNLP-1850 Layered Term model: Term, TermAnalyzer (2b)
omit de75f594f OPENNLP-1850 Review: uax29 javadoc pass, IAE guards on
WordTokenizer, loader consistency, new tests
omit a62439329 OPENNLP-1850 Perf: hoist the per-char volatile reads in
WordBreakProperty/ExtendedPictographic
omit 721f08b79 OPENNLP-1850 Review: drop lazy-init justification comments
in WordBreakProperty/ExtendedPictographic
omit a44de61c7 OPENNLP-1850 Review nits: ExtendedPictographic fail-loud
parity + doc; WordType heuristic note (tokenizer)
omit 11ea9367e OPENNLP-1850 Fail loud on a Word_Break line missing its ';'
(tokenizer)
omit 3a993d160 OPENNLP-1850 UAX #29 word tokenizer: WordSegmenter,
WordTokenizer, WordType (2a)
add 47fd462a0 OPENNLP-1859: Add tests for BilouCodec encode/decode and
outcome compatibility (#1135)
add e2ffecd8a OPENNLP-1862: UAX #29 word tokenizer — WordSegmenter,
WordTokenizer, WordType (#1110)
add 2260b55ec Minor: Regenerated NOTICE File for
e2ffecd8a278653969d39045dff4dccfbdc9569c (#1142)
add e67c8bd29 OPENNLP-1850 Layered Term model: Term, TermAnalyzer (2b)
add 69906de04 OPENNLP-1850 Review nits: rename dashes()->dash(); LEMMA
doc+test; soften forward-link (Term)
add 1cb0bb03a OPENNLP-1850 Review nits: TermAnalyzer javadoc references
matchingAnalyzer()
add 37ce9e128 OPENNLP-1850 Term model hardening: fail loud on a
contract-violating Lemmatizer
add 59329f1ed OPENNLP-1850 Per-language NormalizationProfile registry (2c)
add 9646cd145 OPENNLP-1850 Resolve Norwegian nb/nn to the Norwegian
profile (profiles)
add ec916b649 OPENNLP-1850 Review nits: add Turkish profile; derive
coverage from the enum (profiles)
add f84ba8cc1 OPENNLP-1850 Review nits: rename
searchAnalyzer->matchingAnalyzer; drop 'search' framing in profile docs
add efe5506f7 OPENNLP-1850 Offset-safe, Unicode-aware input normalization
in the DL components
add 5bf589a14 OPENNLP-1850 Add OffsetMappingNameFinder capability
interface and a findInOriginal end-to-end test
add ba1668092 OPENNLP-1850 Resolve overlapping chunk spans and compose the
input alignment
add ddfaed50d OPENNLP-1850 Add real-model chunk-boundary eval tests; drop
dead label constants
add a7a419991 OPENNLP-1850 Harden fail-loud paths in the DL components
add 993dce689 OPENNLP-1850 Fail fast on null finder input; fix the GPU
eval test options
add 0e2d54911 OPENNLP-1850 Fail loud on corrupt document-classification
model output
add d4317a4eb OPENNLP-1850 Fully-qualify TokenNameFinder javadoc links in
NameFinderDL
add f4832b714 OPENNLP-1850 Reject non-finite logits in softmax, not just
NaN (dl)
add f7363f385 OPENNLP-1850 Make mergeOverlappingSpans O(n log n) (dl)
add e81f23877 OPENNLP-1850 Review nits: extract testable DL guards;
merge-copy; capitalize msgs; migration note
add e15875b0a OPENNLP-1850 Follow AlignedText.normalized() -> CharSequence
in NameFinderDL
add 018ce5835 OPENNLP-1850 DL hardening: IllegalArgumentException null
contract and null-element guards
This update added new revisions after undoing existing revisions.
That is to say, some revisions that were in the old version of the
branch are not in the new version. This situation occurs
when a user --force pushes a change and generates a repository
containing something like this:
* -- * -- B -- O -- O -- O (125162bd8)
\
N -- N -- N refs/heads/OPENNLP-1850-3-dl (018ce5835)
You should already have received notification emails for all of the O
revisions, and so the following emails describe only the N revisions
from the common base, B.
Any revisions marked "omit" are not gone; other references still
refer to them. Any revisions marked "discard" are gone forever.
No new revisions were added by this update.
Summary of changes:
.../opennlp/tools/namefind/BilouCodecTest.java | 381 +++++++++++++++++++++
opennlp-distr/src/main/readme/NOTICE | 28 +-
2 files changed, 402 insertions(+), 7 deletions(-)