[
https://issues.apache.org/jira/browse/LUCENE-2341?page=com.atlassian.jira.plugin.system.issuetabpanels:all-tabpanel
]
Michał Dybizbański updated LUCENE-2341:
---------------------------------------
Attachment: morfologik-stemming-1.5.0.jar
LUCENE-2341.diff
Hi
This patch introduces stemming filter and analyzer, that use [Morfologik
library|http://morfologik.blogspot.com], developed by Dawid Weiss and Marcin
Miłkowski.
Tokens are stemmed by Morfologik with a dictionary, and current distribution
provides a dictionary for polish language.
The MorfologikFilter yields one or more terms for each token. Each of those
terms is given the same position in the index.
I'm attaching a binary distribution of the library
(morfologik-stemming-1.5.0.jar), that needs to be placed in
modules/analysis/morfologik/lib/ subdirectory.
It is also available as a [Maven
artifact|http://mvnrepository.com/artifact/org.carrot2/morfologik-stemming/1.5.0].
The library is BSD-licensed and a dictionary uses data from [Polish dictionary
for aspell/ispell/myspell (SJP.PL)|http://www.sjp.pl/slownik/en/], which is
licensed under GPL, LGPL, MPL and CC SA licenses.
This is my first contribution to the Lucene project, so please be forgiving :)
Thanks to Dawid for help.
Regards,
Michał
> explore morfologik integration
> ------------------------------
>
> Key: LUCENE-2341
> URL: https://issues.apache.org/jira/browse/LUCENE-2341
> Project: Lucene - Java
> Issue Type: New Feature
> Components: modules/analysis
> Reporter: Robert Muir
> Assignee: Dawid Weiss
> Attachments: LUCENE-2341.diff, morfologik-stemming-1.5.0.jar
>
>
> Dawid Weiss mentioned on LUCENE-2298 that there is another Polish stemmer
> available:
> http://sourceforge.net/projects/morfologik/
> This works differently than LUCENE-2298, and ideally would be another option
> for users.
--
This message is automatically generated by JIRA.
For more information on JIRA, see: http://www.atlassian.com/software/jira
---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]