lucene-dev mailing list archives

Site index · List index
Message view « Date » · « Thread »
Top « Date » · « Thread »
From "Robert Muir (JIRA)" <j...@apache.org>
Subject [jira] Updated: (LUCENE-2503) light/minimal stemming for euro languages
Date Thu, 17 Jun 2010 20:44:25 GMT

     [ https://issues.apache.org/jira/browse/LUCENE-2503?page=com.atlassian.jira.plugin.system.issuetabpanels:all-tabpanel
]

Robert Muir updated LUCENE-2503:
--------------------------------

    Attachment: LUCENE-2503.patch

patch, not ready for committing. only some of these are ready, others need tests (where I
intentionally put a fail() placeholder to indicate they are still untested).

also i didn't implement the finnish one yet, but it contains various implementations for 9
euro languages.


> light/minimal stemming for euro languages
> -----------------------------------------
>
>                 Key: LUCENE-2503
>                 URL: https://issues.apache.org/jira/browse/LUCENE-2503
>             Project: Lucene - Java
>          Issue Type: New Feature
>          Components: contrib/analyzers
>    Affects Versions: 3.1, 4.0
>            Reporter: Robert Muir
>            Assignee: Robert Muir
>            Priority: Minor
>             Fix For: 3.1, 4.0
>
>         Attachments: LUCENE-2503.patch
>
>
> The snowball stemmers are very aggressive and it would be nice if there were lighter
alternatives.
> Some applications may want to perform less aggressive stemming, for example:
> http://www.lucidimagination.com/search/document/5d16391e21ca6faf/plural_only_stemmer
> Good, relevance tested algorithms exist and I think we should provide these alternatives.

-- 
This message is automatically generated by JIRA.
-
You can reply to this email to add a comment to the issue online.


---------------------------------------------------------------------
To unsubscribe, e-mail: dev-unsubscribe@lucene.apache.org
For additional commands, e-mail: dev-help@lucene.apache.org


Mime
View raw message