mahout-dev mailing list archives

Site index · List index
Message view « Date » · « Thread »
Top « Date » · « Thread »
From "Robin Anil (JIRA)" <>
Subject [jira] [Commented] (MAHOUT-1190) SequentialAccessSparseVector function assignment is very slow
Date Fri, 12 Apr 2013 17:00:18 GMT


Robin Anil commented on MAHOUT-1190:

I dont see the edge of SeqSV against RandSV.
> SequentialAccessSparseVector function assignment is very slow
> -------------------------------------------------------------
>                 Key: MAHOUT-1190
>                 URL:
>             Project: Mahout
>          Issue Type: Bug
>            Reporter: Dan Filimon
> Currently when calling .assign() on a SASV with another vector and a custom function,
it will iterate through it and assign every single entry while also referring it by index.
> This makes the process *hugely* expensive. (on a run of BallKMeans on the 20 newsgroups
data set, profiling reveals that 92% of the runtime was spent updating assigning the vectors).
> Here's a prototype patch:

This message is automatically generated by JIRA.
If you think it was sent incorrectly, please contact your JIRA administrators
For more information on JIRA, see:

View raw message