You are viewing a plain text version of this content. The canonical link for it is here.
Posted to dev@nutch.apache.org by "Doğacan Güney (JIRA)" <ji...@apache.org> on 2007/09/06 15:24:31 UTC
[jira] Commented: (NUTCH-530) Add a combiner to improve performance
on updatedb
[ https://issues.apache.org/jira/browse/NUTCH-530?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel#action_12525418 ]
Doğacan Güney commented on NUTCH-530:
-------------------------------------
Andrzej, what do you think about this one in light of Emmanuel's last comment? I am still uneasy about ScoringFilters running twice, but I think Emmanuel is right that semantics don't change.
> Add a combiner to improve performance on updatedb
> -------------------------------------------------
>
> Key: NUTCH-530
> URL: https://issues.apache.org/jira/browse/NUTCH-530
> Project: Nutch
> Issue Type: Improvement
> Environment: java 1.6
> Reporter: Emmanuel Joke
> Assignee: Emmanuel Joke
> Fix For: 1.0.0
>
> Attachments: NUTCH-530.patch
>
>
> We have a lot of similar links with status "linked" generated at the ouput of the map task when we try to update the crawldb based on the segment fetched.
> We can use a combiner to improve the performance.
--
This message is automatically generated by JIRA.
-
You can reply to this email to add a comment to the issue online.