nutch-dev mailing list archives

Site index · List index
Message view « Date » · « Thread »
Top « Date » · « Thread »
From Doğacan Güney (JIRA) <>
Subject [jira] Commented: (NUTCH-530) Add a combiner to improve performance on updatedb
Date Thu, 06 Sep 2007 13:24:31 GMT


Doğacan Güney commented on NUTCH-530:

Andrzej, what do you think about this one in light of Emmanuel's last comment? I am still
uneasy about ScoringFilters running twice,  but I think Emmanuel is right that semantics don't

> Add a combiner to improve performance on updatedb
> -------------------------------------------------
>                 Key: NUTCH-530
>                 URL:
>             Project: Nutch
>          Issue Type: Improvement
>         Environment: java 1.6
>            Reporter: Emmanuel Joke
>            Assignee: Emmanuel Joke
>             Fix For: 1.0.0
>         Attachments: NUTCH-530.patch
> We have a lot of similar links with status "linked" generated at the ouput of the map
task when we try to update the crawldb based on the segment fetched.
> We can use a combiner to improve the performance.

This message is automatically generated by JIRA.
You can reply to this email to add a comment to the issue online.

View raw message