Download of heritrix-0.10.0.zip (heritrix-0.10.0.zip ( external link: SF.net): 12,117,387 bytes) will begin shortly. If not so, click link on the left.
The archive-crawler project is building Heritrix: a flexible, extensible, robust, and scalable web crawler capable of fetching, archiving, and analyzing the full diversity and breadth of internet-accesible content.