The seed for this crawl was a list of every host in the Wayback Machine
This crawl was run at a level 1 (URLs including their embeds, plus the URLs of all outbound links including their embeds)
The WARC files associated with this crawl are not currently available to the general public.
Google Project Hosting powers Project Hosting on Google Code and Eclipse Labs.
It provides a fast, reliable, and easy open source hosting service with the following features: