January data update

Freshmeat, Objectweb, Rubyforge, Savannah, FSF, Tigris - all done, waiting for release

Github collector is broken, have student working on fixing that now. They suddenly made it hard to seed the initial project list so we're trying to figure out a way to get the entire corpus of projects. I'm getting flashbacks of when SF got really big and made it difficult for everyone to work with.

Google Code is plugging along. We're on the 7th out of 8 collection processes. Won't be too much longer on that.

Debian data is being cleaned for release. Have a meeting with a student tomorrow to see the status of that cleaning but it should be released this weekend.

New features in the hopper: (1) auto-generating DOI information at the time of release so we won't be behind, (2) federated search - I know that is huge, but we'll see how it goes.

Data Resources: