New MySQL user accounts

Hi moles,

This message is for those of you who are actively using direct database access on the MySQL server at Teragrid/SDSC.

Our direct access database is being moved off Teragrid at SDSC and onto our own servers at flossdata.syr.edu.

To accomplish this move, I'm going to need to re-create our active user accounts on MySQL.

(This is a good opportunity to clean house anyway since there were a number of user accounts on there for people who have moved on or are no longer active.)

Please respond to this email (send to megansquire) IF you still require direct database access, and let me know the username you were using, if possible. I will recreate your user account and let you know the connection details.

Remember the flat files are still available at: http://code.google.com/p/flossmole/downloads/list and no login is required to use these.

Google Code data released for Nov 2012

The datasource ID is 350, and we've got a fresh run of Google Code project names, thanks to Audris. Also the People collector is fixed up now, so we're getting both the userid and the user name if available. So note the new column in the people table.

Go here to get the data files

November 2012 data released

The November 2012 data has been released to Google Code. Go get it!

Note: we are in the process of copying from the Teragrid, so we will not be putting any more data up there for the time being. More details on direct database access will be forthcoming.

september Launchpad data is released

Huge thank you to our former student and flossmole alumnus Christian Funkhouser for working on re-writing the Launchpad collector to use the API.

Get the code!

The datasource_id for Launchpad in September is 342!

September 2012 data released

Data files have been released for September 2012. Go check it out on our Google Code downloads page or sign up for direct database access.

Special release notes:
--The Google Code, Github, and Launchpad collections are not included this month.
--Alioth is back, and Tigris is back, including emails.

August 2012 data released

Data files have been released for August 2012. Go check it out on our Google Code downloads page or sign up for direct database access.

Special release notes:
--The Google Code, Github, and Launchpad collections are not included this month.

Freecode New Project Registrations (1998-2011) and language tags

This chart shows the new project registrations for each year 1998-2011, and what programming language those projects were tagged with.

For example, 2003 was the highest year for new "C" projects to be registered with Freecode (then called Freshmeat).

(Couple of caveats about the data here: (1) A project could be tagged with one language when it's created, and then the tag could change later. For example a project may have been created as "C" and then later switched its tags to "Java". This chart will show whatever the most current language tag is for that project. I have not calculated the number of times this happens or if it is likely, but I suspect that it is not too common. (2) The languages in the legend are in order of the total number of times they exist as tags for any project. (3) Not every project has a language tag for itself.)

Here is the SQL code used to generate the data sets for this graph:

SELECT YEAR(p.date_added ) , COUNT( DISTINCT p.project_id )
FROM fm_projects p
INNER JOIN fm_project_tags t
ON t.project_id = p.project_id
WHERE p.datasource_id = 316
AND t.datasource_id=316
AND t.tag_name="C"
GROUP BY 1
ORDER BY 1;

Substitute your current datasource_id and any valid programming language tag. I used the following:

C
Java
C++
PHP
Perl
Python
JavaScript

These are in order of total number of times they appeared in the sitewide tag list over time.

Other top languages, in order of frequency the tag is used:

Unix Shell
XML
SQL
HTML
Ruby
Tcl

July 2012 data released

Data files have been released for July 2012. Go check it out on our Google Code downloads page or sign up for direct database access.

Special release notes:
--Alioth data has not made it into the SDSC site (direct db access) but we are working on it. In the meantime, the data is on Google Code site.
--The Google Code collector is still using an old list of projects, but this bug is being worked on.
--Launchpad collector is being re-written
--Alioth collector is being re-written

May 2012 data releases

Data files have been released for May 2012. Go check it out on our Google Code downloads page.

Re-writes:
--Free Software Foundation has been re-written from scratch to match their new layout.
--Google Code collector has been re-written to fix a few bugs (still running)
--Launchpad has been finished and will be re-written for June to fix bugs
--Alioth is being re-written to fix bugs

Syndicate content