The Australian government's open data initiative is in the laudable business of publishing publicly accessible data about the government's actions and spending, in order to help scholars, businesses and officials understand and improve its processes. Read the rest
Microsoft co-founder Paul Allen funded the Allen Brain Observatory, a detailed, rich data-set derived from parts of a mouse-brain: what's striking is that the Allen Institute released all the data into the public domain, at once, as soon as it was available, which is exactly what you'd want the publicly funded alternatives to do, and what they almost never do. Read the rest
In a lead editorial in the current Nature, John Wilbanks (formerly head of Science Commons, now "Chief Commons Officer" for Sage Bionetworks) and Eric Topol (professor of genomics at the Scripps Institute) decry the mass privatization of health data by tech startups, who're using a combination of side-deals with health authorities/insurers and technological lockups to amass huge databases of vital health information that is not copyrighted or copyrightable, but is nevertheless walled off from open research, investigation and replication. Read the rest
The Transatlantic Trade and Investment Partnership is an EU-US "trade agreement" that will allow corporations to sue governments in secret tribunals to force them to repeal their safety, environmental and labor laws. Read the rest
Rogue archivist Carl Malamud sez, "I just finished ripping 30 DVDs from the IRS. This is the monthly feed of nonprofit tax returns. I now have 7,442,564 of these returns spinning on the net. I've had it.
This year, the IRS upped the cost of this feed to $2910. I've already spent $16,137 on this brain dead format. For 2 years, I've been writing to the IRS to suggest better ways. Dropbox anybody? An FTP server?" Read the rest
Rogue archivist Carl Malamud sez,
Read the rest
If you want access to all the tax filings of US nonprofit corporations, the IRS will sell you sets of DVDs for $2580 per year of data. We acquired all of these filings from 2002 to the present, a set of DVDs weighing 98.7 pounds. I'm pleased to report that all 6,461,326 of those returns are now successfully extracted and available on our new bulk data feed.
This data really should be available directly from the IRS at no charge. Accordingly, we've drafted a deed of gift offering the system back to the government.
Until the .gov people do take it over, we're offering access to all 5 TBytes of data using the http, ftp, and rsync protocols. Our hope is that developers will come up with lots of new uses for this information. In order to make the database even more useful, we've started working with Captricity to extract data from the forms and make it available as computable data (e.g., CVS files instead of TIFF images!).
Once search engines such as Google finish indexing the data, the tax filings of nonprofits will show up in the search results. When you search for a nonprofit, the first thing you see ought to be their home page. But, the next thing you ought to see are things like how much they pay their CEO, how much revenue goes for fundraising, and if they spend money to lobby public officials.
Nonprofits in the US had $1.87 trillion in 2009 revenues and it is these periodic filings that make the nonprofit marketplace work properly, just like SEC EDGAR filings help make the corporate markets work properly.
Adam sez, "The first Open-data Cities Conference takes place in Brighton, England next week. It's aimed at local councils and government agencies who want to open up more of their datasets, and giving them ideas and practical help on how to do it. There's some good speakers, including Tom Steinberg from MySociety and Rufus Pollock from the Open Knowledge Foundation."
The high-profile conference – the first of its kind in the United Kingdom – will focus on how publicly-funded organisations can engage with citizens to build more creative, prosperous and accountable communities.
It will be attended by more than 200 people who believe the value of public data is greatest when it is freely and openly shared. They will be leaders from the public sector, arts and cultural organisations, and creative and digital industries.
The focus will be on the opportunities to improve the lives of more than 10 million citizens in the UK’s biggest cities.
Here's a terrific article by Gilles Frydman at e-patients.net advocating for opposition to H.R. 3699, aka The Research Works Act (RWA). The bill before Congress would seriously impede "the ability of patients and caregivers, researchers, physicians and healthcare professionals to access and use critical health-related information in a timely manner." (@timoreilly via @epatientdave) Read the rest