Home » Articles posted by James A Jacobs (Page 230)
Author Archives: James A Jacobs
Job losses in recessions visualized
Recently this graphic was posted on the "The Gavel" blog of the Speaker of the House (What 3.6 Million Jobs Lost Over 13 Months Looks Like, by "Karina," February 6th, 2009). It shows the number of jobs lost (and recovered) in the recessions of 1990, 2001, and 2008. By juxtaposing the three time periods over each other starting with the peak job month and showing employment change by month, it gives a startling comparison that highlights the severity of the current situation. It also implies that we will have a long time to wait before we reach our previous peak month.
I was curious about this graph and did a little follow up that I share below. Data librarians may find this a bit tedious, but for those who have never used raw data, it may be useful as an illustration of the difference between "data" (the raw numbers that you put into statistical software) and "statistics" (the human-viewable tables and graphs that we see in publications).
Unfortunately, as is often the case with statistical information, the source given for the graphic is incomplete: simply "Bureau of Labor Statistics." I could not find the graphic itself on the bls.gov site, so I assume that the chart was constructed from BLS data, specifically, the Current Population Survey or the Current Employment Statistics Survey. These two surveys count employment differently -- one is a survey of individuals and the other is a survey of employers.
There is a similar, but not identical, chart ("Percent change in total nonfarm employment, from beginning of recession) in the January 2009 (released February 6, 2009) Current Employment Statistics Highlights, Monthly (Bureau of Labor Statistics), so my guess is that someone at the Speaker's office built the chart from the raw CES data.
Just out of curiosity, I went to the CES "Most Requested Statistics webpage and downloaded "Total Nonfarm Employment - CES0000000001" for 1990 through the end of 2008. Raw data suitable for analysis even look "raw," not even like a statistical table:
1990,Jan,109151 1990,Feb,109396 1990,Mar,109611 1990,Apr,109651 1990,May,109800 1990,Jun,109817 1990,Jul,109775 1990,Aug,109567 1990,Sep,109485 1990,Oct,109324 1990,Nov,109180 1990,Dec,109120 1991,Jan,109001 1991,Feb,108695 1991,Mar,108535 1991,Apr,108324 1991,May,108196 1991,Jun,108283 ...
Of course, it is relatively easy, using statistical software, to construct tables and graphs from raw data. Here, for example, is a published statistical table with essentially the same raw information (but from CPS, not CES) that I downloaded. (See the full table from Employment from the BLS household and payroll surveys: summary of recent trends, February 6, 2009).
By using the raw data to create a graph, one can tell a story that has more impact than just a table of numbers. It is relatively easy to get these data into statistical software. I used Excel and Stata to create a small time-series data file. I organized it by month (from month "1" to month "48") with each row of the data file having data for 3 recessions. The first row has data for the first month of the three recessions, the second row has data for the second month, etc. The CES data has employment totals in millions. For example, the employment for 2008:
138152 138080 137936 137814 137654 137517 137356 137228 137053 136732 136352 135755 135178
I had to compute a new variable for each recession: the cumulative number of jobs lost. So, for example, 2008:
138152 0 138080 -72 137936 -216 137814 -338 137654 -498 137517 -635 137356 -796 137228 -924 137053 -1099 136732 -1420 136352 -1800 135755 -2397 135178 -2974
The first 12 months with all three recessions (v1, v2, v3) and the computed variables (1990, 2001, 2008) look like this:
month v1 1990 v2 2001 v3 2008 1 109817 0 132530 0 138152 0 2 109775 -42 132500 -30 138080 -72 3 109567 -250 132219 -311 137936 -216 4 109485 -332 132175 -355 137814 -338 5 109324 -493 132047 -483 137654 -498 6 109180 -637 131922 -608 137517 -635 7 109120 -697 131762 -768 137356 -796 8 109001 -816 131518 -1012 137228 -924 9 108695 -1122 131193 -1337 137053 -1099 10 108535 -1282 130901 -1629 136732 -1420 11 108324 -1493 130723 -1807 136352 -1800 12 108196 -1621 130591 -1939 135755 -2397
Here is a complete tab-separated-values version of the data file I constructed. Then I used Stata to build a graph and it looks very much like the one at the Speaker's Blog.
Of course, when one tells one story, one leaves out other stories. This graphic doesn't show that the starting points of the recessions were different: 1990: 109 million 2001: 132 million 2008: 138 million
Open re-usable government informationOne could use the raw data to tell a lot of different stories and analyze the data in many different ways. And that brings me to the connection between all this and why we need to be sure that government information is not just "free as in beer" but also "free as in open."
It is important for statistical agencies to publish statistics to help us understand their raw data. But, it is also essential that they provide us with the raw data so that we can better understand their statistics and do our own analyses. Most of the statistical agencies of the U.S. government do an excellent job of making their raw data easily available. In fact, the rest of government would do well to use statistical agencies as a model for instantiating their information in usable and re-usable formats (in addition to any publishing and presentation of their data/information) so that the information, whether it is text or images or video or sound or numbers, can be used, reused, analyzed, stored, and preserved.
Continue readingPrivacy on the Internet: going or already gone?
Two recent stories in the New York Times summarize the issue of privacy in the digital age.
- Do We Need a New Internet?, By JOHN MARKOFF, New York Times, February 15, 2009.
- As Data Collecting Grows, Privacy Erodes, By NOAM COHEN, New York Times, February 16, 2009.
Still no official word on PACER project
An article in the New York Times adds a little more to the story of the Public Access to Court Electronic Records (PACER) saga. The article also says of Carl Malamud: "Mr. Malamud said his years of activism had led him to set a long-shot goal: serving in the Obama administration, perhaps even as head of the Government Printing Office."
- An Effort to Upgrade a Court Archive System to Free and Easy, By JOHN SCHWARTZ, New York Times, February 13, 2009.
Those courts, with the help of the Government Printing Office, had opened a free trial of Pacer at 17 libraries around the country. Mr. Malamud urged fellow activists to go to those libraries, download as many court documents as they could, and send them to him for republication on the Web, where Google could get to them. Aaron Swartz, a 22-year-old Stanford dropout and entrepreneur who read Mr. Malamud's appeal, managed to download an estimated 20 percent of the entire database: 19,856,160 pages of text. Then on Sept. 29, all of the free servers stopped serving. The government, it turns out, was not pleased. A notice went out from the Government Printing Office that the free Pacer pilot program was suspended, "pending an evaluation." A couple of weeks later, a Government Printing Office official, Richard G. Davis, told librarians that "the security of the Pacer service was compromised. The F.B.I. is conducting an investigation." ...At the administrative office of the courts, a spokeswoman, Karen Redmond, said she could not comment on the fate of the free trial of Pacer, or whether there had been a criminal investigation into the mass download. The free program "is not terminated," Ms. Redmond said. "We'll just have to see what happens after the evaluation."See also Why was PACER suspended? and More on PACER. Continue reading
The Stimulus Deal: The Latest Tally
The reporters at ProPublica have done a great job of comparing the House, Senate, and Conference versions of the Stimulus bill as best they can given that we don't yet have a copy of the final version of the bill. Check out their chart of changes:
- The Stimulus Deal: The Latest Tally, by Michael Grabell, ProPublica, February 12, 2009.
Well, we wanted to tell you what’s in the final, $789 billion stimulus package, but guess what? The bill still hadn't been released as of late Thursday. So the best we can do is this partial account, which is based on summaries released so far. Where an item is blank, it means we don't yet have the figure.Continue reading
Virginia asks citizens for stimulus ideas on web
The Commonwealth has developed a website for citizens, groups, localities, and others to use to share project proposals for potential funding from the expected federal stimulus package. As the stimulus package becomes finalized, more information and details will be made available on this site. ...These suggestions will be posted on the Virginia Stimulus webpage and shared with state government officials. All the information will be public.It has a page of proposals and is making the data available under a Creative Commons Attribution 3.0 License. Continue reading

Latest Comments