<?xml version="1.0"?>
<rss version="2.0">
   <channel>
      <title>The Book of Trogool</title>
      <link>http://scienceblogs.com/bookoftrogool/</link>
      <description>E-research, cyberinfrastructure, data curation, open access... an academic librarian examines how computers change research and libraries.</description>
      <language>en</language>
      <copyright>Copyright 2010</copyright>
      <lastBuildDate>Thu, 01 Apr 2010 08:32:06 -0600</lastBuildDate>
      <generator>http://www.sixapart.com/movabletype/?v=4.32-en</generator>
      <docs>http://blogs.law.harvard.edu/tech/rss</docs> 

      
      <item>
         <title>Introducing... Curatr!</title>
          <description><![CDATA[<p>Not good at organizing your thoughts, much less your research notes? Think publishing your data should be as easy as falling off the couch?</p>

<p>Yes, well, me too. So I've built a new site to do it all for you, and I'm calling it <a href="http://gavialib.com/home/curatr/">Curatr</a>. </p>

<ul><li>Built on all the shiniest and most proprietary technologies, from HyperCard to Flash</li>
<li>Automatically builds the most appropriate storage and interaction models based on computerized analysis of provided data. No documentation needed!</li>
<li>Auto-organizing. Never touch metadata again!</li>
<li>Can be managed by a single graduate student in two hours a week without any prior training</li>
<li>Wholly grant-funded, so you never have to worry about cost or sustainability</li>
<li>Never needs updating. Curatr does all the pesky link-breaking for you, as its underlying technology stack migrates to keep up with the times.</li></ul>

<p>Join the Curatr perpetual beta today!</p> <a href="http://scienceblogs.com/bookoftrogool/2010/04/introducing_curatr.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/04/introducing_curatr.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/04/introducing_curatr.php</guid>
         <category>Tactics</category>
         
         <pubDate>Thu, 01 Apr 2010 08:32:06 -0600</pubDate>
      </item>
      
      <item>
         <title>Tidbits, 30 March 2010</title>
          <description><![CDATA[<p>Tuesday seems a good day for tidbits. (I am head-down in my UKSG presentation and class stuff at the moment, so kindly forgive posting slowness.)</p>

<ul>
<li>One argument I rarely see made for open access that should perhaps be made more often is that it reduces friction in both accessing <em>and</em> providing information. <a href="http://www.govexec.com/dailyfed/0310/031810e1.htm">Want to reduce the overhead of responding to FOIA requests? Post the information online.</a></li>
	<li>Data, data, we love data! <a href="http://www.researchinformation.info/features/feature.php?feature_id=255">Data is at the heart of new science ecosystem</a> and <a href="http://www.symmetrymagazine.org/cms/?pid=1000770">Preserving the Data Harvest</a>. Oh, and if you hadn't noticed, <a href="http://dataspora.com/blog/the-data-singularity-is-here/">The Data Singularity is Here</a>.</li>
	<li>Some good lay-level explanations of digital preservation <a href="http://gizmodo.com/5495191/giz-explains-how-data-dies-and-how-it-can-be-saved">from Gizmodo</a> and from <a href="http://alanake.wordpress.com/so-you-want-to-keep-all-your-stuff/">Alan's notes</a>.</li>
	<li>FXPAL asks <a href="http://palblog.fxpal.com/?p=3177">Whither data privacy?</a> in the wake of reidentification research.</li>
	<li><a href="http://www.istl.org/10-winter/refereed2.html">Are Article Influence Scores Comparable across Scientific Fields?</a>: In a word, "no." More musings on impact: <a href="http://blogs.nature.com/rpg/2009/06/22/on-article-level-metrics-and-other-animals">On article-level metrics and other animals</a>.</li>
	<li><a href="http://blog.okfn.org/2010/03/15/the-cake-test-of-freedom/">The cake test of freedom</a>: Clever! It's only open data if you can paint it on a cake... See also O'Reilly Radar on <a href="http://radar.oreilly.com/2010/03/truly-open-data.html">Truly Open Data</a>.</li>
	<li>The scientist and the librarian should be friends: <a href="http://blogs.lib.utexas.edu/texlibris/2010/03/09/texas-water-researchers-working-with-the-texas-digital-library/">Texas water researchers working with the Texas Digital Library</a>.</li>
	<li>Curation as contextualization: <a href="http://blockslabpillar.com/2010/03/06/the-importance-of-curation-in-a-metadata-data-driven-information-architecture/">The importance of curation in a metadata driven information architecture</a> and <a href="http://derivadow.com/2010/03/11/some-thoughts-on-moving-beyond-the-resource/">Some thoughts on curation &#8211; adding context and telling stories</a></li>
	<li>Did you know there was a <a href="http://bytesizebio.net/index.php/2010/03/10/bioinformatics-blog-carnival-1/">Bioinformatics Blog Carnival</a>? Me neither. Now we both know.</li>
	<li>A rundown on the research-collaboration tool HubZero: <a href="http://www.isgtw.org/?pid=1002354">Doing science on the hub</a>.</li>
	<li>Look what data can do! <a href="http://www.wired.com/wiredscience/2010/03/chile-earthquake-moved-entire-city-10-feet-to-the-west/">Chile Earthquake Moved Entire City 10 Feet to the West</a> But apply data with care: <a href="http://arstechnica.com/science/news/2010/03/were-so-good-at-medical-studies-that-most-of-them-are-wrong.ars">We're so good at medical studies that most of them are wrong</a>.</li></ul>

<p>As always, drop a comment or use the tag "trogool" on del.icio.us to bring something to my attention. Thanks!</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/tidbits_30_march_2010.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/tidbits_30_march_2010.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/tidbits_30_march_2010.php</guid>
         <category>Tidbits</category>
         
         <pubDate>Tue, 30 Mar 2010 17:16:01 -0600</pubDate>
      </item>
      
      <item>
         <title>A personal heroine: Henriette Davidson Avram</title>
          <description><![CDATA[<p>This is my blog post for <a href="http://findingada.com/">Ada Lovelace Day</a>, on which we celebrate technical achievement by women. I'm writing it the day before, and setting it to post at midnight.</p>

<p>I hope someone is writing a biography of Henriette Avram. I will be first in line to buy it. I desperately want to know how she did what she did.</p>

<p>Her achievement is generally, and appropriately, recognized as a technical one: designer and implementer of the MARC (MAchine Readable Cataloging) format still in use in hundreds of thousands of libraries worldwide. If that had been all: <i>dayenu</i>, it would have been enough. For all its baroqueness, its bizarre redundancies, its even more bizarre limitations, I find a lot to like in MARC, especially taking into account the computing environment under which it was developed.</p>

<p>But that is hardly all Henriette Avram did. Consider: she led a team of male engineers. How did <em>that</em> work, exactly, in the 1960s? Just to add to the mystery, Henriette Avram had no college degree. How did she win their respect for her obviously fearsome intellect? I don't know, but I bet there's one hell of a story in it. If winning the respect of male engineers in the 1960s <em>and</em> designing and implementing MARC had been all Henriette Avram did: <i>dayenu</i>, it would have been enough.</p>

<p>But consider: Henriette Avram also succeeded in getting MARC adopted in libraries across the nation and eventually across the globe. This, despite not having a library degree at the time (she was lavished with honorary degrees later), in a profession that (as non-library-degreed software engineers even today will attest) is <em>extremely</em> conscious of its professional degree. This, in an extraordinarily insular profession with a propensity to view anything digital as a diabolical plot! (No, that loathing is <a href="https://www.ideals.illinois.edu/handle/2142/872">not of recent vintage</a>, not one bit. Nor has it disappeared.)</p>

<p>So Henriette Avram designed and implemented MARC, <em>and</em> she led teams of male engineers, <em>and</em> she succeeded in winning adoption of her new system. Dayenu, and dayenu, and dayenu.</p>

<p>How did she do it? How? Please, library historians, <em>write this book</em> before we lose all the people she worked with. I can't even begin to express how important her story is to me, a librarian who's run afoul of both boy's-locker-room software-development types and librarians who fear and loathe the digital.</p>

<p>Henriette Avram died in 2006. A selection of obituaries:</p>

<ul><li><a href="http://www.nytimes.com/2006/05/03/us/03avram.html">New York Times</a></li>
<li><a href="http://www.washingtonpost.com/wp-dyn/content/article/2006/04/27/AR2006042702105.html">Washington Post</a>
<li><a href="http://articles.sfgate.com/2006-05-04/bay-area/17294007_1_henriette-d-avram-library-schools-library-thousands">San Francisco Gate</a></li>
<li><a href="http://www.pla.org/ala/alonline/currentnews/newsarchive/2006abc/april2006ab/avram.cfm">American Libraries</a></li></ul> <a href="http://scienceblogs.com/bookoftrogool/2010/03/a_personal_heroine_henriette_d.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/a_personal_heroine_henriette_d.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/a_personal_heroine_henriette_d.php</guid>
         <category>Miscellanea</category>
         
         <pubDate>Wed, 24 Mar 2010 00:00:28 -0600</pubDate>
      </item>
      
      <item>
         <title>OA publishers: just use HTML!</title>
          <description><![CDATA[<p>I was reading <a href="http://journals.tdl.org/jodi/issue/view/91">the latest issue of the Journal of Digital Information</a> today, and I found myself wishing I could turn <a href="http://lab.arc90.com/experiments/readability/">the Readability bookmarklet</a> loose on half its PDF-only articles.</p>

<p>I'm sorry, authors. I know you tried, but those PDFs are <em>terrible-looking</em>. Times New Roman, really? (The one in Arial is the worst, though.) Could we discuss your line-height and why it's not tall enough? Line-length, and why it's too long?</p>

<p>Sniff at me for an ex-typesetter if you like (I <em>am</em> an ex-typesetter, as it happens), but the on-the-ground reality is that I didn't read as much of those articles as I'd have read if they were, you know, <em>readable</em>. As for JoDI, their lack of a consistent look damages their brand and their credibility among their readers. Like it or not, centuries of print journals have created certain expectations for the quality of typesetting in a PDF.</p>

<p>So what's a shoestring open-access journal that can't afford professional typesetting to do? Believe you me, this is a common and vexing dilemma. It's not as though authors will lift a finger to make a publisher's production or branding job easier, as JoDI trenchantly demonstrates.</p>

<p>My answer: If you're not going to put effort into typesetting, chuck PDF. HTML is where it's at for you. Embrace the Web and its pitifully low standards for typography.</p>

<p>This is, of course, easier to say than to do. It does still take more technical savvy to produce decent HTML than to produce a bad PDF from the most typical manuscript formats. Making a print CSS stylesheet for your journal&#8212;which is also a good idea, to avoid grumbling from the print-dependent&#8212;is also eggheady. If your subject area is math-heavy, you have an entire new suite of problems.</p>

<p>On the whole, though, it's much easier to produce good HTML than good PDF. Moreover, bad PDFs are essentially irredeemable; there's nearly no way (and definitely no easy way) to reflow, re-typeset, or otherwise reformat them. If you go the HTML route, as your skills improve you will (trust me!) learn to fix your bad HTML, and if your content-management system is any good, you'll be able to go back and fix your old articles in a decently automated fashion.</p>

<p>As you rebrand your journal and its look and feel, which you eventually will unless and until the journal dies, you get a bonus: automatic rebranding of your old articles! They never have to look out-of-date, as old-school PDFs often do.</p>

<p>For those of you who have hopes of sending your journal to PubMed Central, there's an even more compelling reason to stick with HTML: PMC demands <a href="http://dtd.nlm.nih.gov/">NLM XML</a>, which you have <em>no hope</em> of producing straight from PDF. (From your typesetting format, perhaps, but <em>you have to know what you're doing</em>.) The skills you will learn from making HTML will transfer. PDF, not so much.</p>

<p>I admit that part of my reason for writing this is that I am hopelessly in love with the Readability bookmarklet and wish I could use it in more contexts. (I can't read Emerald or Informaworld HTML without it.) Still, my advice is heartfelt and I believe it's good.</p>

<p>I don't even have to use the Readability bookmarklet to read <a href="http://journal.code4lib.org/">the code4lib journal</a>. Just sayin'.</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/oa_publishers_just_use_html.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/oa_publishers_just_use_html.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/oa_publishers_just_use_html.php</guid>
         <category>Open Access</category>
         
         <pubDate>Tue, 23 Mar 2010 17:10:14 -0600</pubDate>
      </item>
      
      <item>
         <title>Thank you, OASPA</title>
          <description><![CDATA[<p>OASPA is starting to get its act together, <a href="http://oaspa.org/blog/2010/03/19/oaspa-assessment-of-new-applications-and-complaints-procedures/">posting a concise summary of its membership procedures</a> and making <a href="http://www.oaspa.org/membership.procedures.php">a new procedure for complaints</a> relevant to the quality measures OASPA wishes to maintain among its members.</p>

<p>I think OASPA is right not to offer to police every OA journal in existence. There isn't enough money in the <em>world</em>. It's also a clever stance that invites additional membership.</p>

<p>It's not perfect, however. OASPA had a choice to make between complete transparency&#8212;of accusers, of accused, of the process&#8212;and the sort of hush-hush under-wraps procedures that invite elevated eyebrows. Obviously, I think they made the wrong decision; they'll regret it most when some half-rabid academic sends in scores of complaints and cannot be reined in by public embarrassment.</p>

<p>I'm not entirely happy with "cannot investigate the circumstances surrounding individual editorial decisions, unless there is evidence of systematically flawed processes," either. "Usually will not" I can understand&#8212;again, resources are finite, and half-rabid academics tend to be on about individual editorial decisions!&#8212;but if this statement is their way of saying "won't investigate the Dove Medical Press matter," then I disapprove. One excruciatingly bad editorial decision should suffice to prompt OASPA to <em>look for</em> systematically flawed processes.</p>

<p>Still, it's a decent start. I daresay someone has sent the necessary email about Dove Medical Press, though again, transparency would be a virtue here. I look forward to seeing OASPA live up to its lofty intent.</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/thank_you_oaspa.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/thank_you_oaspa.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/thank_you_oaspa.php</guid>
         <category>Open Access</category>
         
         <pubDate>Mon, 22 Mar 2010 17:02:44 -0600</pubDate>
      </item>
      
      <item>
         <title>Productized what wired into what now?</title>
          <description><![CDATA[<p>First, a small warning: I am having an extremely crowded and busy week, so blogging here (even the catchup I need to do to the many excellent comments on the Battle of the Opens post) will suffer.</p>

<p>Something for folks to chew on in the meantime: can anybody explain to me <a href="http://www.cmio.net/index.php?option=com_articles&view=portal&id=publication:56:article:21274:himss-axolotl-debuts-elysium-open-access-platform&division=cmio#">what this tool (if it is a tool) actually does</a>? I clicked over thinking it might be a good thing to add to a tidbits post, but I confess myself wholly flummoxed by the jargon therein.</p>

<p>Any ideas, anyone? Especially anyone with a health-care background?</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/productized_what_wired_into_wh.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/productized_what_wired_into_wh.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/productized_what_wired_into_wh.php</guid>
         <category>Jargon</category>
         
         <pubDate>Thu, 18 Mar 2010 17:44:39 -0600</pubDate>
      </item>
      
      <item>
         <title>Societies and science</title>
          <description><![CDATA[<p>John Dupuis <a href="http://scienceblogs.com/confessions/2010/03/scholarly_societies_why_bother.php">asks some provocative questions</a>; I thought I'd take a stab at answering them, and I encourage fellow SciBlings to do likewise.</p>

<p>I quite agree with John when he says that the ferment over publishing models disguises a larger question, "the role of scholarly and professional societies in a changing publishing and social networking landscape." My own history with professional societies, I think, bears this out nicely.</p>

<p>John asks first: <strong>What societies do you belong to?</strong></p>

<p>I belong to the <a href="http://asis.org/">American Society for Information Science and Technology</a>. I was a member of the American Library Association for a time as a library-school student, until <a href="http://cavlec.yarinareth.net/2005/06/04/why-ill-quit-ala/">unchallenged racist statements from its then-president Michael Gorman</a> made me <a href="http://cavlec.yarinareth.net/2006/01/10/what-could-ala-do/">reconsider ALA's value proposition</a>; I wound up dropping the membership.</p>

<p>I am also a member/supporter of the <a href="http://creativecommons.org/">Creative Commons</a> and the <a href="http://www.eff.org/">Electronic Frontier Foundation</a> (which reminds me that it's about time I kicked another donation over to the latter). These aren't scholarly or professional societies in the sense John means, <em>but</em> I invite you to consider two things. One is, of course, that professional societies are competing with advocacy groups like CC and EFF for my money, attention, and time. The second: an often-rehearsed refrain justifying joining ALA in particular is the lobbyists that ALA sponsors in Washington, and the other advocacy and education work that ALA does.</p>

<p>I'm not knocking that work. In fact, if I could donate to ALA's <a href="http://www.ala.org/ala/aboutala/offices/oitp/index.cfm">Office of Information Technology Policy</a> (makers of the highly useful <a href="http://www.librarycopyright.net/digitalslider/">Copyright Slider</a>, among other things) and be assured that every penny of my donation would go to OITP's work, I would gladly do that. I'm happy to support advocacy I believe in. I just want to do it without having to support ALA per se, which I <em>don't</em> particularly believe in as presently constituted.</p>

<p>Next question: <strong>What value do you get from your membership?</strong></p>

<p>For a while, I had a pretty good streak going of one ASIST-sponsored conference per year. That streak ended last year, but it's as likely as not to pick up again; of the major library and info-sci organizations, the likeliest one to sponsor a conference I'm interested in (and thus cut me a break on conference fees) is ASIST. (<a href="http://www.acm.org/">ACM</a> is competitive in this regard, but they lost any chance of hooking me when they <a href="http://mybiasedcoin.blogspot.com/2009/04/acm-does-not-support-open-access.html">played games with Harvard over its OA policy</a>. You can stop sending me marketing materials now, ACM. You lose.)</p>

<p>There is also professional-identity value in an ASIST membership. It's a signifier; it signals not only that I'm serious about my profession, but what elements of the profession I'm serious about. Not a few librarians belong to ALA and some of its subsidiary organizations for similar reasons.</p>

<p>Value I don't get from ASIST includes professional-networking value; I do just fine for myself on the interwebs. Because I'm not tenure-track, I also don't have service obligations required of me. If I did, ASIST would unquestionably be the outlet for my labor. Again, the need to demonstrate national-level service is a motivation for many academic librarians who <em>are</em> tenure-track.</p>

<p>I'm also not particularly invested in ASIST's publications. JASIST contains eggheadery on a level I simply can't rise to, and the Bulletin isn't in my experience terribly interesting.</p>

<p>Third question: <strong>Is how you're thinking about your membership and the society's role in your professional life changing?</strong></p>

<p>Not noticeably, but I haven't been in the profession all that long, so it hasn't had much <em>time</em> to change, has it? I will say that I expect personal value out of my ASIST membership that I <em>don't</em> expect from CC and EFF. All I expect CC and EFF to do is keep on keepin' on with their missions, without wasting money (which they don't) or creating huge mission-unrelated scandals (which they haven't). At such time as the signifier value of an ASIST membership drops significantly for me, that membership may be in trouble.</p>

<p>Does this mean that scholarly/professional societies need to think harder about what they <em>do</em> instead of what they <em>are</em>? Quite possibly. Instead of <em>esse quam videri</em> (yes, I grew up in North Carolina), <em>facere quam esse</em>. I'm happier to throw money at <em>doing</em> than at <em>being</em>.</p>

<p>John saves the best for last: <strong>Do you think societies should be in the scholarly publishing business?</strong></p>

<p>Oof. That's a loaded question, because it's <em>different</em> from the question <strong>should scholarly societies publish journals?</strong> I am <a href="http://cavlec.yarinareth.net/2006/09/25/unyielding-opposition/">on record as saying</a> that societies have no particular right to fund their non-publishing activities from their publishing activities at the expense of library budgets. I still believe that.</p>

<p>Still, scholarly societies are in a good place to mobilize much of the labor that underpins journal publishing. The authors, peer reviewers, and acquisitions editors pretty much come to them! (Per <a href="http://www.dlib.org/dlib/march10/king/03king.html">this just-out D-Lib editorial</a>, that's 80% of the total labor cost of journal publishing anyway. Admittedly, that's a bit of a red herring, because all the shouting is really over the other 20%; there are other eyebrow-raisers in that editorial, but let that go for now.) It would be a shame to lose that, and my sense is that online networking cannot presently replace it because of the low participation in online networking by academia generally (with exceptions, of course).</p>

<p>However, I also believe that any journal-publishing operation needs to operate responsibly. In the present environment, <em>it is irresponsible</em> not to use the Internet to reach the widest possible audience. (There are exceptions, but they are <em>vanishingly</em> few.) <em>It is irresponsible</em> to withhold uncompensated knowledge from emerging nations, from non-profit organizations, from practitioners, from governments, from <em>anyone</em> who could benefit from it but cannot pay out-of-pocket for it and does not (for whatever reason) have a proxy such as a library available. <em>It is irresponsible</em> to operate exclusively in the digital world without strong preservation plans in place. <em>It is irresponsible</em> to charge fees or to allow one's publishing partners to charge fees (no matter the business model; this goes for author-side fees as well as subscription charges) that wildly exceed the true out-of-pocket costs of publication.</p>

<p>A <em>whacking</em> lot of society publishers are flagrantly irresponsible by the above criteria. Should they be publishing? I'll say "no." Not until they can get their heads back on straight. If that means they fold, because they put all their revenue eggs in the subscription-journal basket&#8212;I'm not unhappy with that outcome. Whatever they did that is still necessary will resurface; <em>that</em> I believe.</p>

<p>Hope these are the kinds of answers you're interested in, John.</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/societies_and_science.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/societies_and_science.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/societies_and_science.php</guid>
         <category>Praxis</category>
         
         <pubDate>Wed, 17 Mar 2010 16:47:33 -0600</pubDate>
      </item>
      
      <item>
         <title>Battle of the Opens</title>
          <description><![CDATA[<p>I'm committed to a lot of different kinds of "open." This means that I can and do engage in tremendous acts of hair-splitting and pilpul with regard to them. "Gratis" versus "libre" open access? Free-speech versus free-beer software code? I'm your librarian; let's sit down and have that discussion.</p>

<p>Unfortunately, out there in the wild I find a tremendous amount of misunderstanding about various flavors of open, sometimes coming from otherwise perfectly respectable communications outlets. (Pro tip: If you're not completely sure you understand, please find someone to ask. A librarian is a good start!) </p>

<p>Make no mistake, getting these things wrong sometimes does <em>serious harm</em>. The open-access movement exhausts itself contending with the same old misunderstandings over and over again, and I'm sure we're not alone in that.</p>

<p>So here&#8212;free, gratis, libre, and open&#8212;is a brief, simplistic guide to several flavors of open, organized around the following questions:</p>

<ul><li><strong>What</strong> is the target of this movement? What is being made open? As compared to what?</li>
<li>What <strong>legal regimes</strong> are implicated?</li>
<li><strong>How</strong> does openness happen? What are the major variants of open works of this type?</li></ul>

<p>Onward. We'll start with:</p>

<h3>Open source</h3>

<p><strong>What is being made open?</strong> Software, specifically its human-readable "source code." Software that is not open-source is usually distributed solely in non-human-readable "binary" form, and (as copyrighted expression) cannot legally be reverse-engineered or changed.</p>

<p><strong>What legal regimes are implicated?</strong> Copyright, mostly, though patents sometimes rear their ugly heads. The legal tools are copyright licenses specific to source code, such as the <a href="http://www.gnu.org/copyleft/gpl.html">GPL</a> and <a href="http://www.opensource.org/licenses/bsd-license.php">BSD license</a>.</p>

<p><strong>How does openness happen?</strong> Programmers place the source code they have written on the web, associating an open-source license with it. Other programmers are then able to read, use, and change the code. As open-source projects grow, they may have hundreds or thousands of programmers working on the code.</p>

<p>One of the two major ideological variants in the open-source world is the "free software" movement, which holds that opening source code is insufficient without ensuring that those who build upon open source code also make their code open (except when they are using it only privately). This movement produced the GPL. The "open-source software" movement holds that open code can and should be employed in proprietary, closed-source projects, and so tends to prefer licenses like the BSD license, which does not require open release of derivative code.</p>

<h3>Open standards</h3>

<p><strong>What is being made open?</strong> Specifications for how to accomplish particular tasks or build particular (tangible or virtual) objects. Open standards cover everything from computer cables to metadata to the building blocks of websites.</p>

<p><strong>What legal regimes are implicated?</strong> Our old friends copyright and patent. Open standards generally want to be implementable without treading on royalty-requiring copyrighted or patented intellectual property.</p>

<p><strong>How does openness happen?</strong> Generally a "standards body" does the design and outreach work. This may be an ad-hoc collection of engineers (<a href="http://ietf.org/">IETF</a>),  a group of interested commercial and/or nonprofit entities surrounding a particular trade or technical phenomenon (<a href="http://idpf.org/">IDPF</a> or <a href="http://w3.org/">W3C</a>), or a national or international organization whose specific remit is standards (<a href="http://www.iso.org/">ISO</a>, despite quibbles about having to buy their specifications' text).</p>

<h3>Open access</h3>

<p><strong>What is being made open?</strong> The academic literature: specifically, the peer-reviewed journal literature which is not written for royalties or any other direct monetary reward to its authors. (While open-access advocates happily cheer for open access to books and other research media, the different money-flows in these areas mean they are not a focus of the movement.) Open-access literature is in opposition to literature which is not available to be read unless a subscription, per-article, or other fee is paid by the reader or the reader's proxy (e.g. a library).</p>

<p><strong>What legal regimes are implicated?</strong> Copyright, again. Typical practice for the academic article is that its author(s) transfer their copyright in its entirety to the journal publisher, allowing the publisher to control reuse.</p>

<p><strong>How does openness happen?</strong> In two basic ways. Yes, <em>two</em>! One is the soi-disant "gold road," in which authors publish in journals that make their contents available on the Web immediately upon publication without charging reader-side fees. The other is the "green road," in which authors reserve or are granted by the publisher sufficient rights in their article to make some version of it (usually <em>not</em> the final typeset, copy-edited publisher's version) available openly online.</p>

<p>Another division can be drawn between "gratis" open access, in which articles are available freely to be read but require explicit permission for most reuse, and "libre" open access, in which articles are clearly licensed up-front for reuse, often with a Creative Commons license.</p>

<h3>Open educational resources</h3>

<p><strong>What is being made open?</strong> Many sorts of classroom materials, including syllabi, lecture audio/video, assignments, and instructional material such as self-contained web-based "learning objects."</p>

<p><strong>What legal regimes are implicated?</strong> Copyright and the related work-for-hire doctrine, that last because some educational institutions claim copyright in instructional materials created by instructors in the course of their regular job duties.</p>

<p><strong>How does openness happen?</strong> Typically, through institution-based "courseware" programs or learning-object repositories. Some instructors share educational material through consumer web applications such as SlideShare.</p>

<p>The open-textbook movement is worth mentioning here. Though it is logically affiliated with the OER movement, in practice it bears more resemblance to the open-access movement.</p>

<h3>Open (research) data</h3>

<p><strong>What is being made open?</strong> Data resulting from the research process, in a form less "cooked" than the graphs, tables, and charts in journal articles. ("Data" is a vague word, granted.) Ideally, sufficient description of the data and how they were obtained is included for the data to be verifiable and reusable.</p>

<p><strong>What legal regimes are implicated?</strong> In some countries, copyright. For data from industry, trade-secret law.</p>

<p><strong>How does openness happen?</strong> Researchers, with or without help from librarians and IT professionals, make their data open. Some journals and science funders are beginning to demand open data; others demand data-sustainability plans that align well with the open-data movement.</p>

<h3>Open (government) data</h3>

<p><strong>What is being made open?</strong> Information gathered by governments in the course of business: geographical information, demographic information, research data gathered by government agencies, sometimes records.</p>

<p><strong>What legal regimes are implicated?</strong> For pure data, none in the United States; data are not copyrightable. For other works, copyright, sometimes. Though works authored by (employees of) the US federal government are in the public domain, works authored by (employees of) other governments in the US can be copyrighted.</p>

<p><strong>How does openness happen?</strong> Usually, the government in question releases the data online. There is considerable stir and excitement at present over "linked (open) data," which means data expressed in such a way as to be easily and usefully combined with data from other sources.</p>

<h3>Open notebook science</h3>

<p><strong>What is being made open?</strong> The process and progress of a particular research project, analogous to placing a lab notebook on the Web for public view.</p>

<p><strong>What legal regimes are implicated?</strong> Copyright, insofar as making original expression available in tangible form (yes, the Internet counts as "tangible" for copyright purposes) immediately creates copyright in it. Patent, insofar as making a patentable invention available removes patentability (in the US), but also creates prior art such that subsequent patents can be challenged.</p>

<p><strong>How does openness happen?</strong> At present, researchers employ whatever tools come to hand, from wikis to Google Docs to FriendFeed to github, to document their research process on the Web as the research is happening. Some institutions are trying out "electronic lab notebooks" which could facilitate open notebook science if they are not kept behind firewalls, or if researchers have the option to move their workspaces into the open.</p>

<p>---</p>

<p>Any effort such as this will be nitpicked endlessly. That's what the comments are for, so go to it&#8212;but be warned, religious wars and diatribes will be ruthlessly deleted. Emacs and vi are both awful, I don't like Windows <em>or</em> Linux as a desktop environment, and progress in both the green and gold roads to OA makes me happy.</p>

<p><b>Edited to add</b>: I extend many thanks to the commenters on this post, and have revised it in light of their comments. Any remaining errors or infelicities are of course mine. I strongly recommend supplementing this post with <a href="http://sarahglassmeyer.com/?p=413">Sarah Glassmeyer's post on free law</a>.</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/battle_of_the_opens.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/battle_of_the_opens.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/battle_of_the_opens.php</guid>
         <category>Jargon</category>
         
         <pubDate>Mon, 15 Mar 2010 04:57:45 -0600</pubDate>
      </item>
      
      <item>
         <title>Profile: Dryad</title>
          <description><![CDATA[<p>We have a guestblogger today! At my request, Peggy Schaeffer kindly sent me the following introduction to <a href="http://www.datadryad.org/">Dryad</a>, which I reproduce as I received it (save for minor formatting details).</p>

<p>I will happily pass any questions in the comments on to Peggy for response.</p>

<p>----</p>

<p><a href="https://www.nescent.org/wg_dryad/Main_Page">Dryad</a> is a repository for data underlying scientific publications, with an initial focus on evolution, ecology, and related fields.  It's not an institutional repository, or one focused on only a single type of data -- it's designed for the multitudes of data underlying published articles that would otherwise be scattered ineffectively, hard to find, or lost. Dryad enables researchers to archive their data at the time of publication, dedicate it to the public domain, and get a citable DOI for it.  In so doing, Dryad promotes the discovery and reuse of data by others.</p>

<p>The Dryad repository model has these strengths:</p>

<ul><li>all data is associated with a published article (this collection policy provides a qualitative measure and enables links between journal articles and their data)</li>
<li>data archiving is facilitated at the point of publication, when authors' motivation to share is strongest and the data are at hand</li>
<li>Dryad is governed and supported by a growing Consortium of major international journals and societies (<a href="http://www.datadryad.org/repo/partners">see list here</a>)</li>
<li>partner journals support a <a href="http://datadryad.org/repo/jdap">Joint Data Archiving Policy</a> that requires data archiving at the point of publication </li>
<li>all types of data and formats are welcome; journals may specify standards appropriate for particular data types.</li>
<li>data submission is facilitated by Dryad's integration with the manuscript processing systems of its partner journals; authors publishing in these journals don't need to input bibliographic details</li>
<li>data submitted to Dryad will be also served to select specialized repositories (like GenBank), further reducing the burden on authors to submit data to multiple sites</li>
<li>all data is made freely available under the <a href="http://creativecommons.org/about/cc0">Creative Commons Zero waiver</a></li>
<li>data receive DataCite DOIs and receive an independent citation when they are reused</li>
<li>descriptive metadata will be automatically generated from the article and data content, allowing authors and curators to review & select proposed descriptors from multiple ontologies and thesauri.  For more details, see <a href="http://ils.unc.edu/mrc/hive/">the HIVE project page</a>. Don't miss <a href="http://www.youtube.com/watch?v=BmAXxv-8q9U">the nifty video</a>.</li>
<li>Dryad plans to expose its contents through a variety of web standards, to enable metadata harvesting, remote queries, and linked-data applications.</li></ul>

<p>Dryad allows future investigators to validate published findings, explore new analysis methodologies, repurpose the data for research questions unanticipated by the original authors, and perform synthetic studies such as formal meta-analyses.</p>

<p>The repository is being developed at <a href="http://www.nescent.org/">the National Center for Evolutionary Synthesis, or NESCent</a>, in collaboration with the University of North Carolina at Chapel Hill School of Information Science, and the Dryad Consortium of partner journals.  A number of the partner journals have recently announced their intention to require data deposition in a publicly available archive as a condition of publication: </p>

<ul><li>Whitlock, M. C., M. A. McPeek, M. D. Rausher, L. Rieseberg, and A. J. Moore. 2010. Data Archiving. American Naturalist. 175:145-146, doi:10.1086/650340</li>
<li>Rieseberg, L., T. Vines, and N. Kane. Editorial and retrospective 2010. Molecular Ecology. 19:1-22, doi:10.1111/j.1365-294X.2009.04450.x</li>
<li>Rausher, M. D., M. A. McPeek, A. J. Moore, L. Rieseberg, and M. C. Whitlock. Data Archiving. Evolution. doi:10.1111/j.1558-5646.2009.00940.x</li>
<li>Allen J. Moore, Mark A. McPeek, Mark D. Rausher, Loren Rieseberg, Michael C. Whitlock.  The need for archiving data in evolutionary biology. Journal of Evolutionary Biology 2010 Published Online: Feb 9 2010. doi: 10.1111/j.1420-9101.2010.01937.x</li>
<li>Uyenoyama, M. K. (2010). MBE editor's report. Mol Biol Evol, 27(3):742-743.  doi: 10.1093/molbev/msp22</li></ul>

<p>Dryad currently has a staff of about 7 (curator, repository architect, programmer, communications officer, etc.) led by</p>

<ul><li>Project Director: <a href="http://visionlab.bio.unc.edu/">Todd Vision</a>, Associate Professor of Biology,  University of North Carolina at Chapel Hill, and Associate Director for Informatics at  NESCent</li>
<li><a href="http://ils.unc.edu/~janeg/">Jane Greenberg</a>, Professor and Director, SILS Metadata Research Center School of Information and Library Science, University of North Carolina at Chapel Hill.</li></ul>

<p>Funding comes from the National Science Foundation, and the Institute of Museum and Library Services (IMLS) has funded HIVE (<a href="http://ils.unc.edu/mrc/hive/">Helping Interdisciplinary Vocabulary Engineering</a>) a 3-year project that will enhance Dryad's metadata.</p>

<p>For more detailed information, including upcoming features, details of the metadata format, and other development plans, please see <a href="http://www.datadryad.org">the Dryad website</a> and <a href="https://www.nescent.org/wg_dryad/Main_Page">the team Wiki</a>. Also, you can follow Dryad's activities on the Dryad blog: http://blog.datadryad.org/ and Twitter: <a href="http://twitter.com/datadryad">@datadryad</a>.</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/profile_dryad.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/profile_dryad.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/profile_dryad.php</guid>
         <category>Miscellanea</category>
         
         <pubDate>Thu, 11 Mar 2010 16:03:56 -0600</pubDate>
      </item>
      
      <item>
         <title>Sensitive data, linked data, and the &quot;reidentification&quot; phenomenon</title>
          <description><![CDATA[<p>One of the truisms in data curation is "well, of <em>course</em> we don't let sensitive data out into the wild woolly world." We hold sensitive data internally. If we must let it out, we anonymize it; sometimes we anonymize it just on general principles. We're not as dumb as <a href="http://www.businessinsider.com/warning-google-buzz-has-a-huge-privacy-flaw-2010-2">the Google engineers</a>, after all.</p>

<p>Only it turns out that data anonymization can be frighteningly easy to reverse-engineer. We've had some high-profile examples, such as <a href="http://news.cnet.com/2100-1030_3-6103098.html">the AOL search-data fiasco</a> and <a href="http://www.freedom-to-tinker.com/blog/paul/netflixs-impending-still-avoidable-multi-million-dollar-privacy-blunder">the ongoing brouhaha over Netflix data</a>. <a href="http://papers.ssrn.com/sol3/papers.cfm?abstract_id=1450006">Paul Ohm's working paper</a> on the topic is a great way to get up to speed.</p>

<p>We librarians are fairly dogmatic about this sort of thing, owing to our professional-ethics commitment to your freedom to read. We wipe your checkout record clean after you turn your items back in. We do keep passive-voice usage records on our materials: "this book has been checked out X times since Y date." But that's it. (And no, we don't keep track of when you visit the library, so it's not possible to connect a formerly checked-out book with you based on the date of checkout.)</p>

<p>This long-standing design decision is being challenged on social-media grounds; it's hard to build Web 2.0-ish applications around your library behavior if we don't keep records of your library behavior! I used to be on the Web 2.0 side of this particular controversy, but as I've been reading about reidentification, my mind has changed. Information about which local public library one goes to isn't precisely "zip code," but it's awfully, awfully close.</p>

<p>Anyway, the application to human-subjects data of all stripes is, I hope, obvious. It's not as simple as anonymizing data; even aggregating it and only permitting queries may not solve the problem. Certain data breakdowns (e.g. from survey data) may be problematic.</p>

<p>Taking heed of the problem is the first step to solving it&#8212;but only the first. The sooner we have data-release guidelines that take reidentification into account, the happier I will feel about open data in the social sciences and medicine.</p>

<p>Incidentally, are you as sanguine about governments providing "<a href="http://www.w3.org/DesignIssues/GovData.html">linked data</a>" as you were? Because I'm not.</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/sensitive_data_linked_data_and.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/sensitive_data_linked_data_and.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/sensitive_data_linked_data_and.php</guid>
         <category>Praxis</category>
         
         <pubDate>Wed, 10 Mar 2010 16:12:42 -0600</pubDate>
      </item>
      
      <item>
         <title>RFC: Repository platform comparison</title>
          <description><![CDATA[<p>I interrupt your regularly-scheduled blog to ask for some help... comments closed on this post so that you'll comment where it'll do the most good.</p>

<p>---</p>

<p>Apologies for duplication, and please forward/repost as appropriate...</p>

<p>We are working on comparing four digital-repository software packages (DSpace, ePrints, Fedora, and Zentity) in hopes of helping libraries and other institutions select the most appropriate software for their requirements. Read more about our project at<br />
<a href="http://blogs.lib.purdue.edu/rep/ ">http://blogs.lib.purdue.edu/rep/</a>.</p>

<p>We invite anyone who has recently embarked upon planning for a digital repository to tell us what criteria were used to select a software package. (We are not interested in hosted repository services at this time, only repositories managed in-house.) Your input will inform our testing criteria.</p>

<p>Please leave your comments at <a href="http://blogs.lib.purdue.edu/rep/2010/02/25/a-comparative-analysis-of-institutional-repository-software/">http://blogs.lib.purdue.edu/rep/2010/02/25/a-comparative-analysis-of-institutional-repository-software/</a>. Pointers to public planning documents or lists of criteria are equally welcome. Though we will read all comments submitted, we plan to respond only by private email so as not to bias the public comment-stream.</p>

<p>We very much appreciate your assistance!</p>

<p>Siddharth Singh, Michael Witt, and Dorothea Salo</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/rfc_repository_platform_compar.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/rfc_repository_platform_compar.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/rfc_repository_platform_compar.php</guid>
         <category>Miscellanea</category>
         
         <pubDate>Tue, 09 Mar 2010 11:04:14 -0600</pubDate>
      </item>
      
      <item>
         <title>Tidbits, 5 March 2010</title>
          <description><![CDATA[<p>I'm in Urbana-Champaign this weekend to teach an in-person day for my online collection-development class. I'm looking forward to it; every time I teach I am reminded that students are smarter than I am.</p>

<p>For now, tidbits!</p>

<ul>
<li>As world plus dog probably knows already, <a href="http://www.economist.com/specialreports/displaystory.cfm?story_id=15557443">The Economist tackled the data deluge</a>.</li>
<li>Adam Christensen gives us the modest, unassuming <a href="http://asmarterplanet.com/blog/2010/03/data-the-foundation-for-everything-on-an-intelligent-interconnected-instrumented-planet.html">Data. The foundation for everything on an intelligent, interconnected, instrumented planet.</a></li>
<li>Rethinking scholarly communication from the ground up: SciBling John Dupuis asks <a href="http://scienceblogs.com/confessions/2010/03/are_computing_journals.php">Are computing journals too slow?</a> and Dan Cohen muses about <a href="http://www.dancohen.org/2010/03/05/the-social-contract-of-scholarly-publishing/">how best to deconstruct the humanities' reverence for the print codex</a>, while Craig Mod <a href="http://craigmod.com/journal/ipad_and_books/">brilliantly deconstructs book design</a> in an iPad world.</li>
<li>So-called "digital natives" have digital histories; the Library of Congress asks <a href="http://digitalpreservation.gov/videos/students10/index.html">whether and what they think about preserving them</a>. (For more on personal digital preservation, I strongly recommend Microsoft Research's Cathy Marshall. Her <a href="http://www.dlib.org/dlib/march08/marshall/03marshall-pt1.html">two D-Lib</a> <a href="http://www.dlib.org/dlib/march08/marshall/03marshall-pt2.html">articles</a> are wonderful; also keep an eye on <a href="http://code4lib.org/conference/2010/marshall">her recent presentation at the code4lib conference</a>, which there should shortly be video of.)</li>
<li>The city of Vancouver <a href="http://straight.com/article-292923/vancouver/citys-history-safe-thanks-digital-archiving">is taking digital archiving seriously</a>. No "put floppy disks in the fridge" here (no, seriously, I've seen that hailed as innovative archival practice!). I like what I see of <a href="http://archivematica.org/">Archivematica</a>.</li>
<li>Stefano Costa hopes to <a href="http://blog.okfn.org/2010/02/25/open-data-in-archaeology/">make data in archaeology open</a>. While there are serious and legitimate concerns about making location data on some finds and digs public&#8212;my father the anthropologist used to call himself a "grave-robber and junk-picker" in jest, but there are real robbers out there&#8212;in the main, archaeology data is a great target for open.</li>
<li>Sarah Askew once again explains <a href="http://sarahaskew.net/2010/03/01/on-software-in-astronomy/">why the software turned loose on data</a> should be kept and scrutinized, with astronomy as her case study. Good insight into why "one software suite fits all" doesn't work, which should give some web4science developers pause.</li>
<li>Harvard's School of Engineering and Applied Science <a href="http://www.seas.harvard.edu/topics/uncovering-open-access">interviews Stuart Shieber</a>, in a treatment of open access refreshingly free of hyperbole on one side and panic on the other.</li></ul>

<p>As always, if there's a link I should see, comment here or tag it "trogool" on del.icio.us. Thanks!</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/tidbits_5_march_2010.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/tidbits_5_march_2010.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/tidbits_5_march_2010.php</guid>
         <category>Tidbits</category>
         
         <pubDate>Fri, 05 Mar 2010 16:02:23 -0600</pubDate>
      </item>
      
      <item>
         <title>Cognitive dissonance</title>
          <description><![CDATA[<p>One of the latest institutional open-access policies comes from <a href="http://osc.hul.harvard.edu/OpenAccess/policytexts.php#hbs">Harvard Business School</a> (hat tip to <a href="http://blogs.law.harvard.edu/pamphlet/2010/02/28/harvard-business-school-approves-open-access-policy/">Stuart Shieber</a>).</p>

<p>This is the same school that <a href="http://dltj.org/article/ebsco-hbp/">plays horrendous anti-library, anti-education games</a> with their flagship <i>Harvard Business Review</i>. </p>

<p>My head hurts.</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/cognitive_dissonance.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/cognitive_dissonance.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/cognitive_dissonance.php</guid>
         <category>Open Access</category>
         
         <pubDate>Tue, 02 Mar 2010 15:48:15 -0600</pubDate>
      </item>
      
      <item>
         <title>Grey literature considered harmful?</title>
          <description><![CDATA[<p>So the United Nations' Intergovernmental Panel on Climate Change is mired in a <a href="http://www.guardian.co.uk/environment/2010/feb/08/climate-scientists-melting-glaciers">rapidly heating controversy</a> over a report that apparently let some dubious information slip through the cracks. Here's the money quote:</p>

<blockquote><p>The discovery of the glaciers mistake has focused attention on the IPCC's use of so-called grey literature: reports that do not appear in conventional scientific journals, and are instead drawn from sources such as campaign groups, companies and student theses. The IPCC's rules allow such grey literature, but many people have been surprised at the scale of its inclusion.</p></blockquote>

<p>Oh my, oh dear. Let's take that apart a bit at a time, see what's likely going on, and suggest some remedies.</p>

<p>That's a fairish definition of grey literature, but what it leaves out is that the importance and acceptance of certain genres of grey literature varies considerably by discipline (although I think most accept dissertations as quality, citable work, assuming they're recent enough). For example, quite a few social-science disciplines have a flourishing working-papers culture; the expectation is that many (though probably not all) working papers will go through the peer-review wringer at some juncture, but it's important both for authors and readers to circulate the ideas fast.</p>

<p>I don't know of any literature on this specific point (which doesn't mean it doesn't exist, just that I haven't looked), but from my admittedly anecdotal experience, one factor that seems to create a grey-literature&#8211;friendly culture is a desire for influence beyond the academy: influence on practitioners, policymakers, non-profits, or the public generally. The IR I run has several grey-literature collections predicated on <em>precisely</em> this, and I don't think it's any coincidence that a strong vein of public-policy discourse runs through the research blogosphere generally and ScienceBlogs specifically.</p>

<p>Why is grey literature a better choice than the peer-reviewed journal literature in this context? Simply because it's much more <em>accessible</em>. Practitioners, policymakers, and particularly nonprofits have limited if any access to toll-access peer-reviewed journals. If you want them to see it and use it, placing it in a toll-access journal is tantamount to shredding it.</p>

<p>Now go back and read the paragraph from the <i>Guardian</i> again. "Surprised at the scale of [grey literature's] inclusion." Well, I'm not. I think a mixture of two factors is an odds-on bet in this clash of the literatures: a grey-literature&#8211;friendly discipline working with a grey-literature&#8211;unfriendly one, <em>and</em> the simple matter of access I just explained.</p>

<p>Climate science strikes me as grey-literature&#8211;unfriendly for cogent reasons. One is that the peer-review process (one hopes!) does yeoman's work eliminating bad science and bad data in a field where <em>plenty</em> of people grinding axes are happy to pollute the discourse with bad science and bad data. Another is that climate scientists <em>need</em> to limit their discourse population somewhat, or they'll be overwhelmed with axe-grinding bizarrerie from outside the field. I would guess it doesn't trouble them much that their literature isn't open-access, and it may even please them, by way of getting on with their work without having to stop every five seconds to deal with some ignorant lout who Googled their latest paper.</p>

<p>But social scientists, they love their grey-lit&#8212;and thus the clash of cultures. Neither the social scientists nor the climate scientists realized that they didn't have the same standards for reliable, citable previous work. The social scientists applied their normal quality standards and search techniques, not realizing that the climate scientists had different ones, not even <em>thinking to ask</em>. What the social scientists found on the open Web about climate science had a high probability of being junk, given that peer-reviewed climate science is almost entirely toll-access. (I found <a href="http://www.doaj.org/doaj?func=subject&cpid=86">only 22 journals</a> in DOAJ, two of which are Bentham Open and thus liable to be junk; on the green-OA side, I certainly haven't heard that climate scientists are heavy self-archivers.) I expect similar clashes play out in quite a few interdisciplinary collaborations.</p>

<p>So what do we take away from this?</p>

<ul><li>Scientists, if you want people outside your discipline to read your work in preference to whatever they can find on the open Web, <em>make it open access</em>; that may not be sufficient, of course, but it's absolutely necessary. However, all of us need to take into account that the knee-jerk call for everything to be OA removes a shield from a significant set of researchers, those poor hot-spotlighted souls whose professional lives OA would make a living hell.</li>
<li>If you're troubled by bad science and bad data on the open Web, the answer is to <em>make better science and better data open access</em>. There will always be bad science and bad data, and it will seek the path of least resistance to earn the most eyeballs possible. Locking up the broccoli only increases candy consumption.</li>
<li>Interdisciplinary collaborations had better start by establishing common ground for citable literature. Disciplinary cultures differ, and for legitimate reasons; practices do not necessarily transfer.</li>
<li>Post-publication review and commentary don't <em>just</em> offer an alternative to peer review. They may help keep junk science from taking over the wider discourse, to the benefit of all.</li></ul>

<p>I would never recommend that libraries discontinue collecting grey literature, especially in digital form. For one thing, some grey literature gets much more scrutiny than the average peer-reviewed article: consider how often and by how many people dissertations are reviewed, and how many rewrites they get! For another, I see no reason collection practices should be governed by the most restrictive disciplinary cultures out there.</p>

<p>I do acknowledge grey literature for what it is, however, and I hope this culture-clash sparks discussion about how best to manage differing citation practices when collaborating.</p> <a href="http://scienceblogs.com/bookoftrogool/2010/03/grey_literature_considered_har.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/03/grey_literature_considered_har.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/03/grey_literature_considered_har.php</guid>
         <category>Praxis</category>
         
         <pubDate>Mon, 01 Mar 2010 16:26:35 -0600</pubDate>
      </item>
      
      <item>
         <title>Tidbits, 26 February 2010</title>
          <description><![CDATA[<p>It's Friday! Snack on some tidbits.</p>

<ul><li>In the "didn't anyone teach you to show your work in grade school?" department, we have <a href="http://www.scoop.co.nz/stories/SC1002/S00004.htm">NIWA unable to justify official temperature record</a>, as well as <a href="http://blogs.nature.com/nautilus/2010/02/nature_neuroscience_on_gaps_in.html">the radical notion of using actual data</a> to gauge the effectiveness of review boards in stopping unethical research.</li>
<li>In the "open is not a panacea" department, we have Nat Torkington <a href="http://radar.oreilly.com/2010/02/rethinking-open-data.html">rethinking open data</a>, or at least its funding models (hat tip to Trevor Muñoz), and <a href="http://clarionproject.wordpress.com/2010/01/29/principal-investigators-opinions-on-open-data/">JISC's Clarion project trying to convince principal investigators</a> that sharing data is a useful thing to do.</li>
<li>In the "let's kill all the lawyers" department, we have <a href="http://arstechnica.com/science/news/2010/02/dna-data-sharing-a-privacy-conundrum.ars?utm_source=rss">troubling privacy questions about the sharing of personal genome data</a>, and on a happier note, the wonderful <a href="http://pantonprinciples.org/">Panton Principles</a> for making data properly open and reusable once the decision to share them has been made.</li>
<li>In the first-principles department, the redoubtable Carole Palmer tells us that <a href="http://www.physorg.com/news186233278.html">data need to be curated</a>. AAAS wonders <a href="http://blogs.nature.com/news/blog/2010/02/aaas_2010_data_techs_needed_1.html">who's going to do the work</a>, and JISC comes up with <a href="http://www.jiscinfonet.ac.uk/research">a good-practice guide</a> for those willing to dive in.</li>
<li>In the tools-and-toys department, we have Stuart Lewis building <a href="http://blog.stuartlewis.com/2010/02/03/easydeposit-sword-deposit-tool-creator/">a SWORD library in PHP</a> aimed at making quick, easy, even one-off repository-deposit tools. I am all in favor!</li></ul>

<p>As always, tag a delicious link with "trogool" or leave a comment here if you have something tidbit-worthy. Thanks!</p> <a href="http://scienceblogs.com/bookoftrogool/2010/02/tidbits_26_february_2010.php#commentsArea">Read the comments on this post...</a>]]></description>
         <link>http://scienceblogs.com/bookoftrogool/2010/02/tidbits_26_february_2010.php</link>
         <guid>http://scienceblogs.com/bookoftrogool/2010/02/tidbits_26_february_2010.php</guid>
         <category>Tidbits</category>
         
         <pubDate>Fri, 26 Feb 2010 16:39:31 -0600</pubDate>
      </item>
      
   </channel>
</rss>