Monday, June 30, 2008

Search is a combination what you know and who you know

In my first years of academic librarianship, it struck me as really odd that I would have faculty call me up and ask me to teach their undergraduate students how to use the library indexes. Some of them would even confess that they really didn't know how to use these tools themselves.

In time, I came to realize two crucial points. First, faculty keep up with their research by regularly reading their personal journal subscriptions and keeping track of what the most important researchers in their field are doing. When they do use library-provided indexes, they tend to search by author and by following citations. Google Scholar recognizes this; when you search for Hamlet you get an article written by D. Hamlet as the first result returned.

The second point is that the process of learning who the key figures are in a particular field is part of the transformation from beginner to expert. The trouble is, our existing indexes are designed for librarians - not for students who are making this transition from undergraduate to graduate.

Rarely are students told that research is a reiterative process and that it is often necessary to perform a number of literature searches for a paper as one discovers gaps in coverage or realizes that there is something worth dwelling into more detail. But as seasoned librarians and educators know, even if you tell them this hard won nugget of wisdom, most students will still just grab the first set of articles that match their topic (which can be quite random especially if the database sorts results by date) and then try to mash these citations into their paper, often the night before its due.

So then how can we point these students to the best articles on a particular topic ?

One solution is to point beginning researches to sources of review articles. At the reference desk, I try to check Annual Reviews whenever I get a student who can't narrow down their research question to anything beyond their topic. You know the ones: "I need articles on anorexia." Another solution would be for us to take matters in our own hands. Here's my idea: we should buy the ten most popular textbooks for undergraduates in a particular domain, like Biology. Then we should make note of all the citations in the texts and add them into a searchable database.

These two examples still rely on the work of editors and experts to select the research that made a difference and to put such work in context. TopCited tools from Scopus and Web of Knowledge have been developed to automatically quantify the value of articles and researchers by the rate and number of times papers have been cited.

I believe that the adoption of social networking software into the academic sphere will eventually provide another means of establishing who are the experts in a particular field and what pieces of writing and research are the ones that have made the most impact. It very well might replace the traditional way we search and research and keep up with our fields.

Reading blogs is a case in point. You read library blogs on a regular basis rather than searching for the word 'library' in Google on a regular basis, don't you?

Thursday, June 26, 2008

Boolean NOT

I have been thinking about social software and libraries and I suspect that the result of this thinking might be a rash of posts not unlike a bunch I wrote when I couldn't stop thinking about Wikipedia.

But before I can write about how I think social software may affect search and research, I feel its necessary to clear some cognitive space first. So before proceeding, I want to make something very clear: LIBRARIANS HAVE TO STOP TEACHING BOOLEAN SEARCHING.

I know this is scary stuff for many academic librarians who I suspect would be at a loss of what to teach if you took away their ((ANDs) and (ORs)). But teaching about Boolean logic has become a crutch and its time to throw it away. I would go so far to say that banning the word Boolean is the number one way to improve information literacy practice in libraries.

You see, most people don't use Boolean searching. And they are still healthy, happy, and successful people. Sure, they may not be the most searching in the most efficient way but their searching is still effective. Users want to apply their energies to their results and not to search grammar.

post-search filtering

It is far more efficient to create searching interfaces that try to address these behaviour patterns than try to educate everyone of the masses on "the right way" to search.

We are at a state where the bibliographic databases have become so large that even some of the worst, poorly constructed search terms now bring relevent hits among the debris. Unfortunately, one response I've seen first hand is a library assignment in which the student is forced to use boolean logic by requiring that only 10 hits or less be returned from a search string. Any research assignment does not resemble how a real person conducts research is a poor one, to say the least.

Instead teaching that AND means 'and' and 'OR' means 'or', we should try to teach something beyond the mechanics of search. Instead of teaching about Boolean, we should instruct on the importance of language, on the nature of the publication cycle, and of the research process.

And it goes without saying that we have to get the word Boolean out of our library catalogues.

Tuesday, June 24, 2008

If there is no porn in your library catalogue, then its not working

"Based on my Tripod experience, I’d offer the hypothesis that any sufficiently advanced read/write technology will get used for two purposes: pornography and activism. Porn is a weak test for the success of participatory media - it’s like tapping a mike and asking, “Is it on?” If you’re not getting porn in your system, it doesn’t work. Activism is a stronger test - if activists are using your tools, it’s a pretty good indication that your tools are useful and usable." [The Cute Cat Theory Talk at ETech]

Friday, June 20, 2008

reddit.com: Who else is sick of sites hosting research papers that show all their content to Google so it gets indexed, but when people visit, they want you to pay exorbitant fees?


reddit.com: Who else is sick of sites hosting research papers that show all their content to Google so it gets indexed, but when people visit, they want you to pay exorbitant fees?

Some of the comments are quite entertaining:

When did "not getting something for free" become a problem?
When people starting trying to trick users into reading their reports.
If it's not open and free, google wouldn't index it. If google doesn't index it, nobody finds it. Therefore, if somebody found it in google, by definition it should be open and free.

Also interesting is that the discussion immediately proceeds into the realm of IP spoofing and cloaking. Check your library for holdings? Nah - pretend to be Googlebot to get a free pass!

Wednesday, June 18, 2008

My definition of a boombatic library blog

I’ve also heard librarians discussing the same concept in the library community. In library-related articles, blog posts, and presentations I’ve attended and/or read this past year, the presenters/writers have been saying that Web 2.0 and Library 2.0 are all about starting conversations, building community, and telling our stories. But the writer/presenter tends to skip over what I think is the most important part - they never explain how to do it. [David Lee King]

David Lee King makes the case that a library (blog) should invite participation or, if possible, be actionable. While I find this ambition admirable, I don't entire agree with it. I subscribe to the "you use a library, you don't make friends with it" school of library twotopianism.

This is my suggestion on how to do it: use your library (blog) to connect readers with readings. Or, more generally, make every post a connection between your library and the world they live in.

Here are the specifics. First, think about how your readers decide what book or article to read next. Chances are, they don't use your library or its OPAC to decide what to read. Your readers rely on the recommendations of their friends and/or from their reading of newspapers, magazines, journals and even blogs.

So, as you read your magazines or your local or national newspaper, online or otherwise, make note of the books, government reports, and scientific papers mentioned and see if its available through your library. Heck, you can do the same for books that you've head mentioned on the radio. Even if the government report scientific paper you read about is free online, your readers will appreciate the direct link to it as it will save *them* the aggravation of having to navigate the Statistics Canada website, for example.

The benefits of such blogging is multifold. Not only does it help the reader, this regular reading helps the librarian in keeping up with the world. Checking to see if you have the books that are currently being reviewed in the press is a great way to measure how successful your library's approval plan is working (or not). Furthermore, the librarian gains a stronger understanding of their strengths of the library's collection through what I'd like to call, curiousity-based collection development.

Here's an example of it: some months ago I was curious if the library at MPOW had The Pillars of the Earth, a recent book recommended by Oprah. At the time, when I checked what readers had said about the book on Amazon, it told me that folks who bought that book also were interested in Cathedral: The story of its construction by David Macaulay. Following that link, I learned that this was the same man responsible for The Way Things Work (which we had in our library's education collection) and The New Way Things Work (which we did not).

I'm not enirely that fond of the notion that the library blog is a tool of conversation between a library and its users - at least, not in the sense of how The Cluetrain Manifesto uses it. Reading is the real conversation.

Monday, June 16, 2008

Server on a stick

Occasionally, I feel bad that my computer skills have large stagnated.

My excuse, which I would use to comfort myself, was that I didn't have ready access to a server and thus, no ready access to advanced computer languages or advanced applications.

Now, this wasn't much of an excuse: there are web hosting companies that provide access to languages such as PHP and I know computer folks who are generous with their help. But still - access to a server would require effort. And as long as there was that barrier of effort required, I had my excuse for my ignorance.

That is, until I realized that not only can you run applications like WorkPress and Django on a USB stick.

Wednesday, June 11, 2008

Librarianship and journalism - are they connected?

2008jun10. The 35 articles of impeachment introduced by Dennis Kucinich yesterday. Not covered by NYT FOX CBS ABC CBS CNN etc mmm big MSM luv you bet

Anyone who watches The Colbert Report knows that the MSM is both high comedy and high tragedy. The Daily Show is a daily reminder that 'the news' is just directed attention to events construed by a small group of editors. One could say that it has always been this way. Unless news outlets realize that investigative journalism is their critical function, their days of printing money may be over.

I am interested in libraries, blogging, the media, and open government and I think I've thought of one instance where they all intersect. We clearly cannot trust MSM to bring attention to the day's important events and discoveries. That is why it is essential that an archive of the raw material of news - the press releases, the government reports, the committee agendas and minutes, the legislature, the video feeds, the full-text of budgets - should be made available to all for word-of-mouth reporting.

And just like the news media, the library's critical function of being a subsidized source of organized information is also being eroded by the Internet. It used to be that if you wanted to find a newspaper older than a couple of days, you had to go to your public library to get it. We could do worse that provide our citizens with the information that they can't get from the MSM.

Wednesday, May 28, 2008

If they say that is "going green" then I say "goodbye polar bears"

Whenever I read about a library conference that goes green by banning presentation handouts, I don't know whether to laugh or to cry or to release tanks of methane into the auditoriums.

Honestly, how could grown adults act so self-congratulatory over such a insanely pathetic and inane action when compared to the scale of the environmental problems that are at hand.

Oh yeah, we just happen to be a profession that BUYS ENTIRE BOOKS for our communities. Stop us before we kill again!

Tuesday, May 27, 2008

Peer reviewed versus peer to peer

There has been some chatter as of late in the library blogosphere in which there has been talk of supplanting library peer-reviewed journals with library blogs. Dorothea has provided my favourite response to the matter but I do want to add one little tiny thought.

I suspect one of the things that is holding the (inevitable) transition to unmediated online writing through self-publishing software is the the fact that words blog and blogger still sound ridiculous (hover over image for alt-tag).

Tuesday, May 20, 2008

You can and must understand computers NOW

computer lib

One of the reasons why I was so disappointed with the One Laptop Per Child's decision not to support teachers in their supposed mission of educating children is because I know that its not enough to put a computer in front of a child and expect that a savvy computer user will emerge in time.

I know this because when I was young, my gadget-loving dad brought home an Apple IIe and for the life of me, I couldn't make it do anything but use it for word processing and to play games that one my dad's co-workers kindly copied for us. It wasn't that I wasn't interested in computers. By that time I had already taken a couple "computers for young people" classes through the local community college. But my knowledge of BASIC did me no good and the manuals that Apple provided may as well been written in binary. The computer was a black box with big floppy disks.

Fast forward to high school. I was one of 3 or 4 girls in a computer class and we pretty much kept to ourselves. Eventually our classmate Brian would occasionally join us in our corner and amuse us with stories. After we gave him our respective username and passwords ( a strange trust exercise / friendship ritual that kids continue to do today) he somehow created little animated stick drawings that would display when we logged in at our terminals at the beginning of class. Somehow Brian had learned to do all sorts of cool and strange things that were never mentioned in class. When I asked him how he figured all this out, he told me that he and a small group of guys would hang out in the computer lab and over time had picked up tricks from each other and the teacher who was supportive of their interests.

Similarly, I eventually figured out that many computer enthusiasts became what they are through the help of other computer enthusiasts through user groups, or BBS, or Usenet.

In short, it takes a village to raise a geek.

Now, through the miracle that we call the Internet, its even easier to learn how to learn about computing. There are all sorts of manuals, FAQs and tutorials about like the newly resurrected webmonkey which I used some years ago to build my own HTML skills. And more importantly, there are kind people like Dan Chudnov who are willing to help you to learn2code.

The support is there. You just have to figure a project that you want to do and then beat things with rocks until its working.

Monday, May 19, 2008

OLPC : hubris or fraud

Until now, I had chalked up OLPC founder Nicolas Negroponte's disdain for teachers and software support as an extreme strain of faith in the innate computer capabilities of children and the common person.

I'm sad to say that recent evidence suggest that the reality might be much bleaker: that the OLPC program was and is only about the distribution of laptops.

I'm not so much sad over the failure of this particular program as sad over all the volunteer efforts and international goodwill that has been recklessly spent.

Tuesday, May 13, 2008


University of Michigan education professor David Cohen says that no education occurs until what he calls “inert” assets (books, teachers, rooms, curricula, rules, budgets, and so on) interact with each other and with students. Education is interaction. People in educational organizations, he says, often behave as if the inert assets were essential and the interactions expendable. They fight political wars over budgets, space, and personnel, and spend little time defending and perfecting the interactions among these assets through cooperation, communication, teamwork, and knowledge about students.

The above passage struck me as having a direct parallel to some of the recent changes I've noticed in academic librarianship. Putting it in broad strokes, we are going through a fundamental shift in libraries from being collection-focused to being user-focused. What I like about the above quotation is that spending time "perfecting the interactions among these assets through cooperation, communication, teamwork, and knowledge about students" sounds very much like what good Information Literacy practice aspires to be.

Its from an essay that calls for a fundamental organizational change in health care titled Escape Fire: Lessons for the Future of Health Care [pdf]. I read this essay because of this recommendation to do from Brett Bonfield on the ACRL blog. And I would second the recommendation. Its worth reading for its own merits relating to health care and for reminding us how important transparency and access to information can be in a person's life.

Thursday, May 08, 2008

Freedom for our FOIA

While on leave from libraryland, I'm still keeping touch through the library blogosphere. And in doing so, I noticed that none of the Canadian library blogs that I follow had made mention that the Harper government recently told all federal agencies to stop providing monthly updates to the public of all the requests it is answering under the Access to Information Act.

The Freedom of Information Act (FOIA) is one of the few tools that ordinary citizens can use to keep their government accountable. In my hometown, a citizen recently took on our excessively secretive city council with a FOIA and now we know the details of an arena contract between the city of Windsor and the local OHL hockey franchise.

When I was in library school at McGill, I wrote a loving profile of Ken Rubin, another private citizen who uses FOIA much to the chagrin to those in power -- and I am glad to say that I am not the only one to do so. But I think librarians should do much more to support the cause. Here's a start: every library's government document's web presence should provide a link and instructions how to place a Freedom of Information Act.

And if the CAIRS database is not reinstated, then I wish and hope that a library out there will create a Canadian version of WhatDoTheyKnow.

Thursday, May 01, 2008

A Library Fortress of Silentude? | Ask Metafilter

A Library Fortress of Silentude? | Ask Metafilter:
"There is a war going on in the library. This conflict is between students who seek solitary silent study and those who seek to study or work on projects in groups. An individual student's allegiance to a faction can change from day to day based on their current course load. Because the Grouparians have the advantage of numbers, they tend to win out over the Solitarites. Surely the latter group needs a fortress all their own?"

Sunday, April 27, 2008

Kill your television

My sibling and I have had an ongoing conversation about the dearth of meaningful volunteer work and share a frustration that more non-profit groups don't make more use of the Internet's potential.

We're not the only one's who have recognized the untapped potential that exists. Clay Shirky has too - but has quantitfied it and put into a historical context in his recent essay, Gin, Television, and Social Surplus. It's highly recommended reading.

Wednesday, April 16, 2008

XO Laptop - beautiful but tragically flawed

So yesterday I entertained my 2 1/2 year old son by bringing out my XO laptop from OLPC.

The XO laptop is designed for children from ages 6 to 12 but there are some activities that a small child can explore. Mats likes making noise with TamTamMini and making B I G L E T T E R S with Write, its text editing program. But it was the ability to capture video that made him scream and drool with delight (thank goodness for the xo's water resistant keyboard).

But like so many times before, my delight in the XO laptop quickly evaporated into pure frustration. After the little one had gone to bed, I turned on my XO to find that I could not view any of the videos I had just recorded of my little guy. I sighed and then, with a growing sense of dread, I went online to see if someone else had experienced the same problem. I had been down this road before - and it wasn't pretty.

You see, there are multiple sources of support for the XO. There is the official OLPC Wiki - which suffers from a severe lack of editorial work. Don't believe me? This is what I mean: there is a community user guide, a new user guide, a workaround section, the "official FAQ" and the support FAQ, the Ask OLPC A Question section (which has its own archive), and of course, separate pages for the various aspects of hardware and software which often have different information than the guides. And supplementing the wiki, there are two support discussion forums: the official one and the OLPC News Forum. And supplementing those are mailing lists (public, community and official ones) and, if you like your help old-skool-stylee, there are several IRC channels on the subject of the OLPC.

Trouble is, not a single person is dedicated to answering any of the questions. Evidently children are going to do the job but in the meantime, its just volunteers.

You see, the founder of OLPC Nicholas Negroponte, doesn't believe in product support for the XO. According to Negroponte, I should be asking a small child to teach me how to use my laptop because, "with all due respect" to those who believe that education depends teachers, schools, curriculum, or content, success is dependent "on leveraging the children themselves".

Except that the XO is buggy as hell.

I *understand* how video should work. But it doesn't. And after hours of searching wikis and forums, I now know that at least one person has experienced the same problem since January and that no one has offered a possible solution yet.

I have a lot of faith in the power of online communities but unlike the OLPC organization, I don't take such groups for granted. The Open Source community is a beautiful one but they are more likely to take the time to answer a question that can be answered in python than to take the time to do the multi-step analysis required to find out why, for a presumed small group of users, playing recorded video doesn't work.

Non-working video hasn't been the first problem I've encountered. When I try to register my XO nothing happens. Evidently, the Write activity is able to save text as html but when I bring up my code in Browse, it refuses to recognize the code (there are TWO pre-installed activities dedicated to learning programming in Python and nothing to create a simple webpage?!?).

Hardware-wise, the XO laptop is a thing of beauty. But until a core group of XO activities can be developed to perform solidly and consistently, I don't think an entire nation should invest in these laptops for their children.

Tuesday, March 04, 2008

Ontario Scholars Portal – Yours to a Discovery Layer

(This is a little something I wrote to support this work)

What is a “Discovery Layer”?

To me, a Discovery Layer allows a user to search across a library catalogue (or several), an ebook platform (or several), and a source of articles (or several). 


A Discovery Layer could make use of one or more combinations of the following:
  • Federated Searching
    • a query is distributed to multiple sources, and responses are compiled, de-duped, and returned
    • e.g. Sirsi Single Search
  • Metasearching
    • a single, regularly complied index is created from the collection of metadata from multiple sources
    • e.g. Endeca, Google
  • Single host environment
Why a Discovery Layer?
The pursuit of a Discovery Layer seem to be driven by the need to present one, strong and stable user interface over many disparate sources of information. Some benefits of a discovery layer include:
  • users only have to learn one interface, instead of many
  • users don’t have to choose from lists of dozens of indexes
  • users don’t have to repeat searches depending on format (one search for books, then one for dissertations, then one for articles…)
  • users expect simple, effective search tools like Google
What’s the problem?
The challenges that face the construction of a discovery layer include:


  • many of our research tools are very difficult to extract data from as they make use of a multitude of non-standard formats and protocols
  • most of our research tools (especially the library catalogue) generate search results with poor relevance ranking
  • some sources will be rich in text and metadata (articles, ebooks) while other sources will only be represented by metadata (print books)
How much can an improved interface improve things?
At the present time, I would say that there are 3 archetypes of Discovery Layer Interfaces.
How much can an improved interface, improve relevant results?
Coming up with what a user might deem relevant from 2 or 3 keywords is challenging in a regular search environment. Producing consistently relevant results in a federated or metasearch environment is extremely difficult.
Relevance might be improved through one or more of the following:
  • by taking into account the user’s previous searching behaviour
  • by weighing results by the number of times an item has been bookmarked, printed, or saved
  • by using citation information to determine ‘likeness’ (e.g. based on a percentage of shared citations in item’s bibliography)
  • by using user-created lists articles to generate similar items of possible interest
  • by knowing what courses a users is currently taking/teaching and emphasizing relevant resources accordingly
What is a Good Enough Discovery Layer?
Is it realistic to expect a Discovery Layer to serve both the novice researcher and the expert to access a variety of formats in a multitude of disciplines? Can one size fit all? Should we develop several Discovery Layers with one for each discipline? (Arts, Social Sciences, Medicine). Should we develop one interface for undergraduates and one for faculty and graduate students?

How will we know we have reached the Promised Land?
Most discovery layers are still in the earliest stages of their development and by appearances, they seem more alike than unalike. How should we choose what is an acceptable product? One suggestion is to measure the success of a Discovery Layer by comparing its search results to Google.

Thursday, February 28, 2008

Libraries COE

Sun Microsystems, The University of Alberta Libraries and The Alberta Library create Centre of Excellence for Libraries: "SANTA CLARA, CA February 27, 2008 Sun Microsystems of Canada Inc., the University of Alberta Libraries (UAL) and The Alberta Library (TAL) today announced the creation of a new Sun Centre of Excellence for Libraries (COE)."

Wednesday, February 27, 2008

Digg + Local Library Purchases


Digg + Local Library Purchases: "So here’s my idea: take the engine that runs Digg, the “social news” website, and repurpose it as a web application that allows library patrons to collectively decide which books the library system should purchase. Patrons would “login” to “LibraryDigg” with their regular library card number and password, and then could enter books, DVDs, etc. that they want opened up for consideration." [Distant Librarian]

What I really like about this idea is that this service would provide public feedback illustrating what the library community is interested in and what are their unmet desires. I'm positive that this sort of information would be of interest to more than librarians as in my library, one can always see users check out the responses on the library's complaints bulletin board (which one day I would love to put online like Carleton's Dear Library service.)

But I'm not sure that I would use the Digg engine. I check out Digg and Reddit frequently and its not exactly a secret that the system is constantly being gamed (e.g. "Garbage can be turned into oil through a green method. We dispose of enough garbage per year to create a years worth of usable oil! Why aren't we funding these guys millions! UP VOTE THIS!!").

Instead, I would be more inclined to use something more independent engine to determine popularity.
Digg's Design Dilemma - Bokardo : The result of all these factors is that Digg breaks the cardinal rule of voting: independence. As outlined in James Surowiecki’s book The Wisdom of Crowds, independence arises when a person makes a decision (votes, diggs) without the direct influence of others, on their own, by making up their own mind. Of course, there will always be influences on that decision…what others have said, where their political party is leaning, their current situation, but in the end they need to have the privacy of their vote. On Digg, no votes are private, and when you make them you can’t help but notice the way others are voting...
The voting on Digg is in contrast to a site like Del.icio.us, where voting (saving a bookmark) is done more independently, often without having any idea whether or not someone else even viewed it, let alone voted on it... On Del.icio.us, the main value is personal, as people use it to store bookmarks that are valuable to them. On Digg, the bookmarking utility is secondary to the voting, in both the interface and the wording used on the site.

This all being said, I do support the notion that a library community can be trusted to select some of the materials that can be found in *their* library.

Tuesday, February 26, 2008

Riddle me this at your library

The MIT Libraries Puzzle Challenge is a series of puzzles released during the MIT semester. The series began with three puzzles in the Fall 2007 semester, and will continue with three more puzzles during the Spring 2008 semester. Puzzles are produced with the dual goals of being fun for our community and of introducing solvers to resources they might not otherwise know about.

My personal epiphany that there was a natural connection between puzzle games, learning, and libraries happened while I was playing the World Without Oil ARG and read as strangers on a forum thread worked together to identify a foreign language used in a video clue (Bulgarian) and then translate and transcribe the speech within hours.

I'm not the only one to have made this connection.

And MIT has been playing puzzle games with their students for over 25 years.

Saturday, February 23, 2008

Read about Griefers. Then be one

I've had a long standing prejudice that if you rely on WIRED magazine in order to understand Internet culture, then you probably are not particularly WIRED.

But two things have happened since I've made this conclusion : WIRED has become a much better magazine (I'm talking since the days from when they would lovingly profile venture capitalists and worship corporate heroes). And I've become older and (more) out of touch.

The article that made me come to this realization is from February issue of WIRED (16.02) and is called Mutilated Furries, Flying Phalluses: Put the Blame on Griefers, the Sociopaths of the Virtual World. Before this article, I knew that griefers existed but didn't know much more about them than that. After this article, I'm shocked to say, that I have now some small sympathy for their cause : the battle against "The Internet is Serious Business." I can't and won't excuse the worst of their behaviour (the death threats, for one) but now I can see the humour in the crassness and the method in the madness.

This deeper understanding has helped navigate the world of the very offensive and very amusing forumwarz [waxy]. In this role-playing game about the Internet, I'm playing a Troll. Who knew being a griefer was so much fun?

[cross-posted on NJA]

Monday, February 18, 2008

The Big Why About The Big Here

Its early evening. The Toddler is in bed. Right now is a short period of personal time before the promise of sleep will lure me to bed before the rest of the world settles in to watch primetime television. For some days now, this is the time where I begin researching answers to such questions as

Where is the nearest earthquake fault? When did it last move?

I've been meaning to complete The Big Here Quiz ever since I first stumbled upon it in an issue of Whole Earth Review years ago and I've finally hit its half-way point. Its 30 questions are meant to inspire watershed awareness but in doing the quiz I've found that it has generated more than just a stronger curiosity and connection to where I live. It is making me a better librarian.

Finding who can delivery pizza to your house is not hard. Finding out whether "the soil under your feet, more clay, sand, rock or silt?" is surprisingly hard. Determining out how water gets into your tap and tracing where it goes after its sent down the drain is largely dependent on how much information your local utilities care to share with you. The information you need to finish this quiz can largely be found online, but in forms that have been fragmented and rebound depending on government level, jurisdictions, and departmental politics. (Watersheds transcend such boundaries).

By doing this quiz, I have been reintroduced to many government document websites that I have not visited in a long, long time. Many of these sites now feature GIS driven datasets which with a bit of trial and error, can usually be cajoled to produce information but the results are usually disappointingly thin. I'm waiting for the next generation of localized collections of data visualizations but I think it will be a long, long time before EveryBlock comes to Windsor, Ontario.

Like a good homework assignment, it is the curiosity that stretches while pursuing the question which is more important than the "answer". Now I'm wondering about the worms in my backyard, whether I really understand the concept of azimuth, and where is that closest fault line...

Sunday, January 20, 2008

OLPC and the Library : The Talk

If you are in the Windsor / Detroit area, please consider joining a talk sponsored by MPOW called:

One Laptop Per Child: Open Source, Open Access, Open Library
January 28th, 7 pm, Freed-Orman Centre
By attending, you could win your very own XO laptop!

There's an article about the OLPC in the Toronto Sunday Star. What I found very amusing is that I share the journalist's same sensibility towards simple computers: he used to own an eMate and I used to lust after one.

Tuesday, January 08, 2008

Today my guest post at OLPC got published

Today the OLPC News blog kindly printed a guest post of mine about the relationship between libraries and the work of the One Laptop Per Child program.

That's the good news. The bad news is that evidently, there were only 4000 orders for XO laptops from Canada and that due to shipping logistics, the Canadian deliveries have been pushed back until February 15th!

Thursday, January 03, 2008

OLA wants your picutres of libraries from around the world

The Ontario Library Association's Super Conference 2008 is asking
Do you have pictures of photos of libraries you've snapped in your travels? If you’ve got some hidden away on your hard drive or in a photo album please help us by submitting them to OLA. Photos of any kinds of libraries (or librarians) in any location are good...

Photos will be used in a gallery on the OLA Super Conference Web site, at plenary sessions throughout the conference and, most particularly, at the closing plenary on Saturday, Feb. 2nd during the luncheon in which incoming IFLA President Ellen Tise from South Africa will be speaking. Photos can be submitted in any of the
following ways:
  • Send digital photos to superconference@accessola.com (in any format but as high a resolution as possible).
  • If you are a Flickr member, you may submit photos to the “Libraries of the World – OLA Super Conference Pool” Flickr group.Again, as high a resolution as possible.
  • Send print photos to OLA, 50 Wellington St. East, Suite 201, Toronto M5E 1C8. Photos can be picked up at Super Conference from the Information Desk in the Registration Lobby. Any photos not picked up will be mailed to their owners following Super Conference.
I'll be at this year's Super Conference for Friday only and will be speaking on that day with Stacy Allison-Cassin in Session #1318: Scholar's Portage: Leveraging Social Networking Tools and Scholars Portal Data.

Wednesday, January 02, 2008

Books that scientists use - an addendum

Most cited references in Nature 2007 - all books

Some years ago when I started my tenure as Science Librarian at the Leddy Library, I worked on a particularly tedious project that really helped me get a better understanding on what books scientists actually use in their work (as opposed to so many other titles).

Using Web of Science, I would search for a year's worth of articles from Science, Nature and PNAS, download these articles' references into an Excel spreadsheet and try to decipher which ones were books. It was a very very mechanical and time-consuming process. But it made me a better science librarian.

Now, I'm pleased to say, that Scopus makes this same task possible in less than 30 seconds:

1. in search for box type in journal name (e.g. Nature) and select Source Title from drop-down menu
2. limit date range to a particular time period (e.g. 2007)
3. hit search button
4. in the Refine Results box, limit your results to just the journal name in question (e.g. Nature and not Nature Biochemistry)
5. in the Results box, click the Select All box and the hit the References button
6. If your initial search results brought more than 2000 hits, you will be informed that only the first 2000 articles will have their references retrieved.
7. review the list of most cited references of that journal from most cited to least

Here's a list of the 10 most cited books by Nature, Science, and PNAS*. The second number in the list represents where in the journal's list of most cited items can the book be found.

Most Cited Books in Nature, 2007**
1. [1] Molecular cloning: a laboratory manual (1989)
2. [2] Diagnostic Statistical Manual of Mental Disorders (1994)
3. [3] Numerical Recipes (1992)
4. [4] Biometry (1995)
5. [5] Biostastical Analysis (1984)
6. [7] The Rat Brain in Stereotaxic Coordinates (1986)
7. [13] Principles of Optics (1980)
8. [15] Physics of Semiconductor Devices (1981)
9. [19] Intermolecular and Surface Forces (1992)
10. [21] Co-Planar Stereotaxic Atlas of the Human Brain (1988)

Most Cited Books in Science, 2007**
1. [1] Diagnostic Statistical Manual of Mental Disorders (1994)
2. [2] CRC Handbook of Chemistry and Physics (2005)
3. [3] Biometry (1995)
4. [4] Biostastical Analysis (1984)
5. [5] The Rat Brain in Stereotaxic Coordinates (1986)
6. [9] Principles of Optics (1980)
7. [10] Computer Simulation of Liquids (1987)
8. [11] Structured Clinical Interview for DSM-III-R (1990)
9. [12] Advanced Organic Chemistry (1988)
10. [14] An Introduction to Probability Theory and its Applications (1971)

Most Cited Books in PNAS, 2007**
1. [1] Molecular cloning: a laboratory manual (1989)
2. [4] Diagnostic Statistical Manual of Mental Disorders (1994)
3. [6] Numerical Recipes (1992)
4. [8] CRC Handbook of Chemistry and Physics (2005)
5. [9] Biometry (1995)
6. [10] Biostatistical Analysis (1984)
7. [12] The Rat Brain in Stereotaxic Coordinates (1986)
8. [13] Handbook of Mathematical Functions (1972)
9. [22] Stastical Power Analysis for the Behavioral Sciences (1988)
10. [23] Statistical Methods (1980)

*Nature and Science published almost 2000 items in 2007 and PNAS published closer to 3000 items in 2007.

I'm planning to use this method to determine if there are any important books my library is missing by reviewing the references of the key journals in various fields for 2007. Its still a largely mechanical process (although the Foxy Leddy LibX toolbar makes book-checking much faster than typing titles into the library catalogue) but its a good task to slowly start the work of the new year.

**Addendum:
I've been thinking further about these lists and I think I am in grievous error.

For example: the approximately 2000 articles in Nature evidently produce 24,660 references. That's about 200 items in each item's bibliography - which sounds high but its in the realm of possibility. But what confuses me is the cited by column which says that the first item in the list, Molecular Cloning: A Laboratory Manual, has been cited 103, 825 times. So that must refer to how many times the item has been cited within the Scopus database. The fact that the some of the books in the list appear in the same relative order when performing the same procedure using the journals PNAS and Science, means that these lists reflect the most popular science books within Scopus and not necessarily within each journal.

Rats.

Monday, December 17, 2007

Scholars Portal Day Presentation

Stacy Allison-Cassin and myself gave a presentation / daytime talkshow chat on our Scholr 2.0 paper last week at the George Ignatieff Theatre, University of Toronto as a part of Scholars Portal Day : slides video stream (37 min)

Monday, December 10, 2007

Information Literacy Literature Review 2006

For the last handful of years, the journal Reference Services Review has published an annotated round-up of articles on Information Literacy and continuing this tradition, its most recent issue (Volume 35 Issue 4 2007) has the 57 page literature review, Library instruction and information literacy 2006.