Just a reminder, you’ve still got a few hours to post your photos for the Harry Potter book pile contest…
The submissions so far (plus some more listed in the comments of this blog post).
Just a reminder, you’ve still got a few hours to post your photos for the Harry Potter book pile contest…
The submissions so far (plus some more listed in the comments of this blog post).
Labels: book pile, contests, harry potter
So, it looks like Saturday is going to involve lightning, so the barbeque has morphed into a pizza party.
But we’re not getting just ordinary pizza! We’ll be ordering from Emma’s Pizza in Cambridge, MA, which is some seriously gourmet stuff (menu).* And we’ll have plenty of chips and hummus too.**
Everyone is invited. Bring a friend. If you can RSVP, great. If not, that’s fine. We’re probably going to need to order the pizza beforehand. We’re going to double the RSVP list. If you like, you can bring something .
When: Saturday, July 28th, 4pm to whenever.
Where: 15 Gurney Street, Cambridge MA (Google map). It’s about 15 minutes walking from the Square. You can also take the 72 (Huron) bus, and ask for the Fayerweather stop.
Parking: The City of Cambridge has declared Saturday LibraryThing day***. You can park anywhere on Gurney Street and between Gurney and Huron on Fayerweather.
Can’t wait to see everyone!
*When Emma’s was at the foot of Gurney Street, when I was young, it was decidedly less upscale. There were no tables—just a counter nobody used—and the ambiance was comprised of Emma berating her meek husband Greg in angry, staccato Armenian all day long. When the current owners bought it they moved it to Kendall Square, avoided marital conflict, added tables and the goat cheese, sun dried tomatoes and artichoke hearts today’s Cambridge requires. Somehow they managed to preserve what was good about it. It’s an amazing pizza.
**Alas, Tim’s trademark sigara börek will not fit with the rest of the meal.
***Untrue.
Labels: 1
In the spirit of fraternal concern, I post that the Internet Archive is looking for a systems engineer with PHP experience for their book-scanning project. (Also they promised to send us their discards. We need one too.)
The help-wanted has some excellent provisions:
If I had the skills, I’d be tempted to take it. The Internet Archive is a great institution. The people are great, and they have the best office space ever. LibraryThing’s second-story apartment steps from the Portland waterside pales in comparison. They have this adorable jewel-box in San Francisco’s Presidio, with the Golden Gate Bridge right outside the window.
Labels: internet archive, jobs
A tenth book-swapping site has chosen to integrate with LibraryThing, ReadItSwapIt.
ReadItSwapIt is a UK site, and boasts some 125,000 titles available right now. Here’s the page on LibraryThing showing copies of a book on ReadItSwapIt. They’ve got a bang-up buzz page. It’s a credit-free system. This is a bold move, but not unattractive. As they explain it:
“Swapping on ReadItSwapIt is like swapping with friends. If you like each other’s books, you swap. If you don’t, either of you can reject the swap.
Many swap sites operate a credit system. That means, instead of swapping books, these sites allow you to swap credits. Any time anyone wants a book of yours, they give you a credit and you post the book out to them. You can then use this credit later to get a book sent to you by someone else.
This sounds great on the surface. But in practice, the problem with this system is that whenever anyone requests any of the books you have registered, you have to post out that book immediately. It doesn’t matter how inconvenient it might be for you to get to the post office that week. And what if you go on holiday? You have to let the site know. You have no control over the amount of swaps you make. You could end up acquiring loads of credits but not be able to find any books you like on the site. So you’re left with a load of worthless credits, no books and a big postage bill.”
That rings true to my experience. I posted a book to a swap site, and then dithered when I got the email. My wife decided she wanted to read it too. I ended up buying the book on Amazon and having it expressed to the swap person, just to save face. People like me need a more forgiving system.
I’m also impressed by their commitment to accessibility It’s way more than LibraryThing does.
Labels: swap
New England Journal of Medicine article–with fancy animation–explains how social networks cluster… by weight. (Hat tip: David Weinberger)
Labels: Uncategorized
I’ve flogged David Weinberger’s Everything is Miscellaneous before, in blog posts and in my Library of Congress talk. I think it’s something close to the intellectual justification for LibraryThing.
Anyway, I flogged the ARC for so long that, when it came out, LibraryThing bought a small box of hardcovers direct from the publisher–to give out at conferences, to thank people for inviting me to talk, and so forth. I still have half a box. So I’m going to open it up to the whole LibraryThing community.
We’re going to give out ten copies. We like contests*—we have a Harry Potter book photo and another review contest going—so we’re going to make a contest of it.
I’ve created a thread Contest: What does tagging do to knowledge?
I’m not going to pick winners. I’m just going to randomly pick ten members who left comments. But you can’t just say “I want a book.”
*No purchase necessary. Void where prohibited. Also void where discouraged, unseemly or tacky. We pay to mail it. You are responsible for taxes. My taxes.
Labels: Uncategorized
The LibraryThing groups feature turns one tomorrow, followed shortly by Talk. I thought it would be fun to share the news-messages statistics for the Harry Potter Group, Hogwarts Express.
Check out the little boom for the movie (released July 11) and the crazy boom-bust-boom around when the book itself was released. For 24 hours, LibraryThing Harry Potter fans were reading, dammit.

I can report from experience that the rest of the world is still reading it. I went down to New York on business yesterday (and got caught in LaGuardia overnight, but that’s another story). The plane was like Harry Potter study hall.
REMINDER: We’re giving away prizes to 50 Harry Potter reviewers.
Labels: groups, harry potter, statistics
Short version: I’ve just gone live with a new feature called “tagmash,” pages for the intersections of tags. This is a fairly obvious thing to do, but it isn’t trivial in context. In getting past words or short phrases, tagmash closes some of the gap between tagging and professional subject classifications.
For example, there is no good tag for “France during WWII.” Most people just don’t tag that verbosely. Tagmash allows for a page combining the two: France, wwii. If you want to skip the novels, you can do france, wwii, -fiction. The results are remarkably good.
Tagmash pages are created when a user asks for the combination, but unlike a “search” they persist, and show up elsewhere. For example, the tagmash for France, Germany shows France, wwii as a partial overlap, alongside others. Related tagmashes now also show up on select tag and library subject pages, as a third system for browsing the limitless world of books.
Booooring? Go ahead and play a bit:
That’s the short version. But stop here and you’ll never know what Zombie Listmania is!
(full post over at Thingology, “Tagmash: Book tagging grows up”)
Labels: new feature, new features, tagging, tagmash
Tagmash: alcohol, history gets over the fact that almost nobody tags things history of alcohol
Short version: I’ve just gone live with a new feature called “tagmash,” pages for the intersections of tags. This is a fairly obvious thing to do, but it isn’t trivial in context. In getting past words or short phrases, tagmash closes some of the gap between tagging and professional subject classifications.
For example, there is no good tag for “France during WWII.” Most people just don’t tag that verbosely. Tagmash allows for a page combining the two: France, wwii. If you want to skip the novels, you can do france, wwii, -fiction. The results are remarkably good.
Tagmash pages are created when a user asks for the combination, but unlike a “search” they persist, and show up elsewhere. For example, the tagmash for France, Germany shows France, wwii as a partial overlap, alongside others. Related tagmashes now also show up on select tag and library subject pages, as a third system for browsing the limitless world of books.
Booooring? Go ahead and play a bit:
That’s the short version. But stop here and you’ll never know what Zombie Listmania is!
Long version. LibraryThing has shown some of the things that book tags are good for, such as plain language, genre fiction, capturing identity and perspective, academic schools, staying current and changing over time. (Details and examples in footnote.*)
It also demonstrates some of the weaknesses, including:
As I’ve argued elsewhere and in my Library of Congress talk, problems 1, 2 and 3 are mitigated by having LOTS of tags. Idiocy, malice and personal junk fall out statistically. A tag here or there can’t be trusted, but a large body of tags in agreement is different.
Problems 4 and 5 are harder to tackle. Flickr has shown the way with one solution, statistical clustering. The screen shot below shows this–clusters of images related to the tag “bow.”

Some day–when I become a better programer?–I’m going to try this on LibraryThing data. It will help with ambiguity—the secondary tags on the various meanings of “leather” are surely wildly divergent! But I suspect it separates better than it clarifies. Flickr supposes that tags fall into discrete clusters, but subjects interact with books in extremely complex ways. On a more basic level, I am suspicious of the too-quick resort to algorithms against user data.*** After all, if computers are so good at figuring out meaning, why were users necessary in the first place? It smacks of technological revanchism.
So, where Flickr’s clusters are automated, tagmash is a semi-automated process. LibraryThing does the statistics, but users decide what the meaningful clusters are. Some mashes are interesting and useful. Some aren’t. By and large, uninteresting clusters won’t last.****
This certainly helps with ambiguity. Take the problemmatic tag leather, which divides easily into tagmashes like:
Now let’s take the “focusing” power of hierarchy. As mentioned above, there is no good way to get at “france during wwii.” The tag Vichy covers some of the ground, but not enough. Tagmash provides an answer.
The book list is good, and a simple union gets around an imposed hierarchy. Looking at the related LCSHs, for example, one is left in doubt whether France is part of World War II, or World War II part of France—or what:
Of course, both trees are equally artificial. David Weinberger writes how, in the real world, a leaf can be on many branches. But it’s equally true that what’s trunk and what’s branch are largely about where you start–dirt or pinecone. Either way, branching happens. The order of the branches isn’t necessarily important.
Even as it borrows some of the virtues of subject classification, tagmash keeps the strenghts of tagging. Subject systems are pre-built things. Now and then they get larger, but it takes deliberation and effort. What gets “blessed” is often surprising. I would have never predicted the unusually staid LCSH would have embraced:
But tagging has no limits. Think of the tagmash “erotica” and “zombies” and there it is. (Tagmash: erotica, zombies). Want to know what chick lit takes place in Greece? (Tagmash: chick lit, greece.) Young adult books involving horses? (Tagmash: horses, young adult.) Poems from or about San Francisco? (Tagmash: poetry, san francisco). Slavery in Brazil? (Tagmash: brasil, slavery.) Non-fiction books about Narnia? (Tagmash: narnia, -fiction.) The options are endless.
Of course, tagmash only narrows the gap. It doesn’t eliminate it. Tagmash: poetry, San Francisco still can’t distinguish between poetry about and poetry from San Francisco–it involves whatever is tagged “San Francisco” and that’s probably a mixed bag.***** Well-planned and carefully executed subject systems have strengths that no ad hoc, regular-person system can match.
Lastly—let there be no doubt—tagmash needs a very large quantity of tags to work. For tagmash after tagmash, the data is simply insufficient.
You’ve made it to Zombie Listmania! There are some obvious directions this can go:
Amazon calls its static, or dead, lists “Listmania.” All these tend to create a “Zombie Listmania,” lists of books that “won’t stay dead.” Instead, they change over time, as the underlying social and non-social data change. There’s no reason you couldn’t create “Zombie” versions of formal subject headings—a series of tags and other markers which approximated the content of a professionally-assigned subject heading.
Pretty cool idea, I think. We’ll see what we can do about it.
Details.
Footnotes!
*What’s good about tagging:
**I’ve left out one problem, not covered at the LC—how “democratic” weighting can put Angela’s Ashes at the top of the Ireland tag. books. I want to write a blog post on the topic sometime. I think there are ways around it, and algorithmic solutions that nobody has really tried.
Aside: Much LIS anti-tagging polemic focuses on the most trivial of problems—spelling mistakes and “incorrect” tags. The former underestimates technology, the latter insults our intelligence. LibraryThing has dealt with the spelling problem, and has seen very few “wrong” tags. In fact, there are some serious problems with tagging. But you have to understand tags before you can see the problems, and many refuse to get past the idea that people will spell “white” wrong, or tag white horses as black.
***This is half formed. I have a problem with the reflexive “turn” from people-centered data to algorithms. I see this pattern again and again in software. Something transformative happens–something human. But it’s imperfect, so programmers conclude that programs will fix humans. In a way, it’s a reassertion of importance. More often, humans fix humans. To adapt David Weinberger, the answer to user-generated data is MORE user-generated data.
****Probably there’s got to be some system to expire unused clusters.
*****UPDATE: After turning the feature loose I watched what new tagmashes would be created. One was children, cooking. Should I call the police?
Labels: new feature, tagging, tagmash
I’ve added some spice to the home page—a section showing the books added recently. It updates every five seconds. The “last five minutes” section can go above 600. This one was taken during a slow patch.* But, hey—two Harry Potters!
It’s an experiment to see if being more up-front about the size and dynamism of the site will draw more users in. As one user ably described it, LibraryThing’s home page looks about the same now as it did twelve months ago. See more discussion of the experiment here. Whether it stays on the home page, we’re going to playing with features like this.
Tomorrow: Tags grow up.
*Saturday night is our nadir. And this weekend is for reading Harry Potter, not playing with LibraryThing. I note, for example that the Hogwarts Express group, one of our most active, has gone from near hysteria to eerie quiet!
Labels: features
The Library of Congress has just posted a talk I did there back in April, part of the Digital Future and You series.
I cover the basics of LibraryThing and some of what LibraryThing “means” to libraries, including a long section on tagging. It has a short section—a sermon, really—on open data, in anticipation of the launch of Open Library, and another on the upcoming Everything is Miscellaneous.*
To my regret, it ends abruptly. They didn’t include the 20+ minute Q&A**, which went a lot deeper on some of the interesting issues (particularly tagging), and with the nation’s top library talent!
Being asked to talk in front of the LC was a great honor. There aren’t many institutions I hold in higher regard. And it was fun. I got to be myself—PowerPoint-less, off-the-cuff and passionate–and was greeted warmly and given the benefit of the doubt when I pushed the limits. Also, I got to have lunch with some of their top people. It was a blast.
*The subtext of that section is that I just had a lunch conversation about open data, and heard more about the whys, wherefores and finances involved.
**Apparently they felt that they needed permission from everyone who appears on tape, and that the questions were not well miked.
Labels: Uncategorized
Altay’s attempt to insert the CSS version of the old <blink> tag into our upcoming Facebook application, produced this excellent reply from Facebook:

He was in fact kidding.
Labels: facebook