Category Archives: Digital Humanities

Lurking Impactfuls

Today on the NPR I heard someone say, “even more impactfully” [link].  Knowing that anything to do with “impact” is peever-bait [of the “only teeth can be impacted” variety–see BBC Magazine: “Should “impact” ever be used as a verb?“], I was surprised to find that “impactfully” occurs unselfconsciously about once a day on Twitter, and […]

The Lifespan of Words (three ways)

Getting ready for DH2017 this morning, I found myself curious about the lifespan of English words–when they come into the language and when they fall out. So I got all the earliest and latest attestation dates for all the words in OED3, and plotted them out. Here are three graphs (“visualizations,” if you like), all […]

Guest Post: Don’t go breaking (up) my genre: visualizing genre against attributes

Danielle Griffin is a research assistant on her third co-op term at The Life of Words. This is the first of a few posts based on her last work-term report,”Comparative Data Visualizations of Textual Features in the OED and the Life of Words Genre 3.0 Tagging System”. Danielle’s report won the Quarry Integrated Communication Co-op […]

How Indigenous American words came into English

I’ve been deep in the OED documentation of borrowings and loanwords for my look at “tramlines” [see my previous post, and look out for a few more to come] and OED’s treatment of foreign, about to be naturalized, and naturalized words. I got curious about some of the Indigenous American words in my dataset, and […]

Three conferences this summer

After a baby-related travelling hiatus of a couple three years, TLOW is hitting the road this summer, with stops at Ryerson University in Toronto (just barely down the road, really) at the end of May, for the Canadian Society for Digital Humanities meeting at CFHSS Congress; then off to Barbados and the University of the […]

One last round with metadata from Hathi and Underwood

In “Hathi’s Automatic Genre Classifier” and “Hathi Genre Again – Zero Recall“, I ran a couple of experiments comparing genre categories assigned by human taggers working on the Life of Words OED mark-up project to two sources of genre metadata associated with the HathiTrust Digital Library. The first post looked at data from the automatic […]

Shakespeare’s Earliest Citations in the OED

No author’s representation in the OED has received more comment than Shakespeare’s: if you ever come across a mention of OED citation evidence, more than likely it’s being used to substantiate (sometimes challenge or qualify) a claim that Shakespeare invented the most English words, or made up the most new meanings for existing words, or […]

OED Subject Matter

In my last post I described using HathiTrust’s Solr Proxy API to fetch Hathi genre metadata for OED quotations. But genre is not the only metadata that Hathi sends back down the intertubes when I ask it a question. For most works, I also get a Library of Congress Classification code for the volume. This […]

Hathi Genre Again – Zero Recall

In “Hathi’s Automatic Genre Classifier” [17.01.06] I compared the consolidated automatic genre metadata for a subset of HathiTrust Digital Library texts (available here) to the genre classifications arrived at for human-inspected works as part of the OED quotation tagging project under-way at The Life of Words. My process there was pretty closely supervised, but the […]

Guest Post: Magazines and the Dentist Test

Cosmin Dzsurdzsa is a research assistant working on identifying the textual genre of quotations in the OED. Here he writes the first in a series of posts on borderline and difficult genre determinations. Filtering quotation blocks is essential to optimizing our results with the quantity of data we deal with here at LOW. For a […]