Sapping Attention

Digital Humanities: Using tools from the 1990s to answer questions from the 1960s about 19th century America.

Monday, January 31, 2011

Where were 19C US books published?

›
Open Library has pretty good metadata. I'm using it to assemble a couple new corpuses that I hope should allow some better analysis than...
1 comment:
Friday, January 28, 2011

Picking texts, again

›
I'm trying to get a new group of texts to analyze. We already have enough books to move along on certain types of computer-assisted text...
Friday, January 21, 2011

Digital history and the copyright black hole

›
In writing about openness and the ngrams database, I found it hard not to reflect a little bit about the role of copyright in all this. I...
2 comments:
Thursday, January 20, 2011

Openness and Culturomics

›
The Culturomics authors released a FAQ last week that responds to many of the questions floating around about their project. I should, by t...
Tuesday, January 18, 2011

Cluster Charts

›
I'll end my unannounced hiatus by posting several charts that show the limits of the search-term clustering I talked about last week be...
1 comment:
Tuesday, January 11, 2011

Clustering from Search

›
Because of my primitive search engine , I've been thinking about some of the ways we can better use search data to a) interpret historic...
6 comments:
Monday, January 10, 2011

Searching for Correlations

›
More access to the connections between words makes it possible to separate word-use from language. This is one of the reasons that we need a...
1 comment:
Thursday, January 6, 2011

Basic Search

›
To my surprise, I built a search engine as a consequence of trying to quantify information about word usage in the books I downloaded from t...
Wednesday, January 5, 2011

Correlations

›
How are words linked in their usage? In a way, that's the core question of a lot of history. I think we can get a bit of a picture of th...
2 comments:
Thursday, December 30, 2010

Assisted Reading vs. Data Mining

›
I've started thinking that there's a useful distinction to be made in two different ways of doing historical textual analysis. First...
5 comments:
Monday, December 27, 2010

Call numbers

›
I finally got some call numbers. Not for everything, but for a better portion than I thought I would: about 7,600 records, or c. 30% of my b...
3 comments:
Sunday, December 26, 2010

Finding keywords

›
Before Christmas, I spelled out a few ways of thinking about historical texts as related to other texts based on their use of different word...
2 comments:
Thursday, December 23, 2010

What good are the 5-grams?

›
  Dan Cohen  gives the comprehensive Digital Humanities treatment on Ngrams, and he mostly gets it right. There's just one technical poi...

Second Principals

›
Back to my own stuff. Before the Ngrams stuff came up, I was working on ways of finding books that share similar vocabularies. I said at the...
5 comments:
Sunday, December 19, 2010

Not included in ngrams: Tom Sawyer

›
I wrote yesterday about how well the filters applied to remove some books from ngrams work for increasing the quality of year information a...
2 comments:
‹
›
Home
View web version
Powered by Blogger.