Posts

Showing posts with the label code4lib

Presenting at Code4Lib-North

Using Open Source Tools for Visualization and Semantic Mapping in a Large Scale Article Digital Library [slideshare] - Questions: Anyone working in the cloud? Converting 4TB of TIFFS to PDF in 24hrs: Hadoop + EC2 + S3 = Super alternatives for researchers (& real people too!)

code4lib 2009: Day 1+2

Day 1 LibX2. LibX Edition builder . Build custom version of LibX Xtensible Catalog. Drupal, LMS, NCIP, LMS integration: Blackboard. webcast . scriblio: Social Library System Wordpress based plugin enjoysthin.gs . Mark Matienzo. anarchivist.Rich contextual book marklet. Emily Lynema. NCSU Libraries. E-Matrix : Open Source ERM Eric Lease Morgan. Alex4 . Erik Hatcher. Lucid Imagination. Lucene/SOLR. Index of Lucene apache site. Mike Taylor and Mike. Index Data. Translucent record store=="Torus" pazpar2 . Registry of searchable targets? Hard to do. IRSpy :Z39.50 Mike Beccaria from Paul Smith College. Microsoft DeepZoom . "Like microfiche" - audience. Photosynth of library stacks?? Dan Chudnov, LOC. BagIt File Package Format Random things heard and seen: citation style language , Open Vocab , UCSD Libraries Digital Assest Management System , SWORDS . Distributed version control: monotone , mercurial , bzr . Day 2: A new frontier - the Open Library Environment (OLE) -- T...

code4lib: Sebastian Hammer quote

" If you have something to say, you should release it as code... " Sebastian Hammer , Index Data

code4lib update: LuSql talk done; Lucene, Solr links

Gave my LuSql talk today at code4lib2009 and didn't get cut down by any Solr/Lucene dudes! Met Erik Hatcher of Lucene/Solr fame (and now of Lucid Imagination fame) & hopefully we can collaborate on some Lucene/indexing Solr stuff in the future. I also spoke with Tom Burton-West of UMich about Lucene indexing and search performance for their 1M+ Google Books index (they use Solr). These are documents that are a lot longer than the STM articles I work with. They have 220GB sized indexes and - as they have to keep stops words for their Humanities for phrase searching - suffer from poor query performance (despite 32GB RAM). I pointed to some of my previous work on high performance indexing and searching [ 1 , 2 , 3 ]. I'd like to get at their data to examine some performance issues in Lucene, both on the indexing and searching side. I was wondering if Solr is configurable for the initial/max number of IndexSearchers. I couldn't find this in the Solr wiki, but did see in...

"Elvis impersonators as XML documents"

As heard at code4lib2009: " Let's say all of these Elvis impersonators are XML documents... " - Mark A. Matienzo, New Your Public Library

code4lib pre-conference: Linked Data et al...

I am at the exciting and arcane code4lib 2009 conference here in Providence, Rhode Island. Right now at the pre-conference called LinkedData . on Linked Data . I had forgotten that Rhode Island and more specifically Providence, are the old stomping grounds (and location for many short stories and novels) of H.P. Lovecraft . And - this morning - I was talking to Ross Singer about this, and realised how this all made sense: when I first met Ross at an Access conference a number of years ago, the first thing I thought on meeting him was, " Chthulu "! He of course denied being one of the Elder Things and then levitated across the room from me. But I think this explains a lot of things... ;-) We will have to see what other Links I make at this conference. :-) Oh, BTW I will be giving a presentation tomorrow morning on LuSql . Feel free to drop in. :-)