Tuesday, February 19, 2013
Code4Lib 2013 from Jonathan Rochkind
On the first day, I attended and presented at a pre-conference on libraries that have innovated in their web services around "delivery" -- that is, what happens after 'discovery'. I presented on the Umlaut software that powers our Find It service; Umlaut is open source software for which I'm the chief developer. The pre-conference demonstrated that there's a lot of interest in Umlaut at the moment -- several other libraries, including Princeton, are exploring using the Umlaut software themselves. And other libraries are developing their own alternative software inspired by Umlaut. As we've been using Umlaut for 4 or 5 years now, I think we can take pride in being ahead of the curve in paying attention to improving the access/delivery experience in ways Umlaut/Find It are intended to. As I've been trying to raise interest from other libraries in Umalut for several years, it's gratifying to see it starting to happen -- and will result in more sustainability for Umlaut as a software package, to get more involvement.
Once the conference itself began, one trend I noticed was many presentations about projects based on Blacklight. Blacklight is the open source project that our own Catalyst is based on. Blacklight seems to be only gaining in popularity, which again is encouraging for the sustainable future of the Blacklight project.
Many of these Blacklight projects were 'digital repository' type projects. Penn State and other institutions are developing a product called "ScholarSphere", which is based on Blacklight, Fedora, and Hydra (another open source project in the Blacklight eco system). ScholarSphere was impressive for having one of the best User Interfaces in an IR or 'digital repository' product I've seen.
In the past, I always liked Code4Lib for being just about the only conference on library technology that focused on innovation in what I'll call "research services" as opposed to "digital repository" services and products. Accross libraries, I think "digital repository" innovation receives a lot more attention, funding, and resources. It was hard to find the people working on innovation in research services too, which I think is also very important but doesn't get sufficient attention. However, this year, even at Code4Lib fewer presentations than usual were on tech innovations in traditional research services, and more on digital repository services. Lately I have started to wonder if it's simply too late for libraries to succesfully innovate in research services, maybe the future of libraries really is solely in digital repository services.
Also, when I first started going to Code4lib, almost all of my peers there were from similar situations to me, in terms of being one of only one or two developers in the library, sort of eeking out development of innovative web projects in their 'spare' time, amongst support of existing systems. This year, it hit me that quite a few libraries have extensive development staff these days. I couldn't help but be envious of libraries that have 10 or more developers working on developing new software, along with a handful of additional User Expereience staff, and additional completely seperarate staff supporting legacy products like the ILS. Here at JHU we manage to do an extraordinary amount of innovation with our relatively paltry development staff (1-4 developers depending on how you count, who are generally also responsible for extensive user support and legacy system support too) -- but it definitely makes me envious of how much we could accomplish and innovate if we had the resources of those peer institutions who have prioritized extensive software engineering staff resources (examples include U of Michigan, Stanford, and North Carolina State University).
All in all, the Code4Lib conference remains crucial to my professional skill and awareness of trends and possibilities in library technological innovation and relevant technologies.
Tuesday, January 8, 2013
MLA 2013
Highlights:
- Loved the session on "The Literary Lab," even though it turned out to really be about DH center-type lab, not the laboratory as a paradigm for the classroom (which is why I went).
- A roundtable on "The Dark Side of the Digital Humanities" was the source of much subsequent debate. You can read a summary in the Chronicle here.
- "The Digital Humanities and the Future of Scholarly Communication" included amazing talks by Matt Kirschenbaum, Cathy Davidson and Bethany Nowiskie; this panel was one of the special sessions related to the conference theme, "Avenues of Access."
- Michael Berube's Presidential address focused on the precarious place of adjuncts in the profession and the dismal outlook for grad students. He also talked about disability and disability studies. All these topics were linked, again, to the "Avenues of Access" theme. It was quite moving, and the Q & A featured a lot of folks simply thanking him for shining the spotlight on these issues.
- Two good sessions sponsored by the discussion group Libraries and Research in Languages and Literature: a roundtable on "how many copies is enough" re: print collections and a provocative roundtable about joint degrees in library science and literature.
- A session that is going to help me with the Crane exhibition, on William Morris and late 19th century arts and crafts publishing.
- First, for a panel called "Crossed Codes: Print Dreams of the Digital Age, Digital's Memory of the Age of Print," I gave a paper on extra-illustration largely based on materials at Garrett and the Peabody. Looks like it will be published in Textual Cultures this fall.
- I also organized a "master class" session called "Two Tools for Student-Generated Digital Projects: WordPress and Omeka in the Classroom," in which two practiced user-teachers, Amanda French and George Williams, led hands-on instructional activities to a group of about 50 people new to these technologies.
Tuesday, November 20, 2012
Charleston Conference, November 2012
--------------------------------------------------------------------
- Most users don’t read the whole book, like
reference linking, and hate DRM
- Most publishers are still designing print and then converting to electronic
- TRLN Consortium (Triangle Research Libraries Network) worked a deal with OUP – all 4 schools get a given e-book and one shared print copy. (Nancy Gibbs, Head of Acquisitions at Duke)
- Duke had many questions about e-book choices, processing workflow, etc., so they formed an E-book Advocacy group
- Before they started buying, they interviewed patrons, and then wrote a statement about how Duke can be advocates for patrons regarding e-books [PDF on the right]
- Duke also held an “E-book Boot Camp” for technical and reference staff, including overview, hands-on exercise, and how to find usage stats
- Texas A&M canceled their e-books packages and went back to choosing individual titles in GOBI to save money, and also went to PDA. Before doing any of that, they had a meeting of all the stakeholders to discuss.
- All speakers agreed that: (1) hiring new people and training of all personnel are crucial, (2) getting e-book records into the discovery layer as fast as possible, preferably daily, is crucial, (3) workflows will need to be disrupted, but lay it all out ahead of time and make sure that person/people in charge of e-book workflow(s) has project management experience
- Over half of the students didn’t know what e-thing they were looking at
- They were shown a sample of 18 online resources, such as journal articles, e-books (like Springer), open web sites, the catalog, a gov doc
- Springer e-book: 47% said it was a web site; only 28% knew it was an e-book
- Google Books e-book: most recognized what that was, not only because of familiarity but because of the simple uncluttered screen
- ScienceDirect article – 37% said it was a journal article
- The kids are becoming “format agnostic”; they don’t know or care what kind of thing they’re looking at and just want the info
- In 2012 so far, e-textbooks have been $4 billion, which is only 6% of the total e-book market, so there’s much room for growth
- Students still prefer print textbooks; profs reluctant to use because the edition they want isn’t available online or they’re just not interested, but this will change
- E-texts are about 50% cheaper than print
- E-texts are better than print because you can get usage stats, they have links to data/videos, they have self-assessment tools (if you do poorly they create a remedial lesson!), and profs can see what the students read. But these must be accessible across devices, and allow highlighting/notes/copying.
- Companies providing online education and/or textbooks include Boundless, Khan Academy, Cengage Mindtap, Yammer (allows working in groups and communicating with profs; used at ECU Dental School), VitalSource bookshelf, and Moodle
- The librarians took video of what the students did on screen, and audio of them explaining why they did it
- There was a script; e.g., “What’s an e-book? Have you ever used one? Find one, and use one [on various platforms].”
- Depressing results – They didn’t understand most of the language or icons that *we* understand, they love scroll bars (none on eBrary), and most screens were much too busy (e.g., MyiLibrary). Things weren’t intuitive, things were buried, it’s a steep learning curve to use our tools, confused about limits on printing/downloading, confused about browser vs. platform functions
- They can *define* e-books but can’t find them or use them
- About half the kids started looking *outside* the library
- They wanted to print or download the whole book to read later or mark up
- Their wish list for the future: using touch to flip a page and take notes, more e-textbooks, share notes with friends and prof, more intuitive, have audio
- WE need to pressure vendors to include students in their usability testing!
- AFTER these interviews, the students’ opinions of e-books were higher (education is good!)
- Q/A – You’re buying e-books and set approval plan to “e- preferred,” but the kids are lukewarm, so why? Because usage stats are up so they’re voting with their feet despite what they say (or maybe they don’t know they’re using them)
Tuesday, October 9, 2012
HathiTrust UnCamp
The HTRC's goal is to provide computational access to a large portion of the works in the HathiTrust. Being able to run textual analysis tools on a large corpus such as HathiTrust would enable humanists to do some interesting research. The mix of languages and subjects in the Trust closely resembles that of the physical collections of the contributing libraries. For now access is mostly limited to pre-1923 American publications that we know are in the public domain. They are starting to allow researchers to submit requests for analysis of items that are thought to be in copyright. The rationale behind this is known as "non-consumptive research", that is, the researcher would not be reading or "consuming" the text, but instead, would just be analyzing the words contained there. The proposed text-mining could be done only by researchers at institutions who have a signed agreement with Google as they would primarily be mining works scanned by the Google Books project. All of this is very much up in the air at this point.
Tuesday, May 15, 2012
Maryland Library Association Conference, 2012
The thing about this conference is that it's geographically-determined, so the full range of the Library World within that geographic boundary is represented: Research libraries; academic libraries; special libraries, school libraries; public libraries; even prison libraries, all represented in one form or another. As always, the great diversity of Maryland's libraries is impressive; and the fact that they share most of the same challenges despite being differentiated by user group, funding sources, etc., seems always to be the moral of the story.
The Opening Remarks were supposed to have been by Alexander Sanchez, Maryland Secretary of Labor, Licensing, and Regulation -- but he left that post earlier in the week! So his Deputy Secretary, Scott Jensen, appeared instead. It so happens that Jensen is a Philosopher by training and education and is intimately familiar with the great research libraries of New York City. So before he ever entered public service, he was a big fan of libraries. However, now that he serves in the Maryland DLLR he's even more of a fan, noting the great symbiosis between that department and the mission of public libraries across the state, at least with respect to the promotion of job growth and vocational training. He foresees a merging of the various DLLR One Stop centers across the state with local public libraries. He is, in fact, currently working on getting DLLR-mandated GED testing services to be hosted physically by local public libraries.
Anirban Basu, he of WYPR/NPR Fame, gave the keynote, essentially an extended riff on the global economy, the national economy, the state economy, and our local economies, with attention paid, here and there, to where libraries, primarily public libraries, fit in. Basu noted that public libraries are "in the business of promoting job growth." Let's hope that's not all they're in the business of, but good point. He then noted that over the past few years and despite the recession something like 1.8 million jobs were created in this country, mostly in the professional and business services sectors. Public libraries are crucial to this growth. Nevertheless, his numbers indicate that state and local financial support of public libraries is falling, while actual public library visitation and use is sharply climbing. During down times in the economy, people turn, as they should, to their local libraries.
If you ever have a chance to see Basu speak, take it! His delivery is rapid-fire and hilarious.
As I do each year, my notes from this conference are attached below, for whatever they're worth. They give me just enough of a reminder of what I attended and the high points of each presentation just in case I ever need to go back and delve a little deeper.
########################
GENERAL SESSION - Scott Jensen, Deputy Secretary, DLLR
GENERAL SESSION - The Dog Ate My Home, Anirban Basu
Technology Core Competencies for Library Staff
Beth Tribe and Maurice Coleman
Primarily public library staff. Howard County/Harford County public librarians
Library users and random devices
Nook, eReaders, instruction
eBooks on smart phones, apps
"Dedicated eReaders are the eight track tape decks of our time" "One trick ponies"
Public librarians WRESTLING with this issue!
anti-dedicated-eReaders. We need Options.
Tablets, smart phones Tablets Rule.
iOS and Android ecosystems. Part of our jobs
Tablets and HDMI, librarians putting tablet display on large screen
"rooting" the Nook for full-blown tablet
Not everyone has Internet access at home. Libraries to the rescue
Mobile Websites for Library or library catalog
Future is not OS dependent. "Mobile site that lives in the Cloud..."
Must work in ALL browsers
Smart phones and barcode scanners
Pintrest -- photo sharing. Local libraries participating
Quick Response QR code, smart phone scanning
QR code takes you directly to enhanced content
Google QR Generator
QR can send patrons directly to help
QR in the stacks
QR for room reservations
Creator/Maker/Hacker Spaces in public libraries
3D printers
STEM lab in Howard County Public Libraries. Create mobile games. Music production. Peer instruction
Virtual training statewide
Shared training videos in the MD public library systems, eReaders and Overdrive
savedelete.com
teleread.com
mashable.com
howtogeek.com
mediabistro.com/appnewser
webpronews
instructibles
techland
lifehacker
"There are two buttons on the thing -- one of them must turn it on. Press one!"
iPads on loan. Some public libraries doing this. Loaning of Nooks. Some eBooks cost more than the hardware.
Instruction crosses the line to tech support all the time.
Don't Miss the Hidden Treasures! Ideas for Successful Library Outreach
Nedelina Tchangalova, Engineering and Physical Sciences Library, University of Maryland
Gergana Kostova, UMBC Library
Simmona Simmons, UMBC Library
At UMD, Library Award for Undergraduate Research
About 50 US Universities making such awards. Open access to student research
Papers go into IR. Selection for award is based on these submissions.
Evaluation rubric -- assign points to each submission.
Posters, ad in campus newspaper, ebulletin boards, library blog, library Website, social network sites
Subject librarians broadcast
33 applications in first year, including 6 teams
three awards of $1000 each
Most applications from the College of Arts and Humanities
Embedded Librarianship
embedded in department
International Coffee Hour
showcase library resources and services
Speaking of Books
conversations with campus authors
similar to public library book talks
UMBC Outreach:
to new students
attract students to new learning center
Coffee Hour, Mini-sessions, Branded events, Q&A sessions, Library orientations
collaboration with Undergraduate Education; Residential Life; Student Life; International Education
YouTube video: "Don't miss the treasures!"
UMBC outreach to Faculty
high touch; low tech
Welcome email to new faculty, announcements of new resources, attend faculty meetings
Invite faculty to sit on library committees, hiring committees
Escort new faculty through library, quick tour
Making the Most of User Comments, Surveys, and Focus Groups with Qualitative Data Analysis
Patricia McDonald and Shana Gass, Albert S. Cook Library, Towson University
Prove value of libraries, improve services
quantitative vs. qualitative
open ended, free questions, observations
content analysis
Cook Library Assessment Committee
LibQUAL+ open comments at end of survey, ripe for content analysis methods
2500+ responses, 900 free comments
Observation study; focus groups
Steps toward content analysis:
What do you want to know?
Unitize data. break comment down into the smallest categorizable unit
Taxonomy: Categorize topics
Coding scheme: rules to apply taxonomy to comments
Interrater reliabilty: Make sure multiple graders agree on application of taxonomies
Coding
Report findings
Adapted Brown's LibQUAL taxonomy, 2005
Towson Taxonomy
nuggets of meaning within each comment
Software used:
NVivo <-- Towson used this
Atlas.ti
N6
MaxQDA
Weft QDA, open source
Demonstration of coding with NVivo
Interrater agreement of 80% == goal
Coding practice. Interesting disagreements
Content categories must be mutually exclusive, equivalent, and exhaustive
"Wordled" to create tag clouds
MS Word Frequency macro
Through the Users' Eyes
Elias Darraj, Yoni Glaser, Lucy Holman, University of Baltimore
Students are driven away from library Websites due to their unintuitive, complicated Natures
Discovery tools, multiple resources through single index. Preharvested metadata
ExLibris Primo; Encore; Vufind; Summon; WorldCat; Ebsco Discovery Service
USMAI institutions, UMD System, plus St. Mary's and Morgan State
UBalt Interaction Design and Information Architecture program
user research design, user centered design
task based design
three primary audiences: Undergrad; Grad; Faculty
compared four tools across three audience types
six tasks, scenarios based on task
21 participants
UBalt usability lab
eye tracking software -- duration and intensity of participant's gaze. Heat map of gaze. Cool! (yet I have no doubt this would NOT work with my eyes)
two known item searches; two topic searches; item save and retrieval
EDS interface: Intuitive; facets; simple. But "right-side" blindness for tools in right-side pane. Easy search for known items. Difficult for topic searches. Students still going to Google to get basic info about a book! even when the EDS detail screen of that book is in front of them!
Summon: Overwhelming main interface; pretty clear results page; inconsistent detail records. Very confusing how to save results.
Primo: Basic and Advanced on front screen. Clean interface. Simple, consistent interface.
Faceted search. filter based on attributes in collection. supports exploratory searching. Most relevant filters near top of screen. Provide many filtering options. Allow multiple selections of filters before updating. Display faceted search options on left side. Avoid jargon. Visual cues -- what you've clicked.
Summon: Overwhelming main interface; pretty clear results page; inconsistent detail records. Very confusing how to save results.
Need to mimic what Google has done.
Save/Retrieval functionality. Must be clear indication that something has been saved. Visual cues.
USMAI is going to go with EDS. Content and cost figure in. Go Live in Fall?
Google Plus or Google Minus
Patricia Anderson, Joel Shields, J. Shore, Julie Strange
NLM, AskUsNow, Western Maryland Consortium
Google+ more like Twitter than Facebook
Your Circles -- you control where your messages are going
Public and Limited posting
Google+ Hangouts
subject or audience or age-based Circles
sharing Circles with specific people or with other Circles
use as blogging tool
polling mechanism
Hangouts, plugin, similar to Adobe Connect. 10 people at a time in video window. Free video conferencing
Screen sharing, share documents, collaborative editing, meeting in Google Hangout
Hangout On Air feature: Broadcast push to public audience
automatically pushed to YouTube
Help Desk hangouts, Reference hangouts, support group hangouts
Security and privacy: SCARY STUFF. Photos automatically uploaded from cellphone via Google account.
Google reruns its algorithm every 45 minutes based on YOUR data
Custom, focused advertising, personalized based on what's in your Google account
There is a woman sitting next to me who is KNITTING! And when she's not knitting she's taking notes with a PENCIL and paper. This is somewhat distracting -- and the juxtaposition of this old school technology with the glow of the video streaming in from the Google+ Hangout in the front of the room was at first jarring.
And yet, I now find it somehow comforting.
Maybe I should take up whittling wooden fisherman figurines while loading the latest Ubuntu?!
Friday, May 4, 2012
Research Data Access & Preservation Summit 2012
Some highlights from the panels which were centered around the following five themes.
Data management plans and policies - Suzanne Allard with DataOne emphasized the need to focus on developing the tools scientists need in order to shift the culture, in particular she pointed out their continued focus and work on interoperabily and the need for everyone to work together on this. RUCore at Rutgers, built on Fedora, was showcased as one model for data management. They've integrated with the researcher's workflow system, created discipline specific portals, and are developing tools such as RUAnalytics that let researchers annotate video for example. The Texas Digital Library is also looking at workflow and has integrated their repository with the Open Journal System for their community.
Data Citation - It was noted that ASIS&T will be publishing best practices for citing data soon. DataCite's Mark Martin spoke to the confusion around citing data e.g. no real acknowledgement, processed and packed forms, which mirror was used, which edition or version of th data, where to get the data, etc. The idea of citation landing pages garnered a heated discussion ranging from concerns of whether you'd need landing pages for subsets of collections and others noting that that this starts to sounds like a MARC record system. Paul Uhlir from the National Academy of Sciences spoke about philosophical differences between the US and Europe in the approach to developing a citation style.
Curation Services and Models - David Minor of the University of California San Diego spoke to their experiences with high performance computing and storage management and recent pilots in data curation ranging observational data of the human brain to archeology to geological collections. Michael Witt explained how Purdue's PURR is being set up to support data management and the process they plan to put in place for managing data. I was one of the panelists for the Curation Services and Models panel and presented on the services we've been developing and building at Johns Hopkins. There was good interest in the work we're doing particulary in how we've scoped and modeled service provision within the JHU DMS.
Sustainability - The ArXiv.org presentation pointed out how their community is expanding into other scientific domains and the success of this resources. Models for financially sustaining ArXiv.org are being reviewed including a potential membership model. Dryad's Peggy Schaeffer talked about their financial model with involves working with publishers and charging them a fee everytime an author deposits data in Drayd in association with a publication. Fees are kept very low but so is the storage allocated for that fee.
Training Data Management Practitioners - Kirk Borne of George Masson University shared his experiences with training high school students in data mining and the importance of building these skills in young people. Peter Fox talked about the Computer Science Programs at Rensselaer Poloytechnic Institute the need to think about application themes by domain. Jian Qin of Syracuse University talked about their new online training opportunities in data services. She noted that data literacy is not just important for science students but also important for librarians.
Besides panels there was a lively poster session and lots of opportunities to network and learn from others. The models for data management and service provision range widely as the needs of researchers and communities differ from institution to institution. - Barbara Pralle
Wednesday, April 18, 2012
ASERL Spring Meeting
- simply give it away on a thumb drive or put it on a server
- distribute free through iBooks Store--minimal IP agreement with Apple
- charge for iBook on Store--more complex and more checks and balances about who owns copyright
Apple is touting the fact that you can update an iBook at any time in order to keep it current and accurate. Big discussion about how that affects references. Since there is no "edition" or versioning, fact checking for editors can become a nightmare. While the same is true when citing a website, most researchers would expect information in a "book" to remain constant or be updated with successive editions. The Apple rep just did not get it on that issue. The other huge drawback is the fact that the application is only available on the iPad.
Another selling point is that libraries could use the iBooks store to sell books created from their collections. Turn this into a money-making venture. While this bears watching, many of the attendees thought that this app is not a good fit for libraries at this time.
Libraries as Publishers: Current and Best Practices
Representives from the Scholar Commons at the University of South Florida and the Center for Digital Research and Scholarship at Columbia talked about the publishing services that they offer to their community.
Scholar Commons offers a suite of services that includes repository services, eScholarship services, OA publishing services, geoportal & data repository. They started by taking over a journal that was failing at the request from a faculty member. Things quickly grew from there. They use the bepress platform for publishing and help with creating the journal site and training editors. They did a cost study and found that the cost per article for their OA journals is about 10% of the cost of a subscription journal. They currently host six journals and are in discussions with a couple more.
Before taking on a new journal project, they look at three things:
- aim aligns with library
- peer reviewed
- health and active editorial profile
- platform software hosting: OJS, WordPress (most common), blogs and wikis
- consultation: digitization of back issues; copyright, licensing, author agreements; open access and other business models
- ISSN and domain acquisition
- print on demand
- design, content migration
- workshops on journal best practices