TEI now in Wikipedia

From TEI-L, a mention of the new Wikipedia entry on the TEI.

Posted in General | Leave a comment

Firefox Search Plugins for Classics

Classic end-of-term displacement activity: I’ve created a bunch of Firefox search plugins for open content sites (Perseus, the Stoa, the Suda on Line, and the Latin Library), to accompany one for Perseus I found ready-made at the Firefox site. Just unzip the set that you’ll find here, and put the resulting files in locations like the following:

  • OS X — /Applications/Firefox.app/Contents/MacOS/searchplugins/
  • Linux — /usr/share/firefox/searchplugins/
  • Windows — \Program Files\Mozilla\Firefox\searchplugins\

I’m happy to archive more of these if anyone feels like contributing. They aren’t hard to create.

Posted in General | Leave a comment

Taking a wrong turn at the APA

An earlier post to this blog summarizes new NEH-funded work on the problems of digitizing Latin incunabula. The project will disseminate its results very broadly, through publication of data on freely accessible sites like Perseus and in other university digital libraries, application of extremely liberal Creative Commons licenses to program code, and so forth. In taking this approach, Rydberg-Cox and his colleagues have lots of company: a strong consensus has long since formed among classicists with the greatest relevant expertise that Open Access methods represent “best practice” in our field. Experiences from a full decade of scholarly electronic publication online have demonstrated that we can now reach a huge international audience that’s eager to use the resources (texts, images, tools, analyses) we can make available concerning the ancient world.

Against that background, the recent APA decision to create a members-only portion of its web site strikes me as an obvious mistake, to the extent that the APA pushes this as a repository for additional members-only scholarly content. I believe that what the APA has done represents an unimaginative and inadequate response to the opportunities afforded us in our networked world. In the first place, the claim that limiting access to the database of members “gives current members a strong incentive to remain members” is hard to take seriously. Also, in restricting TAPA to this closed portion of the site, the members-only ploy creates a needless dichotomy between a tiny group of insiders with privileged access to information and outsiders (= the entire world) without that. Could the APA come up with any better way than this to perpetuate the ideal of Classics as a 19th century gentlemen’s club? Then too, by putting TAPA (wait, members already get that, right?) and discounts on books from OUP behind the firewall, the APA privileges the most conventional and traditional forms of scholarly communication in Classics but relegates to a lesser status the wonderful variety of innovative work being done online by many of its own members. Finally, the APA looks to be running against an accelerating trend in other disciplines, especially the natural sciences (read about the proposed new requirement for Public Access from the NIH here).

[Update: The New York Times article (14 December 2004) on new digitization projects notes that “The Google effort and others like it that are already under way, including projects by the Library of Congress to put selections of its best holdings online, are part of a trend to potentially democratize access to information that has long been available to only small, select groups of students and scholars.” Compare!]

A second update: Please don’t miss the thoughtful remarks of Alun Salt at The Undoctored Past regarding this post.

A third update: Interesting and important continuations of this discussion at Blogographos, by David Meadows and Alun Salt — not to be missed. I hope to have more to say about it all when I am done with current traveling.

PS: You’ll notice that the comments form below is turned off, simply because it’s hard to prevent blog comments sections from filling up with spam, and I don’t want to waste time clearing it out. But I’m always glad to hear people’s thoughts on this important topic. You can reach me at scaife–AT–gmail.com.

Posted in General | Leave a comment

CHS summer workshops for graduate students

The Center for Hellenic Studies invites applications to two upcoming summer seminars on Greek scholarship and electronic publication. The first seminar (led by professors Casey Dué and Mary Ebbott) will be on “Homer: Research on Homeric Poetry, Emphasizing Textual Criticism” and the second (led by professors Kent Rigsby and Joshua Sosin) will be on “Epigraphy: Greek Inscriptions, Introductions, Methods and Research.” In between the two there will be a common session on “Online Publishing Technologies and Sharing Results.” For this session,

Christopher Blackwell and several guest lecturers will work with participants on state-of-the-art technologies and methods for publishing scholarship in electronic media. Topics will include eXtensible Markup Language (XML), the Text Encoding Initiative’s standards for marking up humanistic texts in XML, the Unicode standard for representing the characters and symbols, the Classical Text Services protocol for distributing texts and fragments of them, and transforming XML documents for print and electronic publication using eXtensible Stylesheet Language Transformations (XSLT).

Posted in General | Leave a comment

Common Sense from Tim Wu

The Future of Digital Media is “a two-month series, sponsored by Orb, that explores how the empowerment of the consumer over his or her media experience, coupled with technological innovation that’s broadly democratizing media creation, is leading to a revolution in the way people access, consume, share and remake content.” Now the series offers an interesting interview with Tim Wu, an associate professor at University of Virginia Law School, who teaches intellectual property and international trade.

My vision of copyright will sound conventional: I think copyright law should serve authors and consumers. But that turns out to be a radical view. Because if we took those ideals of copyright seriously, as opposed to paying them lip service, the law would look a lot different than it does today.

More here.

Posted in General | Leave a comment

Addressing the Problems of Digitizing Latin Incunables

Congratulations to Jeff Rydberg-Cox and others at the University of Missouri at Kansas-City for winning a substantial new grant from NEH Preservation and Access. Here is the summary of the proposal; be sure to note the smart plans for open dissemination of tools and results at the end:

Addressing the Problems of Digitizing Latin Incunables

Early books printed in Latin are a major component of our early modern cultural heritage. Before 1600, considerably more than half of the books printed in England alone were printed in Latin, as were the majority of books traded at international book fairs and marketed internationally. The ability to create digital editions of these texts is, therefore, essential for preserving the greater portion of the intellectual heritage of the early modern period. Digitization of these books, however, poses unique and difficult problems: characters and ligatures are printed using graphs not represented in ASCII or Unicode, figures and pictures that are essential for understanding a passage are embedded within the texts, words that carry from one line to the next may not be hyphenated, and common words and letter combinations are abbreviated with a system of brevigraphs based on medieval handwritten manuscripts. In recent years, projects such as the Making of America have developed techniques for rapid and cost-effective digitization of large corpora of printed works from the nineteenth and twentieth centuries, while Early English Books Online and the Text Creation Partnership have addressed problems of early books printed in English. At the same time, efforts such as the Newton Project and the Digital Scriptorium have developed extensive knowledge about best practices for transcribing and cataloging manuscript material. The unique problems of Latin incunables, however, still remain to be addressed. Our project will have the following specific deliverables:

  1. Digital facsimile editions of a collection of incunables containing texts by Al-Qabisi, Bernard of Gordon, Sebastian Brant, Isidore of Seville, Petrarch, Pliny the Elder, Suetonius, and Jacobus de Voragine. These facsimile editions represent a wide variety of scientific and literary Latin from many time periods and geographic locations, all published in the first fifty years of printing. They will serve both as testbeds for developing our tools and also as demonstrations of the results that we will be able to achieve.
  2. Tools that can automatically or semi-automatically address the typographical difficulties posed by early printed works including:
    1. Tools for the automatic identification of abbreviations and broken words;
    2. Integration of these tools with a text editor allowing for interactive editing and disambiguation of uncertain abbreviations;
    3. A digitized edition of a dictionary essential for reading fifteenth-century Latin based on Du Cange’s standard medieval Latin dictionary;
    4. Extremely flexible look-up tools for dealing with the wide variety of orthographic variation in early printed Latin texts.
  3. Guidelines for data entry and encoding brevigraphs in early printed Latin texts that can be shared with others digitizing similar material.

Results of our work will be disseminated in several ways.

  • First, we will publish high resolution images of our early printed books alongside digital transcriptions on the web using the software infrastructure developed by the Perseus project (https://www.perseus.tufts.edu/).
  • Second, we will return our TEI-conformant XML transcriptions of the texts to the libraries that provide us the images of the books so that they can be disseminated via their own web sites as well as through our e-publication infrastructure.
  • Third, we will release all tools for public use both via the internet and as stand-alone applications so scholars in rare books rooms without easy internet access will be able to use them.
  • Fourth, the source code for our tools will be made available to any interested researcher under a Creative Commons license, allowing other scholars to adapt and extend them for their own work.
  • Finally, we will conduct a detailed analysis of both our workflow and the ways that users interact with our digital texts so that we can provide a clear understanding of the technical requirements for building large digital collections of rare and complicated books. Our ultimate goal will be to produce documentation that details data entry methods and encoding standards for early printed Latin works so that librarians and scholars can digitize their own early printed holdings.
Posted in General | Leave a comment

December issue of SPARC Open Access Newsletter now available

In addition to the usual round-up of news from the past month, it takes a close look at the Congressional approval of the NIH public access plan and the UK government response to the open-access recommendations from the House of Commons Science and Technology Committee. Among the news stories given shorter takes are a series of national OA initiatives launched in November, the Kaufman-Wills study of open-access journals, and Google Scholar.

December issue:

https://www.earlham.edu/~peters/fos/newsletter/12-02-04.htm

Posted in General | Leave a comment

Topic Maps: Searching Smarter, Not Harder

Wired has a short article about topic maps; here’s an excerpt:

Databases and search engines provide instantaneous access to endless information about anyone or anything, but the search results often include as many misses as hits. To generate more-relevant answers, organizations including the federal government are using topic maps to index their data.

Topic maps are smart indices that improve search capabilities by categorizing terms based on their relationships with other things. For example, William Shakespeare is a topic that would be mapped to essays about him, his plays and his famous quotes.

Organizing content with topic maps provides context for words that can have multiple meanings, according to Patrick Durusau, chairman of a topic maps technical committee at OASIS, the Organization for the Advancement of Structured Information Standards.

For example, searching Google for “Franz Ferdinand” mixes results for the alternate rock group and the doomed Austrian archduke for whom the group is named. If topic maps were used to organize the data, the musical and historical links would be separated, Durusau said. “The payoff (of topic maps) from the user standpoint is that you are no longer confronted with everything in the world that is known about the subject,” Durusau said.

A good deal more on this in Steve Pepper, “The Tao of Topic Maps.”

Posted in General | Leave a comment

The Digital Encyclopedia blog gets “Suber’d”

Says Chris Blackall of the welcome new blog Digital Encyclopedia (about pay-for-view and open-access digital encyclopedias):

Suber’d

* Definition: To be blogged in Peter Suber’s Open Access News; the open access equivalent of being slashdotted, that is, mentioned in https://slashdot.org/
* Usage: “Heavens, I’ve been Suber’d!”

I’ve only had this blog going for day or so and Peter Suber has already spotted it. The man has mystical powers for finding stuff.

Another eagle-eyed reader picked up a silly error I made in one of my first posts. And I just thought I was having a conversation with myself.

The Web is an amazing place—really.

Posted in General | Leave a comment

new from Edward Ayers: The Academic Culture and the IT Culture

The Academic Culture and the IT Culture: Their Effect on Teaching and Scholarship

A year ago, my colleague Charles Grisham and I wrote an EDUCAUSE Review article entitled “Why IT Has Not Paid Off As We Hoped (Yet).” In short, we argued that information technology has not yet transformed higher education because the areas of teaching and scholarship, the “heart” of colleges and universities, have remained relatively untouched by the new technologies. In this article, I’d like to continue the discussion and also go further, exploring not only why these two areas continue to be, for the most part, resistant to the changes but also how technology can successfully address these core missions of higher education.

Posted in General | Leave a comment

Inscriptions from the land of Israel

I am writing to announce a new web site, “Inscriptions from the Land of Israel.” The primary goal of this site is to create a searchable database of inscriptions, along with their contextual information (e.g., images and geographical data), of published inscriptions that are roughly within the geographical boundaries of the modern State of Israel and that date from between 500 BCE – 640 CE. The database can be searched according to a broad range of criteria, and the text of the inscriptions can be searched in both the original language (Greek, Latin, Hebrew, Aramaic) and English translation. The site also includes bibliographies and research and teaching resources connected to these inscriptions.

Presently in the database are only some inscriptions from Caeserea and Hammat Gader (none in Hebrew). It will be regularly updated.

The site can be found at:

https://www.stg.brown.edu/projects/Inscriptions/

Your comments are welcome and can be sent to Michael_Satlow@Brown.edu

Posted in General | Leave a comment

a new TEI publishing infrastructure

Here’s an e-mail to the TEI-L list from Eric Lease Morgan of Notre Dame:

In my copious spare time, I wrote a set of object oriented Perl scripts to manage TEI files — my TEI publisher:

https://infomotions.com/musings/tei-publisher/

The system is really a relational database application with a Web front-end. There are tables for authors, subjects, TEI templates, XSL stylesheets, and texts. I use these tables to manage authority lists, controlled vocabulary terms, TEI skeletons, transformation files, the various TEI meta data, and of course the content of my writings such as articles, conference presentations, software, and travel logs.

Once data entry for a particular work has been done, I use the system to build the work, save it to the file system, and then transform the work into something readable for the Web. The underlying CSS files make the collection easily navigable as well as nicely printable. I also take advantage of an indexer (swish-e) to enable robust searching. Because everything is in a database, it is easy to create reports against the content. I create on-the-fly reports for subject lists, title indexes, lists of texts by date, as well as sets of static OAI files for harvesting. Here is the process I used to create a work:

  1. Have an idea.
  2. Write it down.
  3. Mark it up in TEI.
  4. Assign subject terms
  5. Make sure they are in the database.
  6. Add the TEI to the database — do data entry.
  7. Build the file.
  8. Check it for validity.
  9. Transform it into XHTML.
  10. Check it for validity.
  11. Index the entire corpus.
  12. Create OAI reports.
  13. Go to Step #1.

The system is NOT a TEI/XML editor. I use my text editor (BBEdit plus a locally created “glossary”) to mark up the body of my TEI files. Everything is placed into fields of XHTML forms and thus into fields of the database. The system took me about three weeks of very hard and very concentrated work to create, but now I can pump out a rich, consistently formatted, valid TEI and XHTML file in about thirty minutes or less.

Through the process of creating the publisher I had to decide what TEI elements to incorporate and to what degree. I believe one of the challenges of using TEI is deciding on these issues. All of my TEI files are available directly from each page.

Posted in General | Leave a comment

Google Scholar

UPDATE: Here’s a valuable overview and critique of this new service.

from Peter Suber’s indispensable Open Access News blog:

Tomorrow Google will launch the beta version of Google Scholar, although it is online today for use. From the press release: ‘[W]e are excited to announce Google Scholar, a free search service that helps users find scholarly literature such as peer-reviewed papers, theses, books, preprints, abstracts, and technical reports. This service will be available tomorrow morning….Like Google Web Search, Google Scholar orders search results by relevancy to ensure the most useful references appear at the top of the page. This ranking takes into account the full text of each article as well as the article’s author, the publication in which the article appeared, and how often it has been cited in scholarly literature….Whenever possible, Google searches across the full text of a paper, not just the abstract….Google Scholar offers relevant results for a wide range of scholarly materials including research that isn’t yet online. For instance much of Einstein’s work isn’t online, but it is heavily cited by other researchers. Google Scholar leverages these citations to make users aware of important papers or books that are not online, yet may be available in their local library.’

(PS: This is an important development. It will make OA literature even more visible and retrievable than it already is. It will give authors new incentives to make their work OA. It will help readers find what they need. Because it indexes work that is not online, even non-OA publishers will have an incentive to participate, making it more and comprehensive and useful. When you run a search, Google Search labels each hit by the number of citations it has, presumably from other works in the index. It also lets you click through to a new page showing just those citing works. Authors and publishers: see the FAQ for instructions on how to make sure that your work is included.)

Posted in General | Leave a comment

Two new collections in the Stoa Image Gallery

The Ancient World Mappping Center has begun to archive its collection of digital images on www.stoa.org, at https://www.stoa.org/gallery/awmc. And Mark Lehman has begun The Ruth and Louise McCollum Memorial Collection of Ancient Coins at https://www.stoa.org/gallery/lehman.

Posted in General | Leave a comment

Science Commons

Neel Smith just alerted me to Science Commons, a new project of Creative Commons set to launch on January 1, 2005:

The mission of Science Commons is to encourage scientific innovation by making it easier for scientists, universities, and industries to use literature, data, and other scientific intellectual property and to share their knowledge with others. Science Commons works within current copyright and patent law to promote legal and technical mechanisms that remove barriers to sharing.

Posted in General | Leave a comment