Much of the work on Perseus is focused on collecting and converting the
data on which the project is based. At the same time, it is necessary to
provide means of access to the information, in order to make it usable,
and them to investigate how it is used. As we learn more about what
students and scholars from different backgrounds do with Perseus, we can
adjust our data collection, and also modify the system to accommodate
them. In creating a delivery system for general use, we have tried to
avoid favoring any one type of use by allowing multiple forms of access
to and navigation through the system.
The way text is handled exemplifies some of these principles. All text
in Perseus is tagged using SGML, following the guidelines of the Text
Encoding Initiative (TEI). This markup is used to index the text, and
process it so that it can be imported into HyperCard. No SGML markup
remains in the text that reaches the user, because currently it would be
too expensive to create a system that acts on SGML in real time.
However, the regularity provided by SGML is essential for verifying the
content of the texts, and greatly speeds all the processing performed on
them. The fact that the texts exist in SGML ensures that they will be
relatively easy to port to different hardware and software, and so will
outlast the current delivery platform. Finally, the SGML markup
incorporates existing canonical reference systems (chapter, verse, line,
etc.); indexing and navigation are based on these features. This ensures
that the same canonical reference will always resolve to the same point
within a text, and that all versions of our texts, regardless of delivery
platform (even paper printouts) will function the same way.
In order to provide tools for users, the text is processed by a
morphological analyzer, and the results are stored in a database.
Together with the index, the Greek-English Lexicon, and the index of all
the English words in the definitions of the lexicon, the morphological
analyses comprise a set of linguistic tools that allow users of all
levels to work with the textual information, and to accomplish different
tasks. For example, students who read no Greek may explore a concept as
it appears in Greek texts by using the English-Greek index, and then
looking up works in the texts and translations, or scholars may do
detailed morphological studies of word use by using the morphological
analyses of the texts. Because these tools were not designed for any one
use, the same tools and the same data can be used by both students and
scholars.
NOTES:
(5) Perseus is based at Harvard University, with collaborators at
several other universities. The project has been funded primarily
by the Annenberg/CPB Project, as well as by Harvard University,
Apple Computer, and others. It is published by Yale University
Press. Perseus runs on Macintosh computers, under the HyperCard
program.
Eric CALALUCA
Public-domain text, read in full here on John Shaqi.
Reviews
Reviews
No reviews yet
Be the first to share your thoughts on this work.
Elsewhere in the archive
Join the Discussion
Join the discussion
Sign in to leave a comment or review.
Sign InorCreate an account