Encoded text of the artefact

Each of the four digitised pages has been encoded as a separate TEI-XML file. Select a page below to view its encoding. Each file can also be downloaded/ copied individually.

img1.xml
Download ↓
Loading...

Tags used & reasoning

How we handled the text

The original text was written in old Swedish and "Fraktur" (2026) typeface (this is the closest assumption we arrived at, after analysing fonts of the book). This presented a challenge since the language and script are foreign to myself (Sachini) and only partially comprehensible to other member (Erik) despite his native fluency of modern Swedish. Therefore, we used Gemini AI (Google, 2026) and Transcribus (READ-COOP SCE, 2026) tool to get the closest initial transcriptions. We did not rely on the accuracy of these transcriptions alone and each transcription was then manually cross-examined with the original digitised images. To further improve transcription accuracy, I (Sachini) downloaded and installed the Walbaum Fraktur typeface, which allowed me to type and compare individual letters directly against the original text.

After correcting and finalising the transcriptions, we used course materials and TEI guidelines to encode them into xml. The tags we used are listed on the table here. -->

Please note that we are focusing exclusively on Chapter 40, which begins on the bottom half of page 252 and ends at the top of page 255. As a result, image text that falls outside of this chapter has been omitted from the transcription.

What is gained and lost

After remediating the digitised images into TEI-XML, they now have gained searchability, machine-readability and interoperability with other heritage systems. These transcriptions can now be used to perform computational analysis (Dillen, 2026).

TEI-XML files lost the visual appearance of their corresponding digitised images. The unique Fraktur typeface conveys the historical significance. This "aura" is irretrievably lost in the tei-xml files (Bolter & Grusin, 1999).

TEI elements used
<teiHeader>Contains metadata for digitised page.
<titleStmt>Groups title of page and responsibility.
<respStmt>Name the person responsible for transcription.
<publicationStmt>Information about publisher, date, and license.
<sourceDesc>Information about source: The original printed book.
<msDesc>Location and ownership of the book.
<bibl>Bibliographic citation for the book.
<physDesc>Information about material and font/ typeface of the book.
<profileDesc>Describes the language the book was written.
<tagsDecl>Describes the taggs used.
<gap>Indicates a point where material has been omitted in the transcription.
<hi>Indicates a dropcap letter.
<lb/>Marks the beginning of a line.
<fw>Indicates a catchword/ header or page number (Werner, 2012).
<milestone>Indicates horizontal line.