TEI 2026
Creating Connections, Unsettling Practices
August 10-14, 2026
University of British Columbia, Vancouver, BC, Canada
Conference Agenda
Overview and details of the sessions of this conference. Please select a date or location to show only sessions at that day or location. Please select a single session for detailed view (with abstracts and downloads if available).
|
Daily Overview |
| Session | |
|
Public Workshop 2: Navigating and Processing Data from the TEI with XPath and XSLT Location: Koerner Digital Scholarship Lab | |
| Presentation 1 | |
ID: 143
/ Public Workshop 2: 1
Workshop Keywords: XPath, XSLT, HTML, SVG, CSV, TSV, data analysis, data visualization Navigating and Processing Data from the TEI with XPath and XSLT 1: Penn State Erie, United States of America; 2: University of Graz, Austria; 3: Maynooth University, Ireland TEI markup provides structures that are particularly useful for processing data beyond what we can do with so-called “plain text”. A full-day workshop allows us plenty of time to teach the pull-processing of data from XML/TEI with simple, reusable XSLT templates to represent in simple TSV/CSV, HTML tables/charts, and (if time!) simple SVG graphics. Knowing how to locate and explore data in your encoding can help to learn how to work with TEI and XML generally. This half-day workshop is designed for people who have some experience with TEI and seek to learn how to work with XML markup for analysis and research. Participants will gain a working, practical knowledge of the query language XPath and the transformation language XSLT, and learn how these can help to reduce reliance on software, packages and plugins that may become obsolete without warning. Further, XSLT's functional programming can serve as a way of articulating research questions around a document data model expressed in XML.
The emphasis of our workshop is ”pull-processing”, that is, extracting data and metadata from markup documents for analysis, as opposed to providing the reading view of a digital scholarly edition. Markup in documents supplies structures and contexts that are especially useful for processing data, beyond what we can do with so-called "plain text". We will demonstrate some basic XPath navigation and calculation functions, and then show how XPath is applied in XSLT templates to address specific nodes that hold data of interest for visualization.
We will process TEI documents composed in various languages represented by our workshop members' projects, to show that the code we write is transferable to multiple projects across language and cultural borders.
Participants will learn how to "pull" data from TEI and output text formats required for simple online tools, where the structure of the output data is transferable to many different online calculation programs and amenable to statistical processing. During the workshop we will produce some simple structured documents for storing, sharing, and visualizing data: HTML lists and tables as well as plain text tabulated data (CSV or TSV files), and (if we have time) simple SVG bar or line graphs.
We hope to process some participant-supplied XML before, during, and after the workshop. We will carefully document the XSLT that we supply during the workshop to assist participants with revising and adapting the code to their own projects.
The workshop material - including documentation, exercise examples and solutions - will be made available in a publicly and permanently accessible GitHub repository. Outline
Room/Materials Required
| |
