Transcription+

ISO 24624:2016

Language resource management — Transcription of spoken language

Welcome to the Transcription+ homepage!

This site provides information on the Standard ISO 24624:2016 Language resource management — Transcription of spoken language, a standard published in 2016 for digital representation of transcriptions of audiovisual language data. The standard is based on the Guidelines of the Text Encoding Initiative (TEI). It has been succesfully applied in various research projects and research data infrastructures.

The Transcription+ project is a cooperation of the University of Duisburg-Essen and the Text+ consortium of the German National Research Data Infrastructure NFDI. The Academy of Sciences and Humanities in Hamburg, the Center for Sustainable Research Data Management at the University of Hamburg and the Hamburg Center for Language Corpora (HZSK) are partners from the Text+ consortium supporting Transcription+.

The project's aim is to improve documentation and tool support for the standard. To this end, existing documentation will be revised, completed and published here. Resources like XML schemas, XSLT stylesheets and other code will be gathered from different sources, harmonized, documented and assembled in a GitHub repository. An improved version of the multilingual EXMARaLDA demo corpus will be made available in the standard format. The demo corpus will also be made accessible through an instance of the ZuMult platform. An extended version of the TEILicht web services, serving to convert, normalize, annotate or visualize transcript documents in the ISO/TEI format, will be developed.

The project started in January 2026 and will be completed by September 2026. This website is work in progress. The source code for the website is hosted on GitHub. Click on any of the cards below to learn more about the current status of the respective work items. Feel free to contact me with any questions or suggestions.