Editing and Attributing Musical Texts: the Chansonnier du Roi and the Maritem Project

Camps, Jean-Baptiste
École nationale des chartes | PSL, Paris, France
jean-baptiste.camps@chartes.psl.eu

Chaillou-Amadieu, Christelle
CESCM / CNRS, Poitiers, France
christelle.chaillou.amadieu@univ-poitiers.fr

Mariotti, Viola
CESCM / CNRS, Poitiers, France
viola.mariotti.maritem@gmail.com

Saviotti, Federico
Università degli Studi di Pavia, Italy
federico.saviotti@unipv.it

Table of contents

This contribution examines the case of the Chansonnier du Roi, a very important 13th century lyrical manuscript, in the context of the ongoing maritem project. It presents a workflow for the edition and stylometry of musical texts, from text acquisition and encoding to the analysis of text and musical notations.

1. The Chansonnier du Roi and the maritem Project

Compiled during the second half of the 13th century, the Chansonnier du Roi (MS Paris, BnF, fr. 844) contains 602 lyrical compositions from different musical and literary traditions: profane songs of French trouvères and Occitan troubadours, French motets, instrumental works and Latin sacred compositions. Moreover, shortly after the original compilation, some additional pieces also from multiple origins (French rondeaux and motets entés, Occitan dansas and descortz) were transcribed in many of the blank pages and columns. Assembling several different repertoires in a uniform plan (further developed through additions), the Chansonnier du Roi is an ideal resource for the study not only of different and multilingual lyric traditions per se, but also of their unitary reception in the late 13th century Gallo-Romance area. This, together with a small but significant presence of lyrics from the 14th century, explains why we chose this manuscript as the object of our research.

The first purpose of the maritem project (anr) is to produce a dataset and digital edition of text and music, encoded in XML/TEI and XML/MEI. If the digital scholarly edition of texts is a well established practice, the field of edition of medieval music is on the other hand relatively new. The Corpus Monodicum project (Haug / Puppe 2020) developped a software, Monodi+, for the musical transcription of medieval latin songs in MEI. We plan to work with the software with a few adaptations and optimisations with the cooperation of the Corpus Monodicum team. The first challenge will be to link the textual edition (TEI) and the musical edition (MEI). In this aspect, the project is pioneer. The choice to work with only one manuscript was made with the perspective to create a prototype and to develop a methodology for the edition and the indexation of all musical chansonniers in different medieval languages (mainly Old French, Old Occitan and Old German). The edition of music and text will prove very useful to both musicologists and philologists, as both aspects will be addressed with equal accuracy.

It is our belief that involving both complementary constituents of medieval lyric songs will bring new results about the history of the codex, the languages, the link between musical composition and language, the history and link between the different traditions, and notably the attribution of the songs.

2. Data pipeline

Acquisition of the text is done with a pipeline that aims to fully integrate the contributions of human and artificial intelligence, in the spirit of digital philology (Andrews 2012). It builds on the workflow developped for another 13th century French manuscript, the Légendier BnF, fr. 412 (Camps et al. 2020).

Layout analysis was performed with Transkribus (Kahle 2017) default model, with some success for text regions and baselines (estim. F measure for baselines according to Transkribus: 0.71). Zones for music and illumination were added and typed manually. We plan to train a more specific model in the future to better detect musical notations and illumination.

The handwritten text recognition was performed using a model trained on data from MS fr. 412, with good results concerning the main hand (CER around 8%, WER 25%). The prediction was then fully corrected by a human expert, inside Transkribus. In the next steps, we aim to reuse and adapt the pipeline for automatic text segmentation, normalisation and lemmatisation (cf. Camps et al. 2020).

For now, the melodies are human-transcribed, but we hope to be able to train a model for music recognition.

3. Towards a scholarly digital edition

The future edition of the Chansonnier du Roi will be an electronic and interactive edition where it will be possible to consult and to query two different textual levels at the same time: the first one will be the allographetic and graphematic transcription of the manuscript (Robinson / Solopova 1993; Stutzmann 2011; Camps 2016), while the second one will be the normalized edition. Both levels will be accompanied by high definition images of all the pages of the manuscript.

Both transcriptions, allographetic and graphematic, are conceived to reflect different aspects of the text as it was written by medieval copyists. An allographetic transcription is a transcription whose goal is to “give access to every form of every letter or sign” (Stutzmann 2011). We followed the recommendations of the Medieval Unicode Font Initiative (Haugen 2015) in order to reproduce the formal Medieval letter variants and abreviations marks. Allographetic transcriptions respect as well the Medieval word segmentation, and punctuation. We decided to transcribe allographetically about one page and a half per copyist (about 20 copyists for the entire manuscript) in order to study the very special usus copiandi of each scribe of the Manuscrit du Roi, and appreciate the very different hands, both French and Italian, which have composed this precious manuscript. Compared to the former, graphematic transcriptions simply normalise variant letter forms.

The second level of the electronic edition of the Chansonnier du Roi is the editorial one, focused on the normalised edition of the lyrical texts, in order to propose a text approachable even for non specialised readers. It modernises the modern word segmentation, use of punctuation, accents and capital letters, and of course the resolution of all the abbreviation marks.

The edition follows the guidelines of the TEI (TEI Consortium 2020) for the text and MEI for musical notations (Music Encoding Initiative 2020). The melodies are transcribed with the software Monodi+ (Eipert et al. 2019). The musical transcription is easier than the text because the notation is clear and simple. The music edition has two levels: a diplomatic transcription with the keys, the signs and the presentation of the manuscrit; and a modernised transcription with a G key and a verse alignment.

4. Stylometric analysis of text and music

The availability of a complete transcription of this manuscript makes possible the stylometric analysis of the songs of the trouvères and trouveresses, at a level impossible until now. This is a critical issue, because disputed attributions are very numerous inside the Old French Lyrical tradition (Gatti 2019), yet it poses specific challenges to both traditional and stylometric approaches:

  1. individual components are very short (fig. 1);
  2. lyrical idiolect, based on a shared and elitist cultural tradition, is rather homogeneous, apparently leaving little space for personal features, though such a situation is not unheard of in the stylometry of Medieval and Early Modern texts (e.g. Camps / Cafiero 2013; 2019), but still poses significant challenges;
  3. attribution of the text and of the melody have rarely been addressed as a whole. Indeed, the musical elements (for example: musical curbe, modality, intervals), in connection with the texts could be an important contribution for the knowledge of songs attributions;
  4. the manuscript tradition creates noise, with both linguistic and substantial variants due to subsequent copy steps upstream of the manuscript which are quite difficult to retrace.

Figure 1: Distribution of the length in words of the texts

An automatic quantitative analysis of the scriptological data (Goebl 1975; Dees 1987), such as graphic and phonetic allotropes in different texts attributed to the same author or in different authorial corpora, will provide useful hints in the linguistically stratified French scriptae of the two main copyists. Thus, not only will it be possible to try to better determine their origins (it’s already acknowledged that one of them is generally Italian, but no previous study could determine with any plausibility the origin of the other, responsible for the main part of the chansonnier), but also to recover valuable insights about the linguistic habits of the authors themselves.

Specifically, the Chansonnier du Roi offers important documentation on the works of Thibaut de Champagne, perhaps the most prominent of all trouvères (Barbieri 1999). The attribution of several songs to Thibaut is still disputed (Wallensköld 1925; Callahan 2010).

In order to give new insights into these disputed attributions, we performed several stylometric analyses, using features robust to noise and short text length, in particular character 3-grams (Camps et al. 2020). Both exploratory and supervised analyses, the latter using SVM, were performed to shed more light on the attribution of these components.

precisrecallF_1
Not Thib.76.791.000.87
Thibaut1.000.460.63
GT \ PredNot Thib.Thib
Not Thib.860
Thibaut2622

Table 1: Metrics and confusion matrix for the leave-one-out training

For instance, we trained models on a corpus of 140 songs to distinguish Thibaut’s hand from a group of contemporary trouveres (Gace Brulé, Gautier de Dargies and Blondel de Nesle), with a leave-one-out approach. While the models do not attain a perfect accuracy (global 80.6%), precision reaches 100% for attribution to Thibaut (0.46% recall; table 1). These preliminary results (that we plan to substantially extend for the conference) seem to confirm the attribution to Thibaut of two very famous songs, Ausi com l’unicorne sui ( Linker 240,3; RS 2075) and Li dous penser et li dous souvenir (Linker 240,35; RS 1469) (table 2).

titleRSThibaut
Quant fine Amours me proie que je chant306
Sans atente de gueredon1867
Dame, li vostres fins amis1516
Ausi com l’unicorne sui2075X
Tres haute amours, qui tant s’est abessie1098
Li dous penser et li dous souvenir1469X

Table 2: Model results for a sample of disputed pieces

Appendix A

Bibliography
  1. Andrews, Tara (2012): "The third way: philology and critical edition for a digital age", in: Variants: The Journal of the European Society for Textual Scholarship 10 <http://boris.unibe.ch/43071/>.
  2. Barbieri, Luca (1999): "Note sul ‘Liederbuch’ di Thibaut de Champagne", in: Medioevo romanzo 23: 388-416.
  3. Barton, Louis W. G. (2002): "The neumes project: digital transcription of medieval chant manuscripts", in: Second International Conference on Web Delivering of Music, WEDELMUSIC 2002. Darmstadt 211–218 DOI: 10.1109/WDM.2002.1176213.
  4. Cafiero, Florian / Camps, Jean-Baptiste (2019): "Why Molière most likely did write his plays", in: Science Advances 5, 11. American Association for the Advancement of Science: eaax5489 DOI: 10.1126/sciadv.aax5489.
  5. Callahan, Christopher (2010): "Thibaut de Champagne and Disputed Attributions: The Case of MSS Bern, Burgerbibliothek 389 (C) and Paris, BnF fr. 1591(R)", in: Textual Cultures 5, 1. Bloomington: Indiana University Press 111–132 DOI: 10.2979/tex.2010.5.1.111.
  6. Camps, Jean-Baptiste (2016): La ‘Chanson d’Otinel’: édition complète du corpus manuscrit et prolégomènes à l’édition critique. Paris: Paris-Sorbonne PhD thesis, dir. Dominique Boutet DOI: 10.5281/zenodo.1116735. <https://halshs.archives-ouvertes.fr/tel-01664932>.
  7. Camps, Jean-Baptiste / Cafiero, Florian (2013): "Setting bounds in a homogeneous corpus: a methodological study applied to medieval literature", in: Revue Des Nouvelles Technologies de l’information (RNTI), SHS-1: 55–84 <https://halshs.archives-ouvertes.fr/halshs-00765651/>.
  8. Camps, Jean-Baptiste / Clérice, Thibault / Pinche, Ariane (2020): "Stylometry for Noisy Medieval Data: Evaluating Paul Meyer’s Hagiographic Hypothesis", in: ArXiv:2012.03845 [Cs] <http://arxiv.org/abs/2012.03845> [08.01.2021].
  9. Chaillou-Amadieu, Christelle (2016): "Le chant du poète, entre innovation et tradition", in: Saviotti, Federico (ed.): L’expression de l’identité dans la poésie lyrique romane, entre texte et musique. Milan: Presses universitaires de Milan 101-114.
  10. Chaillou-Amadieu, Christelle (2017): "Philologie et musicologie. Les variantes musicales dans les chansons de troubadours", in: Cazaux-Kowalski, Christelle / Chaillou-Amadieu, Christelle / Rillon-Marne, Anne-Zoe / Zinelli, Fabio (eds.): Les noces de Philologie et de musicologie. Textes et musiques du Moyen Âge. Paris: Classiques Garnier 69-95.
  11. Haug, Andreas / Puppe, Frank (2020): Corpus monodicum: Die einstimmige Musik des lateinischen Mittelalters, Digitale Ausgabe. Mainz / Würzburg: Alpha-Version <https://corpus-monodicum.de>.
  12. Dees, Anthonij (1987): Atlas des formes linguistiques des textes littéraires de l’ancien français. Tübingen: Niemeyer.
  13. Eipert, Tim et al. (2019): "Editor Support for Digital Editions of Medieval Monophonic Music", in: Proceedings of the 2nd International Workshop on Reading Music Systems. Delft <https://sites.google.com/view/worms2019/proceedings>.
  14. Gatti, Luca (2019): Repertorio delle attribuzioni discordanti nella lirica trovierica. Roma: Sapienza Università Editrice <http://www.editricesapienza.it/sites/default/files/5850_Gatti_Lirica_trovierica_interior_OA.pdf>.
  15. Goebl, Hans (1975): "Qu’est-ce que la scriptologie?", in: Medioevo romanzo 2: 3-43.
  16. Haugen, Odd Einar (ed.) (2015): MUFI Character Recommendation v. 4.0. Medieval Unicode Font Initiative <https://bora.uib.no/bora-xmlui/handle/1956/10699> [08.01.2021].
  17. Kahle, Philip / Colutto, Sebastian / Hackl, Günter / Mühlberger, Günter (2017): "Transkribus - A service platform for transcription, recognition and retrieval of historical documents", in: 2017 14th IAPR International Conference on Document Analysis and Recognition (ICDAR) 4. IEEE 19–24.
  18. Lavrentiev, Alexei (2007): Systèmes graphiques de manuscrits médiévaux et incunables français : ponctuation, segmentation, graphies. Actes de la journée d’étude de l’ENS LSH, 6 juin 2005. Chambéry: Université de Savoie.
  19. Lyrik des deutschen Mittelalters, database, <http://www.ldm-digital.de>.
  20. Linker, Robert White (1979): A Bibliography of Old French Lyrics. University, MS: Romance Monographs.
  21. Music Encoding Initiative (2020): Guidelines. Mainz <https://music-encoding.org/> [08.01.2021].
  22. Refrain, database <http://refrain.ac.uk>.
  23. Robinson, Peter / Solopova, Elizabeth (1993): "Guidelines for Transcription of the Manuscripts of the Wife of Bath’s Prologue", in: The Canterbury Tales Project Occasional Papers 1: 19–52 <http://www.canterburytalesproject.org/pubs/transguide-MI.pdf>.
  24. RS (1955): G. Raynauds Bibliographie des altfranzösischen Liedes(neu bearbeitet und ergänzt von Hans Spanke). Leiden: Brill.
  25. Stutzmann, Dominique (2011): "Paléographie statistique pour décrire, identifier, dater… Normaliser pour coopérer et aller plus loin?", in: Fischer, Franz / Fritz, Christiane / Vogeler, Georg (eds): Kodikologie und Paläographie im digitalen Zeitalter 2 / Codicology and Palaeography in the Digital Age 2 (= Schriften des Instituts für Dokumentologie und Editorik 3). Norderstedt: Books on Demand 247–277 <https://halshs.archives-ouvertes.fr/halshs-00596970/>.
  26. TEI Consortium (2020): TEI P5: Guidelines for Electronic Text Encoding and Interchange <https://tei-c.org/Guidelines/> [09.05.2015].
  27. Troubadour Melodies Database <http://www.troubadourmelodies.org/melodies>
  28. Wallensköld, Axel (1925): Les chansons de Thibaut de Champagne, roi de Navarre. Édition critique. Paris: Librairie ancienne Édouar Champion <https://gallica.bnf.fr/ark:/12148/bpt6k53236> [08.01.2021).
  29. Wick Christoph / Puppe, Frank (2019): "OMMR4all - a Semiautomatic Online Editor for Medieval Music Notations", in: 2nd International Workshop on Reading Music Systems WoRMS.