scieee AI-readable full text Open interactive document viewer

You're Collating Just Fine and Other Lies You've Been Telling Yourself

Bordalejo, Barbara; Vázquez, Adam

Abstract

Although textual scholars agree that collation is a crucial component of the editing process, it often goes undefined and only briefly explained. This article defines the term, explains different kinds of collation, and explores some of its applications. We emphasize stemmatology and medieval textual traditions. By drawing from editorial examples and the theoretical frameworks of projects centred on works such as the Canterbury Tales, Troilus and Criseyde, Dante’s Commedia and the Greek New Testament, the article seeks to compare manual and computer-assisted approaches to collation methods. We delineate the scope of this activity and argue that computer-assisted collation minimizes the risk of missing out on relevant data. We examine the advantages of full-text collation over sample collation and conclude that no decisions about stemmatically significant variation can be made a priory and that variant distribution is the major factor weighing on significance.

Full text

1 You’re Collating Just Fine and Other Lies You’ve Been Telling Yourself by Barbara Bordalejo and Adam Vázquez 1. Introduction Although most theorists and textual scholars refer to collation in one way or another, they do so in passing, as if anyone coming upon their texts would understand unequivocally their meaning, and no one should ever require further explanations. In this article, we briefly examine the concepts behind the word collation, with a focus on textual collation, stressing the fundamental considerations for the optimization of computer-assisted collation. Because our research interest resides in the investigation of textual filiation, this article emphasizes the stemmatological purposes of our collations, describing and questioning some procedures. To conclude, we restate that editions depend on a limited number of variants and that how researchers identify and select the variants through collation affects their understanding of textual relationships and their perception of the text. In researching this article, we found that a significant number of scholars write the word collation and emphasize the importance of the process without going into further details about the concept. We are not the first ones to point this out. In a piece published in 2017, Elena Spadini states: “If the reasons why we collate are well known, the way we do it, especially when we do it manually, is less documented: handbooks and essays seem to take for granted this delicate task or summarize it in a couple of sentences.” (Spadini 2017, 245). Although Spadini is not strictly right in this assertion (see our discussion of David C. Parker’s advice for manual collation below), the spirit of what she writes resonates with anyone researching collation theory. In 2016, Tara Andrews presented, as part of the preliminary DiXiT activities before the conference of the European Society for Textual Scholarship in Antwerp, a paper which was 2 eventually printed (in the same volume where the article by Spadini can be found) under the title “What we talk about when we talk about collation” (Andrews 2017). With its four printed pages, this is one of the most substantial considerations surrounding collation that we have been able to find. Andrews relies on the entry for “Collation,” as found in the Lexicon of Scholarly Editing (“Collation” 2013), which partly explains the carnivalesque heterogeneity of her references. The authors of record are as diverse as Grésillon, Hockey, and Plachta. They include Biblical scholars (Colwell and Tune), Victorianists (Shillingsburg), Modernists (Eggert), and Armeniologists (Andrews herself). And yet, as Andrews seems to acknowledge, the mere mention of the word collation does not imply an explicit definition of the concept but rather the implicit understanding that scholars know what it is and how to carry it out. Indeed, more often than not, authors mention the process without pausing to reflect on the meaning of collation in the context of their own work. It is left to others to extrapolate what is being said and what is its context. Like the authors of the Lexicon, we have also found references to collation in the context of textual scholarship (Blecua 1983; Gabler 2007; Parker 2008; Waltz 2013; Trovato and Reeve 2014; Bordalejo et al. 2014; Bordalejo 2014; Driscoll and Pierazzo 2016; Bordalejo 2018; Fischer 2019), the vast majority of which live up to Spadini’s description of articles that make a note of collation but never elaborate. 2. What is collation? To collate is to compare by close examination. The word, however, has different specialized meanings, even among textual scholars. One can collate books or collate texts; one may speak of horizontal or vertical collation (Williams and Abbott 2009, 92). It is also possible to collate documents. Although these terms are not in any way obscure for specialists, it is reasonable to include them briefly as part of this article for two reasons: first, they could serve as a reference point for the textually curious; and second, they clarify exactly where the emphasis of our argument lies. Some textual scholars are mainly bibliographers, i.e. their main interest resides in the physical characteristics of the book. Like book historians, they seek to learn about the book as a material 3 object. When bibliographers use the term collation, they are talking about bibliographical description. In the context of analytical and descriptive bibliography, collation refers to the accounting for the quire structure that constitutes the physical form of the book (McKerrow 1927; Greetham 1994; Bowers 1995; Gaskell 2000). This is expressed in a collation formula, “a shorthand note of all the gatherings, individual leaves, and cancels as they occur in the ideal copy” (Gaskell 2000, 328). Bibliographical collation, as fascinating as it is, does not concern us for the purposes of this article. In textual criticism, collation refers to the systematic comparison of two or more texts with the aid of a base-text or without one. Those texts could have come to be by different means: copied by scribes, printed in a manual press, or reproduced by a mechanic press. There are two types of text collation with a very different focus: vertical and horizontal. Horizontal collation occurs while one is comparing different instances of the same print, a technique used by bibliographers to compare different copies of printed books, often using optical collators but now also replicating the optical processes by digital means. The objective is to detect differences that might shed light on the material history of the print to detect stop-press variants or resettings. The variants uncovered by this type of collation are not transmissional (as is the case with manuscripts copied by scribes) but revisional if errors are detected during the printing process, or remedial, when they are the result of a forme resetting due to accidents occurred during the book production process (Plachta 1995, 504-506). An example of revisional variation can be found in Mari Agata’s research. Agata detected stoppress variants in the Gutenberg Bible from which she inferred that the paper print predated the vellum print (Agata 2006; 2011). It makes sense, as the less expensive material was used as a sort of preliminary state before the more expensive production was undertaken. For her research, Agata used digital methods, which included semi-transparent images in Photoshop and fast animation alternating images in Macromedia Director. The first method emulated what purely optical collators (McLeod’s or Hailey’s collators) would do, while the latter produced results similar to the optomechanical alternating lights of the Hinman collator (Hinman 1955). 4 Vertical collation investigates successive textual stages, usually within manuscript traditions or pre-print documents, instances likely to present a less than clear chronology to guide research. There are different reasons to carry out vertical collations, which can be related to linguistic or textual matters. This article focuses on vertical collation and considers some of its potential applications with a particular emphasis on collation for stemmatological and textual analysis. Nearly thirty years of research by the Canterbury Tales Project seek to achieve a thorough understanding of its textual transmission. This complex task requires many steps of which transcription and collation are fundamental. 3. Defining Variation In order to compare texts, we have to define what is a variant. According to Vinaver’s analysis of the “movements” that constitute the act of copying, scribal variants come to be if any of the following steps is ineffective: “(a) the reading of the text; (b) the passage of the eye from the text to the copy; (c) the writing of the copy; and (d) the passage of the eye from the copy back to the text” (Vinaver 1976, 142; Blecua 1983, 16-17). Then, an initial categorization of the nature of variants resides in the now-classic division of substantive and accidental. W.W. Greg coined the terms in his well-known essay, “The Rationale of Copy-Text” (Greg 1950, 21). According to Greg, substantive variants are ones that “affect the author’s meaning,” this division responds to the way in which scribes or compositors “may in general be expected to react” (Greg 1950, 21). Greg posits that one can assume the aim with substantives will be to reproduce exactly those of their copy, while scribes or compositors “will normally follow their own habits or inclination” with accidentals (Greg 1950, 22). João A. Hansen in Volume 5 of his edition to Gregório de Matos e Guerra’s poetry, expands on Greg’s thoughts and hypothesizes that accidentals are more likely to be changed in the copying process because they are not considered to be as related to the author’s will as substantives (Hansen and Moreira 2013). Daniel Paul O’Donnell adapts Greg’s for his edition of Cædmon’s Hymn, where he divides them in orthographic, substantive, and (potentially) stemmatically significant (O’Donnell [2005] 2018, §7.6 to §7.9). Most textual criticism manuals agree with this distinction, yet when it comes to 5 editing projects, the definition of a variant changes. Moreover, other disciplines deal with similar issues, although their goals are different. Linguistics corpora exemplify such disciplines (See below). Greg and others like Ben Salemans seek to present a definition of variation, a priori. They are trying to establish principles that allow anyone to understand what a variant is and apply that notion of variant to any text. This process of defining variation a priori results either in a series of vague suggestions or in a prescription which will not be applicable to every textual tradition. Salemans lists characteristics of non-significant variants in his Building Stemmas with the Computer in a Cladistic, Neo-Lachmannian, Way: differences in capitalization, orthographic variants, dialectological variants, punctuation, word separation, difference in clause headers, ungrammatical sentences, nonsense readings, evident copy mistakes (Karel de Grote vs Krl de Grote), names, archaisms, frequently used word, synonymous parallelism, and inflectional parallelism (Salemans 2000, 68-70). Paolo Trovato celebrates Salemans’ ability to distinguish between variants, that are numerous, polygenetic, and irrelevant, and significant errors, which, according to him, are as a rule few, can derive from previous copies, thus, are useful for the construction of a stemma (Trovato and Reeve 2014, 110). After analyzing Salemans’ list, Trovato considers the following as significant variants: variation in word order, following rhyming conventions in verse, addition or omission of words when they are not small or very common (Trovato and Reeve 2014, 111). Although O’Donnell makes a similar distinction, when he describes significant variants: [The] apparatus entries include only those forms that might be understood as involving a change in metrical, lexical or syntactic “significance” from the editorial lemma: i.e. variation involving the substitution of one lexical form, metrical pattern, or syntactic construction for another, or the irreversible destruction of sense, metre, or syntax. Such substitutions include variation between contextually appropriate alternatives... and contextually inappropriate alternatives and nonsense forms that cannot easily be restored to the archetypal form. (O’Donnell [2005] 2018, §7.8) 6 By referring to “potentially” significant variants, O’Donnell suggests that the judgement on significance is made at a later point; otherwise, there would be no need for the qualifier. Trovato instead subscribes to Caterina Brandolli’s efforts to describe what a polygenetic variant is, that is a variant that scribes produced independently, not from a common ancestor (Brandoli 2007), what we refer to as agreement by coincidence. Since Brandoli’s results largely agree with Saleman’s (Trovato and Reeve 2014, 220). Trovato’s conclusions aspire to be comprehensive and applicable to every textual tradition. It is an argument that favours the judgement of variants a priori, that is, the editor decides what is relevant for stemmatological purposes. In her doctoral dissertation, Bordalejo outlines the criteria she used for collating the texts of Caxton’s printed editions of the Canterbury Tales: “I have considered as significant all additions, deletions and substitutions, all the changes in word-order, all substantive variants [as opposed to Greg’s accidental variants]” (Bordalejo 2002, 104). She agrees with some of Salemans’ categories, but at no point does Bordalejo indicate that these criteria are applicable to any other textual tradition or even to other aspects of the study of the Canterbury Tales. It is more productive to describe the type of variation taken into account during the collation process and how the results of the collation were employed in the creation of the apparatus. There is no use in attempting to create a definitive list of what is relevant genealogical variation for every textual tradition. Vázquez, in his article “Transcribing and Collating for Digital Stemmatology” (Vázquez Forthcoming), shows examples of how some of the items in Salemans’ list can be questioned. Furthermore, Bordalejo’s approach aims to postpone the judgment of variants since the advances on stemmatics, and the use of digital tools do not require the a priori judgment of readings that would likely belong to the archetype (Bordalejo 2002, 98). This emphasis on “potential” stemmatically significant variation is similar in approach, if not in substance, to what O’Donnell proposes. Although the judgment of variants is postponed, it must be emphasized that it is not abandoned; but only after and through the lens of the analysis of all the variants is it possible to make legitimate claims about the genealogy of the witnesses of a textual tradition (Bordalejo 2002, 99). Hence, the difference between prescriptive and descriptive stemmatology: the distinction between assuming that one knows what should be taken into account versus allowing the textual evidence to speak for itself while also being aware of the objectives of an 7 individual project. In the case of the Canterbury Tales Project, research concentrates on stemmatically significant variation that can (potentially) contain genetic information. 4. The purposes of collation In the context of textual criticism focused on medieval materials, collation can have one or more end goals. Scholars might be concerned with the range of variation present in a series of texts, they might want to achieve a better understanding of the relationships between different instances of a text, they might want to explain the historical or linguistic circumstances surrounding a textual tradition, they might want to understand the development of a text, or they might be seeking to isolate readings to include in their editions. These are just a few examples of things that can be achieved by collating texts. Each scholar can choose to focus on some aspects more than others, or they can use multiple simultaneous approaches. Although Andrews states that “[t]he comparison may be done at the word level, at the character level, or at another unspecified syntactic or semantic level, according to the sensibilities of the editor,” (Andrews 2017, 232) we maintain that the choice has little to do with “sensibility” and everything to do with knowledge and the nature of the extant primary documents. a) Apparatus A common goal of collation is the production of a critical apparatus for inclusion in an edition. The type of edition (historico-critical, genetic, reader’s edition) is not as relevant as it is the dialectic relationship between apparatus and text. Marina Buzzoni makes this distinction when she states: The apparatus is indeed different from the descriptive lectio variorum one can get by, say, applying any collation software to the transcription of the witnesses. The apparatus is critical—i.e. interpretative—in that it accommodates certain variant readings, and excludes some others, according to the editorial principles to which the philologist conforms. (Buzzoni 2016, 76) She goes on to explain that while stemmatologists might only record those variants that preserve genetic information (excluding singletons, for example), anyone interested in the linguistic 8 features or language evolution, might include the lectiones singulares as a record of a philologically significant moment (Buzzoni 2016, 76). This was particularly pertinent in light of the discussions of a possible change of name for the TEI Critical Apparatus group. Suggestions like “textual variance” or “textual variants” do not carry the same weight or implications of “critical apparatus.” The relationship of the later with the text it informs, differs from the output of a collation process. b) Stemmatology For stemmatological purposes, transcription and collation form a solid base from which scholars can investigate the relationships between different witnesses within a textual tradition. To research textual filiation, scholars must first decide which variants are going to be deemed stemmatologically significant. And yet, every editorial decision made before that point affects the potential results. A thorough understanding of matters related to the textual tradition builds a trustworthy foundation from which collation decisions can be made. A dependable collation should, in turn, provide the basis for further research. This can be carried out by hand or, as is the case of the Canterbury Tales Project with a combination of phylogenetic analysis and advanced database searches (see Bordalejo’s article on analysis of textual materials using these techniques). Below, we refer to various projects using collation for stemmatological purposes: the Commedia, the Canterbury Tales, Troilus and Criseyde, and the Greek New Testament. A collation for stemmatological purposes requires, at least in the case of medieval texts, a process of regularization and alignment. See below the section on Chaucer’s Canterbury Tales. c) Corpus Linguistics It is clear that linguistics projects are not within the scope of textual criticism. However, there are some parallels regarding the treatment of words, especially the process of regularization and lemmatization since each groups different items under disciplinary criteria to improve the usability of the data. The Diachronic and Diatopic Corpus of American Spanish (Corpus Diacrónico y Diatópico del Español de América referred to as CORDIAM henceforth) is a 9 project recently supported by the Association of the Spanish Language Academies, hosted by the Mexican Academy of the Language, and is possible due to the collaboration and direction of Mexican linguist Concepción Company and Uruguayan linguist Virginia Bertolotti. The aim of this project is to: a) historicize the development of American Spanish, b) to accomplish a historical dialectology of American Spanish, and c) to achieve a complete and rich study of the American Spanish history, without geographical or dialectological parcellations when these are not required, given that to speak and write in Spanish is an integral reality, common to hundreds of millions of Hipanophones.” (Company and Bertolotti 2018, 78) It is unfortunate that this is a needed clarification. In this text, “American Spanish” does not mean ‘the Spanish of the United States of America,’ it means ‘the Spanish of the American continent.´ CORDIAM’s interface asks to write a word so that it can perform a search through the 12,907 texts and 9,644,566 words that it has up to May 21, 2020. A crucial feature of CORDIAM is that 70% of the corpus is lemmatized (Company and Bertolotti 2018, 100), which allows for complex searches. A lemma is “the technical term in lexicography and linguistics for a lexical item as it is presented in a dictionary entry, for the sake of clarity and economy” (Butterfield 2015). Thus, to lemmatize “is to group together varying words or forms of words: e.g. in work on a concordance” (Matthews 2014). One can search the Spanish verb ir (to go) plus a a preposition that indicates direction and another infinitive (infinitives in Spanish end in -ar, -er, -ir), which would be the equivalent to the English construction “to go to + infinitive.” The results include conjugated forms of the verb ir and the rest of the phrase. Since the verb ir is irregular, this can only be done because of the lemmatization process. 16 c. The Commedia, Sanguinetti and Shaw Peter Robinson, in “The Textual Tradition of Dante's Commedia and the Barbi ‘loci,’” reviews the differences between the treatment of variants in Shaw’s and Federico Sanguineti’s editions of the Commedia. According to Robinson: Sanguineti declared that not only could traditional stemmatics be applied to the whole tradition, but he had done it. He had looked at all the Barbi loci in every one of the 800 manuscripts (so achieving on his own, with virtually no support, what scholars had failed to achieve in over a hundred years), and from analysis of the readings at these loci he had created a comprehensive account of the whole tradition, and isolated just seven manuscripts as necessary and sufficient for the creation of a critical text. (Robinson 2012, 5) These witnesses were later called the Sanguinetti seven. Shaw and Robinson created “became a test of Sanguineti’s arguments about the relationships among these seven manuscripts” (Robinson 2012, 6). Sanguinetti argued that out of all the witnesses that have been divided into tradizione α and tradizione β, only the text of Vatican Library ms. Urbinate latino 366 (Urb) was a good representative of tradizione β. On the other hand, Prue Shaw’s edition of the Commedia argues that there is a common ancestor between Urb and Ms. Riccardiano 1005 (Roddewig n. 302) (Rb). This argument renders Sanguinetti’s editorial work unfruitful. The reason for the discrepancy is that a high proportion of the variants considered by Shaw and Robinson as demonstrative of genealogical relevance “would not satisfy Barbi’s criteria” (Robinson 2012, 28). Yet, it is not the a priori appearance of these variants what supports their argument, but the consistency of the agreements between these witnesses which reveals a pattern, or as Petrocchi would put it and Shaw quotes, the “foltezza di statistica” (Robinson 2012, 29). The discrepancy is the result of full-text collation versus the use of loci, as well as the assumptions that give place to both of them, namely, to think that the editor can predetermine what is stematologically relevant vs allowing the textual evidence to show the relationships the witnesses bear. The definition of what is a genetically significant variant shows the difference in method and attitude towards textual criticism. Although Prue and Robinson carried on a computer-assisted collation, this must not be taken to mean that the collation was done automatically. Regularization and alignment are under editorial control to assure that the collation is optimal for stemmatological purposes (See bellow regularization and alignment d) Chaucer’s The Canterbury Tales). 17 Robinson presents five readings that are “likely to have been introduced by the common ancestor of Urb/Rb; none of these five lines appear among the Barbi loci” (Robinson 2012, 27). These are the variants. Inf. i 89: aiutami da lei, famoso saggio, famoso e saggio LauSC Rb Urb FS famoso saggio Ash Ham Mart Triv PET Inf. ii 71: vegno del loco ove tornar disio; di LauSC Rb Urb FS del Ash Mart Triv PET dal Ham Inf. ii 110: a far lor pro o a fuggir lor danno, pro e a Mart-orig Rb Urb FS pro ne a Ash LauSC-c2 Triv prode et a Ham pro [..] a LauSC-orig pro o a Mart-c2 PET Inf. iii 3: per me si va tra la perduta gente. ne la Rb Urb FS tra la Ash Ham LauSC Mart Triv-c1 PET tra Triv-orig Inf. iii 22: Quivi sospiri, pianti e alti guai altri Ash-orig Rb Urb alti Ash-c2 Ham LauSC Mart Triv FS PET (Robinson 2012, 27) Both Robinson and Shaw understand why these variants would not catch Barbi’s attention: they would have the appearance of polygenetic readings to him, “an error liable to arise independently in independent manuscripts” (Robinson 2012, 28). Their work demonstrates that there is no valid a priori judgment; what matters is the distribution of variants across the manuscripts. The way in 18 which variants are judged and the assumptions it carries are at the center of the distinction between Barbi and Petrocchi, but also between Shaw and Sanguinetti. Caterina Brandoli, in her “Due canoni a confronto” examines the passages that Barbi and Petrocchi considered useful to constitute a stemma through the lens of various definitions of what constitutes polygenetic and monogenetic readings. It is an attempt to systematize what is genealogically relevant a priori. She concludes that out of the 396 Barbi loci, 366 fit her definition of monogenetic variants, thus are stemmatologically relevant (Brandoli 2007, 113). As for Pettrochi’s key passages, she states that 282 out of 477 seem to be monogenetic, which is still a majority but not as decisive as Barbi’s. Then she points out that out of those 282, 132 were already considered by Barbi (Brandoli 2007, 122), which would seem to prove that Petrocchi was only able to find 150 truly relevant readings and has a tendency to take into account polygenetic variants that are not of interest and would compromise his results. This difference extends to the discrepancy between Federico Sanguinetti’s edition of the Commedia and Prue Shaw’s and corresponds to the distinction previously made between prescriptive and descriptive stemmatology. The textual analysis conducted by Shaw’s edition rests on full-text transcriptions and computer-assisted fulltext collation, which minimizes the risk of missing any data by trying to analyze the tradition with manual methods and a limited set of variants. Shaw demonstrates that there is no a priori valid method to judge variants, makes the process as transparent as possible and reports the ties that the textual evidence shows, instead of the assumptions of the editor in detriment of what he did not judge to be relevant. d. Chaucer’s The Canterbury Tales For their collation of the Canterbury Tales, John Manly and Edith Rickert (Manly and Rickert 1940) recorded information in some 50.000 cards. Each card registers a number (D162) on the top left corner, corresponding to individual lines of the Canterbury Tales (in this case, “Al this sentence / me liketh euery del”). However, the indication, in the right-hand corner, that this card is one of two explains why we are only dealing with the first half of the line. The bottom of the card explains why some witnesses have left this particular line out, i.e. we find a record of whether the line, the passage or the tale are not present. The sigils of the witnesses with variants (Cn, Ha5, Ad3 and Ph2) are encircled with a pen, and the variation is recorded below. 19 Figure 2. Collation card by Manly and Rickert It is a deceptively simple system, so effective, that it allowed the editors a degree of accuracy that has not been recognized (the Canterbury Tales Project preliminary research suggests a punctilious precision beyond what one might expect from the tools available to the collators). Those of us who have carried out computer-assisted collations of the Tales can attest to the level of work that can be found in the Manly and Rickert volumes. The exactness of this collation illustrates the importance of thoughtful consideration of these most basic matters. The Canterbury Tales Project (along with any other projects its leaders have been involved in) produces complete encoded transcriptions of each witness of the Tales. The project attempts to improve over the manual work carried out by Manly and Rickert in the 1920s and 1930s, not because their edition is imprecise, but because their interpretation of the Canterbury Tales’ vast variants corpus created problems that muddled their understanding of the textual filiation in a sea of variation. 20 The main goal of the Canterbury Tales Project is to understand the textual history of the Tales as fully as possible with the information currently available. Under Bordalejo’s leadership, the project will press on with the production of edited texts which will be included in our publications (the first instance of an edited text appearing in a project publication was Bordalejo’s General Prologue reading text in the CantApp [Chaucer 2020]), but this does not alter the original goal of understanding the relationships between the different texts in the extant fifteenth-century witnesses of the Canterbury Tales. From the beginning, the Canterbury Tales Project’s transcriptions’ were conceived to retain information which would be useful for stemmatic analysis, even though the project has always made an emphasis on retaining scribal spellings which must be regularized during collation in order to produce meaningful stemmatological results (see Bittner and Dase; Bordalejo in this collection and Bordalejo 2016. Our original transcription guidelines (Robinson and Solopova n.d.) retained tails and flourishes that might have stood in place of final e. As the project continued, it became clear that these were merely ornamental. So they were finally excluded from the transcription. Moreover, the implementation of separate encoding based on the one used for the Divine Comedy (Bordalejo 2010) to account for the text of the document and the variant states of the text (Bordalejo 2016) also meant a modification of the transcription guidelines which culminated in its current version (Bordalejo and Robinson 2018). In this way, the project produces rich and detailed transcriptions that record places of variation within each document (with the use of the apparatus element) and which can be published alongside the images, thus allowing readers access to the same resources we use in our analysis. Our work is designed for generosity. We want others to benefit from the many hours spent on the creation of individual transcripts, which is why we make those available for reuse. Since a significant portion of our research was funded with money from various governmental sources, we have the duty to make them available for others to use in their research. Perhaps other scholars are interested in analyzing idiosyncratic peculiarities or quirky spellings. Someone interested, for example, in toponymy, could encode all the place names and repurpose our transcriptions to build maps; or they could encode historical figures or characters for a different type of study. 21 For our own purposes, if our initial transcription fails to record variation, it can be easily modified and processed again. Because we have complete transcriptions, we can choose to publish editions of individual witnesses (Stubbs 2000; Bordalejo 2003) or of multiple ones (Chaucer 2004; 2006). We can rebuild our transcriptions for different purposes, and select what version of the text and in which format will be displayed (see Bittner and Dase’s description of the multiple encoding of our apparatus tag). If there is a downside to the full-text transcription is the trap of the illusion of control. The transcription of complete witnesses postpones the detection of variants. Moreover, the researcher might feel some relief as she waits for the eventual results of the computer-assisted collation. There is safety in not having to jump in a vast sea of variation with decisions being made at the moment. As we collate the Canterbury Tales, we carry out two further processes: regularization and alignment. The aim of these processes is two-fold. On the one hand, we seek to flatten spelling information in order to analyze the collation results with the help of evolutionary biology software. As we have pointed out before, the project seeks to do a genealogical analysis of the textual tradition, and spelling variation obscures the filiation of the witnesses. On the other hand, we intend to produce a readable apparatus that can be easily understood by human reading. In the same way that the project’s transcription guidelines have evolved, so have our regularization ones. The original guidelines for regularization indicated that the different witnesses should be regularized to a very lightly edited version of Hengwrt (National Library of Wales, Peniarth, 392 D), likely to be the oldest manuscript of the Tales and witness to one of the best texts in the tradition. But the words “lightly edited” mean nothing unless further specified. By lightly edited, we meant that all medieval characters, no longer in use in contemporary English, were substituted by modern forms, abbreviations were expanded, and all lines found in any witness not present in our base-text added. This created our base-text for collation, which Collate used for comparison purposes. Later, as the edition was built, the base-text for collation was discarded. 22 Although there would have been little interest in the base-text, this was useful because it recorded most of the spellings that we would use for regularization purposes. Our rule for this was to regularize to the most common spelling of Hengwrt, and we had a list of preferred witnesses in case this turned out not to be possible (in cases in which Hengwrt did not include a particular term, for example). Somewhere else we explain the concept that underlines our view on variation: It is our core conviction, based on decades of work with digital tools, that “significant variants” are defined entirely by how the variants are distributed across the whole tradition. That is: if we find a number of variants which are present, over and over again, in the same distinctive pattern of witnesses, then those variants are significant. (Bordalejo and Robinson 2019, 37) For the purposes of the Canterbury Tales Project, whether variants are polygenetic or monogenetic is not crucial. Instead, we consider variant distribution, which we can only know a posteriori, the essential factor to determine stemmatically significant variation. Our collations carried out with CollateX present similar challenges. They require to show stemmatically significant variation which includes all text present in some witnesses and not others, changes in order, substitutions, and all substantive variation. From our collations, we seek to both produce files that can be used with evolutionary biology (such as PAUP (Swofford 2003) or stemmatological software (RHM [T. Roos and Heikkila 2009] or SemStem [Teemu Roos and Zou 2011]). e. Chaucer’s Troilus and Criseyde There is no stemma for the textual tradition of Troilus and Criseyde. Previous editions and analysis (Chaucer 1926; 1984; 2008) have stated that it is not possible to conduct traditional recension due to the constant changes in filiation of the witnesses. Thus, this is an opportunity for analysis with the aid of digital tools. For the initial stages of the recension of the textual tradition of Chaucer’s Troilus and Criseyde, Adam Vázquez decided to transcribe and collate the 23 first 546 lines of Book one, as well as lines 764-833 Book one, and 490-1225 Book two, from the 16 manuscripts and two early printed versions. The purpose of the project is to conduct phylogenetic analysis. The first 546 lines work as a point of reference since we know thanks to the work of past scholars that, from line 547 to the end of the poem, Wynkyn de Worde follows Caxton’s edition but not for lines 1-546, that is, there is a change in filiation. Then the analysis of lines 764-833 showed Wynkyn changing place in the phylogram. That proved the efficacy of the method. Then, the analysis of lines II. 490-1225 attempts to shed light on this problematic excerpt of Book two, given the changes in filiation of some witnesses that have been investigated previously by various scholars (Root 1926; Hanna 1996). This method postpones judgment. The project does not seek to make general claims on the textual tradition based on the analysis of 1350 lines. The aim is to analyze the selected excerpts rigorously. The full transcription of the excerpts and the full collation of them enables the textual critic to get acquainted with the material. The semi-automatic collation aids the collator and reduces the risk of error. One of the advantages of semi-automatic full-text collation is that even if the collator misses one or two variants, the direction of the collation is less susceptible to change since it relies on the complete body of evidence. The stemmatological aspect of the project does not rest on the selection of a few variants. As seen with the discrepancies between Barbi and Petrocchi, or Shaw and Sanguinetti, variants that would escape the attention of the manual collator and alter the direction of genealogy, are less likely to be missed by a textual critic that uses digital tools. The Troilus project is far from complete, but it shows consistent results so far (Vázquez 2020). By comparing the critical apparatus that Barry Windeatt and Robert K. Root provide in their editions, with a computer-assisted collation we can see that there is a small omission as early as line four. The line reads “fro wo to wele, and after out of ioie” (Chaucer 1984, 84). All the witnesses but one agree on “out of.” Rawlinson reads “on to.” Neither apparatus registers this variant. It is a small variant, it is also not relevant for genealogical purposes since it is not present in any other witness, yet one may argue that since both their editions walk away from making genealogical statements, the apparatus is there to inform the reader of what is present in the textual tradition. On another occasion, Windeatt’s apparatus draws attention to the fact that the 24 text in the Corpus manuscript presents a peculiar spelling in the third line, thus: “auentures] auentuirs Cp (3 minims after t)” (Chaucer 1984, 85). Thus, a lack of interest in small detail is not to blame for Windeatt’s disregard for the R variant: the most likely explanation for this omission is that it is easy to miss when doing manual collation. A digital collation tool that draws variants from full-text transcriptions cannot fail to bring this reading to the collator’s attention. f. The Greek New Testament, Nestle-Aland, and the Editio Critica Maior The Institute for New Testament Textual Research has collected the necessary materials of the textual tradition of the Greek New Testament. It was then necessary to filter the material so that “new views about important manuscripts find their way into the minor editions of the institute, and finally to present an Editio Critica Maior” (Mink 2004, 17). In order to access the relevance of the text of manuscripts, a sampling method was devised and the results published “in the five volumes of the Text und Textwert der Griechischen Handschriften des Neuen Testaments” (Morrill 2012, 7). It meant to separate the manuscripts that contained “the relatively uniform text which was standard at the end of the Byzantine tradition from the still large number of manuscripts which must be considered relevant on account of their deviations from the majority text” (Mink 2004, 17). The results made it possible to select manuscripts that do not contain the uniform text from the end of the tradition. The Claremont Profile Method that was used for the classification of 1385 manuscripts of the Gospel According to St. Luke (Morrill 2012, 23) is another example of sampling collation. The test passages were picked out after “the complete collation of 282 manuscripts. A sample of three chapters, 1, 10, and 20, were selected, and all variation in these chapters was evaluated” (Morrill 2012, 21). After careful consideration, 196 were used to analyze 816 more manuscripts “for a total of 1385 manuscripts in 196 passages” (Morrill 2012, 23), and fourteen groups of witnesses were created after the analysis. The purpose of this collation was not to create an apparatus, not an edition. The reasoning was that the “critical apparatus would adequately represent the manuscript tradition if it included representatives of the groups, plus those manuscripts that did 25 not fall into definable groups” (Morrill 2012, 19). It cannot be denied that the Claremont Profile Method achieved remarkable results, yet it will always be preferable to rely on full-text transcriptions and full-text collations. The Institute for New Testament Textual Research has used computer-assisted collation tools for more than 20 years. First Collate and now CollateX have played a significant role in the development of their editions. Both the Nestle-Aland Greek New Testament and the Editio Critica Maior built their research and apparatuses with the help of semi-automated software after carrying out complete transcriptions of witnesses (Houghton et al. 2020). Bordalejo has before described the Editio Critica Maior as a born-digital printed edition, in reference to the techniques used in its construction and as an example of how similar digital textual critical tools are to their analogue counterparts, despite their speed and higher accuracy (Bordalejo 2013, 65n). What is remarkable about this edition is how its apparatus, based on a computer-assisted collation, differs from that of the Nestle-Aland edition. Despite the same tools being employed by both, the distinct purposes of the editions are made evident in the apparatuses generated from very different collations. 6. Digital Collation Tools There are several tools that can be used in order to conduct a computer-assisted collation. In this section, we examine some of the tools with collation features, their characteristics, and their potential uses. i. TUSTEP TUSTEP (TUebingen System of TExt processing Programs) is a toolbox for scholarly processing textual data. According to Gabler, the “algorithm was devised 30 years ago and is still among the most powerful and efficient of collation” (Gabler 2007, 4). A demonstration of TXSTEP in 2015 shows the algorithm at work. It shows versions of the text, but it seems like one has to program 32 7. Conclusions: Learning from Collation Collation is not an isolated process that happens in a vacuum, but a practice occurring within the wider context of textual-critical research and which, at times, leads to the production of an edition. It is the process of variant identification that allows scholars to further their research agenda. However, editions depend on a limited number of variants. Even in the case of digital editions, where every piece of variation can be included, we are forced to reckon with the reality represented by our limited number of texts. Human intellect is also limited, particularly when it refers to a large number of items and, for this reason, we rely on other systems to help us process the vast number of variants we detect during a regular collation process. We have shown that, although the collation process relies on the same principles, whether it is carried out manually or with the aid of computers, there are significant advantages in using computer-assisted methods over manual ones. Although the preparation of files for computer collation requires a significant investment of time and effort, by creating full-text transcriptions and making them publicly available, we ensure that our work can be evaluated by other scholars and reused in future research. These advantages are more evident in the context of research on large textual traditions, but they do not disappear in reference to briefer or less distributed texts. The success of a critical edition relies on its ability to connect a system of data. With computerassisted collation methods and full-text transcriptions, the process that leads to a critical text becomes comprehensive, thorough, and more transparent to the reader. By consequence, the critical text turns into a window through which we can observe the circumstances and the intervention of many of the agents that made it possible for us to connect/engage with the texts that weave us as part of a community. 33 Exemplum: The Miller’s Tale: Manual vs ComputerAssisted Collation. -Señor conde Lucanor -dixo Patronio-, mucho me plaze desto que dezides, et para que vós mejor lo podades fazer, plazerme ya que sopiésedes lo que consteçió a un muy grand philósopho et mucho ançiano. (Don Juan Manuel, El Conde Lucanor ) This research was carried out in the years after the publication of The Miller’s Tale on CDROM. We offer it here as an example of how much more accurate the computer is. This, naturally, could be an isolated instance of carelessness. Barbara Bordalejo is aware that this text exposes her and her research in ways that are not often privy to others. During the preparation of Bordalejo’s De Montfort University Ph.D. thesis, she used Collate 2 to compare the encoded transcriptions of Caxton’s first (Cx1) and second (Cx2) editions of the Canterbury Tales. Both of the Caxton transcriptions used in Bordalejo’s 2003 study attempted to represent spelling as accurately as possible but ignored form distinctions between ragged and Roman r or long and round s. The system overlooked other distinguishable sorts, including tailed d and ligatures, because it was primarily developed for the transcription of manuscript materials and not for incunables. Because Bordalejo’s work focused on isolating potentially stemmatically significant variants, it required her to separate those from the orthographic and graphetic variation included in the transcriptions by default. The Canterbury Tales Project single-tale editions use software record and save (in a separate file as did Collate or in a database as is the case of CollateX) regularizations of orthographic variants (see Farrell’s article on his work on dialectological variation in The Reeve’s Tale). This process ensures that the transcriptions remain a close reflection of what is found in the source document. 34 At the time, Bordalejo did not carry out a complete regularization of Caxton’s editions. The regularization process would have allowed to hide all accidental variation (such as spelling and punctuation) and would have shown only variation at word level, including modifications in word-order and other substantive changes. She collated Caxton editions using the raw transcriptions (without the help of a regularization file). The result of this process was an extremely long list of differences between Cx1 and Cx2. The lists were printed out and read before deciding whether further analysis of the variants would be required. For example, a collation of both editions, using Cx1 as a base, of the first lines of the General Prologue, we find: Figure 8. 35 The Collate output shows all the differences between the two editions, even those that only represent a difference in the type (bret˙/breth; wit˙/wyth) or encoding ([3orncp]W[/3orncp]Han/[4orncp]W[/4orncp]Han). These variants were eliminated from the analysis, as they were of no use to trace the affiliations of the source for the corrections of Cx2. A great deal of attention was required to detect those variants that could have the potential to be stemmatically significant. In the previous example, such variants are represented by “And the/The” in line 2 and “licoue/lycour” in line 3. Bordalejo read all the variants between the editions attempting to isolate those with potential stemmatic significance. Although the aim was to record the majority of the variants, and great care was devoted to the data gathering, human error remained a consideration. However, the proportion of the potential differences between the electronic and manual data gathering were not clear, not until the regularization process had been completed for the Miller’s Tale on CD-ROM, published in 2003. This publication allowed users to retrieve the information about the differences between both of Caxton’s editions in seconds. The non-computerized search, s ‘manual’ collation, used to gather data for Bordalejo Ph.D. yielded 79 stemmatically significant variants in The Miller’s Tale. This number did not take into account the addition, deletion and substitution of lines, as this was being dealt with in a different section of her work where she established that there were three line substitutions, and seven lines were added in Cx2. No lines were deleted that were on Cx1 without being replaced with an alternative line. Although the data about addition, deletion and substitution of lines were easily retrieved using Collate, individual (word) variants detected were the result of reading and marking an unregularized printed collation. These accounted for 33 variants that Bordalejo was considering separately (reference to thesis) and which included lines 47, 47a, 28, 416-1, 465-1, 466, 577 to 584, and 585. By subtracting 33 from the 150 variants of the computer collation, there 117 variants of which 38 were not accounted for in the manual collation. There were several reasons which could have explained these discrepancies: 1. Transcription mistakes 2. Different segmentation 3. Incorrect regularization 4. Different understanding of variation 36 5. Human error The latest can be further divided into errors helped by the display in Collate, which was not optimized to be read but rather to be processed, and inaccuracy or oversight. 1. Transcription mistakes could have occurred at any point. These could have been input after Bordalejo’s final collation, but before the Miller’s Tale on CD-ROM was finished or they could have been corrected for the publication, in which case the mistake would have remained part of the final collation. 2. The Canterbury Tales Project uses parallel segmentation. This ensures that variant phrases are reduced to their minimum variant expression (sometimes two words but in a few cases six or seven). However, Collate just shows sequences of differences. When evaluating the Cx1/ Cx2 collation, variant phrases were counted as a single variant (even in cases in which these could have been separated into more than one). 3. Because the regularization process is carried out by humans), it is prone to error. Occasionally, there has been overregularization (MI 322, MI 564, MI 619) or under regularization of the variants. There are also simple cases of misregularization when a word goes to the incorrect lemma (MI 64) 4. In very few cases, the regularization policy employed by a particular editor might not be the same that was used when detecting the variants for Bordalejo’s thesis. She might have made a different regularization choice in lines MI 16, MI 37, MI 153. Analyzing these factors it is possible to see that of the 150 original variants which result of the CD-ROM’s search: 33 are line variants considered in a different section of Bordalejo’s thesis. 7 are the result of different segmentation (-7) 4 are the result of incorrect regularization (-4) 3 are the result of a different understanding of variation (-3) This gives a total of 14 variants which we need to subtract to obtain 103 net variants of which 79 had been correctly isolated. This translates into 24 variants which were not detected in the original Caxton collation when reading the Collate output. Of those 24 overlooked variants: 3 are related the Collate display 2 (MI 208 and MI 210) were errors in judgment due to lack of experience. 37 The other 19 even though present in the original collation between the Caxton editions, were not detected. Four of those occurred in lines that have some other kind of variation, and it might be possible that their closeness to other variants might explain why they were overlooked. This leaves a total of fifteen missed variants for which we cannot find any explanation. Those fifteen variants translate into around 20% of the variation. Projecting such a number to the rest of the text would be a total of some 600 undetected variants if one were to carry out a manual collation with a steady error rate. Even with the risks and decisions made during the regularization process, its accuracy rate is much higher. Since its publication in 2004, only three regularization mistakes have been found in the Miller’s Tale. Not taking into account disagreements with the editor or parallel segmentation matters, then the machine aided collation has an inaccuracy rate of only 2%—a much better ratio than the one achieved by Bordalejo. If we assume that these numbers are representative, the case for computer-assisted collation is solid. Although it is possible to produce more accurate manual collations if researchers can dedicate time and effort to their work, it seems evident that the computer’s capacity to count and classify should produce better results when a similar timeframe is allowed and provided that the data is relatively free of errors. Although computer-assisted collation is time-consuming in its data preparation, it is a superior system when limited time is available. Bibliography Agata, Mari. 2006. “Stop-Press Variants in the Gutenberg Bible.” Graduate, Tokyo: Keio University. ———. 2011. “Improvements, Corrections, and Changes in the Gutenberg Bible.” In Scribes, Printers, and the Accidentals of Their Texts, edited by Jacob Thaisen and Hanna Rutkowska, 135–55. Frankfurt am Main; New York: Peter Lang. Andrews, Tara. 2017. “What We Talk about When We Talk about Collation.” In Advances in Digital Scholarly Editing, edited by Peter Boot, Anna Cappellotto, Wout Dillen, Franz Fischer, Aodhán Kelly, Andreas Mertgens, Anna-Maria Sichani, Elena Spadini, and Dirk Van Hulle, 233–36. Leiden: Sidestone Press. https://www.sidestone.com/books/advances-in-digital-scholarly-editing. Blake, Norman Francis. 1996. “The New Lineation System.” Edited by Norman Francis Blake and Peter M. W. Robinson. The Canterbury Tales Project: Occasional Papers II. 38 Blecua, Alberto. 1983. Manual de Crítica Textual. Literatura y Sociedad 33. Madrid: Castalia. Bordalejo, Barbara. 2002. “The Manuscript Source of Caxton’s Second Edition of The Canterbury Tales and Its Place in the Textual Tradition of the Tales.” De Montfort University. https://www.academia.edu/2987324/THE_MANUSCRIPT_SOURCE_OF_CAXTONS_ SECOND_EDITION_OF_THE_CANTERBURY_TALES_AND_ITS_PLACE_IN_THE _TEXTUAL_TRADITION_OF_THE_TALES. ———. , ed. 2003. Caxton’s Canterbury Tales: The British Library Copies. Leicester: Scholarly Digital Editions. ———. 2013. “The Texts We See and the Works We Imagine: The Shift of Focus of Textual Scholarship in the Digital Age.” Ecdotica 10: 64–76. ———. 2014. “Caxton’s Editing of the Canterbury Tales.” Papers of the Bibliographical Society of America 108 (1): 41–60. ———. 2016. “The Genealogy of Texts: Manuscript Traditions and Textual Traditions.” Digital Scholarship in the Humanities 31 (3): 563–577. https://doi.org/10.1093/llc/fqv038. ———. 2018. “Digital versus Analogue Textual Scholarship or The Revolution Is Just in the Title.” Digital Philology: A Journal of Medieval Cultures 7 (1): 7–28. https://doi.org/10.1353/dph.2018.0001. Bordalejo, Barbara, Fiona Maguire, Manuel Moreno, Peter Robinson, and Dorothy Severin. 2014. “An Electronic Corpus of Fifteenth-Century Castilian Cancionero Manuscripts.” Digital Philology: A Journal of Medieval Cultures 3 (1): 11–23. Bordalejo, Barbara, and Peter Robinson. 2018. “New Transcription Guidelines - Wiki - Textual Community.” 2018. http://www.textualcommunities.usask.ca/web/canterbury-tales/wiki/- /wiki/Main/New+Transcription+Guidelines. Bordalejo, Barbara, and Peter M. W. Robinson. 2019. “Manuscripts with Few Significant Introduced Variants.” Ecdotica 15 (March): 37–65. https://doi.org/10.5281/zenodo.3344880. Bowers, Fredson. 1995. Principles of Bibliographical Description. Oak Knoll Press. Brandoli, Caterina. 2007. “Due Canoni a Confronto: I Luoghi Di Barbi e Lo Scrutinio Di Petrocchi.” In Nuove Prospettive Sulla Tradizione Della Commedia, 99–148. Butterfield, Jeremy. 2015. Lemma. Oxford University Press. https://www.oxfordreference.com/view/10.1093/acref/9780199661350.001.0001/acref9780199661350-e-3283. Buzzoni, Marina. 2016. “4. A Protocol for Scholarly Digital Editions? The Italian Point of View.” In Digital Scholarly Editing: Theories and Practices, edited by Matthew James Driscoll and Elena Pierazzo, 59–82. Open Book Publishers. https://doi.org/10.11647/OBP.0095.04. Chaucer, Geoffrey. 1926. The Book of Troilus and Criseyde,. Edited by Robert K Root. Princeton: Princeton University Press. ———. 1984. Troilus & Criseyde: A New Edition of “The Book of Troilus.” Edited by B. A. Windeatt. London ; New York: Longman. ———. 2000. The Hengwrt Chaucer digital facsimile. Edited by Estelle Stubbs. Leicester: Scholarly Digital Editions. ———. 2004. The Miller’s Tale on CD-ROM. Edited by Peter Robinson. Canterbury Tales Project. Leicester: Scholarly Digital Editions. http://www.sdeditions.com/AnaAdditional/miller/images/millerhome.html. 39 ———. 2006. The Nun’s Priest’s Tale on CD-ROM. Edited by Paul Thomas. Leicester: Scholarly Digital Editions. http://www.sd-editions.com/NP/index.html. Chaucer, Geoffrey, and Larry Dean Benson. 2008. The Riverside Chaucer. 3rd ed. Oxford New York: Oxford University Press. Cherchi, Paolo. 2008. “Review Article. Nuove Prospettive Sulla Tradizione Della Commedia. Una Guida Filologico Linguistica Al Poema Dantesco.” Annali d’Italianistica 26: 413– 22. “CollateX – About the Project.” n.d. Accessed June 23, 2020. https://collatex.net/about/. “Collation.” 2013. Lexicon of Scholarly Editing (blog). February 15, 2013. https://lexiconse.uantwerpen.be/index.php/lexicon/collation/. Company, Concepción, and Virginia Bertolotti. 2018. “El Corpus Para América.” In Historia Del Léxico Español y Humanidades Digitales. Berlín: Peter Lang. Driscoll, Matthew James, and Elena Pierazzo, eds. 2016. Digital Scholarly Editing: Theories and Practices. Digital Humanities Series, v. 4. Cambridge: Open Book Publishers. Fischer, Franz. 2019. “Digital Classical Philology and the Critical Apparatus.” In Digital Classical Philology: Ancient Greek and Latin in the Digital Revolution, edited by Monica Berti. De Gruyter Saur. https://www.degruyter.com/view/title/537705. Frenk, Margit. 1971. Entre Folklore y Literatura. México, D.F.: El Colegio de México. ———. 2003. Nuevo corpus de la antigua lírica popular hispánica (siglos XV a XVII) Vol. I. Vol. 1. 2 vols. México, D.F.: Facultad de Filosofía y Letras, UNAM ; Colegio de México; Fondo de Cultura Económica. Gabler, Hans W. 2007. “Remarks on Collation.” https://www.academia.edu/167070/_Remarks_on_Collation_. Gaskell, Philip. 2000. A New Introduction to Bibliography. Oak Knoll Press. Greetham, David C. 1994. Textual Scholarship: An Introduction. Routledge. Greg, W.W. 1950. “The Rationale of Copy-Text.” Studies in Bibliography 3: 19–36. Hanna, Ralph. 1996. Pursuing History: Middle English Manuscripts and Their Texts. Stanford (Calif.: Stanford University Press. Hansen, João A., and Marcello Moreira. 2013. Para que todos entendias. Poesia atribuida a Gregorio de Matos e Guerra. Belo Horizonte: AUTENTICA EDITORA. Hinman, Charlton. 1955. “Mechanized Collaioon at the Houghton Library.” Harvard Library Bulleting 9: 132–34. Houghton, H. A. G., D.C. Parker, Peter Robinson, and Klaus Wachtel. 2020. “The Editio Critica Maior of the Greek New Testament.” Early Christianity 11 (1): 97. https://doi.org/10.1628/ec-2020-0009. Manly, John M., and Edith Rickert. 1940. The Text of the Canterbury Tales Studied on the Basis of All Known Manuscripts, Manly, John M.; Rickert, Edith ; with the Aid of Mabel Dean e. a.; with a Chapter on Illuminations by Margaret Rickert. Chicago (Ill.): University of Chicago press. Matthews, P.H. 2014. “Lemma.” https://doi.org/10.1093/acref/9780199675128.013.1845. McKerrow, Ronald B. 1927. An Introduction to Bibliography for Literary Students. 1960th ed. Oxford: Oxford University Press/ Clarendon Press. Mecca, Angelo E. 2013. “Il Canone Editoriale Dell’antica Vulgata Di Giorgio Petrocchi e Le Edizioni Dantesche Del Boccaccio.” In Nuove Prospettive Sulla Tradizione Della Commedia: Seconda Serie, 2008-2013. Padova: Libreriauniversitaria. Mink, Gerd. 2004. “Problems of a Highly Contaminated Tradition: The New Testament.” In 40 Studies in Stemmatology II, edited by Pieter van Reenen, August den Hollander, and Margot van Mulken, 13–85. Philadelphia, NETHERLANDS, THE: John Benjamins Publishing Company. http://ebookcentral.proquest.com/lib/usask/detail.action?docID=623163. Morrill, Bruce. 2012. “A Complete Collation and Analysis of All Greek Manuscripts of John 18.” Birmingham: University of Birmingham. Nury, Elisa. 2018. “Automated Collation and Digital Editions: From Theory to Practice. Classical Studies.” King’s College London: King’s College. https://hal.archivesouvertes.fr/tel-02493805/document. O’Donnell, Daniel Paul, ed. (2005) 2018. Cædmon’s Hymn: A Multimedia Study, Edition and Archive. 1.1 Internet Reprint. Vol. A.8. SEENET. SEENET. http://caedmon.seenet.org/index.html. Accessed July 2, 2020. https://caedmon.seenet.org/htm/introduction/ch7.html#CH7.1.1.HEAD. Parker, D. C. 2008. An Introduction to the New Testament Manuscripts and Their Texts. 1 edition. Cambridge, UK ; New York: Cambridge University Press. Petrocchi, Giorgio. 1966. “Criteri fondamentali dell’edizione.” In La Commedia secondo l’antica vulgata / Dante Alighieri ; a cura di Giorgio Petrocchi, by Dante Alighieri, 1a ed. Vol. 1. Milano: Mondadori. Plachta, Bodo. 1995. “German Literature.” In Scholarly Editing: A Guide to Research, edited by D. C. Greetham, 504–29. New York: Modern Language Association of America. Robinson, Peter. n.d. Textual Communities. https//:textualcommunities.org. Robinson, Peter M. W. 2012. “The Textual Tradition of Dante’s Commedia and the Barbi’s Loci.” Ecdotica 9: 1–32. ———. 2014. “Collate 2, and the Design for Its Successor: CollateXML (Now, CollateX).” Zenodo. https://doi.org/10.5281/zenodo.3902430. ———. n.d. Collate. Robinson, Peter M. W., and Elizabeth Solopova. n.d. “Original Transcription Guidelines - Canterbury Tales Project 2 - Wiki.” Accessed June 30, 2020. https://wiki.usask.ca/display/CTP2/Original+Transcription+Guidelines. Robinson, Peter Max Wilton. 1994. Collate 2: A User Guide. Oxford: The Computers and Variant Texts Project. Root, Robert K. 1926. “Introduction.” In The Book of Troilus and Criseyde,. Princeton: Princeton University Press. Salemans, Ben J. P. 2000. Building Stemmas with the Computer in a Cladistic, NeoLachmannian, Way: The Case of Fourteen Text Versions of Lanseloet van Denemerken. Nijmegen: Nijmegen Univ. Press. Smith, Catherine. 2019. Itsee-Birmingham/Collation_editor_core 1.0.4. Zenodo. https://doi.org/10.5281/zenodo.3539578. Spadini, Elena. 2017. “The Role of the Base Manuscript in the Collation of Medieval Texts.” In Advances in Digital Scholarly Editing, edited by Peter Boot, Anna Cappellotto, Wout Dillen, Franz Fischer, Aodhán Kelly, Andreas Mertgens, Anna-Maria Sichani, Elena Spadini, and Dirk Van Hulle, 345–49. Leiden: Sidestone Press. https://www.sidestone.com/books/advances-in-digital-scholarly-editing. Timpanaro, Sebastiano. 1985. La Genesi Del Metodo Del Lachmann. 1a ristampa corr. con alcune aggiunte. Saggi 5. Padova: Liviana. 41 Trovato, Paolo, and Michael D. Reeve. 2014. Everything you always wanted to know about Lachmann’s method: a non-standard handbook of genealogical textual criticism in the age of post-structuralism, gladistics, and copy-text. Storie e linguaggie. Padova, Italy: Libreriauniversitaria.it Ed. Vázquez, Adam. 2020. “Transcribing and Collating for Digital Stemmatology.” Digital Studies / Le Champ Numérique. Vergani, Luisa. 1969. “Review Article. La Commedia, Secondo l’antica Vulgata.” Italica 46 (2): 191–93. https://doi.org/10.2307/477953. Vinaver, Eugène. 1976. “PRINCIPLES OF TEXTUAL EMENDATION.” In Medieval Manuscripts and Textual Criticism, edited by CHRISTOPHER KLEINHENZ, 139–66. University of North Carolina Press. www.jstor.org/stable/10.5149/9781469646190_kleinhenz.13. Waltz, Robert B. 2013. The Encyclopedia of New Testament Textual Criticism. Last Preliminary Edition. Williams, William Proctor, and Craig S. Abbott. 2009. An Introduction to Bibliographical and Textual Studies. 4th ed. New York: Modern Language Association of America.