<?xml version="1.0"?>
<feed xmlns="http://www.w3.org/2005/Atom" xml:lang="en">
	<id>https://wiki.tei-c.org/api.php?action=feedcontributions&amp;feedformat=atom&amp;user=Andrew+ollett</id>
	<title>TEIWiki - User contributions [en]</title>
	<link rel="self" type="application/atom+xml" href="https://wiki.tei-c.org/api.php?action=feedcontributions&amp;feedformat=atom&amp;user=Andrew+ollett"/>
	<link rel="alternate" type="text/html" href="https://wiki.tei-c.org/index.php/Special:Contributions/Andrew_ollett"/>
	<updated>2026-10-07T04:42:42Z</updated>
	<subtitle>User contributions</subtitle>
	<generator>MediaWiki 1.46.2</generator>
	<entry>
		<id>https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16175</id>
		<title>SIG:IndicTexts</title>
		<link rel="alternate" type="text/html" href="https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16175"/>
		<updated>2018-04-10T02:58:26Z</updated>

		<summary type="html">&lt;p&gt;Andrew ollett: &lt;/p&gt;
&lt;hr /&gt;
&lt;div&gt;The purpose of the TEI Special Interest Group “Indic Texts” is to&lt;br /&gt;
allow scholars engaged in the study of Indic texts to develop and&lt;br /&gt;
document best practices in applying the TEI’s Guidelines to these&lt;br /&gt;
kinds of texts.  To participate, please join the mailing list at&lt;br /&gt;
http://lists.lists.tei-c.org/mailman/listinfo/indic-texts.&lt;br /&gt;
&lt;br /&gt;
There are several respects in which the applicability of the TEI&lt;br /&gt;
guidelines to these texts is less than obvious. These relate to&lt;br /&gt;
distinctive features of Indic textuality, including:&lt;br /&gt;
&lt;br /&gt;
* the use of syllabic scripts, and the non-coincidence of grapheme- (syllable-) and word-boundaries;&lt;br /&gt;
* the application of phonotactic rules (sandhi) that further obscures the boundaries between words; extensive compounding;&lt;br /&gt;
* the use of distinctive media and writing supports (such as birch bark, palm leaves, and copper plates);&lt;br /&gt;
* distinctive metrical patterns with different types of caesuras;&lt;br /&gt;
* the prominence of the commentary as a genre, and the depth of intertextual relations this implies;&lt;br /&gt;
* the frequent reuse of texts in other texts, which requires careful and deliberate application of the &amp;quot;quoteLike&amp;quot; module.&lt;br /&gt;
&lt;br /&gt;
The expected outcome of the SIG’s work is a practical guide that&lt;br /&gt;
analyzes common cases in the markup of Indic texts and proposes ways&lt;br /&gt;
in which the analytical tools provided by the TEI Guidelines might&lt;br /&gt;
best be used in these cases, discussing benefits and drawbacks of the&lt;br /&gt;
solutions possible. Ideally, this guide will become a part of the TEI&lt;br /&gt;
Guidelines.&lt;br /&gt;
&lt;br /&gt;
&lt;br /&gt;
== Manuscript Transcription ==&lt;br /&gt;
&lt;br /&gt;
Canonically, an akṣara makes up a single &amp;quot;grapheme,&amp;quot; and this is reflected in Unicode representations of Indic scripts, where consonants and independent vowels are encoded first, and then vowel-markers (and dependent consonants like &#039;&#039;anusvāraḥ&#039;&#039; and &#039;&#039;visargaḥ&#039;&#039;) are encoded subsequently as combining characters. Unless marked with a combining vowel character, or a cancellation character, consonants are understood to have an inherent vowel &#039;&#039;a&#039;&#039;. The sequence of consonants within conjuncts is also canonically the same as their phonological sequence. Thus in the conjunct &amp;quot;rg&amp;quot;, the &amp;quot;r&amp;quot; is represented before the &amp;quot;g&amp;quot; in transliteration, in Devanagari र्ग (0930 + 094D + 0917) and in Kannada ರ್ಗ (0CB0 + 0CCD + 0C97), although it is rendered on top of the &amp;quot;g&amp;quot; in Devanagari and to the right of the &amp;quot;g&amp;quot; in Kannada. &lt;br /&gt;
&lt;br /&gt;
=== Cancelling dependent vowels ===&lt;br /&gt;
In manuscripts, dependent vowel markers can be cancelled, and the consonant is then read with the inherent vowel &#039;&#039;a&#039;&#039;. &lt;br /&gt;
&lt;br /&gt;
[[File:Ondondarol.png]]&lt;br /&gt;
&lt;br /&gt;
&lt;br /&gt;
If you want to encode this kind of change, there are technical problems, whether one is using an Indic script or an alphabetic transliteration system (like IAST or ISO-15919):&lt;br /&gt;
* In Indic scripts, rendering problems are likely if the cancelled vowel marker is enclosed within the &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;del&amp;amp;gt;&amp;amp;lt;/del&amp;amp;gt;&amp;lt;/code&amp;gt; tags, since it is a combining character;&lt;br /&gt;
* In transliteration, the deletion of one vowel must be accompanied by the addition of the inherent vowel, although there is no addition marked as such in the manuscript.&lt;br /&gt;
&lt;br /&gt;
There are two possible solutions that came up on the list. Both involve the use of the &amp;amp;lt;subst&amp;amp;gt; element, which contains the akṣara that is subject to scribal modification, and within it, the &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;add&amp;amp;gt;&amp;lt;/code&amp;gt; and &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;del&amp;amp;gt;&amp;lt;/code&amp;gt; elements.&lt;br /&gt;
&lt;br /&gt;
The first, and most straightforward, solution is to treat the akṣara, and not the &amp;quot;akṣara part,&amp;quot; as the smallest unit of variation in the manuscript, and thus to include the consonant in the &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;add&amp;amp;gt;&amp;lt;/code&amp;gt; and &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;del&amp;amp;gt;&amp;lt;/code&amp;gt; elements. The correction can thus be read as changing &amp;quot;ḷo&amp;quot; into &amp;quot;ḷa&amp;quot;. &lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
&amp;amp;lt;subst&amp;amp;gt;&amp;amp;lt;del type=&amp;quot;cancelled&amp;quot;&amp;amp;gt;ḷo&amp;amp;lt;/del&amp;amp;gt;&amp;amp;lt;add&amp;amp;gt;ḷa&amp;amp;lt;/add&amp;amp;gt;&amp;amp;lt;/subst&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
The other option involves putting only vowel markers in the &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;add&amp;amp;gt;&amp;lt;/code&amp;gt; and &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;del&amp;amp;gt;&amp;lt;/code&amp;gt;. This is more precise, but since a dependent &amp;quot;a&amp;quot; cannot actually be represented in Indic scripts, it requires that the transcription be displayed in Roman transliteration (or otherwise some ad-hoc processing will be necessary).&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
&amp;amp;lt;subst&amp;amp;gt;ḷ&amp;amp;lt;del type=&amp;quot;cancelled&amp;quot;&amp;amp;gt;o&amp;amp;lt;/del&amp;amp;gt;&amp;amp;lt;add place=&amp;quot;implicit&amp;quot;&amp;amp;gt;a&amp;amp;lt;/add&amp;amp;gt;&amp;amp;lt;/subst&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
Of course not all projects will require markup of this granularity. &lt;br /&gt;
&lt;br /&gt;
=== Floating consonants ===&lt;br /&gt;
When an orthographically dependent consonant is separated from another consonant, for instance by a binding hole, it can generally be transcribed without any special markup:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
śa&amp;amp;lt;space type=&amp;quot;binding-hole&amp;quot;/&amp;amp;gt;ḥ&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
But when the orthographic sequence of consonants differs from the canonical sequence of consonants, this is not possible. &lt;br /&gt;
&lt;br /&gt;
[[File:Margam.png]]&lt;br /&gt;
&lt;br /&gt;
If necessary the out-of-sequence consonant could be specifically marked as such:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
māgg&amp;amp;lt;space type=&amp;quot;binding-hole&amp;quot;/&amp;amp;gt;&amp;amp;lt;g ref=&amp;quot;#floating-r&amp;quot;&amp;amp;gt;r&amp;amp;lt;/g&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
In this case a processor should be able to &amp;quot;fix&amp;quot; the sequence of characters given this markup. But the group advised that this level of markup was not generally required.&lt;br /&gt;
&lt;br /&gt;
=== Floating vowel markers ===&lt;br /&gt;
Floating vowel markers present an analagous case to floating consonants and should probably be encoded similarly if required.&lt;/div&gt;</summary>
		<author><name>Andrew ollett</name></author>
	</entry>
	<entry>
		<id>https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16174</id>
		<title>SIG:IndicTexts</title>
		<link rel="alternate" type="text/html" href="https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16174"/>
		<updated>2018-04-10T02:46:08Z</updated>

		<summary type="html">&lt;p&gt;Andrew ollett: &lt;/p&gt;
&lt;hr /&gt;
&lt;div&gt;The purpose of the TEI Special Interest Group “Indic Texts” is to&lt;br /&gt;
allow scholars engaged in the study of Indic texts to develop and&lt;br /&gt;
document best practices in applying the TEI’s Guidelines to these&lt;br /&gt;
kinds of texts.  To participate, please join the mailing list at&lt;br /&gt;
http://lists.lists.tei-c.org/mailman/listinfo/indic-texts.&lt;br /&gt;
&lt;br /&gt;
There are several respects in which the applicability of the TEI&lt;br /&gt;
guidelines to these texts is less than obvious. These relate to&lt;br /&gt;
distinctive features of Indic textuality, including:&lt;br /&gt;
&lt;br /&gt;
* the use of syllabic scripts, and the non-coincidence of grapheme- (syllable-) and word-boundaries;&lt;br /&gt;
* the application of phonotactic rules (sandhi) that further obscures the boundaries between words; extensive compounding;&lt;br /&gt;
* the use of distinctive media and writing supports (such as birch bark, palm leaves, and copper plates);&lt;br /&gt;
* distinctive metrical patterns with different types of caesuras;&lt;br /&gt;
* the prominence of the commentary as a genre, and the depth of intertextual relations this implies;&lt;br /&gt;
* the frequent reuse of texts in other texts, which requires careful and deliberate application of the &amp;quot;quoteLike&amp;quot; module.&lt;br /&gt;
&lt;br /&gt;
The expected outcome of the SIG’s work is a practical guide that&lt;br /&gt;
analyzes common cases in the markup of Indic texts and proposes ways&lt;br /&gt;
in which the analytical tools provided by the TEI Guidelines might&lt;br /&gt;
best be used in these cases, discussing benefits and drawbacks of the&lt;br /&gt;
solutions possible. Ideally, this guide will become a part of the TEI&lt;br /&gt;
Guidelines.&lt;br /&gt;
&lt;br /&gt;
&lt;br /&gt;
== Manuscript Transcription ==&lt;br /&gt;
&lt;br /&gt;
Canonically, an akṣara makes up a single &amp;quot;grapheme,&amp;quot; and this is reflected in Unicode representations of Indic scripts, where consonants and independent vowels are encoded first, and then vowel-markers (and dependent consonants like &#039;&#039;anusvāraḥ&#039;&#039; and &#039;&#039;visargaḥ&#039;&#039;) are encoded subsequently as combining characters. Unless marked with a combining vowel character, or a cancellation character, consonants are understood to have an inherent vowel &#039;&#039;a&#039;&#039;. The sequence of consonants within conjuncts is also canonically the same as their phonological sequence. Thus in the conjunct &amp;quot;rg&amp;quot;, the &amp;quot;r&amp;quot; is represented before the &amp;quot;g&amp;quot; in transliteration, in Devanagari र्ग (0930 + 094D + 0917) and in Kannada ರ್ಗ (0CB0 + 0CCD + 0C97), although it is rendered on top of the &amp;quot;g&amp;quot; in Devanagari and to the right of the &amp;quot;g&amp;quot; in Kannada. &lt;br /&gt;
&lt;br /&gt;
=== Cancelling dependent vowels ===&lt;br /&gt;
In manuscripts, dependent vowel markers can be cancelled, and the consonant is then read with the inherent vowel &#039;&#039;a&#039;&#039;. &lt;br /&gt;
&lt;br /&gt;
[[File:Ondondarol.png]]&lt;br /&gt;
&lt;br /&gt;
&lt;br /&gt;
If you want to encode this kind of change, there are technical problems, whether one is using an Indic script or an alphabetic transliteration system (like IAST or ISO-15919):&lt;br /&gt;
* In Indic scripts, rendering problems are likely if the cancelled vowel marker is enclosed within the &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;del&amp;amp;gt;&amp;amp;lt;/del&amp;amp;gt;&amp;lt;/code&amp;gt; tags, since it is a combining character;&lt;br /&gt;
* In transliteration, the deletion of one vowel must be accompanied by the addition of the inherent vowel, although there is no addition marked as such in the manuscript.&lt;br /&gt;
&lt;br /&gt;
The consensus seems to be: wrap the consonant, to which these modifications are referred, in the &amp;amp;lt;subst&amp;amp;gt; element, and use the &amp;amp;lt;@place=&amp;quot;implicit&amp;quot;&amp;amp;gt; attribute on &amp;amp;lt;add&amp;amp;gt; in reference to the vowel, as follows:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
&amp;amp;lt;subst&amp;amp;gt;ḷ&amp;amp;lt;del type=&amp;quot;cancelled&amp;quot;&amp;amp;gt;o&amp;amp;lt;/del&amp;amp;gt;&amp;amp;lt;add place=&amp;quot;implicit&amp;quot;&amp;amp;gt;a&amp;amp;lt;/add&amp;amp;gt;&amp;amp;lt;/subst&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
(Of course projects might not require this degree of markup.)&lt;br /&gt;
&lt;br /&gt;
=== Floating consonants ===&lt;br /&gt;
When an orthographically dependent consonant is separated from another consonant, for instance by a binding hole, it can generally be transcribed without any special markup:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
śa&amp;amp;lt;space type=&amp;quot;binding-hole&amp;quot;/&amp;amp;gt;ḥ&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
But when the orthographic sequence of consonants differs from the canonical sequence of consonants, this is not possible. &lt;br /&gt;
&lt;br /&gt;
[[File:Margam.png]]&lt;br /&gt;
&lt;br /&gt;
If necessary the out-of-sequence consonant could be specifically marked as such:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
māgg&amp;amp;lt;space type=&amp;quot;binding-hole&amp;quot;/&amp;amp;gt;&amp;amp;lt;g ref=&amp;quot;#floating-r&amp;quot;&amp;amp;gt;r&amp;amp;lt;/g&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
In this case a processor should be able to &amp;quot;fix&amp;quot; the sequence of characters given this markup. But the group advised that this level of markup was not generally required.&lt;br /&gt;
&lt;br /&gt;
=== Floating vowel markers ===&lt;br /&gt;
Floating vowel markers present an analagous case to floating consonants and should probably be encoded similarly if required.&lt;/div&gt;</summary>
		<author><name>Andrew ollett</name></author>
	</entry>
	<entry>
		<id>https://wiki.tei-c.org/index.php?title=File:Ondondarol.png&amp;diff=16173</id>
		<title>File:Ondondarol.png</title>
		<link rel="alternate" type="text/html" href="https://wiki.tei-c.org/index.php?title=File:Ondondarol.png&amp;diff=16173"/>
		<updated>2018-04-10T02:45:11Z</updated>

		<summary type="html">&lt;p&gt;Andrew ollett: &lt;/p&gt;
&lt;hr /&gt;
&lt;div&gt;&lt;/div&gt;</summary>
		<author><name>Andrew ollett</name></author>
	</entry>
	<entry>
		<id>https://wiki.tei-c.org/index.php?title=File:Margam.png&amp;diff=16172</id>
		<title>File:Margam.png</title>
		<link rel="alternate" type="text/html" href="https://wiki.tei-c.org/index.php?title=File:Margam.png&amp;diff=16172"/>
		<updated>2018-04-10T02:44:20Z</updated>

		<summary type="html">&lt;p&gt;Andrew ollett: &lt;/p&gt;
&lt;hr /&gt;
&lt;div&gt;&lt;/div&gt;</summary>
		<author><name>Andrew ollett</name></author>
	</entry>
	<entry>
		<id>https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16165</id>
		<title>SIG:IndicTexts</title>
		<link rel="alternate" type="text/html" href="https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16165"/>
		<updated>2018-04-04T19:29:59Z</updated>

		<summary type="html">&lt;p&gt;Andrew ollett: &lt;/p&gt;
&lt;hr /&gt;
&lt;div&gt;The purpose of the TEI Special Interest Group “Indic Texts” is to&lt;br /&gt;
allow scholars engaged in the study of Indic texts to develop and&lt;br /&gt;
document best practices in applying the TEI’s Guidelines to these&lt;br /&gt;
kinds of texts.  To participate, please join the mailing list at&lt;br /&gt;
http://lists.lists.tei-c.org/mailman/listinfo/indic-texts.&lt;br /&gt;
&lt;br /&gt;
There are several respects in which the applicability of the TEI&lt;br /&gt;
guidelines to these texts is less than obvious. These relate to&lt;br /&gt;
distinctive features of Indic textuality, including:&lt;br /&gt;
&lt;br /&gt;
* the use of syllabic scripts, and the non-coincidence of grapheme- (syllable-) and word-boundaries;&lt;br /&gt;
* the application of phonotactic rules (sandhi) that further obscures the boundaries between words; extensive compounding;&lt;br /&gt;
* the use of distinctive media and writing supports (such as birch bark, palm leaves, and copper plates);&lt;br /&gt;
* distinctive metrical patterns with different types of caesuras;&lt;br /&gt;
* the prominence of the commentary as a genre, and the depth of intertextual relations this implies;&lt;br /&gt;
* the frequent reuse of texts in other texts, which requires careful and deliberate application of the &amp;quot;quoteLike&amp;quot; module.&lt;br /&gt;
&lt;br /&gt;
The expected outcome of the SIG’s work is a practical guide that&lt;br /&gt;
analyzes common cases in the markup of Indic texts and proposes ways&lt;br /&gt;
in which the analytical tools provided by the TEI Guidelines might&lt;br /&gt;
best be used in these cases, discussing benefits and drawbacks of the&lt;br /&gt;
solutions possible. Ideally, this guide will become a part of the TEI&lt;br /&gt;
Guidelines.&lt;br /&gt;
&lt;br /&gt;
&lt;br /&gt;
== Manuscript Transcription ==&lt;br /&gt;
&lt;br /&gt;
Canonically, an akṣara makes up a single &amp;quot;grapheme,&amp;quot; and this is reflected in Unicode representations of Indic scripts, where consonants and independent vowels are encoded first, and then vowel-markers (and dependent consonants like &#039;&#039;anusvāraḥ&#039;&#039; and &#039;&#039;visargaḥ&#039;&#039;) are encoded subsequently as combining characters. Unless marked with a combining vowel character, or a cancellation character, consonants are understood to have an inherent vowel &#039;&#039;a&#039;&#039;. The sequence of consonants within conjuncts is also canonically the same as their phonological sequence. Thus in the conjunct &amp;quot;rg&amp;quot;, the &amp;quot;r&amp;quot; is represented before the &amp;quot;g&amp;quot; in transliteration, in Devanagari र्ग (0930 + 094D + 0917) and in Kannada ರ್ಗ (0CB0 + 0CCD + 0C97), although it is rendered on top of the &amp;quot;g&amp;quot; in Devanagari and to the right of the &amp;quot;g&amp;quot; in Kannada. &lt;br /&gt;
&lt;br /&gt;
=== Cancelling dependent vowels ===&lt;br /&gt;
In manuscripts, dependent vowel markers can be cancelled, and the consonant is then read with the inherent vowel &#039;&#039;a&#039;&#039;. If you want to encode this kind of change, there are technical problems, whether one is using an Indic script or an alphabetic transliteration system (like IAST or ISO-15919):&lt;br /&gt;
* In Indic scripts, rendering problems are likely if the cancelled vowel marker is enclosed within the &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;del&amp;amp;gt;&amp;amp;lt;/del&amp;amp;gt;&amp;lt;/code&amp;gt; tags, since it is a combining character;&lt;br /&gt;
* In transliteration, the deletion of one vowel must be accompanied by the addition of the inherent vowel, although there is no addition marked as such in the manuscript.&lt;br /&gt;
&lt;br /&gt;
The consensus seems to be: wrap the consonant, to which these modifications are referred, in the &amp;amp;lt;subst&amp;amp;gt; element, and use the &amp;amp;lt;@place=&amp;quot;implicit&amp;quot;&amp;amp;gt; attribute on &amp;amp;lt;add&amp;amp;gt; in reference to the vowel, as follows:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
&amp;amp;lt;subst&amp;amp;gt;ḷ&amp;amp;lt;del type=&amp;quot;cancelled&amp;quot;&amp;amp;gt;o&amp;amp;lt;/del&amp;amp;gt;&amp;amp;lt;add place=&amp;quot;implicit&amp;quot;&amp;amp;gt;a&amp;amp;lt;/add&amp;amp;gt;&amp;amp;lt;/subst&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
(Of course projects might not require this degree of markup.)&lt;br /&gt;
&lt;br /&gt;
=== Floating consonants ===&lt;br /&gt;
When an orthographically dependent consonant is separated from another consonant, for instance by a binding hole, it can generally be transcribed without any special markup:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
śa&amp;amp;lt;space type=&amp;quot;binding-hole&amp;quot;/&amp;amp;gt;ḥ&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
But when the orthographic sequence of consonants differs from the canonical sequence of consonants, this is not possible, and it necessary the out-of-sequence consonant could be specifically marked as such:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
māgg&amp;amp;lt;space type=&amp;quot;binding-hole&amp;quot;/&amp;amp;gt;&amp;amp;lt;g ref=&amp;quot;#floating-r&amp;quot;&amp;amp;gt;r&amp;amp;lt;/g&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
In this case a processor should be able to &amp;quot;fix&amp;quot; the sequence of characters given this markup.&lt;br /&gt;
&lt;br /&gt;
=== Floating vowel markers ===&lt;/div&gt;</summary>
		<author><name>Andrew ollett</name></author>
	</entry>
	<entry>
		<id>https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16164</id>
		<title>SIG:IndicTexts</title>
		<link rel="alternate" type="text/html" href="https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16164"/>
		<updated>2018-04-04T19:29:30Z</updated>

		<summary type="html">&lt;p&gt;Andrew ollett: &lt;/p&gt;
&lt;hr /&gt;
&lt;div&gt;The purpose of the TEI Special Interest Group “Indic Texts” is to&lt;br /&gt;
allow scholars engaged in the study of Indic texts to develop and&lt;br /&gt;
document best practices in applying the TEI’s Guidelines to these&lt;br /&gt;
kinds of texts.  To participate, please join the mailing list at&lt;br /&gt;
http://lists.lists.tei-c.org/mailman/listinfo/indic-texts.&lt;br /&gt;
&lt;br /&gt;
There are several respects in which the applicability of the TEI&lt;br /&gt;
guidelines to these texts is less than obvious. These relate to&lt;br /&gt;
distinctive features of Indic textuality, including:&lt;br /&gt;
&lt;br /&gt;
* the use of syllabic scripts, and the non-coincidence of grapheme- (syllable-) and word-boundaries;&lt;br /&gt;
* the application of phonotactic rules (sandhi) that further obscures the boundaries between words; extensive compounding;&lt;br /&gt;
* the use of distinctive media and writing supports (such as birch bark, palm leaves, and copper plates);&lt;br /&gt;
* distinctive metrical patterns with different types of caesuras;&lt;br /&gt;
* the prominence of the commentary as a genre, and the depth of intertextual relations this implies;&lt;br /&gt;
* the frequent reuse of texts in other texts, which requires careful and deliberate application of the &amp;quot;quoteLike&amp;quot; module.&lt;br /&gt;
&lt;br /&gt;
The expected outcome of the SIG’s work is a practical guide that&lt;br /&gt;
analyzes common cases in the markup of Indic texts and proposes ways&lt;br /&gt;
in which the analytical tools provided by the TEI Guidelines might&lt;br /&gt;
best be used in these cases, discussing benefits and drawbacks of the&lt;br /&gt;
solutions possible. Ideally, this guide will become a part of the TEI&lt;br /&gt;
Guidelines.&lt;br /&gt;
&lt;br /&gt;
&lt;br /&gt;
== Manuscript Transcription ==&lt;br /&gt;
&lt;br /&gt;
Canonically, an akṣara makes up a single &amp;quot;grapheme,&amp;quot; and this is reflected in Unicode representations of Indic scripts, where consonants and independent vowels are encoded first, and then vowel-markers (and dependent consonants like &#039;&#039;anusvāraḥ&#039;&#039; and &#039;&#039;visargaḥ&#039;&#039;) are encoded subsequently as combining characters. Unless marked with a combining vowel character, or a cancellation character, consonants are understood to have an inherent vowel &#039;&#039;a&#039;&#039;. The sequence of consonants within conjuncts is also canonically the same as their phonological sequence. Thus in the conjunct &amp;quot;rg&amp;quot;, the &amp;quot;r&amp;quot; is represented before the &amp;quot;g&amp;quot; in transliteration, in Devanagari र्ग (0930 + 094D + 0917) and in Kannada ರ್ಗ (0CB0 + 0CCD + 0C97), although it is rendered on top of the &amp;quot;g&amp;quot; in Devanagari and to the right of the &amp;quot;g&amp;quot; in Kannada. &lt;br /&gt;
&lt;br /&gt;
=== Cancelling dependent vowels ===&lt;br /&gt;
In manuscripts, dependent vowel markers can be cancelled, and the consonant is then read with the inherent vowel &#039;&#039;a&#039;&#039;. If you want to encode this kind of change, there are technical problems, whether one is using an Indic script or an alphabetic transliteration system (like IAST or ISO-15919):&lt;br /&gt;
* In Indic scripts, rendering problems are likely if the cancelled vowel marker is enclosed within the &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;del&amp;amp;gt;&amp;amp;lt;/del&amp;amp;gt;&amp;lt;/code&amp;gt; tags, since it is a combining character;&lt;br /&gt;
* In transliteration, the deletion of one vowel must be accompanied by the addition of the inherent vowel, although there is no addition marked as such in the manuscript.&lt;br /&gt;
&lt;br /&gt;
The consensus seems to be: wrap the consonant, to which these modifications are referred, in the &amp;amp;lt;subst&amp;amp;gt; element, and use the &amp;amp;lt;@place=&amp;quot;implicit&amp;quot;&amp;amp;gt; attribute on &amp;amp;lt;add&amp;amp;gt; in reference to the vowel, as follows:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
&amp;amp;lt;subst&amp;amp;gt;ḷ&amp;amp;lt;del type=&amp;quot;cancelled&amp;quot;&amp;amp;gt;o&amp;amp;lt;/del&amp;amp;gt;&amp;amp;lt;add place=&amp;quot;implicit&amp;quot;&amp;amp;gt;a&amp;amp;lt;/add&amp;amp;gt;&amp;amp;lt;/subst&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
(Of course projects might not require this degree of markup.)&lt;br /&gt;
&lt;br /&gt;
=== Floating consonants ===&lt;br /&gt;
When an orthographically dependent consonant is separated from another consonant, for instance by a binding hole, it can generally be transcribed without any special markup:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
&amp;amp;lt;śa&amp;amp;lt;space type=&amp;quot;binding-hole&amp;quot;/&amp;amp;gt;ḥ&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
But when the orthographic sequence of consonants differs from the canonical sequence of consonants, this is not possible, and it necessary the out-of-sequence consonant could be specifically marked as such:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
māgg&amp;amp;lt;space type=&amp;quot;binding-hole&amp;quot;/&amp;amp;gt;&amp;amp;lt;g ref=&amp;quot;#floating-r&amp;quot;&amp;amp;gt;r&amp;amp;lt;/g&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;br /&gt;
&lt;br /&gt;
In this case a processor should be able to &amp;quot;fix&amp;quot; the sequence of characters given this markup.&lt;br /&gt;
&lt;br /&gt;
=== Floating vowel markers ===&lt;/div&gt;</summary>
		<author><name>Andrew ollett</name></author>
	</entry>
	<entry>
		<id>https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16163</id>
		<title>SIG:IndicTexts</title>
		<link rel="alternate" type="text/html" href="https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16163"/>
		<updated>2018-04-04T18:58:34Z</updated>

		<summary type="html">&lt;p&gt;Andrew ollett: &lt;/p&gt;
&lt;hr /&gt;
&lt;div&gt;The purpose of the TEI Special Interest Group “Indic Texts” is to&lt;br /&gt;
allow scholars engaged in the study of Indic texts to develop and&lt;br /&gt;
document best practices in applying the TEI’s Guidelines to these&lt;br /&gt;
kinds of texts.  To participate, please join the mailing list at&lt;br /&gt;
http://lists.lists.tei-c.org/mailman/listinfo/indic-texts.&lt;br /&gt;
&lt;br /&gt;
There are several respects in which the applicability of the TEI&lt;br /&gt;
guidelines to these texts is less than obvious. These relate to&lt;br /&gt;
distinctive features of Indic textuality, including:&lt;br /&gt;
&lt;br /&gt;
* the use of syllabic scripts, and the non-coincidence of grapheme- (syllable-) and word-boundaries;&lt;br /&gt;
* the application of phonotactic rules (sandhi) that further obscures the boundaries between words; extensive compounding;&lt;br /&gt;
* the use of distinctive media and writing supports (such as birch bark, palm leaves, and copper plates);&lt;br /&gt;
* distinctive metrical patterns with different types of caesuras;&lt;br /&gt;
* the prominence of the commentary as a genre, and the depth of intertextual relations this implies;&lt;br /&gt;
* the frequent reuse of texts in other texts, which requires careful and deliberate application of the &amp;quot;quoteLike&amp;quot; module.&lt;br /&gt;
&lt;br /&gt;
The expected outcome of the SIG’s work is a practical guide that&lt;br /&gt;
analyzes common cases in the markup of Indic texts and proposes ways&lt;br /&gt;
in which the analytical tools provided by the TEI Guidelines might&lt;br /&gt;
best be used in these cases, discussing benefits and drawbacks of the&lt;br /&gt;
solutions possible. Ideally, this guide will become a part of the TEI&lt;br /&gt;
Guidelines.&lt;br /&gt;
&lt;br /&gt;
&lt;br /&gt;
== Manuscript Transcription ==&lt;br /&gt;
&lt;br /&gt;
Canonically, an akṣara makes up a single &amp;quot;grapheme,&amp;quot; and this is reflected in Unicode representations of Indic scripts, where consonants and independent vowels are encoded first, and then vowel-markers (and dependent consonants like &#039;&#039;anusvāraḥ&#039;&#039; and &#039;&#039;visargaḥ&#039;&#039;) are encoded subsequently as combining characters. Unless marked with a combining vowel character, or a cancellation character, consonants are understood to have an inherent vowel &#039;&#039;a&#039;&#039;. The sequence of consonants within conjuncts is also canonically the same as their phonological sequence. Thus in the conjunct &amp;quot;rg&amp;quot;, the &amp;quot;r&amp;quot; is represented before the &amp;quot;g&amp;quot; in transliteration, in Devanagari र्ग (0930 + 094D + 0917) and in Kannada ರ್ಗ (0CB0 + 0CCD + 0C97), although it is rendered on top of the &amp;quot;g&amp;quot; in Devanagari and to the right of the &amp;quot;g&amp;quot; in Kannada. &lt;br /&gt;
&lt;br /&gt;
=== Cancelling dependent vowels ===&lt;br /&gt;
In manuscripts, dependent vowel markers can be cancelled, and the consonant is then read with the inherent vowel &#039;&#039;a&#039;&#039;. If you want to encode this kind of change, there are technical problems, whether one is using an Indic script or an alphabetic transliteration system (like IAST or ISO-15919):&lt;br /&gt;
* In Indic scripts, rendering problems are likely if the cancelled vowel marker is enclosed within the &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;del&amp;amp;gt;&amp;amp;lt;/del&amp;amp;gt;&amp;lt;/code&amp;gt; tags, since it is a combining character;&lt;br /&gt;
* In transliteration, the deletion of one vowel must be accompanied by the addition of the inherent vowel, although there is no addition marked as such in the manuscript.&lt;br /&gt;
&lt;br /&gt;
The consensus seems to be: wrap the consonant, to which these modifications are referred, in the &amp;amp;lt;subst&amp;amp;gt; element, and use the &amp;amp;lt;@place=&amp;quot;implicit&amp;quot;&amp;amp;gt; attribute on &amp;amp;lt;add&amp;amp;gt; in reference to the vowel, as follows:&lt;br /&gt;
&lt;br /&gt;
&amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&lt;br /&gt;
&amp;amp;lt;subst&amp;amp;gt;ḷ&amp;amp;lt;del type=&amp;quot;cancelled&amp;quot;&amp;amp;gt;o&amp;amp;lt;/del&amp;amp;gt;&amp;amp;lt;add place=&amp;quot;implicit&amp;quot;&amp;amp;gt;a&amp;amp;lt;/add&amp;amp;gt;&amp;amp;lt;/subst&amp;amp;gt;&lt;br /&gt;
&amp;lt;/code&amp;gt;&lt;/div&gt;</summary>
		<author><name>Andrew ollett</name></author>
	</entry>
	<entry>
		<id>https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16162</id>
		<title>SIG:IndicTexts</title>
		<link rel="alternate" type="text/html" href="https://wiki.tei-c.org/index.php?title=SIG:IndicTexts&amp;diff=16162"/>
		<updated>2018-04-04T17:31:49Z</updated>

		<summary type="html">&lt;p&gt;Andrew ollett: &lt;/p&gt;
&lt;hr /&gt;
&lt;div&gt;The purpose of the TEI Special Interest Group “Indic Texts” is to&lt;br /&gt;
allow scholars engaged in the study of Indic texts to develop and&lt;br /&gt;
document best practices in applying the TEI’s Guidelines to these&lt;br /&gt;
kinds of texts.  To participate, please join the mailing list at&lt;br /&gt;
http://lists.lists.tei-c.org/mailman/listinfo/indic-texts.&lt;br /&gt;
&lt;br /&gt;
There are several respects in which the applicability of the TEI&lt;br /&gt;
guidelines to these texts is less than obvious. These relate to&lt;br /&gt;
distinctive features of Indic textuality, including:&lt;br /&gt;
&lt;br /&gt;
* the use of syllabic scripts, and the non-coincidence of grapheme- (syllable-) and word-boundaries;&lt;br /&gt;
* the application of phonotactic rules (sandhi) that further obscures the boundaries between words; extensive compounding;&lt;br /&gt;
* the use of distinctive media and writing supports (such as birch bark, palm leaves, and copper plates);&lt;br /&gt;
* distinctive metrical patterns with different types of caesuras;&lt;br /&gt;
* the prominence of the commentary as a genre, and the depth of intertextual relations this implies;&lt;br /&gt;
* the frequent reuse of texts in other texts, which requires careful and deliberate application of the &amp;quot;quoteLike&amp;quot; module.&lt;br /&gt;
&lt;br /&gt;
The expected outcome of the SIG’s work is a practical guide that&lt;br /&gt;
analyzes common cases in the markup of Indic texts and proposes ways&lt;br /&gt;
in which the analytical tools provided by the TEI Guidelines might&lt;br /&gt;
best be used in these cases, discussing benefits and drawbacks of the&lt;br /&gt;
solutions possible. Ideally, this guide will become a part of the TEI&lt;br /&gt;
Guidelines.&lt;br /&gt;
&lt;br /&gt;
&lt;br /&gt;
== Manuscript Transcription ==&lt;br /&gt;
&lt;br /&gt;
Canonically, an akṣara makes up a single &amp;quot;grapheme,&amp;quot; and this is reflected in Unicode representations of Indic scripts, where consonants and independent vowels are encoded first, and then vowel-markers (and dependent consonants like &#039;&#039;anusvāraḥ&#039;&#039; and &#039;&#039;visargaḥ&#039;&#039;) are encoded subsequently as combining characters. Unless marked with a combining vowel character, or a cancellation character, consonants are understood to have an inherent vowel &#039;&#039;a&#039;&#039;.&lt;br /&gt;
&lt;br /&gt;
=== Cancelling dependent vowels ===&lt;br /&gt;
In manuscripts, dependent vowel markers can be cancelled, and the consonant is then read with the inherent vowel &#039;&#039;a&#039;&#039;. If you want to encode this kind of change, there are technical problems, whether one is using an Indic script or an alphabetic transliteration system (like IAST or ISO-15919):&lt;br /&gt;
* In Indic scripts, rendering problems are likely if the cancelled vowel marker is enclosed within the &amp;lt;code language=&amp;quot;xml&amp;quot;&amp;gt;&amp;amp;lt;del&amp;amp;gtl;&amp;amp;lt;/del&amp;amp;gt;&amp;lt;/code&amp;gt; tags, since it is a combining character;&lt;br /&gt;
* In transliteration, the deletion of one vowel must be accompanied by the addition of the inherent vowel, although there is no addition marked as such in the manuscript.&lt;/div&gt;</summary>
		<author><name>Andrew ollett</name></author>
	</entry>
</feed>