{"title":"Anshuman Pandey","link":[{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/","rel":"alternate"}},{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/feeds\/all.atom.xml","rel":"self"}}],"id":"https:\/\/pandey.github.io\/unicode\/","updated":"2014-05-09T16:43:39.794000-07:00","subtitle":"Anshuman Pandey","entry":[{"title":"Gujarati signs for transliterating\u00a0Arabic","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2014-05-09-gujarati-arabic.html","rel":"alternate"}},"published":"2014-05-09T16:43:39.794000-07:00","updated":"2014-05-09T16:43:39.794000-07:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2014-05-09:\/unicode\/posts\/2014-05-09-gujarati-arabic.html","summary":"<p>Usage of new diacritic marks in Ismaili&nbsp;orthography<\/p>","content":"<p>Ismaili communities such as the Ithnashari Khoja (&#8220;Twelver Shia&#8221;) and Agakhani Khoja \nuse a convention for transliterating Arabic into the Gujarati script. The \nconvention allows for the representation of Arabic letters and signs for which \nthere are no corresponding characters in Gujarati. The diacritics used in \nthe convention are written with Gujarati letters that mostly closely \napproximate the Arabic sound being represented. \nThe creation of the full set and the first documented printing of\nthese signs was undertaken by the Ithnashari Khoja publisher Gul\u0101mal\u012b Ism\u0101\u02beil of\nBhavnagar, Gujarat in 1901. The characters are now standard elements of the \nGujarati orthography used by these communities.\nSome of these diacritics are shown below, with \ncolor coding added, in an excerpt of a printed version of the Qur\u02be\u0101n in the \nGujarati&nbsp;script:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/gujarati_arabic.jpg\"><\/p>\n<p>I&#8217;ve written a <a href=\"http:\/\/std.dkuug.dk\/JTC1\/SC2\/WG2\/docs\/n4574.pdf\">proposal<\/a> \nto encode these characters as combining signs in the Gujarati block of the Unicode \nstandard. I am interested in communicating with users about these and other signs \nused in Gujarati for similar purposes. Thank you to Iqbal Akhtar for informing me \nabout these&nbsp;signs.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"gujarati"}}]},{"title":"A \u201cNew\u201d Khojki\u00a0Inscription","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2014-01-05-new-khojki-inscription.html","rel":"alternate"}},"published":"2014-01-05T15:39:00-08:00","updated":"2014-01-05T15:39:00-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2014-01-05:\/unicode\/posts\/2014-01-05-new-khojki-inscription.html","summary":"<p>It&#8217;s sort of a reverse Rorschach Test: if I see\na blank slate, my mind will fill it with orthographic \nimprints. Usually ethereal calligraphy, more recently, \ntranslucent epigraphy. I&#8217;ve always marvelled\nat the craftsmen who meticulously engrave complex scripts upon \nmarble facades and copper plates; bringing a blank \u2026<\/p>","content":"<p>It&#8217;s sort of a reverse Rorschach Test: if I see\na blank slate, my mind will fill it with orthographic \nimprints. Usually ethereal calligraphy, more recently, \ntranslucent epigraphy. I&#8217;ve always marvelled\nat the craftsmen who meticulously engrave complex scripts upon \nmarble facades and copper plates; bringing a blank slate \nto life with visible language. I recently had a chance to \nmaterialize my orthographic&nbsp;fantasies.<\/p>\n<p>In early December 2013, I attended an\nopen house at Equinox Studios in the Georgetown neighborhood of\nSeattle. The artists had flung open the doors\nof their studios for the public to experience\nthe spaces where inspiration meets the workbench. \nIn one space an artist swung a wrecking ball into massive, ornate glass \nsculptures. In another, iron-wrights lit up the sky with \npyrotechnic displays emanating from giant, intricate lattices \nand delicate cut-metal trees. While watching these \nindustrial artists cast light and sound into the night, a \nmore silent performance in one corner of the studio caught my&nbsp;ears.<\/p>\n<p>As I walked over I saw stacks of sandstone blocks \nand people hunched over tables. I peered over the shoulder of a \ncouple and saw them carving the logo of the\nSeattle Seahawks into one of these blocks. A few feet away, \nmetal workers fed wood into the blazing mouth of a \nfurnace. I watched as a pair of workers reached into the furnace \nwith a heavy metal yoke and extracted a crucible filled with \nbubbling liquid. Another worker collected finished blocks with \ncarvings and placed them on a metal rack resting upon a bed of\nblack sand. At her signal, the bearers of the crucible inched\nover to the rack and steadily poured the molten liquid into each\nblock. Further along the perimeter, a worker cracked the cooled \nblocks to release a metal tile, each bearing an \nembossed image of what has been etched into the&nbsp;block.<\/p>\n<p>I was engrossed by this process of sandstone etchings being\nformed into pictographs on steel and nearly forgot about the\nrest of the open house. Some of my companions wandered onto the\nother installations as I grabbed a sandstone block and \nfound a spot at the table. I began thinking about what \nI would want to see wrought in steel. Luckily, my friend \nsaved me from writer&#8217;s, um, block by exclaiming\n&#8220;Oh, write my last name in&nbsp;Hindi!&#8221;.<\/p>\n<p>For the next 45 minutes I stood hunched over my block,\nsketching her name first in pencil, then lightly tracing it with a\ncarving tool, then going over the grooves repeatedly until the\ndepths reached a quarter of an inch. Then the devil of details\narrived and he sat on my shoulder as I tried to coax the stone\nto yield curves that resemble those of inked letters. When\nI finished, I brushed off my shoulder and handed the finished\nblock to my friend, who ran her finger over the grooved of the\netched imprint of her name in reverse. Thumbs up. She passed it\nonto the metal worker, who sprayed it with a graphite coat and\nset it on the rack. We watched as the two crucible bearers\narrived with a fresh trove of molten steel and poured it onto\nthe&nbsp;carving&#8230;<\/p>\n<p>The experience at Equinox loomed in my mind. I still wasn&#8217;t \nsure what I wanted to see cast in metal. A few days later it \nstruck me while I was researching variant letter-forms used in\nKhojki manuscripts while listening to Raageshwari Loomba&#8217;s\nrendition of the Ismaili <i>ginan<\/i> &#8220;Aaye Rahim Raheman&#8221; by\nImam Begam. I read the <i>ginan<\/i> in a printed Khojki book earlier in\nthe&nbsp;month:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/aaye_rahem_raheman_khojki.jpg\"><\/p>\n<p>I wondered if I could reproduce the\nletter-forms that Laljibhai Devraj had cut in Germany in 1903\nfor the first ever Khojki metal types, which he used at his\nKhoja Sindhi Printing Press in Bombay. I thought it\nmight be nice to do an etching of &#8220;Aaye Rahim Rahman&#8221;, the\n<i>ginan<\/i> which inspired me to think about a Khojki etching\nin the first place, but that seemed a bit ambitious. After all,\nI had only done one etching. So, I thought of the next best\nthing I would want to etch in&nbsp;Khojki&#8230;<\/p>\n<p>But first, I would need my blank slate. My friend contacted\nAlair Wells, a talented sculptress and metal artist, whose\nstudio <a href=\"http:\/\/tinderheartmetals.com\/\">Tinder Heart\nMetals<\/a> at Equinox was hosting the metal working during the \nopen house. I described my idea to Alair, who was excited \nto help. We produced a wooden cast for the block by\nusing four 2&#8221; x 4&#8221;s and affixing it to a plywood board. Then\nAlair cut an 8&#8221; x 10&#8221; block from a pine board and we fixed that\nto the plywood. We poured Washington grade 7 white silica into\nthe cast to measure the amount of sand we would need. Then\nafter we placed the silica into a bucket, Alair and my\nfriend mixed a catalyst and bonding agent into it, mixed\nvigorously for minutes, and then poured the preparation into the\nblock. The 25 lb. sandstone block turned out&nbsp;beautifully:<\/p>\n<p><img alt=\"image\" class=\"image-process-article-image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/02_silica_block_with_tools.jpg\"><\/p>\n<p>Now, I had my blank slate. Then, I chose my tools: \na <span class=\"st\">\u215c<\/span>&#8221; nib, a <span\nclass=\"st\">\u215b<\/span>&#8221; nib, and a needle point:<\/div><\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/08_tools.jpg\"><\/p>\n<p>The first step was to sketch out the Khojki text using a pencil.\nThe text must be etched as a mirror image so that the embossing\non the tile will have the correct orientation. I had thought of\nmaking a stencil by printing out the Khojki text in reverse and\ncutting out the letters using a fine blade, but the glyphs of\nthe only Khojki font I have do not possess the sort of modulated\nstrokes I desired for the etching. So, I decided to do a free\nhand rendering of the text in&nbsp;reverse:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/03_pencil_sketch.jpg\"><\/p>\n<p>I then performed an initial etching. I held my breath quite a\nbit during this part. Unlike writing on paper or composing in\ntypesetting software, you really cannot &#8216;erase&#8217; or &#8216;undo&#8217; an\nerron without having extra\nsilica and bonding agent on hand. I lightly scraped the <span\nclass=\"st\">\u215c<\/span>&#8221; nib over the sketch I&nbsp;made:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/04_etching_started.jpg\"><\/p>\n<p>The first phase of the etching of the upper portion of the block\nis shown below. The grooves are shallow. I would eventually\ndeepen the&nbsp;routes:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/05_etching_finished.jpg\"><\/p>\n<p>Then time to sketch out the text for the bottom portion of the\netching. I could have sketched out the entire text first, but I\nhad a feeling that my palm would just smudge the&nbsp;graphite.<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/06_more_sketches.jpg\"><\/p>\n<p>Below is the first run of the&nbsp;etching:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/07_all_etched.jpg\"><\/p>\n<p>I used the broad nib for the text of the upper portion. It \nproduced a very pleasant width and modulation that\nresembled the style of Khojki I wanted to&nbsp;emulate.<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/09_broad_nib_line.jpg\"><\/p>\n<p>Etching the tail of a vowel sign:<\/div><\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/10_broad_nib_modulation.jpg\"><\/p>\n<p>For the bottom portion, I used the medium&nbsp;nib. <\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/11_small_nib_line.jpg\"><\/p>\n<p>I used the needle point to take care of the small details, like\nthe dots and the terminals of&nbsp;letters:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/14_needle_hook.jpg\"><\/p>\n<p>Working on the&nbsp;details:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/12_needle_line.jpg\"><\/p>\n<p>I was concerned about the size of the dots. I feared that if I\nspaced them too closely, then the molten metal might just\nobliterate the sandstone in between and I&#8217;d end up with a blob\ninstead of three&nbsp;dots:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/13_needle_dots.jpg\"><\/p>\n<p>Here&#8217;s the before picture of the silica&nbsp;block:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/01_silica_block.jpg\"><\/p>\n<p>and the after&nbsp;picture:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/16_final_etching.jpg\"><\/p>\n<p>And a close up of the text: <\/div><\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/17_final_etching_face.jpg\"><\/p>\n<p>On the morning of New Year&#8217;s Eve, I returned the finished block\nto Alair, who did another pouring in Tacoma that evening. This\nis how it turned&nbsp;out:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/18_metal_etching.jpg\"><\/p>\n<p>Exactly the best Khojki text to fill that blank&nbsp;slate!<\/p>\n<p>While I was I developed my proposal for encoding Khojki in Unicode in\n2009, I was contacted by Irfan Gowani, who was enthusiastic\nabout the future encoding. We spoke intermittently, but I did\nnot meet Irfan until I returned to Seattle in 2013. During the\npast year Irfan and his wife Shelina have become wonderful\nfriends of mine. They are inspirational, creative, and lovely\npeople. And they are damned good cooks. One evening, a few days\nbefore Christmas, as I left their home after dinner Shelina\nhanded me two jars of homemade quince and fig jam, sourced from\nthe fruit of their own trees. Apart from the kababs they feed\nme, this was yet another testament of their giving nature. It\ngot me to thinking: what in the world could I offer them as an\nexpression of my&nbsp;friendship?<\/p>\n<p>What else would express my gratitude for their\npresence in my life more uniquely than an 8&#8221; x 10&#8221; x<span class=\"st\">\n\u00be<\/span>&#8221; metal plaque weighing 17 lb, which bears the names of\ntheir family embossed in Khojki by my own hands?! I have yet to\npresent it to them. But, I have a feeling that it will always remind \nthem of me&#8230; especially whenever they need to move&nbsp;it&#8230;<\/p>\n<p>Merry Christmas and Happy New Year, Irfan and&nbsp;Shelina!<\/p>\n<p><i>\u0101ye rahem rahem\u0101n, ab to rahem karoge,<\/i><br>\n<i>\u0101ye rahem rahem\u0101n, ab to rahem&nbsp;karoge.<\/i><\/p>\n<p>\n\n<i>ej\u012b tana mana dhana guru ne arpa\u1e47a k\u012bje<\/i><br>\n<i>to gin\u0101ne gin\u0101ne gin\u0101n, ab to rahem karoge.<\/i>\n<p>\n\n<i>ej\u012b d\u0101na sakh\u0101vat har dam k\u012bje, <\/i><br>\n<i>to d\u0101ne d\u0101ne d\u0101n, ab to rahem karoge.<\/i>\n<p>\n\n<i>ej\u012b saba gha\u1e6d ekaja rahem\u0101n k\u012bse,<\/i><br>\n<i>to \u015b\u0101ne \u015b\u0101ne \u015b\u0101n, ab to rahem karoge.<\/i>\n<p>\n\n<i>ej\u012b kahet \u012bm\u0101m begam mer\u0101 p\u012br hasan sh\u0101h,<\/i><br>\n<i>\u012bm\u0101ne \u012bm\u0101ne \u012bm\u0101n, ab to rahem&nbsp;karoge.<\/i>","category":{"@attributes":{"term":"articles"}}},{"title":"Siddham headstroke: abode of the inherent\u00a0vowel?","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2013-11-15-siddham-headstroke.html","rel":"alternate"}},"published":"2013-11-15T12:54:53.127000-08:00","updated":"2013-11-15T12:54:53.127000-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2013-11-15:\/unicode\/posts\/2013-11-15-siddham-headstroke.html","summary":"<p>Native Japanese analysis of the Indic inherent&nbsp;vowel<\/p>","content":"<p>The below excerpt from the <em>Shittan Hidenki<\/em> [\u6089\u66c7\u7955\u50b3\u8a18] (<em>Taish\u014d Shinsh\u016b Daiz\u014dky\u014d<\/em>, \nvol. 84, no. 270) of Shinpan [\u4fe1\u7bc4] shows an analysis of independent vowel letters, \ntheir dependent forms, and the combinations of these dependent forms with the Siddham \nconsonant letter <span class=\"caps\">KA<\/span>:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_myoten_inherent_a.jpg\"><\/p>\n<p>It also shows a horizontal stroke associated with Siddham vowel letter A that correlates \nwith the dependent forms of the other vowel letters. In Japanese, this stroke is known as \nthe \u30a2\u70b9 <em>a-ten<\/em> &#8220;<em>a<\/em> mark&#8221; and is considered to be an elongated form of the \u547d\u70b9 \n<em>my\u014d-ten<\/em> &#8220;life mark&#8221;. The <em>my\u014d-ten<\/em> is the initial brush stroke used in the writing \nof all Siddham letters. The <em>a-ten<\/em> is produced by extending the <em>my\u014d-ten<\/em> to the right. \nIt is not a true vowel sign; it is the headstroke of each consonant&nbsp;letter.<\/p>\n<p>My hunch is that the concept of the <em>a-ten<\/em> was developed by Japanese \nscholars as a way of explaining \nthe inherent vowel \/a\/ possessed by every consonant letter. All Indic vowels \nhave both independent and dependent forms, except for the letter A, which has only an \nindependent form. The <em>a-ten<\/em> raises several philosophical questions regarding the \nsound and forms of letters: How does one capture this inherent sound, which is part of the \nphonetic identity of each consonant letter, but which is graphically unmarked? Is it contained \nsomehow in the letter form? If so, where in the glyph does it reside? As it represents the \ninitial brush stroke used for writing the vowel letter A, it may be said to contain the graphical \nand phonetic essence of the&nbsp;letter.<\/p>\n<p>I raise the matter of the Siddham <a href=\"https:\/\/pandey.github.io\/posts\/2013-11-13-siddham-myoten.html\"><em>my\u014d-ten<\/em><\/a> \nand <em>a-ten<\/em> because it is significant from an ideographic perspective, as are the other \nelemental strokes identified in pedagogical texts. I am \ncurrently investigating the potential of encoding these elemental strokes, which I \nmentioned in my Siddham proposal as being out of scope for the basic encoding. I \nwelcome any information on the <em>my\u014d-ten<\/em> from users familiar with its philosophical \ninterpretations and its use in Siddham&nbsp;pedagogy.<\/p>\n<!-- toscheNovember 16, 2013 at 2:08 PM\n\nIn Japan, Siddham writing survived much more strongly than Sanskrit speech (which stopped \nrather quickly after Japan stopped trade with India), so I don't think \u30a2\u70b9 and \u547d\u70b9 are \nimportant in terms of phonetics. Looking at Siddham books in Japan, I can find that almost \nall strokes have their own names (obviously the most important of all is \u30a2\u70b9). So I agree \nwith you; I think they were named so because the Japanese monks needed to develop terminology \nin order to teach writing, and found that this stroke is found in every letter.\n\nAs most of the books that discusses Siddham mention \u30a2\u70b9 and \u547d\u70b9 in the text, it may be nice \nto have it encoded.\n\n-->","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"siddham"}}]},{"title":"Siddham my\u014d-ten: the essence of a\u00a0character?","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2013-11-13-siddham-myoten.html","rel":"alternate"}},"published":"2013-11-13T12:46:06.879000-08:00","updated":"2013-11-13T12:46:06.879000-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2013-11-13:\/unicode\/posts\/2013-11-13-siddham-myoten.html","summary":"<p>Notes on analysis of Siddham character&nbsp;strokes<\/p>","content":"<p>A short downward sloping horizontal stroke is shown in some historical\nand modern Siddham handbooks in connection with <span class=\"caps\">SIDDHAM<\/span> <span class=\"caps\">LETTER<\/span> A. In\n<em>Sacred Calligraphy of the East<\/em> (1981), John Stevens calls it a\n&#8216;variation&#8217; of the vowel letter&nbsp;A:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_myoten_variation_stevens.png\"><\/p>\n<p>In the excerpt, below, from the <em>Zusetsu Bonji<\/em> [\u56f3\u8aaa\u68b5\u5b57] (1974) of Kijun\nTokuzan [\u5fb3\u5c71\u6689\u7d14], the stroke is shown as a form of the vowel letter A,\non par with the dependent forms of the vowel letters A, I, and <span class=\"caps\">II<\/span>.<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_myoten_tokuzan.png\"><\/p>\n<p>Giry\u016b Kodama [\u5150\u7389\u7fa9\u9686] also shows it as a form of A in his <em>Bonji Hikkei<\/em>\n[\u68b5\u5b57\u5fc5\u643a]&nbsp;(1991):<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_myoten_kodama.jpg\"><\/p>\n<p>Although it is correlated with Siddham vowel letter A, it is not a\ntrue vowel sign, but represents the initial brush stroke used in\nwriting the vowel letter. Below is another excerpt from Tokuzan\n(1974), which shows the nine brush strokes used for constructing the\nvowel letter A, with the initial stroke&nbsp;highlighted:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_myoten_stroke_order.png\"><\/p>\n<p>The same stroke is used in the creation of all\nSiddham letters. The stroke is shown below in the forms for Siddham\nletters A, <span class=\"caps\">KHA<\/span>, <span class=\"caps\">HA<\/span>, <span class=\"caps\">RA<\/span>, <span class=\"caps\">VA<\/span>:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_a-ten_tokuzan_.png\"><\/p>\n<p>This stroke is known in Japanese as \u547d\u70b9 <em>my\u014d-ten<\/em> &#8220;life mark&#8221;. It\ncorresponds to the Chinese basic stroke \u9ede <em>di\u01cen<\/em> &#8220;dot&#8217;&#8221;, which is\nencoded in Unicode as \u31d4U+31D4 <span class=\"caps\">CJK<\/span> <span class=\"caps\">STROKE<\/span>&nbsp;D.<\/p>\n<p>I have not been able to conduct much research into the meaning of\nSiddham <em>my\u014d-ten<\/em>, but I have a hunch that it embodies the phonetic\npower of a character. I would theorize that the philosophy behind the\n<em>my\u014d-ten<\/em> arose as a way of explaining the absence of a mark for the\ninherent vowel. All Indic vowels have both independent and dependent\nforms, except for the letter A, which has only an independent form.\nThis raises several questions from a philosophical angle: How does one\ncapture this inherent sound, which is part of the graphical structure\nof each consonant letter, but which is unmarked? Is it contained\nsomehow in the letter form? If so, where in the glyph does it reside?\nAs it represents the initial brush stroke used for writing the vowel\nletter A, it may be said to contain the graphical and phonetic essence\nof the&nbsp;letter.<\/p>\n<p>I raise the issue of the Siddham <em>my\u014d-ten<\/em> because it is significant\nfrom an ideographic perspective, as are the other elemental strokes\nidentified in pedagogical texts. I am currently investigating the\npotential of encoding these elemental strokes, which I mentioned in my\nSiddham proposal as being out of scope for the basic encoding. I\nwelcome any information on the <em>my\u014d-ten<\/em> from users familiar with its\nphilosophical interpretations and its use in Siddham&nbsp;pedagogy.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"siddham"}}]},{"title":"Siddham nukta: a curious\u00a0innovation","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2013-11-05-siddham-nukta.html","rel":"alternate"}},"published":"2013-11-05T09:58:42.109000-08:00","updated":"2013-11-05T09:58:42.109000-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2013-11-05:\/unicode\/posts\/2013-11-05-siddham-nukta.html","summary":"<p>Usage of a modern Indic diacritic in&nbsp;Siddham<\/p>","content":"<p>The combining sign  \u093c   <span class=\"caps\">NUKTA<\/span> is used in Indic scripts for transcribing sounds\nfor which distinct characters do not natively exist in a writing\nsystem. The name of the character is derived from the Arabic word \u0646\u0642\u0637\u0629\n<em>nuq\u1e6dah<\/em> (simplified in Indic romanization as <em>nukta<\/em>)\n&#8220;dot&#8221;. The <span class=\"caps\">NUKTA<\/span> is used in Devanagari and related scripts of northern\nIndia for expressing sounds not commonly used in Indo-Aryan languages,\nmostly those that originate from Arabic and Persian. It is also used\nnatively in scripts such as Bengali and Tirhuta for distinguishing\nbetween characters that have nearly identical graphical structures.\nAlthough now used quite frequently in modern Indic scripts, the <span class=\"caps\">NUKTA<\/span>\nis not part of the traditional character repertoire based upon the\nrepresentation of Sanskrit phonology. As one might expect, the <span class=\"caps\">NUKTA<\/span>\nis not attested in historical Siddham materials as the usage of\nSiddham is largely restricted to the representation of&nbsp;Sanskrit.<\/p>\n<p>I was, therefore, surprised when, during research for my\nproposal to encode the script in Unicode, I stumbled across the \n<a href=\"http:\/\/www.mandalar.com\/BonjiDigitalDictionarySAMPLE\/member\/_Tattoo\/00Tattoo.html\">following sample<\/a> \nat the <a href=\"http:\/\/www.mandalar.com\">Mandalar<\/a> site \nshowing usage of <span class=\"caps\">NUKTA<\/span> in Siddham text&nbsp;as:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_nukta_mandalar_tattoo.jpg\"><\/p>\n<p>The above excerpt provides the\ntranslation or transliteration of English &#8220;tattoo&#8221; in Sanskrit,\nEnglish\/Hindi, and Japanese. The Sanskrit \u0935\u0947\u0927 <em>vedha<\/em> does\nnot truly translate as &#8220;tattoo&#8221;, but carries the connotation of a\n&#8220;piercing&#8221;, &#8220;puncturing&#8221;, or &#8220;perforation&#8221; (\u221a \u0935\u094d\u092f\u0927\u094d); the\nEnglish\/Hindi \u0924\u0924\u0942  <em>tat\u016b<\/em> is simply a transliteration of English &#8220;tattoo&#8221;\n(which also might be rendered \u091f\u0948\u091f\u0942 <em>t\u0323ait\u0323\u016b,<\/em> using retroflex\nletters instead of dentals); and \u0907\u0930\u0947\u095b\u0941\u092e\u093f <em>irezumi<\/em>, which\nis the transliteration into Siddham of the Japanese word for &#8220;tattoo&#8221;,\n\u5165\u308c\u58a8 <em>irezumi<\/em>.<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_nukta_mandalar_tattoo_irezumi.jpg\"><\/p>\n<p>Of interest is the usage of <span class=\"caps\">NUKTA<\/span> in\nwriting the word \u0907\u0930\u0947\u095b\u0941\u092e\u093f. In general, the <span class=\"caps\">NUKTA<\/span> is written with a\nletter that has the closest phonetic proximity to the target sound.\nHere, the <span class=\"caps\">NUKTA<\/span> is combined with <span class=\"caps\">SIDDHAM<\/span> <span class=\"caps\">LETTER<\/span> <span class=\"caps\">JA<\/span> (\/\u02a4\/) for\nrepresenting \/z\/. When <span class=\"caps\">NUKTA<\/span> occurs with a consonant to which a vowel\nsign is attached, then it is ordered in encoded text immediately after\nthe consonant and before the combining vowel sign: <span class=\"caps\">JA<\/span> + <span class=\"caps\">NUKTA<\/span> + <span class=\"caps\">VOWEL<\/span> \n<span class=\"caps\">SIGN<\/span> U. The positioning\nof the <span class=\"caps\">NUKTA<\/span> with regard to the base letter depends upon the shape of\nthe letter and the presence of any below-base vowel&nbsp;signs.<\/p>\n<p>The Mandalar site has a <a href=\"http:\/\/www.mandalar.com\/DisplayJ\/Bonji\/index2.html\">chart<\/a>\nthat shows the usage of <span class=\"caps\">NUKTA<\/span> with other consonants for representing\nfricatives and other&nbsp;sounds:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_nukta_consonants_mandalar.jpg\"><\/p>\n<p>The introduction of the <span class=\"caps\">NUKTA<\/span> in Siddham is a\nsurprising innovation. I have yet to investigate the issue fully, but\nat first glimpse the usage of <span class=\"caps\">NUKTA<\/span> suggests that modern users of\nSiddham seek to extend the script by adopting features used in Indic\nscripts. Such borrowings show that Siddham is quite alive in Japan and\nthat its usage is continually being extended beyond traditional\ncontexts. Indeed, the Mandalar site displays these innovations under\nthe banner of \u73fe\u4ee3\u6089\u66c7 <em>gendai shittan<\/em> or &#8220;modern&nbsp;Siddham&#8221;:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_mandalar_gendai_shittan.jpg\"><\/p>\n<p>In order to support such usage, I\nproposed the character for encoding as U+115C0 <span class=\"caps\">SIDDHAM<\/span> <span class=\"caps\">SIGN<\/span> <span class=\"caps\">NUKTA<\/span>. It\ncorresponds to characters such as U+093C <span class=\"caps\">DEVANAGARI<\/span> <span class=\"caps\">SIGN<\/span> <span class=\"caps\">NUKTA<\/span> and\npossesses the same&nbsp;properties.<\/p>\n<p>I am curious to know the\nhistory of <span class=\"caps\">NUKTA<\/span> in Siddham and the rationale of the <em>gendai shittan<\/em> \nuser(s) who first used <span class=\"caps\">NUKTA<\/span> in the script. Was the idea motivated\nby the usage of <span class=\"caps\">NUKTA<\/span> in modern Indic orthographies? Was it influenced\nby Unicode? Does the &#8216;dot&#8217; have ideographic&nbsp;interpretations?<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"siddham"}}]},{"title":"Siddham ornaments: beyond\u00a0punctuation","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2013-11-04-siddham-ornaments.html","rel":"alternate"}},"published":"2013-11-04T17:10:41.817000-08:00","updated":"2013-11-04T17:10:41.817000-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2013-11-04:\/unicode\/posts\/2013-11-04-siddham-ornaments.html","summary":"<p>Notes on ornamental signs used in&nbsp;Siddham<\/p>","content":"<p>Siddham has several characters that are used in manuscripts for marking the \nend of text sections. Some of these are shown in column \u2466&nbsp;below: <\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_section_marks_1.jpg\"><\/p>\n<p>Palaeographically, these section marks are similar to characters in other \nIndic scripts, such as \ud804\udc4d U+1104D <span class=\"caps\">BRAHMI<\/span> <span class=\"caps\">PUNCTUATION<\/span> <span class=\"caps\">LOTUS<\/span>, which are ornamental \nmarks. In general, these marks do not possess phonetic values or semantic \nmeaning beyond their function as terminations. However, as the above excerpt \nshows, in the Japanese analysis of Siddham, these marks additionally represent \nthe syllable&nbsp;&#8220;a\u1e43&#8221;.<\/p>\n<p>The extant sources contain several other ornaments beyond the three\nshown above. My research identified at least 14, which I <a href=\"http:\/\/std.dkuug.dk\/JTC1\/SC2\/WG2\/docs\/n4336.pdf\">proposed for\ninclusion<\/a> in the\nSiddham&nbsp;block:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_section_marks_proposed.png\"><\/p>\n<p>The section marks may be classified into five groups based upon their\ngraphical&nbsp;structure.<\/p>\n<p>In addition to their representation of &#8220;a\u1e43&#8221;, according to certain Japanese \nBuddhist traditions, these marks have esoteric connotations that offer \ninsights into the textual passages after which they are written. Moreover, \nthe graphical structure indicates other philosophical meanings, of which I \nam not fully aware or to which I am not privy! These characters would be \nconsidered ornaments or glyphic variants of a basic set of section marks \naccording to the character-glyph model of Unicode, but on account of the \nsemantics they possess beyond their function as punctuation, they have been \nencoded&nbsp;independently. <\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"siddham"}}]},{"title":"Siddham digits: any good\u00a0evidence?","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2013-11-02-siddham-digits.html","rel":"alternate"}},"published":"2013-11-02T01:07:47.574000-07:00","updated":"2013-11-02T01:07:47.574000-07:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2013-11-02:\/unicode\/posts\/2013-11-02-siddham-digits.html","summary":"<p>Notes on the status of digits in&nbsp;Siddham<\/p>","content":"<p>As I conducted research for my <a href=\"http:\/\/std.dkuug.dk\/JTC1\/SC2\/WG2\/docs\/n4294.pdf\">proposal to encode\nSiddham<\/a> in Unicode,\nI began to notice that the primary and secondary sources I examined\nall lacked a set of characters that are typically found in related\nIndic scripts that emerged during the 10th&nbsp;century&#8230;<\/p>\n<p>Digits.<\/p>\n<p>Well, this is not entirely true. As I discussed in section 3.13 of the\nproposal, Siddham contains two characters, that based upon their form\nand function, appear to be the remaining representatives of a set of\ndigits that were once likely part of the script. These characters\ncorrespond to a palaeographic form of an Indic digit <span class=\"caps\">TWO<\/span>. But, in\nSiddham this sign has lost its numeric value and it lives on only\ngraphically. Well, this is not entirely true, either. The sign may\nhave lost its numeric value, but the vestigial character retains its\nnumerical semantics in its function. This character is the repetition\nmark shown&nbsp;below:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/siddham_repetition_mark_example.jpg\"><\/p>\n<p>The sign in question appears twice in the excerpt (middle row,\ncharacters 3 and 4 from the left) and is transliterated as &#8220;h\u016b\u1e43&#8221;. This\ngloss is slightly misleading. The character itself does not represent\nthe syllable &#8220;h\u016b\u1e43&#8221;; rather, it represents a &#8216;ditto&#8217; mark. Moreover,\nalthough it is graphically based upon a palaeographic digit <span class=\"caps\">TWO<\/span>, it\nsimply indicates that the word or syllable that precedes it is to be\nrepeated; not twice, but just once. The two instances of the mark mean\nthat &#8220;h\u016b\u1e43&#8221; is to be recited three times, as shown. It appears that the\neditor of the original text provided the intended reading instead of\n&#8216;ditto&#8217; for the sake of&nbsp;clarity.<\/p>\n<p>This Siddham repetition mark corresponds glyphically to Devanagari\ndigit two (\u0968), but more so to the Gurmukhi digit two (\u0a68). It seems\nthat as Siddhamatrika made its way from south to east Asia, the\nBuddhist scribes and scholars who used it slowly discarded from the\nscript features that they didn&#8217;t really find necessary or useful. So,\naway went the digits 0-1 and 3-9. The digit two (\u0a68) seems to have\nstruck someone as a handy way of indicating a &#8220;doubling&#8221; when writing\nor copying texts. So, what we have now in Siddham is essentially an\neast Asian ideographic interpretation of an Indic character that was\noriginally and palaeographically a sign for a digit. The &#8216;\u0a68&#8217; is the\nonly numerical sign that I have seen in Siddham&nbsp;texts.<\/p>\n<p>Well, that is not entirely true either&#8230; In the latter stages of my\nresearch I found the following chart on a Japanese&nbsp;site:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/bonji_numerals_mandalar.jpg\"><\/p>\n<p>These &#8220;Bonji numerals&#8221; appear to be modern creations based upon\nDevanagari digits. I submitted a\n<a href=\"http:\/\/std.dkuug.dk\/JTC1\/SC2\/WG2\/docs\/n4467.pdf\">proposal<\/a> to encode\na set of digits in the Siddham block based upon these signs, but the\nUnicode Technical Committee (<span class=\"caps\">UTC<\/span>) requested additional evidence of the\nuse of digits in the&nbsp;script.<\/p>\n<p>Does anyone know of Siddham manuscripts, pedagogical texts, or other\nrecords that contain examples of digits? Hmmm, at this point, I may\neven accept \u5165\u308c\u58a8 <em>irezumi<\/em> as evidence, even if the <span class=\"caps\">UTC<\/span> might&nbsp;not!<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"siddham"}}]},{"title":"A distinctive Khojki letter for\u00a0qa?","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2013-10-12-khojki-qa.html","rel":"alternate"}},"published":"2013-10-12T19:40:36.542000-08:00","updated":"2013-10-12T19:40:36.542000-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2013-10-12:\/unicode\/posts\/2013-10-12-khojki-qa.html","summary":"<p>Is a character used for \/q\/ a distinctive letter or a glyphic&nbsp;variant?<\/p>","content":"<p>The following excerpt from a printed Khojki book contains two forms of\n<em>ka<\/em>. The <em>ka<\/em> highlighted in blue is the representative glyph of\nU+11208 <span class=\"caps\">KHOJKI<\/span> <span class=\"caps\">LETTER<\/span> <span class=\"caps\">KA<\/span> (to be published in Unicode 7), while the red\n<em>ka<\/em> is the alternate&nbsp;form.<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/khojki_ka_variant.jpg\"><\/p>\n<!-- _kala\u0304\u0303m_ -->\n\n<p>Here, the red <em>ka<\/em> is used in the word <em>kal\u0101m<\/em>, while the blue <em>ka<\/em> is\nused in the word <em>kal\u0101\u1e43m<\/em>, both are representations\nof the Arabic word \u06a9\u0644\u0627\u0645 <em>kal\u0101m<\/em> &#8220;composition, work&#8221;; the <em>anusv\u0101ra<\/em> in\n<em>kal\u0101\u1e43m<\/em> is not semantically&nbsp;significant.<\/p>\n<p>The alternate form is used in other sources for representing <em>ka<\/em>,\nsuggesting that the red <em>ka<\/em> is a glyphic&nbsp;variant:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/asani_1992_p283_variant_ka.jpg\"><\/p>\n<p>However, it also appears in other source, where it is used for\nrepresenting <em>qa<\/em> (the Arabic \u0642 <em>q\u0101f<\/em>), suggesting that it may indeed\nbe a distinctive letter, a possible *<span class=\"caps\">KHOJKI<\/span> <span class=\"caps\">LETTER<\/span> <span class=\"caps\">QA<\/span>:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/asani_1992_p294_qaf.jpg\"><\/p>\n<p>I have not yet seen this work, so I do not know the context in which\nthe letter is used for <em>q\u0101f<\/em>.<\/p>\n<p>The shape of <span class=\"caps\">KHOJKI<\/span> <span class=\"caps\">LETTER<\/span> <span class=\"caps\">KA<\/span> is fairly uniform across printed\nmaterials and manuscripts. It is of interest that both forms occur in\nthe same document, as shown in the first image, and in such close\nproximity. The excerpt is from a printed book and is likely a faithful\nreproduction of a manuscript. Analysis of several manuscript sources\nis necessary in order to determine whether the red <em>ka<\/em> is a glyphic\nvariant of <span class=\"caps\">KA<\/span> or a distinct Khojki letter <span class=\"caps\">QA<\/span>.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"khojki"}}]},{"title":"A Graphite font for\u00a0Khojki","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2013-10-10-khojki-graphite.html","rel":"alternate"}},"published":"2013-10-10T00:05:32.506000-07:00","updated":"2013-10-10T00:05:32.506000-07:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2013-10-10:\/unicode\/posts\/2013-10-10-khojki-graphite.html","summary":"<p>A Graphite font for&nbsp;Khojki<\/p>","content":"<p>As mentioned in \n<a href=\"2013-10-09-script-encoding-paradox.html\">&#8220;The Imminent Paradox of New Scripts in Unicode&#8221;<\/a>, \nI have begun to develop Graphite fonts for several of the historical and minor scripts\nthat I have proposed for inclusion in the Unicode Standard. The first\nof these fonts is for Khojki, which I \n<a href=\"http:\/\/std.dkuug.dk\/JTC1\/SC2\/WG2\/docs\/n3978.pdf\">proposed for encoding<\/a> in Unicode a few\nyears ago and will likely appear in The Unicode Standard, version 7.0,\nplanned for release in&nbsp;2014.<\/p>\n<p>Khojki is a medieval Indic script from the Sindh region, situated in\npresent day Pakistan, which is used for liturgical purposes by the\nNizari Ismaili community. It has been preserved by this minority\ncommunity for six centuries, both at home and in their diaspora. The\nscript has been used since the 15th century for manuscripts and since\nthe early 20th century for the printing of books. There are\norganizations, such as the <a href=\"http:\/\/www.iis.ac.uk\/\">Institute of Ismaili\nStudies<\/a> (<span class=\"caps\">IIS<\/span>) in London, which possesses a\nlarge collection of materials in Khojki, and several scholars\nworldwide that conduct research on these materials, that will benefit\nfrom this&nbsp;effort.<\/p>\n<p>I developed this font on a volunteer basis to meet the needs of the\n<span class=\"caps\">IIS<\/span>, which has a requirement for the input and display Unicode Khojki\nin order to carry out their cataloguing efforts. The font, named\nKhojkiGraphite, is based upon the KhojkiJiwa font designed by Pyarali\nJiwa of the United Kingdom years ago. Mr. Jiwa&#8217;s font is the only\nTrueType font for Khojki; however is it based upon a legacy encoding.\nIts glyph repertoire contains several consonant-vowel ligatures;\nhowever, there are several inconsistencies with these glyph. I&#8217;ve\nedited the glyphs to bring them into structural and optical alignment,\nfixed metrics, and added Graphite rendering rules in order to create\nthe first Unicode-encoded Khojki font. At present it works in\nOpenOffice and LibreOffice, in XeLaTeX and XeLaTeX, as well as in\nother applications that support&nbsp;Graphite.<\/p>\n<p>Here is an excerpt of a <em>gin\u0101n<\/em> typed in LibreOffice using&nbsp;KhojkiGraphite:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/khojki_graphite_specimen.jpg\"><\/p>\n<p>In order to provide users with an effective means of using the font,\nWafi Momin of <span class=\"caps\">IIS<\/span> has developed a basic keyboard using the Microsoft\nKeyboard Layout Creator (<span class=\"caps\">MSKLC<\/span>). So far, the combination of the font\nand the keyboard is helping <span class=\"caps\">IIS<\/span> to carry out its projects using modern\ntechnologies and to make their materials available in digital media\nusing common&nbsp;standards.<\/p>\n<p>I will release the KhojkiGraphite font and keyboard for public testing\nin the next few months. My goal is to release the package in\nconjunction with the publication of Unicode 7.0, so that users have\nimmediate support for Khojki, even though in limited&nbsp;environments.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"khojki"}}]},{"title":"The paradox of encoding new scripts in\u00a0Unicode","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2013-10-09-script-encoding-paradox.html","rel":"alternate"}},"published":"2013-10-09T00:09:18.590000-08:00","updated":"2013-10-09T00:09:18.590000-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2013-10-09:\/unicode\/posts\/2013-10-09-script-encoding-paradox.html","summary":"<p>The need for full support for each Unicode&nbsp;version.<\/p>","content":"<p>Over the past several years I have developed Unicode standards for\nnumerous historical and minor\n<a href=\"http:\/\/pandey.github.io\/unicode\">scripts<\/a> of South Asia. Some of my\nformally proposed encodings have been approved and are making \ntheir way through various stages of balloting as part of the\nprocess established by the International Organization for\nStandardization (<span class=\"caps\">ISO<\/span>) &#8212;- Unicode is also known as <span class=\"caps\">ISO<\/span>\/<span class=\"caps\">IEC<\/span> 10646 &#8212;-\nwhile others have been formally published in the&nbsp;standard. <\/p>\n<p>The inclusion of a script in Unicode is a major milestone. Native user and\nscholarly communities await the arrival of these new encodings from\nthe day that I contact them &#8212;- if not for years in the making &#8212;- \nand discuss my plans for developing an encoding for a particular script. \nIt takes nearly two years for an encoding to be published in Unicode after \nbeing approved by both the Unicode Technical Committee and <span class=\"caps\">ISO<\/span>&#8217;s <span class=\"caps\">WG2<\/span> group. \nBut, once a script is in Unicode the general perception of what &#8220;inclusion\nin the standard&#8221; means is far removed from the reality that is\nsuggested by such an&nbsp;&#8220;inclusion&#8221;.<\/p>\n<p>The disjunction between expectation and reality is truly a paradox: The\nencoding of a script in Unicode does not automatically imply that it will be\nsupported out of the box. As is the case with other historical and minor scripts\nencoded in Unicode, it is unlikely that major software houses will\nrush to provide support for these scripts right away, or if they do so\nat&nbsp;all!<\/p>\n<p>If these companies do not implement support for these scripts,\nusers will be unable to produce documents in their writing systems \nusing common applications such as Microsoft Word and Adobe InDesign. Users \nwill be unable to view websites containing text in these scripts using \nInternet Explorer, Firefox, and other browsers. Moreover, typographers \nwishing to design new, modern typefaces for these scripts will be hampered \nbecause their creations will be unusable in ubiquitous platforms such as Windows.\nThe reason for this is that OpenType, the common font standard, is\ndependent upon Microsoft&#8217;s Uniscribe rendering engine and a font will\nnot work in Windows until Uniscribe can handle the underlying script.\nCertainly, those who use Linux systems can rely on HarfBuzz, an open\nsource OpenType rendering engine; but, ultimately, from a practical\nperspective, Uniscribe support is important because Windows is the\nmost-commonly used operating system&nbsp;worldwide.<\/p>\n<p>What good is a Unicode standard for a script if there is no way for\nthe general public to actually make use of it? What good is a Unicode\nstandard for a script if implementation of it is dependent upon the\neconomic cost-benefit decisions of major software companies? Although\nhistorical and minority scripts and related linguistic ecologies may\nnot be profit generators, they are certainly valuable from humanist&nbsp;perspectives.<\/p>\n<p>So, what is the reality if Microsoft, Apple, Google, and others don&#8217;t support \nhistorical and minor scripts in their operating systems when a new\nversion of the Unicode standard is published&#8230; or ever? For now,\nthankfully, there we can provide font solutions using the Graphite\nrendering engine developed by <span class=\"caps\">SIL<\/span>. However, Graphite fonts are not\nsupported in Windows and, therefore, will not work in Microsoft\nOffice, Internet Explorer, and other commonly-used applications.\nApplications like LibreOffice and XeTeX support Graphite, but the user\nbase for these is limited. In the end, at least Graphite exists and\nthose of us who work with lesser-used writing systems owe gratitude to\n<span class=\"caps\">SIL<\/span> for having the foresight and imagination to provide a means for\nrendering writing systems independent of the whims of software&nbsp;house.<\/p>\n<p>I maintain the hope that Microsoft will continue to expand Uniscribe\nin order to support all scripts in the Unicode standard. But until the\ntime comes when there is mandatory synchronization between releases of\nthe standard and implementation, we must turn to other, albeit\nlimited, solutions. As part of this solution, I will build Graphite\nfonts for all the scripts I have brought into Unicode in order to\nprovide a basic level of support for these scripts in the Windows\nenvironment, even if the range of usage is limited. I am not a\nprofessional typopgrapher nor do I pretend to be, but such efforts are\njust an extension of the work I&#8217;ve&nbsp;started.<\/p>\n<p>But, I am an individual with limited resources and numerous\nconstraints. I encourage Microsoft, Apple, Google and others to meet me half way.\nAfter all, what good is a standard if no one cares to support&nbsp;it?<\/p>","category":{"@attributes":{"term":"unicode"}}},{"title":"Fraction signs in\u00a0Oriya","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2013-05-08-oriya-fraction-signs.html","rel":"alternate"}},"published":"2013-05-08T11:19:00-08:00","updated":"2013-05-08T11:19:00-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2013-05-08:\/unicode\/posts\/2013-05-08-oriya-fraction-signs.html","summary":"<p>Six fraction signs historically used in Oriya have been included in\nUnicode 6.0. These characters appear in written and printed documents,\nand were part of at least four different sets of Oriya metal fonts.\nThe fraction signs were commonly used until 1958, at which time the\nGovernment of India \u2026<\/p>","content":"<p>Six fraction signs historically used in Oriya have been included in\nUnicode 6.0. These characters appear in written and printed documents,\nand were part of at least four different sets of Oriya metal fonts.\nThe fraction signs were commonly used until 1958, at which time the\nGovernment of India adopted the decimal system for currency and the\nmetric system for weights and measures. The characters are now&nbsp;obsolete.<\/p>\n<p>The fraction signs are illustrated by <span class=\"caps\">R. J.<\/span> Grundy in the <span\nstyle=\"font-style:italic;\">Concise Oriya-English Dictionary<\/span>\n(Cuttack: Orissa Mission Press,&nbsp;1928):<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/oriya_fractions_grundy.jpg\"><\/p>\n<p>Amos Sutton shows only the quarter fractions in the <span\nstyle=\"font-style:italic;\">Introductory Grammar of the Oriya\nLanguage<\/span> (Cuttack: Baptist Mission Press,&nbsp;1831):<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/oriya_fractions_sutton.jpg\"><\/p>\n<p>These fraction signs are now part of the <a\nhref=\"http:\/\/www.unicode.org\/charts\/PDF\/U0B00.pdf\">Oriya block<\/a> in&nbsp;Unicode:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/oriya_fractions_unicode.jpg\"><\/p>\n<p>The Oriya fractions are related to the &#8216;currency numerator&#8217; signs used\nin Bengali for writing&nbsp;fractions:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/oriya_fractions_comp.jpg\"><\/p>\n<p>More information on the Oriya fraction signs is available in the \n<a href=\"http:\/\/www.dkuug.dk\/jtc1\/sc2\/wg2\/docs\/n3471.pdf\">original proposal to encode these signs<\/a>. \nInformation about the &#8216;Common Indic&#8217; fraction signs shown\nin the above image may be found in the \n<a href=\"http:\/\/std.dkuug.dk\/jtc1\/sc2\/wg2\/docs\/n3367.pdf\">proposal<\/a> \nto encode &#8216;Common Indic Number Forms&#8217; in&nbsp;Unicode.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"oriya"}}]},{"title":"Signs for representing Kashmiri vowels in\u00a0Sharada","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2011-05-12-sharada-kashmiri-signs.html","rel":"alternate"}},"published":"2011-05-12T12:31:00-07:00","updated":"2011-05-12T12:31:00-07:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2011-05-12:\/unicode\/posts\/2011-05-12-sharada-kashmiri-signs.html","summary":"<p>Innovations for Kashmiri vowels in&nbsp;Sharada<\/p>","content":"<p>The Sharada script was used primarily for writing Sanskrit. In order\nto represent Kashmiri, three additional signs were added to the&nbsp;script.<\/p>\n<p>A below-base slash that resembles the Devanagari <em>vir\u0101ma<\/em>:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/sharada_kashmiri_slash.jpg\"><\/p>\n<p>An above-base bar that represents the Devanagari <em>anudatta<\/em>:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/sharada_kashmiri_bar.jpg\"><\/p>\n<p>An underdot that is similar to the Devanagari <em>nukta<\/em>:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/sharada_kashmiri_underdot.jpg\"><\/p>\n<p>I am currently researching these signs. Please contact me if you have\ninformation about the use of these signs for writing&nbsp;Kashmiri.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"sharada"}}]},{"title":"Change of block name: \u2018Sindhi\u2019 to\u00a0\u2018Khudawadi\u2019","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2011-01-30-sindhi-name-change.html","rel":"alternate"}},"published":"2011-01-30T14:48:00-08:00","updated":"2011-01-30T14:48:00-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2011-01-30:\/unicode\/posts\/2011-01-30-sindhi-name-change.html","summary":"<p>Details on change of proposed block name to&nbsp;Khudawadi<\/p>","content":"<p>My work on the Landa-based scripts of Sindhi continues. However, there\nhas been a <a href=\"http:\/\/std.dkuug.dk\/jtc1\/sc2\/wg2\/docs\/n3957.pdf\">change<\/a>\nto the name of the &#8216;Sindhi&#8217; script that I proposed, and which I\nbriefly discussed in <a\nhref=\"http:\/\/anshumanpandey.blogspot.com\/2010\/08\/standard-sindhi-script.html\">a\npost in August 2010<\/a>. The script is now called &#8220;Khudawadi&#8221;. An\nimage of it is given&nbsp;below:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/grierson_khudawadi_excerpt.png\"><\/p>\n<p>The Khudawadi script was initially proposed under named &#8216;Sindhi&#8217;\nbecause its character repertoire is based upon the &#8216;Standard Sindhi&#8217;\nscript. Standard Sindhi is itself based upon the Khudawadi script,\nwhich was the most well-known and complete of the Landa-based scripts\nused in Sindh. The intent of the <a href=\"http:\/\/std.dkuug.dk\/JTC1\/SC2\/WG2\/docs\/n3871.pdf\">original\nproposal<\/a> was to\ndevelop a standard that might also be used for representing the minor\nLanda scripts of Sindh, some of which are unsuitable for independent\nencoding. As &#8216;Standard Sindhi&#8217; was developed from Khudawadi in the\n1860s with the same idea, the rationale is valid, but the choice of\nthe name &#8216;Sindhi&#8217; is&nbsp;not.<\/p>\n<p>The generic name &#8216;Sindhi&#8217; is problematic for several reasons. It is\ngenerally used for referring to the class of Landa-based scripts of\nSindh, ie. the &#8216;Sindhi scripts&#8217;. It is not the proper name of any\nwriting system that belongs to this script family. In fact, several\nother scripts, such as Khojki and Shikarpuri, are also known as\n&#8216;Sindhi&#8217;. Even &#8216;Standard Sindhi&#8217; was not known by the generic name\n&#8216;Sindhi&#8217;, but as &#8216;Hindi Sindhi&#8217;, etc. In modern India and Pakistan,\nthe &#8216;Sindhi script&#8217; is commonly understood to be the Arabic-based\nscript used for writing the Sindhi&nbsp;language.<\/p>\n<p>The name &#8216;Khudawadi&#8217; is the most appropriate name for the script. It\nis the proper name of the script and is well attested in primary and\nsecondary literature, such as the <em>Linguistic Survey of India<\/em>. The\nname is also used for &#8216;Standard Sindhi&#8217;, which is a reformed variant\nof Khudawadi and is referred to as such, ie. &#8216;improved Khudawadi&#8217;, in\nvarious&nbsp;sources.<\/p>\n<p>The change of name from &#8216;Sindhi&#8217; to &#8216;Khudawadi&#8217; will provide greater\nsemantic and taxonomic clarity in identifying the various scripts used\nfor writing Sindhi and the scripts used in Sindh. For example, it is\nmore appropriate to refer to &#8220;the Sindhi scripts &#8216;Khudawadi&#8217; and\n&#8216;Khojki&#8217;&#8221;, rather than to &#8220;the Sindhi scripts &#8216;Sindhi&#8217; and &#8216;Khojki&#8217;&#8221;.\nMoreover, the use of the name &#8216;Khudawadi&#8217; will enable users to\ndistinguish between the Landa-based and Arabic-based&nbsp;scripts.<\/p>\n<p>More details are available in the\n<a href=\"http:\/\/std.dkuug.dk\/jtc1\/sc2\/wg2\/docs\/n3957.pdf\">document<\/a> requesting\nthe formal change of name. The final proposal for Khudawadi is almost\nready for submission to the Unicode Technical&nbsp;Committee.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"khudawadi"}}]},{"title":"Alternate letter for \u1e0da in\u00a0Devanagari","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2010-12-10-devanagari-dda.html","rel":"alternate"}},"published":"2010-12-10T16:22:00-08:00","updated":"2010-12-10T16:22:00-08:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2010-12-10:\/unicode\/posts\/2010-12-10-devanagari-dda.html","summary":"<p>Another letter for writing a retroflex stop in&nbsp;Devanagari<\/p>","content":"<p>In Devanagari, the phoneme [\u0256] is generally represented with \u0921 U+0921\n<span class=\"caps\">DEVANAGARI<\/span> <span class=\"caps\">LETTER<\/span> <span class=\"caps\">DDA<\/span>. In historical Devanagari orthography for the\nMarwari language of Rajasthan, [\u0256] is represented with a different\nletter. In Marwari, \u0921 is instead used for representing [\u027d], an\nallographic variant of [\u0256], which is generally written in Hindi as \u0921\u093c\nU+095C <span class=\"caps\">DEVANAGARI<\/span> <span class=\"caps\">LETTER<\/span> <span class=\"caps\">DDDHA<\/span>.<\/p>\n<p>In the <em>Linguistic Survey of India<\/em>, George Grierson describes the\nusage of these characters in Marwari, as shown in the image&nbsp;below:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/grierson_marwari_dda.jpg\"><\/p>\n<p>The character is also shown by Mathias Metzger (<em>Die Sprache der\nVak\u012bl-Briefe aus R\u0101jasth\u0101n<\/em>, W\u00fcrzburg: Ergon Verlag,&nbsp;2003):<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/metzger_retroflexes.jpg\"><\/p>\n<p>This Marwari <span class=\"caps\">DDA<\/span> cannot be treated as a glyphic variant of \u0921 and\nshould be encoded as an independent letter. I have submitted a\n<a href=\"http:\/\/www.dkuug.dk\/JTC1\/SC2\/WG2\/docs\/n3970.pdf\">proposal<\/a>, which\ncontains additional information about the letter and examples of&nbsp;use.<\/p>\n<p>The forms shown by Grierson and Metzger differ slightly. My attempt at\nrepresenting the two forms, in adherence with Monotype Devanagari, is\nshown&nbsp;below:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/marwari_dda_glyphs.jpg\"><\/p>\n<p>I recently learned that this letter was also used in Devanagari orthography \nfor languages of Sindh. The following specimen shows use of \nthe letter in&nbsp;Kutchi:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/kutchi_dda_color.jpg\"><\/p>\n<p>Its usage contrasts with the regular letter <span class=\"caps\">DDA<\/span>:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/kutchi_dda_gray.jpg\"><\/p>\n<p>I would be interested to learn if <span class=\"caps\">MARWARI<\/span> <span class=\"caps\">LETTER<\/span> <span class=\"caps\">DDA<\/span> is used in the \northographies for other&nbsp;languages.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"devanagari"}}]},{"title":"An inscription in the Khudawadi\u00a0script","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2010-09-01-khudawadi-inscription.html","rel":"alternate"}},"published":"2010-09-01T13:13:00-07:00","updated":"2010-09-01T13:13:00-07:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2010-09-01:\/unicode\/posts\/2010-09-01-khudawadi-inscription.html","summary":"<p>Details on an inscription bearing text in&nbsp;Khudawadi<\/p>","content":"<p>The <a href=\"http:\/\/std.dkuug.dk\/JTC1\/SC2\/WG2\/docs\/n3766.pdf\">Landa-based scripts of\nSindh<\/a> were generally\nnot used for much more than routine writing and general commercial\nactivity. Michel Boivin of L&#8217;\u00c9cole des hautes \u00e9tudes en sciences\nsociales (<span class=\"caps\">EHESS<\/span>), Paris, recently sent me photographs that show the\nuse of the Khudawadi script in an&nbsp;inscription.<\/p>\n<p>Below is a photograph of a Khudawadi inscription in an <em>ex-voto<\/em> at\nthe shrine of <a href=\"http:\/\/maps.google.com\/maps?f=q&amp;source=s_q&amp;hl=en&amp;geocode=&amp;q=odero+lal+sindh+pakistan&amp;sll=25.698903,68.561978&amp;sspn=0.020302,0.038581&amp;ie=UTF8&amp;hq=&amp;hnear=Odero+Lal+Village,+Matiari,+Sindh,+Pakistan&amp;t=h&amp;ll=25.700938,68.565674&amp;spn=5.195627,9.876709&amp;z=7\">Udero Lal in Hyderabad,\nSindh<\/a>:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/khudawadi_exvotos_udero_lal.jpg\"><\/p>\n<p>The image below shows the detail of the Khudawadi&nbsp;inscription:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/khudawadi_exvotos_udero_lal_detail.jpg\"><\/p>\n<p>In addition to Khudawadi, this <em>ex-voto<\/em> has inscriptions in four\nother scripts: Latin, Devanagari, Arabic, and&nbsp;Gurmukhi.<\/p>\n<p>A proposal to encode Khudawadi in the Unicode standard is being&nbsp;prepared.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"khudawadi"}}]},{"title":"Forms of candrabindu in\u00a0Sharada","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2010-08-31-sharada-candrabindu.html","rel":"alternate"}},"published":"2010-08-31T03:03:00-07:00","updated":"2010-08-31T03:03:00-07:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2010-08-31:\/unicode\/posts\/2010-08-31-sharada-candrabindu.html","summary":"<p>Varying representations of candrabindu in&nbsp;Sharada<\/p>","content":"<p>The regular form ? of the <span class=\"caps\">CANDRABINDU<\/span> in Sharada resembles \nan inverted form of Devanagari ?. Several Sharada manuscripts \nshow an inverted form of <span class=\"caps\">CANDRABINDU<\/span>, which co-occurs with \nthe regular form, sometimes on the same line.\nIn the following image, the inverted <span class=\"caps\">CANDRABINDU<\/span> sign is used for\nwriting <span class=\"caps\">OM<\/span>, while the regular <span class=\"caps\">CANDRABINDU<\/span> is used in other&nbsp;contexts:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/sharada_candrabindus_1.jpg\"><\/p>\n<p>The above specimen might lead one to think that <span class=\"caps\">OM<\/span> is written with an\ninverted <span class=\"caps\">CANDRABINDU<\/span>. Is it perhaps because it is considered a\n&#8216;special&#8217;&nbsp;symbol?<\/p>\n<p>In the following image both regular and inverted <span class=\"caps\">CANDRABINDU<\/span> are used\nfor writing <span class=\"caps\">OM<\/span> (sorry, the beginning of the second line is smudged in\nthe&nbsp;original):<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/sharada_candrabindus_2.jpg\"><\/p>\n<p>Is there a semantic distinction between the regular and inverted forms\nof <span class=\"caps\">CANDRABINDU<\/span> in Sharada? It does not seem so, especially when\nconsidering the use of both forms in writing <span class=\"caps\">OM<\/span>, which has a fairly\nfixed&nbsp;meaning&#8230;<\/p>\n<p>Do we just chalk it up to scribal&nbsp;idiosyncrasy? <\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"sharada"}}]},{"title":"Modi accounting signs and number\u00a0forms","link":{"@attributes":{"href":"https:\/\/pandey.github.io\/unicode\/posts\/2010-08-28-modi-accounting-signs.html","rel":"alternate"}},"published":"2010-08-28T15:40:00-07:00","updated":"2010-08-28T15:40:00-07:00","author":{"name":"Anshuman Pandey"},"id":"tag:pandey.github.io,2010-08-28:\/unicode\/posts\/2010-08-28-modi-accounting-signs.html","summary":"<p>Signs used for numerical notation in&nbsp;Modi<\/p>","content":"<p>Before the standardization of currency and monetary units \nthat followed the formation of the Republic of India \nand other nation states in South Asia in the middle of \nthe 20th century, various forms of numerical notation \nwere used in the&nbsp;region.<\/p>\n<p>Modi documents contain unique signs that are used for marking monetary\namounts. Two such signs are shown in the image&nbsp;below:<\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/modi_number_forms.jpg\"><\/p>\n<p>The character \ua837 is called \u0906\u0933\u0940  <em>\u0101\u1e37\u012b<\/em> in Marathi. It represents\nthe absence of numbers and is written after whole amounts to indicate\nthat there are no fractions. In the image, the first highlighted\nexample is \u096e\u0966\u0966\ua837 &#8216;800\/-&#8216; and the third is \u0967\u096a\u0966\ua837 &#8216;140\/-&#8216;. The \ua837 sign is\nencoded in Unicode as U+A837 <span class=\"caps\">NORTH<\/span> <span class=\"caps\">INDIC<\/span> <span class=\"caps\">PLACEHOLDER<\/span> <span class=\"caps\">MARK<\/span>. (More\ninformation on this and other signs is provided in my <a href=\"http:\/\/std.dkuug.dk\/jtc1\/sc2\/wg2\/docs\/n3367.pdf\">proposal to\nencode Common Indic number&nbsp;forms<\/a><\/p>\n<p>The other sign shown resembles the \u00f7 &#8216;division sign&#8217; or <em>obelus<\/em>. It\nis used in the highlighted examples two and four in the above image.\nBased upon the context in which it is used, it could represent a\n&#8216;remainder&#8217; sign, eg. the amount written and then some; or a \n<em>p\u0101val\u012b<\/em> `quarter&#8217;. <!-- .|. --><\/p>\n<p><img alt=\"image\" src=\"https:\/\/pandey.github.io\/unicode\/images\/modi_pavali.jpg\"><\/p>\n<p>If there is anyone who is familiar with the use of these signs in\nModi, I would be happy to hear from&nbsp;you.<\/p>","category":[{"@attributes":{"term":"unicode"}},{"@attributes":{"term":"modi"}}]}]}