App-Lingua-BO-Wylie-Transliteration

 view release on metacpan or  search on metacpan

README  view on Meta::CPAN

NAME
    App::Lingua::BO::Wylie::Transliteration

VERSION
    version 0.1.0

SYNOPSIS
    Wylie transliterate can be used to transliterate words from Wylie
    Transliteration to (Classical) Tibetan (dbu med)

BACKGROUND
    When you have one (foreign) alphabet and would like to display it in a
    different alphabet (example: Russian to Latin alphabet), you will want
    to use a certain transliteration scheme.

    Just compare all the different names you can find "Dostojevski"
    transliterated to, to see what enourmous differences there will be.

    Now for the Classical Tibetan "dbu med" alphabet there exist two main
    transliteration schemes:

    *   Library of Congress Transliteration

    *   Wylie Transliteration

    Classical Tibetan alphabet itself works in a really interesting way.

    First, let's have a look at the table of the individual "characters"
    with their Wylie transliterations:

    <http://en.wikipedia.org/wiki/Tibetan_alphabet>

    A few key observations:

    *   (Almost) all letters represent a consonant, carrying an inherent
        vocal: "a"

    *   These letters/syllables are sorted according to tonality and
        aspiration (in pronunciation)

    *   Other vocals will be achieved by adding certain vocal-symbols in the
        proper places:

       * i.e. you can build the following syllables by adding a vowel sign:
         * ka -> ko
         * ka -> ku
         * ka -> ki
         * ka -> ke

    The latter is a process of merging symbols to form new symbols (with the
    merging taking place in the proper places -- 'e' is on top, 'u' will be
    inserted at the bottom)

    This merging process can be seen as building "ligatures", of which even
    more exist.

    As we have seen, the vocal symbols (for all vocals except 'a', which is
    inherent) need to be added in the proper places.

    For the rest of the symbols that can be added, the scheme looks the
    following:

             b  s  g   r   u   b  s
             |  |  |   |   |   |  |
             1  2  3   4   5   6  7

    With the places being the following:

    1) Prescript 2) Superscript 3) The Center piece (carriyng the inherent
    vocal, mandatory) 4) Subscript 5) The vocal sign 6) Postscript1 7)
    Postscript2

    Note: Except from the Center piece (3), all other signs are optional.

    Note: Optional character can form ligatures with the character they are
    combined with.

TECHNICAL BACKGROUND (UNICODE)
    The Unicode consortium had to decide what they want their code points to
    look like:

    a) either each altered base syllable is represented (i.e. ka, ko, ke,
    ki, ku) as a separate character (code point) b) the base syllables are
    represented and the altered syllables will be merged

    Since it was chosen for the latter, this has a few consequences:

    *   The graphical representation will depend on building ligatures and
        thus on the font you are using.



( run in 2.349 seconds using v1.01-cache-2.11-cpan-364913b4093 )