Virama
Virama (Sanskrit: विराम/हलन्त, romanized: virāma/halanta ्, IPA:[ʋiraːmɐ,ɦɐlɐn̪t̪ɐ] (see #Names for other terms) is a Sanskrit phonological concept to suppress the inherent vowel that otherwise occurs with every consonant letter, commonly used as a generic term for a codepoint in Unicode, representing either
- halanta, hasanta or explicit virāma, a diacritic in many Brahmic scripts, including the Devanagari and Bengali scripts, or
- saṃyuktākṣara (Sanskrit: संयुक्ताक्षर) or implicit virama, a conjunct consonant or ligature.
Unicode schemes of scripts writing Mainland Southeast Asia languages, such as that of Burmese script and of Tibetan script, generally do not group the two functions together.
Names
The name is Sanskrit for "cessation, termination, end". As a Sanskrit word, it is used in place of several language-specific terms, such as:
| Name in English books | Language | In native language | Form | Notes |
|---|---|---|---|---|
| halant | Hindi | हलन्त, halant | ् | |
| halanta | Punjabi | ਹਲੰਤ, halanta | ੍ | |
| Marathi | हलंत, halanta | ् | ||
| Nepali | हलन्त, halanta | ् | ||
| Odia | ହଳନ୍ତ, hôḷôntô | ୍ | ||
| Gujarati | હાલાંત, hālānta | ્ | ||
| hosonto | Bengali | হসন্ত, hôsôntô | ্ | |
| Assamese | হসন্ত, hoxonto / হছন্ত, hosonto | ্ | ||
| Sylheti | ꠢꠡꠘ꠆ꠔꠧ, hośonto | ◌ ꠆ | ||
| pollu | Telugu | పొల్లు, pollu | ్ | |
| pulli | Tamil | புள்ளி, puḷḷi | ் | |
| chandrakkala | Malayalam | ചന്ദ്രക്കല, chandrakkala / വിരാമം, virāmaṁ | ് | Unlike other virama diacritics, it is pronounced [ə̆] word-finally. |
| ardhakshara chihne | Kannada | ಅರ್ಧಾಕ್ಷರ ಚಿಹ್ನೆ, ardhākṣara cihne / ಸುರುಳಿ, suruḷi | ್ | |
| hal kirima | Sinhalese | හල් කිරිම, hal kirīma | ් | |
| a that | Burmese | အသတ်, a.sat, IPA:[ʔa̰θaʔ] | ် | lit. "nonexistence" |
| viream | Khmer | វិរាម, viream | ៑ | |
| toandeakheat | ទណ្ឌឃាដ, toandeakheat | ៍ | ||
| karan, thanthakhat | Thai | การันต์, kārạnt[1][2] / ทัณฑฆาต, thanthakhat[3][4] | ◌์ | Thanthakhat is the name of the diacritic, while karan refers to the character that was marked. These two terms are often used interchangeably. It is used to mark as silent vowels or consonants that were originally pronounced, but have become silenced in Thai pronunciation (mostly from Sanskrit and Old Khmer). This diacritic is sometimes used in loanwords from European languages to mark final consonants in consonant clusters (e.g. want as วอนท์). |
| pinthu | พินทุ, pinthu | ◌ฺ | Pinthu is akin to Sanskrit bindu, and means "point" or "dot". It is used to mark a syllable as closed, and it is only used in Thai script when writing Pali or Sanskrit. | |
| nikkhahit | นฤคหิต / นิคหิต | ◌ํ | Nikkhahit represents what was originally anusvāra in Sanskrit. Like pinthu, it is also only used when writing Pali or Sanskrit in Thai script. It marks a syllable as nasalized, realized in Thai as a nasal closed consonant following the vowel. | |
| rahaam | Northern Thai (Lanna) | ᩁᩉ᩶ᩣ᩠ᨾ, rahaam[5] | ◌᩺ | |
| Tai Khün | ◌᩼ | |||
| Tai Lue | ◌᩼ | |||
| wirama | Kawi | 𑼮𑼶𑼬𑼴𑼪, wirāma | ◌𑽁 | |
| pangkon | Javanese | ꦥꦁꦏꦺꦴꦤ꧀, pangkon | ◌꧀ | Other variations of the pangkon exist if the suppressed syllable is in the middle of a word (a subjoined pasangan form to stack the next consonant instead). |
| adeg-adeg | Balinese | ᬳᬤᭂᬕ᭄ᬳᬤᭂᬕ᭄, adəg-adəg | ◌᭄ | |
| pangolat | Mandailing | ᯇᯝᯬᯞᯖ᯲, pangolat | ◌᯲ | |
| Pakpak | ||||
| Toba | ||||
| penengen | Karo | ᯇᯧᯉᯧᯝᯧᯉ᯳, pənəngən | ◌᯳ | |
| panongonan | Simalungun | ᯈᯉᯬᯝᯬᯉᯉ᯳, panongonan | ||
| pamaeh | Sundanese | ᮕᮙᮆᮂ, pamaeh | ◌᮪ | |
| bunuhan | Rejang | ꤷꥈꤵꥈꥁꥐ, bunuhan | ꥓ | |
| sukun | Dhivehi | ސުކުން, sukun | ް◌ | Derives from Arabic "sukun" |
| Srog med | Tibetan | Srog med | ྄ | Only used when transcribing Sanskrit |
Usage
In Devanagari and many other Indic scripts, a virama is used to cancel the inherent vowel of a consonant letter and represent a consonant without a vowel, a "dead" consonant. For example, in Devanagari,
- क is a consonant letter, ka,
- ् is a virāma; therefore,
- क् (ka + virāma) represents a dead consonant k.
إذا تبع الحرف k حرف ساكن آخر، مثلاً ṣa ष، فقد تبدو النتيجة क्ष ، والتي تُمثل kṣa كـ ka + (المرئي) virāma + ṣa . في هذه الحالة، يُوضع العنصران k क् و ṣa ष ببساطة جنبًا إلى جنب. بدلاً من ذلك، يمكن كتابة kṣa أيضًا كحرف مُركب क्ष ، وهو الشكل المُفضل. عمومًا، عند دمج حرف ساكن ميت C1 مع حرف ساكن آخر C2 ، قد تكون النتيجة:
- ربط كامل لـ C 1 + C 2 ؛
- نصف ملتصقة—
- C 1 - الربط: شكل معدل (نصف شكل) من C 1 متصل بالشكل الأصلي (الشكل الكامل) من C 2
- C 2 -conjoining: شكل معدل من C 2 متصل بالشكل الكامل لـ C 1 ؛ أو
- غير مرتبط: الأشكال الكاملة لـ C1 و C2 مع فيراما مرئية. [ 6 ]
إذا كانت النتيجة متصلة كليًا أو جزئيًا، فإنّ الفيراما (المفهومية) التي جعلت C 1 غير قابلة للقراءة تصبح غير مرئية، ولا توجد منطقيًا إلا في نظام ترميز الأحرف مثل ISCII أو Unicode . أما إذا لم تكن النتيجة متصلة، فإنّ الفيراما تكون مرئية، ومتصلة بـ C 1 ، ومكتوبة فعليًا.
باختصار، هذه الاختلافات مجرد تنويعات في شكل الحروف، والأشكال الثلاثة متطابقة دلاليًا . مع أن لكل لغة شكلًا مفضلًا لمجموعة حروف ساكنة معينة، وبعض الخطوط لا تحتوي على أي نوع من الوصلات أو الأشكال النصفية، فإنه من المقبول عمومًا استخدام شكل غير متصل بدلًا من شكل متصل، حتى وإن كان الأخير مفضلًا، إذا لم يكن الخط يحتوي على رمز للوصلة. في بعض الحالات الأخرى، يُعد استخدام الوصلة من عدمه مسألة ذوق شخصي.
وبالتالي، قد يعمل الحرف virāma في التسلسل C 1 + virāma + C 2 كحرف تحكم غير مرئي لربط C 1 و C 2 في نظام يونيكود. على سبيل المثال،
- كا क + virāma + ṣa ष = kṣa क्ष
هو ربط متصل تمامًا. من الممكن أيضًا ألا يربط الفيراما بين C1 و C2 ، تاركًا الصيغ الكاملة لـ C1 و C2 كما هي:
- كا क + فيراما + ṣa ष = kṣa क्ष
يُعد مثالاً على هذا الشكل غير المرتبط.
The sequences ङ्क ङ्ख ङ्ग ङ्घ [ṅkaṅkhaṅɡaṅɡha], in common Sanskrit orthography, should be written as conjuncts (the virāma and the top cross line of the second letter disappear, and what is left of the second letter is written under the ङ and joined to it).
End of word
The inherent vowel is not always pronounced, in particular at the end of a word (schwa deletion). No virāma is used for vowel suppression in such cases. Instead, the orthography is based on Sanskrit where all inherent vowels are pronounced, and leaves to the reader of modern languages to delete the schwa when appropriate.[7]
Glyph comparison
| Comparison of virama in different scripts |
|---|
ملحوظات
|
يونيكود
- U+094D ् DEVANAGARI SIGN VIRAMA
- U+09CD ্ BENGALI SIGN VIRAMA
- U+0A4D ੍ GURMUKHI SIGN VIRAMA
- U+0ACD ્ علامة جوجاراتية VIRAMA
- U+0B4D ୍ ORIYA SIGN VIRAMA
- U+0BCD ் علامة تاميلية فيراما
- U+0C4D ్ لافتة تيلوجو فيراما
- U+0CCD ್ KANNADA SIGN VIRAMA
- U+0D3B ഻ MALAYALAM SIGN VERTICAL BAR VIRAMA
- U+0D3C ഼ MALAYALAM SIGN CIRCULAR VIRAMA
- U+0D4D ് MALAYALAM SIGN VIRAMA
- U+0DCA ් علامة السنهالية اللاكونا
- U+0E3AฺTHAI CHARACTER PHINTHU
- U+0E4C์THAI CHARACTER THANTHAKHAT
- U+0E4E๎THAI CHARACTER YAMAKKAN
- U+0EBA຺LAO SIGN PALI VIRAMA
- U+0ECC໌LAO CANCELLATION MARK
- U+0ECE໎LAO YAMAKKAN
- U+0F84྄TIBETAN MARK HALANTA
- U+1039္MYANMAR SIGN VIRAMA
- U+103A်MYANMAR SIGN ASAT
- U+1714᜔TAGALOG SIGN VIRAMA
- U+17CD៍KHMER SIGN TOANDAKHIAT
- U+17D1៑KHMER SIGN VIRIAM
- U+1B44᭄BALINESE ADEG ADEG
- U+1BAA᮪SUNDANESE SIGN PAMAAEH
- U+1BAB᮫SUNDANESE SIGN VIRAMA
- U+1BF2᯲BATAK PANGOLAT
- U+1BFF᯿BATAK SYMBOL BINDU PANGOLAT
- U+A8C4꣄SAURASHTRA SIGN VIRAMA
- U+A8F3ꣳDEVANAGARI SIGN CANDRABINDU VIRAMA
- U+A8F4ꣴDEVANAGARI SIGN DOUBLE CANDRABINDU VIRAMA
- U+A953꥓REJANG VIRAMA
- U+A9C0꧀JAVANESE PANGKON
- U+10A3F𐨿KHAROSHTHI VIRAMA
- U+11046𑁆BRAHMI VIRAMA
- U+110B9𑂹KAITHI SIGN VIRAMA
- U+111C0𑇀SHARADA SIGN VIRAMA
- U+11235𑈵KHOJKI SIGN VIRAMA
- U+1134D𑍍GRANTHA SIGN VIRAMA
- U+11C3F𑰿BHAIKSUKI SIGN VIRAMA
- U+11442𑑂NEWA SIGN VIRAMA
- U+114C2𑓂TIRHUTA SIGN VIRAMA
- U+115BF𑖿SIDDHAM SIGN VIRAMA
- U+1163F𑘿MODI SIGN VIRAMA
- U+116B6𑚶TAKRI SIGN VIRAMA
- U+1172B𑜫AHOM SIGN KILLER
- U+11839𑠹DOGRA SIGN VIRAMA
- U+1193E𑤾DIVES AKURU VIRAMA
- U+1193D𑤽DIVES AKURU SIGN HALANTA
- U+119E0𑧠NANDINAGARI SIGN VIRAMA
- U+11A34𑨴ZANABAZAR SQUARE SIGN VIRAMA
- U+11D44𑵄MASARAM GONDI SIGN HALANTA
- U+11D45𑵅MASARAM GONDI VIRAMA
- U+11D97𑶗GUNJALA GONDI VIRAMA
- U+11F41𑽁KAWI SIGN KILLER
- U+11F42𑽂KAWI CONJOINER
See also
- Sukun, a similar diacritic in Arabic script
- Zero consonant
References
- ↑"คำศัพท์ การันต์ แปลว่าอะไร?". Longdo Dict.
- ↑th:การันต์
- ↑"คำศัพท์ ทัณฑฆาต แปลว่าอะไร?". Longdo Dict.
- ↑th:ทัณฑฆาต
- ↑"Tai Tham"(PDF). The Unicode Standard. Retrieved 30 July 2022.
- ↑Constable, Peter (2004). "Clarification of the Use of Zero Width Joiner in Indic Scripts"(PDF). Unicode, Inc. Retrieved 2009-11-19.
- ↑Akira Nakanishi: Writing Systems of the World, ISBN 0-8048-1654-9, pp. 48.
External links
- Brahmic diacritics
