From mboxrd@z Thu Jan 1 00:00:00 1970 Path: news.gmane.io!.POSTED.blaine.gmane.org!not-for-mail From: =?UTF-8?Q?=E0=A4=B8=E0=A4=AE=E0=A5=80=E0=A4=B0_?= =?UTF-8?Q?=E0=A4=B8=E0=A4=BF=E0=A4=82=E0=A4=B9?= Sameer Singh Newsgroups: gmane.emacs.bugs Subject: bug#55303: 29.0.50; Bengali characters \u09F0 and \u09FE are not properly displayed in Emacs Date: Sat, 7 May 2022 21:50:13 +0530 Message-ID: Mime-Version: 1.0 Content-Type: multipart/alternative; boundary="000000000000bd090405de6e5a6e" Injection-Info: ciao.gmane.io; posting-host="blaine.gmane.org:116.202.254.214"; logging-data="38406"; mail-complaints-to="usenet@ciao.gmane.io" To: 55303@debbugs.gnu.org Original-X-From: bug-gnu-emacs-bounces+geb-bug-gnu-emacs=m.gmane-mx.org@gnu.org Sat May 07 18:21:13 2022 Return-path: Envelope-to: geb-bug-gnu-emacs@m.gmane-mx.org Original-Received: from lists.gnu.org ([209.51.188.17]) by ciao.gmane.io with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.92) (envelope-from ) id 1nnNBA-0009pl-M1 for geb-bug-gnu-emacs@m.gmane-mx.org; Sat, 07 May 2022 18:21:12 +0200 Original-Received: from localhost ([::1]:48952 helo=lists1p.gnu.org) by lists.gnu.org with esmtp (Exim 4.90_1) (envelope-from ) id 1nnNB9-0001Mb-HY for geb-bug-gnu-emacs@m.gmane-mx.org; Sat, 07 May 2022 12:21:11 -0400 Original-Received: from eggs.gnu.org ([2001:470:142:3::10]:38644) by lists.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.90_1) (envelope-from ) id 1nnNB0-0001M8-E5 for bug-gnu-emacs@gnu.org; Sat, 07 May 2022 12:21:02 -0400 Original-Received: from debbugs.gnu.org ([209.51.188.43]:58932) by eggs.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_128_GCM_SHA256:128) (Exim 4.90_1) (envelope-from ) id 1nnNB0-0004Lg-5U for bug-gnu-emacs@gnu.org; Sat, 07 May 2022 12:21:02 -0400 Original-Received: from Debian-debbugs by debbugs.gnu.org with local (Exim 4.84_2) (envelope-from ) id 1nnNAz-0006cj-PW for bug-gnu-emacs@gnu.org; Sat, 07 May 2022 12:21:01 -0400 X-Loop: help-debbugs@gnu.org Resent-From: =?UTF-8?Q?=E0=A4=B8=E0=A4=AE=E0=A5=80=E0=A4=B0_?= =?UTF-8?Q?=E0=A4=B8=E0=A4=BF=E0=A4=82=E0=A4=B9?= Sameer Singh Original-Sender: "Debbugs-submit" Resent-CC: bug-gnu-emacs@gnu.org Resent-Date: Sat, 07 May 2022 16:21:01 +0000 Resent-Message-ID: Resent-Sender: help-debbugs@gnu.org X-GNU-PR-Message: report 55303 X-GNU-PR-Package: emacs X-Debbugs-Original-To: bug-gnu-emacs@gnu.org Original-Received: via spool by submit@debbugs.gnu.org id=B.165194043825391 (code B ref -1); Sat, 07 May 2022 16:21:01 +0000 Original-Received: (at submit) by debbugs.gnu.org; 7 May 2022 16:20:38 +0000 Original-Received: from localhost ([127.0.0.1]:52823 helo=debbugs.gnu.org) by debbugs.gnu.org with esmtp (Exim 4.84_2) (envelope-from ) id 1nnNAc-0006bT-AE for submit@debbugs.gnu.org; Sat, 07 May 2022 12:20:38 -0400 Original-Received: from lists.gnu.org ([209.51.188.17]:58012) by debbugs.gnu.org with esmtp (Exim 4.84_2) (envelope-from ) id 1nnNAa-0006bM-Je for submit@debbugs.gnu.org; Sat, 07 May 2022 12:20:36 -0400 Original-Received: from eggs.gnu.org ([2001:470:142:3::10]:38634) by lists.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.90_1) (envelope-from ) id 1nnNAa-00019A-F5 for bug-gnu-emacs@gnu.org; Sat, 07 May 2022 12:20:36 -0400 Original-Received: from mail-qt1-x836.google.com ([2607:f8b0:4864:20::836]:41802) by eggs.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_128_GCM_SHA256:128) (Exim 4.90_1) (envelope-from ) id 1nnNAY-0004L6-V3 for bug-gnu-emacs@gnu.org; Sat, 07 May 2022 12:20:36 -0400 Original-Received: by mail-qt1-x836.google.com with SMTP id y3so8152592qtn.8 for ; Sat, 07 May 2022 09:20:34 -0700 (PDT) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=gmail.com; s=20210112; h=mime-version:from:date:message-id:subject:to; bh=/xEsJYI6CfI4I02A2hwjZ+6dZCayt3QJr/rsntYLUPE=; b=JTmKztRHyEz1OwmOrGvHAVh7XYSxoLbiW2L0QhNKAEUxtqOgP4omdcx9CQfDAJ5jNP /s317MpFur98XMOYsV3rN1+YgPAtavgHI7GdkzaEY4KbZ9yvhMtJtJRYd5CuWnVprdZh 7Ikc1w9vw/PJ10ZEXW5XrRFIhLi6pgQQ+blTJZ1zlcK0V5GVFv6M2lqbOG3hxAoGxbAa sha41Onq36bBHc+zCPiThj0HSbwc6eK1jjUtCwT54lz24lvpjuyrRHu7subKdWVIbF4R Lt6K+JGKtqVpNnRxd1IAyG8M/H3L2V9V1oeFSBIonqxXuAlhPI0tlFndX38XnaDKQGTy /DAg== X-Google-DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20210112; h=x-gm-message-state:mime-version:from:date:message-id:subject:to; bh=/xEsJYI6CfI4I02A2hwjZ+6dZCayt3QJr/rsntYLUPE=; b=U2rljvlrRc7T1z2vJ/gXhICh5dhLAjLzcxRqdLHXufn7Keh3gzpI0hy1qDfPaYv4Q8 C9335gqze5+zzPBUYVX/w9O+rVc0T+2Wh/Fhd25ZS8JcykwnEYEgCTGVs1ubi17l/VEk y+sVSX6j2oO7GjTlHeFVVzPnGRy+ZvV+GpO7fCqyYizptl+xoka3QRtQSscD38oDtnF5 AuxIndZd8GjR3jUntYGWOh9GktXUttZiHQoI/t0sb6yP5Zk3+Yca1Yu+QAiyRJJoYA3g 6B/dU75mLwi+mj56DNNdD7aX93kKXNtf0mWPNY5dzqb1uD02BhzRTSvzS8lTFdu6C5/N HYIg== X-Gm-Message-State: AOAM5305XHx19VuSATGW7nwd8Ug343LWh4MWXxwRu6N9gYpz8MynQvQD D1ylVY1d8nvMVAA/dH0qE3wlJyV2yiCZs1utmEE7Kz64+E3WUA== X-Google-Smtp-Source: ABdhPJwaF1cLGg6ZyZFZZWDHJ1bBnMyPtb5ty/KJUAr9UAITzTFuU58nuqzRblIpqhsLjKc0aD1KJM/Q/pHmwvxs6xE= X-Received: by 2002:a05:622a:250:b0:2f3:cfd5:45aa with SMTP id c16-20020a05622a025000b002f3cfd545aamr2611431qtx.676.1651940433659; Sat, 07 May 2022 09:20:33 -0700 (PDT) Received-SPF: pass client-ip=2607:f8b0:4864:20::836; envelope-from=lumarzeli30@gmail.com; helo=mail-qt1-x836.google.com X-Spam_score_int: -18 X-Spam_score: -1.9 X-Spam_bar: - X-Spam_report: (-1.9 / 5.0 requ) BAYES_00=-1.9, DKIM_SIGNED=0.1, DKIM_VALID=-0.1, DKIM_VALID_AU=-0.1, DKIM_VALID_EF=-0.1, FREEMAIL_ENVFROM_END_DIGIT=0.25, FREEMAIL_FROM=0.001, HTML_MESSAGE=0.001, RCVD_IN_DNSWL_NONE=-0.0001, SPF_HELO_NONE=0.001, SPF_PASS=-0.001, T_SCC_BODY_TEXT_LINE=-0.01 autolearn=ham autolearn_force=no X-Spam_action: no action X-BeenThere: debbugs-submit@debbugs.gnu.org X-Mailman-Version: 2.1.18 Precedence: list X-BeenThere: bug-gnu-emacs@gnu.org List-Id: "Bug reports for GNU Emacs, the Swiss army knife of text editors" List-Unsubscribe: , List-Archive: List-Post: List-Help: List-Subscribe: , Errors-To: bug-gnu-emacs-bounces+geb-bug-gnu-emacs=m.gmane-mx.org@gnu.org Original-Sender: "bug-gnu-emacs" Xref: news.gmane.io gmane.emacs.bugs:231601 Archived-At: --000000000000bd090405de6e5a6e Content-Type: text/plain; charset="UTF-8" Content-Transfer-Encoding: quoted-printable There is no composition rule for \u09F0 (=E0=A7=B0) and \u09FE (=E0=A7=BE) = in emacs, therefore they are not rendered properly. Steps to reproduce: 1. emacs -Q 2. Type: =E0=A7=B0=E0=A6=BE=E0=A6=AE =E0=A6=95=E0=A7=B0=E0=A7=8D=E0=A6=AE =E0= =A7=B0=E0=A7=82=E0=A6=AA =E0=A6=B2=E0=A6=BE=E0=A7=BE=E0=A6=A8=E0=A7=81 =E0=A6=A8=E0=A6=BE=E0=A7= =BE The following patch will fix the problem: diff --git a/lisp/language/indian.el b/lisp/language/indian.el index b240403b0a..b6fcdbb348 100644 --- a/lisp/language/indian.el +++ b/lisp/language/indian.el @@ -194,13 +194,14 @@ bengali-composable-pattern '(("a" . "\u0981") ; SIGN CANDRABINDU ("A" . "[\u0982\u0983]") ; SIGN ANUSVARA .. VISARGA ("V" . "[\u0985-\u0994\u09E0\u09E1]") ; independent vowel - ("C" . "[\u0995-\u09B9\u09DC-\u09DF\u09F1]") ; consonant + ("C" . "[\u0995-\u09B9\u09DC-\u09DF\u09F0\u09F1]") ; consonant ("B" . "[\u09AC\u09AF\u09B0\u09F0]") ; BA, YA, RA ("R" . "[\u09B0\u09F0]") ; RA ("n" . "\u09BC") ; NUKTA ("v" . "[\u09BE-\u09CC\u09D7\u09E2\u09E3]") ; vowel sign ("H" . "\u09CD") ; HALANT ("T" . "\u09CE") ; KHANDA TA + ("S" . "\u09FE") ; SANDHI MARK ("N" . "\u200C") ; ZWNJ ("J" . "\u200D") ; ZWJ ("X" . "[\u0980-\u09FF]")))) ; all coverage @@ -209,7 +210,7 @@ bengali-composable-pattern ;; syllables with an independent vowel, or "\\(?:RH\\)?Vn?\\(?:J?HB\\)?v*n?a?A?\\|" ;; consonant-based syllables, or - "Cn?\\(?:J?HJ?Cn?\\)*\\(?:H[NJ]?\\|v*[NJ]?v?a?A?\\)\\|" + "Cn?\\(?:J?HJ?Cn?\\)*\\(?:H[NJ]?\\|v*[NJ]?v?a?A?S?\\)\\|" ;; another syllables with an independent vowel, or "\\(?:RH\\)?T\\|" ;; special consonant form, or --000000000000bd090405de6e5a6e Content-Type: text/html; charset="UTF-8" Content-Transfer-Encoding: quoted-printable
There is no composition rule for \u09F0 (=E0=A7=B0) a= nd \u09FE (=E0=A7=BE) in emacs, therefore they are not rendered properly.
Steps to reproduce:
1. emacs -Q
2. Type:
=
=C2=A0=C2=A0=C2=A0=C2=A0 =E0=A7=B0=E0=A6=BE=E0=A6=AE =E0=A6=95=E0=A7= =B0=E0=A7=8D=E0=A6=AE =E0=A7=B0=E0=A7=82=E0=A6=AA
=C2=A0=C2=A0=C2= =A0=C2=A0 =E0=A6=B2=E0=A6=BE=E0=A7=BE=E0=A6=A8=E0=A7=81 =E0=A6=A8=E0=A6=BE= =E0=A7=BE

The following patch will fix the problem= :

diff --git a/lisp/language/indian.el b/lisp/lang= uage/indian.el
index b240403b0a..b6fcdbb348 100644
--- a/lisp/langu= age/indian.el
+++ b/lisp/language/indian.el
@@ -194,13 +194,14 @@ ben= gali-composable-pattern
=C2=A0 '(("a" . "\u0981"= ;) ; SIGN CANDRABINDU
=C2=A0 =C2=A0 ("A" . "[\u0982\u09= 83]") ; SIGN ANUSVARA .. VISARGA
=C2=A0 =C2=A0 ("V" . &q= uot;[\u0985-\u0994\u09E0\u09E1]") ; independent vowel
- =C2=A0 (&q= uot;C" . "[\u0995-\u09B9\u09DC-\u09DF\u09F1]") ; consonant+ =C2=A0 ("C" . "[\u0995-\u09B9\u09DC-\u09DF\u09F0\u09F1]= ") ; consonant
=C2=A0 =C2=A0 ("B" . "[\u09AC\u09AF\= u09B0\u09F0]") ; BA, YA, RA
=C2=A0 =C2=A0 ("R" . "= [\u09B0\u09F0]") ; RA
=C2=A0 =C2=A0 ("n" . "\u09BC= ") ; NUKTA
=C2=A0 =C2=A0 ("v" . "[\u09BE-\u09CC\u0= 9D7\u09E2\u09E3]") ; vowel sign
=C2=A0 =C2=A0 ("H" . &qu= ot;\u09CD") ; HALANT
=C2=A0 =C2=A0 ("T" . "\u09CE&= quot;) ; KHANDA TA
+ =C2=A0 =C2=A0 =C2=A0 =C2=A0 =C2=A0 ("S" = . "\u09FE") ; SANDHI MARK
=C2=A0 =C2=A0 ("N" . &qu= ot;\u200C") ; ZWNJ
=C2=A0 =C2=A0 ("J" . "\u200D&qu= ot;) ; ZWJ
=C2=A0 =C2=A0 ("X" . "[\u0980-\u09FF]")= ))) ; all coverage
@@ -209,7 +210,7 @@ bengali-composable-pattern
=C2= =A0 =C2=A0 =C2=A0 =C2=A0;; syllables with an independent vowel, or
=C2= =A0 =C2=A0 =C2=A0 =C2=A0"\\(?:RH\\)?Vn?\\(?:J?HB\\)?v*n?a?A?\\|"<= br>=C2=A0 =C2=A0 =C2=A0 =C2=A0;; consonant-based syllables, or
- =C2=A0 = =C2=A0 =C2=A0"Cn?\\(?:J?HJ?Cn?\\)*\\(?:H[NJ]?\\|v*[NJ]?v?a?A?\\)\\|&qu= ot;
+ =C2=A0 =C2=A0 =C2=A0"Cn?\\(?:J?HJ?Cn?\\)*\\(?:H[NJ]?\\|v*[NJ]= ?v?a?A?S?\\)\\|"
=C2=A0 =C2=A0 =C2=A0 =C2=A0;; another syllables wi= th an independent vowel, or
=C2=A0 =C2=A0 =C2=A0 =C2=A0"\\(?:RH\\)?= T\\|"
=C2=A0 =C2=A0 =C2=A0 =C2=A0;; special consonant form, or
<= /div> --000000000000bd090405de6e5a6e--