From mboxrd@z Thu Jan 1 00:00:00 1970 Path: news.gmane.io!.POSTED.blaine.gmane.org!not-for-mail From: Visuwesh Newsgroups: gmane.emacs.bugs Subject: bug#73846: [PATCH] Make djvused emit UTF-8 encoded text Date: Thu, 17 Oct 2024 09:42:30 +0530 Message-ID: <87y12n1i6p.fsf@gmail.com> Mime-Version: 1.0 Content-Type: multipart/mixed; boundary="=-=-=" Injection-Info: ciao.gmane.io; posting-host="blaine.gmane.org:116.202.254.214"; logging-data="13834"; mail-complaints-to="usenet@ciao.gmane.io" User-Agent: Gnus/5.13 (Gnus v5.13) Cc: "Tassilo Horn" To: 73846@debbugs.gnu.org Original-X-From: bug-gnu-emacs-bounces+geb-bug-gnu-emacs=m.gmane-mx.org@gnu.org Thu Oct 17 06:15:58 2024 Return-path: Envelope-to: geb-bug-gnu-emacs@m.gmane-mx.org Original-Received: from lists.gnu.org ([209.51.188.17]) by ciao.gmane.io with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.92) (envelope-from ) id 1t1HvZ-0003QI-Sf for geb-bug-gnu-emacs@m.gmane-mx.org; Thu, 17 Oct 2024 06:15:58 +0200 Original-Received: from localhost ([::1] helo=lists1p.gnu.org) by lists.gnu.org with esmtp (Exim 4.90_1) (envelope-from ) id 1t1HvM-00056V-F2; Thu, 17 Oct 2024 00:15:44 -0400 Original-Received: from eggs.gnu.org ([2001:470:142:3::10]) by lists.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.90_1) (envelope-from ) id 1t1HvK-00056D-Gl for bug-gnu-emacs@gnu.org; Thu, 17 Oct 2024 00:15:42 -0400 Original-Received: from debbugs.gnu.org ([2001:470:142:5::43]) by eggs.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_128_GCM_SHA256:128) (Exim 4.90_1) (envelope-from ) id 1t1HvK-0002Sc-7c; Thu, 17 Oct 2024 00:15:42 -0400 DKIM-Signature: v=1; a=rsa-sha256; q=dns/txt; c=relaxed/relaxed; d=debbugs.gnu.org; s=debbugs-gnu-org; h=MIME-Version:Date:From:To:Subject; bh=lio7kJiOuA8CmWPW1GWsIYNeZOnOcjvhLIBU1Kjtoyg=; b=RX9CXW9MIkMAuctal4F9EmCSW/Nk8yV2mOcxKrQ85tJYbvnvizlToLdKHSAPLAJkHjlwzEPiVP7w9Ne0sbzftMAxvKuQX9rIw4Cgoo0jzmG4qMAwbfjMx8L14o/wGV9IpDrdVZnbGIXumCLU9tSXNG4Ux77DFV+dZIxEKQqO+2xLagB1lXq6nxteKr5JuBsV8RylqOHM9IAZrh0M8wzAbuQKKFiReG7GxtN3ZEq7c0U0hbqR+pxhapLIc3VOtCDwB1NDM0L9PSHqENPj8uef1Jpi9tkYnJkU4gBL3UPuFAXutoCwCDSSKUXJRTMd0FYrarz+KvxsR7duFX9+L7CPEQ==; Original-Received: from Debian-debbugs by debbugs.gnu.org with local (Exim 4.84_2) (envelope-from ) id 1t1Hve-00086i-Hm; Thu, 17 Oct 2024 00:16:02 -0400 X-Loop: help-debbugs@gnu.org Resent-From: Visuwesh Original-Sender: "Debbugs-submit" Resent-CC: tsdh@gnu.org, bug-gnu-emacs@gnu.org Resent-Date: Thu, 17 Oct 2024 04:16:02 +0000 Resent-Message-ID: Resent-Sender: help-debbugs@gnu.org X-GNU-PR-Message: report 73846 X-GNU-PR-Package: emacs X-GNU-PR-Keywords: patch X-Debbugs-Original-To: bug-gnu-emacs@gnu.org X-Debbugs-Original-Xcc: "Tassilo Horn" Original-Received: via spool by submit@debbugs.gnu.org id=B.172913851230739 (code B ref -1); Thu, 17 Oct 2024 04:16:02 +0000 Original-Received: (at submit) by debbugs.gnu.org; 17 Oct 2024 04:15:12 +0000 Original-Received: from localhost ([127.0.0.1]:32930 helo=debbugs.gnu.org) by debbugs.gnu.org with esmtp (Exim 4.84_2) (envelope-from ) id 1t1Huq-0007zh-7m for submit@debbugs.gnu.org; Thu, 17 Oct 2024 00:15:12 -0400 Original-Received: from lists.gnu.org ([209.51.188.17]:40324) by debbugs.gnu.org with esmtp (Exim 4.84_2) (envelope-from ) id 1t1Hup-0007zZ-17 for submit@debbugs.gnu.org; Thu, 17 Oct 2024 00:15:11 -0400 Original-Received: from eggs.gnu.org ([2001:470:142:3::10]) by lists.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.90_1) (envelope-from ) id 1t1HsN-0004t1-0N for bug-gnu-emacs@gnu.org; Thu, 17 Oct 2024 00:12:39 -0400 Original-Received: from mail-pg1-x543.google.com ([2607:f8b0:4864:20::543]) by eggs.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_128_GCM_SHA256:128) (Exim 4.90_1) (envelope-from ) id 1t1HsL-000259-AY for bug-gnu-emacs@gnu.org; Thu, 17 Oct 2024 00:12:38 -0400 Original-Received: by mail-pg1-x543.google.com with SMTP id 41be03b00d2f7-7ea8de14848so390063a12.2 for ; Wed, 16 Oct 2024 21:12:36 -0700 (PDT) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=gmail.com; s=20230601; t=1729138355; x=1729743155; darn=gnu.org; h=mime-version:user-agent:message-id:date:subject:to:from:from:to:cc :subject:date:message-id:reply-to; bh=lio7kJiOuA8CmWPW1GWsIYNeZOnOcjvhLIBU1Kjtoyg=; b=Lbip054ejuesVc7pslMHgMMmWIdBcUmcmRTtVBtVkIn2w9BlSChJk0zB0R0vEbn6er hI06CoGkDEPdi8lQDfvBLhkEgtxbBWTNpbvoFVY4+zeBBxmi9qinV/SCr1GGWjMeKIf+ AeU6E88veYJmBm66x4TRZXGqr8Ah+ndG9JiY/6e9tIMi33o1nRU4yX/8u+X34kcNonyY TEo+h1+/SEtObtpBPNb0m8ufBajJzHnyjaUkLyBqRwdy/F++2OI6QDvR/I4BH97PJTd6 KKFnV8D1VA7/1wVX4yJr67IxGWRkIxKHr3HhG7EpqaKBtclimctR8RveZa4nI4IQWUnq o6lA== X-Google-DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20230601; t=1729138355; x=1729743155; h=mime-version:user-agent:message-id:date:subject:to:from :x-gm-message-state:from:to:cc:subject:date:message-id:reply-to; bh=lio7kJiOuA8CmWPW1GWsIYNeZOnOcjvhLIBU1Kjtoyg=; b=KU3XYAwwuHItV0Enar1BUXv8j7Qfb+2dM9erIvTp/J98UqUuaUtMyhAnZn7dwLGmsB dhJYeQ9QIH009GQhqTrnUp+ie2z7qkbX2FSs+FNMKjQRqADv7hOF11o7cw1+FSZQqZpD unsJp3sOYRHQP7MrmEq7fE+fjMQUOrxx2n3NcnDHrbT7mvTigJ5I9ihyfzG4HsJklgVA jSB/XNHMV2ajkkWpp8vUHk4aBE1AvCThnHlDz0mYbAoEv2dQpM+FZTNFKsUOow5Ox8i0 l7fOHZ7217P+tyWUetF1ak33V9mu6YGdSkceg+6E7F0jZZTdEcVPB1ysOz2a6KSw+B+A aBUQ== X-Gm-Message-State: AOJu0Yyz4/v0Qh+jA6gJzmLpIBomYihZPv6EP90XXSLHO1OneaHMDt4R F1IlnBou/7tHku4PJex4t0ddI6iTvM/BY5OcXKhAA45x3RScbVW2XcWjI6KU X-Google-Smtp-Source: AGHT+IHZ9KBfDZirPf0db48YsI2D1KfDE44cNs7yqeODnL4Ck1Vca550fa5ab/C1lZGeEr2+WmryqQ== X-Received: by 2002:a05:6a21:1643:b0:1d5:2f56:9fe5 with SMTP id adf61e73a8af0-1d8bcfb272bmr30208638637.39.1729138354919; Wed, 16 Oct 2024 21:12:34 -0700 (PDT) Original-Received: from localhost ([115.240.90.130]) by smtp.gmail.com with ESMTPSA id 98e67ed59e1d1-2e3e094bcafsm721027a91.54.2024.10.16.21.12.33 for (version=TLS1_3 cipher=TLS_AES_256_GCM_SHA384 bits=256/256); Wed, 16 Oct 2024 21:12:34 -0700 (PDT) Received-SPF: pass client-ip=2607:f8b0:4864:20::543; envelope-from=visuweshm@gmail.com; helo=mail-pg1-x543.google.com X-Spam_score_int: -20 X-Spam_score: -2.1 X-Spam_bar: -- X-Spam_report: (-2.1 / 5.0 requ) BAYES_00=-1.9, DKIM_SIGNED=0.1, DKIM_VALID=-0.1, DKIM_VALID_AU=-0.1, DKIM_VALID_EF=-0.1, FREEMAIL_FROM=0.001, RCVD_IN_DNSWL_NONE=-0.0001, SPF_HELO_NONE=0.001, SPF_PASS=-0.001 autolearn=ham autolearn_force=no X-Spam_action: no action X-BeenThere: debbugs-submit@debbugs.gnu.org X-Mailman-Version: 2.1.18 Precedence: list X-BeenThere: bug-gnu-emacs@gnu.org List-Id: "Bug reports for GNU Emacs, the Swiss army knife of text editors" List-Unsubscribe: , List-Archive: List-Post: List-Help: List-Subscribe: , Errors-To: bug-gnu-emacs-bounces+geb-bug-gnu-emacs=m.gmane-mx.org@gnu.org Original-Sender: bug-gnu-emacs-bounces+geb-bug-gnu-emacs=m.gmane-mx.org@gnu.org Xref: news.gmane.io gmane.emacs.bugs:293700 Archived-At: --=-=-= Content-Type: text/plain Tags: patch Hi Tassilo, This is a small patch to make djvused emit UTF-8 encoded text. In the djvu test file that I sent you, outline in the appendix have non-ASCII characters which are written as octal escapes. Rather than unescaping them on Emacs side, we can request djvused to use UTF-8 directly which this patch does. The attached patch does just that. In GNU Emacs 31.0.50 (build 13, x86_64-pc-linux-gnu, X toolkit, cairo version 1.18.0, Xaw scroll bars) of 2024-10-06 built on astatine Repository revision: 500f5da5fb62cd0bbded8df754d93e3147d1d847 Repository branch: master Windowing system distributor 'The X.Org Foundation', version 11.0.12101011 System Description: Debian GNU/Linux trixie/sid Configured using: 'configure --with-sound=alsa --with-x-toolkit=lucid --without-xaw3d --without-gconf --without-libsystemd --with-cairo CFLAGS=-g3' --=-=-= Content-Type: text/patch Content-Disposition: attachment; filename=0001-Make-djvused-emit-UTF-8-encoded-text.patch >From 8e21167c6e01ab76b76e15fa84bd198bc8df59b4 Mon Sep 17 00:00:00 2001 From: Visuwesh Date: Thu, 17 Oct 2024 09:40:34 +0530 Subject: [PATCH] Make djvused emit UTF-8 encoded text * lisp/doc-view.el (doc-view--djvu-outline): Pass -u to djvused to make it emit UTF-8 encoded text rather than using octal escapes for non-ASCII string. --- lisp/doc-view.el | 2 +- 1 file changed, 1 insertion(+), 1 deletion(-) diff --git a/lisp/doc-view.el b/lisp/doc-view.el index bbfbbdec925..018c4eddd34 100644 --- a/lisp/doc-view.el +++ b/lisp/doc-view.el @@ -2027,7 +2027,7 @@ doc-view--djvu-outline (unless file-name (setq file-name (buffer-file-name))) (with-temp-buffer (call-process doc-view-djvused-program nil (current-buffer) nil - "-e" "print-outline" file-name) + "-u" "-e" "print-outline" file-name) (goto-char (point-min)) (when (eobp) (setq doc-view--outline 'unavailable) -- 2.45.2 --=-=-=--